-
Xi'an Jiaotong University
- Xi'an, Shaanxi, China
-
18:43
(UTC +08:00) - https://scholar.google.com/citations?user=MpG_ZDEAAAAJ&hl=en
- @NingnanWang
Highlights
- Pro
Lists (12)
Sort Name ascending (A-Z)
Stars
InternRobotics' open platform for building generalized navigation foundation models.
Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)
[Official] AstraNav-Memory: Contexts Compression for Long Memory. An image-centric memory framework for lifelong embodied navigation via visual context compression and Qwen2.5-VL. SOTA on GOAT-Benc…
[AAAI 2026] Official repository of "Expand Your SCOPE, Semantic Cognition Over Potential-based Exploration for Embodied Visual Navigation"
A project page template for academic papers. Demo at https://eliahuhorwitz.github.io/Academic-project-page-template/
The repository provides code associated with the paper VLFM: Vision-Language Frontier Maps for Zero-Shot Semantic Navigation (ICRA 2024)
PyTorch code and models for VJEPA2 self-supervised learning from video.
Official code release for ConceptGraphs
[CVPR 2025] Source codes for the paper "3D-Mem: 3D Scene Memory for Embodied Exploration and Reasoning"
A flexible, high-performance 3D simulator for Embodied AI research.
[ICLR 2022] code for "How Much Can CLIP Benefit Vision-and-Language Tasks?" https://arxiv.org/abs/2107.06383
Code of the paper "NavCoT: Boosting LLM-Based Vision-and-Language Navigation via Learning Disentangled Reasoning" (TPAMI 2025)
[ECCV 2022] This is the official implementation of BEVFormer, a camera-only framework for autonomous driving perception, e.g., 3D object detection and semantic map segmentation.
[CVPR24] Volumetric Environment Representation for Vision-Language Navigation
[🎉IEEE TGRS'24] The official code for paper "CAMP: A Cross-View Geo-Localization Method using Contrastive Attributes Mining and Position-aware Partitioning"
[TCSVT'24] Enhancing Cross-View Geo-Localization with Domain Alignment and Scene Consistency