NavGPT: Explicit Reasoning in Vision-and-Language Navigation with Large Language Models Paper • 2305.16986 • Published May 26, 2023 • 1
NaVid: Video-based VLM Plans the Next Step for Vision-and-Language Navigation Paper • 2402.15852 • Published Feb 24, 2024 • 1
Learning Goal-Oriented Language-Guided Navigation with Self-Improving Demonstrations at Scale Paper • 2509.24910 • Published Sep 29, 2025 • 5
VLNVerse: A Benchmark for Vision-Language Navigation with Versatile, Embodied, Realistic Simulation and Evaluation Paper • 2512.19021 • Published Dec 22, 2025 • 1
LiveWorld: Simulating Out-of-Sight Dynamics in Generative Video World Models Paper • 2603.07145 • Published Mar 7 • 7
Qwen-RobotWorld Technical Report: Unifying Embodied World Modeling through Language-Conditioned Video Generation Paper • 2606.17030 • Published Jun 15 • 49
Qwen-RobotManip Technical Report: Alignment Unlocks Scale for Robotic Manipulation Foundation Models Paper • 2606.17846 • Published Jun 17 • 36
Qwen-RobotNav Technical Report: A Scalable Navigation Model Designed for an Agentic Navigation System Paper • 2606.18112 • Published Jun 18 • 31
Embodied Agents Take Control: Minimal-Interface Zero-Shot Agents Rival Industrial-Scale Policies in Vision-and-Language Navigation Paper • 2607.26148 • Published Jul 28 • 2
Scaffolding Foundation Models into Physical-World Agents Pushes the Frontier of Long-Horizon Navigation Paper • 2608.30396 • Published Aug 31 • 12
Qwen-VLA: Unifying Vision-Language-Action Modeling across Tasks, Environments, and Robot Embodiments Paper • 2605.30280 • Published May 28 • 146
On Token's Dilemma: Dynamic MoE with Drift-Aware Token Assignment for Continual Learning of Large Vision Language Models Paper • 2603.27481 • Published Mar 29 • 33
Self-Evaluation Unlocks Any-Step Text-to-Image Generation Paper • 2512.22374 • Published Dec 26, 2025 • 17
Rethinking Training Dynamics in Scale-wise Autoregressive Generation Paper • 2512.06421 • Published Dec 6, 2025 • 7
Around the World in 80 Timesteps: A Generative Approach to Global Visual Geolocation Paper • 2412.06781 • Published Dec 9, 2024 • 24