Ting Huang

I am Ting Huang, a visiting researcher in Prof. Hao Tang's group at Peking University. I received my M.S. from Shanghai University of Engineering Science (SUES) and work on embodied AI.

My research focuses on spatial intelligence and generalist embodied systems, including 3D vision-language and vision-language-action (VLA) models for mobile robots. My long-term goal is to build systems that understand and interact with the physical world. Outside research, I enjoy sports and music.

Spatial Intelligence Embodied AI Generalist Robotics

πŸ”₯ News

  • 2026.06πŸŽ‰πŸŽ‰ MobileVLA-R1 MobileVLA-R1 GitHub stars, ConsiSpace, and OpenGround OpenGround GitHub stars were accepted to ECCV 2026.
  • 2026.05πŸŽ‰πŸŽ‰ Our RL Position Paper RL Position Paper GitHub stars was accepted to ICML 2026 Position Papers.
  • 2026.02πŸŽ‰πŸŽ‰ Multigranularity-3DQA was accepted to Expert Systems with Applications.
  • 2026.01πŸ’ΌπŸ’Ό I started a research internship at SRI-Robot, working on embodied AI.
  • 2026.01πŸŽ‰πŸŽ‰ We released 3D CoCa v2 3D CoCa v2 GitHub stars, a generalizable 3D captioning framework.
  • 2025.11πŸŽ‰πŸŽ‰ We released MobileVLA-R1 MobileVLA-R1 GitHub stars, an embodied model for mobile robots.
  • 2025.11πŸŽ‰πŸŽ‰ 3D CoCa 3D CoCa GitHub stars was accepted to 3DV 2026.
  • 2025.09🐧🐧 I joined Prof. Hao Tang's group at Peking University as a visiting researcher.
  • 2025.09πŸŽ‰πŸŽ‰ We released Nav-R1 Nav-R1 GitHub stars, an embodied foundation model.
  • 2025.08πŸŽ‰πŸŽ‰ We released 3D-R1 3D-R1 GitHub stars, an open-source generalist model for unified 3D scene understanding.

πŸ“ Publications

arXiv Preprint
3D CoCa v2 method overview

3D CoCa v2: Contrastive Learners with Test-Time Search for Generalizable Spatial Intelligence

Hao Tangβ€‘β˜…, Ting Huangβ˜…, Zeyu Zhangβ˜…

  • 3D CoCa v2 extends 3D CoCa with an inference-only test-time search (TTS) module and an external LLM judge.
ECCV 2026
Selected Publication

MobileVLA-R1: Reinforcing Vision-Language-Action for Mobile Robots

Ting Huangβ˜…, Dongjian Liβ˜…, Rui Yangβ˜…, Zeyu Zhangβ˜…β€ , Zida Yang, Hao Tang‑

MobileVLA-R1 GitHub stars

  • MobileVLA-R1 bridges language-guided high-level reasoning and continuous low-level control for mobile robots.
  • Leveraging a large VLA chain-of-thought dataset and reinforcement learning to produce interpretable plans and robust real-world execution.
arXiv Preprint

Nav-R1: Reasoning and Navigation in Embodied Scenes

Qingxiang Liuβ˜…, Ting Huangβ˜…, Zeyu Zhangβ˜…β€ , Hao Tang‑

  • Nav-R1 unifies dialogue, reasoning, planning, and navigation in a single embodied foundation model.
  • Using Nav-CoT-110K and GRPO-based RL rewards plus a Fast-in-Slow paradigm to achieve coherent long-horizon reasoning with low-latency control in 3D environments.
arXiv Preprint
3D-R1 method overview

3D-R1: Enhancing Reasoning in 3D VLMs for Unified Scene Understanding

Ting Huangβ˜…, Zeyu Zhangβ˜…β€ , Hao Tang‑

3D-R1 GitHub stars

  • A pioneering 3D-VLM leverages reinforcement learning and dynamic view selection to enhance reasoning capabilities in 3D scene understanding.
3DV 2026
3D CoCa method overview
Selected Publication

3D CoCa: Contrastive Learners are 3D Captioners

Ting Huangβ˜…, Zeyu Zhangβ˜…β€ , Yemin Wangβ˜…, Hao Tang‑

3D CoCa GitHub stars

  • Proposes 3D CoCa, a unified framework that jointly performs contrastive 3D-text alignment and 3D caption generation within one architecture, instead of relying on a two-stage "proposal-then-caption" pipeline.

πŸ₯‡ Awards & Scholarships

  • Dec. 2024 β€” Outstanding Master's Student Scholarship.
  • Dec. 2023 β€” Graduate Entrance Scholarship.
  • Oct. 2020 β€” First Place, ROBOCON National College Student Robot Competition.

πŸ“– Education

  • Sep. 2023 – Jul. 2026 β€” M.S., Shanghai University of Engineering Science.
  • Sep. 2018 – Jun. 2022 β€” B.S., Hebei University of Engineering.

😊 Academic Services

Conference Reviewer: AAAI 2026, AAAI 2027, 3DV 2026, ICRA 2026