Avatar of Jiting Liu

Jiting Liu

The University of Hong Kong

Research Intern at The University of Hong Kong with interests in embodied intelligence and spatial reasoning.

  • About
  • Publications
  • CV

Publications

Selected work on spatially grounded VLA models, depth estimation, and embodied perception.

Visual summary for Evo-Depth: A Lightweight Depth-Enhanced Vision-Language-Action Model
Co-first author Evo-Depth: A Lightweight Depth-Enhanced Vision-Language-Action Model
arXiv
#Embodied AI #VLA #Spatial Perception #Depth #3D Understanding

A lightweight VLA model that incorporates depth cues to improve spatial grounding and embodied manipulation.

Visual summary for Evo-1: Lightweight Vision-Language-Action Model with Preserved Semantic Alignment
Evo-1: Lightweight Vision-Language-Action Model with Preserved Semantic Alignment
CVPR
#Embodied AI #VLA #Robot Learning #Semantic Alignment

A lightweight VLA model that preserves semantic alignment while improving efficiency for embodied tasks.

Visual summary for Focusable Monocular Depth Estimation
Focusable Monocular Depth Estimation
May 2026 arXiv
#Depth Estimation #Spatial Perception #3D Understanding #Embodied Perception

A region-aware monocular depth estimation framework that focuses depth modeling on target regions while preserving coherent global geometry.

View
Visual summary for Evo-0: Vision-Language-Action Model with Implicit Spatial Understanding
Evo-0: Vision-Language-Action Model with Implicit Spatial Understanding
arXiv
#Embodied AI #VLA #Spatial Perception #Robot Manipulation

A VLA model that improves embodied manipulation through implicit spatial understanding.

© 2026 Jiting Liu.
Built with Academic Portfolio Astro