AI.info
Yufan Chen
Explore Yufan Chen on AI.info.
- Do LiDAR Language Models Really Understand Spatio-temporal Relationships?
- PROVIA: Procedure State Tracking for Online Mistake Detection in Egocentric Videos
- GuideFetch: A Task Coordination Framework for Concurrent Navigation and Object Retrieval in Assistive Robot Dogs
- X$^2$Localizer: Cross-grained Alignment for Progressive Cross-view Video Geo-localization
- $M^2$-Occ: Resilient 3D Semantic Occupancy Prediction for Autonomous Driving with Incomplete Camera Inputs
- HybriDLA: Hybrid Generation for Document Layout Analysis
- RefAtomNet++: Advancing Referring Atomic Video Action Recognition using Semantic Retrieval based Multi-Trajectory Mamba
- EReLiFM: Evidential Reliability-Aware Residual Flow Meta-Learning for Open-Set Domain Generalization under Noisy Labels
- Situat3DChange: Situated 3D Change Understanding Dataset for Multimodal Large Language Model