AI.info
Ziwei Liu
Explore Ziwei Liu on AI.info.
- SenseNova-U1.5: Towards Native Unified Visual Intelligence
- VBVR-Pro: A Scalable and Verifiable Suite for Native Visual Reasoning
- Visual Representation Matters: Exploiting Temporal Differences in Video-to-Audio Generation
- Data Pyramid for Embodied Manipulation
- A Very Big Video Reasoning Suite
- DynamicVLA: A Vision-Language-Action Model for Dynamic Object Manipulation
- Continual GUI Agents
- Vision-Language-Action Models for Autonomous Driving: Past, Present, and Future
- WorldLens: Full-Spectrum Evaluations of Driving World Models in Real World
- Light-X: Generative 4D Video Rendering with Camera and Illumination Control
- Zero-Shot Video Translation and Editing with Frame Spatial-Temporal Correspondence
- U4D: Uncertainty-Aware 4D World Modeling from LiDAR Sequences
- Scaling Spatial Intelligence with Multimodal Foundation Models
- 3EED: Ground Everything Everywhere in 3D
- The Quest for Generalizable Motion Generation: Data, Model, and Evaluation
- IGGT: Instance-Grounded Geometry Transformer for Semantic 3D Reconstruction
- From Spatial to Actions: Grounding Vision-Language-Action Model in Spatial Foundation Priors
- RealDPO: Real or Not Real, that is the Preference
- Uni-MMMU: A Massive Multi-discipline Multimodal Unified Benchmark
- VideoLucy: Deep Memory Backtracking for Long Video Understanding