AI.info
Xuelong Li
Explore Xuelong Li on AI.info.
- WNM-3D: A World Navigation Model with 3D Scene Conditioning for Closed-Loop VLN
- QwenStyle: Content-Preserving Style Transfer with Qwen-Image-Edit
- ViewMask-1-to-3: Multi-View Consistent Image Generation via Multimodal Discrete Diffusion Models
- Aetheria: A multimodal interpretable content safety framework based on multi-agent debate and collaboration
- Multimodal Continual Learning with MLLMs from Multi-scenario Perspectives
- Information Capacity: Evaluating the Efficiency of Large Language Models via Text Compression
- CAS-Spec: Cascade Adaptive Self-Speculative Decoding for On-the-Fly Lossless Inference Acceleration of LLMs
- A Parameter-Efficient Mixture-of-Experts Framework for Cross-Modal Geo-Localization
- Gestura: A LVLM-Powered System Bridging Motion and Semantics for Real-Time Free-Form Gesture Understanding
- UWBench: A Comprehensive Vision-Language Benchmark for Underwater Understanding
- CompassNav: Steering From Path Imitation To Decision Understanding In Navigation
- FastUMI-100K: Advancing Data-driven Robotic Manipulation with a Large-scale UMI-style Dataset