AI.info
Yan Lu
Explore Yan Lu on AI.info.
- Mage-Flow: An Efficient Native-Resolution Foundation Model for Image Generation and Editing
- SciForma: Structure-Faithful Generation of Scientific Diagrams
- Temperature as a Meta-Policy: Adaptive Temperature in LLM Reinforcement Learning
- Closing the Modality Reasoning Gap for Speech Large Language Models
- InfiniteWeb: Scalable Web Environment Synthesis for GUI Agent Training
- Generative Latent Coding for Ultra-Low Bitrate Image Compression
- Probing Scientific General Intelligence of LLMs with Scientist-Aligned Workflows
- CustomX: Unified Character, Action, and Scene Customization in Video World Models
- Single-step Diffusion-based Video Coding with Semantic-Temporal Guidance
- VideoVLA: Video Generators Can Be Generalizable Robot Manipulators
- Divide, then Ground: Adapting Frame Selection to Query Types for Long-Form Video Understanding
- CoD: A Diffusion Foundation Model for Image Compression
- VFM-VAE: Vision Foundation Models Can Be Good Tokenizers for Latent Diffusion Models