AI.info
Xuming He
Explore Xuming He on AI.info.
- Wiki-R1: Incentivizing Multimodal Reasoning for Knowledge-based VQA via Data and Sampling Curriculum
- DA-DPO: Cost-efficient Difficulty-aware Preference Optimization for Reducing MLLM Hallucinations
- Probing Scientific General Intelligence of LLMs with Scientist-Aligned Workflows
- GUI-Rise: Structured Reasoning and History Summarization for GUI Navigation
- NoisyGRPO: Incentivizing Multimodal CoT Reasoning via Noise Injection and Bayesian Estimation