AI.info
Qi Han
Explore Qi Han on AI.info.
- onPanda: Efficient Annotation of On-Policy Alignment Data for LLMs and Agents via Token-Level Correction
- $Φ$-Bench: Can Large Language Models Engineer the Infrastructure That Powers Them?
- PRIME: A Process-Outcome Alignment Benchmark for Verifiable Reasoning in Mathematics and Engineering
- R-Align: Enhancing Generative Reward Models through Rationale-Centric Meta-Judging