Skip to content
AI.info
News
Tools
Research
Learn
AI.info
Yinmin Zhang
Explore Yinmin Zhang on AI.info.
$Φ$-Bench: Can Large Language Models Engineer the Infrastructure That Powers Them?
PRIME: A Process-Outcome Alignment Benchmark for Verifiable Reasoning in Mathematics and Engineering
R-Align: Enhancing Generative Reward Models through Rationale-Centric Meta-Judging
Page 1