AI.info
Pang Wei Koh
Explore Pang Wei Koh on AI.info.
- Knowledge Distillation During Mid-Training Favors Reasoning over Factual Recall
- Buy versus Build an LLM: A Decision Framework for Governments
- Anchored Decoding: Provably Reducing Copyright Risk for Any Language Model
- Reliable and Responsible Foundation Models: A Comprehensive Survey
- Privasis: Synthesizing the Largest "Public" Private Dataset from Scratch
- DR Tulu: Reinforcement Learning with Evolving Rubrics for Deep Research
- RLVE: Scaling Up Reinforcement Learning for Language Models with Adaptive Verifiable Environments