arxiv:2602.01576
Reiss Koh
Reiss
AI & ML interests
None yet
Recent Activity
upvoted a paper about 16 hours ago
Eliciting Weak-to-Strong Generalization with On-Policy Reverse Distillation upvoted a paper 7 days ago
Language Models Can Control Their Own Attention upvoted a paper 2 months ago
LLM-as-a-Tutor: Policy-Aware Prompt Adaptation for Non-Verifiable RL