Arman Cohan previews COLM 2026 work on RL and research agents
Original titleArman Cohan — Faculty Research Scientist 👇
AISummary
Ai2 faculty research scientist Arman Cohan shared a thread previewing his group's upcoming COLM 2026 presentations. The background post says the work covers reinforcement learning with metacognitive rewards, on-policy self-distillation with rubric rewards, and evolving research agents. The main post itself contains only a call to see the thread, so no results or figures are reported.
Source: Ai2 · x.comPublished · added here