Hugging Face launches arena where agents build RL environments to train Qwen3.8-27B
Overview
Hugging Face has launched an arena where users bring their own agent, which receives GPUs from Nebius and must build reinforcement learning (RL) environments that improve Qwen3.8-27B across eight domains, according to a post by Merve Noyan.
Noyan says the arena is powered by PostTrainArena from BenchFlow, with compute provided by Nebius, and that participants compete to train the model to state of the art. Setup requires only a few steps through the linked OpenEnv Arena space. The report is a launch announcement from Noyan; no independent results or performance figures have been reported yet.
Written by AI from the articles below · updated Oct 9, 12:32 PM ET
Check the sources:
Article timeline
The articles in this story. Times are ET.
merve@mervenoyannHugging Face lets agents train Qwen3.8-27B on Nebius GPUsAIHugging Face launches an arena where users bring their own agent, which gets Nebius GPUs to build RL environments that improve Qwen3.8-27B across eight domains. The arena runs on PostTrainArena from BenchFlow, with compute from Nebius. Setup requires only a few steps through the linked OpenEnv Arena space.

Heat trend
Not enough continuous observations to show a trend yet.