Skip to content
Trending storyDeveloping

Hugging Face launches arena where agents build RL environments to train Qwen3.8-27B

1 article1 sourcesince Oct 9Last article 2h ago ·

Overview

AISummary of 1 article

Hugging Face has launched an arena where users bring their own agent, which receives GPUs from Nebius and must build reinforcement learning (RL) environments that improve Qwen3.8-27B across eight domains, according to a post by Merve Noyan.

Noyan says the arena is powered by PostTrainArena from BenchFlow, with compute provided by Nebius, and that participants compete to train the model to state of the art. Setup requires only a few steps through the linked OpenEnv Arena space. The report is a launch announcement from Noyan; no independent results or performance figures have been reported yet.

Written by AI from the articles below · updated Oct 9, 12:32 PM ET

Check the sources:

Article timeline

The articles in this story. Times are ET.

Oct 9
  1. merve
    Hugging Face lets agents train Qwen3.8-27B on Nebius GPUs

    AIHugging Face launches an arena where users bring their own agent, which gets Nebius GPUs to build RL environments that improve Qwen3.8-27B across eight domains. The arena runs on PostTrainArena from BenchFlow, with compute from Nebius. Setup requires only a few steps through the linked OpenEnv Arena space.

    Video from @mervenoyann's post

Heat trend

Not enough continuous observations to show a trend yet.