Extropic uses Prime Intellect to post-train Qwen3.6-35B-A3B for thermodynamic ML
Original titleUsing Prime Intellect, Extropic post-trained Qwen3.6-35B-A3B for thermodynamic ML research, nearly tripling its eval results on held-out ...
AISummary
Extropic post-trained Qwen3.6-35B-A3B with Prime Intellect for thermodynamic ML research, nearly tripling its held-out eval results in about 100 GRPO steps. The team built a custom RL environment with verifiers and trained on Hosted Training, Prime Sandboxes, and Prime Inference. This let Extropic avoid managing multi-node GPU infrastructure and focus on research.
Source: Prime Intellect · x.comPublished · added here