Tencent Hunyuan releases ExplorationBench to test AI rule discovery
Original titlemore 🔛 https://github.com/Tencent-Hunyuan/ExplorationBench
AISummary
Tencent Hunyuan, with Fudan and Tsinghua researchers, released ExplorationBench, a benchmark that tests whether AI systems can discover rules through experiments in verifiable alien worlds. Across 10 frontier systems, feedback from experiments raised the best AlienCode score to 89.0% after four rounds, while closed-book runs without feedback stayed at 0.5–11.0%.
Source: Tencent Hunyuan · x.comPublished · added here