Synthszr Charts — die großen AI-Marken im Wettkampf ums Podium
synthszr charts

Unknown · v3 · depuis 25. März 2026 · 2× · vu le 02 oct. 2026

18
Momentum

ARC-AGI-3 is an interactive reasoning benchmark by the ARC Prize Foundation that tests AI agents in novel, instruction-free game environments. Instead of static puzzles, agents must explore rules, infer goals, build world models, and learn continuously across levels without guidance. The benchmark comprises hundreds of handcrafted turn-based environments; humans score 100% while frontier AI scored below 1% at launch. Access is via an open REST API and Python toolkit (MIT license), with results published on a public leaderboard.

Historique du momentum
04.07.02.10.

Fonctionnalités

Key Benchmark (%)Best score (leaderboard, Sep 2026): GPT-6 Astra 62.7% (up to 99.9% depending on harness); humans 100%; frontier AI at launch 0.51%
LicenseARC-AGI Toolkit/benchmarking code: MIT license; competition submissions must be CC0/MIT-0; private test set restricted
MultimodalityVisual/interactive: 2D pixel-grid game environments (e.g., 64x64), no language instructions, agents act via real-time actions
PlatformWeb-based game at arcprize.org/tasks; agent access via REST API and Python toolkit (arcprize/ARC-AGI on GitHub)
Release DateMarch 25, 2026 (official launch)

Plus de produits dans cette catégorie: Modèles de reasoning

Preuves (2)

Subscribe free. Unsubscribe the second it sucks.

High-signal news across AI, business, UX, and tech. Every morning.