← Back to explorer

[schema]: Frontier Models with the Right Harness Achieve ~99% on ARC-AGI-3 Public

Type
paper
Venue
Impossible Research
Year
2026
Source
web
Access
free
Language
en
Added
2026-08-14T19:35:00Z
Verified
2026-08-14T19:35:00Z

Summary

Schema harness (not weight changes) for ARC-AGI-3: observe→deliberate→execute→record with an append-only timeline. Deliberation writes step(state,action), run_backtest over all recorded transitions, run_bfs inside the certified program, commit_actions only. Self-reported Public-set RHAE 98.98% with Opus 4.8+Fable 5 fallback (vs 42.83% same pairing in Claude Code) and 95.35% with GPT-5.6 Sol xhigh/max fallback. Official Sol max was 13.33% Public / 7.78% Semi-private — not a matched harness ablation. Scores unverified by ARC Prize. 50 released trajectories (25+25). Residual: Claude 19/25 games at 100 RHAE, six remaining 89.87–99.10; Sol 20/25 at 100.

Keywords

schema · arc-agi-3 · harness · world-model · rhae · impossible-research · blog · x

Topics

ARC-AGI, agent harnesses, symbolic world models

Research notes

  • Primary: Impossible Research blog (cite schema2026). Discord/X https://x.com/HavenFeng/status/2077770348876247502 via fxtwitter. Traces https://huggingface.co/datasets/schema-harness/arc-agi-3-schema-traces → datasets_local ARC-AGI-3 Schema Gameplay Trajectories. Earlier VIGA (state grounding) / WorldCoder (transitions) cited as one-sided predecessors. Self-reported public-set only.