Grok 4.7 scores 1.8% on ARC-AGI-3 standard harness
Original title@SpaceXAI On ARC-AGI-3, Grok 4.7 scores 1.8% (vs Grok 4.7's 2.1%) in the standard harness, which lets models carry forward notes between ...
AISummary
Grok 4.7 scored 1.8% on ARC-AGI-3 in the standard harness, which lets models carry notes between turns, slightly below the 2.1% reported for Grok 4.7 in that setting. In a new provider adapter harness that preserves opaque reasoning and enables auto compaction, the score rose to 10.0%.
Source: ARC Prize · x.comPublished · added here