A North Star for open AGI. Co-founders: @fchollet @mikeknoop. President: @gregkamradt. We're hiring mission-driven builders: arcprize.org/jobs

Earth
Pinned Tweet
Announcing ARC-AGI-3 The only unsaturated agentic intelligence benchmark in the world Humans score 100%, AI <1% This human-AI gap demonstrates we do not yet have AGI Most benchmarks test what models already know, ARC-AGI-3 tests how they learn
261
571
4,346
783,112
Gemini 3.8 Flash from @Google on ARC-AGI (Verified): - ARC-AGI-3: 10.4%, $4.4K (standard harness), 35.0%, $4.5K (provider adapter harness) - ARC-AGI-2: 89.2%, $0.40/task - ARC-AGI-1: 98.5%, $0.21/task Gemini 3.8 Flash stands out for its low cost and high scores.
18
31
540
53,603
Gemini 3.8 Flash is the first Gemini model tested on v3 with both our Standard harness, which carries forward notes, and a Provider Adapter harness, which preserves opaque reasoning and compacts older context. Full results: arcprize.org/results/google-…
2
44
4,815
New ARC Prize 2026 - ARC-AGI-3 High Score 19.4% by Lord Han Solo (new 1st place)
2
3
120
6,585
Opus 5.5 from @AnthropicAI on ARC-AGI (Verified): - ARC-AGI-2: 93.3%, $0.41/task - ARC-AGI-1: 98.5%, $0.16/task Opus 5.5 outscored Opus 5 by 2.9 percentage points on v2 and 1.0 on v1, at roughly 80% lower evaluation cost.
10
42
717
34,813
During ARC-AGI-3 testing, our API requests were frequently mistakenly classified as reverse engineering attempts, which prevented us from completing testing prior to model release. We're working with Anthropic to resolve the issue. Full results: arcprize.org/results/anthrop…
6
6
115
7,216
GPT-6 Luna from @OpenAI on ARC-AGI (Verified): - ARC-AGI-3: 0.19%, $241 (standard harness), 0.59%, $237 (provider adapter harness) - ARC-AGI-2: 59.3%, $0.062/task - ARC-AGI-1: 86.7%, $0.018/task GPT-6 Luna achieved similar scores to GPT-5.6 Luna at roughly 62% lower cost.
24
42
883
60,223
On ARC-AGI-3, GPT-6 Luna scores 0.19% using the Standard harness (which lets models carry notes between turns) and 0.59% using the Provider Adapter harness (which preserves opaque reasoning and uses compaction). Full results: arcprize.org/results/openai-…
2
1
50
7,649
ARC Prize 2026 - ARC-AGI-3 Progress Prize - 9 day left $37,500 in prizes are being awarded on *Sept 30th* to the top open source solutions Current standings: 1. Tufa Labs 2. Lord Han Solo 3. NVARC3 Which top 3 will claim the open source prize?
8
14
175
75,903
Dots3-Note Preview from @dotsstudioai on ARC-AGI (Verified): - ARC-AGI-2: 76.8%, $0.08/task Dots3-Note Preview sets a new SOTA score for open-weight models on our verified ARC-AGI-2 leaderboard.
12
16
302
47,441
We tested the Dots3-Note Preview model on a dedicated deployment set up by @Baseten. The estimated $0.08/task is based on Dots AI's listed token prices of $0.14/M input and $0.28/M output, rather than Baseten hosting charges.
1
25
4,081
ARC-AGI-4 will be a benchmark for autonomous open-ended innovation. It will continue our commitment to open-source, giving the research community a shared target for progress that benefits all of humanity. Despite rapid model progress, humans still significantly outperform AI at open-ended invention. This is the meta-skill that unlocks progress across every field of technology. Advanced AI capable of scientific innovation will lead to tremendous new technology, knowledge, and understanding. This is a positive-sum future. We are deeply committed to advancing it. Open source is the foundation for that progress. The knowledge behind frontier AI, not just the technology itself, should be broadly distributed among researchers, academics, and organizations. Any coordinated effort by the AI industry to reduce openness or concentrate access to frontier AI would undermine that positive-sum future. We are committed to advancing a future where everyone can contribute to and benefit from AI progress.
We Must Pace the Frontier: I’ve written a new essay on why the AI industry should slow down, with a three-part plan for doing so. Anthropic is unilaterally committing to the first of these steps. We’ll provide third-party evaluators with permanent, employee-level access to our systems, so that they can verify adherence to our safety measures, report on incidents, and assess models’ alignment during training. You can read the full post here: darioamodei.com/post/we-must…
94
240
2,666
325,231
Final call for applications to present at ARC Prize Research Summit in Boston on Oct 23rd Application deadline: September 14th, 2026, 11:59pm anywhere in the world Travel expenses will be covered for selected presenters
ARC Prize Research Summit 2026 October 23, 2026 - Boston, MA The ARC Prize community comes together to share new work, challenge ideas, and advance the frontier of machine reasoning towards general intelligence
3
6
32
6,469