Programmatic memory for long-horizon LLM agents: the harness appends everything to one log, and the agent searches it with code. 97.4% on ARC-AGI-3 (arXiv:2607.20064)
-
Updated
Aug 21, 2026 - Python
Programmatic memory for long-horizon LLM agents: the harness appends everything to one log, and the agent searches it with code. 97.4% on ARC-AGI-3 (arXiv:2607.20064)
The history files when recording human interaction while solving ARC tasks
My writings about ARC (Abstraction and Reasoning Corpus)
An agent skill that plays ARC-AGI-3. One rule: say what an action will do before you spend it. Claude Code on Opus 5 finished all 25 public games at 100.00 RHAE in 7,645 actions.
A Mimetic Procedural Benchmark Generator for the Abstraction and Reasoning Corpus
Enjoy puzzle-solving directly in your browser.
ARC gym: a data generation framework for the Abstraction & Reasoning Corpus
Artificial intelligence with a network of connected neurons
A python framework to streamline your ARC challenge solutions. From graphical displays to optimized Kaggle submissions
Source code for Reason to Play [NeurIPS '26]
SHYRS: A proof of concept for specifying how humans solve ARC-AGI-2 puzzles
Abstract Reasoning Chain-of-Thought (AR-CoT): is a neuro-symbolic cognitive architecture (AR-CoT) for LLM abstract reasoning. Solves perception & generalization bottlenecks via symbolic grounding, advanced geometric priming, and principle induction for the Kaggle ARC-AGI Prize 2025.
Prompts for solving ARC (Abstraction and Reasoning Corpus) with GPT4 or similar
demo benchmark for sample-blind evaluation
To associate your repository with the arc-agi topic, visit your repo's landing page and select "manage topics."