A 150M param recurrent model scores 29.5% on ARC-AGI-1 at $0.0007 per task
Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.
| Not a transformer. It's a recurrent latent reasoning setup that keeps "thinking" in latent space before answering. Sits completely outside the published cost/accuracy frontier for ARC-AGI, and something this size runs on basically anything. Paper is from the Pathway team, dropped 4 days ago. I want to see it scaled to 1-3B before getting too excited, but the shape of the result is wild. [link] [comments] |
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.