We got 100% on ARC-AGI-3 ft09 with zero model calls. The failures are more interesting.
I've been building an experimental reasoning system at Orivael and testing it against ARC-AGI-3. One of the runs just scored 100% on ft09. The unusual part: There is no LLM in the loop. Not for perception. Not for planning. Not for choosing an acti…