I replaced our agent’s CLAUDE.md with a POMDP-style state-action graph, +16 to +20pts task success
Ran into this problem: a flat markdown file (or even a GraphRAG "brain") gives agents better context, but when a chain of actions fails, you cannot point to which stpe broke or what to fix. No explicability, a slow convergence, and no ceiling…