Discussion about this post

User's avatar
Saja's avatar

Genuinely enjoyed reading this! have a look at my lasted and first substack on AI and remaining human if you'd like :)

Vasanth's avatar

Good breakdown. The AGENTS.md point stands out to me: most teams treat it as a one-time README when it should be maintained like a living spec, and that's probably where real advantage shows up over time. I'd push back a bit on 'the model wars are over' though. SWE-bench convergence doesn't mean real-world reliability has converged too. Hoping Part 2 digs into failure modes, not just workflow fit.

1 more comment...

No posts

Ready for more?