Latest BlogSWE-InfraBench + Kiro: What Happens When a Coding Agent Tackles IaC?
Builder Center
Language models solved only 34% of SWE-InfraBench's infrastructure-as-code tasks in one attempt. We ran Kiro over all 100 and found that a full agent loop with Sonnet 5 reaches 82%, with partially correct answers converging once the agent can re-run the tests itself.
View all 29 blog posts