Builder Center2026SWE-InfraBench + Kiro: What Happens When a Coding Agent Tackles IaC?
Language models solved only 34% of SWE-InfraBench's infrastructure-as-code tasks in one attempt. We ran Kiro over all 100 and found that a full agent loop with Sonnet 5 reaches 82%, with partially correct answers converging once the agent can re-run the tests itself.




























