Base44 reports fewer iterations across 118 app builds
In a customer testimonial published by Anthropic, Base44’s Gabriel Grinberg reports that Sonnet 5.5 matched Opus 5’s scores across 118 app builds, averaging 3.6 iterations per build versus 7.7 for Opus 5.
Éléments favorables
Extrait original
“Across 118 real app builds, Claude Sonnet 5.5 produced apps that scored level with Opus 5. It got there in 3.6 iterations per build on average, where Opus 5 took 7.7. It had the fewest failed tool calls of any model we compared. It also rarely stopped mid-build to ask the user a question, so fewer builds stall waiting on someone to answer.”