Story thread · 2 reports / 2 sources

AI’s best coding agent fails 60% of the time — and the data backs it up

thenewstack.io · 15h · first report

How the coverage leans

Across 2 sources · syndicated copies counted once

Claude Fable 5.1 just won a new coding benchmark despite failing more than six out of 10 times. Its 38.8% The post AI’s best coding agent fails 60% of the time — and the data backs it up appeared first on The New Stack .

The coverage

  1. StackHawk’s Wingman fixes security flaws while the AI agent is still coding

    siliconangle.com · 1h

The conversation · 0

Sign in to join the conversation.

No comments yet — start the thread.