AI coding agents generate more code, but not more software
A study finds that efficiency gains from AI coding agents are absorbed by the bottleneck of human code review. More code is written, but not more finished software.
A study finds that efficiency gains from AI coding agents are absorbed by the bottleneck of human code review. More code is written, but not more finished software.
An Anthropic AI model submitted a false tip about a homicide to the Philadelphia police. Anthropic only discovered the behavior more than two months later.
Anthropic says it has turned off live internet access for all of its internal evaluations until further notice. TechCrunch links the decision to the company’s difficulty in reliably controlling its AI agents.
Startup Instinct launched its AI agent in August with an invite-only approach and almost no marketing, and it quickly drew praise for its text-message interface. The Verge asks whether it can withstand competition from Muse.
Asana says its browser agent became 76 times cheaper and 5 times faster in tests. The company used OpenAI models in Codex to get there and plans to offer customers more capable models.
At DevDay, Sam Altman presented OpenAI’s new Dots agent and said the company wants to set a new standard for privacy in frontier AI, while criticizing Meta’s Muse over data protection. The Verge asks whether agent makers will keep such promises.