CoreWeave opens Agent Lens preview for tracing agent failures
Image: Primary
Image: Primary
GitHub released a research preview of ReviewBench, a public benchmark that lets developers evaluate AI code reviewers against a common dataset and scoring method. It contains 219 pull requests from 187 public repositories spanning...
Anthropic has moved the virtual machine that executes Cowork's tool calls from users' computers into the cloud, Anthropic's Felix Rieseberg said. Model inference already ran in the cloud; the earlier design ran the virtual machine...
Anthropic's subscriptions offer ~5x more API-equivalent value per month than OpenAI's for agentic workloads using Claude Opus 5.5 versus GPT-6.1 Sol, SemiAnalysis finds. Its analysis tests subscription limits across Anthropic, Ope...
Meta and Microsoft are working to reduce employees' use of Claude, The Information reported, citing sources. Both companies are among Anthropic's biggest corporate customers. The number of Meta employees using Claude Code has fall...
AWS says Claude Opus 5.5 and Claude Sonnet 5.5 are available through Amazon Bedrock in AWS GovCloud (US), supporting AI coding workflows for organizations with regulatory requirements. Both models hold FedRAMP Class D certificatio...
The U.S. Department of Defence has stopped using Anthropic's AI tools, an official told the BBC. Sources told the broadcaster that Claude was still being used as recently as last week, including in military operations against Iran...