Introducing Claude Opus 4.5
Anthropic releases Claude Opus 4.5 with improved coding, agents, computer use, and token efficiency.
Search the full wire by company, model, lab, or keyword. Every story we have ever aggregated.
Anthropic releases Claude Opus 4.5 with improved coding, agents, computer use, and token efficiency.
Notion rebuilt AI layer on GPT-5 to enable autonomous agents that reason and adapt across workflows in Notion 3.0.
Gemini Robotics 1.5 enables physical AI agents to perceive, plan, and execute multi-step tasks with tool use in real environments.
OpenAI releases AgentKit, expanded evals, and reinforcement fine-tuning tools for agent development.
SafetyKit product leverages GPT-5 for content moderation and compliance enforcement with improved accuracy over legacy systems.
Basis built AI agents using o3, o3-Pro, GPT-4.1, and GPT-5 delivering 30% time savings for accounting firms.
Anthropic publishes framework for developing safe and trustworthy autonomous agents with specified governance principles.
Outtake uses GPT-4.1 and OpenAI o3 agents to detect security threats 100x faster.
Model ML CEO discusses AI-native infrastructure and autonomous agents for financial services transformation.
Genspark built $36M ARR no-code agent product in 45 days using GPT-4.1 and OpenAI Realtime API.
Devstral: Mistral AI open-source model optimized for autonomous coding agents and software development.
OpenAI introduces BrowseComp benchmark for evaluating web browsing agent capabilities.
PaperBench: new benchmark measuring AI agents' ability to replicate state-of-the-art research papers.
OpenAI shifts from intent-based bots to proactive AI agents architecture.
Hebbia's AI platform claims to automate 90% of finance and legal work tasks using OpenAI models.
OpenAI released advanced text-to-speech and speech-to-text APIs with customizable voice instructions for voice agents.
xAI unveils early preview of Grok 3, emphasizing advanced reasoning and agentic capabilities.