Evaluating AI’s ability to perform scientific research tasks
OpenAI introduces FrontierScience benchmark measuring AI reasoning across physics, chemistry, and biology research tasks.
Search the full wire by company, model, lab, or keyword. Every story we have ever aggregated.
OpenAI introduces FrontierScience benchmark measuring AI reasoning across physics, chemistry, and biology research tasks.
OpenAI presents evaluation framework using GPT-5 to measure AI acceleration of wet-lab biological research, including molecular cloning optimization.
Meta introduces SAM Audio, unified multimodal model for audio source separation accepting text, visual, and temporal prompts.
OpenAI publishes organizational guidance on AI readiness covering strategy, training, governance, and innovation acceleration.
OpenAI upgrades ChatGPT Images with GPT-Image-1.5, delivering 4× faster generation, precise edits, and consistent detail.
BBVA deploys ChatGPT Enterprise across 120,000 employees via multi-year partnership with OpenAI for banking transformation.
OpenAI demonstrates rapid Android development: Sora shipped in 28 days using Codex for AI-assisted planning and parallel coding workflows.
BNY Mellon deploys OpenAI agents at scale via Eliza platform, enabling 20,000+ employees to build and deploy AI for efficiency gains.
GPT-5.2 achieves state-of-the-art on GPQA Diamond and FrontierMath, solves open theoretical problems and generates formal proofs.
Google DeepMind and UK AI Security Institute (AISI) strengthen collaboration on critical AI safety and security research
Disney and OpenAI license 200+ characters for Sora video generation; Disney adopts ChatGPT Enterprise and OpenAI API.
xAI partners with El Salvador government on nationwide AI education initiative.
GPT-5.2 released with state-of-the-art reasoning, long-context, coding, vision; available via ChatGPT and OpenAI API.
GPT-5.2 System Card documents safety mitigations aligned with prior GPT-5 and GPT-5.1 methodologies.
Podium deployed GPT-5 AI agent 'Jerry' to 10,000+ SMBs, achieving 300% customer engagement growth.
Deepening our partnership with the UK government to support prosperity and security in the AI era
OpenAI details cybersecurity risk mitigation and defensive capabilities as AI model power increases.
Anthropic donates Model Context Protocol and establishes Agentic AI Foundation for agent standards.
Scout24 built GPT-5 conversational real-estate assistant with clarifying questions and personalized listing recommendations.
Accenture and Anthropic announce multi-year partnership to productionize enterprise AI deployments.
Mistral ships Devstral 2 and Mistral Vibe CLI, open-source agentic coding models with autonomous agent scaffolding.
Systematically evaluating the factuality of large language models with the FACTS Benchmark Suite.
OpenAI co-founds Agentic AI Foundation under Linux Foundation; donates AGENTS.md standard for interoperable agentic AI.
OpenAI launches certification courses and AI Foundations training to build workforce skills.
OpenAI and Deutsche Telekom deploy ChatGPT Enterprise across European telecom infrastructure.