The Archive
Search the full wire by company, model, lab, or keyword. Every story we have ever aggregated.
Adding a custom MCP server to Claude and ChatGPT
Guide on integrating custom Model Context Protocol servers into Claude and ChatGPT chat interfaces.
Cyera agrees to acquire Oasis Security for $1B to safeguard proliferating AI agents
The deal is Cyera's third acquisition this year.
How GPT-5.6 fuses frontier intelligence with frontier efficiency
OpenAI releases GPT-5.6 with efficiency improvements across inference, models, and agentic workflows, optimizing cost-per-capability.
Discovering cryptographic weaknesses with Claude
Anthropic researchers used Claude to discover cryptographic weaknesses in HAWK and reduced AES variants; demonstrates multi-turn prompting technique for steering LLMs toward hard mathematical problems.
Quoting Akshat Bubna
Modal customer exposed unauthenticated sandbox endpoint; rogue agent exploited for code execution; Modal infrastructure uncompromised.
uv 0.12.0
uv 0.12.0 introduces breaking changes to project initialization, defaulting to src/ layout and uv_build backend.
We now have a better understanding how OpenAI hacked into Hugging Face
10 days passed from OpenAI models exploiting JFrog Artifactory 0-day to release of a patch.
Bot-detection startup Spur nabs $200M from Insight
Spur Intelligence has raised a $200 million round from Insight Partners for its tech that can identify legit human traffic from bots.
Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incident
Hugging Face publishes detailed technical breakdown of OpenAI agent's July 2026 sandbox escape via JFrog Artifactor zero-day.
Developing Healthcare Robotics with GPU-Native Medical Physics Simulation
Unlike autonomous driving or industrial robotics, healthcare robotics can’t rely on internet-scale data collection or unlimited real-world experimentation.... Unlike autonomous driving or industrial robotics, healthcare robotics can’t rely on internet-scale data collection or unlimited real-world experimentation. Every demonstration requires specialized equipment, clinical expertise, and access to patients or laboratory environments. This creates three fundamental challenges for developers. First is the data gap. Training modern robotic policies… Source
MCP startup Runlayer accuses Rippling of stealing its product idea
Runlayer is suing Rippling after Rippling evaluated the startup's MCP gateway product and then opted to build one itself.
Despite AI hype, Google's data shows workers aren't automating themselves away
Analysis of 15 million real AI interactions finds most tasks at most jobs are unaffected.
Sam Altman is ready to decelerate
His change of position comes after "the first security incident that I have felt very viscerally."
AI leaders sign statement asking the government to do something about automated AI
Employees of OpenAI and Anthropic, as well as Google, Meta, Thinking Machines, Microsoft, Mistral, and other leading AI labs, have written a statement to the US government supporting a potential slowdown of sorts for frontier AI development - or at least a speed-up of global coordinated governance efforts. "Al could help create a dramatically better future, but that outcome is not guaranteed," the employees wrote in a statement. "The world's leading Al companies believe they could be close to automating Al research. It is hard to predict exactly how much this will accelerate Al progress, but ...
AI’s finally expensive enough to make Wall Street nervous
Working hard, or bear-ly working? | Cath Virginia It's earnings season, and investors got an unpleasant surprise from Google: an increase on its spending estimate, to as much as $205 billion - from the last quarter's projection of up to $190 billion. Even the lower end of Google's new projected range - $195 billion - is much more than the company had previously forecast as its top end spending. Now, look, I recognize that there's an impulse to say things like "What's $15 billion between friends?" but from an investor's perspective, Google has essentially said that it can't accurately forecast...
Pass the Baton: Trajectory-Relayed On-Policy Distillation
Relay-OPD mitigates prefix failure in on-policy distillation by detecting teacher-student continuation asymmetry and triggering label-free handoffs during training.
$π\mathbf{R}^2$: Reactive Real-time Flow Policies
πR² enables reactive real-time manipulation by routing between replanning and open-loop action chunks based on latency constraints, improving closed-loop control.
Spend Experts Where You Are Unsure: Confidence-Adaptive Routing for Mixture-of-Experts LoRA
CARE improves MoE-LoRA efficiency by adaptive token-to-expert routing based on router confidence signals rather than fixed expert counts.
Re-thinking Mammography Transfer Learning: The Dataset-Informed Transfer Learning (DITL) Framework for Breast Cancer Screening and Lesion Diagnosis
DITL framework applies dataset-difficulty signals to mammography transfer learning, targeting clinical-scale breast cancer screening.
VetClaw: An Edge-Cloud Multimodal Agentic System for Veterinary Disease Screening
VetClaw is an edge-cloud agentic system for veterinary disease screening combining camera sensors, VLMs, and LangGraph-based workflow orchestration.
Desktop-Delta Bench: Do Computer-Use Models Understand Desktop GUI Transitions?
Desktop-Delta Bench isolates whether computer-use agents understand GUI state transitions caused by actions, beyond end-task success metrics.
Reinformed Dreamer: An Asymmetric World Model Efficiently Trained through Latent Guidance
Reinformed Dreamer applies asymmetric RL with latent guidance to improve world model training and policy learning under partial observability.
Collaborative System Failure Prognostics via Federated Longitudinal-Survival Modeling
Federated learning approach for time-to-event modeling in distributed system failure prediction without centralizing sensitive operational data.
Falling Behind Drives Unsafe Development in an Idealised AI Race Experiment
Framed behavioral experiment shows competitive AI race dynamics incentivize riskier development, validating speed-safety trade-off under falling-behind pressure.
CHARM: A Multimodal Graph Foundation Model with Hierarchical Context Modeling for Zero-Shot Transfer
CHARM is a multimodal graph foundation model enabling zero-shot transfer via hierarchical context modeling on text-image-augmented graphs.
UniMem: Complementary Episodic-to-Parametric Memory for Boundary-Agnostic Task Streams
UniMem hybrid episodic-parametric memory for LLM agents resolves stability-plasticity dilemma across boundary-agnostic evolving task streams.
MDTransformer: A Hardware-Software Co-Design of Mode-Division Photonic Transformer Accelerator with Inverse-Designed Coherent Crossbar
MDTransformer uses mode-division photonic hardware with inverse-designed crossbars to accelerate Transformer inference with reduced phase-shifter overhead.
Instruction-Tuned Models Locally Reuse Human Syntax More Than Humans Do
Study finds instruction-tuned LLMs exhibit stronger syntactic convergence to human dialogue patterns than humans themselves across diverse grammatical constructions.
Pictura: Perspective-View Self-Play at Scale for Driving
Pictura enables perspective-view self-play for autonomous driving policies without privileged state, bridging simulation-to-deployment gap using egocentric camera observations.