The Archive
Search the full wire by company, model, lab, or keyword. Every story we have ever aggregated.
Quoting Anthropic
US Department of Commerce lifts export controls on Anthropic's Claude Fable 5 and Mythos 5; access restoration begins tomorrow.
Ahmad Osman on why local AI is catching up
Ahmad Osman argues local AI inference is rapidly closing performance/cost gap vs. cloud APIs across consumer and enterprise deployments.
Nano Banana 2 Lite
Google releases Gemini 3.1 Flash Lite, optimized for fast, low-cost image generation; author tests visual search capability.
OpenClaw is finally available on Android and iOS
The free open source agentic program is finally invading your phone.
Claude Science is Anthropic’s newest flagship product
At an event for pharmaceutical executives, biotech founders, and researchers on Tuesday, Anthropic announced Claude Science, a major new product intended to support scientific research in the same way that Claude Code supports software engineering. Like Claude Code, Claude Science can autonomously carry out meaningful work when given concise, high-level instructions, and it has access…
What's new in Claude Sonnet 5
Claude Sonnet 5 launched with performance near Opus 4.8 at lower cost; includes cyber-task restrictions aligned with Opus 4.7/4.8 safeguards.
The DeepMind trio who built a poker AI are now making money for quant hedge funds
EquiLibre Technologies, a Prague-based AI lab founded by three ex-DeepMind researchers, is now valued at more than $500 million.
New attack provides one more reason why AI browsers are a bad idea
Telling an LLM that 2 + 2 = 5 is enough to make it follow forbidden instructions.
Google’s NotebookLM can sum up your research in a TikTok-style clip
Google's NotebookLM is adding a new way to catch up on your notes: TikTok-style AI videos. The new feature is rolling out to Google AI Ultra and Pro subscribers, allowing NotebookLM to generate 60-second vertical AI clips based on the sources you upload to the app. The example shared by Google details Australia's unsuccessful war on emus, pairing paper cutout-style AI art of emus with narration. It adds to some of the other ways NotebookLM lets you interact with your research, including by generating AI podcasts, cinematic videos, and visual explainers. Doom scrolling but make it educational ...
Google introduces a faster, cheaper image generator with Nano Banana 2 Lite
Google is updating its image generator to make it faster and cheaper, making it a more useful tool for creators looking to make AI content.
Google's new Nano Banana 2 Lite image model is its fastest and cheapest yet
They may not look as good, but Nano Banana 2 Lite images only take a few seconds to create.
Nvidia competitor Etched hits $5B valuation, $1B in sales for AI chip
Nvidia AI chip competitor Etched says it has already booked $1 billion under contract for the inference systems powered by its chip.
Anthropic launches Claude Sonnet 5 as a cheaper way to run agents
Anthropic’s Claude Sonnet 5 brings stronger agentic capabilities, lower pricing, and improved safety, positioning the model as a cheaper alternative to Opus, GPT-5.5, and Gemini Pro.
Introducing Claude Sonnet 5
Anthropic releases Claude Sonnet 5, a frontier model optimized for coding, agents, and professional workflows at scale.
Introspective Coupling: Self-Explanation Training Tracks Behavioral Change Despite Fixed Supervision
LMs trained on counterfactual explanations from earlier checkpoints produce explanations more faithful to current behavior than source models, advancing mechanistic interpretability.
QVal: Cheaply Evaluating Dense Supervision Signals for Long-Horizon LLM Agents
QVal proposes efficient evaluation method for dense supervision signals in long-horizon LLM agents without expensive end-to-end training.
Reinforcement Learning with Metacognitive Feedback Elicits Faithful Uncertainty Expression in LLMs
RL with metacognitive feedback trains LLMs to accurately express uncertainty and improve performance through self-monitoring.
When LLMs Read Tables Carelessly: Measuring and Reducing Data Referencing Errors
Systematic evaluation reveals data referencing errors (1.7B-20B models misread tables) as overlooked reliability issue beyond final-answer accuracy.
Freeform Preference Learning for Robotic Manipulation
Freeform Preference Learning enables robot policy training from natural-language preference axes instead of binary comparisons.
AdaJEPA: An Adaptive Latent World Model
AdaJEPA adapts latent world models at test time via MPC loop to handle distribution shift in robotic planning.
Generative Skill Composition for LLM Agents
Generative Skill Composition uses LLM-guided retrieval to select and compose skills for complex agent tasks without exposing full skill library.
Acti puts AI agents directly into your smartphone keyboard
Startup Acti is betting the smartphone keyboard is the next home for AI assistants. Its new keyboard for iOS and Android works across apps and lets users create custom AI-powered shortcuts using natural language.
FLORA: A deep learning approach to predict forest attributes from heterogeneous LiDAR data
FLORA applies deep learning to LiDAR data for forest attribute prediction under heterogeneous acquisition conditions.
SemRF: A Semantic Reference Frame for Residual-Stream Dynamics in Language Models
Semantic Reference Frames (SemRF) standardize residual-stream analysis across LM layers to distinguish computation from measurement drift.
Automated Background Swapping for Robustness against Spurious Backgrounds
Automated Background Swapping reduces classifier reliance on spurious background correlations through synthetic data augmentation.
TRIAGE: Role-Typed Credit Assignment for Agentic Reinforcement Learning
TRIAGE framework improves credit assignment in agentic RL by classifying action segments into semantic roles, addressing limitations of uniform outcome-based reward.
FedLAB: Traceable Semantic Codebooks for Federated Multimodal Graph Foundation Learning
FedLAB enables federated learning of multimodal graph foundation models with semantic traceability while preserving privacy across decentralized clients.
Scalable Behaviour Cloning on Browser Using via Skill Distillation
Skill distillation approach scales browser agent training via behavior cloning from implicit priors in human browsing traces rather than low-level operation.