Latent Space
RSS Feed · ANALYST
[AINews] Megakernels are so dead and so back
Cursor product launch and megakernel GPU engineering tradeoffs discussed in Latent Space newsletter roundup.
Unpacking ChatGPT Work: the Agent for a Billion Users
Technical reconstruction of ChatGPT Work's agent architecture: memory, scheduling, browser automation, plugins, and tool composition.
[AINews] Qwen 3.8 Max(2.4T) and 27B, new open weights models for Coding and Cowork
Alibaba releases Qwen 3.8 Max (2.4T params) and 27B open-weight models optimized for coding and collaboration tasks.
The Inference Engineering Masterclass — Philip Kiely & Ali Taha, Baseten
Baseten Series F funding and technical deep-dive on autoregressive and diffusion inference optimization strategies.
[AINews] GPT 5.6 price cut by 20%-80%: Cost of GPT 5.4 Intelligence dropped 13x in 4 months due to GPT 5.6 recursive self-optimization
GPT 5.6 pricing drops 20–80% via recursive self-optimization; equivalent GPT 5.4 intelligence now 13x cheaper in 4 months.
Ontologies Are So Back: Why AI Agents Are Reviving the Semantic Web
AI engineers adopt ontologies to constrain probabilistic agents within deterministic logical boundaries, reviving semantic web techniques.
[AINews] AI is eating Finance; AIE NYC now open
Latent Space observes AI adoption accelerating in financial services as the next major vertical after software engineering.
Codex from 0 to 10M Users: Building ChatGPT Work — Akshay Nathan, OpenAI
OpenAI's Akshay Nathan details ChatGPT Work product strategy: Sites, memory, subagents, finance, no-code tools scaling from 0 to 10M users.
[AINews] Much ado about Open Weights
Kimi K3 open-weights model released amid broader industry discussion on open model availability and strategy.
[AINews] Claude Opus 5: Fable-level performance at Opus price (half Fable)
Anthropic releases Claude Opus 5 matching Fable performance at half the cost, demonstrating efficiency gains in model distillation.
[AINews] Black Forest Labs FLUX 3 - Multimodal Flow Models that beat Seedance 2.0, Gemini Omni and Grok Imagine, and FLUX-mimic video-action robotics model
Black Forest Labs releases FLUX 3 multimodal model with reported improvements over Gemini 2.0, Grok Imagine, and includes video-action robotics variant.
[AINews] "Laguna S 2.1 Released: Cheaper than Deepseek v4 Flash, Better than V4 Pro"
Laguna S 2.1, a 118B MoE model from Poolside AI, achieves Deepseek v4 Pro performance at lower cost than v4 Flash.
Inside the Model Factory — Eiso Kant, Poolside AI
Poolside AI co-CEO Eiso Kant describes building a model factory enabling efficient training of 118B MoE models competitive with 1T open-weight alternatives.
[AINews] AI Cybersecurity becomes top of mind
Latent Space observes emerging trend in AI cybersecurity coverage without detailing specific breakthroughs or novel attacks.
Causal Models Need Causal Data - Xaira’s X-Cell model for Drug Discovery (Bo Wang & Ci Chu, Chief Discovery Officer & Chief AI Scientist)
Xaira Therapeutics builds causal models for drug discovery using synthetic data generation; Bo Wang and Ci Chu discuss data requirements for model training.
[AINews] Kimi K3 2.8T-A50B: the largest open model ever released; Opus 4.8-class at Sonnet 5 pricing
Kimi K3 2.8T-A50B released as largest open-weight model with Opus 4.8-class performance at Sonnet 5 pricing.
🔬 The Lab of the Future Should Feel Like a Data Center — Andy Beam & Rafa Gómez-Bombarelli, Lila Sciences
Lila Sciences argues scientific labs as data sources for AI training, positioning robotics and experimental workflows as frontier training data beyond internet corpora.
[AINews] Thinky's Inkling: 975B-A41B multimodal, new best American Apache 2.0 open model (with Inkling-Small, 276B-A12B)
Thinky releases Inkling, a 975B multimodal open-weights model under Apache 2.0, with a smaller 276B variant.
[AINews] not much happened today
Codex user base growing at 1M users daily; limited concrete details on adoption or product impact.
5 Trends That Defined AI Engineering at World’s Fair 2026
AIE World's Fair 2026 identified shift from agent-centric tooling to systems-level architecture design patterns.
[AINews] Codex usage up >10x in 6 months to 7M users, +1M in the past ~day; did Codex overtake Claude Code??
Codex usage grew 10x to 7M users in 6 months; article questions whether it has outpaced Claude Code amid sparse adoption metrics.
[AINews] not much happened today
Commentary noting an absence of major announcements following a week of model releases.
[AINews] SpaceXAI launches Grok 4.5, first Opus-class model post Cursor acquisition
SpaceXAI continues to move faster than any other frontier lab on earth.
Why AI Infrastructure must evolve for Agent Experience — Akshat Bubna, Modal CTO
Modal CTO Akshat Bubna discusses infrastructure requirements for agentic AI systems, covering lessons from building agent-native cloud platform.
[AINews] Lilian Weng summarizes 35 papers on Harness Engineering for RSI
Lilian Weng summarizes 35 papers on harness engineering for RSI; meta-analysis of recent research without new findings.
[AINews] The Field Guide to Fable
Commentary on a recent model launch framed as historically significant, lacks specifics on model capabilities, architecture, or performance.
AIEWF Daily Dispatch: The great loops debate and the state of AI engineering
AI Engineer World's Fair concludes with debate on agentic loops and report on engineering practices; keynotes address development priorities.
Vercel's Andrew Qu on why agents are a new kind of software
Vercel's Andrew Qu discusses eve agent framework, emphasizing skills, sandboxes, and agent-readable web design as architectural primitives.
The website of the future may assemble itself for every visitor
Adobe experiments with agentic sites that dynamically generate pages based on individual user intent, signaling shift toward personalized web experiences.
Skill engineering and the case against one-shot AI design
Paul Bakaus on Impeccable discusses skill engineering, human-in-the-loop agent design, and limitations of one-shot prompting approaches.
[AINews] not much happened today
Newsletter post noting absence of significant AI industry announcements on a given day.
AIEWF Daily Dispatch: Autoresearch and the tension between AI and human agency
AIEWF speakers debate autoresearch and software factory vision, raising concerns about human agency and control in AI-driven development.
Autoresearch: The feedback loop behind self-improving agents
Introspection co-founder explains autoresearch loops, agent recipes, and self-improving systems while arguing humans remain essential to AI software development.
How Cursor deploys AI inside the enterprise
Cursor's Forward Deployed Engineers help enterprises implement AI agents as software factories, per Pauline Brunet.
🔬 The Coolest Diffusion Research Isn't in LLMs — Evan Feinberg & Sergey Edunov, Genesis Molecular AI
Evan Feinberg and Sergey Edunov discuss diffusion models for drug discovery at Genesis Molecular AI, including PEARL's OpenBind performance and protein co-folding advances.
Warp CEO Zach Lloyd on why software factories are the next phase of coding
Warp CEO Zach Lloyd argues automated software factories will become standard for major projects, outlining preparation strategies for engineers.
AIEWF Daily Dispatch: Loops, Software Factories & Forward Deployed Engineers
AI Engineer World's Fair coverage: agent loops, software factories, forward-deployed engineering, and open model adoption emerging as key themes.
[AINews] Sonnet 5 today, and Fable 5 tomorrow
Latent Space newsletter teases upcoming model releases (Sonnet 5, Fable 5) with minimal detail; appears to be placeholder or speculative content.
Forward Deployed Engineers and the future of software engineering
Sierra's Natalie Meurer discusses convergence of product and forward-deployed engineers in software development.
Ahmad Osman on why local AI is catching up
Ahmad Osman argues local AI inference is rapidly closing performance/cost gap vs. cloud APIs across consumer and enterprise deployments.
[AINews] not much happened today
Commentary on a slow news day in AI; no substantive developments or announcements.
[AINews] OpenAI GPT-5.6 Sol / Terra / Luna — restricted to trusted partners
Oddly tiered releases to both OAI and ANT on the same day.
[AINews] OpenAI reports median internal Codex output tokens grew 56x in Research, 32x in Customer Support, 27x in Engineering, and 13x in Legal since November 2025.
OpenAI internal Codex usage surged 13–56x across departments since Nov 2025, with Research leading adoption.