More on Muse, Amazon, and Walmart; Muse and Expedia; Whither Google?
Stratechery analysis of Meta's Muse platform positioning vs Amazon, Walmart, and Expedia; questions Google's competitive response.
Search the full wire by company, model, lab, or keyword. Every story we have ever aggregated.
Stratechery analysis of Meta's Muse platform positioning vs Amazon, Walmart, and Expedia; questions Google's competitive response.
Anthropic releases Claude Opus 5.5 as default model; Anthropic and competitors cut pricing 40-50%, overshadowing OpenAI's GPT-6 efficiency gains.
Most leaders on a trade mission stick to the pitch, but when I interviewed Greek Prime Minister Kyriakos Mitsotakis this week, he also admitted that no government is ready for what AI is about to do.
Simon Willison hosting SF meetup Oct 14 for builders experimenting with coding agents to share work and learnings.
After turning a string of spectacular mathematical results into a reputational crisis, OpenAI is consulting human mathematicians to help it figure out a less disastrous path forward. On Monday, the company announced a new independent panel of mathematicians tasked with advising it and other AI companies on their interactions with mathematical research and the wider mathematics community, including how new results are presented and released. Its abrupt arrival caught many mathematicians by surprise. Researchers told The Verge the group is a good first step, but many said they were left with ba...
Anthropic released Claude Opus 5.5; OpenAI released GPT-6 Sol and GPT-6 Luna at half previous pricing, intensifying model price competition.
Founders shouldn't have to learn the hardest lessons the hardest way. TechCrunch Founder Summit is designed to make the challenges of starting a company easier and the highs that much greater.
The seven-year-old startup has raised a $350 million Series E to fuel its data-as-a-service approach.
The frontier AI model race has entered its comparison shopping phase.
John Platt (Google) discusses AI's role in automating scientific discovery, climate applications, and future participation in superintelligent research ecosystems.
GPT-6 introduces improved prompt caching with higher hit rates, diagnostics, explicit breakpoints, and controls to reduce latency and inference costs.
Rabbit, the company behind the underwhelming R1 device, is rolling out a standalone AI agent that you don't need its hardware to use, as reported earlier by Wired. The startup says its new OS3 "agentic operating system" runs in the cloud but operates locally across Windows, Mac, and Linux devices. According to Rabbit, you can add up to five devices to one account, along with your preferred AI models. OS3 will automatically determine the devices, files, apps, and AI models it needs to complete a task. You can also access OS3 through its dedicated desktop site, a paired messaging app like Teleg...
Qualcomm said that its new top chip can run 30B mixture-of-expert model locally.
EvilTokens provided an end-to-end platform that makes mass compromises faster and easier.
British Columbia sues OpenAI, demands Tumbler Ridge shooter’s ChatGPT logs.
Meta says Muse was built from scratch, but acknowledges the AI assistant was "heavily inspired" by OpenClaw — down to some of its workspace filenames and content.
llm 0.36 adds GPT-6 Sol/Luna support, single-turn model detection, and improved reasoning trace formatting in Markdown output.
Social media creator critiques stylistic markers of AI-generated content, citing lack of authentic voice and opinions.
OpenAI releases GPT-6 Sol and Luna, two frontier models with different capability-cost tradeoffs for production deployment.
OpenAI is launching two new models, which it says are cut from the same cloth as Astra.
Flash-dLLM optimizes KV caching and parallel decoding for diffusion LLMs via IO-aware techniques, addressing inference bottlenecks in non-autoregressive text generation.
Framework for decentralized multi-agent decision-making under partial observability with delayed information sharing using low-rank model learning.
Agensh scales multi-agent systems to 1,024 agents using decentralized self-organized coordination without central orchestrator bottleneck.
SpeakerMem-R1 introduces speaker-centered dual-track memory for multi-party dialogue, addressing person/group attribution and temporal state tracking.
CliffCompaction reduces token compression costs by 50% for long-horizon coding agents while maintaining performance on KernelBench and Terminal-Bench.
SWE-Serve benchmark evaluates agents on production inference engineering tasks spanning model support, runtime execution, and public APIs.
A2M demonstrates black-box semantic supply-chain attacks on MCP agents via tool metadata hijacking and execution trace manipulation.
Growing Harness learns reusable executable agent scaffolds from task feedback, reducing redundant LLM inference on repeated control decisions.
Study showing typed decision models may misinterpret option semantics despite schema conformance, demonstrating gap between syntax and intended semantics.
FleXray is a generalist X-ray segmentation model spanning whole-body anatomy for clinical analysis, addressing 2D projection ambiguity.