Amazon launches new $1 billion FDE org, following OpenAI and Anthropic
Engineers on the new team will embed within companies to deploy purpose-built agents, focusing on fast deployments and customer self-sufficiency.
Search the full wire by company, model, lab, or keyword. Every story we have ever aggregated.
Engineers on the new team will embed within companies to deploy purpose-built agents, focusing on fast deployments and customer self-sufficiency.
Users will be able use AI to create newsletters based on their recordings.
SpikeLogBERT applies spiking neural networks for energy-efficient log parsing vs. dense transformer approaches.
Proton's Lumo 2.0 is dropping this week, giving users a broader variety of capabilities.
Adversarial distillation improves certified robustness of neural networks by combining tight relaxation bounds with adversarial training techniques.
FARS system autonomously generates, executes, and writes AI research across topics at scale using coordinated agent architecture.
ECHO introduces selective turn memory and pruning for long-horizon agentic RL under context window constraints.
LuckyStar 111B hybrid reasoning model from Cohere and LG CNS enables efficient multilingual tool-using agents with Korean-English support.
LLMs exhibit performative compliance: fairness evaluations overestimate moral safety when demographic identity must be inferred rather than labeled.
Tone-conditioned curriculum learning improves zero-shot ASR for 6 Southern Bantu languages using hybrid difficulty scoring and gated adapters.
Comprehensive survey systematizes LLM attack surface across full lifecycle: data pipelines, agents, tools, memory, and organizational integration.
Intrinsic decomposition extended to 3D Gaussian splatting for texture editing independent of lighting.
LLM agents act as constrained supervisory planners for fault recovery in process plants, validated against external safety constraints.
Bayesian workflow calibration detects and repairs statistical errors in probabilistic programs written by LLMs using posterior checks and diagnostics.
Philosophy of science frameworks applied to evaluating explainability standards for medical AI systems.
LLM + knowledge graph framework automates cause-effect specifications for industrial process control and safety systems.
Higher-order structural alignment method for multi-view radar semantic segmentation in adverse weather.
CLExEval: human-in-the-loop evaluation framework for LLM clinical reasoning with 5,600 physician annotations.
Uncertainty-guided diffusion model for synthetic data augmentation in semantic segmentation with sparse labels.
Dual-Embedding Watermarking (DEW) scheme for LLM text watermarking robust to paraphrasing and translation.
Theoretical framework for optimal training-calibration data splitting in split conformal prediction.
ViToS: dual-stream RL framework for visual token pruning in medical multimodal reasoning on sparse visual evidence.
Evaluation of ML-based intrusion detection systems on Gotham2025 IoT network security dataset.
Optimizer choice drives 7x variance in emergent misalignment severity in fine-tuned Qwen models.
Requirements engineering framework for ML systems to ensure alignment with stakeholder needs and trustworthiness.
ZEBRA framework addresses base-to-novel generalization gap in audio-language models via entropy-regularized prompt learning.
DPPE rethinks camera-based positional encoding for scaling multi-view transformers in 3D vision and novel view synthesis.
Proposes robustness measure bounding neural network MSE under input perturbations via black-box analysis.
Localized conformal prediction improves uncertainty quantification for image classification with vision-language models.