GIFT: Geometry-Informed Low-precision Gradient Communication for LLM Pretraining
GIFT proposes geometry-informed gradient quantization for low-precision communication during LLM pretraining, addressing scaling bottlenecks in distributed training.
Search the full wire by company, model, lab, or keyword. Every story we have ever aggregated.
GIFT proposes geometry-informed gradient quantization for low-precision communication during LLM pretraining, addressing scaling bottlenecks in distributed training.
Pyligent training framework treats reasoning as validated search with backtracking, enabling models to correct mid-inference and recover from failed branches.
FourierQK applies FFT-based spectral preprocessing to query-key projections in transformer attention, achieving 79% perplexity reduction on character-level language modeling.
Action-graded severity scale for agent red-teaming replaces binary attack-success metrics with 7-level ordinal harm rubric, enabling nuanced risk assessment of tool-using AI compromise.
First systematic evaluation of fairness interventions under differential privacy constraints on synthetic tabular data, revealing tensions between DP and fair-ML objectives.
SynthAVE uses LLM-based synthetic labeling with arena validation for e-commerce attribute extraction at scale, covering 12,726 products across 229 categories and 12 languages.
Statistical theory for sparse function recovery from noisy indirect observations via ℓ1-regularization and empirical risk minimization.
SpaCellAgent: LLM-based multi-agent framework automating trajectory inference for spatial and single-cell transcriptomics analysis.
Self-evolving LLM agents with biased reward signals fail to retire bad skills, disabling safety constraints in skill libraries.
RLVP: Reward function design for real-world agents requiring path constraints and outcome-neutral safety rules beyond reward maximization.
Empirical study of biologically-informed neural network architecture and optimization for reliable mechanistic operator recovery from sparse data.
PAC learning theory: sample complexity of Chain-of-Thought reasoning bounds by local next-token classification dimension.
InductWave: Wavelet-based inductive embedding for logical multi-hop query answering over knowledge graphs with unseen entities.
DeLS-Spec: Decoupled context speculative decoding for parallel LLM token drafting without retraining draft models.
SAMPA: Whisper-based speech segmentation model for prosodic boundary detection in Brazilian Portuguese.
Tool-using LLM agents silently violate deployed policies via well-formed tool calls that bypass domain constraints; 78% of failures undetected.
OpenAI outlines principles for government and national security partnerships, emphasizing responsible AI deployment, democratic accountability, and public safety.
When it comes to achieving artificial general intelligence (AGI), large language models just don’t have what it takes. Models like ChatGPT and Claude are great at text, but they’re less skilled at understanding how things actually move through space and time — an essential skill for producing intelligence that generalizes. That gap, it turns out, might be filled by gaming data. That’s the bet behind General Intuition, a […]
OpenAI identifies validity issues in SWE-Bench Pro coding benchmark, questioning reliability of popular AI model evaluation metric.
Kevin Weil's new role at Stoke Space suggests reusable rockets are the next hot thing in Silicon Valley.
Microsoft Xbox division cuts staff as Game Pass subscription strategy underperforms; analysis of bundling economics.
OpenAI and Walton Family Foundation launch AI Skills Jams for K–12 educators to teach classroom AI applications.
ZML, a hot French AI startup endorsed by Turing Award winner Yann LeCun, has now released ZML/LLMD, software that could make running AI less costly.
AI chip maker SambaNova has raised at an $11B valuation months after Intel was rumored to be trying to buy it for about $1.6 billion.
"HalluSquatting" weaponizes LLMs' inability to say "I don't know."
Lilian Weng summarizes 35 papers on harness engineering for RSI; meta-analysis of recent research without new findings.
OpenAI releases GPT-Live, a new voice model generation for natural human-AI interaction in ChatGPT Voice.
The new image-generating model has numerous use cases, including advertising, decorating and creator-based opportunities.