Databricks’ former AI chief thinks he can cut AI’s power bill by 1,000x
Un0 is an image-generation system tool that shows for the first time how the company's technology can replicate conventional AI systems.
Search the full wire by company, model, lab, or keyword. Every story we have ever aggregated.
Un0 is an image-generation system tool that shows for the first time how the company's technology can replicate conventional AI systems.
Generative AI workloads are rapidly outgrowing the memory and compute budget of single GPUs. For inference developers building media generation pipelines, the... Generative AI workloads are rapidly outgrowing the memory and compute budget of single GPUs. For inference developers building media generation pipelines, the challenge is scaling across multiple devices without sacrificing the critical optimizations—like kernel fusions, memory planning, and quantization—that NVIDIA TensorRT delivers for production deployments. Multi-device inference support… Source
AI companions in games have long been constrained by scripted behavior trees and fixed dialogue. PUBG Ally is a different kind of system. Built by KRAFTON for... AI companions in games have long been constrained by scripted behavior trees and fixed dialogue. PUBG Ally is a different kind of system. Built by KRAFTON for PUBG: BATTLEGROUNDS, this AI teammate is powered by NVIDIA ACE and its suite of efficient models and tooling. PUBG Ally uses automatic speech recognition, a 2B-parameter small language model, and text-to-speech to understand player… Source
Framework for persistent embodied agents combining cyber-physical action spaces with autonomous failure recovery in unstructured environments.
RSPC benchmark: 1,799 Reddit posts annotated by psychiatrists for mental health conditions in relational contexts.
Theoretical analysis of GAN training dynamics with structured latent covariance and class-dependent discriminators.
Training-free source selection for LLMs using Fisher alignment metrics at vocabulary scale for scientific domains (SMILES, proteins).
Mechanistic analysis revealing LMs encode factual knowledge in task-specific rather than unified manner, with limited cross-task consistency.
Large-scale study of AI-generated non-consensual sexual imagery on 4chan: 24,105 items, 55.8% non-celebrity targets, shift from celebrity focus.
Conceptual framework for analyzing dialogue in human-AI and multi-agent collaborative problem-solving with hierarchical coding scheme.
CARVE: content-aware recurrent architecture with improved value gating enabling WY-form chunk-parallel training competitive with Transformers.
Formal semantics framework modeling co-evolution of lexical meanings and composition functions under simplicity and accuracy pressures.
BINEVAL: LLM evaluation framework decomposing criteria into atomic binary questions for interpretable, multi-dimensional scoring of open-ended outputs.
Hierarchical Muon reduces Newton-Schulz optimizer complexity via tiled updates, cutting dense-layer training compute from O(r²sK) to subquadratic.
GAversary uses genetic algorithms to craft adversarial text attacks on NLP classifiers, targeting semantically similar token swaps.
AIMS dataset (1.7K prompts) shows intent-aware safety classifiers outperform standard SFT across DPO, reasoning distillation, and RL training.
Syntactic belief-update model predicts garden-path sentence processing difficulty better than lexical surprisal alone.
Survey connects GNN expressiveness to Weisfeiler-Leman hierarchy, clarifying when graph structure justifies computational overhead.
Feature-induced information flow method explains Event-based Temporal GNNs by tracing contributions from embeddings and event-induced variables.
The new Google Finance is coming out of beta and launching a new Android app.
General Intuition has raised $320 million to scale AI trained on millions of hours of gameplay, betting action data can help AI develop something closer to human intuition.
Sparse autoencoders reveal LLMs isolate time-aware vs. look-ahead-biased features; feature steering improves cross-domain forecasting generalization.
Process harness wraps legacy workflows with agentic reasoning layer via Task-Decision-Flow model, enabling Agentic BPM without engine replacement.
HarmVideoBench evaluates LVLMs on multi-layered harmful video detection beyond binary classification, requiring explicit explanatory rationales.
VLM-PBRS uses vision-language model guidance to learn potential-shaping heuristics for sparse-reward RL, avoiding naive reward hacking.
BERT-based multi-task framework for FDA medical device recall triage, severity assessment, and root-cause analysis on 54K historical records.
Model-assisted sampling framework reduces variance in stochastic gradient estimation for deep learning without excessive computational overhead.
Vision-language-action policy with RL achieves 1st place (online) and 2nd place (offline) in LeHome Challenge 2026 bimanual garment folding task.
TOPS framework prunes visual tokens in MLLMs via principled optimization, balancing instruction-relevance and diversity for faster inference.