PipeNetwork/minimax-h3-mlx
MiniMax releases H3, a multimodal generative model generating 15-second video from text/image/audio input; MLX port enables inference on Apple Silicon.
Every story tagged with this topic, ordered by date.
MiniMax releases H3, a multimodal generative model generating 15-second video from text/image/audio input; MLX port enables inference on Apple Silicon.
Google announces July 2026 AI updates; article lacks specific details on models, features, or benchmarks.
Alibaba releases Qwen 3.8 Max (2.4T params) and 27B open-weight models optimized for coding and collaboration tasks.
Analysis of 18 open-source LLMs showing cultural bias in mythology knowledge; models encode cross-cultural distinctions in residual streams but fail to decode non-Western traditions.
Antares: compact LLMs (350M–3B) for agentic vulnerability localization via SFT and RL on cybersecurity reasoning over code.
Microsoft-led open letter signed by 235 AI companies including NVIDIA and OpenAI argues against US government restrictions on open-weight models on safety grounds.
Gaokerena: compact Persian-language medical LLM family trained on 90M-token corpus for low-resource healthcare deployment.
Tevatron 3.0 integrates Megatron-Core for efficient MoE reranker training, enabling billion-scale cross-encoder + distillation workflows on academic budgets.
DeepSeek releases V4-Flash-0731, a 304B parameter model with enhanced agentic capabilities, outperforming larger competitors at $0.14/$0.27 per million tokens.
Podcast discussion on open-weight model competitiveness, cybersecurity risks, and AI leadership policy letters signed by industry leaders.
Simon Willison releases smevals, an open eval framework for benchmarking models, prompts, and inference harnesses across configurations.
DenseOn and LateOn: open-source 149M-parameter retrieval models trained on 1.88M supervised pairs; competitive on multilingual and code search.
Method to detect CSAM-generating LoRAs from weight fingerprints (singular vectors) without generating outputs, enabling safer moderation.
Kimi K3 open-weights model released amid broader industry discussion on open model availability and strategy.
Moonshot releases Kimi K3 weights (2.8T params, 1.56TB) with modified MIT license requiring attribution for products >100M MAU.
Anthropic publishes official stance on open-weights model releases, addressing trade-offs between transparency, safety, and competitive positioning.
Causal-TS: open-source Python library for causal discovery in high-dimensional nonstationary time series with GPU-accelerated conditional independence testing.
Kimi K3: 2.8T MoE model with 104B active params, 1M context window, Delta Attention, 2.5x scaling efficiency over K2.
ELMOD: 2.7B German-language model optimized for mobile deployment using public data and morphology-aware preprocessing.
Byte-Prefix Marginalization method for cross-tokenizer on-policy distillation of open-weight LLMs with incompatible vocabularies.
OpenForgeRL enables end-to-end training of harness-native agents with open infrastructure, addressing limitation of complex inference harnesses like Claude Code.
DONDO releases 26 open w2v-BERT speech recognition models for African languages spanning six countries, trained on religious text corpora.
Open-source evaluation framework for open-weight LLM agents on longitudinal data tasks, addressing privacy constraints in research deployments.
Laguna S 2.1, a 118B MoE model from Poolside AI, achieves Deepseek v4 Pro performance at lower cost than v4 Flash.
Poolside AI co-CEO Eiso Kant describes building a model factory enabling efficient training of 118B MoE models competitive with 1T open-weight alternatives.
Security researcher Thomas Ptacek claims open-weights 2025 models could execute sandbox escapes and network reconnaissance without frontier capabilities.
18 medical image encoders on 650k radiographs show self-supervision drives representational convergence more than clinical labels.
Orchestrated open-weight small LLMs achieve malware analysis performance competitive with frontier closed-weight models at lower computational cost.
CircuitKIT open-source library unifies circuit discovery, evaluation, and intervention workflows for mechanistic interpretability with automated contrastive prompts.
Ben Thompson proposes US law to legalize model distillation and data collection as fair use, addressing licensing hypocrisy and competitiveness vs. Chinese models.
Stratechery argues U.S. frontier labs face minimal threat from Chinese models; policy should prioritize open-weight domestic alternatives instead.
Sam Altman email (Oct 2022) reveals OpenAI planned GPT-3-class open-weight model for consumer hardware to preempt Stability AI.
Study evaluates open-weight LLMs for extracting structured CVE threat data from autonomous vehicle vulnerability text.
Kimi K3 2.8T-A50B released as largest open-weight model with Opus 4.8-class performance at Sonnet 5 pricing.
Moonshot AI releases Kimi K3 (2.8T params), claims top performance vs. Claude Opus 4.8 Max and GPT-5.5, promises open-weight release by July 2026.
Thinking Machines Lab releases Inkling, a 975B-parameter open-weights MoE multimodal model trained on 45T tokens.
Thinky releases Inkling, a 975B multimodal open-weights model under Apache 2.0, with a smaller 276B variant.
Pythia multi-agent system for autonomous clinical symptom extraction using open-weights LLMs without fine-tuning.
Open-weight reasoning models fine-tuned via RLVR for thermal energy storage control, achieving building-scale load shifting with 30 prompts.
Cohere releases Tiny Aya Expedition, a multilingual model supporting 70+ languages for on-device and educational AI applications.
Soofi S 30B-A3B: open-source MoE-Mamba hybrid for German/English with 3B active parameters, matches 14-27B dense models on benchmarks.
LLM4SDM evaluates open-source smaller models on clinical decision-making assessment, comparing privacy-preserving local deployment vs. commercial models.
Cohere releases open-source Arabic speech recognition model for enterprise transcription across Arabic dialect variants.
Tencent releases Hy3, a 295B-param MoE model with 21B active params under Apache 2.0, claiming performance parity with 2-5x larger open-source competitors.
Linear probes decode remaining output length from LLM hidden states across 7-8B open-weight models, revealing internal response-length estimation.
Current AI launches Gap Map v0.1, an index of 421 open-source AI products across models, tools, datasets, and hardware, backed by $400M committed capital.
Simon Willison's June 2026 newsletter covers Claude Fable 5, GPT-5.6, GLM-5.2 open weights, and US export restrictions.
HaloGuard 1.0 releases open-weights constitutional safety classifier achieving state-of-the-art multilingual prompt-safety performance at 1/10 model size.
MultiSynt/MT releases 4.8 trillion tokens of open synthetic parallel pre-training data across 36 European languages via Tower+ and OPUS-MT translation.
AI Engineer World's Fair coverage: agent loops, software factories, forward-deployed engineering, and open model adoption emerging as key themes.