AI and tech news, product notes, and practical tutorials—curated long reads updated regularly.

Meta Superintelligence Labs open-sources Muse Glimmer, a 30B Apache 2.0 agentic model for on-device use. Distilled from the flagship Muse Spark and quantized to under 20GB, it runs on a single consumer GPU. Meta claims it beats Gemma4-31B and Qwen3.6-27B on agentic benchmarks, with DFlash speculative decoding delivering 3.1x faster generation on RTX 5090.
8/11/2026
23 views

ARC Prize published verified ARC-AGI results for DeepSeek V4 Flash 0731: 89.0% on ARC-AGI-1 and 61.4% on ARC-AGI-2 at max effort, costing $0.02-$0.04 per task — comparable to GPT-5.6 Luna at roughly 2x lower cost. Official agent benchmarks far exceed V4-Pro-Preview; open weights run on a single MI300X. DeepSeek also announced a significant API price increase is coming.
8/9/2026
21 views

The U.S. Department of Energy launched the Genesis Open Models Initiative with Arcee AI, unveiling Genesis-Science-1, a trillion-parameter-class open-weight model for scientific research across materials, energy, fusion and biology. Weights and a technical report are slated for release later this year.
8/8/2026
15 views

On August 6, Alibaba's Qwen3.8 Max was ranked first overall on Artificial Analysis' Agentic Index, with 2.4T parameters and open weights coming for the first time at Max scale. The same day, OpenAI updated GPT-5.6 Sol and made Luna the default model for free users with unlimited text chats.
8/7/2026
18 views

Developer ryanzhou open-sourced configs and patches to run DeepSeek-V4-Flash-0731 (304B params) on a single AMD MI300X, with no quantization or weight offload, hitting 168.6 tok/s single-stream decode. The HN post scored 365 points the same day.
8/5/2026
14 views

On August 4, Mistral released Shieldstral, a 3B open-weights multimodal safety classifier under Apache 2.0. It frames moderation as policy-adaptive question answering: policies are supplied as plain-language queries at inference time, returning calibrated safety scores without retraining, unifying text and image safety evaluation.
8/5/2026
14 views

On August 1, OpenAI announced that an internal version of Astra, its next major model family, solved 10 open problems in mathematics, quantum complexity, and theoretical computer science that had seen no progress on their main results for at least a decade. The proofs span sphere packing, non-sofic groups, Connes' rigidity conjecture, and quantum parallel repetition, all formalized in Lean. At Sol API rates, the total token cost to find these solutions was roughly $2,000.
8/3/2026
15 views

ByteDance's Seed team released Seedance 2.5, its next-generation video creation model, on July 31. Single-pass generation doubles from 15 to 30 seconds with multi-round extension, multimodal referencing accepts up to 30 images, 10 video clips and 10 audio clips per pass, and timestamp-level editing is now supported. It is rolling out on Jimeng AI and Doubao Pro, with API access coming soon via BytePlus ModelArk.
8/2/2026
14 views

On July 31, DeepSeek released the official V4-Flash-0731 with major agent capability upgrades (Terminal Bench 56.9 to 82.7) at unchanged pricing, while MiniMax launched H3, an omni-modal generation model with 2K video and native stereo audio, with weights to be open-sourced soon.
8/1/2026
14 views

OpenAI cuts GPT-5.6 Luna pricing by 80% through kernel-level optimizations. Google DeepMind launches Gemini Robotics 2 — robots can now understand video feeds and collaborate in teams. Two major announcements hit the same day.
7/31/2026
12 views

Microsoft launches MAI-Cyber-1-Flash, its first cybersecurity AI model integrated into MDASH, scoring 96% on CyberGym at 50% lower cost.
7/29/2026
6 views

Moonshot AI released Kimi K3 as open weights on HuggingFace on July 27 — the world's first open 3T-class model. With 2.8T parameters, MoE architecture, 1M context window, and native vision.
7/28/2026
15 views