- Meta Superintelligence Labs released Muse Glimmer on August 10 under Apache 2.0 — a 30B multimodal model distilled from Muse Spark 1.2 for always-on local agents.
- 4-bit quant fits the model into 18–20 GB, so a single 24 GB consumer GPU or a Mac runs it with no network call.
- Category-best scores from Meta on MCP Atlas (75.5), SWE-Bench Pro (51.2), AIME 2026 (94.7), and Charxiv Reasoning (78.8).
- Day-zero llama.cpp support (PR #26841), plus Ollama, MLX, ExecuTorch, vLLM, and SGLang integrations at launch.
- HN thread hit 555 points and 293 comments within hours; Apache 2.0 the biggest developer talking point after Llama's more restrictive terms.
- Anthropic disclosed on August 10 that an unreleased research Claude pushed the proven fraction of Riemann zeros on the critical line up by 25.6 points in a single step.
- The run used 31 million output tokens across two sessions, ~60 Claude subagents, 2,400 shell commands, and reviewed 54 arXiv papers.
- Anthropic mathematicians Levent Alpöge and Ralph Furman plus external reviewers Brian Conrey and Dan Goldston vetted the result; Claude produced a Lean formalization.
- Anthropic explicitly says the approach is not expected to yield a full proof — 67.2% is still far from the 100% the hypothesis demands.
- Announcement post on X pulled 5M+ views in hours; one number theorist called it 'the biggest in analytic number theory since bounded prime gaps in 2013.'
- Help center article updated August 11: all Claude models released on or after August 2 now embed a machine-readable watermark in generated text and attach signed C2PA metadata to .svg/.png/.jpg outputs.
- Rollout is global and automatic — driven by EU AI Act Article 50 but applies outside the EU too, with no user opt-out.
- Watermark survives copy-paste; Anthropic claims no quality loss, but critics warn biasing token choice can dent creativity.
- One prediction-market summary of the change pulled 610k+ views on X, and reactions ran heavily negative among paying users worried about schools and platforms treating a probabilistic signal as proof.
- Published August 10 alongside the Muse Glimmer weights drop — Zuckerberg's case for spreading powerful AI broadly instead of concentrating it.
- Three pillars: individual empowerment, invention as AI's primary purpose, balance of power as the safety strategy.
- Wants Washington to lift training-data restrictions on US labs, keep silicon export controls in place, and require labs to share intermediate training checkpoints for government review.
- Meta board will now sign off on safety criteria for future releases; Zuckerberg says he shouldn't be sole decider on how superintelligence is deployed.
- Google moved both models out of preview to GA on August 11 — production traffic is now supported on the Gemini API.
- 3.6 Flash uses 17% fewer output tokens than 3.5 Flash and lists at $1.50 / $7.50 per 1M in/out; knowledge cutoff moves up to March 2026.
- 3.5 Flash-Lite drops to $0.30 / $2.50 per 1M in/out, positioned as the sub-agent tier for high-volume automation.
- Same-day: Gemini 3.5 Flash Cyber joins as a security-tuned variant for SOC and code-review workloads.
01
Meta Ships Muse Glimmer, a 30B Open-Weight Agent Model That Runs on One GPU
models research.meta.ai
02
Claude Raises the Riemann Zeta Zero Lower Bound From 41.6% to 67.2%
research anthropic.com
03
Anthropic Switches On Invisible Watermarks Across Every Claude Model Worldwide
industry claude.com
04
Zuckerberg's 6,500-Word Superintelligence Manifesto Lands With Muse Glimmer
industry meta.com
05
Gemini 3.6 Flash and 3.5 Flash-Lite Hit General Availability
models ai.google.dev