- Oct 7: $0.10/$0.50 per M tokens under 100k context, $0.50/$2.50 above — roughly 75% cheaper than Haiku 4.5.
- Can be invoked as a subagent by Opus 5.5 or Sonnet 5.5 for high-volume routing, extraction, and classification.
- Same drop lowers Sonnet 5.5 cache-read pricing to $0.10/M and adds a monthly API credit for Max/Team subscribers.
- HN front page: ~692 points, 346 comments; frames the launch as a direct answer to GPT-6 Luna's economy tier.
- New 'Intelligent UI' renders charts, forms, tappable buttons, maps, and side-by-side comparisons inline in responses.
- GPT-6 Sol lit up for paid tiers Oct 7; GPT-6 Luna reached Free and Go tiers Oct 8.
- GPT-6 Instant is 44% faster on web-search answers vs GPT-5.6, per OpenAI's own benchmark post.
- HN discussion: ~467 points, 294 comments; r/ChatGPT viral on the new inline components.
- 1.05T-parameter MoE with 49B active, 1M context, and a 1.6B multimodal vision encoder.
- API and Mistral Studio only for now; open-weights release scheduled for ~Oct 27-31 under Apache-style terms.
- Launch pricing $0.68/$2.09 per M tokens (half-off for the first month).
- HN front page: 1,895 points, 1,142 comments — the week's biggest r/LocalLLaMA thread.
- mattpocock/skills (280k stars, +1,403 today) ships practical skills — `/grill-me`, `/tdd`, `/diagnosing-bugs` — targeting agent failure modes.
- addyosmani/agent-skills (103k stars, +677 today) is the Chrome DevRel counterpart, pitched as 'production-grade engineering skills.'
- ayghri/i-have-adhd (55k, +619) forces answers-up-top formatting; thedotmack/claude-mem (97.9k, +578) adds cross-session memory.
- Trend extends to vendors: cloudflare/security-audit-skill cleared 26k stars with +576 today, the first big first-party skill drop.
- Microsoft's forecast Anthropic bill cut by more than a third from $1B+; per-employee monthly cap dropped from $100k to ~$10k.
- Meta's Claude Code user count halved from ~60k to ~30k, despite spending >$105M on Claude Code in a recent 28-day window.
- Memos frame it as cost control plus distillation risk ahead of Anthropic's rumored IPO window.
- HN thread: ~305 points; r/ClaudeAI debating whether this slows Anthropic's enterprise curve or just rationalizes it.
- 740M-parameter model maps text, code, image, video, and audio into one 768-dim space under Apache 2.0.
- ~567MB RAM full, 191MB text-only on a Pixel 11 Pro; 8,192-token context with Matryoshka dims (768/512/256/128).
- Weights live on HuggingFace and Kaggle; Google positions it as the on-device retrieval backbone for Gemini Nano apps.
- r/LocalLLaMA threads already benchmarking it against Nomic-Embed v3 and Qwen3-Embed on BEIR.
01
Anthropic Ships Claude Haiku 5.5 With 75% Price Cut and Subagent Role
models anthropic.com
02
ChatGPT Goes Visual as OpenAI Rolls Out GPT-6 Sol and Luna
models techcrunch.com
03
Mistral Debuts Large 4 'Le Chonk' — 1.05T MoE With Open Weights Due Late October
models artificialanalysis.ai
04
Named-Author 'Skill Packs' Dominate GitHub Trending for Coding Agents
open-source github.com
05
Meta and Microsoft Slash Internal Claude Spend Pre-IPO
industry thestreet.com
06
DeepMind Drops EmbeddingGemma 2 — 740M Multimodal Embeddings Under 200MB
open-source unite.ai