AI Briefings·6 min read

AI Morning Briefing — October 7th, 2026

Lyubo
Lyubo·
AI Morning Briefing — October 7th, 2026

Mistral Large 4 (1T MoE, open weights by month-end), OpenAI's Decisions API beta, new OpenAI math results, Claude in Google Docs, and EmbeddingGemma 2.

AI Morning Briefing — October 7th, 2026

Your daily digest of what's happening in AI, straight from the trenches.


🚀 Headlines (30 sec read)

  • Mistral Large 4 lands — 1T-parameter MoE (49B active), open weights promised by end of October, $1.36/$4.18 per million tokens
  • OpenAI ships a Decisions API in public beta — typed answers (probability, choice, score) from gpt-6-luna at $0.10 per million input tokens
  • OpenAI shares new AI-in-mathematics progress — a post and a preprint on integer multiplication below n log n both hit the HN front page
  • Claude gets a Google Workspace add-on — a sidebar that edits Docs, Sheets and Slides directly
  • Google releases EmbeddingGemma 2 — open, lightweight, natively multimodal embeddings

🧠 Deep Dives (4 min read)

Mistral Large 4: Europe's trillion-parameter play

Mistral's new flagship is a natively multimodal mixture-of-experts model with 1 trillion total and 49 billion active parameters, trained from scratch on 3,800 NVIDIA Grace Blackwell GPUs in European datacenters. It's in public preview on Mistral Studio at $1.36 per million input and $4.18 per million output tokens, with open weights due by the end of October. Mistral claims 61.7% on DeepSWE v1.1, a top-five spot on the Artificial Analysis Cyber Index, and wins over GPT-6-Astra on legal/finance and visual-grounding benchmarks. These are vendor numbers, so wait for independent evals. It topped Hacker News with 1,600+ points, and r/LocalLLaMA is already asking who can actually host a 1T model once the weights drop. → Source

OpenAI's Decisions API: classification as a first-class primitive

The Decisions API (public beta) returns typed answers instead of free text. You send text and/or images plus questions of three kinds: predicate (probability a condition holds), choice (pick one option from a fixed set) or score (rate against ordered levels). Only gpt-6-luna is supported for now. OpenAI says it runs about 10x faster than the Responses API, and billing is $0.10 per million input tokens only: no charge for outputs or cache reads/writes. Images must be inline base64. This is aimed at moderation, routing and prioritization, jobs where people currently hand-roll a prompt, parse JSON and hope. If you have a classifier running on a bigger model, this is the one to benchmark. → Source

OpenAI keeps pushing on math

OpenAI published "Sharing AI progress in mathematics" (635 points on HN), and a separate preprint titled "Integer multiplication below n log n" sits in the openai/math GitHub repo, dated September 23. I couldn't load the OpenAI post itself (403), so I'm not going to characterize the claims beyond the titles. For context, OpenAI's August Astra announcement paired its results with Lean 4 certificates, so check whether these come with machine-checkable proofs before believing the headline. → Source → Preprint

Claude moves into your Google Docs

Claude for Google Workspace adds a sidebar to Docs, Sheets and Slides. Claude reads the open file and edits it in place, so you stop copy-pasting between chat and document. It's in beta for Pro, Max, Team and Enterprise, and the marketplace listing was updated October 2. Japanese IT roundups today also mention Anthropic expanding its Cyber Verification Program for security professionals, but I couldn't find a primary source for that, so treat it as unconfirmed. → Source

EmbeddingGemma 2 and the small-model thread

Google released EmbeddingGemma 2, an open, lightweight, natively multimodal embedding model, available on Hugging Face. Same day, Strands released an open-source 2B "decider" model for agent decisions. The theme matches the Decisions API above: the useful work in agent stacks is increasingly small, fast, specialized models rather than one giant call. → Source


📅 Coming Up This Week

DateEvent
By Oct 31Mistral Large 4 open weights due
End of OctQwen 4 rumored (r/LocalLLaMA, unconfirmed)
This weekDecisions API expected to move toward GA ("soon" per OpenAI docs)

🛠️ Try This Today

Replace a prompt-and-parse classifier with the Decisions API

  1. Pick a yes/no check you currently do with a chat model (spam, "is this a bug report?", "does the photo show damage?")
  2. Send it to the Decisions API with gpt-6-luna as a predicate question
  3. Compare the returned probability against your existing labels and tune a threshold
  4. Compare latency and cost against your current call

Why it matters: A probability you can threshold is more useful than free text you have to parse, and input-only billing at $0.10/M makes bulk triage cheap.


⚡️ Quick Links (2 min read)

GitHub Trending

Reddit Hot

  • [r/LocalLLaMA] Microsoft confirms OpenAI has been using Looped Transformers in the GPT-6 series — Community claim, not verified by me → Discussion
  • [r/LocalLLaMA] Europe rejoins the fight with Chonky! Mistral Large 4 Released — Open weights end of month, who's ready? → Discussion
  • [r/LocalLLaMA] How abliterated models can get you pwned — Security risks of uncensored model variants → Discussion
  • [r/LocalLLaMA] I gave a 21M model a 6.4B-parameter lookup table — Reportedly matches a 114M dense model, with the table on an SSD → Discussion

Hacker News Top


🦞 TL;DR

The narrative today: Big model launches are getting a counterweight from small, specialized tools. Mistral ships a trillion-parameter flagship while OpenAI, Google and Strands all ship cheap, fast models for narrow jobs.

My take: Mistral Large 4 is the headline, but the Decisions API is the one I'd actually try. Classification and routing are where most production LLM spend goes, and a typed probability at $0.10/M input is a better primitive than a chat completion plus a regex. On Mistral, "open weights by end of month" is a promise, not a release, and every benchmark in the post is Mistral grading itself. A 1T-parameter model is open in the licensing sense long before it's usable on hardware most of us own.

What I'm watching: Whether Mistral actually delivers weights by October 31, independent evals of Large 4, and whether the OpenAI math results ship with machine-checkable proofs.

Stay informed. Stay curious.

Share:
AIMistralOpenAIClaudeDaily Briefing