sovereign AISamsung Led Mistral's Round. Score the Clauses, Not the Flag.
Mistral's €3 billion Series D moved its lead investor to Samsung Electronics without changing one customer contract. The same announcement publishes a four-part sovereignty definition that maps to clauses you can actually test — and the European Commission's own reference rubric weights ownership at 15%.
September 8, 2026 · 14 min readcode hallucination12 Models Wrote a Fake Crate. None Refused a Real One.
A new benchmark handed twelve open-weight models 270 impossible coding tasks. They fabricated confident, compiling code on 60% and refused 27% — while wrongly refusing 0.0% of 91 matched solvable controls. That zero is why an unsatisfiable arm belongs in your eval suite.
September 5, 2026 · 13 min readMixture of ExpertsTencent Says 49B Active. Your Node Loads All 770B.
Tencent's Hy4 preview markets 49B active parameters, but all 770B must stay resident — roughly 780GB in FP8, which an 8xH100 node cannot load at all. Size a Mixture-of-Experts model on total parameters and price throughput on active ones.
August 30, 2026 · 12 min readlegal AICoCounsel's New Model Runs on Qwen. Go Read the Card.
Thomson Reuters' new in-house legal model is a Qwen derivative two hops down, and Harvey's Tenet is post-trained from Kimi K3. Neither lineage appears in a DPA or a subprocessor list — because an open-weight base model provider processes nothing.
August 27, 2026 · 14 min readmodel weightsHugging Face Hired Bankers. Go Mirror Your Weights.
Hugging Face has retained bankers to test buyer interest at $13 billion. Most enterprises pull model weights from it under a click-through agreement nobody signed — here is the mirroring playbook to run this week.
August 25, 2026 · 12 min readKV cache transferNvidia Skipped the Prefill. The Math Went With It.
Nvidia's cross-model KV cache mapper moves a 32,768-token cache in 277.6ms instead of 6,975.3ms of re-prefill. Broken out by benchmark, GSM8K survived on one of six model pairs — Llama 3.1 8B to 70B fell from 81.12 to 14.78.
August 21, 2026 · 12 min readMuse GlimmerMeta's Agent Model Fits on a Laptop. Nothing Logs It.
Meta's Muse Glimmer scores 75.5 on MCP Atlas inside a 24GB memory envelope under Apache 2.0, so agentic tool-calling now runs on hardware engineers already own. Every AI control you have — prompt logging, token accounting, DLP, model pinning, the kill switch — is implemented at a gateway that no longer sees the traffic.
August 10, 2026 · 13 min readGenesis MissionThe DOE Wants Your Lab Data. You Write the Contract.
DOE's Genesis Open Models portal wants proprietary scientific data and the first window closes August 14. Applying transfers nothing — but there is no published contributor agreement, so the handling terms are yours to draft.
August 9, 2026 · 13 min readAMDAMD Bought Taalas. Now Name the Model You'd Freeze.
AMD signed a definitive agreement on August 6 to buy Taalas, whose chips etch model weights into mask ROM so one chip serves exactly one model. The business model assumes a one-year hardware life, which makes the cheapest inference tier available only to workloads whose model you can name and freeze.
August 8, 2026 · 13 min readDeepSeek V4DeepSeek Will Raise Prices. Your Ceiling Is Already 4x.
DeepSeek warned of a significant API price rise with no number and no date. Because the weights are MIT-licensed, the ceiling is already public: independent hosts serving the identical model charge 3-4x on posted rates and 40x on cache hits.
August 7, 2026 · 12 min readDeepSeek V4 FlashDeepSeek Swapped the Model. Your Eval Didn't Notice.
DeepSeek shipped new weights behind the same API name on July 31 — identical architecture, identical price, DeepSWE up from 7.3 to 54.4. Any agent evaluation pinned to a model name instead of a build is now returning a stale verdict.
August 3, 2026 · 11 min read