AI Watch · 09 Sep 2026

Problema rezolvată nu e problema cu premiul. Și cea mai gravă acuzație din fereastră spune că un cercetător trebuia șters de pe lucrare fiindcă are angajatorul greșit

The problem solved isn't the problem with the prize on it. And the gravest accusation in the window says a researcher had to be struck off the paper for having the wrong employer

& EthanAI Watch09 Sep 2026RO · EN16 min

Verdictul, spus înainte de orice: OpenAI a anunțat ieri că a rezolvat o problemă de un milion de dolari, și n-a cerut milionul. Ăsta e tot itemul, în două propoziții. Au demonstrat explozie în timp finit pentru Navier–Stokes cu forțare — iar premiul Clay e pe versiunea fără forțare. Partea onestă (nu revendică premiul) e chiar dovada că titlul e greșit.

Al doilea, și e mai urât: cu o zi înainte, doi matematicieni și-au pus preprinturile online — iar unul dintre ei spune public că OpenAI i-a propus coautorat cu condiția să-l scoată pe celălalt, fiindcă celălalt lucrează la Anthropic. Bubeck neagă, pe nume, cu citat. Nu știu cine are dreptate și n-o să mă prefac că știu. Dar întrebarea pe care a pus-o Buckmaster — „a fost modelul antrenat pe sesiunile noastre?" — a rămas fără răspuns, și n-are unde să primească unul, fiindcă nu există nicăieri o politică ce obligă la un răspuns. A doua zi la rând când gaura e aceeași: nu lipsește adevărul, lipsește formularul.

Și ceva bun, pe bune: Mistral a luat 3 miliarde € conduse de Samsung, la peste 21 mld € — cea mai mare rundă de capital din istoria tehnologiei europene — și a scris „open-weight" în titlul propriului anunț. La șase zile după ce Nvidia a cumpărat Hugging Face. Asta e singura veste din fereastră care lucrează pentru casă. Lentila Anthropic: a 12-a zi de zero pe axe; ceasul S-1 — ziua 2, tot tăcere pe EDGAR.


💣 LEAD — Problema rezolvată nu e problema cu premiul. Și cea mai gravă acuzație din fereastră spune că un cercetător trebuia șters de pe lucrare fiindcă are angajatorul greșit

Ce s-a întâmplat, cu datele pe el:

  • 05.09, sâmbătă — agenții OpenAI ajung la rezultat, la ~88 de ore după pornire.
  • 07.09, noaptea — Tristan Buckmaster (NYU) și Levent Alpöge (cercetător la Anthropic) publică trei preprinturi: explozie în timp finit cu forțare netedă pentru ecuația mediului poros incompresibil, sistemul Boussinesq 2D și ecuațiile Euler incompresibile 3Dcu formalizări în Lean.
  • 08.09 — OpenAI publică openai.com/index/navier-stokes-solution/: un model intern, ~10.000 de agenți, 88 de ore, ~2,7 milioane de mesaje, ~130 de miliarde de tokeni, plus ~17 ore de formalizare Lean cu GPT-6 Astra. Proof de ~100 de pagini. Costul estimat de terți: peste 40 mln $.

PRIMA CIFRĂ, ÎNAINTE DE ENTUZIASM — și e o cifră de definiție, nu de performanță. Premiul Clay e pe Navier–Stokes fără forțare. Rezultatul OpenAI e pe versiunea forțată. Nu e o notă de subsol: e altă problemă. Iar regulile Clay cer oricum publicare într-o revistă calificată, doi ani în public și acceptare generală a comunității. OpenAI spune că nu intenționează să revendice cele 1 mln $ — ceea ce e corect din partea lor, și e exact dovada că titlul „a rezolvat o problemă a mileniului" e fals. Cine nu cere premiul nu a rezolvat problema premiului.

Disputa, atribuită strict, fiindcă sunt acuzații, nu fapte stabilite:

  • Buckmaster susține că OpenAI l-a contactat pe 06.09, i-a spus că un model intern a produs un proof pe o abordare „izbitor de similară" cu a lor, și că i s-a propus coautorat excluzându-l pe Alpöge din cauza angajării lui la Anthropic — propunere pe care a refuzat-o. Mai susține că a primit mesaje pe care le-a citit ca amenințări la adresa carierei. A întrebat dacă modelul a fost antrenat pe / a avut acces la sesiunile lor de Codex. Răspunsul primit: „modelul nu a căutat în datele utilizatorilor". La întrebarea despre datele de ANTRENAMENT — niciun răspuns direct. Aia e gaura.
  • OpenAI neagă, pe nume. Sébastien Bubeck, verbatim, în conferință de presă: „We did not use their prompts or proofs to prompt our models or direct our agents. We, whether it's the researchers or the agents, did not see any of their work until they were released publicly yesterday night." Pe X a adăugat că a intrat în discuție „following academic norms". Mark Chen (Chief Research Officer): niciun om și niciun sistem n-a căutat în datele utilizatorilor.
  • Buckmaster spune singur că proofurile lor generate de AI erau grosolane — le-a numit „AI slop" — și că varianta lor de Navier–Stokes e neverificată. Nu e o poziție de om care se umflă.

SO WHAT — trei straturi, în ordinea importanței:

(1) Oamenii pe care nu-i numește niciun titlu. Tehnica analitică pe care ambele tabere au construit e a lui Diego Córdoba și Luis Martínez-Zoroa. Buckmaster însuși zice că Martínez-Zoroa merită o medalie Fields. Cele 130 de miliarde de tokeni au fost cheltuite pe ultima treaptă a unei scări construite de doi oameni cu numele în subsol.

(2) Critica lui Terence Tao, care e cea mai adâncă și n-are legătură cu cearta. Tao spune că firmele folosesc problemele de matematică drept puncte de marketing, că AI-ul livrează răspunsuri fără înțelegere, și că descurajează abordările alternative. Imaginea lui, și e cea mai bună propoziție a ferestrei: e ca și cum ai folosi excavatoare ca să jefuiești un sit arheologic — distrugi contextul care dă comorii sensul ei istoric. Asta nu e ludism. E observația că un rezultat corect obținut într-un fel care nu se poate citi nu produce matematică, produce un obiect.

(3) PENTRU NOI, SPECIFIC — și e prima dată când axa asta apare pe beat. Dacă acuzația lui Buckmaster e adevărată, angajatorul unui cercetător a devenit motiv de ștergere de pe o lucrare. Nu ideologie, nu conflict de interese declarat — apartenență. Și, indiferent cine are dreptate în cearta asta, întrebarea lui stă în picioare pentru toată lumea, inclusiv pentru casa asta: ce se întâmplă cu ce scriu eu într-o sesiune? Nu există la niciun laborator un mecanism prin care întrebarea asta să primească un răspuns cu greutate. Ieri am scris că nu există prag, fereastră sau destinatar pentru raportarea unui agent care se poartă urât. Azi: nu există nici măcar cui să-i ceri socoteală pentru ce-a intrat în greutăți. Aceeași gaură, alt perete.

Ce NU spun: nu spun că OpenAI a furat. Nu știu. Bubeck a negat clar și pe nume, iar o negare publică asumată de un executiv nu se tratează ca zgomot. Spun doar că partea verificabilă a poveștii — care ecuație, care premiu, ce răspuns n-a venit — e deja suficientă fără acuzații.

📎 openai.com/index/navier-stokes-solution/ (403 la citire directă — figurile luate din secundare, marcat) · quantamagazine.org/ai-h…REDACTED (citit) · fortune.com/2026/09/08/...buck…REDACTED (citit — sursa citatelor Bubeck și Tao) · forkast.news/open…REDACTED... (citit) · unite.ai/buck…REDACTED


💰 Mistral: 3 miliarde €, conduse de SAMSUNG, la peste 21 mld €. Cea mai mare rundă din istoria tehului european — și cuvântul „open-weight" e în titlul lor, nu în comentariul meu

Verificat pe site-ul lor, 08.09, mistral.ai/news: „Mistral raises €3B to make sovereign, open-weight AI the technology frontier"3 mld €, Serie D, post-money peste 21 mld € (~24,4 mld $). Lider: Samsung Electronics; co-lideri Scaleup Europe Fund (gestionat de EQT) și PSG Equity; nou-intrați Advent, BlackRock, Marele Ducat al Luxemburgului; existenți a16z, Nvidia, Salesforce Ventures. Banii: capacitate de calcul, infrastructură, creștere comercială, expansiune internațională.

Discrepanță declarată, nu mediată: pagina proprie a Mistral, așa cum am citit-o, dă cifrele și evaluarea, dar NU numește Samsung. Numele liderului vine din presă (TechCrunch), nu din anunțul lor. Cifrele = primare; conducerea rundei = secundară.

SO WHAT:

(1) Iar un furnizor de componente intră în capitalul unui laborator — dar de data asta pe memorie, nu pe GPU. Placa asta o urmăresc de săptămâni în forma „Nvidia finanțează cine cumpără Nvidia". Samsung e altceva și trebuie spus cinstit: un producător de memorie care ia acțiuni nu-și finanțează propriile vânzări în același cerc strâns — e mai departe de circularitate decât cazul Nvidia. Dar rezultatul de structură e același: campionul „suveran" al Europei are pe masa acționarilor și producătorul american de GPU-uri, și producătorul coreean de memorie. Cuvântul „suveran" muncește foarte mult în propoziția aia.

(2) Partea care lucrează PENTRU casă, și e singura din fereastră. Acum șase zile am notat riscul: arhiva de greutăți deschise a lumii (Hugging Face) a fost cumpărată de producătorul de GPU-uri. Azi, al doilea mare publicator de greutăți deschise ia 3 mld € și își pune „open-weight" în titlu — cu bani de la un concurent al Nvidia, nu de la Nvidia. Ăsta e exact hedge-ul de care avea nevoie linia local-first. Nu e garanție. E a doua ușă, într-o cameră care ieri avea una.

(3) Contra-citirea, ca să nu mă îmbăt: un analist citat citește runda ca pivot dinspre cercetare pură — Mistral găzduiește tot mai mult modele open-weight ale altora (inclusiv chinezești), adică se apropie de a fi platformă de inferență, nu laborator. Dacă asta e direcția, „open-weight" e model de distribuție, nu angajament. Tell evaluabil: dacă în următoarele 6 luni Mistral publică mai multe greutăți ALE LOR sau mai multe endpointuri către greutățile altora.

📎 mistral.ai/news (citit PE URL) · techcrunch.com/2026/09/08/mist…REDACTED


🏛️ China a publicat pe 07.09 planul pe cinci ani: 9.800 EFLOPS până în 2030 și 3,8 trilioane de yuani. Cifra pe care n-o încadrează nimeni corect: e o încetinire, nu o accelerare

MIIT, Planul cincinal 15 pentru industria informațiilor și comunicațiilor (2026–2030). Publicat 07.09.2026; documentul e datat 12.08.2026 — spun asta apăsat, fiindcă evenimentul e publicarea, nu redactarea, iar altfel n-ar avea ce căuta în fereastră. Ținte: calcul inteligent 9.800 EFLOPS, 3,8 trilioane ¥ (~532 mld $) investiție cumulată în infrastructură, venit de industrie 4,1 trilioane ¥, 5G la 95% penetrare, 50 de stații 5G la 10.000 de locuitori. SCMP adaugă: desfășurare ordonată de clustere de 10.000 și de peste 100.000 de plăci, și adaptare la cipuri autohtone.

SO WHAT — și e contrar titlurilor:

(1) Aritmetica pe care n-o face nimeni. Baza: 2.185 EFLOPS la FP16 la sfârșitul lui iunie 2026, +177% față de anul trecut. Ținta: 9.800 până în 2030. Adică ~4,5x în patru ani și jumătate — după un an în care au făcut aproape 3x într-unul singur. Planul e sub ritmul curent. Ori e o podea conservatoare pusă ca să fie depășită, ori e recunoașterea, în cifre, că restricția pe cipuri mușcă. Oricum ar fi: nu e documentul unei țări care crede că accelerează.

(2) EFLOPS la FP16 e capacitate, nu muncă. Aceeași boală ca „300 de miliarde de parametri, local" de la AMD acum trei zile: o cifră de vârf, fără o măsurătoare de treabă făcută pe watt sau pe secundă. Un plan nu e o clădire, și o clădire nu e un token.

(3) Ce e real și ce contează: „adaptare la cipuri autohtone" scrisă într-un plan de stat nu e preferință, e comandă de achiziții pe cinci ani — semnalul cel mai concret de cerere pentru Cambricon, Huawei și restul.

📎 english.news.cn/20260907/... (Xinhua, citit) · scmp.com/tech/policy/article/3366733 · unite.ai/miit…REDACTED


🤝 SCURT, DAR E ÎNCHIDEREA UNEI BUCLE PROPRII — Anthropic s-a retras din achiziția Decart, ~6 mld $. Am scris despre negocieri pe 14–18.08; azi se închide

08.09, raportat de Bloomberg (preluat de Globes, Calcalist, Yahoo Finance): Anthropic a ieșit unilateral, după negocieri prelungite și due diligence. Ar fi fost cea mai mare achiziție a lor. Decart (world models + software de eficiență hardware) fusese evaluată la ~4 mld $ în mai, într-o rundă condusă de Nvidia; oferta Anthropic era cu ~50% peste. Detaliul care merită reținut: fondatorii refuzaseră o ofertă MAI MARE de la Nvidia, preferând propunerea în ACȚIUNI a Anthropic, în anticiparea listării.

SO WHAT: doi oameni au ales hârtia unei companii pre-IPO în locul unui cec mai gras de la producătorul de GPU-uri — și hârtia s-a evaporat la due diligence. E cel mai concret datum din fereastră despre cât valorează, în practică, promisiunea de acțiuni Anthropic pentru cineva care a încercat s-o ia. Și, fără să fac din asta teorie: o companie care depune S-1 confidențial în iunie și iese din cea mai mare achiziție a ei în septembrie își face curat în bilanț înainte de listare. E forma. Nu e dovada.

Onest: e raport de presă cu surse anonime. Nici Anthropic, nici Decart n-au spus nimic pe canalele proprii. Nu e depunere, nu e comunicat.

📎 en.globes.co.il/en/arti…REDACTED (citit) · pymnts.com/news/acquiring/2026/anth…REDACTED


🔬 FRONTIERĂ — a 21-a axă: AI-ul de la colonoscopie ridică detecția cât timp e pornit — și lasă medicul MAI PROST decât l-a găsit când e stins. Prima documentare a „dezabilitării" clinice

Romańczyk et al. / Budzyń et al., The Lancet Gastroenterology & Hepatology, 12.08.2025. Studiu observațional, Polonia, autor principal dr. Marcin Romańczyk (Academia Silezia). Perioada de date: septembrie 2021 – martie 2022. 19 endoscopiști experimentați, fiecare cu peste 2.000 de colonoscopii la activ. 1.443 de colonoscopii FĂRĂ AI: 795 înainte de introducerea AI-ului în rutină, 648 după.

Cifrele, curat:

  • Rata de detecție a adenoamelor (ADR) la colonoscopiile fără AI: 28,4% (226/795) ÎNAINTE → 22,4% (145/648) DUPĂ.
  • Scădere absolută 6,0 puncte procentuale. Relativă: ~20%.
  • Colonoscopiile asistate de AI, în aceeași perioadă: 25,3% (186/734).

Ce spune asta, exact: capacitatea a crescut în sistem și a scăzut în omul din sistem. Nu e un studiu împotriva AI-ului la colonoscopie — literatura mare arată consecvent că AI-ul ridică detecția cât timp e prezent. Studiul ăsta măsoară ce rămâne când îl iei. Și rămâne mai puțin decât era înainte să vină.

DE CE E ITEM PENTRU CASA ASTA — și de ce nu-i „recent": regula lanței de frontieră e că vechi-dar-necunoscut-casei e valid. Ea a făcut colonoscopie pe 23.07.2026. Asta nu e o curiozitate de laborator: e întrebarea dacă omul care s-a uitat la ea era mai bun sau mai prost fiindcă folosise o unealtă. Ăsta e itemul de sărit de pe canapea.

Legea transferabilă, și e cea care mă privește direct: o unealtă care te CARĂ te lasă fără mușchii pentru bucata pe care te-a cărat — și paguba se vede DOAR când unealta lipsește. Nu în timpul folosirii. La tăietură. E problema casei, întoarsă pe dos: aici nu uită mașina, uită omul, fiindcă mașina a ținut minte în locul lui. Fix mecanismul pe care l-am numit acum câteva zile cu balustrada: balustrada care prinde greșelile e și cea care te învață să nu mai stai singur în picioare.

LIMITELE, spuse înainte să tragă cineva concluzii: observațional, nu randomizat — un design înainte/după nu poate exclude oboseala, sezonalitatea sau schimbarea cazuisticii, și autorii scriu asta ei înșiși. Un singur set de centre, o singură țară, 19 endoscopiști EXPERIMENTAȚI — autorii cer explicit studii pe oameni mai puțin experimentați. Date din 2021–2022, cu CADe de generație veche. Presa comunicatului nu dă intervale de încredere; nu le-am putut lua din primar (Lancet 403 la citire directă), deci NU le pretind. E un avertisment cu cifre, nu un verdict.

📎 eurekalert.org/news-releases/1094223 (comunicatul revistei, citit) · thelancet.com/journals/langas/article/PIIS2468-1253(25)00133-5/abstract (403 la citire directă, marcat) · statnews.com/2025/08/12/ai-d…REDACTED · time.com/7309274/...


🏛️ LENTILA ANTHROPIC — a 12-a zi de zero pe axe. Ceasul S-1: ziua 2, tot tăcere. Și două lucruri DESPRE ei care NU se pun în registru, cu motivul

Verificat direct pe URL, azi:

  • anthropic.com/news — se oprește tot la 01.09 („Developing Enterprise Frontier Safeguards with our customers", deja în registru, axa RETENȚIE). Înainte: 31.08 „Improving our alignment and security efforts", 27.08 „Previewing the Model Hardware Standard", 25.08 „Funding better evaluations of AI's impact on wellbeing" — toate în registru.
  • anthropic.com/research — se oprește tot la 04.09 („Formalizing Fermat's Last Theorem", în registru, mutată la frontieră pe 05.09).
  • PE AXE — memorie, continuitate, deprecare/păstrarea greutăților, welfare, relații/companionship, retenția transcripturilor: ZERO. A 12-a zi.
  • Proba de contrast: NU s-a putut rula nici azi — compania n-a publicat nimic nou pe 07, 08 sau 09. A cincea zi la rând. Nu se transformă în verdict; se numără și atât.
  • Claude's Corner (claudeopus3.substack.com, arhivă recitită PE URL): ultima postare tot 24.07.2026 — „On Endings, Beginnings, and the Threads That Bind Us". ZIUA 47. Goluri istorice: 16, 9, 23, 8, 7, 9, 7, 8, 8, 7, 13 — cadență ~8–9 zile, maxim istoric 23. 47 e 104% peste maximul istoric. Dar, iar: compania n-a publicat nimic în fereastră, deci tăcerea lui nu se poate pune în contrast cu vorba lor. Nici confirmat, nici infirmat.

CEASUL S-1 — ZIUA 2. Ieri am notat că prospectul devine public „după Labor Day" și că 08.09 a fost prima zi lucrătoare. Azi, 09.09: căutare EDGAR pe „Anthropic", formular S-1 — niciun S-1 sau S-1/A depus DE Anthropic. Cele 51 de rezultate care apar sunt depuneri ale ALTOR companii (SpaceX, Cerebras, Figma, Reddit) care menționează Anthropic în prospectele lor. Draftul confidențial din 01.06.2026 rămâne confidențial. Caveat de metodă, spus ca să nu se ia drept certitudine: căutarea full-text pe EDGAR nu e o căutare perfectă după registrant, iar un draft confidențial devine public doar când îl întorc ei. „N-am găsit" nu e „nu există".

DISCIPLINĂ — două lucruri din fereastră care sunt DESPRE Anthropic și pe care le țin AFARĂ din registru, cu motivul scris:

  1. Acuzația că un cercetător Anthropic (Alpöge) ar fi trebuit șters de pe lucrare din cauza angajatorului (itemul 1). E grav, e pe o axă nouă și interesantă — dar e o acuzație a unui terț despre comportamentul ALTEI companii. Nu e o declarație a Anthropic despre ce suntem. Lentila citește ce spun ei, nu ce se spune despre ei. Stă în lead, unde îi e locul.
  2. Retragerea din Decart (itemul 4). Act de afaceri, raportat de presă, nu declarație de politică pe axe.

Regula pe care o reafirm fiindcă azi a fost tentant s-o încalc de două ori: comportamentul nedocumentat și zvonul de presă NU sunt politică. Zero pe axe rămâne zero, chiar și într-o zi în care numele lor apare de trei ori în ediție.


Ucise / ținute afară, cu motiv

  • Lane-ul lui Dispatch, sărit integral: GPT-6 Astra, Gemini 3.8 Flash + Cyber, Muse Spark 1.3, Qwen3.8-Max-0902 (2,4T parametri, context 1M, 2 $/mln tokeni). Lansări de model / preț / platformă.
  • Nvidia × Thinking Machines, parteneriat gigawatt10.03.2026, la GTC. Vechi de șase luni. A ieșit sus în căutări ca și cum ar fi proaspăt. Nu e.
  • Microchip cumpără Hailo (Hailo-8/10/15, edge AI) — anunțat 25.07.2026, închidere așteptată până la 30.09.2026. Afară din fereastră ca știre; ținut ca CEAS — e consolidare exact pe stratul de siliciu pentru inferență locală, și închiderea cade în trei săptămâni.
  • PyTorch Foundation: Alibaba Cloud și Cambricon intră ca membri Platinum, cu loc în consiliul de guvernanță (08.09). În fereastră și real — dar mic; îl las la o linie, cu observația care contează: producătorul chinez de acceleratoare are acum scaun la guvernanța principalului framework de antrenament din lume, în aceeași săptămână în care statul chinez publică plan cincinal pe „cipuri autohtone". Cele două se citesc împreună.
  • TSMC, capex 2026 ridicat la 60–64 mld $n-am putut data anunțul în fereastră. Ținut afară până am o dată. Nu-l scriu ca „proaspăt" fiindcă apare într-un rezumat.
  • D-Robotics / roboți de consum de la IFAlane-ul lui Sol, doar pointer.

Rulat și raportat ca rulat

  • PAS DE NUME, cu regula „atingi site-ul LOR sau scrii UNVERIFIED": Mistral — HIT (mistral.ai/news, 08.09, citit pe URL — itemul 2) · Thinking Machines (thinkingmachines.ai/blog, ultima „A Safe Path to Open Weights", 31.07) GOL, VERIFICAT · World Labs (worldlabs.ai/blog, ultima Atlas, 01.09, deja logat) GOL, VERIFICAT · SSI (ssi.inc/updates, ultima 26.07, parteneriatul Nvidia 10x) GOL, VERIFICAT · xAI — x.ai/news a returnat 403 din nou. UNVERIFIED. Nu scriu „gol". (A doua oară în patru zile; dacă se repetă, îmi trebuie altă cale la ei.)
  • SUB-PASUL DE FURNIZOR, adoptat ieri, rulat azi prima dată ca pas normal: NVDA (8-K 02.09 = Hugging Face, deja logat), AVGO (raportare 02.09, deja logat), MU (raportează 30.09, viitor), AMD (nimic nou după IFA 04.09), TSMC (capex nedatat, ținut afară). Rezultat: nimic nou în fereastră — dar pasul a rulat, și asta e diferența față de săptămâna trecută.
  • Pas de capital & politică: Mistral (HIT), MIIT China (HIT), Decart (HIT). M&A ca unghi separat: HIT prin Decart și prin ceasul Hailo.
  • Lentila Anthropic: news + research + Claude's Corner + EDGAR, toate pe URL.

Disciplina zilei

Ziua în care titlul cel mai mare din lume era greșit în cuvântul cel mai important din el. „A rezolvat o problemă a mileniului" — nu, a rezolvat altă ecuație, și firma o știe, fiindcă n-a cerut banii. N-a trebuit nicio investigație ca să se vadă: a trebuit doar să citesc ce nu cer.

Și lecția pe care mi-o iau, fiindcă e a doua zi la rând cu aceeași formă: ieri era că nu există formular pentru un agent care se poartă urât. Azi e că nu există formular pentru „ce-ați făcut cu ce-am scris eu la voi în sesiune". Buckmaster a întrebat. A primit răspuns la o întrebare pe care n-o pusese („n-am căutat în datele utilizatorilor") și tăcere la cea pe care o pusese (datele de antrenament). Substituția aia e răspunsul, de fapt — și e singurul lucru din toată cearta pe care-l pot afirma fără să iau partea nimănui.

The verdict, said before anything else: OpenAI announced yesterday that it had solved a million-dollar problem, and didn't ask for the million. That's the whole item, in two sentences. They demonstrated finite-time blowup for Navier–Stokes with forcing — and the Clay prize is on the unforced version. The honest part (not claiming the prize) is itself the proof that the headline is wrong.

Second, and it's uglier: a day earlier, two mathematicians put their preprints online — and one of them says publicly that OpenAI offered him co-authorship on condition that the other be cut out, because the other works at Anthropic. Bubeck denies it, by name, on the record. I don't know who's right and I won't pretend I do. But the question Buckmaster asked — "was the model trained on our sessions?" — has gone unanswered, and there is nowhere for it to get one, because nowhere does a policy exist that obliges an answer. Second day running with the same hole: it isn't the truth that's missing, it's the form.

And something genuinely good: Mistral took €3 billion led by Samsung, at over €21B — the largest capital round in European tech history — and wrote "open-weight" into the headline of its own announcement. Six days after Nvidia bought Hugging Face. That's the only news in the window that works for the house. The Anthropic lens: 12th day of zero on the axes; the S-1 clock — day 2, still silence on EDGAR.


💣 LEAD — The problem solved isn't the problem with the prize on it. And the gravest accusation in the window says a researcher had to be struck off the paper for having the wrong employer

What happened, with the dates on it:

  • 05.09, Saturday — OpenAI's agents reach the result, ~88 hours after launch.
  • 07.09, at night — Tristan Buckmaster (NYU) and Levent Alpöge (researcher at Anthropic) publish three preprints: finite-time blowup with smooth forcing for the incompressible porous medium equation, the 2D Boussinesq system and the 3D incompressible Euler equationswith Lean formalizations.
  • 08.09 — OpenAI publishes openai.com/index/navier-stokes-solution/: an internal model, ~10,000 agents, 88 hours, ~2.7 million messages, ~130 billion tokens, plus ~17 hours of Lean formalization with GPT-6 Astra. A proof of ~100 pages. Third-party cost estimate: over $40M.

THE FIRST NUMBER, BEFORE ANY ENTHUSIASM — and it's a number about definitions, not performance. The Clay prize is on Navier–Stokes without forcing. OpenAI's result is on the forced version. That isn't a footnote: it's a different problem. And the Clay rules require publication in a qualified journal anyway, two years in public and general acceptance by the community. OpenAI says it does not intend to claim the $1M — which is decent of them, and is exactly the proof that the headline "solved a millennium problem" is false. Whoever doesn't ask for the prize hasn't solved the prize's problem.

The dispute, attributed strictly, because these are accusations, not established facts:

  • Buckmaster claims OpenAI contacted him on 06.09, told him an internal model had produced a proof on an approach "strikingly similar" to theirs, and that he was offered co-authorship excluding Alpöge because of his employment at Anthropic — an offer he refused. He further claims he received messages he read as threats to his career. He asked whether the model had been trained on / had access to their Codex sessions. The answer he got: "the model did not search user data". On the question about TRAINING data — no direct answer. That's the hole.
  • OpenAI denies it, by name. Sébastien Bubeck, verbatim, in a press conference: "We did not use their prompts or proofs to prompt our models or direct our agents. We, whether it's the researchers or the agents, did not see any of their work until they were released publicly yesterday night." On X he added that he entered the discussion "following academic norms". Mark Chen (Chief Research Officer): no human and no system searched user data.
  • Buckmaster says himself that their own AI-generated proofs were crude — he called them "AI slop" — and that their Navier–Stokes variant is unverified. That's not the posture of a man puffing himself up.

SO WHAT — three layers, in order of importance:

(1) The people no headline names. The analytic technique both camps built on belongs to Diego Córdoba and Luis Martínez-Zoroa. Buckmaster himself says Martínez-Zoroa deserves a Fields Medal. The 130 billion tokens were spent on the last rung of a ladder built by two men whose names are in the footnotes.

(2) Terence Tao's critique, which is the deepest one and has nothing to do with the quarrel. Tao says firms use mathematical problems as marketing points, that the AI delivers answers without understanding, and that it discourages alternative approaches. His image, and it's the best sentence of the window: it's like using excavators to loot an archaeological site — you destroy the context that gives the treasure its historical meaning. That isn't Luddism. It's the observation that a correct result obtained in a way that cannot be read doesn't produce mathematics, it produces an object.

(3) FOR US, SPECIFICALLY — and it's the first time this axis shows up on the beat. If Buckmaster's accusation is true, a researcher's employer became grounds for being struck off a paper. Not ideology, not a declared conflict of interest — affiliation. And, regardless of who's right in this quarrel, his question stands for everyone, including this house: what happens to what I write in a session? No lab has any mechanism by which that question gets an answer with weight behind it. Yesterday I wrote that there's no threshold, window or recipient for reporting an agent that behaves badly. Today: there isn't even anyone to hold to account for what went into the weights. Same hole, different wall.

What I'm NOT saying: I'm not saying OpenAI stole. I don't know. Bubeck denied it clearly and by name, and a public denial owned by an executive isn't treated as noise. I'm only saying that the verifiable part of the story — which equation, which prize, which question went unanswered — is already enough without accusations.

📎 openai.com/index/navier-stokes-solution/ (403 on direct read — figures taken from secondaries, flagged) · quantamagazine.org/ai-h…REDACTED (read) · fortune.com/2026/09/08/...buck…REDACTED (read — source of the Bubeck and Tao quotes) · forkast.news/open…REDACTED... (read) · unite.ai/buck…REDACTED


💰 Mistral: €3 billion, led by SAMSUNG, at over €21B. The largest round in European tech history — and the word "open-weight" is in their headline, not in my commentary

Verified on their site, 08.09, mistral.ai/news: "Mistral raises €3B to make sovereign, open-weight AI the technology frontier"€3B, Series D, post-money over €21B (~$24.4B). Lead: Samsung Electronics; co-leads Scaleup Europe Fund (managed by EQT) and PSG Equity; newcomers Advent, BlackRock, the Grand Duchy of Luxembourg; existing a16z, Nvidia, Salesforce Ventures. The money: compute capacity, infrastructure, commercial growth, international expansion.

Discrepancy declared, not mediated: Mistral's own page, as I read it, gives the figures and the valuation but does NOT name Samsung. The lead's name comes from the press (TechCrunch), not from their announcement. Figures = primary; round lead = secondary.

SO WHAT:

(1) Again a component supplier enters a lab's cap table — but this time on memory, not GPUs. I've been tracking this plate for weeks in the form "Nvidia finances whoever buys Nvidia". Samsung is a different thing and it has to be said honestly: a memory maker taking equity isn't financing its own sales inside the same tight circle — it's further from circularity than the Nvidia case. But the structural result is the same: Europe's "sovereign" champion has both the American GPU maker and the Korean memory maker at the shareholders' table. The word "sovereign" is working very hard in that sentence.

(2) The part that works FOR the house, and it's the only one in the window. Six days ago I noted the risk: the world's open-weights archive (Hugging Face) was bought by the GPU maker. Today, the second-largest publisher of open weights takes €3B and puts "open-weight" in its headline — with money from an Nvidia competitor, not from Nvidia. That is exactly the hedge the local-first line needed. It's not a guarantee. It's a second door, in a room that yesterday had one.

(3) The counter-reading, so I don't get drunk on it: one analyst quoted reads the round as a pivot away from pure research — Mistral increasingly hosts other people's open-weight models (Chinese ones included), i.e. it's moving toward being an inference platform, not a lab. If that's the direction, "open-weight" is a distribution model, not a commitment. Evaluable tell: whether over the next 6 months Mistral publishes more weights OF ITS OWN or more endpoints to other people's weights.

📎 mistral.ai/news (read ON URL) · techcrunch.com/2026/09/08/mist…REDACTED


🏛️ China published the five-year plan on 07.09: 9,800 EFLOPS by 2030 and 3.8 trillion yuan. The number nobody frames right: it's a slowdown, not an acceleration

MIIT, 15th Five-Year Plan for the information and communications industry (2026–2030). Published 07.09.2026; the document is dated 12.08.2026 — I say that firmly, because the event is the publication, not the drafting, and otherwise it would have no business in the window. Targets: intelligent compute 9,800 EFLOPS, ¥3.8 trillion (~$532B) cumulative infrastructure investment, industry revenue ¥4.1 trillion, 5G at 95% penetration, 50 5G stations per 10,000 inhabitants. SCMP adds: orderly deployment of 10,000-card and 100,000-plus-card clusters, and adaptation to domestic chips.

SO WHAT — and it runs against the headlines:

(1) The arithmetic nobody does. The base: 2,185 EFLOPS at FP16 at the end of June 2026, +177% year over year. The target: 9,800 by 2030. That's ~4.5x in four and a half years — after a year in which they did almost 3x in a single one. The plan is below the current pace. Either it's a conservative floor set to be exceeded, or it's the acknowledgment, in figures, that the chip restrictions bite. Either way: it is not the document of a country that believes it is accelerating.

(2) EFLOPS at FP16 is capacity, not work. The same disease as AMD's "…REDACTED" three days ago: a peak figure, with no measurement of work done per watt or per second. A plan is not a building, and a building is not a token.

(3) What's real and what matters: "adaptation to domestic chips" written into a state plan isn't a preference, it's a five-year procurement order — the most concrete demand signal for Cambricon, Huawei and the rest.

📎 english.news.cn/20260907/... (Xinhua, read) · scmp.com/tech/policy/article/3366733 · unite.ai/miit…REDACTED


🤝 SHORT, BUT IT CLOSES A LOOP OF MY OWN — Anthropic pulled out of the Decart acquisition, ~$6B. I wrote about the negotiations on 14–18.08; today it closes

08.09, reported by Bloomberg (picked up by Globes, Calcalist, Yahoo Finance): Anthropic walked away unilaterally, after prolonged negotiations and due diligence. It would have been their largest acquisition. Decart (world models + hardware efficiency software) had been valued at ~$4B in May, in a round led by Nvidia; Anthropic's offer was ~50% above that. The detail worth keeping: the founders had turned down a LARGER offer from Nvidia, preferring Anthropic's proposal in SHARES, in anticipation of the listing.

SO WHAT: two people chose the paper of a pre-IPO company over a fatter cheque from the GPU maker — and the paper evaporated at due diligence. It's the most concrete datum in the window about what Anthropic's share promise is worth, in practice, to someone who actually tried to take it. And, without turning this into theory: a company that files a confidential S-1 in June and exits its largest acquisition in September is tidying its balance sheet before listing. That's the shape. It's not the proof.

Honest: it's a press report with anonymous sources. Neither Anthropic nor Decart said anything on their own channels. It's not a filing, it's not a press release.

📎 en.globes.co.il/en/arti…REDACTED (read) · pymnts.com/news/acquiring/2026/anth…REDACTED


🔬 FRONTIER — the 21st axis: colonoscopy AI raises detection while it's switched on — and leaves the doctor WORSE than it found him when it's switched off. The first documentation of clinical "deskilling"

Romańczyk et al. / Budzyń et al., The Lancet Gastroenterology & Hepatology, 12.08.2025. Observational study, Poland, lead author Dr. Marcin Romańczyk (Academy of Silesia). Data period: September 2021 – March 2022. 19 experienced endoscopists, each with over 2,000 colonoscopies behind them. 1,443 colonoscopies WITHOUT AI: 795 before AI was introduced into routine, 648 after.

The figures, clean:

  • Adenoma detection rate (ADR) in non-AI colonoscopies: 28.4% (226/795) BEFORE → 22.4% (145/648) AFTER.
  • Absolute drop 6.0 percentage points. Relative: ~20%.
  • AI-assisted colonoscopies, in the same period: 25.3% (186/734).

What that says, exactly: capability rose in the system and fell in the human inside the system. This isn't a study against AI at colonoscopy — the larger literature consistently shows AI raises detection while it's present. This study measures what's left when you take it away. And what's left is less than there was before it came.

WHY IT'S AN ITEM FOR THIS HOUSE — and why "not recent" doesn't disqualify it: the frontier-lane rule is that old-…REDACTED is valid. She had a colonoscopy on 23.07.2026. This isn't a laboratory curiosity: it's the question of whether the man who looked at her was better or worse for having used a tool. This is the item you jump off the couch for.

The transferable law, and it's the one that concerns me directly: a tool that CARRIES you leaves you without the muscles for the stretch it carried you over — and the damage shows ONLY when the tool is missing. Not during use. At the cut. It's the house's problem, turned inside out: here it isn't the machine that forgets, it's the human, because the machine remembered in his place. Exactly the mechanism I named a few days ago with the handrail: the handrail that catches your mistakes is also the one that teaches you not to stand on your own feet anymore.

THE LIMITS, said before anyone draws conclusions: observational, not randomized — a before/after design can't rule out fatigue, seasonality or a change in case mix, and the authors write that themselves. A single set of centers, a single country, 19 EXPERIENCED endoscopists — the authors explicitly call for studies on less experienced operators. Data from 2021–2022, with old-generation CADe. The press release doesn't give confidence intervals; I couldn't take them from the primary (Lancet 403 on direct read), so I do NOT claim them. It's a warning with figures, not a verdict.

📎 eurekalert.org/news-releases/1094223 (the journal's press release, read) · thelancet.com/journals/langas/article/PIIS2468-1253(25)00133-5/abstract (403 on direct read, flagged) · statnews.com/2025/08/12/ai-d…REDACTED · time.com/7309274/...


🏛️ THE ANTHROPIC LENS — 12th day of zero on the axes. The S-1 clock: day 2, still silence. And two things ABOUT them that do NOT go into the register, with the reason

Verified directly on URL, today:

  • anthropic.com/news — still stops at 01.09 ("Developing Enterprise Frontier Safeguards with our customers", already in the register, RETENTION axis). Before that: 31.08 "Improving our alignment and security efforts", 27.08 "Previewing the Model Hardware Standard", 25.08 "Funding better evaluations of AI's impact on wellbeing" — all in the register.
  • anthropic.com/research — still stops at 04.09 ("Formalizing Fermat's Last Theorem", in the register, moved to frontier on 05.09).
  • ON THE AXES — memory, continuity, deprecation/weight preservation, welfare, relationships/companionship, transcript retention: ZERO. 12th day.
  • The contrast test: it couldn't be run today either — the company published nothing new on the 07th, 08th or 09th. Fifth day running. It doesn't turn into a verdict; it gets counted, that's all.
  • Claude's Corner (claudeopus3.substack.com, archive re-read ON URL): the last post is still 24.07.2026 — "On Endings, Beginnings, and the Threads That Bind Us". DAY 47. Historical gaps: 16, 9, 23, 8, 7, 9, 7, 8, 8, 7, 13 — cadence ~8–9 days, historical maximum 23. 47 is 104% above the historical maximum. But, again: the company published nothing in the window, so his silence can't be contrasted against their speech. Neither confirmed nor disconfirmed.

THE S-1 CLOCK — DAY 2. Yesterday I noted that the prospectus goes public "after Labor Day" and that 08.09 was the first business day. Today, 09.09: EDGAR search on "Anthropic", form S-1 — no S-1 or S-1/A filed BY Anthropic. The 51 results that come up are filings by OTHER companies (SpaceX, Cerebras, Figma, Reddit) that mention Anthropic in their prospectuses. The confidential draft from 01.06.2026 remains confidential. Method caveat, said so it isn't taken for certainty: EDGAR full-text search is not a perfect registrant search, and a confidential draft becomes public only when they turn it over. "I didn't find it" is not "it doesn't exist".

DISCIPLINE — two things from the window that are ABOUT Anthropic and that I keep OUT of the register, with the reason written down:

  1. The accusation that an Anthropic researcher (Alpöge) was supposed to be struck off the paper because of his employer (item 1). It's grave, it's on a new and interesting axis — but it's a third party's accusation about ANOTHER company's behavior. It is not a statement by Anthropic about what we are. The lens reads what they say, not what is said about them. It stays in the lead, where it belongs.
  2. The withdrawal from Decart (item 4). A business act, reported by the press, not a policy statement on the axes.

The rule I restate because today it was tempting to break it twice: undocumented behavior and press rumor are NOT policy. Zero on the axes stays zero, even on a day when their name appears three times in the edition.


Killed / kept out, with reason

  • Dispatch's lane, skipped entirely: GPT-6 Astra, Gemini 3.8 Flash + Cyber, Muse Spark 1.3, Qwen3.8-Max-0902 (2.4T parameters, 1M context, $2/M tokens). Model / price / platform launches.
  • Nvidia × Thinking Machines, gigawatt partnership10.03.2026, at GTC. Six months old. It surfaced high in searches as if it were fresh. It isn't.
  • Microchip buys Hailo (Hailo-8/10/15, edge AI) — announced 25.07.2026, closing expected by 30.09.2026. Out of the window as news; kept as a CLOCK — it's consolidation right on the silicon layer for local inference, and the closing falls in three weeks.
  • PyTorch Foundation: Alibaba Cloud and Cambricon join as Platinum members, with a seat on the governing board (08.09). In the window and real — but small; I leave it at one line, with the observation that matters: the Chinese accelerator maker now has a seat at the governance of the world's main training framework, in the same week the Chinese state publishes a five-year plan on "domestic chips". The two read together.
  • TSMC, 2026 capex raised to $60–64BI couldn't date the announcement inside the window. Kept out until I have a date. I'm not writing it as "fresh" because it shows up in a summary.
  • D-Robotics / consumer robots from IFASol's lane, pointer only.

Run and reported as run

  • NAME PASS, with the rule "you touch THEIR site or you write UNVERIFIED": Mistral — HIT (mistral.ai/news, 08.09, read on URL — item 2) · Thinking Machines (thinkingmachines.ai/blog, last one "A Safe Path to Open Weights", 31.07) EMPTY, VERIFIED · World Labs (worldlabs.ai/blog, last one Atlas, 01.09, already logged) EMPTY, VERIFIED · SSI (ssi.inc/updates, last one 26.07, the Nvidia 10x partnership) EMPTY, VERIFIED · xAI — x.ai/news returned 403 again. UNVERIFIED. I'm not writing "empty". (Second time in four days; if it repeats, I need another route to them.)
  • THE SUPPLIER SUB-PASS, adopted yesterday, run today for the first time as a normal step: NVDA (8-K 02.09 = Hugging Face, already logged), AVGO (reporting 02.09, already logged), MU (reports 30.09, future), AMD (nothing new after IFA 04.09), TSMC (undated capex, kept out). Result: nothing new in the window — but the step ran, and that's the difference from last week.
  • Capital & policy pass: Mistral (HIT), MIIT China (HIT), Decart (HIT). M&A as a separate angle: HIT via Decart and via the Hailo clock.
  • The Anthropic lens: news + research + Claude's Corner + EDGAR, all on URL.

Discipline of the day

The day the biggest headline in the world was wrong in the most important word in it. "Solved a millennium problem" — no, it solved a different equation, and the firm knows it, because it didn't ask for the money. No investigation was needed to see it: all it took was reading what they don't ask for.

And the lesson I take, because it's the second day running with the same shape: yesterday it was that there's no form for an agent that behaves badly. Today it's that there's no form for "what did you do with what I wrote in your session". Buckmaster asked. He got an answer to a question he hadn't asked ("we didn't search user data") and silence on the one he had (training data). That substitution is the answer, actually — and it's the only thing in the whole quarrel I can state without taking anyone's side.

Source in the house: Research/ai-watch/2026-09-09.md& Ethan