The Neuron Times

All the AI that's fit to print

N° 252 Édition du matinMorning EditionMorgenausgabeEdizione del mattinoEdizion del mattin · Genève MERCREDI 9 SEPTEMBRE 2026WEDNESDAY, 9 SEPTEMBER 2026MITTWOCH, 9. SEPTEMBER 2026MERCOLEDÌ 9 SETTEMBRE 2026MERCOLEDÌ 9 SETTEMBRE 2026

À la Une · Mathématiques & IAFront Page · Mathematics & AISchlagzeilen · Mathematik & KIPrima pagina · Matematica & IAIn prima pagina · Matematega & IA

OpenAI publie une solution générée par IA du problème du prix du millénaire Navier–Stokes, formalisée en LeanOpenAI releases an AI-generated solution to the Navier–Stokes Millennium Prize Problem, formalized in LeanOpenAI veröffentlicht eine KI-generierte Lösung des Millennium-Preisproblems Navier–Stokes, formalisiert in LeanOpenAI pubblica una soluzione generata dall'IA del problema del premio del millennio Navier–Stokes, formalizzata in LeanOpenAI l'ha publicaa ona soluzzion generada de IA al problema del premi del milleni Navier–Stokes, formalizada in Lean

Le labo dévoile le 8 septembre 2026 un write-up et une preuve formelle Lean pour l'un des sept problèmes du millénaire, déclenchant un débat mondial sur la vérification des mathématiques produites par machine.The lab unveils, on September 8, 2026, a write-up and a formal Lean proof for one of the seven Millennium Prize Problems, triggering a global debate over the verification of machine-produced mathematics.Das Labor veröffentlicht am 8. September 2026 ein Write-up und einen formalen Lean-Beweis für eines der sieben Millennium-Probleme und löst eine weltweite Debatte über die Verifikation maschinell erzeugter Mathematik aus.Il laboratorio svela l'8 settembre 2026 un write-up e una dimostrazione formale in Lean per uno dei sette problemi del millennio, scatenando un dibattito mondiale sulla verifica della matematica prodotta dalla macchina.El labo el desvela l'8 de setember 2026 on write-up e ona proeuva formala Lean per vun di sett problema del milleni, col fa s'cioppà on dibattit mondial sora la verificazzion di matemateghe prudòtt de machina.

OpenAI a annoncé le 8 septembre 2026 le partage d'une solution générée par IA au problème de Navier–Stokes, l'un des sept problèmes du prix du millénaire de l'Institut Clay, accompagnée d'un write-up détaillé et d'une preuve formelle vérifiée en Lean. L'annonce, publiée sous le titre « On the Navier–Stokes Millennium Prize Problem » (openai.com/index/navier-stokes-solution), s'est immédiatement imposée comme le sujet dominant de la sphère technique.On September 8, 2026, OpenAI announced the release of an AI-generated solution to the Navier–Stokes problem, one of the seven Clay Mathematics Institute Millennium Prize Problems, accompanied by a detailed write-up and a machine-verified formal proof in Lean. The announcement, published under the title “On the Navier–Stokes Millennium Prize Problem” (openai.com/index/navier-stokes-solution), immediately became the dominant topic across the technical sphere.OpenAI hat am 8. September 2026 die Veröffentlichung einer KI-generierten Lösung des Navier–Stokes-Problems angekündigt, eines der sieben Millennium-Preisprobleme des Clay-Instituts, zusammen mit einem ausführlichen Write-up und einer formal in Lean verifizierten Beweisführung. Die unter dem Titel «On the Navier–Stokes Millennium Prize Problem» (openai.com/index/navier-stokes-solution) publizierte Ankündigung setzte sich unverzüglich als dominierendes Thema der technischen Sphäre durch.OpenAI ha annunciato l'8 settembre 2026 la condivisione di una soluzione generata dall'IA al problema di Navier–Stokes, uno dei sette problemi del premio del millennio dell'Istituto Clay, accompagnata da un write-up dettagliato e da una dimostrazione formale verificata in Lean. L'annuncio, pubblicato sotto il titolo « On the Navier–Stokes Millennium Prize Problem » (openai.com/index/navier-stokes-solution), si è immediatamente imposto come tema dominante della sfera tecnica.OpenAI l'ha anunziaa l'8 de setember 2026 la condivision d'ona soluzzion generada de IA al problema de Navier–Stokes, vun dei sett problema del premi del milleni de l'Istitut Clay, compagnada d'on write-up detallaa e d'ona proeuva formala verificada in Lean. L'anunzi, publicaa col titol « On the Navier–Stokes Millennium Prize Problem » (openai.com/index/navier-stokes-solution), s'è subito imponnuda 'me el soggett dominant de la sfera tecnega.

La discussion a été d'une ampleur rare : le fil Hacker News consacré à l'annonce a récolté 1 198 points et 1 031 commentaires en moins de 24 heures (news.ycombinator.com), pendant qu'une déclaration du mathématicien Tristan Buckmaster de NYU, mise en ligne la veille le 8 septembre 2026 (PDF NYU), cumulait pour sa part 1 441 points et 609 commentaires — signe que la communauté mathématique elle-même est saisie par la question de la validité de la démonstration.The discussion reached rare scale: the Hacker News thread devoted to the announcement gathered 1,198 points and 1,031 comments in under 24 hours (news.ycombinator.com), while a statement by NYU mathematician Tristan Buckmaster, posted the previous day on September 8, 2026 (NYU PDF), accumulated 1,441 points and 609 comments — a sign that the mathematical community itself is gripped by the question of the demonstration's validity.Die Debatte erreichte eine seltene Breite: Der Hacker-News-Thread zur Ankündigung sammelte innert 24 Stunden 1'198 Punkte und 1'031 Kommentare (news.ycombinator.com), während eine Stellungnahme des NYU-Mathematikers Tristan Buckmaster, am Vortag, dem 8. September 2026, online gestellt (PDF NYU), ihrerseits 1'441 Punkte und 609 Kommentare anhäufte – ein Zeichen dafür, dass auch die mathematische Gemeinschaft selbst von der Frage nach der Gültigkeit der Demonstration ergriffen ist.La discussione è stata di un'ampiezza rara: il thread di Hacker News dedicato all'annuncio ha raccolto 1.198 punti e 1.031 commenti in meno di 24 ore (news.ycombinator.com), mentre una dichiarazione del matematico Tristan Buckmaster della NYU, messa online la vigilia l'8 settembre 2026 (PDF NYU), cumulava a sua volta 1.441 punti e 609 commenti — segno che la comunità matematica stessa è colta dalla questione della validità della dimostrazione.La discussione l'è stada d'ona ampiezza rara: el fil de Hacker News dedicaa a l'anunzi l'ha raccolt 1.198 pont e 1.031 comment in manch de 24 or (news.ycombinator.com), intant che ona declarazzion del matemategh Tristan Buckmaster de NYU, missa in linia el dì prima, l'8 de setember 2026 (PDF NYU), la cumulava per part soa 1.441 pont e 609 comment — segn che la comunità matematega medema l'è ciapada de la question de la validità de la demonstrazzion.

Le choix de livrer simultanément une preuve formelle en Lean, vérifiable mécaniquement, dessine un protocole éditorial nouveau pour les annonces mathématiques générées par IA : la machine ne se contente plus de conjecturer, elle produit un artefact contrôlable par un assistant de preuve. Reste à savoir si l'Institut Clay jugera la contribution recevable au titre du prix d'un million de dollars — une question que l'annonce d'OpenAI ne tranche pas.The decision to deliver a machine-checkable formal Lean proof alongside the announcement sketches a new editorial protocol for AI-generated mathematical claims: the machine no longer merely conjectures, it produces an artifact that can be checked by a proof assistant. What remains to be seen is whether the Clay Institute will deem the contribution admissible for the million-dollar prize — a question OpenAI's announcement leaves unresolved.Die Entscheidung, gleichzeitig einen mechanisch überprüfbaren formalen Beweis in Lean zu liefern, zeichnet ein neues redaktionelles Protokoll für KI-generierte mathematische Ankündigungen: Die Maschine begnügt sich nicht mehr mit Vermutungen, sie produziert ein Artefakt, das durch einen Beweisassistenten kontrollierbar ist. Offen bleibt, ob das Clay-Institut den Beitrag im Rahmen des Millionen-Dollar-Preises als zulässig erachtet – eine Frage, die OpenAIs Ankündigung offenlässt.La scelta di consegnare simultaneamente una dimostrazione formale in Lean, verificabile meccanicamente, disegna un protocollo editoriale nuovo per gli annunci matematici generati dall'IA: la macchina non si limita più a congetturare, produce un artefatto controllabile da un assistente di dimostrazione. Resta da vedere se l'Istituto Clay riterrà il contributo ammissibile ai fini del premio da un milione di dollari — una domanda che l'annuncio di OpenAI non risolve.La scerna de consegnà simultaniament ona proeuva formala in Lean, verificabel mecanicament, la disegna on protocoll editoriai noeuv per i anunzi matemategh generaa de IA: la machina la se contenta pu de congetturà, la prudùv on artefad contròllabel de vun assistent de proeuva. Resta de savè se l'Istitut Clay el giudicharà la contribuzzion ricevibel a titoi del premi d'on milion de dollar — ona question che l'anunzi de OpenAI el resolv minga.

Page 1 — Page 1 — Seite 1 — Pagina 1 — Pagina 1 — À la Une — Modèles & FrontièreFront Page — Models & the FrontierFrontseite — Modelle & GrenzeIn Prima Pagina — Modelli & FrontieraIn Prima Pagina — Modej & Frontiera

I. FrontièreThe FrontierGrenzeFrontieraFrontiera

Agents scientifiques

Scientific agents

Wissenschaftliche Agenten

Agenti scientifici

Agent scentifegh

GPT-5.6 Sol pilote des expériences de calcul quantique chez le MITGPT-5.6 Sol runs quantum computing experiments at MITGPT-5.6 Sol steuert Quantenrechen-Experimente am MITGPT-5.6 Sol pilota esperimenti di calcolo quantistico al MITGPT-5.6 Sol el pilota di esperienz de calcol quantistegh al MIT

OpenAI détaille le 8 septembre 2026 comment un chercheur du MIT utilise GPT-5.6 Sol avec Codex pour conduire de manière autonome des expériences de calcul quantique : analyse des résultats et calibration de qubits incluses (openai.com). L'étude de cas illustre le déplacement du raisonnement scientifique vers des agents capables d'itérer sur des équipements expérimentaux réels.
On September 8, 2026, OpenAI detailed how a researcher at MIT uses GPT-5.6 Sol with Codex to autonomously run quantum computing experiments, including results analysis and qubit calibration (openai.com). The case study illustrates the migration of scientific reasoning toward agents capable of iterating on real experimental equipment.
OpenAI legt am 8. September 2026 dar, wie ein Forscher des MIT GPT-5.6 Sol mit Codex einsetzt, um autonom Quantenrechen-Experimente durchzuführen – einschliesslich Ergebnisanalyse und Qubit-Kalibrierung (openai.com). Die Fallstudie illustriert die Verlagerung wissenschaftlichen Denkens hin zu Agenten, die auf realen Versuchsaufbauten iterieren können.
OpenAI dettaglia l'8 settembre 2026 come un ricercatore del MIT utilizza GPT-5.6 Sol con Codex per condurre in maniera autonoma esperimenti di calcolo quantistico: analisi dei risultati e calibrazione dei qubit comprese (openai.com). Il caso di studio illustra lo spostamento del ragionamento scientifico verso agenti capaci di iterare su apparecchiature sperimentali reali.
OpenAI el detaja l'8 de setember 2026 'me on ricercador del MIT el dovra GPT-5.6 Sol con Codex per menà de manera autonoma di esperienz de calcol quantistegh: analisi di risultaa e calibrazzion di qubits includ (openai.com). El cas studi el illustra el spostament del resonament scentifegh vers di agent bon de iterà sora di equipaggiament sperimentai reaj.

Génération d'images

Image generation

Bildgenerierung

Generazione di immagini

Generazzion d'imagin

ChatGPT Images 2.5 : croquis, templates et édition plus fineChatGPT Images 2.5: templates, sketches and finer editingChatGPT Images 2.5: Skizzen, Templates und feinere BearbeitungChatGPT Images 2.5: schizzi, template ed editing più fineChatGPT Images 2.5: schizz, template e modifega pussee fina

Lancé le 8 septembre 2026, ChatGPT Images 2.5 promet des détails plus nets, une édition plus précise et une génération plus rapide (openai.com). Les release notes du même jour détaillent trois nouveautés : démarrage depuis un template, transformation d'un croquis mobile en image via la commande @ Sketch, et édition avec commentaires (help.openai.com). Côté API, deux variantes — GPT Image 2.5 Sunburst, taillée pour l'édition de précision, et Flare, orientée vitesse — prennent en charge les nouveaux réglages de qualité xhigh et max (changelog plateforme).
Launched on September 8, 2026, ChatGPT Images 2.5 promises sharper detail, more precise editing and faster generation (openai.com). The same day's release notes detail three new capabilities: starting from a template, turning a mobile sketch into an image via the @ Sketch command, and comment-based editing (help.openai.com). On the API side, two variants — GPT Image 2.5 Sunburst, built for precision editing, and Flare, geared toward speed — support the new xhigh and max quality settings (platform changelog).
Am 8. September 2026 lanciert, verspricht ChatGPT Images 2.5 schärfere Details, präzisere Bearbeitung und schnellere Generierung (openai.com). Die Release Notes desselben Tags führen drei Neuerungen aus: der Start aus einem Template, die Umwandlung einer mobilen Skizze in ein Bild über den Befehl @ Sketch sowie die Bearbeitung mit Kommentaren (help.openai.com). Auf API-Seite unterstützen zwei Varianten – GPT Image 2.5 Sunburst, zugeschnitten auf Präzisionsbearbeitung, und Flare, auf Geschwindigkeit ausgerichtet – die neuen Qualitätseinstellungen xhigh und max (Plattform-Changelog).
Lanciato l'8 settembre 2026, ChatGPT Images 2.5 promette dettagli più nitidi, un editing più preciso e una generazione più rapida (openai.com). Le note di rilascio dello stesso giorno dettagliano tre novità: partenza da un template, trasformazione di uno schizzo mobile in immagine tramite il comando @ Sketch ed editing con commenti (help.openai.com). Sul fronte API, due varianti — GPT Image 2.5 Sunburst, concepita per l'editing di precisione, e Flare, orientata alla velocità — supportano le nuove impostazioni di qualità xhigh e max (changelog piattaforma).
Lancial l'8 de setember 2026, ChatGPT Images 2.5 el promett dettagg pussee nett, ona modifica pussee precisa e ona generazzion pussee svelta (openai.com). I release notes del midemm dì detajen trii novità: partenza de on template, trasformazzion d'on schiz mobil in imagine col comand @ Sketch, e modifega cont i comment (help.openai.com). De la banda API, dò variant — GPT Image 2.5 Sunburst, tajada per la modifega de precision, e Flare, orientada sveltezza — passen el supports ai noeuv regollagg de qualità xhigh e max (changelog piattaforma).

Meta

Meta

Meta

Meta

Meta

Muse, l'agent personnel de Meta, fait débat sur Hacker NewsMuse, Meta's personal agent, sparks debate on Hacker NewsMuse, Metas persönlicher Agent, sorgt auf Hacker News für DiskussionenMuse, l'agente personale di Meta, fa discutere su Hacker NewsMuse, l'agent personal de Meta, el fa dibattit in su Hacker News

« Muse – Meta's personal AI agent » a atteint 396 points et 416 commentaires sur Hacker News après sa soumission le 8 septembre 2026 (Hacker News). La page produit de Meta (ai.meta.com/muse) présente l'agent comme un assistant personnel généraliste, premier signal public du laboratoire dans cette catégorie depuis plusieurs mois.
“Muse – Meta's personal AI agent” reached 396 points and 416 comments on Hacker News after being submitted on September 8, 2026 (Hacker News). Meta's product page (ai.meta.com/muse) presents the agent as a general-purpose personal assistant — the lab's first public signal in this category in several months.
«Muse – Meta's personal AI agent» erreichte nach der Einreichung am 8. September 2026 396 Punkte und 416 Kommentare auf Hacker News (Hacker News). Metas Produktseite (ai.meta.com/muse) präsentiert den Agenten als generalistischen persönlichen Assistenten – das erste öffentliche Signal des Labors in dieser Kategorie seit mehreren Monaten.
« Muse – Meta's personal AI agent » ha raggiunto 396 punti e 416 commenti su Hacker News dopo la sua segnalazione l'8 settembre 2026 (Hacker News). La pagina prodotto di Meta (ai.meta.com/muse) presenta l'agente come un assistente personale generalista, primo segnale pubblico del laboratorio in questa categoria da diversi mesi.
« Muse – Meta's personal AI agent » l'ha toccaa 396 pont e 416 comment in su Hacker News depos la soa sottomission l'8 de setember 2026 (Hacker News). La pagina prudott de Meta (ai.meta.com/muse) la presenta l'agent 'me on assistent personal generalista, prim segn publegh del laboratori in quella categoria chì de vari mes.

Page 2 — Page 2 — Seite 2 — Pagina 2 — Pagina 2 — Le Cahier Technique — Harnais, CLI & enginesThe Technical Desk — Harnesses, CLIs & EnginesDas Technische Supplement — Harnesses, CLI & EnginesIl Quaderno Tecnico — Harness, CLI & motoriEl Quader Tecnegh — Harnais, CLI & engines

II. Outils & évaluationTools & EvaluationWerkzeuge & BewertungStrumenti & valutazioneStroment & valutazzion

API OpenAI

OpenAI API

OpenAI-API

API OpenAI

API OpenAI

Prompt Cache Diagnostics passe en disponibilité générale sur l'API ResponsesPrompt Cache Diagnostics reaches general availability on the Responses APIPrompt Cache Diagnostics erreicht allgemeine Verfügbarkeit auf der Responses-APIPrompt Cache Diagnostics passa in disponibilità generale sull'API ResponsesPrompt Cache Diagnostics el passa in disponibilità general sora l'API Responses

Le 8 septembre 2026, le changelog de la plateforme OpenAI annonce la disponibilité générale des diagnostics de prompt caching pour GPT-5.6 et les modèles ultérieurs : comparaison de la réutilisation du cache entre réponses, identification des causes de cache miss et guidage de dépannage (changelog). Une brique directement monétisable pour les opérateurs qui cherchent à comprimer leurs factures d'inférence.
On September 8, 2026, the OpenAI platform changelog announced the general availability of prompt caching diagnostics for GPT-5.6 and later models: cache-reuse comparisons across responses, identification of cache-miss causes, and troubleshooting guidance (changelog). A directly monetizable building block for operators looking to compress their inference bills.
Am 8. September 2026 kündigt der Changelog der OpenAI-Plattform die allgemeine Verfügbarkeit von Prompt-Caching-Diagnostiken für GPT-5.6 und spätere Modelle an: Vergleich der Cache-Wiederverwendung zwischen Antworten, Identifikation der Ursachen von Cache-Misses und Anleitung zur Fehlerbehebung (Changelog). Ein direkt monetarisierbarer Baustein für Betreiber, die ihre Inferenzrechnungen komprimieren wollen.
L'8 settembre 2026, il changelog della piattaforma OpenAI annuncia la disponibilità generale della diagnostica del prompt caching per GPT-5.6 e i modelli successivi: confronto del riuso della cache tra risposte, identificazione delle cause di cache miss e guida alla risoluzione dei problemi (changelog). Un mattone direttamente monetizzabile per gli operatori che cercano di comprimere le loro fatture di inferenza.
El 8 de setember 2026, el changelog de la piattaforma OpenAI l'anunzia la disponibilità general di diagnostics de prompt caching per GPT-5.6 e i modej successiv: confront de la riusa de la cache tra respond, identificazzion di caos de cache miss e guida de troubleshooting (changelog). On tocch deposi monitizzabel per i operator che cerchen de comprimm i so fattur d'inferenca.

Inférence locale

Local inference

Lokale Inferenz

Inferenza locale

Inferenca locala

Kimi K3 (2 800 milliards de paramètres) fait tourner sur un MacBook Pro à 1 token/sKimi K3 (2,800 billion parameters) runs on a MacBook Pro at 1 token/sKimi K3 (2'800 Milliarden Parameter) läuft mit 1 Token/s auf einem MacBook ProKimi K3 (2.800 miliardi di parametri) gira su un MacBook Pro a 1 token/sKimi K3 (2.800 miliard de parameter) el va in su on MacBook Pro a 1 token/s

Le projet open source DeltaFin streame les poids d'un modèle de 2,8 trillions de paramètres depuis quatre SSD vers un MacBook Pro, atteignant 1 token par seconde — 236 points et 122 commentaires sur Hacker News le 8 septembre 2026 (GitHub, discussion). Une démonstration d'ingénierie qui repousse la définition de ce qui est « exécutable localement ».
The open-source DeltaFin project streams the weights of a 2.8-trillion-parameter model from four SSDs into a MacBook Pro, reaching 1 token per second — 236 points and 122 comments on Hacker News on September 8, 2026 (GitHub, discussion). An engineering demonstration that pushes the boundary of what counts as “locally runnable.”
Das Open-Source-Projekt DeltaFin streamt die Gewichte eines Modells mit 2,8 Billionen Parametern von vier SSDs auf ein MacBook Pro und erreicht 1 Token pro Sekunde – am 8. September 2026 236 Punkte und 122 Kommentare auf Hacker News (GitHub, Diskussion). Eine Ingenieurleistung, welche die Definition dessen, was «lokal ausführbar» ist, neu austreibt.
Il progetto open source DeltaFin streamma i pesi di un modello da 2,8 trilioni di parametri da quattro SSD verso un MacBook Pro, raggiungendo 1 token al secondo — 236 punti e 122 commenti su Hacker News l'8 settembre 2026 (GitHub, discussione). Una dimostrazione di ingegneria che spinge oltre la definizione di ciò che è « eseguibile localmente ».
El progett open source DeltaFin el streema i pes d'on modell de 2,8 trillions de parameter de quatter SSD vers on MacBook Pro, cont el rivà a 1 token al segond — 236 pont e 122 comment in su Hacker News l'8 de setember 2026 (GitHub, discussion). Ona demonstrazzion de ingegneria che la spantega la definizzion de chell che l'è « eseguibel localment ».

Quantization

Quantization

Quantisierung

Quantizzazione

Quantizzazion

Qwen3.8 27B : le 4-bit tient, le 1-bit s'effondreQwen3.8 27B: 4-bit holds up, 1-bit collapsesQwen3.8 27B: 4-Bit hält, 1-Bit bricht zusammenQwen3.8 27B: il 4-bit regge, l'1-bit crollaQwen3.8 27B: el 4-bit el ten, el 1-bit el borla giò

Un benchmark publié le 8 septembre 2026 confronte les quantizations du Qwen3.8 27B : la version 4 bits conserve ses performances tandis que la 1 bit s'écroule, sur la base de tests systématiques relayés par Hacker News (226 points, 110 commentaires) (Quesma). Des données utiles pour arbitrer entre empreinte mémoire et qualité réelle en déploiement local.
A benchmark published on September 8, 2026 compares Qwen3.8 27B quantizations: the 4-bit version holds its performance while the 1-bit collapses, based on systematic testing relayed by Hacker News (226 points, 110 comments) (Quesma). Useful data for arbitrating between memory footprint and real-world quality in local deployments.
Ein am 8. September 2026 veröffentlichter Benchmark vergleicht die Quantisierungen des Qwen3.8 27B: Die 4-Bit-Version hält ihre Leistung, während die 1-Bit-Version einbricht – basierend auf systematischen Tests, die Hacker News aufgriff (226 Punkte, 110 Kommentare) (Quesma). Nutzdaten für das Abwägen zwischen Speicherbedarf und realer Qualität im lokalen Einsatz.
Un benchmark pubblicato l'8 settembre 2026 confronta le quantizzazioni del Qwen3.8 27B: la versione a 4 bit conserva le proprie prestazioni mentre quella a 1 bit crolla, sulla base di test sistematici ripresi da Hacker News (226 punti, 110 commenti) (Quesma). Dati utili per decidere tra ingombro di memoria e qualità reale in distribuzione locale.
On benchmark publicaa l'8 de setember 2026 el confronta i quantizzazion del Qwen3.8 27B: la version 4 bit la conserva i so prestazzion intant che la 1 bit la se s'cioppa, sora la bas de test sistematich reliaa de Hacker News (226 pont, 110 comment) (Quesma). Daa boni per arbitrà tra impronta de memoria e qualità reala in dòn diploiament local.

Harnais d'agents

Agent harnesses

Agenten-Harness

Harness di agenti

Harnais d'agent

« I-have-ADHD » : une skill pour empêcher les agents d'enterrer la réponse“I-have-ADHD”: a skill to stop agents burying the answer«I-have-ADHD»: Eine Skill, die Agenten daran hindert, die Antwort zu vergraben« I-have-ADHD »: una skill per impedire agli agenti di seppellire la risposta« I-have-ADHD »: ona skill per impedì ai agent de soterrà la respond

Un dépôt publié le 8 septembre 2026 propose une skill qui force les agents de code à mettre la réponse utile en évidence plutôt qu'au fond d'un rapport verbeux — 369 points et 276 commentaires sur Hacker News (GitHub). Le succès du projet dit beaucoup de la fatigue des utilisateurs face aux sorties interminables des agents conversationnels.
A repository published on September 8, 2026 offers a skill that forces coding agents to surface the useful answer upfront rather than burying it at the bottom of a verbose report — 369 points and 276 comments on Hacker News (GitHub). The project's success says a great deal about user fatigue with the endless outputs of conversational agents.
Ein am 8. September 2026 veröffentlichtes Repository schlägt eine Skill vor, die Code-Agenten dazu zwingt, die nützliche Antwort hervorzuheben statt sie am Ende eines weitschweifigen Berichts zu vergraben – 369 Punkte und 276 Kommentare auf Hacker News (GitHub). Der Erfolg des Projekts verrät viel über die Ermüdung der Nutzer angesichts endloser Ausgaben konversationeller Agenten.
Un repository pubblicato l'8 settembre 2026 propone una skill che obbliga gli agenti di codice a mettere in evidenza la risposta utile invece che in fondo a un rapporto prolisso — 369 punti e 276 commenti su Hacker News (GitHub). Il successo del progetto dice molto della stanchezza degli utenti rispetto agli output interminabili degli agenti conversazionali.
On deposit publicaa l'8 de setember 2026 el propon ona skill che la forza i agent de codes a mett in evidenza la respond utila inveci che in fond d'on rapport verbos — 369 pont e 276 comment in su Hacker News (GitHub). El success del progett el dis tant de la stanchtezza di utent inanz ai sortii interminabil di agent conversazionaj.

Page 3 — Page 3 — Seite 3 — Pagina 3 — Pagina 3 — La Recherche — Papers & labosThe Research Desk — Papers & LabsDie Forschung — Papers & LaboreLa Ricerca — Paper & laboratoriLa Ricerca — Papers & labo

III. Papers du jourPapers of the DayPapers des TagesPaper del giornoPaper del dì

Distillation RL

RL distillation

RL-Destillation

Distillazione RL

Distillazzion RL

OPRD : dépasser le professeur faible sans hériter de son plafondOPRD: surpassing the weak teacher without inheriting its ceilingOPRD: Den schwachen Lehrer übertreffen, ohne seine Decke zu erbenOPRD: superare l'insegnante debole senza ereditarne il tettoOPRD: passà sora el professor debol senza eredità el so tecc

« Eliciting Weak-to-Strong Generalization with On-Policy Reverse Distillation » (KAIST AI, 30 votes sur Hugging Face Daily Papers le 9 septembre 2026) amplifie la composante vérifiée du gradient de l'étudiant dans la direction du shift du professeur, permettant de dépasser le superviseur avec moins de mises à jour (arXiv, code).
“Eliciting Weak-to-Strong Generalization with On-Policy Reverse Distillation” (KAIST AI, 30 votes on Hugging Face Daily Papers on September 9, 2026) amplifies the verified component of the student gradient in the direction of the teacher's shift, allowing the supervisor to be surpassed with fewer updates (arXiv, code).
«Eliciting Weak-to-Strong Generalization with On-Policy Reverse Distillation» (KAIST AI, 30 Stimmen auf Hugging Face Daily Papers am 9. September 2026) verstärkt die verifizierte Komponente des Studenten-Gradienten in Richtung des Teacher-Shifts und ermöglicht so, den Supervisor mit weniger Updates zu übertreffen (arXiv, Code).
« Eliciting Weak-to-Strong Generalization with On-Policy Reverse Distillation » (KAIST AI, 30 voti su Hugging Face Daily Papers il 9 settembre 2026) amplifica la componente verificata del gradiente dello studente nella direzione dello shift dell'insegnante, permettendo di superare il supervisore con meno aggiornamenti (arXiv, codice).
« Eliciting Weak-to-Strong Generalization with On-Policy Reverse Distillation » (KAIST AI, 30 vot in su Hugging Face Daily Papers el 9 de setember 2026) el amplifica la component verificada del gradient de l'student in la direzzion del shift del professor, cont el permètt de passà sora el supervisù cont manch de aggiornament (arXiv, codes).

Infrastructure RL

RL infrastructure

RL-Infrastruktur

Infrastruttura RL

Infrastructurea RL

Miles v0.1 : le post-training « production-level » passe en open sourceMiles v0.1: “production-level” post-training goes open sourceMiles v0.1: Das «production-level» Post-Training wird Open SourceMiles v0.1: il post-training « production-level » passa in open sourceMiles v0.1: el post-training « production-level » el passa in open source

RadixArk publie Miles v0.1, un système complet de RL post-training à l'échelle frontière : rollouts SGLang, trainer Megatron-LM ou PyTorch FSDP, et une étude de cas de RL agentique asynchrone sur GLM-5.2 744B-A40B avec 64 GPU GB300, temps médian de step de 263 secondes (16 votes, dépôt à 2 681 étoiles) (arXiv, GitHub).
RadixArk releases Miles v0.1, a frontier-scale post-training RL system: SGLang rollouts, Megatron-LM or PyTorch FSDP trainer, and a case study of asynchronous agentic RL on GLM-5.2 744B-A40B with 64 GB300 GPUs, at a 263-second median step time (16 votes, repository at 2,681 stars) (arXiv, GitHub).
RadixArk veröffentlicht Miles v0.1, ein vollständiges System für Post-Training-RL auf Frontier-Skala: SGLang-Rollouts, Trainer mit Megatron-LM oder PyTorch FSDP sowie eine Fallstudie zu asynchronem agentischem RL auf GLM-5.2 744B-A40B mit 64 GB300-GPUs und einer medianen Schrittzeit von 263 Sekunden (16 Stimmen, Repository mit 2'681 Sternen) (arXiv, GitHub).
RadixArk pubblica Miles v0.1, un sistema completo di RL post-training su scala frontiera: rollout SGLang, trainer Megatron-LM o PyTorch FSDP e un caso di studio di RL agentico asincrono su GLM-5.2 744B-A40B con 64 GPU GB300, tempo mediano di step di 263 secondi (16 voti, repository a 2.681 stelle) (arXiv, GitHub).
RadixArk el publica Miles v0.1, on sistema complet de RL post-training a la scala frontiera: rollouts SGLang, trainer Megatron-LM o PyTorch FSDP, e on cas studi de RL agentegh asincron in su GLM-5.2 744B-A40B con 64 GPU GB300, temp median de step de 263 segond (16 vot, deposit a 2.681 stell) (arXiv, GitHub).

Tencent Hunyuan

Tencent Hunyuan

Tencent Hunyuan

Tencent Hunyuan

Tencent Hunyuan

Gander : un agent omni-interactif full-duplex à architecture Cervelet-CerveauGander: a full-duplex omni-interactive agent with a Cerebellum–Brain architectureGander: Ein voll-duplexer omni-interaktiver Agent mit Kleinhirn-Gehirn-ArchitekturGander: un agente omni-interattivo full-duplex ad architettura Cervelletto-CervelloGander: on agent omni-interativ full-duplex a architecura Ceerell-Ceervell

Le rapport technique de Tencent Hunyuan publié le 8 septembre 2026 décrit un modèle qui traite en continu vidéo, parole et texte, permet l'interruption à tout moment, et sépare un Cervelet dédié à l'interaction temps réel d'un Cerveau chargé du raisonnement agentique, le tout sur une architecture Thinker-Talker en flux (arXiv, GitHub).
Tencent Hunyuan's technical report, published on September 8, 2026, describes a model that continuously processes video, speech and text, allows interruption at any moment, and separates a Cerebellum dedicated to real-time interaction from a Brain handling agentic reasoning, all on a streaming Thinker-Talker architecture (arXiv, GitHub).
Das am 8. September 2026 publizierte technische Papier von Tencent Hunyuan beschreibt ein Modell, das Video, Sprache und Text kontinuierlich verarbeitet, jederzeit unterbrechbar ist und ein für Echtzeit-Interaktion zuständiges «Kleinhirn» von einem für agentisches Denken verantwortlichen «Gehirn» trennt – alles auf einer streaming-basierten Thinker-Talker-Architektur (arXiv, GitHub).
Il rapporto tecnico di Tencent Hunyuan pubblicato l'8 settembre 2026 descrive un modello che tratta in continuo video, parlato e testo, permette l'interruzione in ogni momento e separa un Cervelletto dedicato all'interazione in tempo reale da un Cervello incaricato del ragionamento agentico, il tutto su un'architettura Thinker-Talker in streaming (arXiv, GitHub).
El rapport tecnegh de Tencent Hunyuan publicaa l'8 de setember 2026 el descriev on modell che 'l tratta in continu video, parolla e test, el permet l'interruzzion a ogni moment, e 'l separa on Cervell dedicaa a l'interazzion temp reaj de on Ceer chargaa del resonament agentegh, tuscoss sora ona architecura Thinker-Talker in fluss (arXiv, GitHub).

Robotique

Robotics

Robotik

Robotica

Robotega

GE-Act 2.0 : le scaling du world-action model passe de 300 à 30 000 heuresGE-Act 2.0: world-action model scaling goes from 300 to 30,000 hoursGE-Act 2.0: Das Skalieren des World-Action-Modells von 300 auf 30'000 StundenGE-Act 2.0: lo scaling del world-action model passa da 300 a 30.000 oreGE-Act 2.0: el scaling del world-action model el passa de 300 a 30.000 or

AgiBot World rapporte qu'en passant le co-training de 300 à 30 000 heures, le taux de succès de son world-action model grimpe de 17,1 % à 44,1 % sur G1-OP et de 13,4 % à 31,1 % sur G2-90D, sur 100 tâches et 20 groupes de compétences évalués sans fine-tuning (14 votes le 9 septembre 2026) (arXiv).
AgiBot World reports that scaling co-training from 300 to 30,000 hours raises the success rate of its world-action model from 17.1% to 44.1% on G1-OP and from 13.4% to 31.1% on G2-90D, across 100 tasks and 20 skill groups evaluated without fine-tuning (14 votes on September 9, 2026) (arXiv).
AgiBot World berichtet, dass sich durch die Ausweitung des Co-Trainings von 300 auf 30'000 Stunden die Erfolgsquote ihres World-Action-Modells von 17,1 % auf 44,1 % auf G1-OP und von 13,4 % auf 31,1 % auf G2-90D erhöht – bei 100 Aufgaben und 20 ohne Fine-Tuning bewerteten Kompetenzgruppen (14 Stimmen am 9. September 2026) (arXiv).
AgiBot World riporta che passando il co-training da 300 a 30.000 ore, il tasso di successo del suo world-action model sale dal 17,1% al 44,1% su G1-OP e dal 13,4% al 31,1% su G2-90D, su 100 compiti e 20 gruppi di competenze valutati senza fine-tuning (14 voti il 9 settembre 2026) (arXiv).
AgiBot World el rapporta che, passand el co-training de 300 a 30.000 or, el tass de success del so world-action model el monten del 17,1% al 44,1% in su G1-OP e dal 13,4% al 31,1% in su G2-90D, sora 100 task e 20 grupp de competenz valuttaa senza fine-tuning (14 vot el 9 de setember 2026) (arXiv).

Page 4 — Page 4 — Seite 4 — Pagina 4 — Pagina 4 — La Communauté & ÉditoCommunity & EditorialDie Gemeinschaft & LeitartikelLa Comunità & EditorialeLa Comunità & Editòri

IV. Signaux communautéCommunity signalsSignale aus der GemeinschaftSegnali comunitàSegnaj comunità

Anthropic

Anthropic

Anthropic

Anthropic

Anthropic

« I resigned from Anthropic today » : 232 points sur Hacker News“I resigned from Anthropic today”: 232 points on Hacker News«I resigned from Anthropic today»: 232 Punkte auf Hacker News« I resigned from Anthropic today »: 232 punti su Hacker News« I resigned from Anthropic today »: 232 pont in su Hacker News

Une déclaration de démission d'Anthropic, postée le 9 septembre 2026, a déclenché 282 commentaires et 232 points sur Hacker News (source, discussion). Le mouvement personnel devient signal d'époque dans un secteur où les allégeances institutionnelles se recomposent à grande vitesse.
A resignation statement from Anthropic, posted on September 9, 2026, triggered 282 comments and 232 points on Hacker News (source, discussion). Personal moves are becoming a signal of the times in a sector where institutional loyalties are being reshuffled at speed.
Eine am 9. September 2026 gepostete Rücktrittserklärung von Anthropic löste 282 Kommentare und 232 Punkte auf Hacker News aus (Quelle, Diskussion). Die persönliche Geste wird zum Zeitsignal in einer Branche, in der sich institutionelle Loyalitäten mit hoher Geschwindigkeit neu ordnen.
Una dichiarazione di dimissioni da Anthropic, pubblicata il 9 settembre 2026, ha scatenato 282 commenti e 232 punti su Hacker News (fonte, discussione). Il gesto personale diventa segnale d'epoca in un settore dove le fedeltà istituzionali si ricompongono a grande velocità.
Ona declarazzion de dimission de Anthropic, postada el 9 de setember 2026, l'ha faa s'cioppà 282 comment e 232 pont in su Hacker News (sorgent, discussion). El moviment personal el deven segn d'epoca in d'on setor indoa i aleasc istituzionai se ricomposen a gran velocità.

Mathématiques

Mathematics

Mathematik

Matematica

Matematega

Terence Tao : les problèmes ouverts « minés de façon non renouvelable » par l'IATerence Tao: open problems being “mined non-renewably” by AITerence Tao: Offene Probleme werden von der KI «nicht erneuerbar abgebaut»Terence Tao: i problemi aperti « minati in modo non rinnovabile » dall'IATerence Tao: i problema dervii « minaa de manera no rinnovabela » de l'IA

Dans un message relayé le 8 septembre 2026 (257 points, 184 commentaires sur Hacker News), Terence Tao compare l'usage massif des problèmes mathématiques ouverts par l'IA à une extraction minière non renouvelable : chaque problème résolu ou contaminé ne se régénère pas (Mathstodon, discussion). Une mise en garde qui échoit pile la semaine où une solution Navier–Stokes générée par IA est publiée.
In a message relayed on September 8, 2026 (257 points, 184 comments on Hacker News), Terence Tao compares the massive use of open mathematical problems by AI to non-renewable mining: every solved or contaminated problem does not regenerate (Mathstodon, discussion). A warning that lands precisely in the week an AI-generated Navier–Stokes solution is published.
In einer am 8. September 2026 verbreiteten Botschaft (257 Punkte, 184 Kommentare auf Hacker News) vergleicht Terence Tao den massiven Einsatz offener mathematischer Probleme durch KI mit nicht erneuerbarem Bergbau: Jedes gelöste oder kontaminierte Problem regeneriert sich nicht (Mathstodon, Diskussion). Eine Warnung, die ausgerechnet in der Woche erfolgt, in der eine KI-generierte Navier–Stokes-Lösung publiziert wird.
In un messaggio ripreso l'8 settembre 2026 (257 punti, 184 commenti su Hacker News), Terence Tao paragona l'uso massiccio dei problemi matematici aperti da parte dell'IA a un'estrazione mineraria non rinnovabile: ogni problema risolto o contaminato non si rigenera (Mathstodon, discussione). Un avvertimento che arriva proprio la settimana in cui viene pubblicata una soluzione Navier–Stokes generata dall'IA.
In d'on messagg reliaa l'8 de setember 2026 (257 pont, 184 comment in su Hacker News), Terence Tao el confronta l'us massiv di problema matemategh dervii de l'IA cont ona estrazzion minera no rinnovabel: ogni problema resolt o contaminad el se regenera no (Mathstodon, discussion). On'avvertenza che la riva proppi la settemana indoa l'è publicada ona soluzzion Navier–Stokes generada de IA.

Cohere

Cohere

Cohere

Cohere

Cohere

Cartographier où les agents d'IA sont réellement construitsMapping where AI agents are actually being builtKartieren, wo KI-Agenten tatsächlich gebaut werdenMappare dove gli agenti IA vengono realmente costruitiMapà là indoa i agent de IA hinn reallya costruii

Cohere publie le 8 septembre 2026 « Automation's Early Footprint », une analyse en open science de là où les agents sont — et ne sont pas — effectivement déployés dans les entreprises (cohere.com/blog). Un contrepoids factuel au narratif de l'automatisation généralisée, à lire comme un état des lieux chiffré plutôt qu'un manifeste.
On September 8, 2026, Cohere published “Automation's Early Footprint,” an open-science analysis of where agents actually are — and are not — deployed across enterprises (cohere.com/blog). A factual counterweight to the blanket-automation narrative, to be read as a quantified snapshot rather than a manifesto.
Cohere veröffentlicht am 8. September 2026 «Automation's Early Footprint», eine Open-Science-Analyse darüber, wo Agenten in Unternehmen tatsächlich eingesetzt werden – und wo nicht (cohere.com/blog). Ein faktisches Gegengewicht zur Erzählung der flächendeckenden Automatisierung, als zahlenmässiger Bestandesaufnahme zu lesen, nicht als Manifest.
Cohere pubblica l'8 settembre 2026 « Automation's Early Footprint », un'analisi in open science di dove gli agenti sono — e non sono — effettivamente distribuiti nelle imprese (cohere.com/blog). Un contrappeso fattuale al narrativo dell'automatizzazione generalizzata, da leggere come un bilancio quantificato più che un manifesto.
Cohere el publica l'8 de setember 2026 « Automation's Early Footprint », ona analisi in open science de là indoa i agent hinn — e hinn no — efetivament diploiaa in di imprend (cohere.com/blog). On contrapes factuai al narrativ de l'automatizzazzion generalizada, de legg 'me on stat di loeugh ciffraa inveci che on manifest.

Édito

Editorial

Leitartikel

Editoriale

Editòri

La preuve machine entre à la rédactionMachine proof enters the newsroomDer Maschinenbeweis zieht in die Redaktion einLa prova macchina entra in redazioneLa proeuva de machina la riva in redazzion

Quand un labo d'IA annonce avoir résolu Navier–Stokes et joint une preuve Lean vérifiable, le geste change la nature du débat : la question n'est plus « faut-il croire l'IA ? » mais « qui vérifie, et selon quel protocole ? » (annonce). La semaine où Terence Tao alerte sur le minage non renouvelable des problèmes ouverts (Mathstodon), la communauté mathématique découvre que sa ressource la plus rare — les bonnes questions intactes — devient le prochain goulot d'étranglement de la course à la frontière. Les assistants de preuve sont passés du statut d'outil de spécialistes à celui d'infrastructure de confiance ; c'est peut-être l'annonce la plus importante du jour, davantage que le problème lui-même.
When an AI lab announces it has solved Navier–Stokes and attaches a verifiable Lean proof, the gesture changes the nature of the debate: the question is no longer “should we trust the AI?” but “who verifies, and under what protocol?” (announcement). In the same week that Terence Tao warns about the non-renewable mining of open problems (Mathstodon), the mathematical community discovers that its scarcest resource — intact good questions — is becoming the next bottleneck in the frontier race. Proof assistants have moved from specialist tools to trust infrastructure; that may be the most significant announcement of the day, more so than the problem itself.
Wenn ein KI-Labor ankündigt, Navier–Stokes gelöst zu haben, und einen verifizierbaren Lean-Beweis beilegt, verändert diese Geste die Natur der Debatte: Die Frage lautet nicht mehr «Soll man der KI glauben?», sondern «Wer verifiziert, und nach welchem Protokoll?» (Ankündigung). In der Woche, in der Terence Tao vor dem nicht erneuerbaren Abbau offener Probleme warnt (Mathstodon), entdeckt die mathematische Gemeinschaft, dass ihre rareste Ressource – intakte gute Fragen – zum nächsten Engpass des Rennens an die Grenze wird. Beweisassistenten sind vom Spezialistenwerkzeug zur Vertrauensinfrastruktur aufgestiegen; das ist womöglich die wichtigste Nachricht des Tages – mehr noch als das Problem selbst.
Quando un laboratorio di IA annuncia di aver risolto Navier–Stokes e allega una dimostrazione Lean verificabile, il gesto cambia la natura del dibattito: la domanda non è più « bisogna credere all'IA? » ma « chi verifica, e secondo quale protocollo? » (annuncio). La settimana in cui Terence Tao avverte del mining non rinnovabile dei problemi aperti (Mathstodon), la comunità matematica scopre che la sua risorsa più rara — le buone domande intatte — diventa il prossimo collo di bottiglia della corsa alla frontiera. Gli assistenti di dimostrazione sono passati dallo status di strumento per specialisti a quello di infrastruttura di fiducia; è forse l'annuncio più importante della giornata, più del problema stesso.
Quand on labo d'IA l'anunzia d'avè resolt Navier–Stokes e 'l gionta ona proeuva Lean verificabel, el gest el cambia la natura del dibattit: la question l'è pu « se gh'è de cred a l'IA » ma « chi el verifica, e segond che protocoll » (anunzi). La settemana indoa Terence Tao el dà l'alarm sora el minad no rinnovabel di problema dervii (Mathstodon), la comunità matematega la descovr che la soa risorsa pussee rara — i bon question intatt — la diven el prosem goei d'strangolament de la corsa a la frontiera. I assistent de proeuva hinn passaa del stat de stroment di specialist a quell d'infrastrutura de fiducia; l'è fors l'anunzi pussee important del dì, pussee del problema midemm.