The Neuron Times

All the AI that's fit to print

N° 245 Édition du matinMorning EditionMorgenausgabeEdizione del mattinoEdizion del mattin · Genève MERCREDI 2 SEPTEMBRE 2026WEDNESDAY, 2 SEPTEMBER 2026MITTWOCH, 2. SEPTEMBER 2026MERCOLEDÌ 2 SETTEMBRE 2026MERCOLEDÌ 2 SETTEMBRE 2026

À la Une · ModèlesFront Page · ModelsSchlagzeilen · ModellePrima pagina · ModelliIn prima pagina · Modej

Anthropic lance Claude Fable 5.1 et Claude Mythos 5.1, ses nouveaux modèles frontière pour le code et la rechercheAnthropic launches Claude Fable 5.1 and Claude Mythos 5.1, its new frontier models for code and researchAnthropic lanciert Claude Fable 5.1 und Claude Mythos 5.1, seine neuen Frontier-Modelle für Code und RechercheAnthropic lancia Claude Fable 5.1 e Claude Mythos 5.1, i suoi nuovi modelli di frontiera per il codice e la ricercaAnthropic el lansa Claude Fable 5.1 e Claude Mythos 5.1, i sò noeuv modej frontiera per el codes e la ricerca

La double release du 1er septembre 2026, saluée pour le code et le travail de connaissance, promet aussi « un aperçu précoce » de la contribution des modèles au progrès scientifique.The dual release of 1 September 2026, praised for code and knowledge work, also promises "an early glimpse" of models' contribution to scientific progress.Die Doppelveröffentlichung vom 1. September 2026, gelobt für Code und Wissensarbeit, verspricht zudem «einen frühen Einblick» in den Beitrag der Modelle zum wissenschaftlichen Fortschritt.Il doppio rilascio del 1° settembre 2026, salutato per il codice e il lavoro della conoscenza, promette anche «un'anticipazione» del contributo dei modelli al progresso scientifico.La doble release del primm de setember 2026, aprezada per el codes e el laurà de conoscenza, la promett anca « ona vista in anticip » del contribut di modej al progress scentifegh.

Anthropic a annoncé le 1er septembre 2026 la sortie de deux nouveaux modèles jumeaux, Claude Fable 5.1 et Claude Mythos 5.1, présentés comme « nos modèles les plus avancés pour le code et le travail de la connaissance ». Le laboratoire précise que leurs capacités de recherche offrent un aperçu précoce de la contribution des modèles d'IA au progrès scientifique, selon l'annonce publiée sur la newsroom officielle d'Anthropic.On Tuesday, 1 September 2026, Anthropic announced the release of two new twin models, Claude Fable 5.1 and Claude Mythos 5.1, presented as "our most advanced models for code and knowledge work". The lab notes that their research capabilities offer an early glimpse into how AI models may contribute to scientific progress, according to the announcement published on Anthropic's official newsroom.Anthropic kündigte am 1. September 2026 die Veröffentlichung zweier neuer Zwillingsmodelle an, Claude Fable 5.1 und Claude Mythos 5.1, die als «unsere fortschrittlichsten Modelle für Code und Wissensarbeit» präsentiert werden. Das Labor präzisiert, dass ihre Recherche-Fähigkeiten einen frühen Einblick in den Beitrag von KI-Modellen zum wissenschaftlichen Fortschritt bieten, gemäss der Ankündigung auf der offiziellen Newsroom von Anthropic.Anthropic ha annunciato il 1° settembre 2026 il rilascio di due nuovi modelli gemelli, Claude Fable 5.1 e Claude Mythos 5.1, presentati come «i nostri modelli più avanzati per il codice e il lavoro della conoscenza». Il laboratorio precisa che le loro capacità di ricerca offrono un'anticipazione del contributo dei modelli di IA al progresso scientifico, secondo l'annuncio pubblicato su la newsroom ufficiale di Anthropic.Anthropic l'ha anunziaa el primm de setember 2026 l'essida de duu modej zemell, Claude Fable 5.1 e Claude Mythos 5.1, presentaa comè « i noster modej pussee avanzaa per el codes e el laurà de la conoscenza ». El laboratori el precisa che i sò capacitaa de ricerca dann ona vista in anticip del contribut di modej de IA al progress scentifegh, segond l'anunzi publicaa sora la newsroom ofizziala d'Anthropic.

L'annonce a immédiatement déclenché un débat d'ampleur : la discussion « Claude Fable 5.1 and Claude Mythos 5.1 » sur Hacker News a atteint 1 042 points et 970 commentaires en moins de douze heures, les participants décortiquant la documentation « What's new in Fable 5.1 » et la System Card publiées par le laboratoire.The announcement immediately triggered a wide-ranging debate: the "Claude Fable 5.1 and Claude Mythos 5.1" discussion on Hacker News reached 1,042 points and 970 comments in under twelve hours, with participants dissecting the "What's new in Fable 5.1" documentation and the System Card published by the lab.Die Ankündigung löste umgehend eine Debatte von grossem Ausmass aus: Die Diskussion «Claude Fable 5.1 and Claude Mythos 5.1» auf Hacker News erreichte in weniger als zwölf Stunden 1042 Punkte und 970 Kommentare, wobei die Teilnehmenden die Dokumentation «What's new in Fable 5.1» und die vom Labor veröffentlichte System Card sezierten.L'annuncio ha subito innescato un dibattito di ampia portata: la discussione « Claude Fable 5.1 and Claude Mythos 5.1 » su Hacker News ha raggiunto 1 042 punti e 970 commenti in meno di dodici ore, con i partecipanti che analizzano minuziosamente la documentazione « What's new in Fable 5.1 » e la System Card pubblicate dal laboratorio.L'anunzi l'ha subet faa sù on dibattit de grandi proporzion: la discussion « Claude Fable 5.1 and Claude Mythos 5.1 » sora Hacker News l'ha razonzaa 1.042 pont e 970 comment in manch de dodes or, cont i partecipant che descomponeven la documentazzion « What's new in Fable 5.1 » e la System Card publicaa del laboratori.

Page 1 — Page 1 — Seite 1 — Pagina 1 — Pagina 1 — À la Une — Modèles & FrontièreFront Page — Models & FrontierTitelseite — Modelle & GrenzenIn Prima Pagina — Modelli & FrontieraIn Prima Paggina — Modej & Frontiera

I. Modèles & FrontièreModels & FrontierModelle & GrenzenModelli & FrontieraModej & Frontiera

Google DeepMind

Google DeepMind

Google DeepMind

Google DeepMind

Google DeepMind

Gemini passe à la compréhension vidéo agentiqueGemini moves to agentic video understandingGemini wendet sich der agentischen Videoanalyse zuGemini passa alla comprensione video agenticaGemini el passa a la comprension video agentiga

Google DeepMind a annoncé le 1er septembre 2026 le lancement de la compréhension vidéo agentique « à travers nos derniers modèles Gemini », promettant une meilleure exactitude et une réduction des coûts et de la consommation de tokens, selon l'annonce officielle. Le modèle ne se contente plus d'analyser passivement une vidéo : il peut la décomposer et interroger son contenu de façon itérative.
Google DeepMind announced on Tuesday, 1 September 2026 the launch of agentic video understanding "across our latest Gemini models", promising improved accuracy along with reduced costs and token consumption, according to the official announcement. The model no longer passively analyses a video: it can break it down and iteratively query its content.
Google DeepMind kündigte am 1. September 2026 die Einführung agentischer Videoanalyse «mit unseren neuesten Gemini-Modellen» an und verspricht bessere Genauigkeit sowie geringere Kosten und Token-Verbrauch, gemäss der offiziellen Ankündigung. Das Modell analysiert ein Video nicht mehr nur passiv: Es kann es zerlegen und seinen Inhalt iterativ abfragen.
Google DeepMind ha annunciato il 1° settembre 2026 il lancio della comprensione video agentica «attraverso i nostri ultimi modelli Gemini», promettendo una migliore accuratezza e una riduzione dei costi e del consumo di token, secondo l'annuncio ufficiale. Il modello non si limita più ad analizzare passivamente un video: può scomporlo e interrogarne il contenuto in modo iterativo.
Google DeepMind l'ha anunziaa el primm de setember 2026 el lanzament de la comprension video agentiga « cont i noster ultem modej Gemini », cont la promessa de ona mej precision e de ona riduzzion di cost e del consum de tokens, segond l'anunzi ofizzial. El modej el analizza pu domà passivament on video: el pò descomponnel e interrogà el sò contegnuu de manera iterativa.

OpenAI · Sûreté

OpenAI · Safety

OpenAI · Sicherheit

OpenAI · Sicurezza

OpenAI · Sùretaa

Path to Astra : OpenAI franchit un seuil critique en cybersécuritéPath to Astra: OpenAI crosses a critical cybersecurity thresholdPath to Astra: OpenAI überschreitet eine kritische Schwelle in der CybersicherheitPath to Astra: OpenAI supera una soglia critica in cybersicurezzaPath to Astra: OpenAI el passa ona soglia crìtega in cybersecurity

Premier modèle OpenAI à franchir le seuil « Critique » de capacité en cybersécurité du Preparedness Framework, Astra s'accompagne de garde-fous renforcés pour sa mise en production, détaille l'analyse « Path to Astra » publiée le 1er septembre 2026. La discussion sur Hacker News (110 points, 50 commentaires) s'est concentrée sur la méthodologie d'évaluation des capacités offensives.
The first OpenAI model to cross the "Critical" cybersecurity capability threshold of the Preparedness Framework, Astra ships with reinforced guardrails for its production rollout, as detailed in the "Path to Astra" analysis published on Tuesday, 1 September 2026. The Hacker News discussion (110 points, 50 comments) focused on the methodology for evaluating offensive capabilities.
Als erstes OpenAI-Modell überschreitet Astra die Schwelle «Kritisch» der Cybersecurity-Fähigkeiten des Preparedness Framework und wird von verstärkten Schutzmassnahmen für den Produktionseinsatz begleitet, wie die am 1. September 2026 veröffentlichte Analyse «Path to Astra» ausführt. Die Diskussion auf Hacker News (110 Punkte, 50 Kommentare) konzentrierte sich auf die Methodik der Bewertung offensiver Fähigkeiten.
Primo modello OpenAI a superare la soglia «Critica» di capacità in cybersicurezza del Preparedness Framework, Astra è accompagnato da garanzie rafforzate per il suo rilascio in produzione, come dettaglia l'analisi «Path to Astra» pubblicata il 1° settembre 2026. La discussione su Hacker News (110 punti, 50 commenti) si è concentrata sulla metodologia di valutazione delle capacità offensive.
Primm modej OpenAI a passà la soglia « Crìtega » de capacità in cybersecurity del Preparedness Framework, Astra el gh'ha adree di salvaguard renforzaa per el sò despiegh in produzzion, comè el detaja l'analisi « Path to Astra » publicada el primm de setember 2026. La discussion sora Hacker News (110 pont, 50 comment) la s'è concentrada sora la metodologia de valutazzion di capacità offensiv.

xAI

xAI

xAI

xAI

xAI

xAI publie « Biosecurity at the frontier »xAI publishes "Biosecurity at the frontier"xAI veröffentlicht «Biosecurity at the frontier»xAI pubblica «Biosecurity at the frontier»xAI el publica « Biosecurity at the frontier »

xAI a publié le 1er septembre 2026 un billet intitulé « Biosecurity at the frontier », portant sa communication sur les mesures de biosécurité appliquées à Grok. Une convergence notable : OpenAI évoquait le même jour ses seuils critiques dans un autre domaine de risque, celui de la cybersécurité.
xAI published on Tuesday, 1 September 2026 a post entitled "Biosecurity at the frontier", centring its communications on the biosecurity measures applied to Grok. A notable convergence: that same day, OpenAI was discussing its critical thresholds in another risk domain, cybersecurity.
xAI veröffentlichte am 1. September 2026 einen Beitrag mit dem Titel «Biosecurity at the frontier», in dem das Unternehmen seine Kommunikation auf die bei Grok angewandten Biosicherheitsmassnahmen ausrichtete. Eine bemerkenswerte Konvergenz: OpenAI erwähnte am selben Tag seine kritischen Schwellen in einem anderen Risikobereich, jenem der Cybersicherheit.
xAI ha pubblicato il 1° settembre 2026 un post intitolato «Biosecurity at the frontier», incentrando la propria comunicazione sulle misure di biosicurezza applicate a Grok. Una convergenza notevole: OpenAI evocava lo stesso giorno le proprie soglie critiche in un altro dominio di rischio, quello della cybersicurezza.
xAI l'ha publicaa el primm de setember 2026 on post intitulaa « Biosecurity at the frontier », centrand la soa comunicazzion sora i mesur de biosicurezza aplicaa a Grok. Convergenza notabil: OpenAI el tratava el midemm dì i soeu soglij crìtegh in d'on alter domini de ris'c, quell de la cybersecurity.

II. Produits & déploiementProducts & deploymentProdukte & EinsatzProdotti & diffusioneProdott & despiegh

Google Workspace

Google Workspace

Google Workspace

Google Workspace

Google Workspace

Google Pics arrive dans Workspace, propulsé par Nano BananaGoogle Pics arrives in Workspace, powered by Nano BananaGoogle Pics kommt in Workspace, angetrieben von Nano BananaGoogle Pics arriva in Workspace, spinto da Nano BananaGoogle Pics el riva in Workspace, spingiuu de Nano Banana

Bâti sur le dernier modèle Nano Banana, Google Pics — l'outil de création et d'édition d'images de Workspace — est désormais disponible, annonce Google le 1er septembre 2026 dans son billet de lancement. Le même jour, l'entreprise publiait sa revue des annonces IA d'août 2026.
Built on the latest Nano Banana model, Google Pics — Workspace's image creation and editing tool — is now available, Google announced on Tuesday, 1 September 2026 in its launch post. The same day, the company published its review of AI announcements for August 2026.
Aufgebaut auf dem neuesten Nano-Banana-Modell ist Google Pics — das Werkzeug zur Bilderstellung und -bearbeitung in Workspace — ab sofort verfügbar, kündigt Google am 1. September 2026 in seinem Launch-Beitrag an. Am selben Tag veröffentlichte das Unternehmen seine Zusammenfassung der KI-Ankündigungen vom August 2026.
Basato sull'ultimo modello Nano Banana, Google Pics — lo strumento di creazione e modifica di immagini di Workspace — è ormai disponibile, annuncia Google il 1° settembre 2026 nel suo post di lancio. Lo stesso giorno, l'azienda pubblicava la sua rassegna degli annunci IA di agosto 2026.
Faa su sora l'ultem modej Nano Banana, Google Pics — l'strument de creazzion e modifega de imagin de Workspace — l'è adess disponibij, comè el nunzia Google el primm de setember 2026 in del so post de lanzament. El midemm dì, l'azienda la publicava la soa rassegna di anunzi IA de avost 2026.

OpenAI · Santé

OpenAI · Health

OpenAI · Gesundheit

OpenAI · Salute

OpenAI · Salut

ChatGPT se connecte aux dossiers de santé électroniquesChatGPT connects to electronic health recordsChatGPT verbindet sich mit elektronischen GesundheitsaktenChatGPT si collega alle cartelle cliniche elettronicheChatGPT el se conliga ai cartellett sanitari eletronegh

Les organisations de santé peuvent désormais connecter des dossiers médicaux électroniques (EHR) et des données sectorielles à ChatGPT, annonçait OpenAI le 1er septembre 2026, permettant aux cliniciens d'accéder de façon sécurisée au contexte patient (annonce). Dans la foulée, les notes de version de ChatGPT du 1er septembre détaillent un plugin « Healthcare Public Data » regroupant neuf applications de recherche en sources publiques, en lecture seule, sans accès aux dossiers patients.
Healthcare organisations can now connect electronic health records (EHRs) and sector data to ChatGPT, OpenAI announced on Tuesday, 1 September 2026, allowing clinicians to securely access patient context (announcement). Following this, the ChatGPT release notes of 1 September detail a "Healthcare Public Data" plugin bundling nine research applications over public sources, in read-only mode, with no access to patient records.
Gesundheitsorganisationen können ab sofort elektronische Patientenakten (EHR) und Branchendaten mit ChatGPT verbinden, kündigte OpenAI am 1. September 2026 an, womit Klinikerinnen und Kliniker sicher auf den Patientenkontext zugreifen können (Ankündigung). Im Anschluss beschreiben die Versionshinweise von ChatGPT vom 1. September ein Plugin «Healthcare Public Data», das neun Recherche-Anwendungen auf öffentliche Quellen bündelt — nur lesend, ohne Zugriff auf Patientenakten.
Le organizzazioni sanitarie possono ormai collegare cartelle cliniche elettroniche (EHR) e dati settoriali a ChatGPT, annunciava OpenAI il 1° settembre 2026, permettendo ai clinici di accedere in modo sicuro al contesto del paziente (annuncio). A seguire, le note di versione di ChatGPT del 1° settembre dettagliano un plugin «Healthcare Public Data» che raggruppa nove applicazioni di ricerca su fonti pubbliche, in sola lettura, senza accesso alle cartelle dei pazienti.
I organizazzion sanitari poden adess conligà di cartellett medegh eletronegh (EHR) e di dacc del settor a ChatGPT, comè l'anunziava OpenAI el primm de setember 2026, permettand ai clenegh de agh acced de manera segura al contest del pazient (anunzi). In del seguit, i not de version de ChatGPT del primm de setember detajen on plugin « Healthcare Public Data » che 'l mett insema nœu applicazzion de ricerca in di font publegh, in domà letura, senza access ai cartellett di pazient.

Page 2 — Page 2 — Seite 2 — Pagina 2 — Pagina 2 — Le Cahier TechniqueThe Technical SectionDas Technik-DossierIl Quaderno TecnicoEl Quadernegh Tecegh

III. Infra & outillageInfrastructure & toolingInfrastruktur & WerkzeugeInfra & strumentiInfra & strumentazzion

Inférence navigateur

Browser inference

Browser-Inferenz

Inferenza nel browser

Inferenzi browser

Hugging Face libère 200+ kernels WebGPU pour l'IA localeHugging Face open-sources 200+ WebGPU kernels for local AIHugging Face veröffentlicht über 200 WebGPU-Kernels für lokale KIHugging Face libera 200+ kernel WebGPU per l'IA localeHugging Face el libera 200+ kernels WebGPU per l'IA local

Hugging Face a publié le 1er septembre 2026 @huggingface/kernels, une bibliothèque de plus de 200 kernels WebGPU prêts à l'emploi pour faire tourner des modèles directement dans le navigateur et en local. L'initiative abaisse la barrière d'entrée de l'inférence locale côté client, historiquement freinée par la fragmentation des backends graphiques.
Hugging Face released on Tuesday, 1 September 2026 @huggingface/kernels, a library of more than 200 ready-to-use WebGPU kernels for running models directly in the browser and on-device. The initiative lowers the barrier to client-side local inference, historically held back by the fragmentation of graphics backends.
Hugging Face veröffentlichte am 1. September 2026 @huggingface/kernels, eine Bibliothek mit über 200 einsatzbereiten WebGPU-Kernels, um Modelle direkt im Browser und lokal laufen zu lassen. Die Initiative senkt die Einstiegshürde für clientseitige lokale Inferenz, die historisch durch die Fragmentierung der Grafik-Backends gebremst wurde.
Hugging Face ha pubblicato il 1° settembre 2026 @huggingface/kernels, una libreria di oltre 200 kernel WebGPU pronti all'uso per far girare modelli direttamente nel browser e in locale. L'iniziativa abbassa la barriera d'ingresso all'inferenza locale lato client, storicamente frenata dalla frammentazione dei backend grafici.
Hugging Face l'ha publicaa el primm de setember 2026 @huggingface/kernels, ona biblioteca de pussee de 200 kernels WebGPU pront per fà girà di modej direttamente in del browser e in local. L'inizziativa la sbassa la barriera de ingress per l'inferenzi local lato client, storegada storegament de la frammentazzion di backend grafegh.

Inférence locale

Local inference

Lokale Inferenz

Inferenza locale

Inferenzi local

slotstream : un modèle de 104 Go sur un Mac de 48 Go à 12 tokens/sslotstream: a 104 GB model on a 48 GB Mac at 12 tokens/sslotstream: ein 104-GB-Modell auf einem 48-GB-Mac mit 12 Tokens/sslotstream: un modello da 104 GB su un Mac da 48 GB a 12 token/sslotstream: on modej de 104 GB sora on Mac de 48 GB a 12 tokens/s

Un développeur a publié le 1er septembre 2026 slotstream, un outil Mac natif (MLX, Swift) qui fait tourner Qwen3.8-Flash-Next en 4 bits — 125 milliards de paramètres, 104 Go de poids — sur un Mac de 48 Go à environ 12 tokens/seconde, grâce à l'expert-offloading et au streaming SSD, à partir de 16 Go de mémoire, selon le dépôt GitHub. La discussion Show HN (181 points, 90 commentaires) a débattu du compromis mémoire/vitesse de l'auto-mode ; le portage du MTP pour le décodage spéculatif est annoncé comme prochaine étape.
A developer released on Tuesday, 1 September 2026 slotstream, a native Mac tool (MLX, Swift) that runs Qwen3.8-Flash-Next in 4-bit — 125 billion parameters, 104 GB of weights — on a 48 GB Mac at roughly 12 tokens per second, thanks to expert-offloading and SSD streaming, starting from 16 GB of memory, according to the GitHub repository. The Show HN discussion (181 points, 90 comments) debated the memory/speed trade-off of auto-mode; a port of MTP for speculative decoding is announced as the next step.
Ein Entwickler veröffentlichte am 1. September 2026 slotstream, ein Mac-natives Werkzeug (MLX, Swift), das Qwen3.8-Flash-Next in 4 Bit ausführt — 125 Milliarden Parameter, 104 GB Gewichte — auf einem Mac mit 48 GB Arbeitsspeicher mit rund 12 Tokens pro Sekunde, dank Expert-Offloading und SSD-Streaming, ab 16 GB Speicher, gemäss dem GitHub-Repository. Die Diskussion Show HN (181 Punkte, 90 Kommentare) diskutierte den Speicher-/Geschwindigkeits-Kompromiss des Auto-Modus; die Portierung des MTP für spekulative Dekodierung ist als nächster Schritt angekündigt.
Uno sviluppatore ha pubblicato il 1° settembre 2026 slotstream, uno strumento nativo per Mac (MLX, Swift) che fa girare Qwen3.8-Flash-Next in 4 bit — 125 miliardi di parametri, 104 GB di pesi — su un Mac da 48 GB a circa 12 token al secondo, grazie all'expert-offloading e allo streaming su SSD, a partire da 16 GB di memoria, secondo il repository GitHub. La discussione Show HN (181 punti, 90 commenti) ha dibattuto il compromesso memoria/velocità dell'auto-mode; il porting dell'MTP per il decoding speculativo è annunciato come prossimo passo.
On desvilupador l'ha publicaa el primm de setember 2026 slotstream, on strument nativ Mac (MLX, Swift) che 'l fa girà Qwen3.8-Flash-Next a 4 bit — 125 miliard de parameter, 104 GB de pes — sora on Mac de 48 GB a circa 12 tokens al segond, grazie a l'expert-offloading e al streaming SSD, a partì de 16 GB de memoria, segond el repository GitHub. La discussion Show HN (181 pont, 90 comment) l'ha dibattuu el compromess memoria/velocità de l'auto-mode; el port del MTP per el decodifegh speculativ l'è anunziaa comè prossim pass.

Agentique

Agentic

Agentisch

Agentico

Agentigh

L'app ChatGPT/Codex embarque une copie complète de LibreOfficeThe ChatGPT/Codex app bundles a full copy of LibreOfficeDie ChatGPT/Codex-App bringt eine vollständige Kopie von LibreOffice mitL'app ChatGPT/Codex incorpora una copia completa di LibreOfficeL'app ChatGPT/Codex la porta adree ona copia completa de LibreOffice

Simon Willison a constaté le 1er septembre 2026 que l'application ChatGPT/Codex embarque une copie complète de LibreOffice, confirmée par la discussion Hacker News (308 points, 143 commentaires). Les agents y trouvent un moteur bureautique complet pour manipuler documents et classeurs — un choix architectural qui en dit long sur la matérialisation des outils agentiques.
Simon Willison found on Tuesday, 1 September 2026 that the ChatGPT/Codex app bundles a complete copy of LibreOffice, confirmed by the Hacker News discussion (308 points, 143 comments). Agents gain a full office suite engine to manipulate documents and spreadsheets — an architectural choice that says a great deal about the materialisation of agentic tooling.
Simon Willison stellte am 1. September 2026 fest, dass die Anwendung ChatGPT/Codex eine vollständige Kopie von LibreOffice mitbringt, bestätigt durch die Diskussion auf Hacker News (308 Punkte, 143 Kommentare). Die Agenten finden darin eine vollständige Office-Engine, um Dokumente und Tabellen zu bearbeiten — eine architektonische Entscheidung, die viel über die Materialisierung agentischer Werkzeuge aussagt.
Simon Willison ha constatato il 1° settembre 2026 che l'applicazione ChatGPT/Codex incorpora una copia completa di LibreOffice, confermato dalla discussione Hacker News (308 punti, 143 commenti). Gli agenti vi trovano un motore office completo per manipolare documenti e fogli di calcolo — una scelta architetturale che la dice lunga sulla materializzazione degli strumenti agentici.
Simon Willison l'ha constataa el primm de setember 2026 che l'applicazzion ChatGPT/Codex la gh'ha denter ona copia completa de LibreOffice, confermada de la discussion Hacker News (308 pont, 143 comment). I agent chì troeuven on motor de biofissa complet per manipolà document e foeuj de calcol — ona scerna architettoniga che la dis propi tant sora la materializazzion di strument agentigh.

Analyse

Analysis

Analyse

Analisi

Analisi

La frontière efficiente de l'inférence LLM, cartographiée par BasetenThe efficient frontier of LLM inference, mapped by BasetenDie effiziente Grenze der LLM-Inferenz, kartiert von BasetenLa frontiera efficiente dell'inferenza LLM, mappata da BasetenLa frontiera effiçienta de l'inferenzi LLM, mappada de Baseten

Baseten publie le 1er septembre 2026 une analyse de la « frontière efficiente » de l'inférence LLM, relayée par une discussion Hacker News (72 points). Le texte cartographie les compromis débit/latence/coût entre configurations d'inférence, utile dans un contexte où les optimisations locales et cloud se multiplient.
Baseten published on Tuesday, 1 September 2026 an analysis of the "efficient frontier" of LLM inference, relayed by a Hacker News discussion (72 points). The piece maps throughput/latency/cost trade-offs across inference configurations, useful in a context where local and cloud optimisations proliferate.
Baseten veröffentlicht am 1. September 2026 eine Analyse der «effizienten Grenze» der LLM-Inferenz, aufgegriffen von einer Diskussion auf Hacker News (72 Punkte). Der Text kartiert die Kompromisse zwischen Durchsatz, Latenz und Kosten über Inferenz-Konfigurationen hinweg — nützlich in einem Umfeld, in dem sich lokale und Cloud-Optimierungen häufen.
Baseten pubblica il 1° settembre 2026 un'analisi della «frontiera efficiente» dell'inferenza LLM, ripresa da una discussione Hacker News (72 punti). Il testo mappa i compromessi throughput/latenza/costo tra configurazioni d'inferenza, utile in un contesto in cui le ottimizzazioni locali e cloud si moltiplicano.
Baseten el publica el primm de setember 2026 ona analisi de la « frontiera effiçienta » de l'inferenzi LLM, relexada de ona discussion Hacker News (72 pont). El test el mapa i compromess debit/latenza/cost tra i configurazzion de inferenzi, util in d'on contest indova che i otimizazzion localegh e cloud se moltiplegen.

Page 3 — Page 3 — Seite 3 — Pagina 3 — Pagina 3 — La RechercheThe Research SectionDie ForschungLa RicercaLa Ricerca

IV. Papers du jourPapers of the dayPapers des TagesPaper del giornoPaper del dì

Architecture

Architecture

Architektur

Architettura

Architetura

SMELT : boucler les couches médianes d'un MoE économise jusqu'à 18 % de FLOPs d'entraînementSMELT: looping a MoE's middle layers saves up to 18% of training FLOPsSMELT: Das Durchlaufen der mittleren Schichten eines MoE spart bis zu 18 Prozent Trainings-FLOPsSMELT: chiudere in loop gli strati mediani di un MoE fa risparmiare fino al 18 % di FLOP di addestramentoSMELT: sarà su i facc median d'on MoE el fa risparmià fin al 18% de FLOPs de addestrament

Des chercheurs de ByteDance Seed proposent SMELT (15 upvotes sur HF Daily Papers le 2 septembre 2026) : boucler deux fois la moitié médiane des couches d'un Transformer MoE, à budget égal en FLOPs par token, paramètres et cache KV, économise 6,8 à 18,0 % de FLOPs d'entraînement sur la frontière compute-optimale, avec des gains les plus marqués en code et croissants avec la longueur de contexte. La recette est validée jusqu'à 54 milliards de paramètres hors embeddings.
Researchers at ByteDance Seed propose SMELT (15 upvotes on HF Daily Papers on Wednesday, 2 September 2026): looping the middle half of a MoE Transformer's layers twice, at equal budget in FLOPs per token, parameters and KV cache, saves 6.8 to 18.0 per cent of training FLOPs on the compute-optimal frontier, with the largest gains on code and growing ones with context length. The recipe is validated up to 54 billion parameters excluding embeddings.
Forschende von ByteDance Seed schlagen SMELT vor (15 Upvotes auf HF Daily Papers am 2. September 2026): Das zweifache Durchlaufen der mittleren Hälfte der Schichten eines MoE-Transformers spart — bei gleichem Budget an FLOPs pro Token, Parametern und KV-Cache — 6,8 bis 18,0 Prozent Trainings-FLOPs auf der compute-optimalen Grenze, mit den grössten Gewinnen bei Code und zunehmend mit der Kontextlänge. Das Rezept ist bis 54 Milliarden Parameter ohne Embeddings validiert.
Ricercatori di ByteDance Seed propongono SMELT (15 upvote su HF Daily Papers il 2 settembre 2026): eseguire due volte in loop la metà mediana degli strati di un Transformer MoE, a parità di budget in FLOP per token, parametri e cache KV, fa risparmiare dal 6,8 al 18,0 % di FLOP di addestramento sulla frontiera compute-ottimale, con i guadagni più marcati nel codice e crescenti con la lunghezza del contesto. La ricetta è validata fino a 54 miliardi di parametri esclusi gli embedding.
Di ricercador de ByteDance Seed proponen SMELT (15 upvote sora HF Daily Papers el 2 de setember 2026): fà vegnì serial duu voeult la mitaa mediana di facc d'on Transformer MoE, con budget compagn in FLOPs per token, parameter e cache KV, el fa risparmià del 6,8 al 18,0% de FLOPs de addestrament sora la frontiera compute-ottemala, cont i guadagn pussee fort in codes e crescent con la longhezza del contest. La ricetta l'è validada infina a 54 miliard de parameter foeura di embedding.

Benchmarks

Benchmarks

Benchmarks

Benchmark

Benchmark

E-Commerce Bench : sur un an de commerce simulé, aucun modèle ne domine tous les autresE-Commerce Bench: over a simulated year of commerce, no model dominates all othersE-Commerce Bench: Über ein Jahr simulierten Handels dominiert kein Modell alle anderenE-Commerce Bench: su un anno di commercio simulato, nessun modello domina tutti gli altriE-Commerce Bench: sora on ann de commerzi simulaa, nissun modej el domina tucc i alter

Le benchmark E-Commerce Bench, ouvert par l'équipe Qwen, simule une année complète (365 jours) de gestion de boutiques en ligne avec négociations fournisseurs et chocs d'offre. Sur les 18 modèles frontière évalués, aucun ne domine : GPT-5.6 Sol transforme un capital initial de 100 000 en 1 431 425, mais se classe 16e sur 18 en détection de fraude ; côté poids ouverts, Qwen3.8-Max-Preview mène avec 416 252, soit 38 % de plus que GLM 5.2 (high).
The E-Commerce Bench benchmark, released by the Qwen team, simulates a full year (365 days) of online shop management with supplier negotiations and supply shocks. Of the 18 frontier models evaluated, none dominates: GPT-5.6 Sol turns an initial capital of 100,000 into 1,431,425, but ranks 16th of 18 in fraud detection; among open weights, Qwen3.8-Max-Preview leads with 416,252, 38 per cent ahead of GLM 5.2 (high).
Der Benchmark E-Commerce Bench, vom Qwen-Team eröffnet, simuliert ein ganzes Jahr (365 Tage) der Führung von Online-Shops mit Lieferantenverhandlungen und Angebotsschocks. Von den 18 bewerteten Frontier-Modellen dominiert keines: GPT-5.6 Sol verwandelt ein Startkapital von 100 000 in 1 431 425, belegt aber bei der Betrugserkennung nur Rang 16 von 18; bei den offenen Gewichten führt Qwen3.8-Max-Preview mit 416 252 — 38 Prozent mehr als GLM 5.2 (high).
Il benchmark E-Commerce Bench, reso pubblico dal team Qwen, simula un anno intero (365 giorni) di gestione di negozi online con negoziazioni con i fornitori e choc dell'offerta. Su 18 modelli di frontiera valutati, nessuno domina: GPT-5.6 Sol trasforma un capitale iniziale di 100 000 in 1 431 425, ma si classifica 16° su 18 nel rilevamento delle frodi; sul fronte dei pesi aperti, Qwen3.8-Max-Preview guida con 416 252, ovvero il 38 % in più di GLM 5.2 (high).
El benchmark E-Commerce Bench, dervid de l'equip Qwen, el simula on ann complet (365 dì) de gestion de bottegh in linia con negoziazzion con i fornitór e sicc de offerta. Tra i 18 modej frontiera valutaa, nissun el domina: GPT-5.6 Sol el transforma on capital inizzial de 100.000 in 1.431.425, ma el se classifega 16esim su 18 in rilevament di frott; de la banda di pes dervii, Qwen3.8-Max-Preview el guida con 416.252, el 38% pussee de GLM 5.2 (high).

Agents GUI

GUI agents

GUI-Agenten

Agenti GUI

Agent GUI

UI-Venus-2 : un agent GUI open source pour 170+ apps mobiles et le desktopUI-Venus-2: an open-source GUI agent for 170+ mobile apps and desktopUI-Venus-2: ein quelloffener GUI-Agent für über 170 mobile Apps und den DesktopUI-Venus-2: un agente GUI open source per 170+ app mobili e il desktopUI-Venus-2: on agent GUI open source per 170+ app mobil e el desktop

Ant Group publie le rapport technique d'UI-Venus-2 (39 upvotes, papier le plus voté du jour sur HF Daily Papers le 2 septembre 2026), un agent GUI multimodal open source couvrant plus de 170 applications mobiles multilingues et les systèmes desktop natifs, avec vérification multi-modèles pour des signaux RL fiables et des mécanismes de sécurité pour les actions à conséquences.
Ant Group published the technical report for UI-Venus-2 (39 upvotes, the most upvoted paper of the day on HF Daily Papers on Wednesday, 2 September 2026), an open-source multimodal GUI agent covering more than 170 multilingual mobile applications and native desktop systems, with multi-model verification for reliable RL signals and safety mechanisms for consequential actions.
Ant Group veröffentlicht das technische Bericht zu UI-Venus-2 (39 Upvotes, meistgewähltes Paper des Tages auf HF Daily Papers am 2. September 2026), einem quelloffenen multimodalen GUI-Agenten, der mehr als 170 mehrsprachige mobile Anwendungen sowie native Desktop-Systeme abdeckt, mit Multi-Modell-Verifizierung für zuverlässige RL-Signale und Sicherheitsmechanismen für folgenreiche Aktionen.
Ant Group pubblica il rapporto tecnico di UI-Venus-2 (39 upvote, paper più votato del giorno su HF Daily Papers il 2 settembre 2026), un agente GUI multimodale open source che copre oltre 170 applicazioni mobili multilingue e i sistemi desktop nativi, con verifica multimodello per segnali RL affidabili e meccanismi di sicurezza per le azioni con conseguenze.
Ant Group el publica el rapport tecegh d'UI-Venus-2 (39 upvote, paper pussée votaa del dì sora HF Daily Papers el 2 de setember 2026), on agent GUI multimodal open source che 'l quatta pussee de 170 applicazzion mobil multilingov e i sistema desktop nativ, con verifica multimodej per di segnaj RL fidabij e di meccanism de sicurezza per i azzion con conseguenz.

Sûreté & RL

Safety & RL

Sicherheit & RL

Sicurezza & RL

Sùretaa & RL

Safin-1 : la sécurité comme état interne du modèle, et DiagEvo apprend de ses erreursSafin-1: safety as an internal model state, and DiagEvo learns from its mistakesSafin-1: Sicherheit als interner Zustand des Modells — und DiagEvo lernt aus seinen FehlernSafin-1: la sicurezza come stato interno del modello, e DiagEvo impara dai propri erroriSafin-1: la sicurezza comè stat intern del modej, e DiagEvo el impara di soeu error

Le papier Safin-1 du Shanghai AI Laboratory (12 upvotes) défend une « sécurité de l'intérieur » : via une architecture MARCH de routage mémoire à travers l'historique de contexte, la sûreté devient un état natif et adaptable du modèle, sans re-entraîner le backbone. Toujours le 2 septembre 2026, DiagEvo de Meituan LongCat (5 upvotes) montre qu'un curriculum d'auto-évolution peut être dérivé de la propre histoire d'échecs du solveur : 72,3 % de précision moyenne sur cinq benchmarks mathématiques avec Qwen3-8B, 4,5 points au-dessus de R-Zero.
The Safin-1 paper from the Shanghai AI Laboratory (12 upvotes) argues for "safety from within": via a MARCH architecture routing memory through the context history, safety becomes a native, adaptable state of the model, without retraining the backbone. Also on Wednesday, 2 September 2026, DiagEvo from Meituan LongCat (5 upvotes) shows that a self-evolution curriculum can be derived from the solver's own failure history: 72.3 per cent average accuracy across five mathematical benchmarks with Qwen3-8B, 4.5 points above R-Zero.
Das Paper Safin-1 des Shanghai AI Laboratory (12 Upvotes) plädiert für «Sicherheit von innen»: Über eine MARCH-Architektur zur Speicher-Routing durch den Kontextverlauf wird Sicherheit zu einem nativen, anpassungsfähigen Zustand des Modells — ohne das Backbone neu zu trainieren. Ebenfalls am 2. September 2026 zeigt DiagEvo von Meituan LongCat (5 Upvotes), dass ein Curriculum der Selbstevolution aus der eigenen Fehlergeschichte des Solvers abgeleitet werden kann: 72,3 Prozent durchschnittliche Genauigkeit auf fünf mathematischen Benchmarks mit Qwen3-8B — 4,5 Punkte über R-Zero.
Il paper Safin-1 dello Shanghai AI Laboratory (12 upvote) difende una «sicurezza dall'interno»: tramite un'architettura MARCH di instradamento della memoria attraverso la storia del contesto, la sicurezza diventa uno stato nativo e adattabile del modello, senza riaddestrare il backbone. Sempre il 2 settembre 2026, DiagEvo di Meituan LongCat (5 upvote) mostra che un curriculum di autoevoluzione può essere derivato dalla stessa storia di fallimenti del solver: 72,3 % di precisione media su cinque benchmark matematici con Qwen3-8B, 4,5 punti sopra R-Zero.
El paper Safin-1 del Shanghai AI Laboratory (12 upvote) el sostegn ona « sicurezza de denter»: cont ona architetura MARCH de instradament de memoria travers la storia del contest, la sùretaa la deven on stat nativ e adattabil del modej, senza re-addestrà el backbone. Semper el 2 de setember 2026, DiagEvo de Meituan LongCat (5 upvote) el mostra che on curriculum de auto-evoluzzion el pò vess derivaa de la propia storia di insuccess del solver: 72,3% de precision media su cinch benchmark matemategh con Qwen3-8B, 4,5 pont sora de R-Zero.

Page 4 — Page 4 — Seite 4 — Pagina 4 — Pagina 4 — La Communauté & ÉditoCommunity & EditorialCommunity & EditorialLa Comunità & EditorialeLa Comunitaa & Edeg

V. Signaux communautéCommunity signalsSignale aus der CommunitySegnali dalla comunitàSegnaj comunitaa

Débat public

Public debate

Öffentliche Debatte

Dibattito pubblico

Dibattit publegh

Dan Luu passe au crible les prédictions d'Ed ZitronDan Luu puts Ed Zitron's predictions under the microscopeDan Luu nimmt Ed Zitrons Vorhersagen unter die LupeDan Luu passa al setaccio le previsioni di Ed ZitronDan Luu el passa al setasciû predizzion d'Ed Zitron

Dan Luu a publié le 1er septembre 2026 une vérification factuelle des prédictions du sceptique Ed Zitron, objet d'une discussion Hacker News à 543 points et 629 commentaires. L'exercice, rare dans le débat IA, confronte chaque prophétie pessimiste publiée à ce qui s'est effectivement produit — et alimente un débat plus large sur la responsabilité des commentateurs des deux camps.
Dan Luu published on Tuesday, 1 September 2026 a fact-check of the predictions of sceptic Ed Zitron, prompting a Hacker News discussion with 543 points and 629 comments. The exercise, rare in the AI debate, confronts each published pessimistic prophecy with what actually happened — and feeds a broader debate about accountability among commentators on both sides.
Dan Luu veröffentlichte am 1. September 2026 eine faktenbasierte Überprüfung der Vorhersagen des Skeptikers Ed Zitron, Gegenstand einer Hacker-News-Diskussion mit 543 Punkten und 629 Kommentaren. Diese Übung, selten in der KI-Debatte, konfrontiert jede pessimistische Prophezeiung mit dem, was tatsächlich eingetreten ist — und befeuert eine breitere Debatte über die Verantwortung von Kommentierenden in beiden Lagern.
Dan Luu ha pubblicato il 1° settembre 2026 una verifica fattuale delle previsioni dello scettico Ed Zitron, oggetto di una discussione Hacker News con 543 punti e 629 commenti. L'esercizio, raro nel dibattito sull'IA, confronta ogni profezia pessimista pubblicata con quanto effettivamente accaduto — e alimenta un dibattito più ampio sulla responsabilità dei commentatori di entrambi gli schieramenti.
Dan Luu l'ha publicaa el primm de setember 2026 ona verifica fattuala di predizzion del settegh Ed Zitron, oggett d'ona discussion Hacker News a 543 pont e 629 comment. L'esercizzi, rar in del dibattit IA, el mett a confront ogni profezia pessimista publicada con quell che l'è suceduu propi — e 'l alimenta on dibattit pussee largh sora la responsabilitaa di comentador de tucc duu i camp.

Pratiques

Practice

Praktiken

Pratiche

Prategh

Retour d'expérience : les modèles locaux sur Mac Mini M4 ProHands-on: local models on the Mac Mini M4 ProErfahrungsbericht: lokale Modelle auf dem Mac Mini M4 ProRitorno d'esperienza: i modelli locali su Mac Mini M4 ProRetorn de esperienza: i modej localegh sora Mac Mini M4 Pro

Un bilot détaillé sur une configuration de modèles locaux sur Mac Mini M4 Pro a réuni 116 points et 57 commentaires sur Hacker News le 1er septembre 2026. Complément utile à slotstream, il documente les choix concrets — modèles, quantifications, outils — d'un usage quotidien hors cloud, signe que l'écosystème local Mac atteint une masse critique.
A detailed post on a local model setup on a Mac Mini M4 Pro gathered 116 points and 57 comments on Hacker News on Tuesday, 1 September 2026. A useful companion to slotstream, it documents the concrete choices — models, quantisations, tools — of everyday cloud-free usage, a sign that the local Mac ecosystem is reaching critical mass.
Ein ausführlicher Beitrag über ein Setup lokaler Modelle auf einem Mac Mini M4 Pro versammelte am 1. September 2026 116 Punkte und 57 Kommentare auf Hacker News. Als nützliche Ergänzung zu slotstream dokumentiert er die konkreten Entscheidungen — Modelle, Quantisierungen, Werkzeuge — einer täglichen Nutzung ausserhalb der Cloud; ein Zeichen, dass das lokale Mac-Ökosystem eine kritische Masse erreicht.
Un post dettagliato su una configurazione di modelli locali su Mac Mini M4 Pro ha raccolto 116 punti e 57 commenti su Hacker News il 1° settembre 2026. Utile complemento a slotstream, documenta le scelte concrete — modelli, quantizzazioni, strumenti — di un uso quotidiano fuori dal cloud, segno che l'ecosistema locale su Mac raggiunge una massa critica.
On post detajaa sora ona configurazzion de modej localegh sora Mac Mini M4 Pro l'ha regolt 116 pont e 57 comment sora Hacker News el primm de setember 2026. Complement util a slotstream, el documenta i scerni concret — modej, quantizzazzion, strument — d'on usagg quotidian foeura del cloud, segn che l'ecosistema local Mac el ragonza ona massa crìtega.

Métascience

Metascience

Metawissenschaft

Metascienza

Metascenza

BenchMIRT : ce que les benchmarks LLM mesurent vraimentBenchMIRT: what LLM benchmarks really measureBenchMIRT: Was LLM-Benchmarks wirklich messenBenchMIRT: che cosa misurano davvero i benchmark LLMBenchMIRT: quell che i benchmark LLM mesuren propi

Le Seattle AI Lab (Allen Institute for AI) publie BenchMIRT le 1er septembre 2026, une étude systématique de ce que mesurent réellement les benchmarks LLM en appliquant la théorie psychométrique MIRT. La conclusion interpelle une industrie qui communique presque exclusivement en scores : une part de la variance des classements tient à des dimensions latentes que les métriques publiques confondent.
The Seattle AI Lab (Allen Institute for AI) published BenchMIRT on Tuesday, 1 September 2026, a systematic study of what LLM benchmarks actually measure by applying MIRT psychometric theory. The conclusion challenges an industry that communicates almost exclusively in scores: part of the variance in rankings stems from latent dimensions that public metrics conflate.
Das Seattle AI Lab (Allen Institute for AI) veröffentlicht BenchMIRT am 1. September 2026, eine systematische Untersuchung dessen, was LLM-Benchmarks tatsächlich messen, durch Anwendung der psychometrischen Theorie MIRT. Die Schlussfolgerung stellt eine Industrie infrage, die fast ausschliesslich in Scores kommuniziert: Ein Teil der Varianz der Ranglisten beruht auf latenten Dimensionen, die die öffentlichen Metriken vermengen.
Il Seattle AI Lab (Allen Institute for AI) pubblica BenchMIRT il 1° settembre 2026, uno studio sistematico di ciò che i benchmark LLM misurano realmente applicando la teoria psicometrica MIRT. La conclusione interpella un'industria che comunica quasi esclusivamente tramite punteggi: una parte della varianza delle classifiche è dovuta a dimensioni latenti che le metriche pubbliche confondono.
El Seattle AI Lab (Allen Institute for AI) el publica BenchMIRT el primm de setember 2026, on studi sistematich de quell che mesuren propi i benchmark LLM cont l'aplicà la teoria psiçometrega MIRT. La concluzzion la fa pensà ona industria che la comunica quasi domà cont i pontegg: ona part de la varianza di classifeghe la se là sora di dimension latent che i metriegh publegh i confonden.

VI. ÉditoEditorialEditorialEditorialeEdeg

Édito

Editorial

Editorial

Editoriale

Edeg

Édito — La sûreté devient un argument de vente : bonne nouvelle, à condition qu'elle soit vérifiableEditorial — Safety becomes a selling point: good news, provided it is verifiableEditorial — Sicherheit wird zum Verkaufsargument: eine gute Nachricht, sofern sie überprüfbar istEditoriale — La sicurezza diventa un argomento di vendita: una buona notizia, a condizione che sia verificabileEdeg — La sùretaa la deven on argoment de vendita: bona noeuva, a condizzion che la sìa verificabij

Deux annonces du 1er septembre 2026 racontent la même histoire sous deux angles. D'un côté, OpenAI publie « Path to Astra » : un modèle franchit pour la première fois le seuil « Critique » de cybersécurité de son propre cadre de préparation — et le laboratoire en fait une communication de transparence. De l'autre, xAI publie « Biosecurity at the frontier » : le concurrent aligne sa communication sur le même registre (OpenAI, xAI). Que la sûreté devienne un argument concurrentiel plutôt qu'un fardeau réglementaire est une bonne nouvelle — à une condition : que ces cadres internes s'ouvrent à la vérification externe. Anthropic, qui a publié la System Card complète de Fable 5.1 et Mythos 5.1 dès le premier jour, montre la marche à suivre (annonce). Le risque inverse existe : une course aux annonces de sûreté sans substance, où le vocabulaire du risque servirait de vernis marketing. Les lecteurs jugeront sur pièces — nous nous employons, de notre côté, à continuer de les leur fournir.
Two announcements of Tuesday, 1 September 2026 tell the same story from two angles. On one side, OpenAI publishes "Path to Astra": a model crosses the "Critical" cybersecurity threshold of its own preparedness framework for the first time — and the lab turns it into a transparency exercise. On the other, xAI publishes "Biosecurity at the frontier": the competitor aligns its communications on the same register (OpenAI, xAI). That safety is becoming a competitive selling point rather than a regulatory burden is good news — on one condition: that these internal frameworks open themselves to external verification. Anthropic, which published the full System Card for Fable 5.1 and Mythos 5.1 on day one, is showing the way (announcement). The opposite risk exists: a race to safety announcements devoid of substance, where the vocabulary of risk would serve as a marketing veneer. Readers will judge on the evidence — and we, for our part, intend to keep providing it.
Zwei Ankündigungen vom 1. September 2026 erzählen dieselbe Geschichte aus zwei Blickwinkeln. Einerseits veröffentlicht OpenAI «Path to Astra»: Ein Modell überschreitet erstmals die Cybersecurity-Schwelle «Kritisch» des eigenen Preparedness-Frameworks — und das Labor macht daraus eine Transparenzkommunikation. Andererseits veröffentlicht xAI «Biosecurity at the frontier»: Der Konkurrent richtet seine Kommunikation auf dasselbe Register aus (OpenAI, xAI). Dass Sicherheit zu einem Wettbewerbsargument statt zu einer regulatorischen Bürde wird, ist eine gute Nachricht — unter einer Bedingung: dass sich diese internen Rahmen der externen Überprüfung öffnen. Anthropic, das die vollständige System Card von Fable 5.1 und Mythos 5.1 schon am ersten Tag veröffentlichte, zeigt den Weg (Ankündigung). Das umgekehrte Risiko besteht ebenfalls: ein Wettlauf um substanzlose Sicherheitsankündigungen, in dem der Risikovokabular als Marketingfirnis dient. Die Leserschaft wird anhand der Dokumente urteilen — wir bemühen uns unsererseits, diese weiterhin zu liefern.
Due annunci del 1° settembre 2026 raccontano la stessa storia da due angolazioni. Da un lato, OpenAI pubblica «Path to Astra»: un modello supera per la prima volta la soglia «Critica» di cybersicurezza del proprio quadro di preparazione — e il laboratorio ne fa una comunicazione di trasparenza. Dall'altro, xAI pubblica «Biosecurity at the frontier»: il concorrente allinea la propria comunicazione sullo stesso registro (OpenAI, xAI). Che la sicurezza diventi un argomento competitivo anziché un peso regolamentare è una buona notizia — a una condizione: che questi quadri interni si aprano alla verifica esterna. Anthropic, che ha pubblicato la System Card completa di Fable 5.1 e Mythos 5.1 fin dal primo giorno, mostra la strada da seguire (annuncio). Esiste il rischio inverso: una corsa agli annunci di sicurezza senza sostanza, in cui il vocabolario del rischio servirebbe da vernice marketing. I lettori giudicheranno sulla base dei fatti — da parte nostra, ci adoperiamo per continuare a fornirglieli.
Duu anunzi del primm de setember 2026 disen la midemma storia con duu anglegh diferent. De la banda OpenAI, « Path to Astra »: on modej el passa per la prima voeulta la soglia « Crìtega » de cybersecurity del sò propi quadro de preparazzion — e 'l laboratori el ne fa ona comunicazzion de trasparenza. De l'altra banda, xAI el pubblica « Biosecurity at the frontier »: el concorrent el alinea la soa comunicazzion sul midemm register (OpenAI, xAI). Che la sùretaa la deven on argoment de vendita putoost che on pes regolatori l'è ona bona noeuva — a ona condizzion: che sti quadro intern chì se derven a la verifica esterna. Anthropic, che l'ha publicaa la System Card completa de Fable 5.1 e Mythos 5.1 del primm dì, el mostra la strada de fà (anunzi). El ris'c invers l'esist: ona corsa ai anunzi de sùretaa senza sostanza, indova che 'l vocabolari del ris'c el servarà de vernis de marketing. I lector giudegherà su la robi — numm, de la nostra banda, se femm promett de fornìghela semper.