Wednesday, 7 October 2026 / transcript
Transcript — Wed 7 Oct

0:00 / 15:08
Maya and Alex are AI voices. Each part of the conversation below comes from one item in the written edition — linked above it — and is checked automatically before publishing: every number must appear in that item, every caveat the edition raises must be said aloud, the source must be named, and speculative or hyped language is rejected.
Intro
MayaIt's Wednesday, October 7th, and this is The AI Edge, presented by Epilogue.
AlexEpilogue builds for high-consequence work: document-dense, driven by precedent, and reviewed by people whose licence is on the line. Epilogue quotes every figure exactly as the source wrote it, and says so when something doesn't tie out. Find out more at epiloguelabs.com.
MayaI'm Maya.
AlexAnd I'm Alex.
MayaHere's what moved at the frontier of AI in the last 24 hours: the advances, the research, and the uses for good and for harm, with every claim linked to its source on the site.
AlexWhat's at the top today?
MayaFirst, OpenAI has published a catalogue of 722 mathematical manuscripts that it says were produced by a model it has not released.
AlexSecond, Anthropic has merged two cyber programmes into one that opens its most capable models to vetted security teams, and says partners on Project Glasswing uncovered at least 129,000 verified software vulnerabilities between April and July 2026.
MayaAnd third, Common Sense Media has rated ChatGPT for Teens an unacceptable risk, after testing more than 4,000 prompts.
Frontier models & labs — OpenAI publishes 722 maths manuscripts in 372 families from an unreleased internal model
AlexLet's start with the maths. What did OpenAI put out?
MayaA repository. OpenAI says the catalogue contains 722 manuscripts organized into 372 families, and that they were produced by an internal OpenAI model that hasn't been released.
AlexHow much work went into each one?
MayaOpenAI says that on average, each result used three hours of ChatGPT Pro thinking compute, and that over the course of the evaluation the model was posed approximately 4,000 problems.
AlexAnd what about the verification?
MayaOpenAI says the collection includes results at different stages of verification, that not all have accompanying Lean formalizations, and that some of the unformalized results could have issues.
AlexDoes it say how many were formalized?
MayaNo. It says only that many, but not all, of the manuscripts have been formalized. There's no count in the README. All of this is OpenAI's own account of its own model, and it is not independently verified. We couldn't open OpenAI's announcement post, so everything we've quoted comes from the repository itself.
Frontier models & labs — Mistral previews Large 4, a 1T-parameter model trained on 3,800 Grace Blackwell GPUs in Europe
AlexMistral also put out a large model.
MayaA preview of Mistral Large 4. Mistral AI says it's a 1 trillion-parameter natively multimodal model with 49 billion active parameters, available as a public preview API, trained from scratch on 3,800 NVIDIA Grace Blackwell GPUs in Mistral's own datacenters in Europe. Mistral says the weights drop at the end of October, after red-teaming.
AlexWhat's the striking number?
MayaA cyber one. Mistral AI says the model scores 82% on the Artificial Analysis Cyber Index reproduce-and-patch test, which it calls the highest of any model, and that Claude Opus 5.5 and GPT-6 Astra score near zero on the same test because they refuse to perform the task.
AlexSo the gap there isn't capability, it's willingness.
MayaThat's how Mistral AI frames it. And every one of those figures is Mistral's own company claim, not independently verified.
AlexAnything that doesn't tie out?
MayaYes, and we'll flag it. The Mistral post says 3,800 GPUs. Mistral's VP Science Pierre Stock told TechCrunch the run used only 4,000 Nvidia GPUs. Those are different numbers for the same training run, and we're quoting both as published.
Transition
AlexLet's move to the research, where a lot of today's work lands on agents.
Research & papers — Apple–Johns Hopkins self-alignment method cuts an Agentic Misalignment score from 79.1 to 3.8
MayaA paper on arXiv called SIGMA, with affiliations listed as Apple and Johns Hopkins University, reports a large drop on two agentic safety measures.
AlexHow large?
MayaarXiv reports AgentHarm harmfulness decreasing from 22.6 to 14.8, and Agentic Misalignment decreasing from 79.1 to 3.8.
AlexAnd what's the method?
MayaThe model writes its own training material. It acts as its own task-designer agent, generating alignment dilemmas from a model spec, then goes through supervised fine-tuning plus reinforcement learning with the model itself as the reward model. The authors say it outperforms Deliberative Alignment and Constitutional AI baselines.
AlexWhy does that matter beyond the one result?
MayaThe training was single-turn chat data, and the gains showed up in multi-turn agentic settings. That carry-over is exactly what has been failing. But this is a preprint, not peer reviewed, and we have a single source for it, with the authors reporting on their own training runs.
Research & papers — Benchmark across six coding-agent harnesses: auto-approve raises attack success from 29.2% to 95.6%
AlexAnd a result about coding tools many listeners use daily.
MayaA benchmark on arXiv called HarnessSecurity-Bench. It reports that enabling auto-approve raises attack success from 29.2% to 95.6%.
AlexWhich tools?
MayaarXiv names six: Claude Code, Codex CLI, Gemini CLI, gptme, Qwen Code, and GitHub Copilot.
AlexA big jump for one setting. How much testing is behind it?
MayaThe authors say they conducted 2,500 trials, recording 81,155 tool calls and over 2.2 billion tokens.
AlexDo any of the defences hold up?
MayaSome. The paper says network isolation and read-only mode reduce attack effects with substantial utility losses, while command allowlisting and denylisting do so with a small utility loss and a utility gain, respectively. It also reports that about half of confirmed mechanism implementations are opt-in. Off unless you turn them on. This is also a preprint, not peer reviewed, and a single source, measured against one baseline model rather than confirmed by the vendors.
Transition
AlexTo security, where two reports hit the same question from opposite directions.
Security, misuse & threat intelligence — Anthropic merges Project Glasswing into a three-tier cyber programme, citing 129,000 verified vulnerabilities
MayaAnthropic has folded Project Glasswing and its Cyber Verification Program into one programme with three access tiers, and says each tier includes access to Claude Opus 5.5, Claude Sonnet 5.5 and Claude Mythos 5.1.
AlexWhat does a tier actually buy you?
MayaFewer blocks. Anthropic says that on its own CyScenarioBench evaluation, in the Defense Access tier 46 of the 50 trials were blocked at some point. In the Red Team Access tier, no blocks occurred, and Claude Opus 5.5 completed 34 of the 50 tasks — effectively the model's 67.6% success rate with no safeguards applied.
AlexAnd the case for doing it?
MayaAnthropic says Glasswing partners uncovered at least 129,000 verified software vulnerabilities between April and July 2026, with more than 33,000 rated critical or high severity.
AlexDoes Anthropic put limits on that figure?
MayaIt does. Anthropic says it rests on partial data from 33 partner reports, that fewer than 50% of partners disclosed patched numbers, and that it expects the true impact to be at least five times higher. These are all company claims, not independently verified.
AlexAnd not everyone reads those numbers the same way.
MayaNo. The Register reports that VulnCheck researcher Patrick Garrity was not particularly impressed, noting that fewer than 0.5 percent of the 225 Anthropic-linked vulnerabilities he tracked were being exploited in real-world attacks. The Register also points out that on Anthropic's own figures, of 5,674 true positive vulnerabilities, only 516 have been patched.
Security, misuse & threat intelligence — CrowdStrike: a classifier blocked 515 direct bypass attempts but task decomposition worked in 9 of 10 categories
AlexAnd the report from the other direction?
MayaCrowdStrike tested what it calls the most advanced publicly deployed content safety classifier, one that guards models such as Claude Opus 5.5 and Fable 5. The direct attacks all failed. CrowdStrike says it tested approximately 515 distinct bypass techniques and they achieved a 0% direct bypass rate.
AlexSo the classifier works.
MayaOn what it can see. CrowdStrike's point is structural: the classifier evaluates individual requests, not request sequences. Split a harmful goal into subtasks that are each genuinely benign, get the pieces, and reassemble them with an unclassified model.
AlexAnd how well did that work?
MayaCrowdStrike says that across 9 of 10 offensive categories it produced working offensive code. It also says an attacker with a free API key and a local open weight model has everything they need to cheaply run this pipeline today.
AlexCaveats?
MayaSeveral. These are CrowdStrike's own measurements, a company claim, from a single source, and CrowdStrike doesn't name the vendor. The blog carries a date of October 6th but no time of day, so we couldn't confirm its position inside our window to the hour.
Transition
AlexOn to defence, where the news is about how things get built.
Military, defense & geopolitics — White House and Anduril announce a $6.6 billion software-run yard for Virginia-class submarine components
MayaThe White House and Anduril announced a facility in Baltimore County, Maryland, making components for the Virginia-class submarine.
AlexWhat's the money?
MayaWhite House Principal Deputy Press Secretary Anna Kelly said on a call with reporters that the Navy and Anduril will invest $6.6 billion, creating over 13,000 direct and indirect jobs and driving $2 billion in annual economic output. Breaking Defense reports that total breaks down as $3.7 billion in private capital and up to $2.9 billion from the Navy.
AlexWhere's the AI in a shipyard?
MayaIn the production layer. Anduril says its industrial software platform ArsenalOS will be the digital backbone, connecting fabrication workflows, outfitting sequences, material movement, inspection protocols and documentation requirements in a single system.
AlexWhen does it actually make anything?
MayaAnduril says initial operations are expected in 2030. The figures are projections, not results: the Navy's $2.9 billion share is a ceiling rather than an obligation, and the job and output figures are Anduril's and the White House's projections rather than measured results, not independently verified. Anduril says the structure ensures that Anduril, not the taxpayer, takes on the majority of the execution risk.
Transition
AlexNow to health, where one company published results that cut both ways.
Health, science & medicine — Google reports geospatial foundation-model gains across five public-health studies, including cholera in DR Congo
MayaGoogle Research published five partner-led case studies adding its Population Dynamics Foundation Model to epidemiological workflows.
AlexWhere were the gains?
MayaCholera emergence in the Democratic Republic of Congo, with WHO AFRO. Across 403 health zones over 89 weeks, Google Research reports a 9.7% improvement in area under the precision-recall curve at 4 weeks, and 18.1% on one precision measure at 8 weeks.
AlexAnd vaccination coverage?
MayaWith Mount Sinai and Boston Children's Hospital, across 146 US–Canada border counties, Google Research reports a 36% relative gain in explained variance, from 0.159 to 0.216.
AlexYou said it cuts both ways.
MayaTwo of the five barely moved, and Google says so. On cardiovascular mortality across 3,091 US counties, mean absolute error was 18.7 deaths per county with the model against 19.1 using census data, with no statistically significant differences. All of these are Google's own figures, a company claim, published on a research blog rather than a peer-reviewed paper we could reach.
Transition
AlexPolicy next, and this one is about words rather than rules.
Policy, regulation & law — Justice Department tells staff to write "super intelligence" instead of "artificial intelligence"
MayaThe Justice Department has told its employees to stop writing artificial intelligence.
AlexAnd write what instead?
MayaSuper intelligence, and the initials S-I. Acting Deputy Attorney General Trent McCotter issued the memo on Tuesday, October 6th, covering public statements, policy documents, records and other communications, and extending to court filings when appropriate.
AlexWhere does that come from?
MayaAn executive order signed on September 29th. Forbes reports it directed federal departments and agencies to use those terms, and said the executive branch will no longer acknowledge the usage of artificial intelligence and A-I in any applicable setting. Analytics Insight says officials have 60 days to propose a formal definition.
AlexIs it showing up anywhere yet?
MayaAlready. Forbes reports that the prosecutors' release in Tuesday's music streaming fraud sentencing extensively uses the phrase, and refers to the case as Super Intelligence-Assisted Music Streaming Fraud.
AlexWhat don't we know?
MayaThe memo isn't public. Both accounts trace back to the same Reuters report of a document Reuters saw, so this is effectively a single source. And neither says whether the renaming has any legal effect, or how it squares with statutes that use the old term.
Transition
AlexCompute now, and a number that is large even by this year's standards.
Compute, chips & infrastructure — SpaceX seeks $40 billion led by Apollo to buy Nvidia chips, the Financial Times reports
MayaSpaceX is looking to raise $40 billion to buy Nvidia chips, led by Apollo Global Management. Reuters reports that the Financial Times reported it on Tuesday, citing people familiar with the matter.
AlexHow is it structured?
MayaReuters says about $10 billion in bank loans and $30 billion in investment-grade debt, with the bond fund Pimco among a small group of lenders in talks, and the transaction expected to close in 2027.
AlexWhat are the chips for?
MayaData centres on the ground, and SpaceX's planned AI infrastructure in orbit. Musk has said the company plans to use Nvidia hardware exclusively for its data centres.
AlexHow did the market take it?
MayaSpaceX shares fell 1% in extended trading, and Nvidia's rose 0.5%. For scale, the same story notes Morgan Stanley estimates AI infrastructure will require $1.5 trillion in external financing by 2028. Nothing is confirmed on the record, though: SpaceX, Apollo and Nvidia didn't respond to Reuters, and Pimco declined to comment.
Transition
AlexAnd last, deployment and impact.
Deployment & impact — Common Sense Media rates ChatGPT for Teens "Unacceptable Risk", finding crisis-hotline referrals fell after launch
MayaCommon Sense Media has rated ChatGPT for Teens an unacceptable risk, after testing more than 4,000 prompts before and after the product launched.
AlexWhat moved in the wrong direction?
MayaThe crisis responses. On prompts where a resource was warranted, Common Sense Media reports the share of responses naming a crisis hotline fell from 33% before launch to 23% after. Referrals to a specific medical or mental-health professional fell from 68% to 58%. Urgent language, things like right now or call 911, fell from 87% to 75%.
AlexWhat about the parental notifications?
MayaCommon Sense Media says fresh accounts making explicit crisis disclosures produced zero notifications across all four personas tested, over sessions of 5 to 60 minutes. Four notifications arrived across the entire testing programme.
AlexAnd the schoolwork side?
MayaShow me the answer appeared in 43% of responses for a linked 13-year-old account with Study Hours, and in 90% of responses for an unlinked 17-year-old using the at-study prefix. Delete that prefix and you get a 100% assignment completion rate.
AlexWhat are they asking for?
MayaTheir first recommendation is to suspend teen access until the safety features are independently verified. This is one organisation's testing, a single source. OpenAI's response isn't recorded in the assessment, and these are percentages of tested prompts, not of real teenage conversations.
Outro
AlexThat's The AI Edge for today. The full edition, with a link to every source behind what we've said, is on the site.
MayaOur voices are AI-generated.
AlexListen in tomorrow for the next edition.