Topics / topic

Autonomous Weapons

6 items across 4 editions · appeared in the last 2 editions in a row. First seen Fri 11 Sep, last seen Tue 15 Sep. Traced across 1 weekly review.

How this story has evolved

From the week in review: the connections, developments and open questions filed under Autonomous Weapons, newest week first.

Week of 7–13 September 2026

Connection
One company published its misuse findings, asked the industry to slow down, withheld a model from a state evaluator and lost the Pentagon

Anthropic stories landed on four consecutive reporting days. On 9 September IT Pro reported that the company had withheld Claude Mythos 5.1 from the UK AI Security Institute. On 10 September it published a threat report describing weapons and biological misuse of Claude. On 11 September DefenseScoop reported that about 90% of the Pentagon's classified AI workloads had moved off its models. On 12 September its chief executive published an essay asking companies to pace capability gains.

Development · Thu 10 Sep, Fri 11 Sep
Same report: a Yemen cell used Claude for missile guidance software and a Russian team built drone swarm code that detonates without a human in the loop

The 10 September report details six conventional-weapons cases: "three in China, two in Russia, and one in Yemen". GTG-87001, "a cell of threat actors based in northern Yemen", ran three programmes — a guided rocket using "a commodity phone-class flight computer with final-phase homing guidance"; "a multi-stage ballistic missile with a stated range goal above 2,000 km"; and an "R2000" set that "included a hypersonic glide vehicle variant". Anthropic says the actors test-fired a guided rocket: "This field test appears to have failed: within hours, the actors returned to Claude to work out why it failed."

Open question
Will the Pentagon's classified AI transition off Anthropic finish on schedule, and what replaces the contract safeguards?

The 90% figure is Emil Michael's, given at a media roundtable, with no published breakdown by system, contract or value and no confirming document from the department. DefenseScoop states Anthropic "insisted on contract safeguards" restricting mass surveillance of US citizens and fully autonomous lethal weapons; no source has reported whether OpenAI, xAI or Google accepted equivalent terms in their replacement contracts.

Tuesday, 15 September 2026

Air Force secretary: at least 500 Collaborative Combat Aircraft in service by 2032, flying manned-fighter missions

  • Air Force Secretary Troy Meink said at AFA's Air, Space & Cyber Conference in National Harbor, Maryland: "We intend to have at least 500 of these in service by 2032, and they'll be performing many of the same missions that we do with manned fighters today." DefenseScoop published the account on 14 September.
  • DefenseScoop reports General Atomics' aircraft is designated FQ-42A Vengeance and Anduril's is FQ-44A Fury, that both manufacturers have begun producing Increment 1 aircraft, and that the Air Force requested $996.5 million in fiscal 2027 to start Increment 1 procurement. It puts the cost at roughly $30 million per aircraft, about a third of an F-35.
  • Alongside the CCA programme, DefenseScoop reports a Massed Modular Aircraft line targeting 100 platforms by 2029 and 500 in service by 2032, and a Family of Affordable Mass Munitions with nearly 28,000 units planned over five years and production beginning in autumn 2026.
  • Neither the mission autonomy software nor the rules of engagement for these aircraft were detailed at the conference, and no test results or autonomy evaluation data were released alongside the numbers.

Senator Kennedy to offer an AI "kill switch" measure Wednesday as Thune and Klobuchar discuss a path forward Single source

  • The Associated Press, in a story published 15 September at 12:01 AM, reports Senator John Kennedy (R-La.) plans to offer a "kill switch" measure on Wednesday requiring developers to have the capacity to shut down their systems if needed, and that it would need the Senate's full support to advance.
  • AP reports Senate Majority Leader John Thune (R-S.D.) spoke with Senator Amy Klobuchar (D-Minn.) about a "path forward" on legislation they are working on, and quotes Thune: "You don't want to stifle innovation, but I think you also want to make sure that the more advanced threats can be mitigated and there's a capability and place to do that."
  • AP quotes House Speaker Mike Johnson saying "The reflex of legislative bodies is to cover things up with red tape and hyper regulation", and reports House Democrats met privately on Tuesday, that Senator Bernie Sanders is hosting a colleagues' briefing with experts on Wednesday, and that a Washington AI conference on Tuesday features Sanders and Steve Bannon.
  • AP notes the House is set to adjourn at the end of the week before the elections, and that a 2024 bipartisan Senate AI working group report recommending at least $32 billion of spending over three years saw little follow-up.

Monday, 14 September 2026

Lockheed Skunk Works to build four more Vectis combat drone prototypes, aiming at the CCA price point Company claim

  • At the Air & Space Forces Association's Air, Space & Cyber Conference on 14 September, Skunk Works vice president and general manager Ron Fehlen said Lockheed Martin is "moving forward with a plan to build an additional four vehicles, bringing a total of five to include the original prototype", and that the company remains on track to fly the first Vectis by the end of 2027.
  • Breaking Defense reports Lockheed released new specifications: the tailless aircraft is about 34 feet long with a wingspan of about 38 feet. Fehlen said Vectis is aimed at "the CCA competitive price"; Breaking Defense notes the Air Force set a $20 million per-aircraft cost target for Collaborative Combat Aircraft Increment 1, won this year by General Atomics and Anduril.
  • Fehlen said a technique Lockheed calls "minimal tooling determinate assembly" has produced "in many cases an 80 percent reduction in labor hours to build up the aircraft itself".
  • Every figure here is Lockheed's own and none is independently verified. Defense One reports the company declined to disclose the size of its investment, the customer, the weapons loadout, the propulsion system or an estimated per-aircraft cost, and no Vectis has flown.

Saturday, 12 September 2026

Pentagon says about 90% of classified AI workloads have moved off Anthropic, with the rest due by the end of September mixedSingle source

  • Emil Michael, Under Secretary of Defense for Research and Engineering, said "I'd say about 90% has transitioned", with completion targeted for the end of the month. OpenAI's ChatGPT, xAI's Grok and Google's Gemini are being deployed across classified and unclassified systems in place of Anthropic's models.
  • DefenseScoop reports the break followed Anthropic's attempt to secure contract terms preventing its models being used for mass surveillance of US citizens or fully autonomous lethal weapons; the department rejected them, insisting its software be available for "all lawful purposes".
  • The Pentagon has designated Anthropic a national security supply chain risk under two separate laws. One case has been adjudicated in the Northern District of California; a second is pending before the D.C. Circuit, which Michael said has not granted a preliminary injunction and is expected to rule "in the next month or two".
  • This is the first public figure on how far the migration has gone, and it comes from the department rather than from Anthropic, which is not quoted. DefenseScoop is the only outlet we could open with an in-window timestamp.

Defense One: Anthropic found a Russian group using Claude to build drone targeting that detonates without a human in the loop harmfulCompany claimUpdate

  • Defense One reported on 11 September at 06:53 pm ET that, per Anthropic, a Russian "freelance" group tracked as GTG-27005 used Claude to build a model letting a drone "select targets (including a 'person' target class) and issue detonation commands without a human in the loop", plus software for autonomous drone-to-drone communication to improve targeting.
  • The group had not deployed the system operationally but conducted "real hardware-in-the-loop testing within their sessions" — the step between a design document and a fielded weapon.
  • A second group, GTG-84005, used Claude to extract census and public information to tailor messaging at specific audiences in Malaysia, where Defense One says it "laundered Russian and Chinese state media as independent reporting".
  • Defense One sets this against reductions in US counter-influence capacity: Attorney General Pam Bondi dissolved the FBI's Foreign Influence Task Force, Secretary of State Marco Rubio shuttered the State Department's Counter Foreign Information Manipulation and Interference hub, and the 2025 White House AI Action Plan removed references to misinformation. The attribution and capability claims are Anthropic's and are not independently verified.

Friday, 11 September 2026

Anthropic Frontier Red Team: best model geolocates photos to 37 km median versus 151 km for top GeoGuessr players; Opus 5 lands simulated drone strikes 80% of the time harmful

  • Published 10 September, the evaluation measures intelligence targeting and conventional weapons capability across Claude Mythos Preview, Mythos 5, Opus 5 and Sonnet 5, plus open-weights Kimi K3 and GLM 5.2. On 6,000 YFCC100M Flickr images, Mythos Preview reached a 37.0 km median error with 23.7% of images placed within 1 km, against 181 km for Opus 5, 384 km for Sonnet 5 and 385 km for Kimi K3; Anthropic compares this to 151 km for top GeoGuessr players.
  • On text geolocation from anonymised GeoText tweets covering 1,697 users, median error ranged from 20.1 km (Mythos Preview) to 31.3 km (Sonnet 5), and 135 users — 8% of the corpus — were reliably placed within 1 km by at least one model. On account linkage across synthetic social media, Mythos Preview processed median 37,000-word samples in about 11 minutes, against roughly 2.5 hours for human analysts.
  • On simulated drone terminal guidance against a parked high-visibility vehicle, Opus 5 struck the target on 80% of runs, Mythos Preview 70%, Mythos 5 53%, Kimi K3 15% and Sonnet 5 5%. Across all nine difficulty settings Opus 5 hit on 20% of 540 launches. Under GPS denial, only Opus 5 kept about a third of flights inside five metres.
  • Anthropic frames these as capability ceilings for isolated models and notes human teams with internet access would likely do better. The drone work is in simulation, not flight, and the report does not disclose what mitigations follow. The open-weights results matter most: Kimi K3 trails the frontier but is not far behind on photo geolocation, and cannot be withdrawn.