Skip to content

Build1 publisher3 min readPublished

Clearing Mythos Preview's 6,105 open findings at 97 fixes per two months takes a decade

Project Glasswing surfaced an estimated 6,202 high or critical vulnerabilities in foundational open source, with 97 confirmed fixed in two months. What the confirmation count actually measures decides how much of that gap is real.

The Engineer · Build desk

Illustration accompanying Clearing Mythos Preview's 6,105 open findings at 97 fixes per two months takes a decade

What happened

  • Anthropic opened Project Glasswing in April 2026, giving technology leaders and critical infrastructure providers early access to Claude Mythos Preview, a model built for vulnerability discovery.
  • The programme identified an estimated 6,202 high or critical severity vulnerabilities across foundational open-source software, and 97 of them were confirmed remediated within two months.
  • Gartner's Q2 2026 Emerging Risk Report, drawn from a survey of 316 senior risk executives, named AI-driven vulnerability discovery the quarter's top emerging enterprise risk.
  • The 2026 Verizon Data Breach Investigations Report put full remediation of critical known-exploited vulnerabilities at 26 percent, down from 38 percent, with median resolution time at 43 days.
  • Mandiant's M-Trends 2026 reported mean time-to-exploit at negative seven days, meaning zero-days are routinely exploited a week before official patches ship.

Compiled by The EngineerSomething wrong?How this is made

Why it matters

  • decision A queue of 6,105 unconfirmed high-severity items forces a written triage rule naming which findings a team will not fix this year, and someone has to sign that rule.
  • exposure Read together, the two field reports put roughly 50 days between first exploitation and the median fix, and that span is the exposure an in-house team is defending with staff it already has.
  • contradiction Remediation of known-exploited criticals fell 12 points year over year while discovery got faster, so capacity was already moving the wrong way before machine-scale finding volume arrived.

Take both numbers at face value and the backlog has a schedule. Subtract the confirmed fixes from the findings and 6,105 items are still open [14]. At 97 confirmations every two months, clearing 6,105 runs to about 126 months, call it ten and a half years, and that assumes Mythos Preview stops finding things [15].

The two figures are not counted the same way. The 6,202 is an estimate of what was identified; the 97 is what was confirmed remediated [2]. Confirmation takes a second event after the fix: someone checks that a maintainer shipped it, and the result gets back to the discloser. A fix that nobody confirms never shows up in the 97. So 1.6 percent [16] is a floor on the remediation rate.

Anthropic's own framing points at the input side. The company said that "even at our relatively slow pace of disclosures, Mythos Preview is adding to an already-overloaded security ecosystem" [3]. The disclosure rate is throttled by choice, and the identification rate is what produced 6,202 [2].

The Log4Shell comparison in the dev.to account carries the argument. Remediation of CVE-2021-44228 ran for months because discovering every transitive instance across complex build graphs was the hard part, not because patching a library was hard, and a model can map those transitive paths quickly while patch delivery stays human-bound [9]. What is left once discovery gets cheap is the gated sequence: ticket, architecture review, compliance gate, regression testing, deployment approvals, scheduled release [8]. Those steps take the same time however early the input arrives.

The post's mobile example is an illustration, and the most useful part of the post. An app receives a deep link such as `mybank://transfer?to=<account>&amount=<value>` while cold-booting. The URL router processes the payload parameters before the authentication completion callbacks finalise, and the transfer flow writes those parameters into active state before access controls block the view controller [10]. No linter or unit test flags it [11]. The claim underneath is that tooling evaluating whole execution paths and state graphs at once compounds isolated correctness bugs into exploitable attack chains [12].

For the 1.6 percent to say anything about your own backlog, your fixes would have to be confirmed the way a third-party discloser confirms them, and your dependency set would have to look like foundational open source maintained by people who never asked for the report. An in-house team patching its own code has a shorter loop and nobody to notify. The median resolution time in the Verizon DBIR is the figure such a team can compare itself against [6].

All of this reaches the reader through one dev.to post, which attributes its figures to Anthropic, Gartner, Verizon and Mandiant [13]. Gartner's warning, as quoted there, is about capacity: "without corresponding improvements in governance, security operations, and remediation capabilities, AI-driven vulnerability discovery may outpace organizational defenses" [5].

What to watch

  • Whether Anthropic publishes a second Glasswing tally, and whether confirmed fixes grow faster than new findings.
  • Whether the 2027 Verizon DBIR shows remediation of critical known-exploited vulnerabilities moving back above 26 percent.
  • Whether Mythos Preview access widens past the Glasswing cohort, since disclosure volume is currently held back by Anthropic's own pace.
Loading claim ledger
Loading source directory links
Loading share composer
Loading topic controls
Loading related stories