Skip to content
You can now search across every topic, entity and event.What's new
AI: Jobs, Power & Money
21SEP

Tom's Hardware challenges Mythos zero-day claims

2 min read
16:45UTC

A technical review found Anthropic's marketing relied on 198 manual reviews to support claims of thousands of severe vulnerabilities.

EconomicDeveloping
Key takeaway

Only 198 manual reviews support Anthropic's claim of thousands of zero-day discoveries.

Tom's Hardware published a critical review of Anthropic's Mythos claims on 9 April, noting that the "thousands of zero-days" assertion rested on only 198 manual reviews 1. Many of the flagged vulnerabilities were in outdated software no longer in active use. The gap between Anthropic's marketing language and the verified sample is wide enough to warrant caution.

The Bessent-Powell emergency meeting at Treasury headquarters proceeded regardless of this scrutiny. Challenger data confirmed AI-attributed cuts crossed 107,094 the same month , suggesting federal regulators assessed the systemic risk of AI broadly, beyond Mythos's specific claims. Whether Mythos found hundreds or thousands of exploitable flaws, the CyberGym benchmark score of 83.1% versus 66.6% for its predecessor represents a measurable capability jump that the twelve Glasswing partners will deploy in production environments.

Deep Analysis

In plain English

When Anthropic announced that Claude Mythos had found 'thousands' of serious security flaws in software, it was a dramatic claim. Tom's Hardware, a technology publication, looked at how Anthropic had actually counted those flaws. The answer was: 198 human reviewers manually checked the model's outputs. Many of the flaws it identified were in old software that organisations had already stopped using. The gap between 'thousands of vulnerabilities' and 198 verified reviews is significant. The US Treasury and Federal Reserve held their emergency meeting with bank CEOs regardless of this critique, which suggests the regulators assessed the risk from the model's overall capability trajectory, not just the specific zero-day count.

First Reported In

Update #5 · The model they won't release

Tom's Hardware· 10 Apr 2026
Read original
Causes and effects
This Event
Tom's Hardware challenges Mythos zero-day claims
Independent scrutiny of Mythos's capability claims introduces uncertainty about the model's actual security impact, even as regulators acted on the headline numbers.
Different Perspectives
Salesforce, Synopsys and TD Bank Group
Salesforce, Synopsys and TD Bank Group
Salesforce, Synopsys and TD Bank Group each filed quarterly disclosures in late August booking restructuring charges, or none at all, without naming AI as a cause. Their silence matters because Challenger's tracker shows AI as a stated reason fell to fourth place in August even as the year-to-date AI-cut total still leads at 116,175.
Singapore, South Korea, Taiwan and Indonesia
Singapore, South Korea, Taiwan and Indonesia
Singapore launched its Skills and Workforce Development Agency on 16 September, giving citizens six months of free premium AI tools, while South Korea ring-fenced its AI tax windfall in a new Future Response Fund. Taiwan kept funding its AI build past NT$190bn and Indonesia rewired vocational training around AI literacy, betting state-built skills beat a market-led adjustment.
ver.di, CGT Fonction Publique and CCOO
ver.di, CGT Fonction Publique and CCOO
Germany's ver.di banked a 3.3% pay rise on 1 September and opened talks on a Tarifvertrag Transformation covering dismissal bans and reskilling, while France's CGT rejected Paris's AI negotiating timetable the same week. Spain's CCOO went further on 21 September, proposing to tax companies by the jobs they generate rather than wait for the next bargaining round.
BIS General Manager and Federal Reserve governors
BIS General Manager and Federal Reserve governors
The BIS's General Manager said on 10 September that AI displacement remains limited, even as the BIS's own survey found nearly 80% of firms plan to automate roles. Two Federal Reserve governors made the same point in July, arguing the labour-market data does not yet show a mass-firing event.
Bank of Canada, ONS and ECB
Bank of Canada, ONS and ECB
The Bank of Canada found the job-finding gap between AI-exposed and unexposed occupations widened from 2.2 to 13.9 percentage points since 2015-19, while separations barely moved. That framing, a hiring freeze rather than a firing wave, is echoed by the ECB's finding that euro-area AI use hit 52% of workers in 2026, concentrated among the university-educated.
Office for National Statistics
Office for National Statistics
Deferred its Transformed Labour Force Survey beyond November 2027 and disclosed a May 2026 telephone-collection failure. The ONS carries no AI-attribution layer at all, so Britain sits outside this month's cohort of measuring states by its own admission.