Anthropic expands powerful AI model testing program for more security teams
Read more
CGTN
cgtn.com

Anthropic expands powerful AI model testing program for more security teams

Anthropic is expanding a program that allows vetted cybersecurity specialists to test its most powerful artificial intelligence models with fewer safeguards. This follows the initiative Project Glasswing, which helped uncover over one hundred thousand software vulnerabilities this year.

Partners within Glasswing, a program aimed at securing the world's most critical software, discovered at least 129,000 confirmed vulnerabilities between April and July. Additionally, Anthropic's proprietary open-source scanning identified another 5,500 vulnerabilities between April and October.

More than 33,000 of these vulnerabilities were classified as critical or high severity. Anthropic notes that these figures are likely underestimated and expects the actual damage to be at least five times higher, as the data comes from surveying a limited number of partners.

The updated Cyber Verification Program (CVP), announced on Tuesday, merges two programs Anthropic has run over the past six months. The first, Glasswing, provided organizations involved in securing critical software with access to Claude Mythos—Anthropic's most cyber-resilient model line. The second, the original CVP, offered vetted security teams reduced restrictions on Claude Opus and Sonnet models.

Previously, the introduction of Claude Mythos Preview in April raised concerns that AI could hack software before it was secured. The new program includes three tiers, each with its own verification requirements and security controls. All three tiers provide access to Claude Opus 5.5, Sonnet 5.5, Mythos 5.1, and future models.

The Defense tier is intended for tasks such as malware analysis and incident response. Security teams, critical infrastructure operators, open-source developers, and vulnerability reporting researchers can apply. The Red Team tier involves authorized penetration testing and 'red team' exercises, but only organizations can apply. The Specialized tier has the fewest restrictions and is reserved for a small group of organizations authorized to test security-critical systems, such as power grids, aviation systems, and interbank transfer infrastructure. Anthropic conducts vetting of every participant jointly with the U.S. government, and existing Glasswing participants transition to this tier.

Similar stories

Anthropic makes more robust AI available for security teams seeking flaws before cybercriminals
Read more
olhardigital.com.br

Anthropic makes more robust AI available for security teams seeking flaws before cybercriminals

Anthropic has expanded access to its most sophisticated artificial intelligence (AI) models for cybersecurity specialists. This expansion allows previously validated groups to operate with fewer restrictions regarding certain digital security uses.

This change integrates a revised version of the Cyber Verification Program (CVP). The objective of the CVP is to enable specialized entities to use Anthropic's powerful models in vulnerability detection, incident management, and security testing.

This decision was motivated by a previous company initiative called Project Glasswing. This project helped partners discover at least 129 thousand verified software flaws between April and July 2026. Of these discoveries, more than 33 thousand were categorized as critical or high severity.

However, Anthropic emphasizes that these numbers likely represent only a fraction of the total, estimating that the real impact could be at least five times greater, given that the data was collected from only a portion of program participants.

Access Levels and Applications

The first tier, named Defense, is aimed at defensive tasks. Activities covered include malware analysis, vulnerability investigation, and incident response. Candidates for this level can be security teams, essential infrastructure operators, open-source project maintainers, and researchers with a proven track record of identifying and disclosing flaws.

Red Team

The second level, known as Red Team, increases possibilities, allowing for authorized penetration tests and red teaming exercises. To participate at this level, the condition of being an organization is required.

The third level, Specialized, will have the fewest limitations and will be intended for a small group of organizations authorized to test systems considered vital for security, such as power grids, aviation systems, and infrastructure used in interbank transfers.

All members of each category undergo rigorous verification processes. Anthropic informs that it conducts this evaluation in collaboration with the United States government.

The results achieved by Project Glasswing served as justification for the program's expansion. Between April and July, the initiative's partners identified a minimum of 129 thousand confirmed vulnerabilities. Additionally, Anthropic itself located another 5.5 thousand through its open-source code scans conducted between April and October, totaling over 134 thousand verified vulnerabilities. Of these, more than 33 thousand have already been classified as severe or critical.

The company stresses that the actual volume of flaws could be considerably higher, as the Glasswing data depends on information provided by a limited number of partners.

This decision also highlights a central dilemma in using advanced AI models in cybersecurity. The same capabilities that allow a system to locate a vulnerability and assist an expert in fixing it can be used to exploit that flaw.

Anthropic had already addressed similar concerns when launching Claude Mythos in April. At that time, there were apprehensions that AI systems might infiltrate software before their vulnerabilities were patched.

For this reason, the company is not simply removing all protections from its models for any user. Reduced restriction access will only be granted to organizations and professionals who have passed verification processes, under different degrees of control. The intention is to put more advanced capabilities into the hands of those who actively work to find and fix flaws before they are exploited by criminals.

Popular