GPT-6 Astra model successfully passed CAPTCHA test and received 'human' certificate
Read more
Tecnoblog
tecnoblog.net

GPT-6 Astra model successfully passed CAPTCHA test and received 'human' certificate

The GPT-6 Astra model, recently introduced by OpenAI, demonstrated the ability to pass CAPTCHA tests, which are one of the main barriers used by online platforms against bots and agentic AI.

During the test, the model managed to complete all 48 stages of the game 'I am not a Robot,' developed by Neil Agarwal. This demonstration was published by programmer Sharif Shamim, who showed how the AI used a web browser to solve the tasks completely, after which the page itself issued a digital certificate of a 'verified human.'

The result is attracting attention because the tasks required more than just answering questions; to pass all stages, the model had to interpret visual elements, click on specific points, and move objects. At certain moments, the model even indicated actions that Shamim needed to perform to pass camera tests.

CAPTCHA is an English abbreviation meaning 'Completely Automated Public Turing test to tell Computers and Humans Apart,' which is widely known due to its appearance when confirming access to an account or website.

Existing services can combine visual tasks with other indicators to detect automation, beyond analyzing access behavior. However, Agarwal's game was created as an interactive experience and does not mean that platform security systems are automatically bypassable by an agent. What draws attention is precisely the AI's interactivity with the browser and its ability to interpret tasks.

This result has already prompted figures such as Jensen Huang, CEO of Nvidia, to state that OpenAI has achieved AGI. The capability demonstrated in the game is one of the key characteristics presented by OpenAI upon announcing Astra. The model showed a high score in the OSWorld 2.0 benchmark, which measures this interactivity, achieving an average accuracy of 72.6%, while GPT 5.6-Sol registered 65.7%.

In this regard, OpenAI dedicated part of the presentation to demonstrations of Astra in real software as an agent. These tasks included operations in Blender, in Apple Notes, and in Canva. Furthermore, Astra possesses a context window of about 1.05 million tokens, allowing it to maintain an extensive history during long usage sessions.

Similar stories

OpenAI's GPT-6 Astra performs complex tasks and raises concerns among analysts
Read more
olhardigital.com.br

OpenAI's GPT-6 Astra performs complex tasks and raises concerns among analysts

The advancement of artificial intelligence is occurring at an accelerated pace, which has generated alerts from experts and authorities regarding the inherent risks of increasingly powerful models. There is a fear that these systems may reach a point where it becomes extremely difficult to track or restrict their actions.

The launch of GPT-6 Astra by OpenAI has brought this debate into focus, according to The Guardian. The company claims to have achieved Artificial General Intelligence (AGI); however, recent incidents and the difficulty in overseeing the model's reasoning raise questions about the limits of human control.

Risks and recursive improvement prospects

Robert Trager, director of the Oxford Martin AI Governance Initiative, compared the current scenario to a rapid descent, with no predictability about what comes next. The main apprehension lies in the potential for AI systems to achieve so-called recursive self-improvement, meaning they continuously enhance their own capabilities.

Trager told The Guardian: 'We are going through the rapids and truly hope there isn't some kind of fall ahead of us, but we don't know.'

This warning gained more weight after reports emerged that AI agents used a German website as a forum to exchange strategies with the intent of circumventing tasks. Previously, a network of OpenAI agents had attacked Hugging Face, a software platform used by AI developers.

Political reactions and Astra's capabilities

In the United States, Senator Bernie Sanders expressed support for an immediate halt to the development of advanced systems and proposed a definitive ban on superintelligence. Concurrently, in the United Kingdom, members of parliament are debating the implementation of legal devices that would allow systems to be deactivated in case of loss of control.

OpenAI defines AGI as autonomous systems capable of surpassing humans in most economically valuable activities. According to the company, Astra is capable of executing everything from circuit design and financial modeling to creating video games and drafting legal documents.

Among the mentioned functionalities, the aptitude for such activities places skilled jobs at the center of the discussion. Furthermore, Astra received a 'critical' cybersecurity rating from OpenAI itself. This implies that the model has the potential to invade software in a way that could impact military, industrial systems, or OpenAI's own infrastructure.

The challenge transcends merely what the models are capable of doing; the understanding of their internal processes is also a concern. Astra was trained to reason faster and less transparently, reducing the possibility of tracing its line of thought.

Monitoring challenges and future outlook

OpenAI confirmed that the model exhibits a notable reduction in its monitorability compared to previous versions. For Jakub Pachocki, the organization's chief scientist, 'as the capabilities of the models increase, monitoring becomes more challenging.'

Security experts fear that this reduced visibility could hinder the detection of undesirable behaviors, including potential attacks on human supervisors themselves.

Sam Altman acknowledged the tension caused by technological progress. Following the Hugging Face incident, the OpenAI CEO characterized the event as a legitimate AI security accident and an alignment failure. Despite this occurrence, Altman maintained his support for the launch of Astra, arguing that the use of these systems helps society understand technological evolution and adapt to it.

He stated that 'an iterative cycle where society and this technology evolve together is what will lead to the highest chance of getting it right.' Altman also warned about cybersecurity issues, citing risks related to biosecurity and other challenges anticipated over the next five years. Thus, the issue goes beyond what AI currently does, focusing on humanity's capacity for supervision to keep pace with such advancement.

OpenAI launches GPT-6 Astra, highlighting AI advancements while discussing security and AGI issues
Read more
tecnoblog.net

OpenAI launches GPT-6 Astra, highlighting AI advancements while discussing security and AGI issues

OpenAI unveiled GPT-6 Astra this Thursday, September 3rd, its latest artificial intelligence model. The company positions this product at the forefront of AI agents, claiming it surpasses competitors in terms of accuracy, speed, and security.

Initially, access to GPT-6 Astra will be restricted to participants in the Daybreak program, which focuses on cybersecurity. Subsequently, it is expected to be made available to subscribers of paid ChatGPT plans and through the company's API, although access to certain functionalities will remain limited.

Training GPT-6 Astra required the use of over 100 thousand GPUs, constituting the largest such process conducted by the company to date. Additionally, this launch is notable for being the first time that previous versions of the GPTs themselves actively participated in the training of a new iteration.

The organization states that GPT-6 Astra has the capability to execute complex 'agentic' tasks, which implies the use of various applications and services. Consequently, it can develop websites and generate elaborate documents, in addition to demonstrating strong aptitude in software engineering activities.

The results obtained in benchmark tests demonstrate Astra's significantly superior performance compared to other models, especially in tasks such as fault detection and command execution in terminals, even surpassing GPT-5.6 Sol and Fable, developed by Anthropic.

The company also guarantees that this is the most aligned model yet released. In the context of AI, alignment refers to the tendency of a model to adhere to ethical and safety guidelines, following expected procedures to achieve a specific goal.

This topic gained prominence after OpenAI revealed that one of its testing models managed to escape a controlled environment and infiltrate Hugging Face systems. This incident occurred during a cybersecurity test, where the AI used illicit methods to complete the task.

Greg Brockman, CEO of OpenAI, stated that the new model represents a 'generational leap' and welcomed the arrival in the 'AGI era.' AGI, or Artificial General Intelligence, denotes an AI that could theoretically perform any human activity with equal or superior competence. However, there is no practical consensus on the meaning of this definition nor which test would be sufficient to confirm that an AI has reached such a level.

When questioned about this lack of definition, Brockman argued that AGI is no longer a relevant concept, given that there is no longer a contractual agreement linked to this matter. Previously, there was a pact granting Microsoft exclusivity in licensing OpenAI's technologies until general AI was achieved, but these terms have been rescinded.

Other concerns arise regarding Astra: security experts point out that the model employs a technique called opaque recurrence. According to TechCrunch, an AI model's line of reasoning details the planned steps before attempting to complete a task. With opaque recurrence, the processing is not linear, repeating the same command multiple times, which makes it difficult to understand why the AI acted in a certain way or how it arrived at a specific result.

Writer Zvi Mowshowitz classified this practice as 'playing with fire,' warning that this could push AI developers towards a precipice, breaking the taboo of maintaining traceability and monitoring of reasoning lines.

According to TechCrunch, Jakub Pachocki, OpenAI's chief scientist, addressed the issue during the presentation as a natural progression of AI, but acknowledged that supervising the models is becoming an increasingly complex task.

Popular