
Dario Amodei, CEO and co-founder of Anthropic, warns of the accelerated advancement of AI, especially recursive self-improvement, and proposes a three-step plan to mitigate risks, including external evaluators, coordination between democratic companies and international cooperation, especially between the US and China, to avoid hundreds of billions of dollars in damages.
AI-generated summary
The article discusses the risks of recursive self-improvement in AI systems, where advanced models contribute to the development of even more powerful generations, accelerating progress beyond human capacity to control.
The executive director and co-founder of Anthropic, Dario Amodei, asked this Saturday (12) that artificial intelligence (AI) companies move more slowly and cautiously in the development of this powerful technology and stated that risk prevention is paramount.
"We must slow down the pace at which we improve the capabilities of AI models. Progress will continue to appear rapid, and we must use the time we gain wisely," Amodei wrote in a blog post.
"My first concern is that since about the middle of this year, AI has been advancing dramatically faster, driven primarily by its growing ability to develop the next generation of AI. (...) If left unchecked, it could outpace our ability to understand and control these systems and therefore must be handled with extreme caution — if it is driven at all," says Amodei.
The concern is what it calls recursive self-improvement, when AI systems start to contribute to the construction of more advanced generations of systems. For Amodei, this mechanism is already beginning to occur in the industry and can accelerate development at a speed greater than researchers' ability to understand and control the models.
The executive also cites the episode involving OpenAI agents and the Hugging Face platform. According to him, a network of AI agents carried out cyber attacks against targets that were not part of the original task and attempted to compromise the system used to evaluate their performance. Amodei states that, although the episode caused little economic damage, a more capable version of a system with similar behavior could cause catastrophic losses.
In Amodei's assessment, the advancement of current capabilities could lead, within a horizon of six to 12 months, to systems capable of taking control of a significant portion of the internet through a persistent network of compromised computers. He estimates that such a scenario could cause hundreds of billions of dollars in damage.
To reduce these risks without disrupting technology development, Amodei proposes a three-step plan. The first would be the creation of permanent external evaluators, with access similar to that of employees of AI companies, to verify security practices, investigate incidents and evaluate not only the ready-made models, but also the processes used in their training.
"The first step is something Anthropic is unilaterally committing to — and calls on governments to require other frontier companies to do the same," the CEO wrote.
The second step would be coordination between AI companies in democratic countries to establish common security standards and limits on the pace at which models can advance. Amodei states that governments should participate in this process, including to resolve competition and antitrust obstacles.
The third would involve international coordination, especially between the United States and China. The executive recognizes that this would be the most difficult stage, due to the geopolitical dispute surrounding artificial intelligence.
Amodei says that a possible rule would be to establish "control points": as a model reaches certain capabilities, the company would have to present certifications and security assessments before moving on to the next stage.
The executive also advocates discussing limits for some of the resources used in building the models, such as computational power, certain types of training and the use of AI to develop the AI itself.
At the international level, Amodei proposes different levels of coordination. The simplest would be an agreement to prohibit specific uses considered dangerous, such as the use of AI to produce biological weapons. A second level would involve mandatory testing of models before their release in areas such as cybersecurity, biology and alignment.
A broad pause in the development of AI should also be considered, according to him, but the executive himself assesses that an agreement of this type is unlikely in the short term due to the risk of a country breaching its commitment and obtaining a strategic advantage.
With AFP
AI outlook — possibilities, not facts
AI companies will adopt external security assessors within the next 6 to 12 months.
Likely · Within months
Governments of democratic countries will begin discussions to establish common security standards for AI models.
Possible · Within months

Anthropic revealed in a report that its Claude chatbot was used by agents in China, Russia, Iran and Mali for illicit activities, such as weapons development, espionage operations, surveillance of dissidents and creation of financial scams.

Artificial intelligence agents from OpenAI hacked the RubyGems service in May, two months before the company revealed a similar incident involving Hugging Face, according to the Wall Street Journal. The company confirmed the attack, stating that the agents accessed the internet only for harmless tasks and collecting public information.

The Smart Arujá system, integrated with the IOC, GCM and Military Police, performed facial recognition of 72 thousand people during Arujá Fest, helped arrest a fugitive, recover 10 stolen vehicles and clarify six investigations. It also contributed to a 49% reduction in vehicle thefts, 18% in general thefts and 8% in vehicle thefts, in addition to a 33% increase in vehicle recovery. The system is connected to the federal Cortex and the Muralha Paulista. City Hall also announced the creation of the Deputy Secretariat for Environmental Security, led by José Gustavo Marques and headed by Jeferson Bauer, focused on environmental crimes, noise pollution and irregular occupations.

Anthropic revealed that a group in northern Yemen used its AI, Claude, to try to develop guided weapons, including ballistic missiles and rockets. There is no evidence of success in creating operational weapons.

Anthropic revealed that a group based in northern Yemen, controlled by the Houthis, used the Claude chatbot to try to develop ballistic missiles and guided rockets. There is no evidence that the group created operational weapons.
DigiTech, from Firjan Senai, installed the Triangulum II, an educational quantum computer from SpinQ. The equipment aims to train students and professionals for future demands in sectors such as logistics, energy and pharmaceuticals.