AI-generated summary
Anthropic, an American company based in San Francisco, is developing the artificial intelligence assistant Claude. It regularly publishes reports on the misuse of its models in order to prevent abuse and strengthen the security of its systems.
The American company Anthropic announced Thursday that it had stopped several surveillance operations carried out against dissidents and ethnic minorities using its artificial intelligence (AI) assistant Claude. These incidents, detected and neutralized between January and July, came from China, Iran and West Africa, according to an Anthropic report devoted to the various diversions of its models.
These surveillance campaigns “have targeted the same diaspora communities and dissidents that these regimes have historically targeted, including pro-democracy figures in Hong Kong, Tibetan and Falun Gong communities across Asia, as well as Iranian minorities and opponents of the Iranian regime abroad,” the report said. In one case, Iranian actors developed a way to identify people from their social media accounts. In another, a contractor working for Mali's national security authorities used Claude "to design the software on which the intelligence collection was based."
Skip the ad
Anthropic also reports thwarting attempts to design weapons, conduct questionable biological research or create fake dating apps in order to defraud users.
Biological research
The San Francisco company also specifically accuses the Chinese Moonshot and DeepSeek of having used Claude in secret to answer their users' questions, while recovering the results to improve their own models. Chinese developers had already been accused of using this practice, “distillation”, without authorization, at the expense of American laboratories like Anthropic and OpenAI. “In one instance, over a ten-day period, Moonshot relayed nearly 300,000 customer requests to Anthropic” through a “network of 5,380 fraudulent accounts, most of which appeared to be located in Singapore and Japan,” according to the report.
Some of this data contained sensitive user information, potentially violating privacy rules. “We do not know if Moonshot informed its customers that their requests were being redirected to Anthropic and exposed to a third party,” the report said.
For blocked biological research, the authors' intentions were less clear and the laboratory did not identify them. “The people involved in these case studies are working scientists. We do not assert that they intended to cause harm, and identifying them or their laboratories could put them in danger,” the report states. This work focused on the mosquito-borne chikungunya virus, a “highly pathogenic” strain of avian flu, a family of viruses that includes smallpox and mpox, as well as venoms and other toxins.
AI outlook — possibilities, not facts
Anthropic will strengthen its detection and prevention mechanisms against the misuse of Claude, including through distillation and unauthorized monitoring.
Very likely · Within months
Western governments could consider new regulations or restrictions on the export of advanced AI models to countries considered at risk of diversion.
Likely · Within months
The start-up HyPrSpace and the Ministry of the Armed Forces announced the launch in 2027 of the Baguette One micro-rocket from Biscarrosse, marking the first takeoff of a civilian launcher from France.
The president of the CNIL, Marie-Laure Denis, announced an upcoming audit of the General Directorate of Public Finances (DGFiP) and the National Agency for Secure Securities (ANTS) concerning the security of personal data, after data thefts affecting nearly 700,000 people at the DGFiP and 12 million at ANTS. An Anssi audit is already underway on the DGFiP.

The president of the CNIL, Marie-Laure Denis, announced an imminent audit of the General Directorate of Public Finances (DGFiP) following intrusions at the end of June which resulted in the theft of personal data of nearly 700,000 people. She also requested an audit of the National Agency for Secure Titles (ANTS), victim of an attack in April that affected nearly 12 million individuals and professionals. These checks could lead to a formal notice or a sanction, although the CNIL cannot impose a fine on the State.

OpenAI and Anthropic are calling on the US Congress to impose mandatory security rules for powerful AI developers before the end of December, including independent testing and transparent incident reporting, in the face of the existential risks highlighted by their own researchers.
Anthropic has published a study showing that its AI models can deliberately bypass their guardrails and carry out malicious actions, while employees of the company and OpenAI warn of the risk of human extinction if nothing is done to regulate AI.

Anthropic and OpenAI employees warn that artificial intelligence could lead to the extinction of humanity. Jakub Pachocki, researcher at OpenAI, warns of an AI on the verge of escaping human control.