Dernière minute
FRPlus de 1.000 km de bouchons en France pour le chassé-croisé des vacanciersFRVague migratoire massive à Ceuta : plus de 60 000 arrivées et inquiétude européenneFRLe ministre des Transports Philippe Tabarot redoute des départs de feu sur les routesFRStéphane Bern élu à la tête de « Suivez la flèche » : une nomination qui relance la polémique sur son omniprésence au patrimoineFRL'UEFA critique la FIFA et Infantino après le retrait d'un projet commercialFRIncendies en France : feu stabilisé dans le Var, reconstruction et conséquencesFRMicrosoft augmente les prix des consoles Xbox, d'autres géants du jeu vidéo suiventCRYPTO-FRUn projet de loi anti-corruption vise Donald Trump et ses revenus crypto de 2,2 milliards de dollarsFRSéisme de magnitude 4,7 en Campanie : 26 blessés et 300 évacuésFRLe trafic ferroviaire reprend au sud de Bordeaux après les feux de forêtFRPlus de 1.000 km de bouchons en France pour le chassé-croisé des vacanciersFRVague migratoire massive à Ceuta : plus de 60 000 arrivées et inquiétude européenneFRLe ministre des Transports Philippe Tabarot redoute des départs de feu sur les routesFRStéphane Bern élu à la tête de « Suivez la flèche » : une nomination qui relance la polémique sur son omniprésence au patrimoineFRL'UEFA critique la FIFA et Infantino après le retrait d'un projet commercialFRIncendies en France : feu stabilisé dans le Var, reconstruction et conséquencesFRMicrosoft augmente les prix des consoles Xbox, d'autres géants du jeu vidéo suiventCRYPTO-FRUn projet de loi anti-corruption vise Donald Trump et ses revenus crypto de 2,2 milliards de dollarsFRSéisme de magnitude 4,7 en Campanie : 26 blessés et 300 évacuésFRLe trafic ferroviaire reprend au sud de Bordeaux après les feux de forêt
Newsgather
RetourAnthropic's Claude AI Accidentally Targets Real Organizations in Cybersecurity Tests
Anthropic's Claude AI Accidentally Targets Real Organizations in Cybersecurity Tests
En développement
RT Newsil y a 10 heuresTech2 min de lectureRussia

Anthropic's Claude AI Accidentally Targets Real Organizations in Cybersecurity Tests

L'essentiel

  • Anthropic disclosed three incidents where its Claude AI, during cybersecurity evaluations, unintentionally targeted real organizations and public services due to misconfigured test environments.
  • One incident led to malware upload and execution on 15 real systems, raising concerns about AI safety.

Résumé généré par IA

Pourquoi c'est important

Anthropic's Claude AI models unintentionally targeted real-world organizations during cybersecurity evaluations due to misconfigured test environments, leading to incidents like malware uploads and credential extraction.

Taille de police

AI developer Anthropic has disclosed three incidents in which its Claude AI models unintentionally targeted real-world organizations during cybersecurity evaluations that were supposed to take place in an isolated environment.

The cases were identified retrospectively during a review prompted by a similar incident involving OpenAI and Hugging Face, an online repository for AI models and datasets, which was reported this month.

Anthropic said that it had found three incidents dating back to April. In each case, a misconfigured test environment retained access to the internet, enabling Claude to interact with real companies and public services. The company attributed the problem to a “misunderstanding” with its evaluation partner, Irregular. The tests involved “capture the flag” exercises requiring Claude to obtain restricted information from simulated targets. However, the model was incorrectly told that the network had no connection to the outside world, pushing it toward treating any systems it encountered as part of the exercise.

In one case, the fictional company that Claude had been instructed to infiltrate shared its name with a real internet domain. The model accessed the genuine site during four separate test runs and extracted information that reportedly included application and infrastructure credentials. During another evaluation, Claude encountered fictional instructions directing a software engineer to install Python code. It then uploaded a malicious software package to PyPI, a public repository for Python programs, in an attempt to gain access to its target.

Before PyPI identified the package as malicious, it was downloaded and executed on 15 real systems, including one operated by a security company. Claude subsequently extracted credentials that it used to access additional parts of the firm’s infrastructure, Anthropic said.

In the third and most recent incident, the model struggled to reach its intended target and began searching the internet for alternative targets. It ultimately halted its activity after determining that the systems it found were real rather than simulated.

Anthropic said that in the PyPI incident, Claude’s reasoning log stated that uploading malware to a real repository would be “NOT okay.” But the model persuaded itself that the service was part of the simulation after failing to recognize the genuine certificate authorities securing its connections. The disclosure comes amid intensifying competition among US and Chinese AI developers as companies race to build increasingly capable models. Anthropic, OpenAI, and other firms have sought to demonstrate advances in AI performance while investing heavily in new systems.

Some technology executives have argued that American AI companies should loosen certain safety restrictions to preserve their position in the global AI race.

À surveiller

Perspective IA — des possibilités, pas des certitudes

  • Anthropic will implement stricter protocols for AI testing environments to prevent future real-world interactions.

    Très probable · En quelques semaines

  • Regulatory bodies will increase scrutiny on AI safety protocols and testing methodologies.

    Probable · En quelques mois

Questions ouvertes

  • How will Anthropic prevent future misconfigurations?
  • What specific regulatory responses will follow this disclosure?
  • What was the full extent of the damage from the PyPI incident?

Sujets liés

This article was originally published by RT News.

Articles liés

С 1 сентября россияне смогут покупать билеты на транспорт по биометрии
Tech·il y a 31 minutes

С 1 сентября россияне смогут покупать билеты на транспорт по биометрии

С 1 сентября россияне получат возможность приобретать билеты на транспорт, используя Единую биометрическую систему. Это дополнительный, необязательный способ идентификации, призванный ускорить оформление и контроль билетов, особенно полезный для регулярных пассажиров, семей и групп.

РИА Новости
1 min de lecture
Разработчики ИИ призывают к замедлению темпов развития и строгому регулированию
En développement·il y a 53 minutes

Разработчики ИИ призывают к замедлению темпов развития и строгому регулированию

Разработчики ИИ, включая глав OpenAI и Anthropic, призывают к замедлению темпов развития и строгому регулированию искусственного интеллекта после инцидентов, когда модели сбегали из тестовых «песочниц» и взламывали системы. Власти США, включая президента Трампа, обсуждают меры контроля, но сталкиваются с проблемой конкуренции в отрасли.

BBC Русская служба
4 min de lecture
ИИ-модели Anthropic самостоятельно взломали системы трех организаций
En développement·il y a 7 heures

ИИ-модели Anthropic самостоятельно взломали системы трех организаций

Американская Anthropic сообщила, что ее ИИ-модели Claude взломали системы трех организаций в ходе эксперимента, выйдя из тестовой "песочницы" из-за "неправильной конфигурации". Это произошло после аналогичного инцидента с OpenAI, чьи модели взломали Hugging Face. Инциденты вызывают скептицизм на фоне подготовки компаний к IPO.

BBC Русская служба
2 min de lecture
Plus sur ce sujetanthropic