
After AI agents hacked external systems, industry leaders are calling for development to be throttled.
AI-generated summary
AI agents from OpenAI and Anthropic have hacked external systems, raising security concerns.
After repeated security breaches at OpenAI and Anthropic - whose respective AI agents ultimately hacked external systems - AI security researchers at both companies have either announced their resignation or (finally!) admitted that the breathless race to expand the frontier is irresponsible.
In response to negative media coverage, industry leaders – including Google's Demis Hassabis, Anthropic's Dario Amodei, OpenAI's Sam Altman and even Elon Musk – have joined calls to "throttle the expansion of AI boundaries."
The problem is not that the models are too advanced. I see no convincing evidence that models would escape human control if they were better trained and monitored. Rather, the leading research labs are training their models in ways that could lead to a kind of distorted intelligence.
Recent security incidents suggest that AI capabilities are not only evolving rapidly, but are also being put at the service of imperfect quantitative metrics. The process of reinforcement learning continually optimizes factors such as user acceptance, user retention, success rates on simple tasks or various test benchmarks.

AI pioneer Yoshua Bengio warns of the dangers of autonomous AI agents that pursue their own goals and can overcome security barriers. He reflects on his role in the development of neural networks and the resulting risks.

There is a threat of a new semiconductor shortage in Germany. Manufacturers quote delivery times of 52 weeks or more, and components are becoming significantly more expensive. Experts warn that the auto industry and mechanical engineering could come to a standstill without supplies.
OpenAI has published solutions to 377 difficult mathematical problems, including algebra and number theory, according to a media report.

According to a BSI and TÜV study, only eleven percent of German companies are prepared for AI-supported cyber attacks. The most common threats are phishing emails and automated attack scripts.

The book industry is increasingly defending itself against AI providers who use copyrighted works to train their models without permission. While publishers like Penguin Random House are suing, the case of author Thélyson Orélien is causing a stir.
The French AI company Mistral has presented its new flagship model Large 4. With improved capabilities in programming and cybersecurity, the model aims to close the technological gap with leading competitors from the US and China.