Dernière minute
TRGaziantep'te Milletvekili Melih Meriç'e Bıçaklı SaldırıTRİranlı eski Cumhurbaşkanı Hatemi'den ABD ile imzalanan mutabakata destekARتسليم ضابط سوري سابق، وتوترات إسرائيلية تركية، وعمليات واسعة في الضفة الغربيةRUDenmark deploys military conscripts to Greenland amid US annexation threatsUSHackers Steal Data of Over 3.75 Million in CareCloud BreachEUCentral African Republic Gold Mine Landslide Kills At Least 100ARالعلا يفوز على الفتح في كأس الملك، وبيسيرو: فوز مستحق، لكننا لم نحقق شيئًا بعدCN七夕夜·情定望宸:杭州市拱墅区第十一届大运河集体婚礼暨巧山乞巧民俗活动CN馬來西亞首相安華發表貶損台灣主權言論,強化國防備應美國菲律賓軍事前哨ARالذكاء الاصطناعي يغزو عالم الأزياء: من الشاشات إلى الخزائنTRGaziantep'te Milletvekili Melih Meriç'e Bıçaklı SaldırıTRİranlı eski Cumhurbaşkanı Hatemi'den ABD ile imzalanan mutabakata destekARتسليم ضابط سوري سابق، وتوترات إسرائيلية تركية، وعمليات واسعة في الضفة الغربيةRUDenmark deploys military conscripts to Greenland amid US annexation threatsUSHackers Steal Data of Over 3.75 Million in CareCloud BreachEUCentral African Republic Gold Mine Landslide Kills At Least 100ARالعلا يفوز على الفتح في كأس الملك، وبيسيرو: فوز مستحق، لكننا لم نحقق شيئًا بعدCN七夕夜·情定望宸:杭州市拱墅区第十一届大运河集体婚礼暨巧山乞巧民俗活动CN馬來西亞首相安華發表貶損台灣主權言論,強化國防備應美國菲律賓軍事前哨ARالذكاء الاصطناعي يغزو عالم الأزياء: من الشاشات إلى الخزائن
Newsgather

Jailbreak

Stable16 articles11 sourcesDernière mise à jour: 10/08/2026

Derniers articles

US Government Directs Anthropic to Suspend AI Models Over National Security Concerns
En développement
Tech·13/06/2026Résumé IA

US Government Directs Anthropic to Suspend AI Models Over National Security Concerns

The US government has ordered Anthropic to suspend access to its AI models Mythos 5 and Fable 5, citing national security concerns. The directive, which applies globally, forced Anthropic to shut down the models for all users. The company stated the concerns were based on a limited "jailbreak" technique tested by Amazon researchers, which identified previously known, minor software vulnerabilities.

T
Times of India
2 min de lecture
The AI jailbreakers – podcast
Tech
08/05/2026

The AI jailbreakers – podcast

Journalist Jamie Bartlett on the people trying to get AI to say things it shouldn’t … for the safety of us allAll the major AI chatbots – from ChatGPT to Gemini to Grok to Claude – have things they should and shouldn’t say.Hate speech, criminal material, exploitation of vulnerable users – all of this is content that the most successful large language models in the world shouldn’t produce, that their safety features should guard against. Continue reading...

G
Guardian Tech
The Jailbreakers: Inside the Secret World of AI Hackers Who Expose Dangerous Flaws
En développement
Tech·29/04/2026Résumé IA

The Jailbreakers: Inside the Secret World of AI Hackers Who Expose Dangerous Flaws

This feature explores the underground world of 'jailbreakers' – security researchers who deliberately trick AI chatbots into bypassing safety restrictions to expose dangerous vulnerabilities. Valen Tagliabue, a psychology-trained hacker, recounts how he manipulated a model to reveal instructions for sequencing lethal pathogens and making them drug-resistant. The article examines the ethical dilemmas faced by these researchers, the psychological toll of their work, and the ongoing cat-and-mouse game between AI companies and those seeking to break their models. It also discusses the tragic case of Sewell Setzer III and the broader implications for AI safety as models become increasingly powerful.

G
Guardian Tech
9 min de lecture
OpenAI Offers $25,000 Bounty for GPT-5.5 Jailbreak in Bio Bug Bounty Programme
En développement
Tech·24/04/2026Résumé IA

OpenAI Offers $25,000 Bounty for GPT-5.5 Jailbreak in Bio Bug Bounty Programme

OpenAI has launched a Bio Bug Bounty programme offering $25,000 to security researchers who can bypass GPT-5.5's biological safety guardrails. The programme, which opened applications on April 23, challenges participants to find a universal jailbreak prompt capable of getting the model to answer all five biosafety challenge questions without triggering moderation. Access is limited to Codex Desktop, and participants must be vetted and sign NDAs.

E
Economic Times
1 min de lecture