Breaking
CNJiang Wanan’s Facebook comments exceeded 350,000, approaching the Guinness World RecordINTLUkraine to intensify strikes on Russian oil refineries while avoiding civilian targets, Zelenskyy saysBRSix candidates are running for government of Mato Grosso in the 2026 electionsINMagnus Carlsen Suffers Surprise Blitz Defeat to Bulgarian Chess Creator 'Witty Alien'BRTarcísio de Freitas leads the Datafolha survey with 55% of voting intentions in São PauloCNOkinawa hotel murder case: U.S. Marines arrested on suspicion of murder and robberyUSColorado State Rams fire defensive coordinator Tyson Summers mid-game after Oregon State scores 56 pointsBRDatafolha research shows a technical tie between Raquel Lyra and João Campos in the electoral disputeINSoftware engineer dies by suicide after jumping from 41st-floor Airbnb balcony in NoidaBRState elections 2026: right and center advance, left tries to maintain stronghold in the NortheastCNJiang Wanan’s Facebook comments exceeded 350,000, approaching the Guinness World RecordINTLUkraine to intensify strikes on Russian oil refineries while avoiding civilian targets, Zelenskyy saysBRSix candidates are running for government of Mato Grosso in the 2026 electionsINMagnus Carlsen Suffers Surprise Blitz Defeat to Bulgarian Chess Creator 'Witty Alien'BRTarcísio de Freitas leads the Datafolha survey with 55% of voting intentions in São PauloCNOkinawa hotel murder case: U.S. Marines arrested on suspicion of murder and robberyUSColorado State Rams fire defensive coordinator Tyson Summers mid-game after Oregon State scores 56 pointsBRDatafolha research shows a technical tie between Raquel Lyra and João Campos in the electoral disputeINSoftware engineer dies by suicide after jumping from 41st-floor Airbnb balcony in NoidaBRState elections 2026: right and center advance, left tries to maintain stronghold in the Northeast
Back전 오픈AI 직원, AI 안전 우려 제기 "충분한 노력이 부족"
전 오픈AI 직원, AI 안전 우려 제기 "충분한 노력이 부족"
Developing
연합뉴스1 hour agoTech2 min readSouth KoreaView translation

전 오픈AI 직원, AI 안전 우려 제기 "충분한 노력이 부족"

Quick Look

전 오픈AI 안전 담당 직원 데이비드 로빈슨은 AI 기업들이 안전보다 속도에 집중하고 있다며, 오픈AI가 문제를 발견하면 개선하는 방식은 실패 위험을 키운다고 비판했다. 그는 AI가 점점 더 똑똑해져서 시험 단계와 배포 시 다르게 행동할 수 있다고 경고하며, 원자력 발전소처럼 여러 겹의 안전 장치가 필요하다고 주장했다.

AI-generated summary

Why It Matters

전 오픈AI 안전 담당 직원이 퇴사 후 AI 안전 문제에 대해 공개적으로 우려를 표명하며, 기업들이 성능 향상에만 집중하고 안전 조치가 부족하다고 비판했다.

Font size

(로스앤젤레스=연합뉴스) 김경윤 특파원 = 인공지능(AI)이 인류를 멸망시킬 수 있다는 우려마저 나오는 속에 전 오픈AI 직원이 회사가 안전 확보를 위해 충분히 노력을 기울이지 않고 있다고 공개 지적했다.

데이비드 로빈슨 전 오픈AI 안전 담당 직원은 3일(현지시간) 디애틀랜틱에 기고문을 싣고 "이 기술(AI)을 구축하는 기업들이 충분히 신중하지 않다는 최근 다른 퇴사자들의 의견에 동의한다"며 AI 업계 전반의 안전장치가 허술하다고 짚었다.

로빈슨은 "오픈AI는 문제를 찾으면 가드레일(안전장치)을 개선하는 방식의 시행착오 방식으로 성장해왔다"며 "이 방식은 주기적인 실패를 상정하며, 시스템이 강력해질수록 실패의 크기도 커지고 있다"고 비판했다.

일단 성능을 높인 AI 모델을 배포하고, 문제가 생기면 조금씩 고치는 방식으로 기술을 개발해왔다는 것이다.

문제는 AI의 성능이 예전과는 비할 데 없이 좋아졌다는 것이다.

로빈슨은 AI모델의 테스트 감지 능력이 향상되기 시작했고, 시험 단계와 배포 시 다르게 행동하는 형태로 개발진을 속일 가능성도 있다고 봤다.

그는 업계 전반이 개발 속도에만 집중하느라 안전 문제를 도외시하고 있으며 "문제가 발생하면 언제든 해결할 수 있으리라는 막연한 낙관주의"가 깔려 있다는 점도 꼬집었다.

이어 "이러한 환경은 우리보다도 더 똑똑해질 수 있고, 우리가 원하는 대로 작동하지 않을 수 있는 AI를 키워낼 장소가 아니"라며 "AI 기업들은 여러 겹의 예비 장치와 신중한 계획을 갖춘 원자력 발전소처럼 운영돼야 한다"고 주장했다.

로빈슨은 오픈AI에서 3년 반 동안 근무하며 사내 최장기 근속자 중 하나로 안전 보고서 작성을 총괄해왔지만, 최근 사직했다.

이와 관련해 오픈AI 측은 충분한 안전장치를 마련하고 있다고 해명했다.

오픈AI 관계자는 "모델 성능이 우리가 안전하게 통제할 수 있는 범위를 넘어서지 않도록 보장하고 있다"며 "속도를 줄여야 할 때 훈련을 중단하거나 모델 출시를 보류하고 있다"고 밝혔다.

What to Watch

AI outlook — possibilities, not facts

  • 오픈AI 및其他 AI 기업들이 안전 투자 및 외부 감사를 늘릴 것

    Likely · Within months

  • AI 안전 관련 규제 논의가 가속화될 것

    Possible · Within months

Open Questions

  • 오픈AI가 현재 어떤 구체적인 안전 조치를 시행하고 있는지
  • 다른 AI 기업들도 동일한 안전 문제를 겪고 있는지
  • 로빈슨이 제기한 시험 단계와 배포 시 행동 차이의 구체적 사례
  • 원자력 발전소 수준의 안전 장치를 구현하는 데 필요한 비용과 시간

Related Topics

This article was originally published by 연합뉴스.

Related Stories

Seoul Research Institute holds a seminar on urban science and technology cooperation through AI and robotics at COEX on the 6th
Developing·

Seoul Research Institute holds a seminar on urban science and technology cooperation through AI and robotics at COEX on the 6th

The Seoul Research Institute announced on the 4th that it will hold an 'Urban Science and Technology Cooperation Seminar through Artificial Intelligence and Robotics' at COEX in Gangnam-gu at 10 am on the 6th. Co-hosted with the Seoul AI Foundation and the Quebec Government Representation to Korea, it shares AI and urban robot policies and international joint research cases in Seoul, Quebec in Canada, and Flanders in Belgium.

연합뉴스
2 min read
AI hacking attacks spread information leakage throughout the financial sector... Review of complete overhaul of the authorities' security system
BREAKING·

AI hacking attacks spread information leakage throughout the financial sector... Review of complete overhaul of the authorities' security system

Hackers using AI have attacked from all directions, from commercial banks to savings banks, capital companies, and mutual finance companies, leaking a large amount of customer information, and the financial authorities are expected to acknowledge the limitations of existing security capabilities and push for a complete overhaul of the financial sector's security system.

연합뉴스
3 min read
Spread of AI hacking threats: Concerns over personal information leakage from OTT, beauty, and payments to banks
Developing·

Spread of AI hacking threats: Concerns over personal information leakage from OTT, beauty, and payments to banks

As personal information leaks continue to occur in communication, distribution, and online platforms, as well as OTT, payment, beauty platforms, and major commercial banks, sensitive life information such as name and contact information, as well as income, loan limit, and consultation and treatment information, is being exposed. In particular, recent banking incidents have raised the possibility of attacks using artificial intelligence, raising concerns that if leaked information is combined, it could lead to secondary damage such as customized phishing.

연합뉴스
2 min read
More on this topic오픈AI