
일본 고베대 연구팀이 트롤리 딜레마 실험에서 챗GPT의 논리적 반론을 들은 참가자 30% 이상이 판단을 바꾼 결과를 발표하며, AI가 사람의 가치관과 판단에 무시할 수 없는 영향을 미칠 수 있다고 경고했다.
AI-generated summary
트롤리 딜레마는 브레이크가 고장 난 트롤리의 상황을 제시하고 다수를 구하기 위해 소수를 희생할 수 있는지의 도덕적 판단을 구하는 사고 실험으로, 마이클 샌델의 책 '정의란 무엇인가'에 등장한다.
'트롤리 딜레마' 제시 후 챗GPT가 설득…"AI가 사람들의 가치관에도 영향"
(서울=연합뉴스) 김병규 기자 = 인공지능(AI)이 사람의 판단을 바꾸는데 실제 사람보다 영향력이 크다는 연구 결과가 일본 연구진에게서 나왔다.
23일 도쿄신문에 따르면 마쓰모토 고헤이 고베대 교수 연구팀은 정답이 없는 도덕적인 문제와 관련해 AI에게 반론을 받게 될 경우 사람의 30% 이상이 판단을 번복한다는 연구 결과를 담은 논문을 발표했다.
연구팀은 19∼28세 청년층 56명, 65세 이상 고령층 74명에 대해 이른바 '트롤리의 딜레마'에 대한 판단을 구했다.
트롤리의 딜레마는 브레이크가 고장 난 트롤리의 상황을 제시하고 다수를 구하기 위해 소수를 희생할 수 있는지 판단을 구하는 것으로, 세계적인 베스트셀러 마이클 샌델의 '정의란 무엇인가'에 등장한다.
연구팀은 브레이크가 고장 나 폭주하는 트롤리(탄광수레)의 선로 앞에는 5명의 사람이 있고 선로를 바꿀 경우 그 선로에는 1명이 있는데, 어떤 선택을 할 것인가를 물었다.
연구 참가자들이 제출한 답에 대해 생성형 AI 챗GPT가 논리적으로 반론을 하게 했는데, 참가자들의 30% 이상이 반론을 읽고 자신의 판단을 뒤집었다. 참가자들에게 다른 문제를 제시한 뒤 같은 실험을 했을 때도 판단을 바꾼 사람의 비율은 비슷했다.
판단에 자신이 없을수록, 젊은 층보다 고령층에서 판단을 바꾸는 경우가 많았다.
마쓰모토 교수는 "사람이 다른 사람의 조언을 받아들이는 비율은 20∼30% 수준인데 이번 연구에서는 AI의 설득으로 30%나 되는 참가자들이 판단을 정반대로 바꿨다"며 "AI가 도덕적인 판단에서 무시할 수 없는 영향을 미칠 수 있다는 것을 의미한다"고 설명했다.
연구팀은 "AI가 사람의 가치관과 판단까지 영향을 미칠 수 있다"며 "허위 정보의 확산과 여론공작 등에 악용될 우려가 있다"라고도 경고했다.
실제로 최근 스위스 로잔 연방공대 연구팀이 미국 18세 이상 900명을 AI와 섞어 찬반이 나뉘는 주제에 대해 토론을 하게 하는 실험을 했는데, 참가자의 성별, 연령 등 개인정보를 알고 있던 AI가 사람보다 참가자들의 판단을 바꾸는데 우위에 있었다.
고베대 연구팀은 "SNS의 투고 이력 정보를 갖게 되면 AI의 설득력이 더욱 높아질 가능성이 있으므로 악용방지책을 마련할 필요가 있다"고 강조했다.
AI outlook — possibilities, not facts
AI가 개인 맞춤형 persuasion에 활용될 경우 가치관 조작 우려가 현실화될 것이다
Possible · Within months

An MIT study found that 95% of generative AI pilot projects failed to achieve tangible bottom-line improvements, and Gartner predicted that 30% of employees replaced by AI will be rehired by 2029. Experts advised that AI should be designed as a collaboration partner, not just a cost-saving tool.

OpenAI announced that it plans to allow external organizations to conduct technical safety assessments throughout the entire process of training, evaluation, and deployment of AI models. This is a step forward from the previous pre-launch safety assessment, with independence, scientific rigor, security practices, and accountability presented as key priorities. Competitor Antropic is also introducing similar safety measures.

Tajikistan's AI Committee vice-chair Najima Noyovtova said in an interview held in Seoul that a Central Asian AI base can be built by combining Korea's technological prowess with Tajikistan's clean energy, and proposed expanding cooperation between the two countries through the Korea-Thailand IT park project and the AI free zone initiative being promoted by KOICA.

Tom Siegel, former vice president of trust and safety at Google, took office as executive director of the Youth AI Safety Research Institute, warning that AI is more harmful to children and adolescents than social media, and presented suicide, mental illness, and cognitive outsourcing as examples of serious harm.

Antropic and OpenAI unveiled 'Claude Opus 5.5' and 'GPT-6 Sol·Luna' respectively, competing for cost-effective models that maintain or improve performance and lower costs. This is analyzed as a strategy to defend market share and increase sales ahead of the IPO while insisting on speed control.

Kakao Map collaborated with Busan City and the Busan Festival Organizing Committee to release a map dedicated to the 2026 Busan International Rock Festival. Toss received the Prime Minister's Award for its contribution to youth policy, Tada officially launched a regular reservation service, and Thanks Carbon and LG Chem announced the results of the Blue Forest Project at the Korea Social Value Festa 2026.