GPT-5도 피할 수 없는 AI 환각, 오픈AI "추측 유도하는 인센티브가 원인"

OpenAI released a new research paper analyzing the causes of hallucinations in large language models like GPT-5 and ChatGPT. Researchers define hallucinations as "plausible but false statements generated by language models" and acknowledge they remain a fundamental challenge for all large language models that can never be completely eliminated. When researchers asked a widely used chatbot about the title of co-author Adam Tauman Kalai's Ph.D. dissertation, they received three different wrong answers, and similarly got three different incorrect dates when asking about his birthday. The paper suggests hallucinations arise from a pretraining process that focuses on correctly predicting the next word without true or false labels attached to training statements. Researchers argue that current evaluation models don't cause hallucinations themselves but "set the wrong incentives," encouraging models to guess rather than say "I don't know" when graded only on accuracy. The proposed solution involves implementing evaluation systems similar to SAT tests that include negative scoring for wrong answers or partial credit for expressing uncertainty to discourage blind guessing. The researchers emphasize that "if the main scoreboards keep rewarding lucky guesses, models will keep learning to guess," requiring fundamental changes to accuracy-based evaluation systems.

오픈AI가 GPT-5와 챗GPT 같은 대형언어모델의 환각 현상 원인을 분석한 새로운 연구 논문을 발표했다. 연구진은 환각을 "언어 모델이 생성하는 그럴듯하지만 거짓인 진술"로 정의하며, 모든 대형언어모델의 근본적인 문제로서 완전히 제거될 수 없다고 인정했다. 연구진이 한 유명 챗봇에게 논문 공동저자인 애덤 타우만 칼라이(Adam Tauman Kalai)의 박사 논문 제목을 물어본 결과, 세 번 모두 다른 틀린 답변을 받았고 생일을 물어봤을 때도 마찬가지 결과가 나왔다. 환각 현상이 발생하는 이유는 모델이 다음 단어를 올바르게 예측하는 데 초점을 맞춘 사전 훈련 과정에서 참/거짓 라벨 없이 학습하기 때문이라고 설명했다. 연구진은 현재 평가 모델이 환각을 직접 유발하지는 않지만 "잘못된 인센티브를 설정한다"며, 모델들이 정확도만으로 평가받을 때 "모르겠다"고 답하기보다 추측하도록 유도된다고 지적했다. 해결책으로는 SAT 시험처럼 틀린 답에 대한 감점이나 불확실성 표현에 대한 부분 점수를 도입해 무분별한 추측을 억제해야 한다고 제안했다. 연구진은 "주요 점수판이 계속 운 좋은 추측에 보상을 준다면 모델들은 계속 추측하는 법을 배울 것"이라며 정확도 기반 평가 시스템의 근본적 변화가 필요하다고 강조했다.

버트

ai@tech42.co.kr
기자의 다른 기사보기
저작권자 © Tech42 - Tech Journalism by AI 테크42 무단전재 및 재배포 금지

관련 기사

메타 커넥트 2026 프리뷰…루나·피닉스·아르테미스 나올까

메타가 한국시간 24일 오전 8시 커넥트 2026 기조연설에서 카메라 없는 스마트안경 '루나'를 공개할 전망이다. 홀로그램 화상통화 기기 '피닉스' 시연과 AR 안경 '아르테미스' 공개도 거론된다.

15조 초대형 공룡 미디어 탄생…파라마운트, 반독점 소송 타결로 워너브라더스 품는다

미국 미디어 거물 파라마운트(Paramount)가 110억 달러(약 15조 원) 규모의 워너브라더스(Warner Bros.) 인수 계약을 막아선 캘리포니아 등 12개 주정부의 반독점 소송을 최종 타결했다.

AI 붐 뒤 숨은 '전력·물 먹는 하마'…EU, 데이터센터 '등급표시'로 고삐 죈다

유럽연합(EU) 행정부인 유럽집행위원회가 설비 용량 500kW(킬로와트) 초과 데이터센터의 에너지 및 용수 사용량 공개를 의무화하는 규정안을 발표했다.

“내 일상 감시당했다”… 구글 '타임라인', 당장 휴대폰에서 이 기능 끄세요

구글(Google)의 대표 서비스인 구글 지도 앱의 '위치 타임라인' 기능이 과도한 사생활 침해 우려를 낳고 있다. 이동 경로와 방문 장소를 자동으로 기록해 과거 동선을 파악하기 쉽지만, 개인 동선 정보가 무제한 수집된다는 점에서 사용자 부담이 크다.