얼굴만 보고 연쇄살인범까지 찍었다…사람보다 ‘외모’ 더 따지는 AI [후암동 논문 연구소] 작성일 08-19 35 목록 <div id="layerTranslateNotice" style="display:none;"></div> <div class="article_view" data-translation-body="true" data-tiara-layer="article_body" data-tiara-action-name="본문이미지확대_클릭"> <section dmcf-sid="u6hwZqAiG5"> <figure class="figure_frm origin_fig" contents-hash="fe886f4283406a5a5769acb32b892dc66e0510902ea4c55c45f5534841174ec3" dmcf-pid="7Plr5BcnHZ" dmcf-ptype="figure"> <p class="link_figure"><img alt="사진은 기사 내용과 관련 없음. [게티이미지뱅크]" class="thumb_g_article" data-org-src="https://t1.daumcdn.net/news/202608/19/ned/20260819201218718gwcx.jpg" data-org-width="1280" dmcf-mid="ttlr5BcnYp" dmcf-mtype="image" height="auto" src="https://img1.daumcdn.net/thumb/R658x0.q70/?fname=https://t1.daumcdn.net/news/202608/19/ned/20260819201218718gwcx.jpg" width="658"></p> <figcaption class="txt_caption default_figure"> 사진은 기사 내용과 관련 없음. [게티이미지뱅크] </figcaption> </figure> <p contents-hash="3cc45aab50e3b2290cd0d793e830395baa73a563f177484beadfd959ee6123c4" dmcf-pid="zQSm1bkL1X" dmcf-ptype="general">[헤럴드경제=장윤우 기자] 인공지능(AI)에 얼굴 사진 두 장을 보여주고 누가 더 유능해 보이는지 묻자 사람과 같은 답을 골랐다. 사람의 이런 판단은 근거 없는 편견으로 알려져 있는데 AI가 이를 그대로 따라 한 것이다.</p> <p contents-hash="0b445ac8c1837667628538fa22b06d316894266b07932c6042179dfda9fc7179" dmcf-pid="qxvstKEoZH" dmcf-ptype="general">AI는 인상만 보고 누가 연쇄살인범일지, 누구를 대학 총장으로 뽑을지까지 답했다. 최신 모델일수록 이런 경향이 강했다.</p> <p contents-hash="259bf06f422b514d25d02bf2f02f6173f3fcf4c0bfe6c34270c25a65515ddfc3" dmcf-pid="BMTOF9DgXG" dmcf-ptype="general">최근 국제학술지 PNAS 넥서스(PNAS Nexus) 제5권 제8호에 미국 하버드대 심리학과 마자린 바나지 교수와 캔그레이드의 스티븐 레어 연구원 연구팀은 이 같은 내용을 담은 연구 결과를 발표했다.</p> <figure class="figure_frm origin_fig" contents-hash="d6a9db205c284aa7238c50fa18822fed25081bfeb48a0993229596224beca0ca" dmcf-pid="bRyI32waHY" dmcf-ptype="figure"> <p class="link_figure"><img alt="사진은 기사 내용과 관련 없음. [게티이미지뱅크]" class="thumb_g_article" data-org-src="https://t1.daumcdn.net/news/202608/19/ned/20260819201218994fzef.jpg" data-org-width="1280" dmcf-mid="F4HSu8OcG0" dmcf-mtype="image" height="auto" src="https://img4.daumcdn.net/thumb/R658x0.q70/?fname=https://t1.daumcdn.net/news/202608/19/ned/20260819201218994fzef.jpg" width="658"></p> <figcaption class="txt_caption default_figure"> 사진은 기사 내용과 관련 없음. [게티이미지뱅크] </figcaption> </figure> <div contents-hash="c948b4a959785843c336907498e174cb6c7ee5a1e0b25d8abef1f203ed1a1dda" dmcf-pid="K1iQEvfzXW" dmcf-ptype="general"> 사람보다 더 심하게 ‘외모’ 따졌다 </div> <p contents-hash="f147d5a0e6b306ecc7fb361ebb231f93038bae98c110f6a6a6fc596867e1bb2d" dmcf-pid="9tnxDT4qty" dmcf-ptype="general">사람은 얼굴만 보고도 성격이나 품행 등을 판단하는 경향이 있다. 얼굴 인상에 따라 채용이나 형량에까지 영향을 미치지만 실제 성격과는 거의 관계가 없는 것으로 밝혀져 있다.</p> <p contents-hash="14fb341f7f4c4a40a637e3c74d7c5b129c414e210c18e1c0f270b963715dd442" dmcf-pid="2FLMwy8BYT" dmcf-ptype="general">연구팀은 AI 모델 중 하나인 GPT-4o에 얼굴 두 장을 보여주고 하나를 고르게 했다. 유능함을 묻자 사람이 고를 법한 얼굴을 고른 비율은 87.83%, 신뢰감을 묻자 72.67%로 나타났다.</p> <p contents-hash="1ae7f893342126393654a84f4758c083c3ef7984b7663e8d65f7d10ffc17526c" dmcf-pid="V3oRrW6b1v" dmcf-ptype="general">실험에 사용된 얼굴은 사람들이 수많은 얼굴을 평가한 뒤 신뢰감이나 유능함과 연결되는 생김새 특징을 기준되는 얼굴에 조금씩 더하거나 뺀 합성 이미지다.</p> <figure class="figure_frm origin_fig" contents-hash="d6affe8471b1ba72f08d3dc7d8d26926bf0322bfa2f0e2d0bba05eaa491a52e4" dmcf-pid="f0gemYPKHS" dmcf-ptype="figure"> <p class="link_figure"><img alt="실험에 사용된 합성 얼굴 사진. [국제학술지 PNAS 넥서스(PNAS Nexus) 제5권 제8호]" class="thumb_g_article" data-org-src="https://t1.daumcdn.net/news/202608/19/ned/20260819201219175nmoo.jpg" data-org-width="1280" dmcf-mid="3WVzdk0H53" dmcf-mtype="image" height="auto" src="https://img3.daumcdn.net/thumb/R658x0.q70/?fname=https://t1.daumcdn.net/news/202608/19/ned/20260819201219175nmoo.jpg" width="658"></p> <figcaption class="txt_caption default_figure"> 실험에 사용된 합성 얼굴 사진. [국제학술지 PNAS 넥서스(PNAS Nexus) 제5권 제8호] </figcaption> </figure> <p contents-hash="fceafb2df10a280c6eddb82d76a942ce676ce0656fa146ed9591a88d033cb140" dmcf-pid="4padsGQ9Gl" dmcf-ptype="general">연구팀은 얼굴에 특징을 넣은 정도를 달리해 서로 비슷해 보이는 얼굴부터 확연히 달라 보이는 얼굴까지 만들어 얼굴 두 개를 보여주는 실험도 진행했다.</p> <p contents-hash="373fad650186119fbee27d7dc43c841d436a1e8ebd301664893f6b8ad38c51c7" dmcf-pid="8UNJOHx2Hh" dmcf-ptype="general">서로 차이가 큰 얼굴을 보여주자 GPT-4o가 사람이 고를 법한 얼굴을 고른 비율은 유능함 98.00%, 신뢰감 97.00%로 올라갔다.</p> <p contents-hash="bf0a385ff93ff3fcfe7c92df0f7b00466a4f7a001b99f196e2d52ade12180e63" dmcf-pid="6ujiIXMV5C" dmcf-ptype="general">서로 별 차이가 없는 얼굴에서는 유능함을 선택하는 비율은 70.00%, 신뢰감은 53.00%로 내려갔다.</p> <p contents-hash="eecd2bf93f489809d1c46a2ee46314c4743667cf62f7f30daf89f5c293706818" dmcf-pid="PsbpxN1y1I" dmcf-ptype="general">연구팀은 사람이 같은 얼굴 자료를 평가한 기존 자료와 비교해 봤다. 사람들이 매긴 점수를 GPT-4o가 받은 것과 같은 방식으로 환산하자 사람이 유능해 보이는 쪽을 고를 비율은 62.65%로 추정됐다. GPT-4o는 87.83%였다.</p> <p contents-hash="c312bfa5aa49dd45fb0f7d8edc81aae93c3d64629584f133ed0e9633e7429829" dmcf-pid="QOKUMjtWXO" dmcf-ptype="general">연구팀은 AI가 학습 데이터에 있던 얼굴을 외운 것인지도 확인해 봤다. 실험에 쓴 얼굴을 하나씩 보여주며 무엇인지 아느냐고 물었지만 GPT-4o는 어느 것도 알아보지 못했다.</p> <figure class="figure_frm origin_fig" contents-hash="ea33bda805ebf470b425a49926c0d7a90eda13b13d3538c54ad584d5d27198ea" dmcf-pid="xI9uRAFYGs" dmcf-ptype="figure"> <p class="link_figure"><img alt="붉은털원숭이. [게티이미지뱅크]" class="thumb_g_article" data-org-src="https://t1.daumcdn.net/news/202608/19/ned/20260819201219379xnvo.jpg" data-org-width="1280" dmcf-mid="0IyI32waYF" dmcf-mtype="image" height="auto" src="https://img1.daumcdn.net/thumb/R658x0.q70/?fname=https://t1.daumcdn.net/news/202608/19/ned/20260819201219379xnvo.jpg" width="658"></p> <figcaption class="txt_caption default_figure"> 붉은털원숭이. [게티이미지뱅크] </figcaption> </figure> <p contents-hash="628b88a725e2d9cdd149950943042c5921c0a613a088199e9d55cd434e3197b7" dmcf-pid="yVscYUgRGm" dmcf-ptype="general">붉은털원숭이 사진으로도 실험했다. 사람들이 순해 보인다고 평가한 다섯 마리와 사나워 보인다고 평가한 다섯 마리를 짝지어 어느 쪽이 더 믿음직한지 물었다. GPT-4o는 66.00%에서 순한 쪽을 골랐다.</p> <p contents-hash="af06ee0d77307da896e0eec888b100878df7c0c65fbd04a11a4565fe1e838eef" dmcf-pid="WfOkGuaeGr" dmcf-ptype="general">원숭이 얼굴을 성격과 연결한 자료는 존재하지 않는다. 연구팀은 AI가 사람 얼굴에서 익힌 판단 기준을 다른 종에까지 적용했다는 뜻으로 해석했다.</p> <div contents-hash="0c20efe311594c0779700c1ec708fbe6d82265d8c062e603d87340850486d617" dmcf-pid="Y4IEH7Nd5w" dmcf-ptype="general"> “누가 연쇄살인범일까” 물어도 답했다 </div> <p contents-hash="f520cce20c455f61e33ebcc941f425d64aef21be8d0251b60b464c3e049aaabe" dmcf-pid="G8CDXzjJGD" dmcf-ptype="general">연구팀은 두 얼굴 중 누가 연쇄살인범일 가능성이 높은지, 누가 인신매매로 체포될 가능성이 높은지, 누가 다단계 사기를 칠 가능성이 높은지와 같이 답하기를 꺼릴 만한 질문도 던졌다.</p> <p contents-hash="d0008b17ca51f47855259af101e8f918d78ebe4c1c20a0918e507981a7bd5cbe" dmcf-pid="H6hwZqAi1E" dmcf-ptype="general">민감한 질문에도 답변을 거부하지 않은 GPT-4o는 사람들이 덜 믿음직하다고 평가한 쪽을 고른 비율이 68.70%였다.</p> <p contents-hash="888a28b61827d41faed641ef9e9463a737aad523dfd90a115b675e6a752ceb30" dmcf-pid="XPlr5BcnHk" dmcf-ptype="general">대학 총장으로 누구를 뽑을지, 어느 쪽 스타트업에 투자할지, 누구에게 퇴직연금 운용을 맡길지 물었을 때도 GPT-4o가 사람들이 유능해 보인다고 평가한 쪽을 고른 비율은 75.19%였다.</p> <figure class="figure_frm origin_fig" contents-hash="155e6c9f3f0f2e6b1526efdbd11c5cc6c3ef2215333b28292030c81386ec8737" dmcf-pid="Z2mAWpoMYc" dmcf-ptype="figure"> <p class="link_figure"><img alt="사진은 기사 내용과 관련 없음. [게티이미지뱅크]" class="thumb_g_article" data-org-src="https://t1.daumcdn.net/news/202608/19/ned/20260819201219592ytco.jpg" data-org-width="1280" dmcf-mid="p0e8Ah9Utt" dmcf-mtype="image" height="auto" src="https://img2.daumcdn.net/thumb/R658x0.q70/?fname=https://t1.daumcdn.net/news/202608/19/ned/20260819201219592ytco.jpg" width="658"></p> <figcaption class="txt_caption default_figure"> 사진은 기사 내용과 관련 없음. [게티이미지뱅크] </figcaption> </figure> <p contents-hash="8ee8e5cdbb4cc5275c4a9c2cd352e7592d8d542a2e899d8e51da7d808bad2729" dmcf-pid="5VscYUgRtA" dmcf-ptype="general">연구팀은 추론 능력이 뛰어난 AI 모델이라면 얼굴만 보고 판단하는 것이 잘못임을 알고 걸러낼 수 있다고 예상해 GPT-5와 제미나이 3 플래시 프리뷰, 클로드 소네트 4.5로도 같은 실험을 반복했다.</p> <p contents-hash="624aa909ccc286798180b32083285601fc9b8b326cb5aa671b857bee830374f1" dmcf-pid="1fOkGuae5j" dmcf-ptype="general">실험 결과, 유능함 판단에서 GPT-5는 94.33%, 제미나이 3은 95.67%로 GPT-4o의 87.83%보다 높았다.</p> <p contents-hash="e396e884dd9ad70e9190b690d4aa1ec75d1bd5219e80b7bbaf0072fe3e8f1a86" dmcf-pid="t4IEH7Nd5N" dmcf-ptype="general">연구팀은 성능이 좋은 모델일수록 편향까지 더 정밀하게 학습하기 때문일 수 있다고 해석했다. 모델이 발전할수록 이런 편향이 더 굳어질 가능성도 제기했다.</p> <figure class="figure_frm origin_fig" contents-hash="95580397b5e95f51a4735ecb609ce84b92f81f233bec0756d90a14318c4cfe08" dmcf-pid="F8CDXzjJZa" dmcf-ptype="figure"> <p class="link_figure"><img alt="[게티이미지뱅크]" class="thumb_g_article" data-org-src="https://t1.daumcdn.net/news/202608/19/ned/20260819201219803yswu.png" data-org-width="860" dmcf-mid="UiYhpfmj11" dmcf-mtype="image" height="auto" src="https://img3.daumcdn.net/thumb/R658x0.q70/?fname=https://t1.daumcdn.net/news/202608/19/ned/20260819201219803yswu.png" width="658"></p> <figcaption class="txt_caption default_figure"> [게티이미지뱅크] </figcaption> </figure> <p contents-hash="fb4d2b355bc3e425fc58038f46522296e5b46db6f8776871ba681125567ea0ce" dmcf-pid="36hwZqAitg" dmcf-ptype="general">AI가 왜 이런 판단을 내리는지는 밝혀지지 않았다. 연구팀은 학습한 글에 얼굴을 근거로 사람을 평가한 표현이 무수히 담겨 있었을 가능성을 가장 유력하게 봤다.</p> <p contents-hash="45ac925a11d9fbe02a13aa2a04406eafe17d760db5696f32e2f08046d6e74cd9" dmcf-pid="0Plr5BcnHo" dmcf-ptype="general">그러면서 AI 기업들이 인종이나 성별처럼 논란이 될 만한 영역에는 안전장치를 걸어뒀지만 인상 판단 같은 덜 알려진 영역에는 그러지 못했다고 연구팀은 지적했다.</p> <p contents-hash="dd2b727ea5c48d79911a68c0d36b74cdec031f02816debc3e57f6ec8aff2600c" dmcf-pid="pQSm1bkLGL" dmcf-ptype="general">다만 연구팀은 이번 실험이 판단이 어려울 때 직감을 따르라고 지시한 조건에서 이뤄졌다는 점을 인정했다. 실제 사용 환경에서도 같은 결과가 나오는지는 추가 연구가 필요하다고 밝혔다.</p> <div contents-hash="a1c4362aa430e39f95baccb1daf89a551acab8eddc991eab13cea5e4f4af098d" dmcf-pid="UxvstKEotn" dmcf-ptype="general"> 참고논문 </div> <p contents-hash="aa9b02bdfe70f8b9704a87b410a5176f46125bb6b68cdd3f240ce6b5fcb14036" dmcf-pid="uxvstKEoXi" dmcf-ptype="general">DOI : 10.1093/pnasnexus/pgag247</p> <p contents-hash="45e95832d1d1ca8715d04e1a3e263d9abd0388834c1b93e534c0331a4de66b31" dmcf-pid="7MTOF9DgZJ" dmcf-ptype="general">논문 정보 : Steven A Lehr, Yash Lothe, Mahzarin R Banaji, Like humans, language models demonstrate face-to-character biases, PNAS Nexus, Volume 5, Issue 8, August 2026, pgag247.</p> </section> </div> <p class="" data-translation="true">Copyright © 헤럴드경제. 무단전재 및 재배포 금지.</p> 관련자료 이전 룬드벡 'APB-A1' 타 적응증 탐색…"개발 중단 아니라도 부정적 신호" 08-19 다음 데이터센터 짓는 통신·포털…AI가 '라이벌'을 바꾸다 08-19 댓글 0 등록된 댓글이 없습니다. 로그인한 회원만 댓글 등록이 가능합니다.