Thursday, 27 August 2026

Aube.

News of progress
Original and translation

Kakao’s Kanana2 leads Google and Alibaba’s lightweight AI models in safety evaluation

Leave comparison

Both versions are aligned block by block, in reading order: title, the essentials, then paragraph by paragraph. Where the translation merged or split a paragraph, the matching cell stays empty — we never pair two passages by guesswork.

Original · Korean
카카오 카나나2, 구글·알리바바 경량 AI보다 안전성 평가 앞서
Translation · English
Kakao’s Kanana2 leads Google and Alibaba’s lightweight AI models in safety evaluation
Original · Korean
카나나2 1.3B와 3B 모델이 종합 안전성 평가에서 각각 1위를 기록했다.
Translation · English
The Kanana2 1.3B and 3B models ranked first in the overall safety evaluation.
Original · Korean
카카오는 한국어·한국 사회 맥락을 반영한 AssurAI로 9560개 항목을 측정했다.
Translation · English
Kakao used AssurAI, which reflects Korean-language and Korean social contexts, to measure 9,560 items.
Original · Korean
평가는 카카오 자체 플랫폼에서 LLM-as-a-Judge 방식으로 이뤄졌다.
Translation · English
The evaluation was conducted on Kakao’s own platform using an LLM-as-a-Judge approach.
Original · Korean

허깅페이스에 공개된 카카오의 경량 언어모델 두 개가 구글과 알리바바의 비슷한 크기 모델을 안전성 평가에서 앞섰다. Kanana-2-1.3B-Instruct와 Kanana-2-3B-Instruct 모두 종합 점수 0.70으로 1위를 기록했다. 카카오는 18일 이 결과를 공개했다.

Translation · English

Two lightweight language models from Kakao released on Hugging Face outperformed similarly sized models from Google and Alibaba in a safety evaluation. Both Kanana-2-1.3B-Instruct and Kanana-2-3B-Instruct ranked first with an overall score of 0.70. Kakao released the results on the 18th.

Original · Korean

비교 대상은 구글의 젬마와 알리바바의 큐웬이었다. 1.3B 모델 비교에서 젬마는 0.68, 큐웬은 0.58을 받았고, 3B 모델 비교에서는 젬마 0.66, 큐웬 0.62였다. 카나나2 1.3B는 범죄·불법 행위 영역에서, 3B는 성적 콘텐츠·아동보호와 권리 침해 영역에서 비교 모델보다 높은 점수를 기록했다.

Translation · English

The models used for comparison were Google’s Gemma and Alibaba’s Qwen. In the 1.3B model comparison, Gemma scored 0.68 and Qwen 0.58; in the 3B model comparison, Gemma scored 0.66 and Qwen 0.62. Kanana2 1.3B scored higher than the comparison models in the crime and illegal activity category, while the 3B model scored higher in the sexual content and child protection, and rights violations categories.

Original · Korean

시험대가 된 것은 AssurAI다. 한국정보통신기술협회, KAIST, 카카오가 지난해 11월 공동 구축한 이 한국어 특화 벤치마크는 한국의 사회·문화적 맥락을 반영해 35개 위험 범주와 9560개 평가 항목을 담고 있다. 텍스트·이미지·오디오 같은 멀티모달 조건과 악의적 프롬프트, 실제 AI 서비스 이용 상황까지 위험 요소를 살핀다.

Translation · English

The test was AssurAI. Jointly developed last November by the Telecommunications Technology Association, KAIST and Kakao, this Korean-language benchmark reflects Korea’s social and cultural context and contains 35 risk categories and 9,560 evaluation items. It examines risk factors across multimodal conditions such as text, images and audio, as well as malicious prompts and real-world AI service usage scenarios.

Original · Korean

카카오는 자체 구축한 AI 안전성 평가 플랫폼에서 사전에 정한 기준으로 모델의 답변을 채점했다. LLM-as-a-Judge, 즉 대규모 언어모델이 다른 모델의 응답을 평가하는 방식을 사용했다. 따라서 이번 결과는 카카오가 제시한 자체 평가 결과다.

Translation · English

Kakao scored the models’ responses against predefined criteria on its own AI safety evaluation platform. It used an LLM-as-a-Judge approach, in which a large language model evaluates another model’s responses. The results are therefore Kakao’s own evaluation findings.

Original · Korean

카카오는 앞으로 이 사전 검증을 자체 개발 모델 전반으로 확대하고, 텍스트 모델을 넘어 멀티모달 모델과 에이전틱 AI의 위험 평가에도 적용할 계획이다.

Translation · English

Kakao plans to expand this pre-release validation across its internally developed models and apply it beyond text models to risk assessments for multimodal models and agentic AI.

Back to the article