• Skip to navigation
  • Skip to content

logo

  • 회사소개

    기술과 서비스로
    세상의 모든 가능성을 연결합니다

    • 소개
    • 주요 계열사
    • 주요 연혁
    • 브랜드 리소스
    • 제휴 제안
    • Contact
    NAVER 사옥
  • 서비스

    검색에서 탐색으로 진화
    On-Service AI

    • 네이버 주요 서비스
    • 포털
    • 도구
    • 검색
    • 광고
    • 커머스
    • 클라우드
    • 핀테크
    • 1784
    • 콘텐츠
    • 데이터센터 각
    • 커뮤니티
    • 전체 서비스
    • 지도
    NAVER 서비스
  • 기술

    혁신의 기술을 일상의 서비스로
    Everyday Tech

    • 네이버 주요 기술
    • HyperCLOVA X
    • 공간지능
    • 로보틱스
    • 이머시브 미디어
    NAVER 기술
  • 지속가능성

    네이버의 연결로 만드는
    더 나은 변화

    • 네이버 지속가능성
    • 지속가능경영
    • Tech for People
    • Social
    • Environment
    • Principle
    • ESG 자료실
    NAVER 지속가능성
  • 투자정보

    투자정보

    • IR 뉴스
    • 기업지배구조
    • IR 일정
    • 재무정보
    • 재무정보
    • IR 자료실
    NAVER 투자정보
  • 미디어

    미디어

    • 보도자료
    • 미디어 행사
    • 네이버 리포트
    • AI in NAVER
    NAVER 뉴스룸
  • 스토리

    네이버 스토리

    스토리 전체 보기
    • Leader's View 2026.07.24
      이해진 의장, 엔비디아·브룩필드와 AI 팩토리 공동 구축을 위한 초대형 협력 소식 발표
    • Tech 2026.07.23
      일본으로 간 아크(ARC) – 네이버 1784 스마트빌딩 모델의 첫 글로벌 레퍼런스
  • 채용
통합검색 입력 폼
  • 한눈에 보는 네이버 전체 서비스 소개

    한눈에 보는 네이버
    전체 서비스 소개

  • 네이버 로고 아이덴티티 브랜드 리소스

    네이버 로고 아이덴티티
    브랜드 리소스

  • 5,400만+ 유저를 고객으로 네이버 광고 검색 상품

    5,400만+ 유저를 고객으로
    네이버 광고 검색 상품

  • NAVER Auunal Report ESG Library

    한눈에 보는 네이버
    전체 서비스 소개

  • NAVER Brand Resource Logo and color

    네이버 로고 아이덴티티
    브랜드 리소스

  • NAVER MAP Connecting online and offline

    5,400만+ 유저를 고객으로
    네이버 광고 검색 상품

logo
logo
  • 회사소개
    • 소개
    • 주요 계열사
    • 주요 연혁
    • 브랜드 리소스
    • 제휴 제안
    • Contact
  • 서비스
    • 네이버 주요 서비스
    • 포털
    • 도구
    • 검색
    • 광고
    • 커머스
    • 클라우드
    • 핀테크
    • 1784
    • 콘텐츠
    • 데이터센터 각
    • 커뮤니티
    • 전체 서비스
    • 지도
  • 기술
    • 네이버 주요 기술
    • HyperCLOVA X
    • 공간지능
    • 로보틱스
    • 이머시브 미디어
  • 지속가능성
    • 네이버 지속가능성
    • 지속가능경영
    • Tech for People
    • Social
    • Environment
    • Principle
    • ESG 자료실
  • 투자정보
    • IR 뉴스
    • 기업지배구조
    • IR 일정
    • 재무정보
    • 재무정보
    • IR 자료실
  • 미디어
    • 보도자료
    • 미디어 행사
    • 네이버 리포트
    • AI in NAVER
  • 스토리
  • 채용
Tech

NAVER Cloud Unveils Omnimodal HyperCLOVA X, Introducing a Step-by-Step Expansion Strategy Built from the Ground Up

2025.12.29
공유하기

NAVER Cloud Unveils Omnimodal HyperCLOVA X, Introducing a Step-by-Step Expansion Strategy Built from the Ground Up

공유하기

NAVER Cloud Unveils Omnimodal HyperCLOVA X, Introducing a Step-by-Step Expansion Strategy Built from the Ground Up

- Two models released—an omnimodal model and a reasoning model—to accelerate the development of real-world AI agents

- Omnimodal architecture refined from the basics, with a focus on data differentiation, phased scale-up, and specialized model production

- Omnimodal AI highlighted as a next-generation foundation technology for understanding the real world

December 29, 2025

NAVER Cloud (CEO Kim Yu-won) has unveiled the first outcome of its “Omni Foundation Model” development project, which it is leading under the Ministry of Science and ICT’s “Independent AI Foundation Model” initiative. NAVER Cloud announced the open-source release of the nation’s first foundation model to apply a native omnimodal structure, “Native Omni Model (HyperCLOVA X SEED 8B Omni),” and a high-performance reasoning model, “HyperCLOVA X SEED 32B Think,” which enhances conventional reasoning-based AI with capabilities in vision, voice, and tool use. With these releases, the company is accelerating the development of AI agents that can be readily integrated into daily life and industrial settings.

Laying the groundwork for a “future technology” omni model, accelerating the transition to daily and industrial AI through data differentiation, phased scale-up, and specialized model production

The newly released “HyperCLOVA X SEED 8B Omni” fully adopts a native omnimodal structure, in which different forms of data—such as text, images, and audio—are learned together from the outset in a single model. Omnimodal AI can contextually integrate information across a shared semantic space, regardless of its form. This enables high applicability in real-world environments where speech, text, visual, and audio information interact simultaneously, making it a next-generation AI technology attracting growing attention. Because of this capability, global big tech companies are also positioning omnimodal AI as a core technology pillar in their next-generation foundation model strategies.

To maximize the potential of omnimodal AI, NAVER Cloud is adopting a strategy that extends beyond conventional training on Internet documents or image-based data, focusing on acquiring data that reflects diverse real-world contexts. Sung Nako, Executive Director of Hyperscale AI at NAVER Cloud, stated, “Even if you scale up a model, if the diversity of the data is limited, the AI’s problem-solving ability will inevitably be confined to specific domains or subjects.” He added, “That’s why the process of securing and refining differentiated real-world data, such as undigitized contextual data from daily life or spatial data reflecting regional geographic features, must come first.”

NAVER Cloud plans to begin a phased scale-up by training the model with differentiated data, having now validated its native omnimodal AI development methodology through this release. Unlike traditional multimodal approaches, which combine separate models for text, images, and speech, omnimodal AI features a single-model architecture, making it easier to scale. Building on this structure, the company aims to efficiently expand specialized omnimodal models in various sizes to support services closely integrated into both industry and daily life.

The model also features an omnimodal generation capability, enabling it to generate and edit images based on text prompts. By understanding the context of both text and images, the model produces output that reflects intended meaning, enabling natural execution of text comprehension and image generation/editing within a single model. This functionality, previously offered by global frontier AI models, demonstrates that NAVER Cloud has now achieved comparable multimodal generation capabilities.

​

[Image description] The HyperCLOVA X SEED 8B Omni model generates output by understanding the context of both text and images

​

Combining vision, voice, and tool capabilities with reasoning-based AI to develop omnimodal agents on par with global models

NAVER Cloud has also released “HyperCLOVA X SEED 32B Think” to validate the practical applicability and future potential of omnimodal AI agents. This model combines its reasoning-based AI with capabilities in visual understanding, voice interaction, and tool use to deliver an agent experience capable of understanding complex inputs and requests and solving problems.

According to benchmarking by global AI evaluation agency Artificial Analysis, the model demonstrated a performance range comparable to leading global AI models, based on a composite index covering 10 key benchmarks, including general knowledge, advanced reasoning, coding, and agentic tasks.

In category-specific evaluations, the model showed particular strength in areas closely tied to real-world use. It demonstrated superior performance compared to global models in key capabilities such as General Knowledge (Korean Text), Vision Understanding, and Agentic Task (tool-based problem-solving as an agent), proving its competency in handling complex tasks.

​

[Image description] Category-specific benchmark scores of “HyperCLOVA X SEED 32B Think”

In addition, when applied to this year’s College Scholastic Ability Test (CSAT), the model achieved Grade 1 (top-tier) scores across all major subjects, including Korean, mathematics, English, and Korean history, and earned perfect scores in English and Korean history. The company noted that, unlike many AI models that require converting exam questions into text before input, this model directly understood and solved problems from image inputs, marking a key point of differentiation.

Sung added, “We confirmed that expanding AI’s sensory capabilities horizontally—across text, vision, and audio—while simultaneously enhancing its reasoning and problem-solving skills significantly improves its ability to address real-world challenges.” He continued, “We plan to continue scaling based on this robust foundational structure, believing that gradual expansion is the way to develop not just a large-scale model, but one that is truly practical and usable.”

NAVER Cloud plans to gradually expand the deployment of AI agents based on this omnimodal HyperCLOVA X in various domains, including search, commerce, content, public services, and industrial applications, accelerating the creation of a technology ecosystem that enables “AI for everyone.” (End)

​

AI AgentHyperCLOVA XNAVER CloudOmni Foundation ModelOmnimodal
전체 이미지 다운로드
목록보기
We the Navigators
  • 파트너 지원
    • 네이버 광고센터 새창 열림
    • 스마트스토어 새창 열림
    • 스마트플레이스 새창 열림
    • 비즈니스 스쿨 새창 열림
    • 네이버 임팩트 새창 열림
    • SME 풀케어
  • 개발자 지원
    • 네이버 개발자 센터 새창 열림
    • 오픈 API 새창 열림
    • 오픈소스 새창 열림
    • 네이버 D2 새창 열림
    • 네이버 D2SF 새창 열림
  • 자료실
    • IR 자료실
    • ESG 자료실
    • 네이버 리포트
    • 브랜드 리소스
  • 주요 계열사
    • 네이버클라우드
    • 스노우
    • 네이버랩스
    • 네이버웹툰
    • 네이버파이낸셜
  • blog link
  • naverTV link
  • instagram link
  • youtube link
  • ffinicial link
  • Contact
  • 제휴 제안
  • 고객센터
  • 기업윤리 상담센터 기업윤리 상담센터
    • 네이버 주식회사
    • 비즈니스 파트너
  • 이용약관
  • 운영정책
  • Contact
  • 제휴 제안
  • 기업윤리 상담센터

©NAVER CORP.