K-EXAONE 2.0 Technical Report

Hugging Face Daily Papers Papers

Summary

K-EXAONE 2.0 is an open-weight multilingual MoE foundation model from LG AI Research with 750B total parameters and 37B active, supporting 10 languages and 256K context, with notable gains in agentic coding, long-context understanding, and safety.

This technical report presents K-EXAONE 2.0, an open-weight multilingual foundation model developed by LG AI Research as a step in our effort toward global frontier-scale foundation models. Rather than training from scratch, we upcycle K-EXAONE and expand its architecture, yielding a Mixture-of-Experts (MoE) model with 750B total parameters and approximately 37B activated per token---more than three times the capacity of its predecessor. K-EXAONE 2.0 supports context lengths of up to 256K tokens and expands multilingual coverage from six to ten languages. Its training pipeline combines continual pre-training, difficulty-focused mid-training, and post-training to strengthen reasoning, agentic coding, multilingual capability, and safety grounded in Korean sociocultural contexts. Across nine evaluation categories selected to reflect the conditions of practical use, K-EXAONE 2.0 improves over K-EXAONE and remains competitive with open-weight models, showing its largest gains in agentic coding and long-context understanding and its clearest strengths in long-context retrieval and safety. Released under the Apache 2.0 license, K-EXAONE 2.0 enables the wider AI ecosystem to evaluate, deploy, adapt, and build upon it, while marking the beginning---rather than the endpoint---of our challenge toward the global frontier.
Original Article
View Cached Full Text

Cached at: 08/06/26, 05:49 AM

Paper page - K-EXAONE 2.0 Technical Report

Source: https://huggingface.co/papers/2608.04505 Authors:

,

,

,

,

,

,

,

,

,

,

,

,

,

,

,

,

,

,

,

,

Abstract

ThistechnicalreportpresentsK-EXAONE2.0,anopen-weightmultilingualfoundationmodeldevelopedbyLGAIResearchasastepinourefforttowardglobalfrontier-scalefoundationmodels.Ratherthantrainingfromscratch,weupcycleK-EXAONEandexpanditsarchitecture,yieldingaMixture-of-Experts(MoE)modelwith750Btotalparametersandapproximately37Bactivatedpertoken---morethanthreetimesthecapacityofitspredecessor.K-EXAONE2.0supportscontextlengthsofupto256Ktokensandexpandsmultilingualcoveragefromsixtotenlanguages.Itstrainingpipelinecombinescontinualpre-training,difficulty-focusedmid-training,andpost-trainingtostrengthenreasoning,agenticcoding,multilingualcapability,andsafetygroundedinKoreansocioculturalcontexts.Acrossnineevaluationcategoriesselectedtoreflecttheconditionsofpracticaluse,K-EXAONE2.0improvesoverK-EXAONEandremainscompetitivewithopen-weightmodels,showingitslargestgainsinagenticcodingandlong-contextunderstandinganditscleareststrengthsinlong-contextretrievalandsafety.ReleasedundertheApache2.0license,K-EXAONE2.0enablesthewiderAIecosystemtoevaluate,deploy,adapt,andbuilduponit,whilemarkingthebeginning---ratherthantheendpoint---ofourchallengetowardtheglobalfrontier.

View arXiv pageView PDFAdd to collection

Get this paper in your agent:

hf papers read 2608\.04505

Don’t have the latest CLI?curl \-LsSf https://hf\.co/cli/install\.sh \| bash

Models citing this paper0

No model linking this paper

Cite arxiv.org/abs/2608.04505 in a model README.md to link it from this page.

Datasets citing this paper0

No dataset linking this paper

Cite arxiv.org/abs/2608.04505 in a dataset README.md to link it from this page.

Spaces citing this paper0

No Space linking this paper

Cite arxiv.org/abs/2608.04505 in a Space README.md to link it from this page.

Collections including this paper0

No Collection including this paper

Add this paper to acollectionto link it from this page.

Similar Articles

K-EXAONE 2.0 Technical Report

arXiv cs.CL

LG AI Research presents K-EXAONE 2.0, a 750B-parameter MoE foundation model upcycled from K-EXAONE, supporting 256K context and six languages, with self-speculative decoding for efficient inference.

LGAI-EXAONE/K-EXAONE-2.0-750B-A37B

Hugging Face Models Trending

LG AI Research introduces K-EXAONE 2.0, a frontier-scale multilingual MoE model with 750B total parameters (37B active), upcycled and trained for advanced reasoning, agentic workflows, and long-context understanding, released under Apache 2.0.

LG AI Research releases K-EXAONE 2.0 750B A37B

Reddit r/LocalLLaMA

LG AI Research has open-sourced the 750B parameter AI foundation model 'K-EXAONE 2.0' on Hugging Face. It is available for commercial use under the Apache 2.0 license, supports multiple languages, and shows performance comparable to global leading models.

XingChen-AGI/Xing4.0-29B-A4B MoE

Reddit r/LocalLLaMA

Xing4.0-29B-A4B is a next-generation MoE large language model developed by China Telecom, featuring 29B total parameters with 4B active per token, native support for 256K context length, and optimization for Ascend NPU with agent-oriented architecture for complex engineering tasks.