Startup Motif Tops Korea's AI Model Race in Global Benchmark

Motif 3 Scores 47, Solar 37 in Latest Artificial Analysis Ranking Critics Note Test Excludes Korean and Risks Benchmaxxing

Technology|
|
By Kim Ki-hyuk
||
Lim Jung-hwan, CEO of Motif Technologies. Photo courtesy of Motif Technologies - Seoul Economic Daily Technology News from South Korea
Lim Jung-hwan, CEO of Motif Technologies. Photo courtesy of Motif Technologies

A startup has overtaken larger companies to rank first on a global performance index in the second-round evaluation of South Korea's government-backed project to build homegrown AI foundation models. Still, some caution that the result is difficult to treat as final, since it covers only part of the overall assessment and does not reflect Korean-language performance.

Motif Technologies' "Motif 3" scored 47 on the latest Artificial Analysis Intelligence Index (AAII), the highest among the four companies participating in the homegrown foundation model project, according to Artificial Analysis, an overseas AI performance evaluation firm, on the 13th. Upstage's "Solar Open2" followed with 37, ahead of SK Telecom's "A.X-K2" at 35 and the LG AI Research's "K-EXAONE 2.0" at 31. Artificial Analysis is a global AI analysis firm that comprehensively evaluates AI models across a range of capabilities, including mathematics, science, coding and reasoning.

Motif 3 is a large language model with 314 billion parameters that Motif Technologies developed entirely from scratch, from pretraining through post-processing, in about five months. "In the AI model market, borders are meaningless, and only demonstrating world-class performance gives a homegrown foundation model its true meaning," CEO Lim Jung-hwan said. "Building on the confidence we have gained this time, we will complete a top-tier AI model that can compete on equal footing with frontier models from the United States and China."

The AAII score does not represent the final result of the second-round evaluation, however. In that evaluation, benchmark testing accounts for 40 of a possible 100 points, of which the AAII contributes 25 and a benchmark assessment by the National Information Society Agency (NIA) contributes 15. The remaining points come from expert evaluation (35 points) and user evaluation (25 points).

Industry officials cite the absence of Korean-language benchmarks from the AAII as a limitation. There is also debate over the reliability of benchmark scores themselves, as so-called "benchmaxxing" — intensively boosting scores through data and post-training aimed at specific benchmarks — spreads across the global AI industry.

The government plans to announce the results of the second-round evaluation soon. In the final assessment, three of the four participating teams will survive to advance to the next stage.

Original reporting by Kim Ki-hyuk for Seoul Economic Daily.

AI-translated from Korean. Quotes from foreign sources are based on Korean-language reports and may not reflect exact original wording.

Watch · Seoul Economic Daily

More →
3:12

AI KEY

Preview
Korean Corporate Intelligence HubKOSPI · KOSDAQ · 12 sectors

A live, cap-weighted view of every KOSPI and KOSDAQ sector, with same-day Korean reporting distilled by company — built for foreign investors, correspondents and analysts who need to scan Korea before the next session.

Korea Company Atlas

Preview
Market Ontology · The Feedback LoopKFTC 2025 · 92 groups · 121,954 articles

An English ontology of the Korean market — how companies, the media, the government and the National Assembly move each other in a loop. Korea's named controlling persons and designated business groups are a mechanism, not a risk to be priced blind.

SIGNAL

Now live
English Edition · Capital MarketsM&A · IPO · PE · Fund Flows

SIGNAL English Edition is live — Korea's deal desk reporting in English. M&A, IPOs, private equity and fund flows, covered daily for global institutional investors. Browse free; subscriber-only scoops at the 50% intro rate.