Garp Independent AI & technology journalism
Saturday, August 8, 2026 Sign In · Join Subscribe
Latest Naïve raises $28.5M to automate the grunt work of setting up and running a company

AI news, research, models, robotics, chips, startups, and infrastructure coverage.

Updated daily

Home  /  AI News  /  Kimi’s open model K3 nears GPT-5.6 Sol and Fable 5 while signaling the end of super cheap Chinese AI

AI News

Kimi’s open model K3 nears GPT-5.6 Sol and Fable 5 while signaling the end of super cheap Chinese AI

Kimi’s open model K3 nears GPT-5.6 Sol and Fable 5 while signaling…

Kimi is launching K3, a multimodal model with 2.8 trillion parameters and a context window of one million tokens. In the company’s own benchmarks, it performs on par with leading proprietary models.

According to Kimi, the new flagship model K3 has 2.8 trillion total parameters, processes images and video natively, and supports a context window of one million tokens. Kimi calls K3 the first open model in the roughly 3 trillion parameter range. Full model weights are scheduled for release by July 27. The model targets long-running programming tasks, knowledge work, and complex reasoning. In Kimi’s own benchmarks, K3 still trails the top proprietary models Claude Fable 5 and GPT 5.6 Sol but beats every other system tested, including the Claude Opus models and Chinese rival GLM-5.2. All results come from Kimi and were achieved at maximum or high thinking intensity, according to the company.Ad Across all 35 tests, K3 took first place about seven times and landed second or third in most of the rest. Fable 5 won the most individual tests. In nearly every benchmark, K3 beat Opus 4.8, GPT 5.5, and GLM 5.2 by a wide margin. Depending on the benchmark, one of three agent systems was used: KimiCode, Claude Code, or Codex. That means the results weren’t all collected under identical conditions.AdDEC_D_Incontent-1 Artificial Analysis confirms K3’s strong performance but flags higher hallucination rate Independent testing lab Artificial Analysis has published its first evaluation of Kimi K3. The model scores 57 on the Artificial Analysis Intelligence Index, putting it on par with Opus 4.8 and GPT-5.5 but still behind Fable 5 and GPT-5.6 Sol. That largely lines up with Kimi’s own claims. On agentic tasks, K3 reaches an Elo rating of 1,668 on GDPval v2, a big jump from K2.6’s 1,190. It beats GLM-5.2 (1,514), GPT-5.5 (1,494), and Claude Opus 4.8 (1,600), though it still falls short of Claude Fable 5 (1,760). K3 also takes the top spot on AutomationBench-AA, Artificial Analysis’s version of Zapier’s agentic SaaS workflow evaluation, with a score of 53 percent.Ad On AA-Briefcase, a private long-horizon knowledge work evaluation, K3 reaches an overall Elo of 1,547, up 732 points from K2.6.