Moonshot AI

10 releases covered on ThursdAI · kimi.ai ↗

July 2026

Moonshot AI
New ModelsOpen weights

Kimi K3 open weights

Moonshot releases Kimi K3's full open-weight checkpoints — 2.8T parameters, the largest open model ever

Two weeks after the API launch, Moonshot published Kimi K3's full checkpoints, model code, and technical report: 2.8T total parameters with 104B active (16 of 896 experts), native vision, a 1M-token context window, and roughly 1.56TB of MXFP4 weights. The report details KDA linear attention, attention residuals, NoPE, and a claimed 2.5x scaling-efficiency jump over K2. On the show, Elie Bakouch called it public building blocks scaled superbly, and Baseten's Philip Kiely described serving it day-zero on eight GB300s. The custom license requires branding above 100M MAU or $20M monthly revenue and a signed agreement for model-as-a-service providers — every provider lists the identical $3/$15 price.

2.8T / 104B total / active parameters1.56TB MXFP4 weights — eight GB300s to serve2.5x claimed scaling efficiency over Kimi K2
Moonshot AI
New ModelsOpen weights

Kimi K3

Moonshot's Kimi K3 — 2.8T parameters — launches its API live mid-show, with full open weights following July 27

Kimi K3 went from rumor to released API in the middle of the ThursdAI broadcast: a 2.8-trillion-parameter MoE (16 of 896 experts active, ~60-75B per LDJ's estimate) with Kimi Delta Attention and attention residuals for roughly 2.5x the scaling efficiency of K2 (Moonshot's technical report later confirmed ~104B active), native vision, a 1M-token context window, and pricing around half of Opus 4.8 or GPT-5.6 Sol. It debuted #1 on the Frontend Code Arena above Claude Fable 5 and #3 on Artificial Analysis's Intelligence Index; demand forced Moonshot to pause new API subscriptions. The full weights shipped July 27 under a bespoke open-weight (not OSI) Kimi K3 license — the first open 3T-class model.

2.8T total parameters (16 of 896 experts active)1M context window (tokens)#3 / #1 AA Intelligence Index / Frontend Code Arena debut

June 2026

Moonshot AI
New ModelsOpen weights

Kimi K2.7 Code

Moonshot AI open-sources Kimi K2.7 Code for agentic coding

Moonshot AI open-sourced Kimi K2.7 Code, a trillion-parameter MoE coding model with benchmark jumps over K2.6 and fewer reasoning tokens. On the show it landed as the second half of the open-source coding wave beside GLM-5.2.

1T MoE parameters30% fewer reasoning tokens

April 2026

Moonshot AI
New ModelsOpen weights

Kimi K2.6

Kimi K2.6: 1T MoE open-source SOTA on SWE-Bench Pro

Moonshot AI released Kimi K2.6, a 1-trillion-parameter MoE with 32B active parameters, 384 experts, MLA attention, and a 256K context window under a modified MIT license. It claims open-source state of the art on SWE-Bench Pro at 58.6, and Wolfram called it the best open-source model he has ever tested on his private wolf-bench.

1T MoE Kimi K2.6

January 2026

December 2025

Moonshot AI (Kimi)
New ModelsOpen weights

Kimi K2

Kimi K2: the Chinese open model that earned mainstream respect

Moonshot AI's Kimi K2 dropped in July and earned serious mainstream recognition, marking peak Chinese-lab dominance of open source. It was named in the show's TL;DR as one of the defining open-weights releases of 2025.

November 2025

Moonshot AI
New ModelsOpen weights

Kimi K2 Thinking

Moonshot AI releases Kimi K2 Thinking, an open 1T-param reasoning MoE

Moonshot AI released Kimi K2 Thinking, an open-source 1-trillion-parameter mixture-of-experts reasoning agent with 256K context and large-scale tool-calling capacity. The panel treated it as the open-source centerpiece of the week, focusing on its reasoning quality and coding utility rather than just benchmark screenshots, and as a sign open models keep closing the usability gap with frontier closed models.

October 2025

Moonshot AI (Kimi)
New ModelsOpen weights

Kimi Linear

Kimi Linear: 48B open model with linear attention and 1M context

Moonshot AI released Kimi Linear, a 48B parameter (A3B active) instruct model that uses linear attention to reach a 1M token context window. It is an open-weights bet on efficient long-context architectures from the Kimi team.

48B parameters (3B active)1M token context window

April 2025

Moonshot AI (Kimi)
New ModelsOpen weights

Kimi-VL & Kimi-VL-Thinking

Moonshot drops Kimi-VL and Kimi-VL-Thinking, tiny A3B open vision models

Moonshot AI released Kimi-VL and Kimi-VL-Thinking, compact vision-language models with only ~3B active parameters (A3B MoE). The thinking variant adds reasoning to a tiny VLM, and both are available openly on Hugging Face.

A3B ~3B active parameters (MoE)