Kimi Developers

75 posts

Kimi Developers banner
Kimi Developers

Kimi Developers

@KimiDevs

The official Kimi account for developers building with Kimi Code and the Kimi API.

Katılım Mayıs 2026
4 Takip Edilen67.9K Takipçiler
Sabitlenmiş Tweet
Kimi Developers
Kimi Developers@KimiDevs·
Kimi K3 is here:Built on Kimi Delta Attention and Attention Residuals, this 2.8T-parameter model pairs native vision capabilities with a 1 Million Context, making it the world's first open 3T-class model!
Kimi.ai@Kimi_Moonshot

Introducing Kimi K3: Open Frontier Intelligence 🔹 2.8 Trillion Parameters, 1 Million Context, Native Multimodal 🔹 Kimi Delta Attention enables up to 6.3x faster decoding in million-token contexts 🔹 Attention Residuals deliver ~25% higher training efficiency at <2% additional cost 🔹 Built for long-horizon agentic coding and self-evolving workflows Kimi K3 is now live on on Kimi.com, Kimi Work, Kimi Code, and the Kimi API. Open Weights by July 27, 2026. 🔗 API: platform.kimi.ai 🔗 Tech blog: kimi.com/blog/kimi-k3

English
50
150
2.3K
119.4K
Kimi Developers
Kimi Developers@KimiDevs·
Bug Fixes: ✅Fail fast when account quota or balance is exhausted instead of silently retrying for ~3 minutes. ✅Stop the turn after repeated invalid tool calls instead of retrying indefinitely. ✅web: Fix garbled line numbers in code blocks. kimi.com/code/docs/en/k…
English
1
1
32
3.7K
Kimi Developers
Kimi Developers@KimiDevs·
Kimi Code Changelog 0.30.0 Features: 💡Add a customizable footer status line, configured via [status_line] in tui.toml Polish: ✅Show a quota note after installing official plugins that bill against plan quota (such as Kimi Datasource). ✅Show a notice when an official plugin used in the session has an update available — run /plugins to update. ✅Remove the 50 MB size limit on file uploads to the built-in server.
English
26
25
556
33.3K
Kimi Developers
Kimi Developers@KimiDevs·
🤗
LMSYS Org@lmsysorg

SGLang day-0 speed on Kimi K3: 423 tok/s (measured on gsm8k), plus RL support ready in Miles @radixark! How the largest open-source model runs this fast: we natively implemented and deeply optimized K3’s new architecture with fused KDA decode kernels, DP attention, DSpark, PD disagg, and KDA-aware prefix caching. We've passed Kimi Vendor Verifier and are ready for production! Thanks to @Kimi_Moonshot, @nvidia, @AMD, @KVCache_AI, @modal, and @baseten for building this with us, and to @googlecloud, @nebiustf, @fal, @digitalocean, @runpod, @DeepInfra and @gmi_cloud for serving K3 on SGLang. Blog, cookbook, benchmarks in the comments. P.S. This demo video? Kimi K3 made it itself. Play the game 👇

ART
1
11
267
12.8K
Kimi Developers
Kimi Developers@KimiDevs·
🤗
LightSeek Foundation@lightseekorg

Kimi K3 is now supported in TokenSpeed. We worked closely with the @Kimi_Moonshot team to bring Day 0 support to both NVIDIA Blackwell ((G)B200/(G)B300) and AMD Instinct (MI350X/MI355X) in just one week. ✅ Prefix caching ✅ Speculative decoding ✅ Disaggregated serving ✅ CUDA Graph decode We also share how we optimized KDA, Gated MLA, Stable LatentMoE, AttnRes, and our unified Flat KV architecture. Blog ↓ lightseek.org/blog/tokenspee…

ART
4
8
390
22K
Kimi Developers retweetledi
Kimi.ai
Kimi.ai@Kimi_Moonshot·
Releasing the model weights and technical report of Kimi K3. Kimi K3 is our most capable model: a 2.8T MoE model with native visual understanding and a 1M-token context window. New model architecture: 2.5x the intelligence per unit of compute, not just more params. Alongside Kimi K3, we're opening up more of the stack behind it — high-performance attention kernels, MoE communication library, and infrastructure for running agent environments at scale. Model weights: huggingface.co/moonshotai/Kim… Tech report: github.com/MoonshotAI/Kim… Tech blog: kimi.com/blog/kimi-k3
Kimi.ai tweet media
English
1.5K
7.3K
45.9K
12.8M
Kimi Developers retweetledi
Kimi.ai
Kimi.ai@Kimi_Moonshot·
Kimi K3 (open weights, coming soon)
English
433
1K
13.8K
1.1M
Kimi Developers
Kimi Developers@KimiDevs·
Kimi Code Changelog 0.29.2 --- Bug Fixes 🌟Fix goal pursuit pausing when a goal turn hits the per-turn step limit (loop_control.max_steps_per_turn). 🌟Fix messages sent during goal pursuit being rejected. 🌟Fix /undo to restore conversation history, todo lists, plan mode, and task notifications consistently. 💡web: Fix copying selected chat text over plain HTTP overwriting the clipboard with an event placeholder. kimi.com/code/docs/en/k…
English
21
27
674
35.2K
Kimi Developers
Kimi Developers@KimiDevs·
Bug Fixes Fix loss of thinking content with OpenAI-compatible endpoints that return reasoning under a different field name (e.g. newer vLLM). See more 👀 #secondary-model" target="_blank" rel="nofollow noopener">moonshotai.github.io/kimi-code/en/c…
English
3
0
37
6.1K
Kimi Developers
Kimi Developers@KimiDevs·
Kimi Code CLI 0.29.1 Features 🌟Add global default MCP server timeouts in config.toml and env vars. 🌟Add environment variables to configure the web search and web fetch services without OAuth login. 🌟Add experimental secondary-model bindings for newly spawned subagents, including per-agent model preferences and subagent-only model overrides.
English
40
38
796
67.4K
Kimi Developers
Kimi Developers@KimiDevs·
Agent workloads bring long contexts, bursty traffic, and frequent tool calls, creating new demands for AI serving. Together with @AMD , we rebuilt the stack for Kimi K2.6 on AMD Instinct™ MI355X, with scheduler-aware multi-tier KV caching. Up to 3.2× smaller p99 TTFT and 7.7% higher total-token throughput, with no accuracy tax. Happy to see more Kimi running on AMD chips! Tech blog: amd.com/en/developer/r…
English
55
108
1.7K
118.4K
Kimi Developers
Kimi Developers@KimiDevs·
He gave K3 a 3D model and one prompt, then went to dinner. In about 30 minutes, K3 completed 95 steps to create the first version. K3 found and deployed local ASR and TTS solutions by itself, connected the LLM, and built the scenes and interactions around them.
Kimi Developers tweet media
English
7
12
277
27.8K
Kimi Developers
Kimi Developers@KimiDevs·
A Kimi staff member used K3 on Kimi Code to build a VR companion. She can listen, reply, show different expressions, and move between scenes like a café and a shop.
English
164
351
4.3K
274.5K
Kimi Developers
Kimi Developers@KimiDevs·
Bug Fixes: Correct the YOLO and Auto permission mode descriptions: YOLO auto-approves tool actions but the agent may still ask questions, while Auto is fully autonomous and never asks. Fix the web backend ignoring symbolic links when loading AGENTS.md files and reading files. See more:moonshotai.github.io/kimi-code/en/r…
English
3
0
45
6.1K
Kimi Developers
Kimi Developers@KimiDevs·
Polish: ✅Thinking effort persists only levels below the model's top tier (max). ✅web: Add a note in the model switcher that switching models or thinking effort invalidates the existing prompt cache.
English
4
0
53
7.1K
Kimi Developers
Kimi Developers@KimiDevs·
Kimi Code CLI Changelog 0.28.0 Features: ✅The kimi server command tree is deprecated; use kimi web instead. ✅kimi web now runs in the foreground of the current terminal and opens the browser; stop it with Ctrl+C.
Kimi Developers tweet media
English
50
65
1.2K
85.5K
Aaliya
Aaliya@aaliya_va·
@KimiDevs This is really useful. Less clutter makes AI work better.
English
2
0
6
688
Kimi Developers
Kimi Developers@KimiDevs·
When your application needs a large number of tools, declaring all of them up front in the top-level tools field of every request leads to Tool Definition Bloat: every request carries the descriptions and parameter schemas of all tools, driving up token usage, and the more candidate tools there are, the more likely the model picks the wrong tool or constructs invalid call arguments.
Kimi Developers tweet media
English
48
73
1.5K
129.2K