
Mitesh B Ashar
30.1K posts






Kimi K3 has received far more love than we expected, and our GPUs are feeling it. Over the past 48 hours, demand has pushed close to the limits of our current capacity. To protect the experience of existing subscribers, we're temporarily pausing new subscriptions and prioritizing compute for current members. Existing subscribed users are not affected. We're adding capacity as fast as we can and will reopen new subscription spots in batches. Going forward, we'll also split membership into two more focused plans: Kimi Membership for Kimi Web, App, and Work; and Kimi Code Membership for coding workflows. This will help us match compute more precisely and keep the experience stable. Thank you for your patience and understanding!

Couldn't agree more with this! I stopped filing issues after a certain point of time. Your own AI was supposed to be your leverage/vantage point. And you couldn't use it well enough to address growing user needs, rather letting them rot and auto close - ignored and abandoned!



6 months of Claude Max 20x, on us. We're expanding Claude for Open Source to more of the community. If you're a maintainer, a core contributor, someone landing PRs across the ecosystem, or someone keeping a critical package alive, apply today!


Chinese models are 112x cheaper than Anthropic per million tokens. Chamath laid it out on CNBC: a "barrel of intelligence" costs $56 from Anthropic, $26 from OpenAI, $1.50 from Meta, $1 from xAI and Google, and $0.50 from Chinese models. That is not a pricing quirk. That is the steepest commodity curve any technology has run in recorded history. Oil took 40 years to compress like this. Semiconductors took 20. AI inference is doing it in months. The companies sitting at $26 and $56 are not stupid. They're buying time… betting that trust, safety, and enterprise contracts hold the premium long enough for costs to catch up. What they cannot bet on is the timeline. Because the $0.50 model is not a demo. I've been inside the labs building it.







Two months ago I was fired by Google for creating the Google Workspace CLI. It went viral, hit #1 on Hacker News, gained thousands of GitHub stars and many thousands of actual users in just a couple days. It was an incredible, confusing journey, from directors and leaders asking what they could learn from the tool to getting grilled by legal about why the Google logo and brand colors are on the Google Workspace GitHub code repositories. I think the cause was that Workspace and certain leaders (and projects) were afraid of being disrupted. But the fear wasn't specific to my CLI, it was a broader fear in what agents meant for Workspace. Either way, the irony of my termination was the announcement at Google Cloud Next two days before I was fired that an official Workspace CLI was coming. I want this out there because it is easier for me to explain my story and it is an experience I want to fully own. It's also part of my healing. Nearly 7 years at Google was an incredible opportunity for me and I was fortunate to have wonderful teammates and a manager that fully supported me through these last few months. Thank you.



If Codex wins over Claude Code it will be purely because 1. Claude team truly treats the user interface like shit (they don't fix widely reported bugs and inconveniences for months, idk what does Boris run his infinite token loops for even?) 2. They keep overselling this "coding is solved" when clearly they cannot create a good frontend product across their mobile app, their website or their TUI. Claude mobile app is a horrible product, the desktop app is so buggy, conversations hang, get lost, remain dangling.... it is almost as if no one in the team ever tries their own products for 5 minutes

One of my favorite interview questions is: What is most important capability of a frontier LLM? Most people get it wrong.

