Ryan Feng
1.4K posts


Most startups don’t need to fight A-teams in big tech, as there aren’t many A-teams.
And now you know, A-teams are not even allowed to ship.
Tibo@thsottiaux
@_chenglou I was part of that team. Basically ChatGPT one year before it came out. Called LMChat and then another codename. Google was too nervous to release it and DeepMind was blocked from shipping products that could disrupt Google. I think about this a lot.
English

When you got 16 x GB10s you can CLAIM running future models. "Preparing to run GLM 5.5 and Minimax M4"
llmwish.open@ciprianveg
Preparing 16xGB10 cluster for Kimi K3, Deepseek V4 Pro, and future GLM 5.5 and Minimax M4..
English
Ryan Feng retweetledi

ok f*ck it: I deployed a free public endpoint for DeepSeek-V4-Flash-0731, today’s release, on Hugging Face Inference Endpoints.
Anyone can use it. No token required. OpenAI-compatible API.
It’s a shared community box, so there’s (light) rate limiting. Be nice to your neighbors 🤗
⬇️ read the guide to start using it with Pi & co
English

🚀 DeepSeek-V4-Flash Official API is now LIVE in public beta!
🔷 We’ve massively upgraded its Agent capabilities—benchmark scores are now far surpassing the V4-Pro-Preview. Check out the massive performance leap below! 👇
🔷 The official V4-Flash now natively supports the Responses API format and is fully adapted for Codex!
Check out the configuration details in our official API docs: api-docs.deepseek.com/quick_start/ag…

English

But plus plan can't even use luna and terra. What's the point?
OpenAI@OpenAI
We are committed to pushing the model frontier across cost efficiency, capability, and speed. Starting today, we are reducing prices for GPT-5.6 Luna by 80% and GPT-5.6 Terra by 20% , and offering a faster option for GPT-5.6 Sol in the API. Luna and Terra’s lower prices are reflected in how usage is counted in Codex and ChatGPT Work, so your usage goes further.
English












