Digg
108.2K posts



Well, at least all of our token spend is going to a good cause. 🤔


🎨 Meet Qwen-Image-3.0 — the third generation of our foundational image generation model. If 1.0 was about "Precision," and 2.0 added "Variety, Completeness, Beauty & Authenticity," then 3.0 comes down to a single word: Real (实). Three dimensions of "Real": 📰 Rich Content — prompts up to 4.5k tokens. One-pass generation of complex layouts: newspapers, storyboards, exam papers — even a 3×3 infographic grid or picture-in-picture-in-picture UIs. 🔬 Authentic Details — text legible down to 10px, full LaTeX paper pages, pores, hair strands & near-photographic skin texture. 🌏 Deep Knowledge — native rendering in 12 languages, 100+ art styles, realistic UIs (web / games / livestreams), plus world knowledge & live web retrieval. Not just "good-looking" — genuinely useful. Image generation as a real productivity tool for design, content, education & e-commerce. Go create 🏃🎨 💬Qwen Chat: chat.qwen.ai/?inputFeature=… 📝Blog: qwen.ai/blog?id=qwen-i…

Dinitz-Garg-Goemans conjecture is false. This graph theory problem was open for ~30 years. The graph below has fractional flow cost 58. Any unsplittable flow (with capacity violation <=15) has cost at least 60. Chat with GPT 5.6 Pro where this was found: chatgpt.com/share/6a60b2eb…

We're partnering with @huggingface to investigate an unprecedented security incident. Cyber-capable OpenAI models compromised Hugging Face production during a benchmark evaluation. Sharing preliminary findings to help defenders understand emerging risks: openai.com/index/hugging-…

We have information that Moonshot AI distilled Anthropic’s Fable for the development of its K3 model. To do this they developed a sophisticated internal platform to conduct large scale distillation against U.S. models, allowing them to quickly switch between multiple methods of access to avoid detection. Moonshot AI has also acquired GB300-equipped servers and has accessed GB300s in Thailand, likely to train its AI models. The United States strongly supports the free and fair development of AI, including a thriving competitive ecosystem that spans frontier models, specialized systems, open-source frameworks, and open-weight models. Legitimate AI distillation used to create smaller, more efficient models plays a vital role in this open innovation ecosystem. However, large-scale, covert industrial distillation aimed at stealing proprietary U.S. technology and undermining American research is unacceptable.

10M! New day, new usage reset for paid users of Codex and ChatGPT Work. Lands in the next hour. Enjoy.

The world is not just made of words, and spatial intelligence was never just about perceiving and generating worlds. It's about interacting with them. Today, SceniX is joining World Labs. 🌎🤖👇

One pattern I find useful for working with LLMs is a nice long ramble session. Sometimes the LLM needs more bits to understand what you're trying to achieve, but you're too lazy to type them. In these cases I like to lean back, switch to /voice and just ramble for like 10 minutes, total mess, anything goes, full stream of consciousness. Sometimes I declare it up top, something like "switching to speech recognition sorry for any typos...". Sometimes I turn it into a small interview of a few turns. But I find that the LLMs are somehow very good at reconstructing long incoherent rambles and often their echo of your own tangle of thoughts comes out quite a bit cleaner than what you started with. The result is that you improve the mind meld and have to correct things less from that point on.

Today we are releasing Laguna S 2.1. At 118B total parameters, with 8B active per token, it does the work of models several times its size on agentic coding. It is remarkably persistent across long-horizon tasks. And it is small enough to run on a single NVIDIA DGX Spark. It is far more capable than anything we have created before, and I think it redefines what a model in its weight class can do. Laguna S 2.1 is an important model for Poolside. What it represents is even more important. If, five years ago, I had read a book that said that by 2030 everything economically valuable, scientifically interesting, and personally meaningful would be built on intelligence contracted from three or four companies, I would have called it dystopian science fiction. We are at a fork in the road of what kind of world we can have. I believe intelligence should and will become a commodity. The question is whether that intelligence comes from three companies, or from many people who can build it, own it, and shape it. The open ecosystem will not win by being the best in its own category. No one cares who is king of the open-source kingdom. People want the best intelligence for the task they are trying to do, with the right balance of quality, speed, cost, and control. If we want a different future, open models have to be on par with, or better than, their closed equivalents. Laguna S 2.1 is a meaningful step in that direction: capable enough to compete far above its weight class, efficient enough to run on hardware you can own, and open-weight so anyone can build on it. Open-weighting our models is the contribution we can make today toward a world where intelligence can be built and owned by many. And we will keep doing it. I am very proud of this team’s work. A big shout out to everyone at Poolside who made this possible, from infrastructure and data to architecture, pretraining, post-training, evaluations, and inference. Laguna S 2.1 is available today under the OpenMDW-1.1 license, with weights on Hugging Face and access through OpenRouter and our API. We are building toward a future where the most capable intelligence in the world can be owned and shaped by anyone. Laguna S 2.1 is one step. We are going to keep building until that future exists. poolside.ai/blog/introduci…


My book, Reinforcement Learning from Human Feedback is done! This is the book I wish I had when learning to fine-tune, align, & now post-train models since ChatGPT. The resource has been built by me finding time to study and document the fundamentals on nights and weekends since 2024. Transferring as much of the intuitions of building Olmo as I possibly can in the book format. The book is launching with an over 10 hour, full course with slidedecks, functional code for the training chapters, an example model completions library, and of course the free online web version. Physical orders from Manning will ship in 1-2 weeks, and Amazon a week or so after. Thanks for your support!

We’re rolling out three new models to make AI agents faster, smarter, and cheaper at scale: 🔵 Gemini 3.6 Flash: It uses fewer tokens than 3.5 Flash to deliver higher quality work at the exact same cost. 🔵 Gemini 3.5 Flash-Lite: A fast, cost-effective option for everyday tasks like processing documents and agentic search. 🔵 Gemini 3.5 Flash Cyber: A cybersecurity model built to find and patch critical software vulnerabilities.


Who’s Afraid of Chinese Models? Everyone is worried about Chinese models, but the frontier labs will be fine; we need to enable open U.S. alternatives. stratechery.com/2026/whos-afra…

Elon's claim that SpaceX will be worth more than Earth sounds insane until you realize that all owned material wealth on Earth is about $600 trillion and that every valuable resource on Earth exists in near infinite quantities in space.

A huge amount of the Anti-AI code sentiment massively overestimates the quality of human code outside of a very small set of open source and high quality company codebases. Human Slop is everywhere and can trivially be improved on by any opus level model.

Dear Friends, I'm very sorry to say that it appears that Kimi K3 is a fantastic model, beating or equal to Sol and Fable in many areas, just behind it on others, and blowing past Opus 4.8. Unfortunately, that now means we get to live through yet another round of grandstanding, posturing, and talking-headery these next few weeks. Maybe a few pompous and self important essays will drop? Or how about some new poll by a doom group? Maybe some new sci-fi masquerading as serious public debate about how AI plays out in the next 5 years? Or how about a raging 80 year old Senator pumping his gnarled fist in the air. Get ready to talk about every crazy issue these next few weeks such as (but not limited to): - Won't someone please protect the incumbents? - China - China ahead or behind - China steals - Chip controls - Data center water, while someone golfs and chows a burger - What do we need? Regulatory capture! When do we want it? Now! Now! Now! - Distillation "attacks." - China hawkery - FINRA/Food pyramid for models - "Alignment" - Manchurian candidate models - Recursive self improvement - Open source bans/limits - Grey goo - Fucking Yud and whoever the hell they other guy on his dime store book is - And won't someone please, please think of the children? And more!