郭宇 guoyu.eth
10.4K posts


Announcing OpenWorker! An open-source agent that doesn't just chat with you, but delivers finished work -- like hand you a polished document, send a slack message, or update a calendar entry. Ask it to prepare a customer brief, untangle your calendar, draft a report, or triage a Slack alert. It works across your files and everyday tools, produces the deliverable, and checks in before doing anything consequential. OpenWorker runs on your Mac, with Windows support coming soon. It does not lock you into any one model. Bring your own API key and run it with GPT 5.6 Sol, Claude Fable, Gemini 3.6, an open weight model (like Kimi, GLM, DeepSeek, Inkling), or Ollama to keep your data local. Your data does not leave your machine except through an LLM provider and integrations that you choose. @rohitcprasad and I are building OpenWorker because AI coworkers are an important way to get work done, and we want there to be an open, privacy-preserving, model-independent option. Check it out and let us know what you think! Try it out: openworker.com (requires your own API key) Source code: github.com/andrewyng/open…

Introducing Marble (YC S26) — the autonomous back of house for restaurants. If you own or have worked in a restaurant, you know: 5% margins, and the whole back of house still runs on a clipboard and whatever the GM remembers. Marble counts your stock in minutes with computer vision and forecasts what each location will actually sell. Then, a set of AI agents manage ordering, food prep, scheduling, and much more on autopilot. Live today across some of the top restaurant groups in the US and Canada, learn more at joinmarble.ai P.S Know a restaurant owner? Show them Marble and get $500 or an Oura Ring

发现了 Codex 一个非常牛逼的更新,好像没见他们宣传! Codex 现在有专门给图像 Agent 做的 UI 模式了: 生成图片后,点击图片会单独弹出一个侧边栏预览窗口。在这里面,你可以直接对图片进行评论、擦除和调整大小。这完全是专门给 GPT Image 2.0 做的。 左上角有一个切换按钮,切换以后,聊天流里就只显示图像,不显示那些文案了。 这样你就可以专注于调整图像,快速看到聊天里生成的所有大图。 你可以多选图片,一并添加到输入框里,然后让 GPT 给你进行批量修改。 这个更新感觉要把设计 Agent 的活吃掉一大半了。

Today we’ve raised $52M Seed and we are announcing the public launch of S2.1 Pro. >It can clone a voice from 5 seconds of audio >2x faster than Cartesia & 1/6th the cost of Eleven Labs >most expressive model with word level control over emotion, intonation, pacing etc We support frontier AI companies including HeyGen, LiveKit, Retell, Sanas, and OpenArt all run our model in production. If you're a business and we can't cut your voice AI costs by 50%, we'll give you 1 year of Fish Audio for free. Book a demo: s.fish.audio/tmapke To celebrate our first birthday, we'll give you 1 month of S2.1 Pro for free. Like, retweet, and comment “Fish” to get it.
















































