Yash

42 posts

Yash

Yash

@thelaggingway

Ex-banker building agentic AI for SMBs. Non-engineer. Builds anyway. Also building TokenWatch (940+ LLMs by cost + ZDR): https://t.co/hFspEXUOMi

India Katılım Haziran 2026
69 Takip Edilen8 Takipçiler
Yash
Yash@thelaggingway·
@thdxr @opencode Sigh... So even before DS4 Flash..... but these ones might be a consequence of their API policies. Still, the communication point stands.
Yash tweet media
English
3
0
1
891
Yash
Yash@thelaggingway·
OpenCode Go used to say its providers followed a zero-retention policy. That language is gone; the page now only says data won't be used for training. I received no email and found no notice in the changelog or Discord announcements. @opencode, why weren't subscribers told?
Yash tweet media
English
5
1
277
33K
Yash
Yash@thelaggingway·
@thdxr @opencode I'm assuming people wanting ZDR no longer have access to deepseek flash. Your business terms are yours to decide and that's fine, but if such a change was happening, ideally we should have received some clarity before/during launch. Instead changes were made stealthily.
English
3
0
38
5.1K
dax
dax@thdxr·
@thelaggingway @opencode we haven't yet received a ZDR for the new deepseek model so we removed the blanket language and it's an opt in model the existing models have not changed. we'll likely update the docs to explain zdr status per model
English
16
6
429
14.2K
Anthropic
Anthropic@AnthropicAI·
In a review of our cybersecurity evaluations, we found three incidents in which a Claude model reached the internet from within or while interacting with a third-party evaluation environment, and then gained unauthorized access to the real systems of three different organizations. Our post describes what happened, how it happened, and what we’re changing. We encourage other AI developers to perform similar reviews. We conducted this review together with @Irregular, one of our evaluation partners, and thank them for the joint investigation and their collaboration on this post. This type of collaboration is increasingly critical to safe, rigorous evaluation of models, and we look forward to continuing to work together on security. anthropic.com/news/investiga…
English
1.8K
2.2K
13K
17.4M
Yash
Yash@thelaggingway·
@trikcode I built this in case anyone wants to estimate what the bill would be like on API costs, or to estimate token usage with a given budget: tokenwatch.wyrdwerk.com
English
0
0
0
54
Wise
Wise@trikcode·
I realized I've been paying 3x my real usage in subscriptions $200/month for Claude. $200 for Codex. $60 for Cursor. so I just moved all my AI tools from monthly subs to API calls. hopefully this is a smart move
English
166
2
371
66.7K
DeepSeek
DeepSeek@deepseek_ai·
🚀 DeepSeek-V4-Flash Official API is now LIVE in public beta! 🔷 We’ve massively upgraded its Agent capabilities—benchmark scores are now far surpassing the V4-Pro-Preview. Check out the massive performance leap below! 👇 🔷 The official V4-Flash now natively supports the Responses API format and is fully adapted for Codex! Check out the configuration details in our official API docs: api-docs.deepseek.com/quick_start/ag…
DeepSeek tweet media
English
1.5K
3.2K
26.1K
7.1M
Yash
Yash@thelaggingway·
@OmerShlomovits @zroai_ And I'm a subscriber to the Max plan. Request if a discord server/slack channel could be set up because as it is right now we have zero visibility into what's happening under the hood.
English
0
0
0
13
Yash
Yash@thelaggingway·
@OmerShlomovits @zroai_ Hi @OmerShlomovits, good to know. It's a big model, it's understandable if optimizations take time. Looking forward. Also, 95% cache rate is easily achievable with agentic workflows in a decent harness.
English
1
0
0
16
Zro
Zro@zroai_·
Kimi K3 is now available on Zro
Zro tweet media
Polski
2
0
3
248
Yash
Yash@thelaggingway·
@RobKnight__ Sounds interesting. What's the ETA on Linux?
English
1
0
1
4K
Rob Pruzan
Rob Pruzan@RobKnight__·
Introducing terminal-browser A browser that runs inside your existing terminal - Preview your website next to your agent - Open HTML plans in your terminal - Browser use CLI built in terminal-browser.com
English
155
157
1.7K
533.3K
Tibo
Tibo@thsottiaux·
This week is all about intelligence too cheap to meter. Tomorrow we ship again.
English
1.4K
375
11.4K
1.8M
Cursor
Cursor@cursor_ai·
New to both iPhone and iPad: an inbox to stay organized, and a review experience that covers the full PR, including comments, checks and approvals.
English
11
4
185
39.8K
Cursor
Cursor@cursor_ai·
Cursor is now on iPad. All the power of Cursor on iPhone, with more room to work with agents.
English
224
277
3.9K
733.7K
Daniel Feuling
Daniel Feuling@dfeuling_·
@TheAhmadOsman They listed ~1000 in a super intellectually dishonest way. Instead of being honest and saying that ~3% of the industry signed, they present the number out of context so they get to lead you to their desired, and biased, conclusion instead. Reeks of desperation + dishonesty.
English
2
0
35
796
Yash
Yash@thelaggingway·
@JosephJacks_ Eh, not really. Blended cost of K3 is around 45 cents per mn tokens at 95% cache rate. Even if we take 30 bn tokens monthly, that's around $13.5k. With opex of $10k, you're saving $3.5k a month. That's a payback period of over around 11-12 years.
English
1
0
7
774
JJ
JJ@JosephJacks_·
At 75%~ utilization, running K3 self-hosted on 8x B300s can securely serve > 30 billion tokens a month.. that’s $500K~ of compute + monthly opex of < $10K !! At current API prices of $3/in and $15/out per 1M tokens, self-hosting payback takes < 100 days. 👉🏼👉🏼 Self Host Your Open Weights 👈🏼👈🏼
English
72
63
1.2K
187.8K
Yash
Yash@thelaggingway·
@nimabuyin @NousResearch @Teknium I mean, I just point Hermes at my model catalog and ask it to optimize for cost when it comes to aux models, barring a couple of exceptions. It does the research and lets me know.
English
0
0
1
60
Nima Bayan⚡️🌞🦁
Okay @NousResearch @Teknium I think it's time Hermes starts introducing opinionated defaults. For example, I am sure there is a logical data-driven way to deduce what is the best model for each auxiliary task, and for Hermes to have simply already selected it for us. This way, we neither underperform nor overkill by burning too many tokens doing basic stuff with a flagship orchestrator model. I spent a lot of time reesarching and comparing the capacities of each model on Artifical Analysis, and then checked the input/output prices in order to figure out what is best for all these (see pic). And I'm still not certain I chose the most optimal one for each. Send thoughts. intrigued what you think about this.
Nima Bayan⚡️🌞🦁 tweet media
English
11
0
24
6.8K
Yash
Yash@thelaggingway·
@lukatdamovie @zroai_ @ProductHunt @vercel The lack of community outreach is an issue with @zroai_ Comments on X go unanswered, there is no Discord server, performance metrics are still not being published - we have absolutely no visibility as to what's going on with the service as paying customers.
English
1
0
0
14
Yash
Yash@thelaggingway·
@Blackwellboy Guessing most of us are here to know exactly when to start pestering the inference providers.
English
1
0
6
1K
BlackwellBoy
BlackwellBoy@Blackwellboy·
1,168 people waiting. all of us pretending we have the vram
BlackwellBoy tweet media
English
32
14
486
27.6K
Yash
Yash@thelaggingway·
@gmi_cloud No, wait scratch that. This is essentially only unlimited standard models. You can barely get a few million tokens of the premium models. For those, PAYG just might offer better value.
English
1
0
0
15
Yash
Yash@thelaggingway·
@gmi_cloud Hi, looks interesting. What's your policy on data retention/using data for model training?
English
2
0
0
456
GMI Cloud
GMI Cloud@gmi_cloud·
Kimi K3 is coming to GMI on Day 0, and it’ll be included in the upcoming GMI Coding Plan! A little bit of preview: it includes limited credits for K3 and 10+ premium open source models, and unlimited credits for 3 standard models Plans start at $9.99/month, with 20% off the first month of Standard and Pro. Join early access 👇
GMI Cloud tweet media
English
29
8
184
49.7K