Sid Shanker

692 posts

Sid Shanker banner
Sid Shanker

Sid Shanker

@sidpshanker

tools for AI builders @basetenco. Non-dairy with the glasses

New York, NY Katılım Nisan 2011
1.2K Takip Edilen819 Takipçiler
Sid Shanker retweetledi
Alex Ker 🔭
Alex Ker 🔭@thealexker·
Tutorial on how to use GLM-5.2 in Claude Code (bookmark this) ~4.5x faster & ~5x cheaper compared to Opus 4.8! 1. Install the latest Claude Code npm install -g @anthropic-ai/claude-code 2. Create an account at baseten.co. 3. Grab an API Key from app.baseten.co/settings/api_k…. Save it for the next step. 4. Edit ~/.claude/settings.json. Open with vim or another editor. Paste the following with your key. "env": { "ANTHROPIC_AUTH_TOKEN": “your_baseten_api_key", "ANTHROPIC_BASE_URL": "inference.baseten.co", "ANTHROPIC_DEFAULT_HAIKU_MODEL": "zai-org/GLM-5.2", "ANTHROPIC_DEFAULT_SONNET_MODEL": "zai-org/GLM-5.2", "ANTHROPIC_DEFAULT_OPUS_MODEL": "zai-org/GLM-5.2" } 5. Enjoy GLM-5.2 in CC!
Alex Ker 🔭 tweet media
English
37
37
437
42.6K
Sid Shanker retweetledi
Dhruv Singal
Dhruv Singal@alphatozeta8148·
@baseten model performance team is absolutely cracked. @Zai_org GLM 5.2 is now 4x faster running at full 1M context! Already available to use in your favorite coding harnesses, here it is COOKING in @FactoryAI Droid and @opencode Docs for how to get it in comments
English
9
17
155
38.5K
Sid Shanker retweetledi
Baseten
Baseten@baseten·
GLM 5.2 is live on Baseten. 5.2 is built for agentic engineering: stronger coding, sharper agentic reasoning, and a long context window built to run hours-long tasks. Test it today: baseten.co/library/glm-52/
Baseten tweet media
English
8
8
144
1.9M
Sid Shanker
Sid Shanker@sidpshanker·
Really proud of the team for shipping rolling deploys! This has been a critical part of ensuring our customers can push the latest & greatest perf optimizations into prod fast & without having to spend on a ton of extra GPUs
Sid Shanker@sidpshanker

x.com/i/article/2065…

English
0
0
14
552
Dom Wong
Dom Wong@domwong·
The dirty secret about consumer research? It’s overrun by fraud. Survey bots. People lying to qualify for interviews. Today, Pogo is launching the world’s only AI researcher that lets brands talk to verified buyers of any product - at scale, in hours. Backed by $32M to date. For 6 years, we’ve built the consumer network that makes research with real, verified buyers possible: • 3M+ opted-in U.S. consumers • Visibility into 1 in every 150 U.S. shopping trips • $470B in observed transactions • Ranked #1 Loyalty App in the U.S. by Newsweek, two years running Watch the full story 👇
English
85
56
410
1.2M
Sid Shanker
Sid Shanker@sidpshanker·
@domwong Congrats on the launch! been so cool to see this play out 😎
English
1
0
1
86
Sid Shanker retweetledi
Orb
Orb@useOrb·
At AI infrastructure companies, pricing is part of the product. Sid Shanker, Engineering Manager at @baseten, helps the team launch new models and hardware without turning monetization into a bottleneck. “We launch new models and hardware constantly, and pricing them is half the product decision. Orb lets us move at the speed of our roadmap and prioritize delivering value to our customers.”
Orb tweet media
English
0
2
9
737
Sid Shanker retweetledi
Beff (e/acc)
Beff (e/acc)@beffjezos·
Feeling robbed of my path to citizenship right now after grinding a PhD and contributing to foundational AI + computing technologies for the United States for the past ~ 10 years. Feels like robbing top and technologists like me of the opportunity to achieve the American Dream.
English
1.1K
157
3.4K
5.5M
Sid Shanker retweetledi
Baseten
Baseten@baseten·
Our engineers just shipped the fastest named entity recognition (NER) inference on the market: 1 ms P50 and 3 ms P99 server-side latency, 7.7x faster than an optimized PyTorch baseline. The bottleneck in fast NER is rarely the model: it's tokenization, HTTP parsing, networking, and other overhead. We solved for that and added it to Baseten Embeddings Inference (BEI). Just point it at your model checkpoint, no model changes needed.
Michael Feil@feilsystem

x.com/i/article/2041…

English
4
6
115
13K
Sid Shanker retweetledi
Zed
Zed@zeddotdev·
Your AI code completions in Zed show up in ~200ms. That's Zeta, our Edit Prediction model, running on @baseten. We love partnering with companies who keep the bar high — Baseten is one of them.
English
30
25
813
97.5K
Sid Shanker retweetledi
World Labs
World Labs@theworldlabs·
We’re building foundational world models to power the next era of 3D. From robotics to gaming, spatial intelligence unlocks entirely new worlds. Powered by inference at scale – shoutout to Baseten.
English
11
26
207
19.6K