avv

237 posts

avv

avv

@avithaldas

tech, learning and Econ

Bengaluru, India Katılım Kasım 2009
838 Takip Edilen151 Takipçiler
avv
avv@avithaldas·
@thdxr This is daxxing…
English
0
0
0
73
dax
dax@thdxr·
im never leaving this app
dax tweet media
English
14
1
276
24.7K
avv
avv@avithaldas·
@athyuttamre A huge improvement over the earlier voice mode. Is access through api expected soon?
English
0
0
0
21
Atty Eleti
Atty Eleti@athyuttamre·
GPT-Live is now fully rolled out to all ChatGPT users globally. We're also doubling everyone's voice usage limit for the whole weekend so you can try more of it. Have fun! 🎉
English
244
125
1.9K
273.5K
avv
avv@avithaldas·
🔥 Thats great to see. Just 1 request - OpenClaw is starting to get complicated and is starting to have a huge barrier for people not already familiar with it. A simpler usage variant (even if it has a lower number of integrations / skills) can go a long way in exposing more people to its possibilities.
English
0
0
4
2.7K
Dave Morin 🦞
Dave Morin 🦞@davemorin·
Today we’re introducing the OpenClaw Foundation: a nonprofit home for open, independent personal AI. A full-time team. Great partners. One mission: bring personal AI to everyone. Welcome to the age of the lobster.🦞 openclaw.ai/blog/introduci…
English
99
161
1.9K
856.7K
avv
avv@avithaldas·
@Angaisb_ Already here.. in DE
English
0
0
0
132
Angel 🌼
Angel 🌼@Angaisb_·
GPT Live 1 is so cool I hope it also gets released in the EU!
English
13
1
145
18.6K
avv
avv@avithaldas·
@martian588 Raised a couple of PR's. That said, have some questions about the direction of the project. Possible to DM?
English
1
0
1
184
martian58
martian58@martian588·
Looking for contributors - Devlaner/devlane: Open-source Jira, Linear, Monday, ClickUp and Plane alternative. github.com/Devlaner/devla…
martian58 tweet media
English
23
15
351
909.1K
avv
avv@avithaldas·
@OpenAI Hopefully, Europe too 🙏 🤞
English
0
0
0
20
OpenAI
OpenAI@OpenAI·
Introducing GPT-Live, a new generation of voice models for natural human-AI interaction. Rolling out in ChatGPT starting today. You’ll want to turn the sound on for this one.
English
1.3K
2.2K
24.5K
7.9M
avv
avv@avithaldas·
Sure, there is a monetisation guardrail that would exist for ad-funded consumer apps. But for B2B software, it may end up being a stronger default
English
0
0
0
6
avv
avv@avithaldas·
While it may seem very incremental, I do believe self-generative apps (like Monogram) are likely to be a UX that most companies would make default going forward. Why have one fixed view, when everyone see's exactly what they care most about.
Eren Bali@erenbali

It's time to get out of stealth 👋 Today, we are launching @monogram_ai and announcing our $40m seed round led by DST and Lux Capital. Monogram is the first AI app that was built around a visual interface, from the ground up. We created a technology that generates an entire user interface on the fly, in just a few seconds. Ask anything, and instead of staring at a wall of text, you get an interactive visual response.

English
1
0
0
42
avv
avv@avithaldas·
@wafer_ai thats a crazy number, any idea on time to first token (ttft) yet?
English
0
0
0
232
wafer
wafer@wafer_ai·
🚨 OpenAI is launching GPT-5.6 Sol on Cerebras at up to 750 tps. Here's How kernels work on Cerebras Chips cerebras built a chip with 900,000 cores on a single silicon wafer. CSL (Cerebras Software Language) is a Zig-inspired DSL that gives you direct control over the Wafer-Scale Engine. the SDK is publicly documented and the programming model is genuinely the most alien one we've ever looked at. in CUDA you write one thread's perspective and launch millions. in CSL, there are no threads, no warps, no shared memory, and no kernel launches. you write code for individual Processing Elements: 900,000 independent cores arranged in a 2D mesh on a single silicon wafer. each PE has its own 48 KB of private SRAM, its own program counter, and a 5-port router connecting it to its 4 neighbors. that's actually it. no DRAM. no HBM. no cache hierarchy. 48 KB is your entire world per PE, where code and data must both fit. the programming model is dataflow. data moves between PEs as 32-bit messages called wavelets, traveling along virtual channels called colors. when a wavelet arrives at a PE on a specific color, it activates a task (a chunk of code bound to that color at compile time). tasks run to completion, then hardware picks the next activated task. tasks cannot call each other. they can only be activated. so instead of "launch N threads," you think of it like: "place code on PEs, define routes, let data flow." the memory model is also very different from GPUs. on an H100 you get 80 GB of HBM shared across all SMs. on the WSE-3, memory is 48 KB per PE, and there are 900,000 of them! this gives you 44 GB total on-chip SRAM at 21 PB/s aggregate bandwidth (vs 3 TB/s on H100). every access is single-cycle. no coalescing needed. no bank conflicts. no cache misses. but also no way to access another PE's memory. all inter-PE communication is explicit wavelet routing through the fabric. take a distributed GEMV for example. you would write a layout file that physically routes wavelets across the mesh. two PEs sit side by side. the left PE computes a partial result and routes it eastward. the right PE receives those wavelets from the west and accumulates. routing is defined at compile time. both operations are asynchronous and activate a task when they finish. you're physically routing data across silicon at 1 clock cycle per hop. cerebras is an extremely technically interesting & alien beast. reports 95-210x speedups over H100 on stencil computations. 3,000 tokens/sec inference on gpt-oss-120B. $10B+ deal with OpenAI for 750 MW of inference infrastructure. very exciting times for alternative accelerators! deep dive 1/6 by @gpuemi
wafer tweet media
English
45
129
1.3K
176.7K
avv
avv@avithaldas·
@blader Any feedback on the speed when used with the API? Sam had mentioned it does 750 toks / sec, but no mention of TTFT anyplace yet
English
0
0
0
440
avv
avv@avithaldas·
@thsottiaux Any feedback on the speed when used with the API? @sama had a comment on the 750 toks / sec, but no mention of TTFT anyplace yet (for real time voice turn use cases)
English
0
0
2
334
avv
avv@avithaldas·
@sama Any feedback on the speed - specifically TTFT? 750 toks / sec is awesome, but how would it play with low latency use cases @OpenAIDevs ?
English
0
0
0
211
Sam Altman
Sam Altman@sama·
GPT-5.6 sol launches thursday! happy building
English
1.9K
1.9K
32.2K
2.1M
avv
avv@avithaldas·
@skirano Any feedback on the speed when used with the API? @sama had a comment on the 750 toks / sec, but no mention of TTFT anyplace yet (for real time voice turn use cases)
English
0
0
0
302
avv
avv@avithaldas·
@thsottiaux Visual dashboards : having a visual front to a repeatable process within codes, that doesnt require me to create a app, and view in the browser.
English
0
0
0
464
Tibo
Tibo@thsottiaux·
What is something that you feel is surprising that Codex still can't do well and we should have gotten right a while ago?
English
3K
42
3K
568.1K
avv
avv@avithaldas·
A more intuitive workflow layer: define personas, responsibilities, and handoffs clearly. Instead of managing everything through long prompts, users could create role-based workflows. For example, a designer explores the experience, an architect defines the solution, and a developer builds it. Each persona has a clear responsibility, relationship to the others, and expected output.
English
0
0
0
464