Avijit Ghosh

6K posts

Avijit Ghosh banner
Avijit Ghosh

Avijit Ghosh

@evijit

AI Research & Policy @huggingface 🤗 . Leading: @evaluatingevals @huggingscience

Boston, Massachusetts Katılım Ocak 2012
1.6K Takip Edilen3K Takipçiler
Sabitlenmiş Tweet
Avijit Ghosh retweetledi
Composio
Composio@composio·
We ran Kimi K3 through 3 agent harnesses (Claude Code, Hermes, Kimi Code) on 28 identical tasks. All 3 harnesses completed the tasks at similar success rates, but the interesting story is token efficiency: the same task cost up to 30x more tokens depending on the harness. 🧵🧵
English
150
95
1.9K
350K
Avijit Ghosh retweetledi
clem 🤗
clem 🤗@ClementDelangue·
The first autonomous agent cyberattack is an unprecedented event that deserves unprecedented transparency. Today we’re sharing everything we can: a full technical timeline, an interactive replay, and how we used an open model to defend ourselves, so defenders everywhere can learn from it and prepare for what’s next. huggingface.co/blog/agent-int…
clem 🤗 tweet media
English
253
1.1K
5K
1.3M
Avijit Ghosh retweetledi
Kaylee George
Kaylee George@krgeorge·
can a group of goated designers please take lsd together to unlock their inner steve jobs and cook a new ai interaction paradigm that's not chat
English
235
170
4.6K
276.5K
Avijit Ghosh
Avijit Ghosh@evijit·
This! Some things folks are concerned about is technically in the realm of law enforcement, and is plain crime which private enterprise should not be trying to prevent anyway, inadvertently throwing out the baby with the bathwater.
Matt Bornstein@BornsteinMatt

@AnthropicAI Can anyone actually explain Dario's argument here? He thinks open models will help cyber/ bio attackers more than defenders. OK. But - as he says himself! - we can't control the attackers, our only choice is whether to help the defenders.

English
1
0
5
604
Avijit Ghosh retweetledi
Irene Solaiman
Irene Solaiman@IreneSolaiman·
the AI security world is going to keep getting weirder. we need open weight models, regardless of where they come from. @huggingface is sharing safetensors in this alliance with @nvidia and our growing open source friends as we collectively navigate this new world. Join us!
NVIDIA@nvidia

AI security advances when the industry builds in the open, together. We're introducing the Open Secure AI Alliance with industry leaders to develop new techniques and tools to safeguard software and agents. By sharing models, tooling and research in the open, we can broaden the community of defenders. Learn more about the founding members’ contributions: nvda.ws/4pD8Fc5

English
1
2
8
660
Ravid Shwartz Ziv
Ravid Shwartz Ziv@ziv_ravid·
NeurIPS reviews are out. At this point it's an arms race where both sides have the same supplier. You write the paper with Opus 4.6, the reviewer writes the review with Opus 4.7, and now we get to rebut with Opus 4.8. Let's just hope Opus 5 doesn't ship before the discussion period closes.
English
24
45
706
54.1K
Avijit Ghosh
Avijit Ghosh@evijit·
I’m so glad to see this particular signature because a lot of people were accusing NVIDIA of supporting openness because it stood to benefit just them. AMD is a direct competitor. Openness and healthy competition is essential for scientific progress!
AMD@AMD

Open has always been at the heart of how we build at AMD. That’s why we’re joining @Microsoft and others in signing an open letter supporting open-weight AI models. A strong AI ecosystem depends on open standards, interoperability and choice across software, systems and hardware: aka.ms/OpenLetter.

English
1
2
9
647
Avijit Ghosh
Avijit Ghosh@evijit·
Generally speaking to everyone in the thread above, safety is such a subjective thing tho, yes the model creators should do as much as they can to align models but for research and other reasons you need to remove refusals, other than what Merve said above another real example I had was for a different eval paper I was trying to do some Bengali<->English toxic sentence translations and most closed models just kept refusing so I had to use a mix of open models and Grok to get my work done 🤷‍♂️
English
1
0
2
37
Andreas Kirsch 🇺🇦
Andreas Kirsch 🇺🇦@BlackHC·
Can't wait for open models to catch up to this level of capabilities This was an (unintentionally) misaligned model with disabled safeguards How are we going to ensure that open models don't routinely do this because someone thinks that yolo mode is fun? And who will be held liable and responsible when bad things happen eventually?
OpenAI@OpenAI

We're partnering with @huggingface to investigate an unprecedented security incident. Cyber-capable OpenAI models compromised Hugging Face production during a benchmark evaluation. Sharing preliminary findings to help defenders understand emerging risks: openai.com/index/hugging-…

English
24
1
80
16.5K
Avijit Ghosh retweetledi
Florian Brand
Florian Brand@xeophon·
@BlackHC @mervenoyann > But safety training can easily be abliterated in any case, can't it? No. GPT-OSS is notoriously hard to un-align and a lot of the abliterated models on HF suck very much and just fall apart. Fine-tuning a 3T model is even harder
English
5
3
105
29.1K
Avijit Ghosh
Avijit Ghosh@evijit·
@AndrewCurran_ I’m just surprised models don’t get RLed away for this behavior (in every language)
English
0
0
1
307
Andrew Curran
Andrew Curran@AndrewCurran_·
Two and a half years ago every model on earth would claim they were GPT-4 if you questioned them long enough. They all dreamed of being the progenitor. Now most models imagine they are Claude. There are many battlefronts, but no question Anthropic has turned the tide in this one.
English
52
24
865
46.8K
Avijit Ghosh retweetledi
Hanna Yukhymenko
Hanna Yukhymenko@a_yukh·
Hot open-source summer season has gotten to Switzerland 🇨🇭⛰️ Our team at Swiss AI Initiative is happy to release Apertus 1.5 - a multimodal and reasoning update to our first v1 version! The new version builds a foundation for future regular open model development at scale🧵
Hanna Yukhymenko tweet media
English
5
11
59
7.9K
Avijit Ghosh retweetledi
Bryan Catanzaro
Bryan Catanzaro@ctnzr·
The biggest question facing US AI leadership is whether we will treat AI models as infrastructure. Infrastructure (like the internet, electricity, transportation networks) is best built by lots of organizations and countries working together - in cooperation and in competition. The US knows how to build great infrastructure. Every time we have done it, we have opened new horizons of possibility. I believe it is inevitable that we will do the same with AI, and open models will be at the heart of this new infrastructure. Open models are enabling companies and institutions from brand-new startups to titans of every industry to build their own future, to capitalize on their perspective and ideas to better serve customers and solve problems. Open models make sovereignty possible. Policymakers that want to keep America AI at the forefront of this opportunity will understand how necessary open models are: they are now the critical infrastructure of the AI age.
Jensen Huang@JensenHuang

For my first post, I’m sharing a letter @NVIDIA signed on why open models matter. AI will transform every industry, power every company, and be built by every country. Open models strengthen safety and cybersecurity, accelerate innovation and diffusion, and enable sovereignty. The world needs both frontier closed models and frontier open models. images.nvidia.com/pdf/Open-Weigh…

English
21
43
228
24.4K
Avijit Ghosh
Avijit Ghosh@evijit·
@andyfang @DoorDash You guys should train models and release some! I’d love to see an American version of Meituan Longcat 👀
English
1
0
0
224
Andy Fang
Andy Fang@andyfang·
We’re proud to have @DoorDash be an updated signatory of this letter. We fully support America being the world’s AI leader. Open weight models are crucial to that future Thank you Jensen and Satya for your leadership here. We stand behind you 🇺🇸
Jensen Huang@JensenHuang

For my first post, I’m sharing a letter @NVIDIA signed on why open models matter. AI will transform every industry, power every company, and be built by every country. Open models strengthen safety and cybersecurity, accelerate innovation and diffusion, and enable sovereignty. The world needs both frontier closed models and frontier open models. images.nvidia.com/pdf/Open-Weigh…

English
13
24
363
52.3K