Bill Demirkapi

1.2K posts

Bill Demirkapi banner
Bill Demirkapi

Bill Demirkapi

@BillDemirkapi

preparedness research @openai

San Francisco, CA Katılım Temmuz 2017
368 Takip Edilen22.4K Takipçiler
Sabitlenmiş Tweet
Bill Demirkapi
Bill Demirkapi@BillDemirkapi·
When OpenAI released ChatGPT, I was among the millions captivated by what we (humanity) achieved. Truly honored to join the mission to accelerate human progress safely @OpenAI Preparedness and stand on the shoulder of giants at a pivotal moment for agentic security.
Bill Demirkapi tweet media
Sam Altman@sama

I am extremely excited to welcome @dylanscandinaro to OpenAI as our Head of Preparedness. Things are about to move quite fast and we will be working with extremely powerful models soon. This will require commensurate safeguards to ensure we can continue to deliver tremendous benefits. Dylan will lead our efforts to prepare for and mitigate these severe risks. He is by far the best candidate I have met, anywhere, for this role. He has his work cut out for him for sure, but I will sleep better tonight. I am looking forward to working with him very closely to make the changes we will need across our entire company.

English
2
1
34
11.4K
Bill Demirkapi retweetledi
Noam Brown
Noam Brown@polynoamial·
Long-running models can solve hard open-ended problems, but their persistence can create safety risks that shorter-horizon evaluations miss. We’re sharing what we learned from studying a long-running model, and how those findings are shaping our approach to evaluations, alignment, monitoring, and user control. openai.com/index/safety-a…
English
119
201
1.8K
545.5K
Bill Demirkapi retweetledi
leo 🐾
leo 🐾@synthwavedd·
🧵 DeepSeek appear to have engaged, or be engaging in, a large-scale operation to collect outputs from proprietary models (including Claude Fable 5) for certain requests via their API as part of a distillation effort. After seeing such claims circulating earlier today, we conducted an investigation into them on our Discord. We found that, when "Deepseek V4" was used within OpenCode - via their official API - for complex prompts (i.e. 3D games) and combined with a knowledge-related query, the model provides virtually identical outputs to Fable 5. CoT structure is also very different from what is typically expected from Deepseek models. Both of these behaviours revert to what is expected for V4 when simpler prompts were used. When "Deepseek V4" was asked to incorporate answers to questions related to cyber or bio tasks that we verified hit Fable's classifiers into its 3D games, outputs tanked in quality. This is very difficult to explain unless the request was routed to Fable and fell back after hitting a classifier. For complex code prompts without anything else mixed in, like 3D games, outputs were remarkably similar to those produced by Fable 5. We were able to produce these results most consistently via OpenCode and the official Deepseek API combined with a prompt that specifies a complex code task. Deepseek have continued to modify their routing system since, as we have observed changes in behaviour and CoT style compared to those seen previously. Our investigation was conducted from 7AM-8AM PT.
leo 🐾 tweet media
English
183
115
1.5K
440K
Bill Demirkapi retweetledi
AI Security Institute (AISI)
AI Security Institute (AISI)@AISecurityInst·
Our first public analysis of the open/closed weight gap in frontier cyber capabilities finds it is 4–7 months with GLM-5.2 and DeepSeek V4-Pro, narrowing from 6–10 months through most of 2025. Advanced capabilities are reaching less safeguarded open models faster than before. 🧵
AI Security Institute (AISI) tweet media
English
11
56
240
67.4K
Bill Demirkapi retweetledi
f4mi ‼️
f4mi ‼️@f4micom·
not sponsored or anything i just wanted to say that i tried unleashing 5.6 sol on a few reverse engineering projects i had started ages ago and got stuck on because i wasn’t familiar with powerpc and holy shit this thing just keeps going i haven’t ever seen anything like this
f4mi ‼️ tweet media
English
28
5
459
21.1K
Bill Demirkapi retweetledi
OpenAI
OpenAI@OpenAI·
GPT-5.6 launches with ultra mode, our highest-performance setting to accelerate your most ambitious work by coordinating multiple agents to work in parallel. It trades higher token use for stronger and faster results on demanding tasks.
OpenAI tweet media
English
6
24
435
56.1K
Bill Demirkapi retweetledi
AI Security Institute (AISI)
AI Security Institute (AISI)@AISecurityInst·
Most AI agent evaluations boil capability down to one score. But that number hides a key choice: how much compute the agent was allowed to use. New work from our Science of Evaluation team shows why that matters. 🧵
AI Security Institute (AISI) tweet media
English
13
76
412
120.3K
Bill Demirkapi
Bill Demirkapi@BillDemirkapi·
@__0xhorror__ I agree scaling test time compute generally leads to better outcomes, but the limitation was not our choice. ExploitBench uses a 300 turn budget. Going beyond would violate the construction of the evaluation. exploitbench.ai/exploitbench.p…
Bill Demirkapi tweet media
English
0
0
3
495
_horror
_horror@__0xhorror__·
I see what they are doing here lol. The tuned 5.6 sol's max test time compute to achieve just below mythos but at vastly superior token efficiency. Look at that its a straight line, thy could blow way past it if they inference scaled it.
_horror tweet media
English
27
36
1K
126.8K
Dean W. Ball
Dean W. Ball@deanwball·
I am pleased and honored to announce that, on July 6, I'll be joining @OpenAI as leader of a new team called Strategic Futures. Our mandate will be to help the company's leadership shape frontier AI policy. There is a ton of work to do, and I'm excited to get started.
Dean W. Ball tweet media
English
383
184
3.1K
605.1K
Bill Demirkapi retweetledi
Sam Altman
Sam Altman@sama·
theUSshould lead on AI by continuing to develop the very best models, making sure they're safe, and getting cyber tools into the hands of trusted defenders. the new EO gets the balance right.
English
626
153
2.7K
338.4K
Bill Demirkapi
Bill Demirkapi@BillDemirkapi·
@tszzl yeah… our comms needs work in several dimensions
English
0
0
0
541
roon
roon@tszzl·
after introducing an elite lawyer friend to 5.5 pro the models do not sell themselves and one of Claude’s great successes has been packaging them up and marketing usecases for many verticals
roon tweet mediaroon tweet media
English
130
27
2.1K
329.1K
Bill Demirkapi retweetledi