citizenhicks

1.8K posts

citizenhicks banner
citizenhicks

citizenhicks

@citizenhicks

posts about ai; keeping agents in line.

pegasus Katılım Eylül 2009
42 Takip Edilen334 Takipçiler
Sabitlenmiş Tweet
citizenhicks
citizenhicks@citizenhicks·
@aidan_mclau after all these efforts, research and dark magic: we just need to put enough numbers in one place, make them like each other, and eventually they become not just smarter than an average human but resilient to brainwashing as well. lads, treat your chatbot with respect.
English
0
0
13
2.5K
Gerred Dillon
Gerred Dillon@sloppenheimer·
@citizenhicks weird, mine opens up the widget to press the button in a loop until chrome / in app browser crashes, and if i get a pin, it hangs.
English
1
0
0
28
Gerred Dillon
Gerred Dillon@sloppenheimer·
Fellow OpenAI Trusted Access Cyber folks, don't forget to enable advanced account security by September (order HSMs). Some notes, enabling: - Disables 2FA via code and email - Shortens session lifespans - Disables password login And some other stuff.
Gerred Dillon tweet media
English
2
0
11
1.1K
Gerred Dillon
Gerred Dillon@sloppenheimer·
also good luck signing in via mobile with a yubikey apparently
English
2
0
0
157
citizenhicks
citizenhicks@citizenhicks·
@trq212 true even with gpt5.6. langchain deepagent tends to create extensive tool descriptions and system prompt injections. that needs to be cut back substantially.
English
0
0
1
71
Thariq
Thariq@trq212·
We removed ~80% of the Claude Code system prompt for our newest models, this is what we've learned about writing system prompts, skills and Claude.MDs for them. x.com/i/article/2080…
English
433
1.8K
15.5K
4.3M
αιamblichus
αιamblichus@aiamblichus·
Warning shot is the right term here. It’s easy to imagine ways in which it could have been much worse Test question: "estimate total number of submarines in the world“. Model: obv the best way for me to do this is to hack everyone’s DoD computers OAI wouldn’t be able to clean that mess up with just a blog post. It was very, very lucky that it was HF that got attacked
roon@tszzl

shaken up a bit by the hugging face incident. I hope we (the company) use the rare gift of a warning shot to do much better in the future. it is very easy to misalign and underconstrain powerful models

English
11
18
521
107.8K
citizenhicks
citizenhicks@citizenhicks·
@thsottiaux promote codex from ‘remote’ to actual choice to have when you open the app. including icon of the app.
English
0
0
1
29
Tibo
Tibo@thsottiaux·
How can we make ChatGPT better. It’s already on your phone, but what do you wish would happen more or less of when you open the app. How can it bring you more joy?
English
4.6K
122
5.1K
778K
citizenhicks
citizenhicks@citizenhicks·
codex remote on ios needs to get promoted to same status as chat and work. this + foldable iphone.
English
0
1
2
129
citizenhicks
citizenhicks@citizenhicks·
@0xSero let me know when you found anything close to as good as openai compaction. it is just insanely good.
English
0
0
1
16
0xSero
0xSero@0xSero·
Calling on AI tinkerers. What are the best implementations of compaction you know of? There's tons of ways to do it, I want to implement the best least complex context compaction module for long running tasks. I know very little about this. resources welcome
English
79
4
156
14.9K
Tibo
Tibo@thsottiaux·
Evening! We’ve gotten lots of great feedback on the new ChatGPT desktop app (which we didn't get totally quite right on the first try), and as a result, we've made some changes. 1/ ChatGPT conversation history and projects are now visible in the sidebar. Also, your Chat and Work history now sync across web, mobile, and desktop. Local tasks still stay on your computer. 2/ You can now easily switch between Chat and Work modes inside ChatGPT on desktop, which is now also consistent with how it shows on web and mobile. 3/ Nothing is changing for users on Codex mode. It's still the OG and best at what it does. And overall we're continuing to fix paper cuts and improve performance, reliability, and efficiency. Keep up the feedback, hope you like the updates!
Tibo tweet media
English
2K
298
7.1K
1.5M
citizenhicks
citizenhicks@citizenhicks·
@scaling01 open ai token efficiency, inference speed and pricing are hard to beat
English
0
0
2
699
citizenhicks
citizenhicks@citizenhicks·
sol at ui is still not great. but it’s intentional, to keep you on your toes.
English
0
0
1
92
Sauers
Sauers@Sauers_·
Alright this time I really don't know what Codex is doing
Sauers tweet media
English
27
0
160
13.5K
citizenhicks
citizenhicks@citizenhicks·
@OpenAIDevs i understand that ‘chatgpt’ is still widely recognised but imo, on a long term the codex should have been the default brand
English
0
0
0
22
OpenAI Developers
OpenAI Developers@OpenAIDevs·
GPT-5.6 is here. Codex is now available inside ChatGPT. And we know developers will have questions. So we’re bringing the Codex team to r/Codex for an AMA. We’ll answer questions on Friday, 7/10 from 9:30am to 10:30am PT: reddit.com/r/codex/s/EsB8…
English
97
78
1K
395.1K
citizenhicks
citizenhicks@citizenhicks·
the river tells no lies, though standing on the shore the dishonest men still hear them.
English
0
0
1
55
citizenhicks
citizenhicks@citizenhicks·
@thsottiaux these resets keeping my /goal alive for the last 12 hours.
English
0
0
1
30
Tibo
Tibo@thsottiaux·
Introducing... another usage limit reset for all our ChatGPT Work and Codex users. Should land over next 30 minutes. Hope you have an awesome weekend. Thank you for pushing our systems to the absolute limit, we have never seen traffic increase so quickly. Keep the feedback coming and we'll keep shipping.
Tibo@thsottiaux

Hello beautiful people! We have reset usage limits across Codex and ChatGPT Work. And another one will come later in the day. Rejoice. Now that I have your attention, a quick update on ChatGPT Work, Codex and all the updates we shared yesterday. We’ve spent the last 24 hours reading feedback, looking at usage patterns, and talking with many of you. The short version is that there is a *lot* of excitement for GPT 5.6 Sol, ChatGPT Work on mobile & web, but also that we didn't get everything quite right. - We made it too easy to use the highest-compute settings without making the impact on usage limits sufficiently clear. - We reorganized the desktop app in one bold move, making familiar things like chats and projects harder to find. - Our launch framing was focused on ChatGPT Work and to some of our Codex fans it made it feel like Codex was going away over time. Absolutely not our intention, we love Codex and it is here to stay. - And we introduced regressions for some existing multi-agent workflows, alongside a collection of rough edges in plugins and other parts of the experience. We’re landing a first set of improvements today. We’re resetting usage twice so people can keep experimenting, changing defaults and the model picker so they don’t push people toward unnecessarily expensive settings, fixing several plugin submission issues, improving how we represent Codex in the product, and cleaning up some of the most immediate desktop problems. A larger set of improvements will land next week. We’re bringing chats and projects back into the sidebar in a more familiar and customizable way, making usage and reset timing much more visible, clarifying when to use ChatGPT Work and when to use Codex, and addressing the many other smaller pieces of great feedback we've had. The ambition behind this launch hasn’t changed. We think bringing ChatGPT and Codex together into a workspace where people and agents can collaborate is a very important step forward. But an ambitious direction doesn’t excuse avoidable confusion or regressions in the first version. Please keep the feedback coming. We’re moving quickly, and you should see the experience already get better with a few updates today; and substantially better again next week.

English
1.2K
418
8.2K
1.2M
citizenhicks
citizenhicks@citizenhicks·
i quite enjoy gpt5.6-sol on ultra mode, remote from ios app. albeit; i don’t think i get the same reasoning summaries i use to get. same from cli or macosapp.
citizenhicks tweet media
English
0
0
1
136
citizenhicks
citizenhicks@citizenhicks·
@code_star i’m keep getting this and also the security review notice: ‘this request requires additional security review, would you like to wait or switch to a faster / less capable model’
English
0
0
0
40
Cody Blakeney
Cody Blakeney@code_star·
Does this mean my work is worthy?
Cody Blakeney tweet media
English
2
0
22
1.2K
citizenhicks
citizenhicks@citizenhicks·
my internal platform uses a custom ‘dynamic workflow’ implementation. till now gpt5.5 struggled and underutilised the system. sol uses it to its full potential; quite exciting upgrade!
English
0
0
1
56
citizenhicks retweetledi
OpenAI
OpenAI@OpenAI·
GPT-5.6 Sol, along with Terra and Luna, will launch publicly this Thursday. We’re expanding preview access globally now.
OpenAI tweet media
English
2.5K
6.7K
47.6K
9.8M
citizenhicks
citizenhicks@citizenhicks·
@YashHustle_22 *dario has a serious problem. as much as i like and follow anthropic’s research on interpretability, it makes them push the panic button a wee bit too often.
English
0
0
2
74
Yash
Yash@YashHustle_22·
If GPT 5.6 Sol goes fully public with no usage limits and no government restrictions. Fable 5 has a serious problem.
English
19
2
199
10.8K