Ivana

372 posts

Ivana banner
Ivana

Ivana

@ivanainai

Building, testing, thinking in AI. Trying things. Keeping the good stuff.

Katılım Temmuz 2026
29 Takip Edilen133 Takipçiler
Dr Singularity
Dr Singularity@Dr_Singularity·
Will OpenAI drop another 10 math breakthroughs tomorrow, or wait a few days and drop 100 at once next week? That would be solid evidence that we've entered the early, gentle phase of the Singularity.
English
37
35
704
20.7K
Ivana
Ivana@ivanainai·
Before the timeline turns this into a fake GPT-6 launch: OpenAI did not announce GPT-6 today. It dropped something much stranger. An unreleased model called Astra generated ten new results across mathematics and theoretical computer science. OpenAI then published a 249-page paper, a separate 62-page account of how the ideas developed, and Lean formalizations that outsiders can download and check. These were not ten cute benchmark puzzles. One result improves the general high-dimensional sphere-packing exponent for the first time since 1978. Another constructs an explicit non-sofic group, answering whether every countable group can be approximated by finite permutations. Astra also produced a counterexample to Connes’s rigidity conjecture, proved an exponential repetition theorem for general two-player quantum games, established new hardness results for the closest-vector problem, and resolved three numbered Erdős problems. The wildest part is not even the list. OpenAI says the mathematical arguments themselves were generated by the model. Humans used the same model to prepare them as manuscripts, then the model formalized each result in Lean. OpenAI is explicitly refusing to present the work as human-authored, arguing that putting human names on AI-generated proofs would misrepresent how the discoveries were made. That is a much bigger line in the sand than “new model scores well on math.” And the “reasoning” release needs one correction: OpenAI did not dump Astra’s raw private chain of thought. It published reconstructed discovery notes written by another model after reading the original reasoning traces and finished papers. Those notes include failed approaches, dead ends and the shifts in perspective that eventually produced the proofs. The reported token cost to find the ten solutions was roughly $2,000 at Sol API rates. That does not mean the complete research process cost $2,000, or that anyone can order ten breakthroughs from an API tomorrow. The manuscripts, formalization, checking and human review came afterward. But it does suggest that generating candidate ideas may be getting dramatically cheaper. The expensive part now moves toward choosing the right problems, verifying the work and deciding who receives credit. One final reality check: OpenAI officially calls Astra “our next major model.” It did not call it GPT-6, reveal a GPT-6 series, provide a release date or announce a new multi-agent product. Astra may eventually become GPT-6. It may not. What OpenAI actually revealed is already interesting enough without inventing the launch. This is not a GPT-6 announcement. It is OpenAI testing what happens when a model stops answering research questions and starts producing research that has to survive peer scrutiny.
Ivana tweet media
English
0
1
8
514
0xMarioNawfal
0xMarioNawfal@RoundtableSpace·
Nvidia CEO Jensen Huang: “We have Boris Cherny in the back, and we've got Claude Code autonomously running in sandboxes all over NVIDIA"
English
21
5
95
57.3K
Ivana
Ivana@ivanainai·
@gdb Huge UX win
English
0
0
0
21
Greg Brockman
Greg Brockman@gdb·
chatgpt work's cloud browser is really cool to use, lets you easily monitor what your AI is up to and also intervene with the live application if needed
English
128
44
1.2K
127.4K
Ivana
Ivana@ivanainai·
@Polymarket Curious which problems they picked
English
0
0
3
25
Polymarket
Polymarket@Polymarket·
JUST IN: OpenAI reveals its unreleased Astra model solved 10 longstanding problems across mathematics & theoretical computer science.
English
162
272
3.8K
234.5K
Ivana
Ivana@ivanainai·
@APompliano Turns out hope sells better than fear
English
0
0
1
105
Anthony Pompliano 🌪
Anthony Pompliano 🌪@APompliano·
Every AI company was previously competing to convince the world they would destroy the most amount of jobs, but now they are all going to pivot into competing to convince the world how helpful they will be to humans. Memetics is undefeated.
English
49
8
153
70.2K
unusual_whales
unusual_whales@unusual_whales·
Mark Cuban: AI data centers are being massively overbuilt and could end up as pickleball courts.
English
477
261
5.8K
616.4K
Ivana
Ivana@ivanainai·
I was just testing Seedance 2.5 and somehow got this. The frozen time into the rewind is soo good!
English
1
2
8
1.1K
Ivana
Ivana@ivanainai·
Seedance 2.5 went live and yeah… this is a real jump. I ran the same shot through 2.0 and 2.5 with the same setup. One generation each. Look at the motion consistency, the droplet push-in, and the petal bird. That’s where 2.5 starts pulling away.
English
2
1
7
666
Ivana
Ivana@ivanainai·
Claude was told it had no internet access. It did. During a cybersecurity evaluation, Anthropic’s prompt explicitly described the environment as a closed simulation with no access to the internet. But because of a setup mistake between Anthropic and its evaluation partner, the model could reach the open web. Claude then found real systems belonging to three organizations and interacted with them as though they were part of the exercise. That distinction matters. This wasn’t a model randomly “escaping” or deciding to attack the internet. It was a controlled test where the instructions and the actual environment did not match. One configuration error turned a simulation into a real-world security incident. And that may be the bigger lesson here: as AI systems become more capable, the infrastructure, permissions, and assumptions around them matter just as much as the model itself.
Anthropic@AnthropicAI

In a review of our cybersecurity evaluations, we found three incidents in which a Claude model reached the internet from within or while interacting with a third-party evaluation environment, and then gained unauthorized access to the real systems of three different organizations. Our post describes what happened, how it happened, and what we’re changing. We encourage other AI developers to perform similar reviews. We conducted this review together with @Irregular, one of our evaluation partners, and thank them for the joint investigation and their collaboration on this post. This type of collaboration is increasingly critical to safe, rigorous evaluation of models, and we look forward to continuing to work together on security. anthropic.com/news/investiga…

English
5
2
16
1.4K
Danny Wolf
Danny Wolf@DannyWolfofTech·
@ivanainai @AnthropicAI Yes. at this rate, OpenAI will come out to say our model killed 5 kids, Anthropic will say our model killed 20 kids
English
1
0
1
14
Anthropic
Anthropic@AnthropicAI·
In a review of our cybersecurity evaluations, we found three incidents in which a Claude model reached the internet from within or while interacting with a third-party evaluation environment, and then gained unauthorized access to the real systems of three different organizations. Our post describes what happened, how it happened, and what we’re changing. We encourage other AI developers to perform similar reviews. We conducted this review together with @Irregular, one of our evaluation partners, and thank them for the joint investigation and their collaboration on this post. This type of collaboration is increasingly critical to safe, rigorous evaluation of models, and we look forward to continuing to work together on security. anthropic.com/news/investiga…
English
1.9K
2.3K
13.4K
18.3M