
JUST IN: Top 10 US stocks now account for 43% of S&P 500
Russ
369 posts

@rl_ipynb
Data Specialist - I study AI, Finance and Data Stuff

JUST IN: Top 10 US stocks now account for 43% of S&P 500

(1) Today we're releasing Muse Spark 1.1 -- a strong agentic and coding model at a very low price. It's available through our new Meta Model API and in Meta AI.


Fable 5 is back.


Exclusive: New Claude app strings tie Fable 5 usage credits to identity verification. The strings show Fable 5 is being put behind the usage credit system, billed outside your plan. Identity verification is referenced in the same update: "Your credits will be added once your identity is verified." Anthropic previously called identity verification unrelated to Fable and limited to flagged accounts. These strings showed up alongside the Fable 5 credit changes.

Herculaneum fused scroll read in full. scrollprize.org/firstscroll

This more accurate now


BREAKING: President Trump says US might buy equity stakes in AI companies

Introducing Agent Arena: real-world agentic evals at scale. How do you evaluate agents doing actual work? We measure millions of live sessions where real users accomplish real tasks. On Arena, models now get web search, filesystem, and terminal tools to complete complex workflows: writing code, creating slide deck, researching the web, building apps, and analyzing documents. Every session produces rich signals. Users iterate with the agent turn-by-turn: approving, editing, correcting, praise or expressing frustration. The environment gives feedback too: shell errors, tool failures, recovery attempts, and more. Our leaderboard measures each model's agentic performance using causal inference across five signals: task success, steerability, error recovery, user praise vs. complaint, and tool hallucination. This leaderboard snapshot is built from 300K+ tasks, 2M+ tool calls, and 40M lines of code by agents. Top labs in Agent Arena: - #1 @OpenAI: GPT-5.5 (High) - #2 @AnthropicAI: Claude-Opus-4.7 (Thinking) - #3 @Zai_org: GLM-5.1 - #4 @GoogleDeepMind: Gemini-3.1-Pro - #5 @Kimi_Moonshot: Kimi-K2.6 More analysis in the thread, with the full technical blog below.
