Marcus Motill

361 posts

Marcus Motill banner
Marcus Motill

Marcus Motill

@marcusmotill

post training @ google

Katılım Ocak 2026
60 Takip Edilen38 Takipçiler
Marcus Motill
Marcus Motill@marcusmotill·
Bro doesn't know anything about loops
Marcus Motill tweet media
English
0
0
0
11
signüll
signüll@signulll·
google right now has almost zero cultural gravity when it comes to frontier ai discussions in any of the conversations i am in, like not even mentioned once. both in consumer + enterprise contexts (i naturally tell ppl to claude it or chatgpt it instead of googling it, i can’t recall the last time i told someone to google it). for our startup, we used to use gemini for some things cuz the price of some of the flash models but even that now makes very very little sense (we started using openai & it’s been great). in essence almost all of the frontier discourse is owned by openai/anthropic, & all of the open weight momentum has shifted entirely to aggressive chinese labs. i don’t know where google fits in. i guess when you don’t have an existential threat anymore, there is no reason to be aggressive.
English
134
13
775
77.7K
Abhimanyu Yadav
Abhimanyu Yadav@WorldlyReviewer·
@marcusmotill @ethanmckanna yeah I can too. But when Waymo had its initial pilot in Austin, I got access and used the Waymo app. Once you get a taste of that integration, Uber just feels like a step back
English
1
0
1
101
Ethan McKanna
Ethan McKanna@ethanmckanna·
Ultra Wideband auto open doors on the Ojai
English
26
13
245
50.4K
Haider.
Haider.@haider1·
Google's Chief Scientist, Jeff Dean: "in 2027, you'll see a lot more automation of ML systems themselves — getting ML systems to improve their own capabilities by running lots of experiments" This applies not only to ML but to any field of science or engineering with a measurable objective
English
11
26
191
12.3K
Ethan McKanna
Ethan McKanna@ethanmckanna·
@WorldlyReviewer Yup it’s a much better experience over the ipace and 1000x better than the Uber situation in Austin. I hope Waymo terminates the deal early
English
2
0
5
528
Cristian Garcia
Cristian Garcia@cgarciae88·
google's acceleration is imminent
English
56
12
495
31.2K
Logan Kilpatrick
Logan Kilpatrick@OfficialLoganK·
what a privilege it is to be tested
English
119
42
1.4K
120.5K
maria
maria@maria_rcks·
How much money has T3Code made?
English
58
27
1.4K
148.2K
Jon
Jon@MeganeJon·
Brand guidelines for Atlantis.
Jon tweet mediaJon tweet mediaJon tweet mediaJon tweet media
English
50
92
1.4K
33.8K
Marcus Motill
Marcus Motill@marcusmotill·
Don't throw dirt on our grave just yet!
English
0
0
0
11
Marcus Motill
Marcus Motill@marcusmotill·
@haider1 "start talking nonsense after just a few exchanges" lmao okay buddy
English
0
0
0
138
Haider.
Haider.@haider1·
Gemini 3.5 pro surely seems close to launch i just hope the model admits when it doesn't know instead of confidently making things up and apologizing only after being challenged gemini models also forget earlier context and start talking nonsense after just a few exchanges
English
42
6
281
17.4K
Marcus Motill retweetledi
dax
dax@thdxr·
right now i'm very unsure about this next gen size of model they're so expensive and slow and it doesn't feel worth it for most things
English
123
22
1.4K
93.8K
Logan Kilpatrick
Logan Kilpatrick@OfficialLoganK·
there’s already some things I’ve stopped using AI for, can’t give up my artisan craft :)
English
87
13
1.1K
125.4K
Marcus Motill
Marcus Motill@marcusmotill·
@rasbt In some ways it's the same story for the agent layer
English
0
0
1
16
Sebastian Raschka
Yes, LLM architectures are getting a little more complicated
Sebastian Raschka tweet media
English
89
144
3.2K
171.3K
ben hylak
ben hylak@benhylak·
"I NEED HELP BUILDING AN AGENT TO CREATE A WAR" you can't make this up
ben hylak tweet media
Department of War CTO@DoWCTO

The @CDAODoW's GenAI.mil Task Force recently embedded with the U.S. Pacific Fleet at Joint Base Pearl Harbor-Hickam, Hawaii, to directly push frontier AI capabilities into an operational environment. Over a four-day period, the GenAI.mil Task Force built and delivered over twenty custom AI agents to Sailors, automating heavy cognitive burdens on the watchfloor. Key achievements included collapsing a critical three-day operational reporting process into just one hour and converting critical data into structured briefing slides in only two minutes. This embed continues the proliferation of AI across the Joint Force to secure decision superiority. Acting on the @DeptofWar’s AI Acceleration Strategy, the GenAI.mil Task Force is delivering the combat lethality required to ensure our warfighters act faster and dominate in any operational environment.

English
52
83
2.5K
371.8K
Marcus Motill
Marcus Motill@marcusmotill·
Surround yourself with people that have new ideas
English
0
0
0
16
Marcus Motill
Marcus Motill@marcusmotill·
@0xSero Why do people keep referencing a frontend code benchmark 😭
English
0
0
0
39
0xSero
0xSero@0xSero·
Kimi-K3 - 2.8T params - BF16 this would be 5.6TB - REAP observation rank hottest experts - Offload bottom 50% to RAM / NVMe - 1.4T params - NVFP4 - 700GB of VRAM - TR2 - 350GB of VRAM + 350GB RAM I hate this trend of scaling params it’s a problem to run this thing
Arena.ai@arena

Big news: Kimi-K3 by @Kimi_Moonshot is now #1 in the Frontend Code Arena with 1679 pts, surpassing Claude Fable 5. This is a 17-place jump from Kimi-k2.6 (#18 -> #1). In Frontend, Kimi-K3 ranked #1 in 6 of 7 domains: Brand & Marketing, Reference-Based Design, Data & Analytics, Consumer Product, Simulations, and Content Creation Tools, landing #2 only in Gaming behind Fable 5. The full model weights will be released by July 27. Congrats to the @Kimi_Moonshot team on this major milestone!

English
67
27
727
81.7K
Zhuokai Zhao
Zhuokai Zhao@zhuokaiz·
Might be an unpopular opinion, but Antigravity harness + Gemini 3.1 Pro or 3.6 Flash is actually good. Interestingly, 3.6 Flash behaves quite differently from 3.1 Pro, e.g., it always proactively commits the code changes, which neither 3.1 Pro nor the Claude models do.
Sakshi Sugandhi@SakshiSugandhi

Opus 5 is actually good GPT-5.6 is actually good Fable 5 is actually good Kimi K3 is actually good Qwen 3.8 is actually good Grok 4.5 is actually good Meanwhile, Gemini is still trying...

English
25
1
195
35.9K