Alex Shtoff

11.6K posts

Alex Shtoff banner
Alex Shtoff

Alex Shtoff

@AlexShtf

Ph.D. Principal Scientist @ TII. Ex @YahooResearch. I do machine learning ∩ numerical methods ∩ SW dev. Author of https://t.co/MkW8DDKamf

Israel Katılım Ağustos 2012
288 Takip Edilen1.4K Takipçiler
Sabitlenmiş Tweet
Alex Shtoff
Alex Shtoff@AlexShtf·
New post in my "Eigenvalues as models" series. The series explores a simple but weird predictive model: build a learned symmetric matrix from the input features, then use one of its eigenvalues as the non-linearity. This time I look at converence speed during training, which we saw in previous posts that can be occasionally slow, and it turns out to be related to the expressiveness of the model. The middle eigenvalue is expressive, but when nearby eigenvalues collide, the eigenvectors can change abruptly. Since eigenvalue gradients come from these eigenvectors, training gets jumpy. Somewhat unexpectedly, a path forward is to smooth the sharp eigenvalue objective first using a tool well-known to the optimization community, but somewhat little known to many machine-learning practitioners - the Moreau envelope. alexshtf.github.io/2026/07/01/Spe…
English
1
8
74
4.2K
Alex Shtoff
Alex Shtoff@AlexShtf·
@cremieuxrecueil Or, perhaps, the cause and effect are different - men who earn more tend to get married?
English
0
0
1
42
Crémieux
Crémieux@cremieuxrecueil·
The gender wage gap is mostly about married men doing an incredible job earning more than everybody else.
Crémieux tweet media
English
524
1.2K
17.6K
3M
Alex Shtoff
Alex Shtoff@AlexShtf·
@Bondisrael_ הייתי אומר שגברים שמתחילים בגיל מוקדם יותר להרוויח יותר הם אלו שמתחתנים, ואלה שנוטים להרוויח פחות נשארים רווקים.
עברית
0
0
0
21
gausts (s/acc)
gausts (s/acc)@gausts_pgs·
@AlexShtf @testinprodcap because it would be stupid not to. chances are we get rolling self-attention and another encoder gets bolted on somewhere. why scrap something that almost works?
English
1
0
0
40
Alex Shtoff
Alex Shtoff@AlexShtf·
@jm_alexia What would be a reasonable time to "do research" until a release?
English
0
0
2
1.1K
Alex Shtoff
Alex Shtoff@AlexShtf·
@willccbb Either that, or they want more research (e.g., more experiments, faster feedback for experiments).
English
0
0
1
228
Alex Shtoff
Alex Shtoff@AlexShtf·
@CauraAI I know what they mean. I asked because they contradict your claim about a model being first or third.
English
1
0
1
35
Caura
Caura@CauraAI·
@AlexShtf They're there to show the gaps that actually separate models from the ones that are noise — where intervals overlap, we'd treat the ranking as a tie rather than a result.
English
1
0
1
1.4K
Caura
Caura@CauraAI·
Five frontier models, 100 questions, judged blind — by each other. Opus 5 wins at 8.87. Fable 5 takes third without being beaten: its own guardrail handed in 4 blanks — two were high-school biology. The numbers, the blanks, the self-bias ↓
English
34
86
1.1K
2.5M
Alex Shtoff
Alex Shtoff@AlexShtf·
It appears agentic systems are going backward. First we had function calling. Now we have loops. Will we soon have `goto`?
English
1
0
4
143
Matthew Lutz
Matthew Lutz@MattLutzPhi·
Beginning to suspect that math just isn't very hard.
English
54
21
420
32.1K
Alex Shtoff
Alex Shtoff@AlexShtf·
@Mononofu @JensenHuang Are you aware of the VAST amount of open-source software and open-weights models that came from NVidia?
English
0
0
3
140
Julian Schrittwieser
I’m so excited that @JensenHuang is a believer in open source now, looking forward to the CUDA and GPU driver open source release!
Jensen Huang@JensenHuang

For my first post, I’m sharing a letter @NVIDIA signed on why open models matter. AI will transform every industry, power every company, and be built by every country. Open models strengthen safety and cybersecurity, accelerate innovation and diffusion, and enable sovereignty. The world needs both frontier closed models and frontier open models. images.nvidia.com/pdf/Open-Weigh…

English
1.7K
352
5.4K
6.5M
Alex Shtoff
Alex Shtoff@AlexShtf·
So, why doesnt X on Android work properly with folded phones and rotated screen? What a mess.
English
0
0
0
172
Nadav Finebuch
Nadav Finebuch@NadavFinebooch·
@Hak_Tsu @KseniaSvetlova בניגוד לנראטיב שספרו לנו והאמנתי בו, אמריקע חוללה את המלחמה באוקראינה לא פוטין. אגב אתה מכיר את העובדה שהשחקן המושחת, זלנסקי הוא בעצם יאיר נתניהו קטן, משתולל עם גברים , עושה באף תוך כדי שהוא שודד את המדינה שהוא החריב ?
עברית
2
0
0
239
Ksenia Svetlova كسنيا سفطلوفا
ולנטינה (וליה) סביצקי, רעייתו של איגור סולוביי, חבר אוקראיני יקר, נהרגה אמש מפגיעה של טיל בליסטי רוסי. פגשתי אותה לפני כשנתיים כאשר ביקרו בישראל. דיברנו על המלחמות שלנו, ועל ההכרח לשתף פעולה (בעלה, איגור, ריכז את הפעילות הממשלתית נגד דיסאינפורמציה רוסית). המלחמה באוקראינה נמשכת כבר יותר מארבע שנים והיא גובה מחירים כבדים מהאוקראינים, יום ביומו. הפעם זאת מישהי שזכיתי להכיר.יהי זכרה ברוך.
Ksenia Svetlova كسنيا سفطلوفا tweet media
עברית
36
39
1K
13K
Damien Teney
Damien Teney@DamienTeney·
@andrewgwils I suspect there are also a lot of crap submissions where the authors are unknowingly reinventing the wheel and/or failed to connect their ideas with the existing literature.
English
1
0
2
369
Uncle Bob Martin
Uncle Bob Martin@unclebobmartin·
I’m significantly older than you. I started coding in the late 60s. My current strategy is to not read any of the code written by my agents. That’s the only way I can take advantage of their productivity. What I do instead is to surround the agents with extreme constraints. Unit tests, gherkin tests, QA procedures, quality metrics, mutation testing, test coverage, and a plethora of others. In the end, I have very high confidence in the code they produce because they’ve had to run the gauntlet of all of my constraints and tests.
English
564
1.9K
18.2K
4.9M
Ori Pomerantz
Ori Pomerantz@ori_pomerantz·
I am trying to use Claude to help me write something, but I just don't feel comfortable letting it edit my files. Does anybody else feel the same? If I am responsible for code, I NEED to understand it, psychologically if for no other reason. Started programming in 1983. Old?
English
234
11
636
186K
Alex Shtoff
Alex Shtoff@AlexShtf·
The main problem is that AOI is essentially software, and software is hackable. Of course, there's this issue of nonstandardness. For example, a paper whose objective is presenting a mathematical object that is widely applicable, rather than the standard problem--> solution --> evidence (empirical/theoretical) would probably be simply rejected. I got a few such non-standard papers as a reviewers, and I dont see how AI would perform well.
English
0
0
0
254
Peter Richtarik
Peter Richtarik@peter_richtarik·
Every single review I have handled at NeurIPS as an AC is **much worse** than a well executed AI review. By "much worse" I do not mean 20% or 50%. I mean a factor of about 10-100. At least - 10x more mathematical issues are caught, - 10x more missing key citations are caught, - 10x more novelty claims are invalidated with evidence, - 10x more issues with the experiments are observed, - 10x more typos and grammatical issues are fixed, - 10x more inconsistencies are found, and so on. In many cases I have handled over the last 10 years, the human reviews are so bad in comparison that the improvement factor is closer to 100. (The best human reviews I received over the years are worse than an AI review I can generate today, but still entirely good enough to make the correct decision. On the other hand, an AI review is often an almost comprehensive summary of all key issues --- something a human almost never has time to deliver --- and such feedback is immensely useful to the authors) At this point, we would be much better to just 1) give one very thorough ChatGPT review to all submitted papers (I am mainly talking about my fields - optimization, AI, machine learning), automatically, and 2) keep asking for a revision or two (to keep the duration of the process within some bounds) until the number of issues decreases to an extent when the AI reviewer is satisfied, in a given time-frame (eg, 1 month). 3) At that point, a human AC can make a decision (to keep an eye on this all should anything go wrong). The role of the AC would be merely to observe and manage the process, and make a final decision, based on the trajectory of the revisions. I can't believe I am saying this -- AI reviews were a nonsense idea even a year ago. The current AI reviews are super-human.
English
34
31
308
131.1K
Tibo
Tibo@thsottiaux·
Should we rename ChatGPT Work to ChatGPT Vibe?
English
2.9K
107
4.8K
776K
Shreyamishraa...
Shreyamishraa...@shryexe·
@AlexShtf Is the community actually unwelcoming, or are we just afraid of rejection?
English
1
0
0
52
Shreyamishraa...
Shreyamishraa...@shryexe·
What's stopping more developers from contributing to open source?
English
61
2
77
7.2K