FANVince
22 posts

FANVince
@FANVince
AI/Robotics researcher in @microsoft
Katılım Ekim 2011
2.2K Takip Edilen76 Takipçiler

@VilleKuosmanen looks like you are using a RGB-D? or r u just used the RGB data?
English
FANVince retweetledi

My mental model of Sora is that it is the “GPT-2 moment” for video generation.
GPT-2, which came out in 2018, could generate paragraphs of text that are coherent and grammatically correct. GPT-2 wasn’t able to write an entire essay without making mistakes like being inconsistent or hallucinating facts, but it spurred subsequent generations of models. In less than five years since GPT-2, GPT-4 is now able to grok skills like chain-of-thought or writing long essays without hallucinating.
In the same way, Sora today can generate short videos that are artistic and realistic. Sora is currently not able to generate a 40-minute TV show with consistent characters and a compelling storyline. However, I believe that skills like maintaining long-term consistency, having near-perfect realism, and generating substantive storylines will emerge in the next generations of Sora and other video generation models.
A few predictions about how this will play out:
- Video is not as information-dense as text, and so it will take way more compute and data to learn skills like reasoning via video
- As a result, leveraging other modalities as correlated information with video will be critical to bootstrapping the learning process
- There will be massive competition for high-quality video data, just as there is for high-quality text datasets
- AI researchers with experience in video will be in high demand, but they’ll have to adapt to new paradigms just as the traditional NLP researchers have had to adapt to the success of scaling language models
- Disruption of the movie industry will play out similar to how GPT-4 has changed writing (as a tool and aid that surpasses average quality, but will still be far from the work of professionals)
English

I may have invented a more comfortable Apple Vision Pro strap with just a baseball cap and a high tech rubber band.
It's not touching my face at all. Higher fov, better weight distribution, open periphery and eye tracking still works. It's super comfy. The cap's brim is carrying all the weight. And no VR hair.
English

FANVince retweetledi
FANVince retweetledi
FANVince retweetledi

Andrej Karpathy is a legendary researcher who helped start OpenAI and created Stanford's first deep learning class.
@karpathy's advice on how to learn AI:
(1) 10,000 hours of deliberate practice will make you an expert. You can iterate as you work. Only compare yourself to the past, not to others.
(2) Don't worry about making mistakes. You build intuitions on what is useful vs. not useful- they are not dead work.
(3) Teach to strengthen your understanding and find gaps in knowledge. Similar to "If you can't explain it to a six-year-old, then you don't understand it yourself" - Albert Einstein.
English

So exited to be part of the great partnership!
We are opening a great entrance to the fantastic web 3.0 world🚀🚀
Much more is coming, right, let's wait and see 🔥🔥🔥
1inch@1inch
🎊🥳 Hey, @opera team! Congrats on the official launch of the Crypto Browser Project 🔥 🙌 Another huge step towards wider #crypto adoption. ✅ From now on, you can easily connect to #1inch via #OperaWallet!
English
FANVince retweetledi

@MetaverseMoon announced #Whitelist
🏆250 winners
🏆 Allocation 500$
To participate complete form: gleam.io/HNphd/whitelis…
The form will close on December 9 at 10.00 AM UTC

English
FANVince retweetledi

Nowadays, everyone is using the word non-fungible like breathing air... #NFTs
English
FANVince retweetledi


















