SFMB

1.5K posts

SFMB banner
SFMB

SFMB

@smendozab

Inventor | Global Labs MX/CA | Bizland Factory

Mexico Katılım Ağustos 2012
1K Takip Edilen341 Takipçiler
Kshitij Mishra | AI & Tech
Kshitij Mishra | AI & Tech@DAIEvolutionHub·
I'm deleting this soon because it's a legit cash-printing formula. 𝗣𝗮𝗶𝗱 𝗖𝗼𝘂𝗿𝘀𝗲 𝗙𝗥𝗘𝗘 (PART - 3) 1. Artificial Intelligence + Data Analyst 2. Machine Learning + Data Science 3. Cloud Computing + Web Development 4. Ethical Hacking + Hacking 5. Data Analytics + DSA 6. AWS Certified + IBM COURSE 7. Data Science + Deep Learning 8. BIG DATA + SQL COMPLETE COURSE 9. Python + OTHERS 10 MBA + HANDWRITTEN NOTES (72 Hours only ) Cost About - $500 To get: - 1. Follow (So I can DM you ) 2. Like & retweet 3. Reply " Send "
Kshitij Mishra | AI & Tech tweet media
English
1.2K
857
3K
262.3K
SFMB
SFMB@smendozab·
@Riazi_Cafe_en Thank you, a great channel to follow, real value!
English
0
0
0
285
New Scientist
New Scientist@newscientist·
Google has released a new AI tool that it says will give scientists 'superpowers' by assisting with literature searches and generating new hypotheses, but does it live up to the hype? #Echobox=1740589174" target="_blank" rel="nofollow noopener">newscientist.com/article/246907…
English
1
13
27
10.8K
SFMB
SFMB@smendozab·
@kibancloud Buenas tardes, he tratado de comunicarme con uds, sería posible tener una llamada? Estamos en México, gracias @kibancloud !
Español
0
0
0
10
kiban
kiban@kibancloud·
Le pusimos una chuleada extrema a nuestro NIP buró y te traemos otra actualización: caché ne NIP burós: Checa todos los detalles en el link: help.kiban.com/en/articles/21…
Español
1
0
0
15
SFMB retweetledi
Darshan 🦖
Darshan 🦖@darshan·
This neuroscientist worked until she was 103. She also: • Won a Nobel Prize at 77 • Became a senator at 92 • Stayed mentally sharp into her 100s Her secret? 5 daily habits that prevented brain aging: 🧵
Darshan 🦖 tweet mediaDarshan 🦖 tweet media
English
254
4.8K
26.7K
4.3M
SFMB retweetledi
Miles Cranmer
Miles Cranmer@MilesCranmer·
DynamicQuantities.jl has hit version 1.0! What started as a dimensional analysis tool for PySR is now a mature physical units package for the Julia community. DQ emphasises type stability and can accelerate both compilation and runtime performance. Thanks to all contributors!
Miles Cranmer tweet media
English
0
3
40
2.2K
SFMB retweetledi
Andrej Karpathy
Andrej Karpathy@karpathy·
To help explain the weirdness of LLM Tokenization I thought it could be amusing to translate every token to a unique emoji. This is a lot closer to truth - each token is basically its own little hieroglyph and the LLM has to learn (from scratch) what it all means based on training data statistics. So have some empathy the next time you ask an LLM how many letters 'r' there are in the word 'strawberry', because your question looks like this: 👩🏿‍❤️‍💋‍👨🏻🧔🏼🤾🏻‍♀️🙍‍♀️🧑‍🦼‍➡️🧑🏾‍🦼‍➡️🤙🏻✌🏿🈴🧙🏽‍♀️📏🙍‍♀️🧑‍🦽🧎‍♀🍏💂 Play with it here :) #scrollTo=75OlT3yhf9p5" target="_blank" rel="nofollow noopener">colab.research.google.com/drive/1SVS-ALf…
Andrej Karpathy tweet media
English
283
1K
7.5K
560.8K
SFMB retweetledi
Andrej Karpathy
Andrej Karpathy@karpathy·
In 2019, OpenAI announced GPT-2 with this post: openai.com/index/better-l… Today (~5 years later) you can train your own for ~$672, running on one 8XH100 GPU node for 24 hours. Our latest llm.c post gives the walkthrough in some detail: github.com/karpathy/llm.c… Incredibly, the costs have come down dramatically over the last 5 years due to improvements in compute hardware (H100 GPUs), software (CUDA, cuBLAS, cuDNN, FlashAttention) and data quality (e.g. the FineWeb-Edu dataset). For this exercise, the algorithm was kept fixed and follows the GPT-2/3 papers. Because llm.c is a direct implementation of GPT training in C/CUDA, the requirements are minimal - there is no need for conda environments, Python interpreters, pip installs, etc. You spin up a cloud GPU node (e.g. on Lambda), optionally install NVIDIA cuDNN, NCCL/MPI, download the .bin data shards, compile and run, and you're stepping in minutes. You then wait 24 hours and enjoy samples about English-speaking Unicorns in the Andes. For me, this is a very nice checkpoint to get to because the entire llm.c project started with me thinking about reproducing GPT-2 for an educational video, getting stuck with some PyTorch things, then rage quitting to just write the whole thing from scratch in C/CUDA. That set me on a longer journey than I anticipated, but it was quite fun, I learned more CUDA, I made friends along the way, and llm.c is really nice now. It's ~5,000 lines of code, it compiles and steps very fast so there is very little waiting around, it has constant memory footprint, it trains in mixed precision, distributed across multi-node with NNCL, it is bitwise deterministic, and hovers around ~50% MFU. So it's quite cute. llm.c couldn't have gotten here without a great group of devs who assembled from the internet, and helped get things to this point, especially ademeure, ngc92, @gordic_aleksa, and rosslwheeler. And thank you to @LambdaAPI for the GPU cycles support. There's still a lot of work left to do. I'm still not 100% happy with the current runs - the evals should be better, the training should be more stable especially at larger model sizes for longer runs. There's a lot of interesting new directions too: fp8 (imminent!), inference, finetuning, multimodal (VQVAE etc.), more modern architectures (Llama/Gemma). The goal of llm.c remains to have a simple, minimal, clean training stack for a full-featured LLM agent, in direct C/CUDA, and companion educational materials to bring many people up to speed in this awesome field. Eye candy: my much longer 400B token GPT-2 run (up from 33B tokens), which went great until 330B (reaching 61% HellaSwag, way above GPT-2 and GPT-3 of this size) and then exploded shortly after this plot, which I am looking into now :)
Andrej Karpathy tweet media
English
122
739
6.2K
724.9K
SFMB retweetledi
MIT CSAIL
MIT CSAIL@MIT_CSAIL·
This repository on ML operations has free talks, books, papers and more: bit.ly/MLops Developed by @visenger
MIT CSAIL tweet media
English
8
197
859
68.5K
SFMB retweetledi
Andrej Karpathy
Andrej Karpathy@karpathy·
# on shortification of "learning" There are a lot of videos on YouTube/TikTok etc. that give the appearance of education, but if you look closely they are really just entertainment. This is very convenient for everyone involved : the people watching enjoy thinking they are learning (but actually they are just having fun). The people creating this content also enjoy it because fun has a much larger audience, fame and revenue. But as far as learning goes, this is a trap. This content is an epsilon away from watching the Bachelorette. It's like snacking on those "Garden Veggie Straws", which feel like you're eating healthy vegetables until you look at the ingredients. Learning is not supposed to be fun. It doesn't have to be actively not fun either, but the primary feeling should be that of effort. It should look a lot less like that "10 minute full body" workout from your local digital media creator and a lot more like a serious session at the gym. You want the mental equivalent of sweating. It's not that the quickie doesn't do anything, it's just that it is wildly suboptimal if you actually care to learn. I find it helpful to explicitly declare your intent up front as a sharp, binary variable in your mind. If you are consuming content: are you trying to be entertained or are you trying to learn? And if you are creating content: are you trying to entertain or are you trying to teach? You'll go down a different path in each case. Attempts to seek the stuff in between actually clamp to zero. So for those who actually want to learn. Unless you are trying to learn something narrow and specific, close those tabs with quick blog posts. Close those tabs of "Learn XYZ in 10 minutes". Consider the opportunity cost of snacking and seek the meal - the textbooks, docs, papers, manuals, longform. Allocate a 4 hour window. Don't just read, take notes, re-read, re-phrase, process, manipulate, learn. And for those actually trying to educate, please consider writing/recording longform, designed for someone to get "sweaty", especially in today's era of quantity over quality. Give someone a real workout. This is what I aspire to in my own educational work too. My audience will decrease. The ones that remain might not even like it. But at least we'll learn something.
English
658
3.6K
18.5K
2.2M
SFMB retweetledi
Ray Dalio
Ray Dalio@RayDalio·
The organization, like the individual, has to push through to results in order to succeed—this is step five in the 5-Step Process. While recently cleaning up a huge pile of work products from the 1980s and 1990s, I came across boxes and boxes full of research. There were thousands of pages, most covered with my scribbles, and I realized that they represented just a fraction of the effort I'd put in. At our fortieth-year celebration I was given copies of the almost ten thousand Bridgewater Daily Observations that we'd published. Every one of them expressed our deepest thinking and research about markets and economies. I also stumbled across the manuscript of an eight-hundred-page book that I wrote but then got too busy to publish, and countless other memos and letters to clients, research reports, and versions of the book you're reading now. Why did I do all these things? Why do others work so hard to achieve their goals? From what I can see, we do it for different reasons. For me, the main reason is that I can visualize the results of pushing through so intensely that I experience the thrill of success even while I'm still struggling to achieve it. Similarly, I can visualize the tragic results of not pushing through. I am also motivated by a sense of responsibility; I have a hard time letting people I care about down. But that's just what's true for me. Others describe their motivation as attachment to the community and its mission. Some do it for approval and some do it for financial rewards. All these are perfectly acceptable motivations and should be used and harmonized in a way consistent with the culture. The way one brings people together to do this is key. This is what most people call "leadership." What are the most important things that a leader needs to do in order to get their organizations to push through to results? Most importantly, they must recruit individuals who are willing to do the work that success requires. While there might be more glamour in coming up with the brilliant new ideas, most of success comes from doing the mundane and often distasteful stuff, like identifying and dealing with problems and pushing hard over a long time. This was certainly the case with the Client Service Department. Through a lot of relentless hard work in the years since the original problem turned up, the department has become an example to other teams at Bridgewater—and our client satisfaction levels remain consistently high. The great irony of all this is that none of our clients ever even noticed the problems we saw with the memos. Sending out work not up to our standards was bad—and I'm glad it was corrected. But it could've been much worse, tarnishing our reputation for delivering pervasive excellence. Once that happens, it becomes much harder to restore trust. #principleoftheday
Ray Dalio tweet media
English
27
77
500
83.7K
SFMB retweetledi
freeCodeCamp.org
freeCodeCamp.org@freeCodeCamp·
Writing code used to be a distinctly human endeavor. But with AI and code generation, the process of development is changing rapidly. In this tutorial @SonyaMoisset goes over how to use AI-generated code safely while maintaining your skills as a dev. freecodecamp.org/news/how-to-us…
English
1
26
147
37.2K