Ivan Kuleshov

4.1K posts

Ivan Kuleshov banner
Ivan Kuleshov

Ivan Kuleshov

@Merocle

Head of Hardware at JetBrains. Any sufficiently advanced technology is indistinguishable from magic.

Munich, Bavaria Katılım Mart 2011
116 Takip Edilen28.1K Takipçiler
Ivan Kuleshov
Ivan Kuleshov@Merocle·
Where's the Money?
Ivan Kuleshov tweet media
English
2
0
12
1.4K
Ivan Kuleshov
Ivan Kuleshov@Merocle·
The perfect use case for nvidia.com/en-us/products…
Ivan Kuleshov@Merocle

Introducing the Lamark agent, The agent that actually trains the model for you. github.com/merocle/lamark… It's designed with a focus on the “Local First” philosophy. I tested and developed it primarily for the Nvidia DGX Spark to leverage its strengths - specifically, its ability to fine-tune models. The agent learns every night through conversations with you. The focus on tools and the ability to call cloud AI models is implemented as a tool, too. The agent is completely open-source The next step will be to add support for more devices, and of course, the new @NVIDIAAI RTX Spark. It is based on Hermes and supports all of its features. Although personally, I've mostly used Telegram and the CLI I'm Inviting Contributors!

English
0
1
4
2K
Ivan Kuleshov
Ivan Kuleshov@Merocle·
Introducing the Lamark agent, The agent that actually trains the model for you. github.com/merocle/lamark… It's designed with a focus on the “Local First” philosophy. I tested and developed it primarily for the Nvidia DGX Spark to leverage its strengths - specifically, its ability to fine-tune models. The agent learns every night through conversations with you. The focus on tools and the ability to call cloud AI models is implemented as a tool, too. The agent is completely open-source The next step will be to add support for more devices, and of course, the new @NVIDIAAI RTX Spark. It is based on Hermes and supports all of its features. Although personally, I've mostly used Telegram and the CLI I'm Inviting Contributors!
English
2
2
32
4.5K
Ivan Kuleshov
Ivan Kuleshov@Merocle·
When I write “Hello” to Claude
English
2
2
40
4.1K
stvchu
stvchu@stvchu·
@Merocle how did you connect these books?
English
1
0
0
347
Ivan Kuleshov
Ivan Kuleshov@Merocle·
It’s finally up and running! And the result is pretty decent. Now I want to put this 1U platform... 😅
Ivan Kuleshov tweet mediaIvan Kuleshov tweet media
English
19
8
191
86.5K
seslly
seslly@seslly·
@Merocle nice thing is you have a built in battery backup
English
1
0
3
2.7K
Ivan Kuleshov retweetledi
EXO Labs
EXO Labs@exolabs·
We're benchmarking every model, every quant, on every different hardware setup for every price point. All developers, companies, and people will have access to local, open source intelligence. Releasing soon.
EXO Labs tweet media
English
14
26
237
38.5K
Ivan Kuleshov
Ivan Kuleshov@Merocle·
There are some fairly detailed tests here reddit.com/r/LocalLLaMA/c…; overall, I got roughly the same results, but I only tried one model from this list. However, the heat, noise and cursor lags (probably because of high GPU utilization) when running the LLM ruin the whole experience
English
1
0
2
621
Ryan Dale
Ryan Dale@rot13maxi·
@Merocle @exolabs Hows prefill speed? On coding agents that’s what kills me on my m3. Can generate tokens fast enough but hits the wall when its time to read a bunch of files
English
1
0
0
655
Ivan Kuleshov
Ivan Kuleshov@Merocle·
Some real-world test results. 1. 2-node model – simple coding 2. Single-node model – similar task 3. Single-node model – simple chat (Hi, what can you do?) 4. 2-node model – simple chat As for the software @exolabs – it delivered the best results of everything I tested (but the choice for Mac OS isn’t great) A 4-node cluster is next. What’s really cool is that loading the model is significantly faster than on Spark (takes 15–20 seconds, rather than 5–10 minutes). But even with simple chats, the laptop starts making an obscene amount of noise.
Ivan Kuleshov tweet mediaIvan Kuleshov tweet media
English
8
3
128
11.9K
Egor Pochekutov
Egor Pochekutov@EgorPochekutov·
@Merocle Разве не дешевле было бы использовать 4 Mac mini вместо четырех Mac book?
Русский
1
0
0
736
Ivan Kuleshov
Ivan Kuleshov@Merocle·
The "S" in "Local AI" stands for Savings
Ivan Kuleshov tweet media
English
96
86
3.5K
173.2K
Ivan Kuleshov
Ivan Kuleshov@Merocle·
Cluster testing is currently in progress
Ivan Kuleshov tweet mediaIvan Kuleshov tweet media
English
16
14
510
141.7K