Modaic

28 posts

Modaic banner
Modaic

Modaic

@modaicdev

Infra for analytical AI

San Francisco, CA Katılım Eylül 2025
12 Takip Edilen130 Takipçiler
Modaic retweetledi
Farouk
Farouk@FaroukAdeleke3·
for now…🤓
Farouk tweet media
English
2
1
7
4.2K
Modaic retweetledi
Tyrin
Tyrin@ty_todd1·
The more I play around with GEPA the more I realize its not just a prompt optimization algorithm. Its really the most efficient way to have an LLM explore a massive dataset and and make useful insights. This visualizer alone shows just how cool that process is.
English
21
108
1.4K
119.9K
Modaic retweetledi
Erika Shorten
Erika Shorten@eshorten300·
Modaic and Weaviate 1. Modaic and Weaviate: Load the `CrossEncoderRanker` program from the Modaic Hub, as well as `PromptToSignature` (github.com/weaviate/recip…)
Erika Shorten tweet media
English
1
4
8
524
Modaic retweetledi
Farouk
Farouk@FaroukAdeleke3·
About a year ago, @plasticlabs achieved SOTA with DSPy on the OpenToM benchmark. The benchmark tests models’ ability to track and reason about the beliefs, perceptions, intentions, and psychological states of simulated characters (social cognition). @vintrotweets experiments are now packaged and available on @modaicdev with the ability to run one of his optimized programs with your own variables and under your own evaluations. Links below.
Farouk tweet media
English
2
3
14
1.7K
Modaic retweetledi
Farouk
Farouk@FaroukAdeleke3·
Most leaderboards use one fixed zero-shot prompt across all models. Problem: Different LMs have different “prompt ceilings.” When you give them all the same prompt, it becomes less about benchmarking models and more about benchmarking how that model performs on a fixed prompt. Solution: Stanford shows that adding structured prompting (especially zero-shot CoT via @DSPyOSS) lifts performance by ~4 points on average and can even flip model rankings on tasks like MMLU-Pro, GSM8K, and MedCalc.
Farouk tweet media
English
1
2
7
588
Modaic retweetledi
Farouk
Farouk@FaroukAdeleke3·
Extremely underrated paper out of Stanford, including one of the creators of MEDVal! Benchmarking the ceiling of LM's as systems (optimized prompting strategy + LM) instead of generalizing the same prompt over competing LM's proves to be a more holistic evaluation of their capabilities. arxiv.org/pdf/2511.20836
English
1
1
4
138
Modaic
Modaic@modaicdev·
game changer.
Tyrin@ty_todd1

I just created IntelliSense for @DSPyOSS. It's a VSCode extension that looks at your Signatures and gives you type hints for modules and Predictions. Download for VSCode #review-details" target="_blank" rel="nofollow noopener">marketplace.visualstudio.com/items?itemName… To download for cursor paste this link in your browser. cursor:extension/modaic.dspy-intellisense

English
0
0
1
115