Tanishq Mathew Abraham, Ph.D.
21.1K posts

Tanishq Mathew Abraham, Ph.D.
@iScienceLuvr
CEO @SophontAI | Founder @MedARC_AI | PhD at 19 (2023) | ex Research Director Stability AI | Biomed. engineer @ 14 | TEDx talk➡https://t.co/xPxwKTq6Qb




if a space odyssey was made in 2026 with AI


finding out hank green, the guy responsible for teaching me so much science and history and research whose whole thing was using credible sources, is now using chatgpt for "research" is so deeply saddening

Current map of medical AI benchmarks: 🧠 CLINICAL JUDGMENT • SCT-Bench tests whether models revise clinical judgments appropriately as uncertain evidence changes. • CPC-Bench tests complex diagnosis, next-test selection and medical-evidence retrieval. • Sequential Diagnosis tests whether models ask the right questions and order the right tests before reaching a diagnosis. These capabilities are increasingly consolidated through MAST, which currently includes CPC-Bench and several benchmarks below. @drethangoh @vishnuravi @pranavrajpurkar @StanfordMed @harvardmed 🛡️ SAFETY + COMMUNICATION • First Do NOHARM v2 tests whether models avoid harmful recommendations and critical omissions in clinical management. • The HealthBench family, including HealthBench Professional which is where the graph is from, tests realistic patient and clinician conversations using case-specific physician rubrics. @OpenAI @thekaransinghal • MedRiskEval, including PatientSafetyBench, tests patient-facing risks such as unsafe advice, misinformation, overconfidence and bias. @MSFTResearch 👁️ MULTIMODAL REASONING • ReXrank Mini tests interpretation and report generation from radiology images. @pranavrajpurkar • IMCBench tests safe, accurate reasoning in image-grounded, multi-turn medical conversations. @AmazonScience • AgentClinic tests clinical agents that must interact with patients, collect multimodal information and use tools. @npjDigitalMed 🧰 EHR + AGENTIC CARE • MedAgentBench v2 tests whether agents can retrieve information and complete actions inside a FHIR-based EHR. @AndrewYNg @james_y_zou @NEJM_AI • PhysicianBench tests long-horizon physician tasks grounded in longitudinal patient records. • LongMedBench tests memory and temporal reasoning across repeated visits and evolving treatment. • EHR-Complex tests executable reasoning over real clinical databases. 🏥 BROAD CLINICAL WORKFLOWS • MedHELM tests clinical decisions, documentation, patient communication, research and administration. @StanfordHAI @StanfordMed @NatureMedicine • MedMarks provides an open benchmark suite spanning medical reasoning, calculations, information extraction and EHR tasks. I’m sure I missed some, lmk what to add!


yes, nonsofic groups exist: this statement is one of many new beautiful results proved by Astra, our next major model. We're releasing 10 such Astra proofs, complete with lean certificates and CoT walkthroughs for each of them. The results are wide-ranging, from von Neumann algebras (disproof of Connes' Rigidity Conjecture) to better bounds for high dimensional sphere packing, for circuit complexity, for monochromatic triangles in multicolored graphs, and more. More thoughts here: openai.com/index/ten-adva…


yes, nonsofic groups exist: this statement is one of many new beautiful results proved by Astra, our next major model. We're releasing 10 such Astra proofs, complete with lean certificates and CoT walkthroughs for each of them. The results are wide-ranging, from von Neumann algebras (disproof of Connes' Rigidity Conjecture) to better bounds for high dimensional sphere packing, for circuit complexity, for monochromatic triangles in multicolored graphs, and more. More thoughts here: openai.com/index/ten-adva…

Oh my God Jensen Huang is now on Twitter!!!! And what an incredible first post too!

I’m surprised nobody has put up a microsite backed by a for loop and token spend for all outstanding “named” problems in mathematics.

🚨🚨Breaking news from the math world! Open AI just proved nonsofic groups existed! GPT 5.6 Sol continues to replace mathematicians!!!

Here is the full letter Leopold sent to his LPs last night. Rumors of his demise are greatly exaggerated.







