Sabitlenmiş Tweet
Aleksei
82 posts

Aleksei
@pentesti
Building startups in public. Cybersecurity, AI agents, and product experiments.
Katılım Temmuz 2026
185 Takip Edilen37 Takipçiler

one filter i’m using for saas ideas:
what does this problem already cost the customer?
if the answer is “not much,” it’s probably going to be hard to sell.
#buildinpublic
English

@Shivam25mishra OpenAI, but betting against Google for five years feels dangerous
English

@kapilansh_twt no users. at least users give you something to work with.
English
Aleksei retweetledi

i tested 5 model + agent-harness setups on 3 blind security labs:
ocdeep - opencode + deepseek v4 pro
glmcode - claude code + glm-5.2
ocglm - opencode + glm-5.2
kimicode - kimi cli + kimi k2.7 code
ockimi - opencode + kimi k2.7 code
Every setup received the same prompts and benchmark conditions. No benchmark-specific hints or known solutions were provided.
all got 3/3 exact.
winner: ocdeep (opencode + deepseek v4 pro) - 20m31s, $0.14, 97/100
One of the clearest takeaways: the harness mattered almost as much as the model.
the same model produced meaningfully different results depending on the agent wrapper.

English
Aleksei retweetledi

i got tired of security agents forgetting everything between hunts, so i built a local second brain in obsidian for cybersecurity (bug bounty + pentesting):
1,310 notes
13,649 links
492 security techniques
250 portswigger lab solutions
84 CVEs
48 security writeups
37 PoCs
45 methodologies + playbooks
19 of my own real bug bounty reports
the graph looks cool. the value is knowing what worked, what failed, and where every claim came from.
next test: same agent, same test - brain vs no brain.

English


















