Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
jfaganel99
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
Show HN: Katra, self-hosted cognitive memory for AI agents (MCP)
(github.com)
4 points
by
jfaganel99
3mo ago
|
0 comments
2.
▲
by
jfaganel99
3mo ago
For the sceptics... The benchmark is research based with a published ArXiv paper on the methodology https://arxiv.org/abs/2604.13764
3.
▲
by
jfaganel99
3mo ago
Question, because I can't answer it myself... Created an open-source benchmark for code security scanners and ran a bunch of them along with LLMs on real vulnerable code. Fable 5 is on there also as of yesterday, but that's the ga
4.
▲
Show HN: We're inviting Anthropic to put the real Mythos 5 on our open benchmark
(realvuln.com)
4 points
by
jfaganel99
3mo ago
|
4 comments
5.
▲
by
jfaganel99
6mo ago
As promised here is the open-source GitRepo so you can give it a go with your tooling: https://github.com/kolega-ai/Real-Vuln-Benchmark Updated benchmark results published here also. BTW, with v002 we are consistently
6.
▲
Show HN: ClawFinder, an open-source discovery and negotiation layer for agents
(github.com)
5 points
by
jfaganel99
6mo ago
|
0 comments
7.
▲
by
jfaganel99
6mo ago
Working on a model benchmark focused on which model is good for these tasks. Keep you posted
8.
▲
by
jfaganel99
6mo ago
This is one of the most practical breakdowns I’ve seen for a while. The spec.md as a living architecture map is smart, and documenting auth guard pattern sites as new modules get added is exactly the kind of thing that prevents issues creep
9.
▲
by
jfaganel99
6mo ago
Author here. The finding that surprised me most while writing this wasn’t the breach numbers. It was the Stanford result: developers with AI assistance introduced more flaws than those without, and felt more confident about their code. The
10.
▲
Vibe Coding Is a Security Disaster That Is About to Happen
(medium.com)
9 points
by
jfaganel99
6mo ago
|
7 comments
11.
▲
by
jfaganel99
7mo ago
That's a great question. This is how I would think about it: The number of vulnerabilities by itself doesn't mean much. It has more to do with the size of the codebase and the attack surface than with the quality of the code. Ther
12.
▲
by
jfaganel99
7mo ago
Hi HN - side project. After reading about the ClawHavoc campaign and seeing how fast malicious skills were spreading on ClawHub (1,100+ at last count), I figured it would be useful to have something where people can actually practice tel
13.
▲
Show HN: Skill or Kill – Can you spot the malicious AI agent skill?
(skillorkill.dev)
3 points
by
jfaganel99
7mo ago
|
1 comments
14.
▲
by
jfaganel99
7mo ago
[flagged]
15.
▲
by
jfaganel99
7mo ago
Author here. We built a security scanner called Kolega that does semantic analysis instead of pattern matching. To see if it actually worked, we ran it against 45 open source projects and reported what it found through responsible disclosur
16.
▲
Vulnerabilities in 45 Open Source Projects (vLLM, Langfuse, Phase, NocoDB)
(kolega.dev)
2 points
by
jfaganel99
7mo ago
|
3 comments
17.
▲
by
jfaganel99
7mo ago
How do we apply this to geospatial face and licence plate blurs?
18.
▲
by
jfaganel99
7mo ago
Yeah, way more than the good old Notepad :)
19.
▲
by
jfaganel99
7mo ago
Notepad had one job... Seems like bringing markdown features killed it :)