Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
novaleaf
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
4 ms
·
1.
▲
by
novaleaf
1mo ago
> "Adversarial red-teaming made auto mode stronger" Has to be said: they (and OpenAi) don't let you red-team your own creations, pisses me off to no end. </rant>
2.
▲
by
novaleaf
1mo ago
opus 5 doesn't follow instructions that great. probably needed that extra "creativity" to benchmax. if you are stuck on anthropic, try Opus 4.8 or Fable5 a try with the same prompts. very different results.
3.
▲
by
novaleaf
1mo ago
that's pretty nice illustration work, I'd agree, keep doing it :D
4.
▲
by
novaleaf
2mo ago
Frontier models can hack you, we should have access to tools assisting defense. I ranted about this in a prior thread [1] Claude doesn't have a "Security whitelist" for small biz. Codex does, but they never replied to my app
5.
▲
by
novaleaf
2mo ago
as others mentioned, it sounds like "Off to Be the Wizard". Unfortunately I wouldn't call it a great series. It starts off strong but drags. IIRC I got through half of the second book before stopping.
6.
▲
by
novaleaf
2mo ago
With Openai and Anthropic, you can opt-out of training. Kimi you can't.
7.
▲
by
novaleaf
2mo ago
> I guess part of it is also that I don't mind doing 'hand-edits' like for example LLMs love to say "// so and so removed" I just go and remove that manually later rather than being like "don'
8.
▲
by
novaleaf
2mo ago
Agree, out of the two, I can get Codex to design and implement security systems (Fable just refuses to discuss). I've only used K3 a little bit, but I found that being able to discuss attack vectors and their mechanics gives me detail
9.
▲
by
novaleaf
2mo ago
Yeah it's not cheap, it's just the normal $200 subs for each.
10.
▲
by
novaleaf
2mo ago
I applied for individuals. I did their verification steps and answered questions in a few minutes, so that was indeed easy. The problem is that was all that happened. No followup, no access, no denial. When I tried to reapply it tells me
11.
▲
by
novaleaf
2mo ago
keep in mind that Kimi trains on your data, so if you care about that, sandbox what it has access to.
12.
▲
by
novaleaf
2mo ago
I think security is an integral part of any production software, and if you get value from LLM's in software development, it seems likely they can be useful in security too. My point and frustration is that gatekeeping in the name of S
13.
▲
by
novaleaf
2mo ago
I don't think it's nefarious, but the end result leads to a pretty frustrating experience by anybody needing actual security work. (and without the organizational deep pockets to obtain SOC 2 attestation)
14.
▲
by
novaleaf
2mo ago
Some irony: I subscribe to Claude and Codex (20x plans), and now Kimi. Why Kimi? because K3 is the only frontier model I can have a serious conversation with about my product's security. (I did apply for OpenAi's Cyber Pilot but
15.
▲
by
novaleaf
2mo ago
Reminds me of the gain-of-function, COVID lab leak hypothesis. It seems like humanity just can't stay away from Pandora's box.
16.
▲
by
novaleaf
2mo ago
frontier vs "not quite" :D
17.
▲
by
novaleaf
2mo ago
makes me wonder if the starbucks story is fiction too :P
18.
▲
by
novaleaf
2mo ago
no, I didn't get access to sol until a few hours ago. I just have my claude protocol files linked to inside codex. trying Sol this morning for the first time, so I can't really comment on that vs gpt5.5. However you can do what
19.
▲
by
novaleaf
2mo ago
I sub both codex and claude at 20x. I like opus+fable more than gpt5.5 because it seems gpt tries to finish tasks by leaving any ambiguity unresolved. claude seems better at surfacing open questions. This is using the same AGENTS.md promp
20.
▲
by
novaleaf
2mo ago
I think the story was Starbucks -> Seattle's Best Coffee, not McD's --> BK. but it does work I guess.
21.
▲
by
novaleaf
2mo ago
reminds me of SBC's (Seattle's Best Coffee) strategy, which was decidedly not Nash: put a store across the street from every Starbucks.
22.
▲
by
novaleaf
2mo ago
it works by ppl harvesting low ranked submissions and resubmitting them via a (unofficial) voting ring.
23.
▲
by
novaleaf
3mo ago
I used to live in Thailand, and over the last 10 or so years it seems that the kind of DIY content you find on Youtube has had a positive effect. Before then it was difficult to find a building maintenance team that understood the purpose
24.
▲
by
novaleaf
3mo ago
yeah, for example, just send a hash of the domain used. but then maybe people would say anthropic is spying on everyone, instead of targeted spying...
25.
▲
by
novaleaf
3mo ago
obligatory link: https://en.wikipedia.org/wiki/Parallel_construction
26.
▲
by
novaleaf
3mo ago
it was released a few hrs ago as "Fable 5". it's an incremental improvement over Opus 4.8.
27.
▲
by
novaleaf
3mo ago
just yesterday I felt that claude code was being aggressive in it's defense, so I lead my response with "Spicy Take! Here's why I think the bug is happening...." Because of syncopathy it took my "Spicy Take"
28.
▲
by
novaleaf
4mo ago
the xAi revenue comes from renting their infrastructure to their biggest competitor. Anthropic is going to IPO for aprox 1/2 the valuation, is profitable, and can cancel the contract with 90 days notice.
29.
▲
by
novaleaf
4mo ago
I like to joke that if you look at every Feng Shui rule through the lens of "to reduce the risk of assassination" it all really makes sense. Maybe it's not so much of a joke....
30.
▲
by
novaleaf
4mo ago
Max 20x is for individuals only. (could probably have emps get it themselves, and reimburse)
More ›