Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
oleczek
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
11 ms
·
1.
▲
by
oleczek
2mo ago
Fair point - I overstated it. Thanks for the correction.
2.
▲
by
oleczek
2mo ago
Thanks for vouching. Appreciate the heads-up.
3.
▲
by
oleczek
2mo ago
Yeah, original is clearer. Just went with the shorter one so more people would actually click.
4.
▲
by
oleczek
2mo ago
yeah English isn't my first language so i used AI to clean up the writing. the research, the data and the analysis are all mine and all open (in the linked repo)
5.
▲
by
oleczek
2mo ago
Yeah, the optimism/hedging part lines up nicely with that framing. The bit that still puzzles me is that the same edit moved expressed confidence in opposite directions — down for Gemma, up for Qwen. If it were simply removing one shar
6.
▲
by
oleczek
2mo ago
Author here. Quick version: “abliteration” (basically removing the direction in the model that causes it to refuse) is the go-to method people use to make open models uncensored. Most people treat it like a clean surgical cut - it just kill
7.
▲
"Uncensored" open LLMs are measurably more optimistic than their base models
(arxiv.org)
43 points
by
oleczek
2mo ago
|
22 comments