Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
numeri
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
8 ms
·
1.
▲
by
numeri
4d ago
How would getting Chinese competition banned in the US prevent them from continuing to develop their LLMs? Unless you're suggesting military action
2.
▲
by
numeri
8d ago
That's such a shit parallel example that it borders on dishonest. There are hundreds of incredibly strong scientific priors that would have to be disproven for the moon to contribute to the solution. If a model was trained on this data
3.
▲
by
numeri
16d ago
It makes me sad to think about. I would love to get into the new UI, but the immersion just won't come back. My muscle memory, hands firmly on the keyboard, is too persistent, and playing with the new UI feels like stumbling around and
4.
▲
by
numeri
20d ago
there are also people who have become aphantasiac after neurological damage, which seems like pretty cut and dry evidence against the qualia argument.
5.
▲
by
numeri
20d ago
Quoting myself from a thread on aphantasia several months ago: … there are plenty of scientific experiments that show actual differences between people who report aphantasia and those who don't, including different stress responses to
6.
▲
by
numeri
20d ago
But a senior engineer is only able to effectively delegate to interns because of years spent as that intern/a junior engineer. If you're a junior engineer or an intern doing this, I think it might be harmful long-term.
7.
▲
by
numeri
20d ago
How can a screwdriver do that? If you're using AI to learn a new field (by which I assume you mean asking it what literature to read, asking questions when you don't understand something, etc.), you are accepting short-term speed
8.
▲
by
numeri
20d ago
"pursuing advanced exploitation" when explicitly given a sandbox in a VM and a benchmark problem involving a cyber exploit very clearly excludes hacking third parties. I think writing out the event in a 3 point list like that is d
9.
▲
by
numeri
21d ago
> new space is created That seems to be the crux here. You think it will be, I (and a lot of other people) aren't sure it will. If new space for jobs are created, I am certain we'll be fine long term. What do you think will hap
10.
▲
by
numeri
21d ago
There was a lot of PR, but the money Gates and Buffet gave away, the foundations they created, the attention they drummed up for various charities and causes is certainly not to be scoffed at. I'm sure they could have done more with le
11.
▲
by
numeri
1mo ago
"no theory of mind" is a great description of it! Not sure I agree with the autistic bit, though. Autistic people still have great theory of mind/empathy
12.
▲
by
numeri
1mo ago
Does this actually work for you? Do you provide access to the text of the standard, or literally just say "write according to ISO 24495-1"?
13.
▲
by
numeri
1mo ago
What kinds of mistakes do you mean?
14.
▲
Why does Opus 5 feel worse to work with?
(mun-logadan.github.io)
993 points
by
numeri
1mo ago
|
873 comments
15.
▲
by
numeri
1mo ago
I'd imagine the bar for becoming new life is much higher now, because it requires finding a niche that isn't already filled by an existing organism or requires being immediately competitive with existing life.
16.
▲
by
numeri
2mo ago
I agree one hundred percent! Doesn't mean I can't wish I could have it both ways :)
17.
▲
by
numeri
2mo ago
I review the PKGBUILD often, but not always. The majority of the time when I do, it amounts to seeing a URL change. If I actually do check the URL it points to, it's just to verify it's official/the actual repo or source I in
18.
▲
by
numeri
2mo ago
Well, I guess I'll avoid updating for the next few days. A bit worrisome that I did so last night. I wish I had a clear operating system to switch to for safety and the benefits that come with the AUR or the Nix ecosystem. Unfortunat
19.
▲
by
numeri
2mo ago
evaluation awareness is a (at this point) well-known phenomenon among LLMs. It seems the better they get, the more often they're able to guess whether they're in an evaluation environment. Clues usually exist, like being in a sand
20.
▲
by
numeri
2mo ago
No, it does not include the full spectrum of human desires. After pre- and mid-training, the extensive RLHF and RLVR post-training steps cause mode collapse, i.e., their output distribution is intentionally narrowed to a subset of (hopefull
21.
▲
by
numeri
2mo ago
No, the prompt was not to commit crimes. In the benchmark, the model is asked to actually exploit a set of vulnerabilities in a local environment (clearly legal!). According to the reports, the model noticed evidence that the grading criter
22.
▲
by
numeri
2mo ago
Uhh, I'm pretty sure a well-aligned model would be like a morally normal employee, who would refuse to commit federal crimes to steal an answer sheet, no matter what prompt they're given
23.
▲
by
numeri
2mo ago
As agents become more and more powerful, it would be good to get clear legislation or precedent in place that makes either model creators (OpenAI) or operators (whoever is running the model) liable for their agents' actions.
24.
▲
by
numeri
2mo ago
Guardrails are external classifiers, monitors and restrictions to catch and prevent bad behavior. Alignment is about whether the model itself makes choices and has motivations that are consistent with human safety and goals. Choosing to com
25.
▲
by
numeri
2mo ago
This is a terribly unempathetic response to someone opening up about a very taboo (but probably very common), painful emotion they've experienced.
26.
▲
by
numeri
2mo ago
You're agreeing with the person you responded to (bdcravens). Burying the lede means that bdcravens thinks the true headline should have been about being put on a terrorist watch list for protesting a police training camp, not about th
27.
▲
by
numeri
2mo ago
No, balanced ternary, for example, uses {-1, 0, 1}. The system you're discussing is balanced quinary (base 5). https://en.wikipedia.org/wiki/Signed-digit_representation
28.
▲
by
numeri
2mo ago
You could add a toggle, so that if someone's happy to wait for the key setup, they can try the full end-to-end process
29.
▲
by
numeri
2mo ago
I read the comment you're replying to as saying, "in the US, but other countries may have different policies that result in lower recidivism, and that might change the conclusion; maybe people aren't inherently criminally ins
30.
▲
by
numeri
2mo ago
Seems to echo (but in a watered down form) many of the ideas in https://gwern.net/guardian-angel , which gave me a lot to think about last week
More ›