Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
solarwindy
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
solarwindy
2mo ago
Heh, one vestigial bit of code, and they all are. Mind you, it's quite a creaky codebase, so it's forgivable to keep finding these appendices and calling them out as such. Useful, even.
2.
▲
by
solarwindy
8mo ago
The possibility to continue to sound like yourself after permanently losing your voice (e.g. from motor neurone syndrome) is one. Perhaps almost the only one.
3.
▲
by
solarwindy
10mo ago
Quoting Troy from a thread beneath the article: > The easiest approach in that case is to take out the subscription, then immediately cancel it. It'll still last the full month, more here: https://support.haveibeenpwned.c
4.
▲
by
solarwindy
11mo ago
Uh, no, most crypto wallet addresses have either a checksum or some other means of typo detection / prevention.
5.
▲
by
solarwindy
11mo ago
100 billion a quarter is Alphabet, right? Given how much click fraud there is, and that every org and business under the sun is held to ransom to feature on the SERP for their own name even — it’s tempting to say Google’s become a private
6.
▲
How to Fix Any Bug
(overreacted.io)
1 points
by
solarwindy
11mo ago
|
0 comments
7.
▲
by
solarwindy
11mo ago
Outstanding video, thank you. No wonder this took months’ worth of research and animation to make.
8.
▲
by
solarwindy
11mo ago
Well, yes. Rather than that being a takedown, isn’t this just a part of maturing collectively in our use of this technology? Learning what it is and is not good at, and adapting as such. Seems perfectly reasonable to reinforce that legal an
9.
▲
by
solarwindy
11mo ago
FWIW, Claude Sonnet 4.5 and ChatGPT 5 Instant both search the web when asked about this case, and both tell the cautionary tale. Of course, that does not contradict a finding that the base models believe the case to be real (I can’t curre
10.
▲
by
solarwindy
11mo ago
I’m finding that whether this process works well is a measure (and a function) of how well-factored and disciplined a codebase is in the first place. Funnily enough, LLMs do seem to have a better time extending systems that are well-enginee
11.
▲
by
solarwindy
1y ago
How about a narwhal spacewalking from the ISS, with Earth visible below (specifically the Niger delta)? https://claude.ai/public/artifacts/f3860a8a-2c7d-404f-978b-e... Requesting an ‘extravagantly detailed’ versio
12.
▲
The subjective experience of coding in different programming languages (2023)
(interconnected.org)
47 points
by
solarwindy
1y ago
|
56 comments
13.
▲
by
solarwindy
1y ago
VLLMs are incredibly good at decoding math from screenshots, if you’re working from a PDF textbook. ChatGPT especially, and since it’s conversant in LaTeX, it can respond directly in the notation you don’t recognize to break it down for you
14.
▲
A model of Boy's surface in constructive solid geometry
(math.univ-toulouse.fr)
2 points
by
solarwindy
1y ago
|
0 comments
15.
▲
by
solarwindy
1y ago
Isn’t GP’s point, that it’s already enough for those two to have solved it? Not every country with a civil nuclear program needs its own waste containment, it’s just such a small absolute quantity.
16.
▲
by
solarwindy
1y ago
Also Antartica → Antarctica
17.
▲
by
solarwindy
1y ago
Relative to humans, these models sure have ungodly amounts of knowledge, but they also kinda have a lobotomy, in never having moved through the world. It’s remarkable they work as well as they do trained chiefly on text, but being so unteth
18.
▲
by
solarwindy
1y ago
What is a visualisation? Our rod and cone cells could just as well be wired up in any other configuration you care to imagine. And yet, an organisation or mapping that preserves spatial relationships has been strongly preferred over bil
19.
▲
by
solarwindy
1y ago
I think the version hosted on the book's website would work fine on smaller screens (and also seems to have been updated more recently): https://www.sscardapane.it/assets/alice/Alice_book_volume_1....
20.
▲
by
solarwindy
1y ago
I remembered once, in Japan, having been to see the Gold Pavilion Temple in Kyoto and being mildly surprised at quite how well it had weathered the passage of time since it was first built in the fourteenth century. I was told it hadn’t wea
21.
▲
by
solarwindy
1y ago
When framed like this, it's quite unsurprising that LLMs struggle to emulate reasoning through programming problems: there's just not that much signal out there. We tend to commit what already works, without showing much (if any)
22.
▲
by
solarwindy
1y ago
Because 'solar' seems to many like a generic descriptor, not specifically related to our star, Sol.
23.
▲
by
solarwindy
1y ago
The relevant research field is known as mechanistic interpretability. See: https://arxiv.org/abs/2404.14082 https://www.anthropic.com/research/mapping-mind-language-mod...
24.
▲
by
solarwindy
1y ago
https://archive.ph/4MEnQ
25.
▲
by
solarwindy
1y ago
For anyone wondering: https://www.mun.ca/biology/scarr/MGA2_02-07.html
26.
▲
by
solarwindy
1y ago
Not a bad idea. For an effective ruse, there ought to be real company formation records, website, job listings, press mentions, and so on. Stepping back for a second though, doesn’t this all underline the safety researchers’ fears that we
27.
▲
by
solarwindy
1y ago
It’s role play until it’s not. The authors acknowledge the difficulty of assessing whether the model believes it’s under evaluation or in a real deployment—and yes, belief is an anthropomorphising shorthand here. What else to call it, tho
28.
▲
by
solarwindy
1y ago
Is that necessarily a blocker? As others in this thread have pointed out, this probably becomes possible only once sufficient compute is available for some form of non-public retraining, at the individual user level. In that case (and hand-
29.
▲
by
solarwindy
1y ago
> Russia is the most pro-nuclear country in the world Not France? > France derives about 70% of its electricity from nuclear energy, due to a long-standing policy based on energy security. — https://world-nuclear.org/i
30.
▲
by
solarwindy
1y ago
> I think that's... up for debate Been trying to inform myself on how these models work, and it’s pretty interesting, I have to say. Came across this paper from Anthropic, Scaling Monosemanticity [0], where they’re extracting feat
More ›