Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
mistymountains
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
mistymountains
2y ago
Again, the problem is custom kernels in CUDA. It’s not straightforward for many applications (LLMs are probably the most straightforward).
2.
▲
by
mistymountains
2y ago
It doesn’t sound like having a house will magically make you feel better. Plenty of people are just as exposed to markets as you yet respond differently. I suggest exercise, nutrition, nature, and extended travel if those are not already a
3.
▲
by
mistymountains
2y ago
These kinds of comments make me think few people have actually tried. My experience has been 1 work day of getting things set up to work the same as before for training and testing (PyTorch).
4.
▲
by
mistymountains
2y ago
Unless you develop in CUDA, you can easily train code (e.g. PyTorch) written for training on Nvidia hardware on AMD hardware. You can even keep the .cuda() calls.
5.
▲
by
mistymountains
2y ago
I’m a AI Scientist and train a lot of models. Personally I think AMD is undervalued relative to Nvidia. No, chips aren’t as fast as Nvidia’s latest and yes, there are some hoops to get things working. But for most workloads in most industri
6.
▲
by
mistymountains
2y ago
Why would you want (versus need) a Tesla? It’s no longer an aspirational product, it’s an appliance.
7.
▲
by
mistymountains
3y ago
For all you know he had a hard adjustment to the college workload. Maybe his high school was not serious and nobody really challenged him. You all could have interrogated why he may have struggled and shared your strategies for success, lif
8.
▲
by
mistymountains
3y ago
My (biotech, mostly remote) company does this. They may not love it but they realize they have to hire from SF, Boston, NYC etc to get the best ML talent and people expect market salaries / don’t want to up and move if they don’t have
9.
▲
by
mistymountains
3y ago
The issue is that actually, outside of technical paths, businesses do like to hire business grads as it allows them to do even less training and they usually don’t care if someone is well read or opinionated (possibly prefer the opposite).
10.
▲
by
mistymountains
3y ago
Or, just maybe, AGI is a mirage with the bulk of its current utility as a marketing tool for much more realistic, if ultimately mundane, applications. OpenAI, of course, knows this.
11.
▲
by
mistymountains
3y ago
You won’t teach yourself by running papers through Claude, and you won’t need to if you went from first principles rather than rushing.
12.
▲
by
mistymountains
3y ago
You gave your 9 year old a smartphone?
13.
▲
by
mistymountains
3y ago
It’s annoying how much things have shifted now that you can’t really own a performance car without worrying if someone will mess with it.
14.
▲
by
mistymountains
3y ago
Yeah we all need to do better and trust China more! Lol.
15.
▲
by
mistymountains
3y ago
Is this your diagnosis? That things are more corporate in terms of outcome? Genuinely curious, as a Bay Area native, why UT/Texas can’t begin to compete given how many issues face SF currently.
16.
▲
by
mistymountains
3y ago
Cool it with the italics.
17.
▲
by
mistymountains
3y ago
Comment datasets are valuable for conversational AI, it’s the same reason Reddit locked down the API I imagine.
18.
▲
by
mistymountains
3y ago
It’s probably the best from their training standpoint. The city is quite small relative to others and their only hope is to essentially memorize San Francisco in the neural networks. IMO you could not take a cruise and drop it into any othe
19.
▲
by
mistymountains
3y ago
They’re too deep, their networks are overfit to San Francisco at this point. Making it work in other cities would require the insane training hours to basically memorize that city.
20.
▲
by
mistymountains
3y ago
I’ve seen this stated a ton and it’s not really true. Once trained, the model (except for decoding) is deterministic, and you can enforce determinism fairly easily. ChatGPT is not deterministic at the chat window but that’s not inherent to
21.
▲
by
mistymountains
3y ago
Welcome to the real world. Move somewhere more affordable.
22.
▲
by
mistymountains
3y ago
Grow up.
23.
▲
by
mistymountains
3y ago
I’d be very careful about this.
24.
▲
by
mistymountains
3y ago
May I ask, are you a VC? This is a complete non sequitur.
25.
▲
by
mistymountains
3y ago
They don’t. They simply assume the model’s most likely output is meaningfully correlated with true rankings even though it was never trained on this task and certainly has not been trained to output the most likely prompt given a prompt in
26.
▲
by
mistymountains
3y ago
Agreed. This is a pretty terrible idea.
27.
▲
by
mistymountains
3y ago
It’s no longer useful / never was that useful. It’s main function was a shiny toy for investors.
28.
▲
by
mistymountains
3y ago
No, you need to show the model is better on this narrow task, not just assert it is because it’s a great general LLM. It’s quite possible you’re correct but just saying GPT is the best, prove me wrong, reeks of a VC or AI bandwagoner.
29.
▲
by
mistymountains
3y ago
The training classes in this case are words in the vocabulary in the context of a sentence.
30.
▲
by
mistymountains
3y ago
It’s not similar other than that attention relates tokens.
More ›