Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
pilooch
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
pilooch
4d ago
That's unless the code produced in the future is much more complex than today's.
2.
▲
Ukraine Delta Software
(nytimes.com)
4 points
by
pilooch
21d ago
|
0 comments
3.
▲
by
pilooch
1mo ago
The risk is they are a magnet for many more talented ML scientists to leave google.
4.
▲
by
pilooch
2mo ago
Daily sports is the way, purges the body, rests the mind.
5.
▲
by
pilooch
3mo ago
Yes, full ft or lora https://github.com/NVIDIA-NeMo/Automodel/blob/main/docs/guid...
6.
▲
by
pilooch
4mo ago
Yes, all my emails gyer sorted out by a finetuned gemma. There are turned into images passes to the model, as multimodal is so practical.
7.
▲
DeepSeek-V4-Flash (official FP8) running across 2x DGX Spark
(forums.developer.nvidia.com)
4 points
by
pilooch
4mo ago
|
1 comments
8.
▲
by
pilooch
4mo ago
Reminded me of the recent, and excellent, Canadian tv series Empathy, with the main character is found in a garbage can by his adoptive parents.
9.
▲
by
pilooch
4mo ago
AlphaEvolve couples map-elites with LLMs. It's an key step in machine learning, in the vein of DQN for reinforcement learning. AE brings diversity from the genetic algorithms community to large scale optmized deep learning and RL model
10.
▲
Anon: Extrapolating Adaptivity Beyond SGD and Adam
(anonymous.4open.science)
2 points
by
pilooch
4mo ago
|
0 comments
11.
▲
The Podcast Where You Can Eavesdrop on the A.I. Elite
(nytimes.com)
4 points
by
pilooch
5mo ago
|
0 comments
12.
▲
by
pilooch
5mo ago
Slides, publications and tech reports, very handy for figures !
13.
▲
by
pilooch
5mo ago
It's useful when using prism, and for exploratory research & code.
14.
▲
by
pilooch
6mo ago
Revolting and so inevitable though I believe: we're sort of running these already in our minds, we'll be outrun here too.!
15.
▲
by
pilooch
7mo ago
Good catch. Cancer treatment scheduling is hard as well as mixes need tombe prepared in advance and cancelles appointments are hard to fill.
16.
▲
by
pilooch
7mo ago
Hello! Not commenting on content or functionality. Scheduling in AI is a very dense field. An a past researcher in AI decision making, I got confused by the 'Scheduling solved' slogan. FYI recent AI for scheduling include GNNs an
17.
▲
by
pilooch
7mo ago
Try Flashback, it's darker but genius as well, maybe more approachable.
18.
▲
by
pilooch
8mo ago
Synthetic data for human (machine) learning... We should spend more time outside, we will!
19.
▲
Photoroom T2i Open Model
(huggingface.co)
2 points
by
pilooch
10mo ago
|
0 comments
20.
▲
by
pilooch
10mo ago
The 'Eclipse' album is a classic.
21.
▲
A Simple Definition of Intelligence
(minimoog.substack.com)
2 points
by
pilooch
11mo ago
|
0 comments
22.
▲
by
pilooch
1y ago
I like the glasses path, well I do wear glasses, but some elements remain unclear to me: - are prescription glasses available for display ? I guess not ? - these glasses need to be online, I guess they do so with a phone and bluetooth conne
23.
▲
by
pilooch
1y ago
Congrats, this solution resembles AlphaEvolve. Text serves as the high-level search space, and genetic mixing (map-elites in AE) merges attemps at lower levels.
24.
▲
by
pilooch
1y ago
Losing the mental map is the number one issue for me. I wonder if there could be a way to keep track of it, even at a high level. Keeping the ability to dig in is crucial.
25.
▲
by
pilooch
1y ago
Hello, very interested in the scrollback! I've used mosh for 10+ years and it still runs my 100+ opened terminals to this day ! Would love to try your alternative
26.
▲
by
pilooch
1y ago
Exactly, for real time applications VTO, simulators,...), i.e. 60+FPS, diffusion can't be used efficiently. The gap is still there afaik. One lead has been to distill DPM into GANs, not sure this works for GANs that are small enough fo
27.
▲
Jet-Nemotron
(github.com)
1 points
by
pilooch
1y ago
|
0 comments
28.
▲
by
pilooch
1y ago
Is this a bit similar to what tensorrt does, but in a more opened manner ?
29.
▲
by
pilooch
1y ago
Because they'd never hire, but subcontract down to the bone.
30.
▲
by
pilooch
1y ago
I'd be interested in what implementation of D3PM was used (and failed). Diffusion model are more data efficient than their AR LLM counterpart but les compute efficient at training time, so it'd be interesting to know whether with
More ›