Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
fabmilo
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
62 ms
·
1.
▲
by
fabmilo
4mo ago
Location: San Francisco, CA, USA Remote: Yes Willing to relocate: No Technologies: Python, Go, TypeScript, PyTorch Linkedin: http://www.linkedin.com/in/fabmilo Hi, I’m Fabrizio Milo, a senior AI/ML engineer, large
2.
▲
by
fabmilo
5mo ago
Writing it thinking. We developed our brain together with our hands. It feels slow but is actually faster for the end goal.
3.
▲
by
fabmilo
6mo ago
I am fascinated by this example of using AI to improve AI. I won a small prize using this technique on helion kernels at a pytorch hackathon in SF. The next step are: - give the agent the whole deep learning literature research and do tree
4.
▲
by
fabmilo
6mo ago
Was thinking the same thing. probably once a day would be more than enough. if you really want a minute by minute probably a delta file from the previous day should be more than enough.
5.
▲
by
fabmilo
7mo ago
indeed. make a loom showing us why is better.
6.
▲
by
fabmilo
7mo ago
There is tons of good advice. This blog post can be easily turned into a skill for agents.
7.
▲
Generative Modeling via Drifting
(arxiv.org)
2 points
by
fabmilo
7mo ago
|
1 comments
8.
▲
by
fabmilo
7mo ago
New generative modeling using a single inference step
9.
▲
by
fabmilo
7mo ago
Very impressive work from Waymo. The driving with a tornado in the horizon example kind of struck my imagination, many people actually panic in such scenarios. I wonder though the compute requirements to run these simulations and producing
10.
▲
by
fabmilo
8mo ago
because of the principle: you only understand what you can create. You think you know something until you have to re-create it from scratch.
11.
▲
by
fabmilo
10mo ago
VAE for real time video generation, WAN 2.1 / Matrix Game 2.0
12.
▲
by
fabmilo
11mo ago
How much would cost to produce these ?
13.
▲
by
fabmilo
11mo ago
nice, didn't knew this tool either
14.
▲
by
fabmilo
1y ago
Yeah I totally agree, we need time to completion of each step and the number of steps, sizes of prompts, number of tools, ... and better visualization of each run and break down based on the difficulty of the task
15.
▲
by
fabmilo
1y ago
How does it work? is just a documentation specification like spec kit?
16.
▲
by
fabmilo
1y ago
I was just reflecting on this blog post after reading it this morning. What do you think on code mode after implementing it? At this point would not be better to just have a sandboxed api environment with customizable api/tools endpoin
17.
▲
by
fabmilo
1y ago
one axis that is missing from the discussion is how fast they are improving. We need ~35 years to get a senior software engineer (from birth to education to experience). These things are not even 3.5 years old. I am very interested in this
18.
▲
by
fabmilo
1y ago
I like zotero, I started vibe coding some integration for my workflow, the project is a bit clunky to build and iterate the development specially with gemini & claude. But I think that is the direction to take instead of reinvent from s
19.
▲
by
fabmilo
1y ago
reference to the library: https://trafilatura.readthedocs.io/en/latest/ for the curious: Trafilatura means "extrusion" in Italian. | This method creates a porous surface that distinguishes pasta trafilat
20.
▲
by
fabmilo
1y ago
more excited about the rust impl than the typescript one.
21.
▲
by
fabmilo
1y ago
The interesting delta here is that this proves that we can distribute the training and get a functioning model. The scaling factor is way bigger than datacenters
22.
▲
by
fabmilo
1y ago
I read the paper and the results don't really convince me that is the case. But the problem still remains of being able to use information from different part of the model without squishing it to a single value with the softmax.
23.
▲
by
fabmilo
1y ago
We have to move past tokenization for the next leap in capabilities. All this work done on tokens, specially in the RL optimization contest, is just local optimization alchemy.
24.
▲
by
fabmilo
2y ago
I will believe reasoning architectures when the model knows how to store parametric information in an external memory out of the training loop.
25.
▲
by
fabmilo
2y ago
was genuinely excited when I read this but the github repo does not have any code.
26.
▲
by
fabmilo
2y ago
I was just about to submit this link and redirected me to this page. I am shocked that it received only four comments. If you are working in the LLMs/Agent space ( you are, right?) and you don't understand the significance of this
27.
▲
by
fabmilo
2y ago
Happy new year to everyone, hacker news is more than my home page. This community is awesome!
28.
▲
by
fabmilo
2y ago
Doesn't make enough drama.
29.
▲
by
fabmilo
2y ago
Thanks to make this open source! Doesn't seem to have any AI enabled search feature does it? Honestly I think you gave up too early, there are solid foundations in the app but definitely needs some more polishing for practical use. I
30.
▲
by
fabmilo
2y ago
I am gonna read this paper and the other latent sentence later today. I always advocated for this kind of solutions together with latent sentence search should get to the next level of AI. Amazing work from Meta
More ›