Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
meffmadd
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
9 ms
·
1.
▲
Recurrent Looped Transformer
(yifanzhang-pro.github.io)
2 points
by
meffmadd
3d ago
|
0 comments
2.
▲
by
meffmadd
1mo ago
Hopefully they release the 35B-A3B version alongside it.
3.
▲
by
meffmadd
1mo ago
Of course, but that was not part of my workflow and I found it quite strange that this was simply not possible especially when docker-compose can easily do this
4.
▲
by
meffmadd
1mo ago
I tried Docker Sandboxes but last time I checked you could not configure custom volume mounts, making more complex setups impossible. For work I need two directories for context for the agent to have access to…
5.
▲
by
meffmadd
2mo ago
This doesn’t hold up imo. Your phone analogy fits the description of a harness. That is the thing you interact with. The LLM is more like the processor in that phone. As long as it gets the work done, I will choose the cheapest/fastest
6.
▲
by
meffmadd
4mo ago
That is very true but I was surprised by how clear the “signal” was. Only Gemini really confidently solved all levels. But yeah the goal is now to include harder levels as well!
7.
▲
by
meffmadd
4mo ago
Yes that was the post that inspired me to build this. While I did implement a more comprehensive harness with path finding tools etc. the models themselves have improved significantly.
8.
▲
by
meffmadd
4mo ago
I found LLMs to be surprisingly good at puzzle games like Baba Is You: https://meffmadd.github.io/samplesurium/posts/baba_is_agent/
9.
▲
by
meffmadd
4mo ago
…or basically the plot of twin peaks
10.
▲
Can LLMs Play Baba Is You?
(meffmadd.github.io)
5 points
by
meffmadd
4mo ago
|
0 comments
11.
▲
by
meffmadd
6mo ago
I think this view assumes no human will/should ever read the code. This is considered bad practice because someone else will not understand the code as well whether written by a human or agent. Unless 0% human oversight is needed anymo
12.
▲
by
meffmadd
6mo ago
Genuine question to people more knowledgeable: Why are politicians/technocrats doing this? Also generally speaking e.g. in relation to chat control and so on. Do they think this is what the people actually want because of lobbying or a
13.
▲
by
meffmadd
7mo ago
I am amazed that the IDM is able to produce enough high quality annotations for the downstream FDM to work, even matching the ground truth contractor annotations!
14.
▲
by
meffmadd
7mo ago
As an EU citizen I really hope we can gain some meaningful distance to the US asap. I hope my leaders feel the same. And if everything works out I think this will be great for the EU. This is really some sort of diplomatic Streisand effect.
15.
▲
by
meffmadd
7mo ago
I really wonder what motivates a seven year old to persistently work on that „one thing“ and not get distracted/bored. I guess he knew he was special?
16.
▲
by
meffmadd
7mo ago
Have you ever used an open model for a bit? I am not saying they are not benchmaxxing but they really do work well and are only getting better.
17.
▲
by
meffmadd
7mo ago
Here is what I don’t get tho: you have UX designers/engineers creating a new interface. What do you tell them? Just to do whatever? They probably spent months designing the new interface but why not fix this? They must have seen it is
18.
▲
by
meffmadd
7mo ago
It will be tough to run on our 4x H200 node… I wish they stayed around the 350B range. MLA will reduce KV cache usage but I don’t think the reduction will be significant enough.
19.
▲
Ask HN: What is the most complicated Algorithm you came up with yourself?
3 points
by
meffmadd
7mo ago
|
9 comments
20.
▲
by
meffmadd
9mo ago
From what I have heard „Bodycam“ uses scans of actual locations for its maps.
21.
▲
by
meffmadd
10mo ago
Yeah true, but also this is a bit like saying the lock screen of your phone should not become a "one stop shop" for all push notifications. I actually do not own an Apple TV but I just imaging you have a list of shows from differe
22.
▲
by
meffmadd
10mo ago
Wow that is quite anti-consumer! Surely a monopoly on streaming will help them realize this. /s
23.
▲
How Attention Got So Efficient [GQA/MLA/DSA] [video]
(youtube.com)
2 points
by
meffmadd
10mo ago
|
0 comments
24.
▲
by
meffmadd
1y ago
As an EU citizen hosting LLMs for researchers and staff at the university I work at, this is hits home. Without Chinese models we could not do what we do right now. IMO, in the EU (and anywhere else for that matter), we should be grateful
25.
▲
Show HN: Aqueduct AI Gateway – A Self-Hostable AI Gateway Without the "Taxes"
(github.com)
1 points
by
meffmadd
1y ago
|
0 comments