Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
nharziro
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
26 ms
·
1.
▲
by
nharziro
4d ago
We're probably not as special as you think.
2.
▲
by
nharziro
6d ago
lol, me too. I still have the paper i wrote from 20+ years ago on this subject.
3.
▲
by
nharziro
6d ago
they are doing this with auto-recharge of credits as well causing people to rack up thousands of dollars unintentionally. https://github.com/openai/codex/issues/31987
4.
▲
by
nharziro
7d ago
This is like the mile record. Everytime someone breaks the existing one a bunch of people surpass the old record shortly after.
5.
▲
by
nharziro
9d ago
Well that's news to me because my model 3 in autopilot definitely does not stop for traffic lights or stop signs.
6.
▲
by
nharziro
11d ago
What exactly do you mean when you say llm generated code? Are people prompting llms for changes and features without reviewing the code or iterating on it and then comparing that to what human writes? Because if so it's not surprising
7.
▲
by
nharziro
12d ago
I don't understand what you mean by novel intelligence or what people actually expect from these kinds of "Ai" but what novel intelligence can humans claim? Everything we know or learn is based on what someone else figured ou
8.
▲
by
nharziro
20d ago
I am interested, so please link it. I agree that our judgment is important, but I think these skills are themselves a form of judgment. The agent uses them to make decisions. It will never be perfect, but it is already pretty good. My issue
9.
▲
by
nharziro
21d ago
That’s determined automatically by the agent and the harness. Certain skills trigger on their own depending on the context and what the agent is doing. For example, I have a '$rest-api-design' skill that acts as a guardrail whenev
10.
▲
Encoding Myself into the System
4 points
by
nharziro
21d ago
|
4 comments
11.
▲
by
nharziro
26d ago
why do people find comments like these necessary? What did you gain by making this comment?
12.
▲
by
nharziro
26d ago
I never said it was easy but it's clear that you're not coping well with new world order. Unfortunately for a lot people, it's here to stay do get used to it.
13.
▲
by
nharziro
26d ago
When I read stuff like this people I feel like people haven't yet accepted reality. I think knowing how to write code made a lot of people feel very special. Like the could do something magical, and now feel like that's being take
14.
▲
by
nharziro
27d ago
This comment aged poorly
15.
▲
by
nharziro
1mo ago
I built a benchmark myself to track this and confirmed the same thing https://gist.github.com/nharziro/aed0c364ce2f295a493494c6f1b... Very similar performance to 4.6 and codex 5.3 but slow and token inefficient. Still
16.
▲
by
nharziro
1mo ago
I was genuinely surprised because it's quite a leap from where 3.6 was an as far as I understand this isn't a new model, it's the same model that's been post trained, so I don't quite understand what they did to imp
17.
▲
by
nharziro
1mo ago
I do agree that Qwen 3.8 27B is excellent but slow and very token inefficient. My benchmark places it near opus 4.6 and codex 5.3 performance. 3.6 27B couldn't even complete the benchmark. Please see below for details: https:/&#x
18.
▲
by
nharziro
2mo ago
he's asking for the most powerful models to be in the hands of the few while the majority will be at their mercy.
19.
▲
by
nharziro
2mo ago
How does huggingface fit into all of this if this is marketing? Their security was faked? What are you suggesting??
20.
▲
by
nharziro
2mo ago
Uhh north Korea?
21.
▲
by
nharziro
2mo ago
where is it? Still not accessible...