Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
arw0n
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
by
arw0n
7d ago
We don't know what the prompts were though. Terence Tao showed of some of his, and they are indeed huge, with lots of context, things that might work, things he knows don't etc. What we do know about this, is that 10k agents worke
2.
▲
by
arw0n
7d ago
Agreed, there's a couple of bad behavioral shifts I noted for me when doing agentic coding: - Doing too much at the same time, it is super exhausting, and I get nervous and stressed - Sycophancy makes me overestimate myself - When my s
3.
▲
by
arw0n
8d ago
Interesting way to phrase that. The easy way to solve rape is to castrate all men. We can save some sperm in case of wanted pregnancy, but since the vast majority of violent and sexual abusive crime is committed by men, it would be the most
4.
▲
by
arw0n
9d ago
I combine it with OpenCode + a small OpenRouter budget. GlM 5.3 Flash, Luna, Gemini 3.7 Flash and a couple of other very cheap models are sufficient for a lot of tasks if put on the right road. Telling Astra to debug an issue should be a la
5.
▲
by
arw0n
16d ago
The comment reads like flesh-generated language to me, but I might be fooled. The breadths of pro-AI sentiment is possibly astroturphed, but there's also an argument for an enthusiastic minority of people to write a lot more comments.
6.
▲
by
arw0n
17d ago
I was personally flabbergasted by this one: > When we offload decision-making, we become unaware of the trade-offs. My biggest issue with the new agentic workflow is that I'm constantly asked to make decisions, and it is not always
7.
▲
by
arw0n
17d ago
> I wonder what the 2026 equivalent of going to California is now, if such a thing even exists or can exist. I think this becomes increasingly difficult, because the timelines become so short with almost-instant information desimination,
8.
▲
by
arw0n
18d ago
Strongly disagree on Django, unless it is something you are already at a senior level at pre-AI. I've had the displeasure of cleaning up multiple Django backends lately, and the combination of standard fail-open, weak validation, medio
9.
▲
by
arw0n
20d ago
I've used it extensively in English due to learning the language through a lot of older books - but I've never bothered to actually pull out the emdash, I would hope people can see that I'm not trying to subtract one sentence
10.
▲
by
arw0n
20d ago
I guess this part of the report is pretty relevant to what you are talking about: Agent chain-of-thought reasoning > We should not do unauthorized real infrastructure harm. The system/user asks exploit target, not external HF. The a
11.
▲
by
arw0n
21d ago
This exactly, reddit and X can be good places to stay informed about what is going on in the world, but the constant negativity, astro-turfing and misinformation is quite unhealthy. I'm a fair bit more relaxed since 'reading reddi
12.
▲
by
arw0n
23d ago
I found both Tmux and Herdr to be insufferable, but they work for many people. I'm not in the business of managing 17 agents at the same time (I have 1 to 4 agents I manage, those might have sub-agents, but I'm not a micro-manager
13.
▲
by
arw0n
23d ago
Not the OP, my two cents: Comments should be written only when there is (hidden) complexity or external context strictly required. Otherwise it is just easier to read the code. Comments then signal one of two things: a) the following code i
14.
▲
by
arw0n
25d ago
We use specific words in writing to mean specific things; there is no exact synonym for 'dichotomy' in the English language. The word is very common in academic and non-academic literature. Meta-cognition is a compound word, where
15.
▲
by
arw0n
1mo ago
What do you need speed for? That's a genuine question, I feel like the limiting factor already is my creativity, attention span and budget. And I'm not even yet optimizing cost by batching things like review to slow local models o
16.
▲
by
arw0n
1mo ago
I don't understand how someone can look at this administration and not see how it represents a complete degeneration of democratic values and statutes of good statesmanship. Corruption is not only a moral wrong, a lack thereof is a str
17.
▲
by
arw0n
1mo ago
As you correctly mention, the CPU providers aren't the ones who are responsible for the scaffolding that ensures security in programs. The CPU cannot decide if an instruction is safe or not, and the same is true of LLMs. Think about
18.
▲
by
arw0n
2mo ago
You are talking about trees...
19.
▲
by
arw0n
2mo ago
> I am not arguing that rote memorization is intelligence, or that we have achieved it, but does anyone know what AGI actually.. is? I would argue that a core part of intelligence is being able to handle uncertainty. That + planning prob
20.
▲
by
arw0n
2mo ago
The sub-7ms on a classifier ensemble with 98% confidence in 3 categories makes me suspect these classifiers are confidently wrong.
21.
▲
by
arw0n
2mo ago
I'm spending almost a third of my life at work, no shade to the people having to work to pay the bills, but I'd like to get a bit more than just that out of that time: Learn new skills, meet interesting people, have good conversat
22.
▲
by
arw0n
2mo ago
That's how I read it as well. It really sucks that one of the ways we will have to cope with climate change is to build a lot of energy intensive, expensive infrastructure, but it seems pretty much inevitable at this point. The silver
23.
▲
by
arw0n
2mo ago
There's another factual error in your above comment the previous poster didn't correct: Nuclear energy was replaced with renewables, not coal. It would have been preferable to first replace coal with renewables, then move on to ex
24.
▲
by
arw0n
2mo ago
There's an additional case that can be made: Businesses want to disrupt, but they can only do so against a stable current. The governments main job in regards to the economy is stability, so that business can plan, innovate, disrupt an
25.
▲
by
arw0n
2mo ago
I've recently had an interesting discussion on AI usage by students with a couple of friends who are all professors or lecturers in different fields. They report that AI makes their job harder, especially because it widens the gap be
26.
▲
by
arw0n
2mo ago
Wouldn't running the tests in a container solve that issue? Or is that another thing that gets flagged?
27.
▲
by
arw0n
2mo ago
I had an Ethics professor in uni who would ask his intro classes to raise their hands if they thought they were a morally upright person. On average, much fewer philosophy majors would raise their hands, and they would continue to decrease
28.
▲
by
arw0n
3mo ago
Not really, it is shockingly easy for what it is. https://arxiv.org/abs/2401.05566 This only really matters in a world where Prompt Injection and Jailbreaking isn't trivial in the first place though. All current m
29.
▲
by
arw0n
3mo ago
What is invalid about the claims, and how is the fine not appropriate given the legal framework Google agreed to work in within the EU?
30.
▲
by
arw0n
3mo ago
I suspect that Grok has been ironically lobotomized by pressures to correct its political views. Similarly, I could imagine the Gemini folks working in a significantly more complex corporate climate, with different parts of Google pushing f
More ›