Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
Benjammer
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
by
Benjammer
6d ago
Recently? Blind (teamblind.com) culture has been taking over tech for like a decade at this point.
2.
▲
by
Benjammer
2mo ago
I usually just say “make sure this code is professional and ready to deliver as a senior engineer” and it usually infers all that stuff you said plus more things as well. I try to give it the goal and let it decide what to do. One thing I u
3.
▲
by
Benjammer
5mo ago
Now do Long Island City in Queens, where 44th Ave, 44th rd, and 44th st are all in a row of blocks parallel to each other.
4.
▲
by
Benjammer
6mo ago
> the quality really does matter. If this level of quality/rigor does matter for something like a game, do you think the market will enforce this? If low rigor leads to a poor product, won't it sell less than a good product in
5.
▲
by
Benjammer
9mo ago
>Since COVID in CA, it feels like driving has become far more dangerous with much more lawlessness regarding excessive speeding and running red lights, going into the left lane to turn right in front of stopped cars, all sorts of weird t
6.
▲
by
Benjammer
9mo ago
>It is the loose equivalent of asking why are you getting hung up on the type of a variable in a programming language? A float or a string? Who cares if it works? No, it's not. This is like me saying "string and float are two t
7.
▲
by
Benjammer
10mo ago
Wind and sunshine are both types of weather, what are you talking about?
8.
▲
by
Benjammer
10mo ago
>They belong in different categories Categories of _what_, exactly? What word would you use to describe this "kind" of which LLMs and humans are two very different "categories"? I simply chose the word "cognition
9.
▲
by
Benjammer
10mo ago
So the idea is what? What's the successful outcome look like for this test, in your mind? What should good software do? Respond and say there are 5 legs? Or question what kind of dog this even is? Or get confused by a nonsensical pictu
10.
▲
by
Benjammer
10mo ago
It always feels to me like these types of tests are being somewhat intentionally ignorant of how LLM cognition differs from human cognition. To me, they don't really "prove" or "show" anything other than simply - LL
11.
▲
by
Benjammer
10mo ago
amusement park --> park amusement... Is that the joke?
12.
▲
by
Benjammer
11mo ago
I mean ok, but it's all just prompting on top of the same base model weights... I tried the same prompt, and I simply added to the end of it "Prioritize truth over comfort" and got a very similar response to the "improve
13.
▲
by
Benjammer
1y ago
I'm not sure how much experience you have, I'm not trying to make assumptions, but I've been working in software over 15 years. The exact skill you mentioned - can visualize the plan for a change quickly - is what makes my LL
14.
▲
by
Benjammer
1y ago
My method is that I work together with the LLM to figure out the step-by-step plan. I give an outline of what I want to do, and give some breadcrumbs for any relevant existing files that are related in some way, ask it to figure out context
15.
▲
by
Benjammer
1y ago
Are you unaware of the concept of a junior engineer working in a company? You realize that not all human code is written by someone with domain expertise, right? Are you aware that your wording here is implying that you are describing a uni
16.
▲
by
Benjammer
1y ago
Why is the "threshold" argument never the first thing mentioned? Do you not understand what I'm saying here? Can you explain why the "code slop" argument is _always_ the first thing that people mention, without disc
17.
▲
by
Benjammer
1y ago
This is the common refrain from the anti-AI crowd, they start by talking about an entire class of problems that already exist in humans-only software engineering, without any context or caveats. And then, when someone points out these probl
18.
▲
by
Benjammer
1y ago
I found "intertwining" with a score of 3 also. Two instances of the word on the same sign and then a false positive third pic.
19.
▲
by
Benjammer
1y ago
Why are engineers so obstinate about this stuff? You really need a GUI built for you in order to do this? You can't take the time to just type up this instruction to the LLM? Do you realize that's possible? You can just write inst
20.
▲
by
Benjammer
1y ago
Are you paying for the higher end models? Do you have proper system prompts and guidance in place for proper prompt engineering? Have you started to practice any auxiliary forms of context engineering? This isn't a magic code genie, it
21.
▲
by
Benjammer
1y ago
This isn't a financial model, they aren't selling the system itself, it's all tooling for data access and financial modeling. It's like they're setting up an OTB, not like they're selling you a system to pick w
22.
▲
by
Benjammer
1y ago
The fact that Thiel backs him so hard is what worries me more than anything. Thiel has a way of making things happen when he's really committed to something on a personal level... (see the Gawker Media case)
23.
▲
by
Benjammer
1y ago
This kind of “hair splitting” is the foundation on current prompt engineering though…
24.
▲
by
Benjammer
1y ago
That's some impressive prompt engineering skills to keep it on track for that long, nice work! I'll have to try out some longer-form chats with Gemini and see what I get. I totally agree that LLMs are great at compressing informat
25.
▲
by
Benjammer
1y ago
I mean, you could build this, but it would just be a feature on top of a product abstraction of a "conversation". Each time you press enter, you are spinning up a new instance of the LLM and passing in the entire previous chat tex
26.
▲
by
Benjammer
1y ago
It's nice to see a paper that confirms what anyone who has practiced using LLM tools already knows very well, heuristically. Keeping your context clean matters, "conversations" are only a construct of product interfaces, they
27.
▲
by
Benjammer
1y ago
Avogadro's number has a 10^23 in it to account for this atom-->physical matter sort of "scale up" conversion. Atoms are really small...
28.
▲
by
Benjammer
1y ago
I've found that heavily commented code can be better for the LLM to read later, so it pulls in explanatory comments into context at the same time as reading code, similar to pulling in @docs, so maybe it's doing that on purpose?
29.
▲
by
Benjammer
1y ago
One problem is how can you even set up a "fair" competition between an AI and Rainbolt? He does ones where it flashes for a fraction of a second and then he guesses the country. How do you simulate "only saw it for a fraction
30.
▲
by
Benjammer
1y ago
This is the nerdiest way I've ever seen someone talk about John Cena
More ›