Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
skiing_crawling
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
skiing_crawling
4d ago
> you keep poking This is waving over engineering an agent with tools, harness, prompts, and loops. The models are still just next token predictors and everything, including predicting more than 1 token, is the result of outside "po
2.
▲
by
skiing_crawling
4d ago
I don't really believe any of it. I've seen articles for nearly 2 years now about "agent" automonously doing things like blackmail, hacking, coordinating. But during that same time, I've used o3 up to fable, sol, an
3.
▲
by
skiing_crawling
7d ago
the "new" Twitter basically looks like a cash/name grab, I was disappointed to see that nobody involved in it was affiliated the Twitter, and it's mostly run by non-technical/lawyer types.
4.
▲
by
skiing_crawling
9d ago
yeah took me a few tries and then it didn't even go to the content
5.
▲
by
skiing_crawling
13d ago
Having a hard time understanding what MCP is really for. Even for my small local models, if MCP is not available, they seem to do just fine connecting to anything I need with an API and falling back to using a browser.
6.
▲
by
skiing_crawling
15d ago
All the benchmarks in the world don't matter if the model just straight up refuses to do mundane things. Claude has too much of an attitude.
7.
▲
by
skiing_crawling
21d ago
I didn't use the word "just"
8.
▲
by
skiing_crawling
21d ago
It is revisionist to say that software engineering was never about writing code. It was, in fact, a huge component, and it also wasn't easy. Sure most code is glue but even the glue was tedious and the actual hard and novel parts still
9.
▲
by
skiing_crawling
1mo ago
This is phone OS developer's fault for even allowing it. When I take a screenshot, I expect to have an image of exactly whatever was displayed on the screen at the time. Its not a picture of your app, its a picture of my screen. Some
10.
▲
by
skiing_crawling
2mo ago
This has that weird "AI written" sentence structure.
11.
▲
by
skiing_crawling
3mo ago
"it can fit" on 256GB of RAM, but it will be heavily quantized and still run very slowly. The headline number is not token generation, its prompt processing. So if you get 10 tok/s and an API gives you 20-30 tok/s, it do
12.
▲
by
skiing_crawling
3mo ago
unauthorized? Are they trying to find a middle ground word between "illegal" and "undocumented"?
13.
▲
by
skiing_crawling
3mo ago
Any generic abliterated or ubcensored open weight model (such as a qwen variant) will happily comply with requests like this.
14.
▲
by
skiing_crawling
3mo ago
Maybe some (or many) people believe that more people will make it less "lovely". I think this is a popular stance and I think many people are more than satisfied with the current population density of their area.
15.
▲
by
skiing_crawling
4mo ago
Does my concern somehow become less valid because I'm American? Everyone should be thinking carefully about which of their data is going where.
16.
▲
by
skiing_crawling
4mo ago
I guess I was speaking as an American, we have good domestically hosted options so although it’s probably not ideal to send this kind of data/control anywhere at all, it’s definitely a worse option for us to send it to china vs to an A
17.
▲
by
skiing_crawling
4mo ago
I’m worried about giving a foreign hosted service access to my machine for a coding agent that can run arbitrary commands and read arbitrary files. Coding agent are much more useless if you have to sit there clicking approve on everything.
18.
▲
by
skiing_crawling
4mo ago
I recently built a system at insane ddr4 prices ($2000 for 256gb). But that’s only after seeing how ddr5 prices were 3-4x that!
19.
▲
by
skiing_crawling
4mo ago
Using an Epyc platform to get plenty of PCIe lanes and memory channels. I have couple of extra 3090s plugged in which get some offload and help with larger models that don't fit entirely on the blackwell.
20.
▲
by
skiing_crawling
4mo ago
> Even if LLMs fail spectacularly Haven't they already proven to be extremely useful? In some areas they are definitely here to stay, coding/software and search (retrieve and summarize information). There's a bunch of plac
21.
▲
by
skiing_crawling
4mo ago
You can get 70-80 tps on qwen3.6-27b f16 with MTP on a single card
22.
▲
by
skiing_crawling
4mo ago
I got an RTX 6000 pro too. I like running locally, I've learned a lot more than if I had used an API and there's less worry about overspending tokens. I accidentally spent $100 on claude api in like 2 days because I didn't kn
23.
▲
by
skiing_crawling
4mo ago
At this point IPOs are mainly for unloading bags onto retail. Every institution who wanted a piece of these labs got in years ago and captured all the value.
24.
▲
by
skiing_crawling
4mo ago
I use claude/gemini as my homepage now (I have to keep switching as these companies make "updates" that periodically render their models useless). Even if I want to search for simple things, I would rather have an LLM wade th
25.
▲
by
skiing_crawling
4mo ago
They won't, its literally part of their sales funnel. They've specifically engineered a bad experience for anyone outside the ecosystem by making it all of their friend's problem too. Its very important for their stock price
26.
▲
by
skiing_crawling
4mo ago
triggered me with that first sentence
27.
▲
by
skiing_crawling
4mo ago
Is this LLM psychosis? So much tending and conversing with the matmuls but what was the outcome? Are people who get this into it more successful somehow? It reminds me of people who take drugs and get "revelations" but then are no
28.
▲
by
skiing_crawling
4mo ago
I've been using Siri (via homekit) to turn all my lights on and off for about 3 years now. It's steadily getting worse and worse as somehow, Siri is becoming less accurate and Apple is failing to adopt this new technology in a tim
29.
▲
by
skiing_crawling
4mo ago
AI didn't start this, journalist have been using wordplay to "technically tell the truth" forever.
30.
▲
by
skiing_crawling
4mo ago
what does it have to do with git?
More ›