Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
randomblock1
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
randomblock1
5d ago
You know how there's a router mode to use the cheapest provider? That only takes into account uncached rates, last I checked. Make another one that takes into account effective rates (the ones that include cache).
2.
▲
by
randomblock1
6d ago
Oh I tried in cloud. I'll give it another shot
3.
▲
by
randomblock1
6d ago
So then they DIDN'T "learn the secret to cracking the problem". They simply knew that part of the problem was solved. Knowing a problem can be solved and knowing the solution are not the same thing.
4.
▲
by
randomblock1
6d ago
I just gave it a try and it doesn't appear to be free, it used up some of my on demand usage. It does say 75% off though. Seems like for Pro subscribers SWE-1.7 is free, maybe SWE-2 is free for them?
5.
▲
by
randomblock1
7d ago
It's the website's "grid-bg" effect. It's glitching out and also normally barely visible at all.
6.
▲
by
randomblock1
13d ago
The writer is the CEO of... whatever this is: https://auraspark.com/ Completely slopped up website, zero human touch. I would even go so far as to call them a hyprocrite, for getting AI to do all that for them instead of, y
7.
▲
by
randomblock1
14d ago
I think it's still a useful data point. For example, omp, which is pi with some default extensions, scores worse. I do agree that adding more configurations of Pi would help though.
8.
▲
by
randomblock1
14d ago
The retain it for "the duration necessary to achieve the intended purposes", which could mean forever.
9.
▲
by
randomblock1
16d ago
It's not that far off anymore. On my 7900 XTX 24GB, I can run Qwen3.8 27B with 131K context at Q4_K_M (55 tok/s with MTP). Excluding hardware cost, it's about $0.02 tok/M in and $0.40 tok/M out (cached in $0.0001).
10.
▲
by
randomblock1
20d ago
To me, this seems like OPC UA / SiLA but instead of being software <-> machine semantics/control it's AI <-> machine semantics/control. Or in simpler terms, an AI-facing hardware abstraction / device-des
11.
▲
by
randomblock1
21d ago
Anything with motion. Sports, games, vlogs, and so on experience immense improvements. And it's not like it takes 2x the bandwidth, because inter-frame compression can be smarter about it. It's like 30-50% more bandwidth.
12.
▲
by
randomblock1
22d ago
I find the tokenizers most compelling. That's what the model is trained on, it's an immutable fact of the model and its architecture. You know for a fact that the model is at least related to other models that way. And if a tokeni
13.
▲
by
randomblock1
22d ago
Even at the hardware level, if it was a separate chip that the camera data passed through or something, that's not really good enough either, people have broken TPMs before. It'd have to be baked into the camera sensor. Even then,
14.
▲
by
randomblock1
23d ago
Yeah seems like it doesn't account for anti-fingerprinting at all.
15.
▲
by
randomblock1
25d ago
They even go into the negative numbers. https://minusonelabs.com/ https://minus3labs.com/ https://minus9labs.com/ (borked but existed at one point) It's getting ridiculous, frankly
16.
▲
by
randomblock1
28d ago
Yes: https://huggingface.co/collections/ornith-ai/ornith-15
17.
▲
by
randomblock1
29d ago
It's not just the chassis, it's also the screen, storage, RAM, ports, speakers, battery, and so on. And there is lots of demand for old mainboards, they sell pretty quick on Ebay. More sustainable to only buy and sell the part you
18.
▲
by
randomblock1
1mo ago
How exactly is this different from something like v86? It's definitely easier to embed but also not as customizable. Like if I don't need bioinformatics stuff, can I just exclude that?
19.
▲
by
randomblock1
1mo ago
TLDR: It keeps ~20k tokens of recent conversations, then hands the rest of the conversation to another model with a special system & user prompt. This then fills out a template with relevant information. See: https://github.c
20.
▲
by
randomblock1
1mo ago
The ENTIRE thing is AI generated. I'm not talking about the article. I'm talking about the entire website, the entire "product". https://0.mk/blog/zero-humans Also it's super easy to tell by lo
21.
▲
by
randomblock1
1mo ago
I think it's just meant to make it more competitive, Gemini has kinda been behind in everything except maybe multimodal. It's only 3 weeks after Flash 3.6, so if they really wanted to, they could probably do a 3.8 Flash before the
22.
▲
by
randomblock1
1mo ago
> afaik the PAYG subscribers are not subject to this lower limit ... In other words, if you want the old provisioning limits for free you just have to put your credit card info in. Always Free can mean different things: to a tenancy like
23.
▲
by
randomblock1
1mo ago
Most of the friction is just JS overhead for the computations, a compiled solver is like 1000x faster. If Anubis ever gets popular enough that scrapers care, it would be trivial to defeat. And last I checked you could bypass it by just modi
24.
▲
by
randomblock1
1mo ago
It supports anything with ACP. So it can actually run Codex and Claude Code, not just the Zed Agent. Looks like omp supports ACP, so all you have to do is specify a Custom Agent in Zed and it should just work. https://omp.sh/
25.
▲
by
randomblock1
1mo ago
By your definition a Lamborghini is affordable, because someone can afford it. I don't think that holds up.
26.
▲
by
randomblock1
1mo ago
What about AMD? I'm guessing it's not supported, which is a shame, because they're better value for VRAM.
27.
▲
by
randomblock1
2mo ago
> Complexity cannot be deduced from the prompt alone. Let’s take an example: “evaluate the tests for the repo $GIT_REPO and improve them” can be a very simple task if you mention a personal website written in plain HTML5; or an incredibl
28.
▲
by
randomblock1
2mo ago
I bet they'll switch over pretty soon. They always make free users use the older models for a little bit, probably to try to push people to upgrade. You can actually use Luna without reasoning (set it to "none"). So if they w
29.
▲
by
randomblock1
2mo ago
I don't see why in the future this couldn't be applied to a semi-flexible backing. Probably will be fragile and hard to make but seems possible.
30.
▲
by
randomblock1
2mo ago
Default to 1h. Allow setting it to longer or shorter, granularly. Add /pause to mark it for eviction, for a token refund.
More ›