Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
_aavaa_
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
_aavaa_
6d ago
What’s better about it the say oh-my-pi?
2.
▲
by
_aavaa_
7d ago
Yes, and? This is their current off-peak pricing for their flash model [0]: $0.007 cache, $0.22 input, $0.66. 0.007 -> 0.003 0.22 -> 0.15 0.66 -> 0.60 Each one is now cheaper. [0]: https://api-docs.deepseek.com/quic
3.
▲
by
_aavaa_
7d ago
I don't see them turning people back from buying tokens through the API. Until then, I don't see why we should follow this argument.
4.
▲
by
_aavaa_
7d ago
They might need to step up their product offerings and offer cheaper.
5.
▲
by
_aavaa_
7d ago
What are you talking about? Current flash prices are 0.66 for output, this is dropping it to 0.60.
6.
▲
by
_aavaa_
8d ago
They’re trying to kidnap what we have rightfully stolen.
7.
▲
by
_aavaa_
12d ago
Whether on net they turn a profit as company overall is neither here nor there.. My point is that they are selling API tokens at a profit (or if being pedantic, then at a price higher than the cost to serve them ignoring research costs). An
8.
▲
by
_aavaa_
12d ago
Unlikely, api pricing includes a healthy profit margin (as near as we can tell from the outside) which they wouldn’t charge themselves.
9.
▲
by
_aavaa_
12d ago
Sure privacy (or legality) considerations are valid, and depending on the subscription or the API you use, you may or may not get that. But we are not talking about the same product anymore. Your $12k homelab does not provide the same produ
10.
▲
by
_aavaa_
12d ago
You are misinformed on token amounts. As one example, z.ai [0] gives you somewhere in the ballpark of 150-300 Mtok/week, so 1-2x your amounts. Plus the model has a 1M window, and will be smarter. [0]: https://docs.z.ai/
11.
▲
by
_aavaa_
12d ago
I don't think that math will work out. If you are okay with using only 7M tokens per day, and 7M from a small model, then you don't have very demanding needs. If you don't have demanding needs, I think you're unlikely
12.
▲
by
_aavaa_
12d ago
So instead of paying $20 or maybe 100$ a month for the equivalent amount of output, I can spend $10k on GPUs (plus a few more thousands for related hardware), then for ongoing electricity, and then have space to store these things. You'
13.
▲
by
_aavaa_
13d ago
I really like the 3D version, but I strongly believe you need to consider the number of tokens required to complete a task, it heavily impacts the results for certain models that rely heavily on test time compute (Glm-5.3-flash is the newes
14.
▲
by
_aavaa_
13d ago
Artificial analysis, the place where they get this data from, has a cost vs time chart. Go to https://artificialanalysis.ai/ and scroll down to the second graph under “Speed & Latency”. I think this is the most import g
15.
▲
by
_aavaa_
14d ago
Shame. Thanks for the info.
16.
▲
by
_aavaa_
14d ago
Do they officially support you using your subscription in other harnesses?
17.
▲
by
_aavaa_
14d ago
I disagree. The y-axis is some arbitrary intelligence score that we use as a proxy for performance on whatever our specific task happens to be. So it doesn't matter if a model is a 0, 1, or 20 along this axis, they are all useless for
18.
▲
by
_aavaa_
14d ago
Do they officially support you use their AI Pro subscription (or whatever the heck it's called this month, the one that gives you models in antigravity) in a 3rd party harness?
19.
▲
by
_aavaa_
14d ago
> That conveniently ignores the drilling, extracting, transporting, refining, and transporting again to make "polyethylene resin". It also ignores the energy requirements to grow the cotton. This isn't the gotcha you think
20.
▲
by
_aavaa_
19d ago
The original comment I replied to states: “A 50% reduction would change nothing, reduction to 0 would change nothing.” When talking about individual data centers. My point is that this is true of basically any since source of pollution. Y
21.
▲
by
_aavaa_
19d ago
It very well could be faster, but right now it isn’t.
22.
▲
by
_aavaa_
19d ago
It’s cheaper sure, but it’s very slow. It’s not a drop in replacement
23.
▲
by
_aavaa_
20d ago
Can't wait for a society dependent on spice and which is still controlled those with political (or super human) powers.
24.
▲
by
_aavaa_
20d ago
That’s an argument about sequencing, not of necessity. And as an analogy it breaks down because we need to reduce/eliminate all sources of pollution, not just the current biggest one.
25.
▲
by
_aavaa_
20d ago
But not when you look at just one ship, or just one datacenter.
26.
▲
by
_aavaa_
21d ago
Let’s not pretend that what is currently in the US is real world capitalism, whatever that means.
27.
▲
by
_aavaa_
21d ago
This is an absurd argument. Any one source is a drop in the bucket.
28.
▲
by
_aavaa_
22d ago
He runs as PR company in addition to this blog. It goes with the territory.
29.
▲
by
_aavaa_
22d ago
> similar to how taxes are added after the fact That is not a positive. I'm fine with splitting up a price if you want to show how much tax gets added, but having to continuously do the mental math of "no this item is 10.99 it&
30.
▲
by
_aavaa_
23d ago
And not just the high level language output, or even the assembly. No no, every single piece of microcode produced.
More ›