Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
byefruit
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
byefruit
10mo ago
I've just found myself using OpenRouter if we need Google models for a project, it's worth the extra 5% just not to have to deal with the utter disaster that is their product offering.
2.
▲
by
byefruit
10mo ago
This is the wrong interpretation of the oxcaml project. If you look at the features and work on it, it's primarily performance or parallelism safety features. The latter going much further than most mainstream languages.
3.
▲
by
byefruit
11mo ago
7.3% return, not bad. As battery prices drop it will get even better.
4.
▲
by
byefruit
1y ago
And even when it does copy other products, it seems to be doing a terrible job of them. Google's AI offering is a complete nightmare to use. Three different APIs, at least two different subscriptions, documentation that uses them inter
5.
▲
by
byefruit
1y ago
How is this different from https://github.com/google-gemini/gemini-cli ? Edit: it seems this is a hosted version. Would be nice if they actually joined up some of their products.
6.
▲
by
byefruit
1y ago
The openrouter rankings can be biased. For example, Google's inexplicable design decisions around libraries and APIs means it's often worth the 5% premium to just use OpenRouter to access their models. In other cases it's abo
7.
▲
by
byefruit
1y ago
Before just accepting this at face value, New Statesman claim this is not the case: https://www.newstatesman.com/politics/2025/07/the-british-we...
8.
▲
by
byefruit
1y ago
You are probably getting downvoted because you don't give any model generations or versions ('ChatGPT') which makes this not very credible.
9.
▲
by
byefruit
1y ago
100% this. We actually use OpenRouter (and pay their surcharge) with Gemini 2.5 Pro just because we can actually control spend via spent limit on keys (A++ feature) and prepaid credit.
10.
▲
by
byefruit
1y ago
Indeed, average in CA is $260/month so $5k pays off very fast in some places.
11.
▲
by
byefruit
1y ago
It's interesting that there's a price nearly 6x price difference between reasoning and no reasoning. This implies it's not a hybrid model that can just skip reasoning steps if requested. Anyone know what else they might be do
12.
▲
by
byefruit
1y ago
"In addition, S3 Express One Zone has reduced the per-GB charges for data uploads and retrievals by 60 percent, and these charges now apply to all bytes transferred rather than just portions of requests greater than 512 KB" It
13.
▲
by
byefruit
1y ago
> Both of our models are trained on top of DeepSeek-R1-Distill-Qwen-7B and DeepSeek-R1-Distill-Qwen-32B. Not to take away from their work but this shouldn't be buried at the bottom of the page - there's a gulf between completel
14.
▲
by
byefruit
1y ago
https://aws.amazon.com/snowball/pricing/ snowball seems to support getting data out of S3 though you still end up paying extortionate egress charges.
15.
▲
by
byefruit
2y ago
I think this is where (in the EU) a Subject Access Request could work.
16.
▲
by
byefruit
2y ago
This has already been proposed by the current government for wind farms: https://inews.co.uk/news/environment/new-energy-bill-discoun... And the switch to zonal energy pricing will likely have a similar effect for
17.
▲
by
byefruit
2y ago
I'm waiting for https://github.com/huggingface/trl/pull/2810 to land. I think this should work with the existing unsloth setup without changes.
18.
▲
by
byefruit
2y ago
> This generally requires thousands of examples created by an expert in the field. Or an AI model pretending to be an expert in the field... (works well in a few niche domains I have used this in)
19.
▲
by
byefruit
2y ago
I look forward to people applying the same standards to the OpenAI's O3 as they did Deepseek's R1 release and paper in the discussions last week.
20.
▲
by
byefruit
2y ago
That's not what I'm saying, they may be hiding their true compute. I'm pointing out that nearly every thread covering Deepseek R1 so far has been like this. Compare to the O1 system card thread: https://news.ycombi
21.
▲
by
byefruit
2y ago
It's amazing how different the standards are here. Deepseek's released their weights under a real open source license and published a paper with their work which now has independent reproductions. OpenAI literally haven't sai
22.
▲
by
byefruit
2y ago
What I don't understand is how you don't end up with a totally impractical number of vectors if you have one per token? Surely nobody is storing that many in any real system?
23.
▲
by
byefruit
2y ago
Do you have any evidence for this accusation? O1's reasoning traces aren't even shown, are you suggesting they've somehow exfiltrated them?
24.
▲
by
byefruit
2y ago
This is pretty harsh on DeepSeek. There are some significant innovations behind behind v2 and v3 like multi-headed latent attention, their many MoE improvements and multi-token prediction.
25.
▲
by
byefruit
2y ago
What would you recommend for building a strong linear algebra foundation?
26.
▲
by
byefruit
2y ago
As sibling comments have said, you can pay per month to use the web interface, pay for use of the API via self-signup or you can even use them via openrouter.
27.
▲
by
byefruit
2y ago
I've seen a couple of talks on DSPy and tried to use it for one of my projects but the structure always feels somewhat strained. It seems to be suited for tasks that are primarily show, don't tell but what do you do when you have
28.
▲
by
byefruit
2y ago
Bit weird to show $ per 1M tokens and not include the actual costs of the systems anywhere. It would be interesting to know the outright prices for those systems as well as their hourly rental rates at the moment.
29.
▲
by
byefruit
2y ago
This is very interesting. Have you got any references describing this approach?
30.
▲
by
byefruit
2y ago
A troll so good it necessitated a change in the law: https://publications.parliament.uk/pa/bills/cbill/58-03/0154... (Page 16, 57A) "A company must not be registered under this Act by a name that, i
More ›