Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
npn
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
npn
3d ago
People do love auditing everyone else’s donations.
2.
▲
by
npn
5d ago
As expected vibecoding bros cannot even read the manual properly. It is pretty trivial to pin a single provider for a model. Better yet, instead of calling the model directly, use presets instead. You can easily change the setting on openro
3.
▲
by
npn
7d ago
it is a way bigger model with extra 200B engram so of course the score improves. can't wait for deepseek v4.1 pro
4.
▲
by
npn
7d ago
Crazy that they still keep the price -- or actually decrease it, even -- despite it is a big improvement. I hope it retains some of the tps speed of the preview release though, 300 tps means gemini flash is no longer "the fastest optio
5.
▲
by
npn
8d ago
yeah raid or not you still get the hard limitation by the pcie lanes it is even worse with 40 macbooks. if 40 macbooks is all that take to serve a 1TB model with decent speed then you would see everyone selling the models for very cheap rig
6.
▲
by
npn
12d ago
luckily I have purchased a used amd mi50 32gb card for pretty cheap back then. while I haven't used it extensively, it feels pretty great having a backup plan that does not depend on any external 3rd parties.
7.
▲
by
npn
14d ago
openai did human crafted chain of thought dataset training. deepseek didn't have the resources so they attempted RL. doing RL correctly is hard because of the risk of model collapsing.
8.
▲
by
npn
14d ago
very important actually. just try to generate code for fresher frameworks/libraries. gemini sucks so bad in real work usage, everything it suggests are outdated and mostly useless.
9.
▲
by
npn
14d ago
Still refuse to search internet for stuff it thinks does not exist lol. And even when searching for internet, it still cannot suggest a up-to-date approach to the problem. For example I'm using crystal, it recently revamped the concurr
10.
▲
by
npn
19d ago
Because you can also use grok and a dozen other models. In fact grok is preferred choice for cursor right now, so obviously the other models get sidelined
11.
▲
by
npn
21d ago
1. they still have revenue though. it might not enough to cover all the r&d but it is surely enough to cover the hardware cost. 2. people tend to ignore this, but the salary budget of a US frontier lab and chinese frontier lab is nowher
12.
▲
by
npn
22d ago
I'm confused? Can you just define some presets and call them instead? With preset you can pinpoint a lot of things, especially the providers
13.
▲
by
npn
23d ago
? Why are you assuming that I don't know about that. Actually have you really design practical webapps before? Because your post only have empty pretty words with no substance. Semantic html tags are just semantic, it is irrelevant for
14.
▲
by
npn
23d ago
what hard is making a fully featured website, with panels no wider than 60ch. typically 60ch equal to 480px (font size 16px), so you need sidebars to the left and the right. which is fine, the holy grail was like that. but if you want to de
15.
▲
by
npn
23d ago
I have Mi Notebook Pro. Not that bad actually. But they have stopped making premium laptop since then.
16.
▲
by
npn
24d ago
I also did some experiment with ch many years ago. I found that 60ch is ideal width for block text for easy reading. too bad it is pretty hard to make websites with only 60ch wide.
17.
▲
by
npn
25d ago
didn't zai already do that with their coding plan? I mean they surely had to pay users to use claude models (paying the differences). they also funded some newapi token resale websites.
18.
▲
by
npn
26d ago
I hope it is glm air. We need more "small" models. Big models are more capable and useful, but for majority of tasks some smaller models can work just fine. It is funny that google gave up on this market, leaving the whole price r
19.
▲
by
npn
26d ago
Ok that antise... We all know who you really want to criticize here.
20.
▲
by
npn
27d ago
No but with 100% clean data you can easily train a model to filter ai generated content.
21.
▲
by
npn
27d ago
I wonder how many posts in this thread are AI generated or shill posted. nobody ever reports that they installed it and ran it on production or something. personally I only use bun to replace yarn as script executor now. Used to follow it a
22.
▲
by
npn
1mo ago
There are like thousands sites with similar features all using newapi core. You can easily find them in Chinese tech forum linux.do
23.
▲
by
npn
1mo ago
> used up internet-scale data yet but it is still contain a lot of trash. you need better models to process those trash and create a curate dataset. this will happen again and again until there is no more juice to squeeze. and I'm s
24.
▲
by
npn
1mo ago
it is partly true, but like I said it is not 2025 anymore. models now get released more often, and still have notable progress so they can safely replace the old models while being faster/cheaper. and thank to chinese models the pricin
25.
▲
by
npn
1mo ago
> * For 3.6 and 3.7 Flash, introductory price expires on December 31, 2026. Starting January 1, 2027, $1.50/1M input tokens and $7.50/1M output tokens will apply. this is hilarious. it is not 2025 any more, by Jan 2027 there wi
26.
▲
by
npn
1mo ago
sorry LLM output is considered public domain in my country.
27.
▲
by
npn
1mo ago
nah, both point to the same model for deepseek. it's just openrouter weirdness.
28.
▲
by
npn
1mo ago
I don't think so. there is a lot of tools with similar usage, some harness even bring their own internal tools for accurately manipulation. also, even if some models claim that they have full 1M context window, some only work effective
29.
▲
by
npn
1mo ago
wait for Deepseek Harness (yes it is the official name) release then try again. for your kind of task, harness tools matter.
30.
▲
by
npn
1mo ago
I still believe this is not the full potential of pro models. I expect they will release another checkpoint later this year.
More ›