Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
DetroitThrow
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
DetroitThrow
8d ago
Yes exactly. Automatic model downgrade seems horrible for a lot of production workloads, even if you are deterministically constraining the behavior of your agents.
2.
▲
by
DetroitThrow
8d ago
I don't think the person describing the paper by Buckmaster as the same as the 2023 paper by Córdoba and Martínez-Zoroa is really discussing this in good faith fwiw. There are some massive advancements within it and if the person was p
3.
▲
by
DetroitThrow
8d ago
>It's unknowable and not possible to prove if any one specific conversation was the key to solving Navier–Stokes. If the conversation was in the training set, there's a high likelihood that the small set of conversations relate
4.
▲
by
DetroitThrow
8d ago
It's unclear if you're suggesting that OpenAI did not train on their input or use their chats as inputs to training on a model that found the solution. Let's not provide an Elizabeth Holmes-esque interview where the question
5.
▲
by
DetroitThrow
8d ago
Given that he has other former collaborators corroborating this horrific behavior, it seems like this a career spanning pattern, and it's interesting to see just how much @sama is willing to lend his support to someone like Bubeck. Sta
6.
▲
by
DetroitThrow
13d ago
404?
7.
▲
by
DetroitThrow
19d ago
Can't check whether star citizen is done yet :/
8.
▲
by
DetroitThrow
19d ago
I don't mind color, the glow effects just hurt the ability to read what's there (this reminds me of the gaming setups teenagers are drawn to)...but there's no reason a utility like a task manager should be closed source or ha
9.
▲
by
DetroitThrow
19d ago
Wow, thought the guy would be a MSFT OG but this looks like it sucks? Vibed, closed source, paid extensions.. Why on earth would anyone use this?
10.
▲
by
DetroitThrow
20d ago
It's a bit of both for Anthropic I think, sometimes cutting edge and quite interesting or just good improvements, sometimes ignoring best practices either recently established or known for decades. Obvious to see where the smart people
11.
▲
by
DetroitThrow
21d ago
Lately I've been throwing tasks at Qwen and a frontier or recently-frontier model (as well as Kimi, GLM, etc) and the smaller parameter models are not really comparable to Opus when it comes to making intelligent decisions about greyer
12.
▲
by
DetroitThrow
1mo ago
It's still not as good as GPT5.6 or Opus5 but it's better than KimiK3. Good job xAI team.
13.
▲
by
DetroitThrow
1mo ago
I've tried it on my "let's run every model in parallel and see which finds more edge cases" type of tasks, and Grok 4.5 was really behind Opus/ChatGPT but ahead of Gemini - despite having a strong showing on benchma
14.
▲
by
DetroitThrow
1mo ago
>Aren't the LLMs trained on a massive corpus of human written texts? If that stands, then they are doing what they were asked, kind of? I think you might be interested in reading in training data generation, training, and post train
15.
▲
by
DetroitThrow
1mo ago
This blogpost didn't really address the elephant in the room that increasingly the bottleneck is no longer the part that jr devs could help out with. Every startup I know that hired some level of jr's and encouraged them to use AI
16.
▲
by
DetroitThrow
1mo ago
They also hike the price so significantly most people stop using it. See meetup.com for example.
17.
▲
by
DetroitThrow
1mo ago
yes, _every_ event I go uses Luma or something else nowadays.
18.
▲
by
DetroitThrow
2mo ago
How many times have they managed to catch it in the last 8 or so missions? There have been a few misses like this already with the boosters
19.
▲
by
DetroitThrow
2mo ago
>FFCS is called holy grail of liquid engines Particularly for reusable liquid engines :)
20.
▲
by
DetroitThrow
2mo ago
Looks like they reset everyone's Fable usage.
21.
▲
by
DetroitThrow
2mo ago
DeepSWE seems to strongly, strongly prefer ChatGPT models. There were also major flaws in its methodology pointed out recently, that overlap strongly with the flaws OpenAI pointed out in its SWE Verified report. I use both ChatGPT and Claud
22.
▲
by
DetroitThrow
2mo ago
>I'll be more peeved if they monetize it FSL (vs a copyleft license or just plain old OSS) implies they want to turn this into a revenue source for themselves ultimately, unfortunately. >Maybe I should put one of those buy me a c
23.
▲
by
DetroitThrow
2mo ago
On top of that, it has a more restrictive license than AmazonBrandFilter. Given this appears to be a very simple AI project, why not just reimplement any missing functionality from AmazonBrandFilter into something under a free license? The
24.
▲
by
DetroitThrow
3mo ago
She's not as big on some of the broader interpretations of the 4th amendment that more civil liberty minded justices would lend credence to.
25.
▲
by
DetroitThrow
3mo ago
He's entitled to his political views and just as we're entitled to potentially use or not use his service because of them :) Not sure why it's such an issue to discuss the political views of the beneficiaries of services we u
26.
▲
by
DetroitThrow
3mo ago
Everyone gets to share but it's also completely within the forum rules to call out irrelevant anecdotes as uninteresting to the discussion. I have no idea why you're making a comparison to a TV show; nothing that was described was
27.
▲
by
DetroitThrow
3mo ago
When performance isn't a concern, I largely agree! Not every financial system can use big decimal as their base, though, too. And HFT isn't the only place in the financial sector where this performance concern might pop up.
28.
▲
by
DetroitThrow
3mo ago
"10% of Americans are uninsured. A US state is pushing to insure all of their residents." "I'm insured!" "Open-source software projects are being spammed with LLM generated PRs. Contributions are becoming more
29.
▲
by
DetroitThrow
3mo ago
Agree with this, working from HFT to payments to account management in the past. You can have the blockchain team be an expert in converting integer cents, or the forex team be an expert in sub-cent conversions. You don't want to requi
30.
▲
by
DetroitThrow
3mo ago
Spam from a bot! I think there are automatic filters to prevent this so I'm surprised it wasn't banned before we all had to see it, maybe @dang can comment on how quickly these are supposed to be resolved
More ›