Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
jpcompartir
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
jpcompartir
1mo ago
I can't be the only person who never finds these made up dialogues funny, right?
2.
▲
by
jpcompartir
3mo ago
The absolute worst name for a model I've seen
3.
▲
by
jpcompartir
3mo ago
After plotnine, with a solid & performant (more than the R versions) Python version of Purrr and Dplyr I might never reach for R again!
4.
▲
by
jpcompartir
3mo ago
They weren't freaked by anything, it's a retaliatory shakedown after ideological differences and Anthropic not doing exactly what they're told/what the Admin wants them to do.
5.
▲
by
jpcompartir
3mo ago
After a day or so this is the first model that really feels next level compared to how Opus 4.5 felt on release
6.
▲
by
jpcompartir
4mo ago
Great person and great company I hope he still gets to do some educative stuff on the side too
7.
▲
by
jpcompartir
5mo ago
Anthropic releases used to feel thorough and well done, with the models feeling immaculately polished. It felt like using a premium product, and it never felt like they were racing to keep up with the news cycle, or reply to competitors. Re
8.
▲
by
jpcompartir
5mo ago
Likewise, I foolishly assumed everybody else was just doing it wrong. But this week I've lost count of the times I've had to say something along the lines of: "Can you check our plan/instructions, I'm pretty sure I
9.
▲
by
jpcompartir
6mo ago
This looks like a Claude-generated SVG to me, is it not?
10.
▲
by
jpcompartir
6mo ago
Fair push back, but I do think the LSTM vs Transformers point kinda supports my position in the limit, not refutes. Once the compute bottleneck is removed, LSTMs scale favourably. https://arxiv.org/pdf/2510.02228 (I b
11.
▲
by
jpcompartir
6mo ago
There are better techniques for hyper-parameter optimisation, right? I fear I have missed something important, why has Autoresearch blown up so much? The bottleneck in AI/ML/DL is always data (volume & quality) or compute. Doe
12.
▲
by
jpcompartir
7mo ago
As a non-US citizen, I'm quite glad in the knowledge that Claude won't be used to kill other non-US citizens with autonomous weapons
13.
▲
by
jpcompartir
7mo ago
"Regardless, these threats do not change our position: we cannot in good conscience accede to their request."
14.
▲
by
jpcompartir
7mo ago
This is great, brings clear benefits to both sides and the rest of us. Always rooting for Hugging Face
15.
▲
by
jpcompartir
7mo ago
Yep, Gemini is virtually unusable compared to Anthropic models. I get it for free with work and use maybe once a week, if that. They really need to fix the instruction following.
16.
▲
by
jpcompartir
7mo ago
Thanks for the long and considered response, but this is a really ugly UX decision. As others have said - 'reading 10 files' is useless information - we want to be able to see at a glance where it is and what it's doing, so t
17.
▲
by
jpcompartir
7mo ago
Yeah 100% This won't change my decision, but it is still impeccable timing
18.
▲
by
jpcompartir
7mo ago
This is great, not 10 minutes before this outage did I present Railway as a viable option for some small-scale hosting for prototypes and non-critical apps as an alternative to the Cloud giants
19.
▲
by
jpcompartir
7mo ago
4.6 is a beast. Everything in plan mode first + AskUserQuestionTool, review all plans, get it to write its own CLAUDE.md for coding standards and edit where necessary and away you go. Seems noticeably better than 4.5 at keeping the codebase
20.
▲
by
jpcompartir
8mo ago
I've been working with a claude-specific directory in Claude Code for non-coding work (and the odd bit of coding/documentation stuff) since the first week of Claude Code, or even earlier - I think when filesystem MCP dropped. It&
21.
▲
by
jpcompartir
11mo ago
I can't remember which paper it's from, but isn't the variance in performance explained by # of tokens generated? i.e. more tokens generated tends towards better performance. Which isn't particularly amazing, as # of tok
22.
▲
by
jpcompartir
11mo ago
Most comments seem to be taking the code seriously, when it's clearly satirical?
23.
▲
by
jpcompartir
1y ago
Assuming you've read OpenAI's paper released this week? https://cdn.openai.com/pdf/d04913be-3f6f-4d2b-b283-ff432ef4a... They attribute these 'compression artefacts' to pre-training, they also refere
24.
▲
by
jpcompartir
1y ago
Polars is great, absolute best of luck with the launch
25.
▲
by
jpcompartir
1y ago
You seem to be responding to a strawman, and assuming I think something I don't think. As of today, 'bad' generations early in the sequence still do tend towards responses that are distant to the ideal response. This is testa
26.
▲
by
jpcompartir
1y ago
Interesting, in the LLM case these compression artefacts then get fed into the generating process of the next token, hence the errors compound.
27.
▲
by
jpcompartir
1y ago
I would echo some caution if using as a reference, as in another blog the writer states: "Backpropagation, often referred to as “backward propagation of errors,” is the cornerstone of training deep neural networks. It is a supervised l
28.
▲
by
jpcompartir
1y ago
^ And if we increase N enough we will be able to find these 'good measurements' and 'statistically significant differences' everywhere. Worse still if we did not agree in advance what hypotheses we were testing, and go l
29.
▲
by
jpcompartir
1y ago
If you say there's no AI-generated code then I retract the original comment, nice work.
30.
▲
by
jpcompartir
1y ago
That is not a disclaimer for generated code, it's referring to the code that generated the simulations/plots. I had read that line before I commented, it was partly what sparked me to comment as it was a clear place for a disclaim
More ›