Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
weiliddat
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
weiliddat
1mo ago
Relevant old school compression benchmarks where people have been using different models (incl. transformers) for compression: https://www.mattmahoney.net/dc/text.html Also interesting the top entry is from fabrice bel
2.
▲
by
weiliddat
1mo ago
I think someone else wrote it better than I can: https://news.ycombinator.com/item?id=49080703 This is worthy feedback to the author (and others), that writing with clear AI tells/register is so distracting that it tak
3.
▲
by
weiliddat
1mo ago
Why?
4.
▲
by
weiliddat
1mo ago
https://www.pangram.com/history/c837bbcb-4aed-43e1-987c-afe2... I know you're being downvoted, because people are tired of hearing "this is AI". But I think your comment "this is AI" is warrant
5.
▲
by
weiliddat
2mo ago
I'm sure you're where you are in some part due to hard work, skill, decisions. But surely you don't believe that it's purely down to skill where one lives/works at?
6.
▲
by
weiliddat
2mo ago
Don't disagree that that's the system not working well. You're lucky to be in a system that works out for you. Not everyone is in such a lucky position. LLMs help automate the tedium out of their job, and hopefully, have a bi
7.
▲
by
weiliddat
2mo ago
Yeah exactly. One more metric isn't going to solve it.
8.
▲
by
weiliddat
2mo ago
I don't necessarily mean reguritating it, but choosing which part of the scripture to surface to the user is already some interpretation/choice. Even the devil can quote scripture (I'm playing the devil's advocate here).
9.
▲
by
weiliddat
2mo ago
Yes it can point you to the right place, but choosing which part of scripture to point you to is already a choice/interpretation
10.
▲
by
weiliddat
2mo ago
Anecdote: was testing Opus 5 and even ASD-STE100 cannot save you from its meta-commentary and signposting all the time. It writes what it wants to write, e.g. "here's where you are right/wrong" instead of "Your inst
11.
▲
by
weiliddat
2mo ago
> They'd be 10x better than us already if task tedium was the problem... the areas where people think that LLMs excel at coding and don't like doing manually it's usually because the human was inclined to slop out repetiti
12.
▲
by
weiliddat
2mo ago
I get where you're coming from, and the intent to make it easier for people to find examples and verses, but there's a fine line with LLMs giving you answers, is that it's interpreting it in some form. Doesn't that run c
13.
▲
by
weiliddat
2mo ago
Kinda, more accurately the current goal for labs is to produce economic value to justify high valuations / capex, thus having to RL on correct answers rather than interesting ones. If you could access the current models as base models
14.
▲
by
weiliddat
2mo ago
idk if you're still actually looking for a solution (I got nerdsniped heh), but I found a seamless way to have it locked into the speakers + room, instead of a specific source/software, is using an iLoud subwoofer that routes to a
15.
▲
by
weiliddat
2mo ago
What you said made me think that it's similar (not entirely the same, but adjacent) to those make-a-website for $XX or buy-a-template era of websites. If we think about that as the intended competition (and not websites corps spent mil
16.
▲
by
weiliddat
2mo ago
I like this argument, but I don't think the article argues against taking urgency into consideration. Do you have anecdotes about companies or people that actually ignore urgency? I've experienced enough business-critical incident
17.
▲
by
weiliddat
2mo ago
Yeah it's a well-known secret, and I think any sort of scoring system misses the point. The scentific method (and system, incl. peer reviews) is a public good (like good governance) that requires moral acknowledgement, the necessary so
18.
▲
by
weiliddat
2mo ago
Hm since the discourse is mostly on how AI changes this, I still want to make a point that it's worth learning programming (and more), even if I'm a huge proponent of AI being economically valuable. Humans at the moment are still
19.
▲
by
weiliddat
2mo ago
Orthogonal to the movie/review, but if I could upvote this 100 times I would. These 3 questions were drilled to me in an elective Study of Visual Arts course in college and changed how I thought about everything and everyone.
20.
▲
by
weiliddat
2mo ago
I read it more as a lament than a real critique. Here's all the ways it was different than the source material. I think that was Nolan's intent all along - a very Nolan, American/Hollywood take on the Odyssey.
21.
▲
by
weiliddat
2mo ago
For most of these benchmarks, I feel like a p50 and p95 (using SWE salary as a proxy?) human benchmark as reference would be interesting. Edit: FWIW the paper the post quoted has repositories as slop baseline https://arxiv.org&#x
22.
▲
by
weiliddat
2mo ago
I couldn't access the post for some reason. Docs: https://octanejs.dev/
23.
▲
by
weiliddat
2mo ago
I read his followup tweet, and your comment, and I'm not fully convinced that open models are decelerationist. Happy to hear other thoughts on this. Open weight AI is decelerationist from the perspective that all capital should be allo
24.
▲
by
weiliddat
2mo ago
One interesting harness thing I saw Cursor do is to give the model access to the entire thread. Even if it doesn't fit in the context window, the model can search through past turns and sanity check if something doesn't seem to be
25.
▲
by
weiliddat
2mo ago
I was curious how have sentiments changed over time. Brief LLM-based analysis: https://ampcode.com/threads/T-019f32ac-3b1e-74be-ad63-5f175d... Overall, seems like it got more nuanced over time - even though it's s
26.
▲
by
weiliddat
3mo ago
> Often, it takes 5-6 broken crappy versions of a thing until you understand that. There is no accelerating the 5-6 broken crappy versions - there’s no agent tech that’s going to help your meat brain avoid thinking time. Fully agreed. Th
27.
▲
$10k bounty to break Pydantic's Python interpreter / sandbox
(hackmonty.com)
3 points
by
weiliddat
4mo ago
|
0 comments
28.
▲
by
weiliddat
4mo ago
No harm done, glad I clarified. I'm generally an optimistic person and very trusting of others. I'd say I'm also a pretty good reader of intentions / listener based on people who are my friends / worked with me (ane
29.
▲
by
weiliddat
4mo ago
Supported is different from doing it well though. You do notice the performance hit even on TVs that playback YouTube videos on AV1. Even on 1080p videos running on AV1 on 1x, the TV system bogs down and any kind of interaction has a variab
30.
▲
by
weiliddat
4mo ago
I'm not sure how to interpret your comment. It could be - a response to my comment saying that I am "illiterate" and cannot differentiate LLM output vs actual human comments (in that case I'm not sure what you're ad
More ›