Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
voxgen
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
19 ms
·
1.
▲
by
voxgen
2mo ago
Fully agree. I can drink coffee until 6pm and soft drinks until 9pm without impact to my sleep. There's little point in me moderating my caffeine intake. Light, sleep time regularity, and exercise make a much more noticeable difference
2.
▲
by
voxgen
2mo ago
Even the official provider on OpenRouter seems to have this issue. Hope it's an easy fix for them.
3.
▲
by
voxgen
2mo ago
Pay attention to what it forgets, and start telling it to proactively note down those things into nominated files while it works, e.g. indexing topics covered by each chapter/section/page of each paper/material. GPT-5.6 is re
4.
▲
by
voxgen
2mo ago
It's a solvable problem if you're willing to throw more tokens at it. Frontier models have gotten very good at cleaning up their own messes. You just need the right skills/loops, and to stick to models that consistently follo
5.
▲
by
voxgen
2mo ago
There is a third case where the other party doesn't realize that the asker lacks the relevant experience to discern good LLM answers from bad answers for that topic. Same solution as case one though - don't be afraid to say "
6.
▲
by
voxgen
3mo ago
Denying tech export to cooperative allies is certainly a move. Many nations are now likely thinking: Why cooperate on international IP enforcement if we get lumped together with adversarial nations anyway?
7.
▲
by
voxgen
4mo ago
As a current free-tier user, this price & transition plan seems completely reasonable. I'll upgrade when the nudge comes. I've got at least that much utility from it so far. Very glad this isn't a subscription.
8.
▲
by
voxgen
5mo ago
I have found Claude to be especially unpredictable. I've mostly switched to GPT-5.4 now - although it's slightly less capable, it's massively more reliable.
9.
▲
by
voxgen
5mo ago
Some of the flak is that issues are often only acknowledged once a fix is in place, and the partial fixes are presented as if they solve the whole problem. The near-instant transition from "there is no problem" to "we already
10.
▲
by
voxgen
5mo ago
Thank you for the reminder. Childhood, at least IME, does a bad job of preparing people for this: fault and blame rarely matter in real interactions. When no teacher is around to play judge, all that matters is that you get a favorable outc
11.
▲
by
voxgen
6mo ago
It works for me Firefox's Cloudflare DNS over HTTP. For clarity, the recent issue[0] likely wasn't intermittent. Cloudflare's malware blocking DNS server now blocks those archive.today sites. Doesn't affect the non-malwa
12.
▲
by
voxgen
7mo ago
This could be an explanation for the drama - LLMs are trained to learn and emulate correlations in text. I'm sure you already have a caricature in mind of the kinds of online posts (and thus LLM training data) that include miscitations
13.
▲
by
voxgen
8mo ago
It's not perfect but it does have a few opt-in security features: running all tools in a docker container with minimal mounts, requiring approvals for exec commands, specifying tools on an agent by agent basis so that the web agent can
14.
▲
by
voxgen
8mo ago
I'm working in AI, but I'd have made this anyway: Molty is my language learning accountability buddy. It crawls the web with a sandboxed subagent to find me interesting stuff to read in French and Japanese. It makes Anki flashcard
15.
▲
by
voxgen
11mo ago
Ratio/quantity is important, but quality is even more so. In recent LLMs, filtered internet text is at the low end of the quality spectrum. The higher end is curated scientific papers, synthetic and rephrased text, RLHF conversations,
16.
▲
by
voxgen
11mo ago
It requires tax increases, and the average earner's UBI will typically balance out the tax increase, meaning they don't directly profit. UBI isn't about giving everyone free money. It's about giving everyone a safety net
17.
▲
by
voxgen
1y ago
That discussion also makes me worry that they may try to use LLMs or LLM-based metrics to measure the size of the gap as a proxy for value of the content. The landlord of the marketplace should probably not dabble in the appraisal of produc
18.
▲
by
voxgen
1y ago
> without punishing regular browsing humans. As a content consumer, I'm also hoping to be part of the ecosystem. I already use Patreon a lot as "AdBlock absolution", but it doesn't fix the market dynamics. Major conte
19.
▲
by
voxgen
1y ago
What makes you think the secrets are small enough to fit inside people's heads, and aren't like a huge codebase of data scraping and filtering pipelines, or a DB of manual labels?
20.
▲
by
voxgen
1y ago
Please consider also describing the business model on the website, even if hidden away on a FAQ. I've so much subscription fatigue now, I just don't try things out if needing a subscription is an inevitability. I'm happy to p
21.
▲
by
voxgen
1y ago
I don't think retrofitting existing languages/ecosystems is necessarily a lost cause. Static enforcement requires rewrites, but runtime enforcement gets you most of the benefit at a much lower cost. As long as all library code is
22.
▲
by
voxgen
1y ago
The last major innovation as a product was PWA support starting in 2016. Browsers used to try new ideas like RSS, widgets, shared and social browser sessions. Interfaces to facilitate low-friction integration with the rest of your life, and
23.
▲
by
voxgen
1y ago
> It's interesting that there are no reasoning models yet This may be merely a naming distinction, leaving the name open for a future release based on their recent research such as coconut[1]. They did RL post-training, and when fed
24.
▲
by
voxgen
1y ago
> Or is Behemoth just going through post-training that takes longer than post-training the distilled versions? This is the likely main explanation. RL fine-tuning repeatedly switches between inference to generate and score responses, and
25.
▲
by
voxgen
2y ago
My thoughts go out to the poor engineers who got put on call because someone scheduled a product release on the day before the biggest holiday of their year.
26.
▲
by
voxgen
2y ago
It's not even "nearly as good as o1". They only compared to the older 4o. You can safely assume Qwen2.5-Max will score worse than all of the recent reasoning models (o1, DeepSeek-R1, Gemini 2.0 Flash Thinking). It'll pro
27.
▲
by
voxgen
2y ago
Vegetarian keto is certainly possible, but vegan would be very tough. Only 2 out of 6 of my regular meals[1] have meat in them, and I'd probably replace these with tofu and mushrooms if I could tolerate them. There's a world of ke
28.
▲
by
voxgen
2y ago
I'm at 3 years with occasional breaks. At a certain point my weight wouldn't go lower and I started feeling terrible. I think I was producing more ketones than I could use. I'm not sure exactly what fixed it, but now I'm
29.
▲
by
voxgen
2y ago
Don't give up! Induction gets easier every time, and you learn lots of tricks/recipes, like keto-ade to feel better during induction, and making oats/flaxmeal tasty for cheap & quick breakfasts. You don't have to com
30.
▲
by
voxgen
2y ago
> it's not clear if that was the author's actual intention The paper[1] doesn't appear to have any other connections to the book/response/memes. A clear distinction is that the UB paper very directly and prominen
More ›