Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
tempusalaria
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
tempusalaria
11mo ago
Most of EA’s revenue comes from franchise games that are way below typical AAA standard. EA’s value is from IP not talent
2.
▲
by
tempusalaria
11mo ago
Lots of situations, here are 2 I’ve faced recently (cannot give too much detail for privacy reasons, but should be clear enough) 1) low latency desired, long user prompt 2) function runs many parallel requests, but is not fired with common
3.
▲
by
tempusalaria
11mo ago
All these things are designed to create lock in for companies. They don’t really fundamentally add to the functionality of LLMs. Devs should focus on working directly with model generate apis and not using all the decoration.
4.
▲
by
tempusalaria
11mo ago
I vastly prefer the manual caching. There are several aspects of automatic caching that are suboptimal, with only moderately less developer burden. I don’t use Anthropic much but I wish the others had manual cache options
5.
▲
by
tempusalaria
1y ago
A lot of the current code and science capabilities do not come from NTP training. Indeed in seems in most language model RL there is not even process supervision, so a long way from NTP
6.
▲
by
tempusalaria
1y ago
Cerebras has very limited scale. Mistral has very few users so they can use cerebra’s in inference whereas OpenAI and Anthropic cannot. If mistral grows a lot they will stop using cerebras
7.
▲
by
tempusalaria
1y ago
Fast tire changes only matter a very limited amount of the time (pretty much only if the extra time drops you a place, so there has to be 1 car/20 in a specific 1 second window on what is typically a 90s lap for 3s (a slow stop) vs 2s
8.
▲
by
tempusalaria
1y ago
I imagine it runs civ 2 pretty well
9.
▲
by
tempusalaria
1y ago
WhatsApp is certainly worth less today than what they paid for it plus the extra funding it has required over time. Let alone producing anything close to ROI. Has lost them more money than the metaverse stuff. Insta was a huge hit for sure
10.
▲
by
tempusalaria
1y ago
SFT is part of the classic RLHF process though
11.
▲
by
tempusalaria
1y ago
Yes this write-up is not about agents. In fact it’s a great illustration of why the hype around agents is misplaced!
12.
▲
by
tempusalaria
1y ago
I understand that calling it ‘agentic’ is nice for marketing, but most of what is described in this blog post is not related to agents. The design patterns you describe are explicitly non-agentic. Many of the use cases described are better
13.
▲
by
tempusalaria
1y ago
They may not be acting in good faith but there is extremely clear evidence that UCLA has engaged in illegal racial hiring and admissions practices and has supported antisemitism on campus. UCLA chose to give them that ammunition.
14.
▲
by
tempusalaria
1y ago
if you are p testing this isn’t the case. A positive result is a much stronger assertion
15.
▲
by
tempusalaria
1y ago
The term agent is just way overloaded. This guy defines it completely differently the the big labs, and I’ve seen half a dozen different definitions in the last few months. In the long run the definition used by OpenAI, Anthropic et al will
16.
▲
by
tempusalaria
1y ago
Even as someone who is skeptical about LLMs, I’m not sure how anyone can look at what was achieved in AlphaGo and not at least consider the possibility that NNs could be superhuman in basically every domain at some point
17.
▲
by
tempusalaria
1y ago
I agree I find claude easily the best model, at least for programming which is the only thing I use LLMs for
18.
▲
by
tempusalaria
2y ago
SemiAnalysis has made up many things. They claim that a small Chinese hedge fund could acquire $1bln in GPUs, with no state support, including many sanctioned chips, then trained a model optimized for a far smaller server compute size, and
19.
▲
by
tempusalaria
2y ago
SemiAnalysis is wrong. They just made their numbers up (among many other things they have invented - they are not to be trusted). I have observed many errors of understanding, analysis and calculation in their writing. Deep Seek R1 is liter
20.
▲
by
tempusalaria
2y ago
DeepSeek v3 (where the training cost claims come from) was announced a month ago and it had no impact outside of a small circle
21.
▲
by
tempusalaria
2y ago
Texas is a world leader in renewable energy. Easy permitting, lots of space, lots of existing grid infrastructure from the o&g industry.
22.
▲
by
tempusalaria
2y ago
1) DPO did exclude some practical aspects of the RLHF method, e.g. pretraining gradients. 2) the theoretical arguments of DPO equivalence make some assumptions that don’t necessarily apply in practice 3) RLHF gives you a reusable reward mod
23.
▲
by
tempusalaria
2y ago
The reality is that these are not culturally significant institutions and most people in London don’t care. Ordinary Londoners rarely use these markets, and they mostly sell to restaurants, and require major financial support. Not everythin
24.
▲
by
tempusalaria
2y ago
many of these labs have more funding in theory than OpenAI. FAIR, GDM, Qwen all are subsidiaries of companies with $10s of billions in annual profits.
25.
▲
by
tempusalaria
2y ago
Airbus was a company setup by consolidating companies controlled by some of the most powerful countries in the world, which sold planes to captive state airlines and militaries controlled by those same governments and their allies. What an
26.
▲
by
tempusalaria
2y ago
This is very similar to how LLMs are taught to understand images in llava style models (the image embeddings are encoded into the existing language token stream)
27.
▲
by
tempusalaria
2y ago
Only if it is reliably correct. Google does offer an AI summary for factual searches and I ignore it as it often hallucinates. Perplexity has the same problem. OpenAI would need to solve that for this to be truly useful
28.
▲
by
tempusalaria
2y ago
Definitely they will. OpenAI’s potential issue is that if Google offers tokens at a 10% gross margin, OpenAI won’t be able to offer api tokens at a positive gross margin at all. Their only chance really is building a big subscription busine
29.
▲
by
tempusalaria
2y ago
I also like 3.5 sonnet as the best model (best ui too) and it’s the one I ask questions to We use Gemini flash in prod. The latency and cost is just unbeatable - our product uses llms for lots of simple tasks so we don’t need a frontier mod
30.
▲
by
tempusalaria
2y ago
It’s not clear this is true because reported numbers don’t disaggregate paid subscription revenue (certainly massively GP positive) vs free usage (certainly negative) vs API revenue (probably GP negative). Most of their revenue is the subsc
More ›