Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
sethkim
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
by
sethkim
20d ago
I appreciate the classic HN sarcasm!
2.
▲
The Analytical AI Handbook
(handbook.sutro.sh)
49 points
by
sethkim
20d ago
|
2 comments
3.
▲
Useful Black Boxes
(sethkim.me)
3 points
by
sethkim
5mo ago
|
0 comments
4.
▲
by
sethkim
6mo ago
This is extremely true. In fact, from what we see many/most of the problems to be solved with LLMs do not have ground-truth values; even hand-labeled data tends to be mostly subjective.
5.
▲
by
sethkim
6mo ago
Feel free to shoot me a note at seth@sutro.sh if you want to check it out!
6.
▲
by
sethkim
6mo ago
We build a product that's somewhat similar in spirit to DSPy, but people come to us for different reasons than the OP listed here. 1) It's slow: you first have to get acquainted with DSPY and then get hand-labeled data for prompt
7.
▲
by
sethkim
11mo ago
Under-discussed superpower of LLMs is open-set labeling, which I sort of consider to be inverse classification. Instead of using a static set of pre-determined labels, you're using the LLM to find the semantic clusters within a corpus
8.
▲
by
sethkim
1y ago
The models you called out at the beginning were all released this year. What do you think is the difference between this generation of models and previous ones?
9.
▲
by
sethkim
1y ago
Yes! Both Llama 3 and Gemma 3 have 128k context windows.
10.
▲
by
sethkim
1y ago
Yes, we're a startup! And LLM inference is a major component of what we do - more importantly, we're working on making these models accessible as analytical processing tools, so we have a strong focus on making them cost-effective
11.
▲
by
sethkim
1y ago
My two cents here is the classic answer - it depends. If you need general "reasoning" capabilities, I see this being a strong possibility. If you need specific, factual information baked into the weights themselves, you'll ne
12.
▲
by
sethkim
1y ago
No doubt prices will continue to drop! We just don't think it will be anything like the orders-of-magnitude YoY improvements we're used to seeing. Consequently, developers shouldn't expect the cost of building and scaling AI
13.
▲
by
sethkim
1y ago
Both great points, but more or less speak to the same root cause - customer usage patterns are becoming more of a driver for pricing than underlying technology improvements. If so, we likely have hit a "soft" floor for now on pric
14.
▲
The End of Moore's Law for AI? Gemini Flash Offers a Warning
(sutro.sh)
113 points
by
sethkim
1y ago
|
75 comments
15.
▲
by
sethkim
1y ago
I run a batch inference/LLM data processing service and we do a lot of work around cost and performance profiling of (open-weight) models. One odd disconnect that still exists in LLM pricing is the fact that providers charge linearly w
16.
▲
by
sethkim
1y ago
Sutro.sh (fka Skysight) | Infrastructure/LLMs & Research Engineering | SF Bay Area | Full-time We are building batch inference infrastructure and a great/user developer experience around it. We believe LLMs have not yet been m
17.
▲
by
sethkim
1y ago
Skysight | Infrastructure/LLMs & Research Engineering | SF Bay Area | Full-time We are building large-scale batch inference infrastructure and a great/user developer experience around it. We believe LLMs have not yet been mean
18.
▲
by
sethkim
1y ago
How "huge" are these datasets? Did you build your own tooling to accomplish this?
19.
▲
by
sethkim
1y ago
Thanks for the reply and the notes. On 4. specifically we've got some thoughts here as well. Will reach out!
20.
▲
Classifying aviation-related posts on Hacker News with SLMs
(skysight.inc)
10 points
by
sethkim
1y ago
|
2 comments
21.
▲
Generating 1M Synthetic Humans
(skysight.inc)
5 points
by
sethkim
1y ago
|
0 comments
22.
▲
Non-Scalar Leverage
(sethkim.me)
1 points
by
sethkim
1y ago
|
0 comments
23.
▲
Model Security with Large-Scale Inference
(skysight.inc)
3 points
by
sethkim
2y ago
|
0 comments
24.
▲
by
sethkim
2y ago
Skysight | Infrastructure/LLMs & Product Engineering | SF Bay Area | Full-time We are building large-scale, data-intensive inference tooling and a great/user developer experience around it. We believe LLMs have not yet been me
25.
▲
by
sethkim
2y ago
What's extremely confusing to me (as a private pilot) is that traffic is almost always routed directly over an airport (midfield), to safely avoid departing and landing traffic. The sense that I get is that it became routine for traffi
26.
▲
by
sethkim
2y ago
This is really cool, and hints at a near-future possibility of building a search engine on top of just about anything. It's clear we've moved past the ability to just search for website url's and webpage content. Anything tha
27.
▲
by
sethkim
2y ago
I figured this comment would get me in trouble :) I recommend doing some instrument lessons if you haven't already. When I got my instrument rating I questioned whether the private requirements are actually enough. The skills that the
28.
▲
by
sethkim
2y ago
In single pilot IFR, an autopilot is often your best friend. It's exactly like you say - when you're busy with everything else you want the plane to fly itself. Isn't that problem already somewhat solved in a sense? Or are yo
29.
▲
by
sethkim
2y ago
> What if this tech made the individual flyer safer? I'd hope that's the case! That's why I put it in the "good" category". > how much could be in the air at one time? Hard to say, but there's a ton
30.
▲
by
sethkim
2y ago
Instrument-rated pilot (and engineer) here. First - congrats on the launch! I think you're working on an interesting set of components that will prove useful to GA aircraft technology. Bringing fly-by-wire, and lowering the cost of mai
More ›