Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
skeptrune
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
The Inference Engineering Skills Map
(skeptrune.com)
1 points
by
skeptrune
6d ago
|
0 comments
2.
▲
We self-host DeepSeek V4 Flash on AWS spot instances
(twitter.com)
2 points
by
skeptrune
2mo ago
|
0 comments
3.
▲
by
skeptrune
2mo ago
based
4.
▲
by
skeptrune
3mo ago
the more awesome thing to me is that you can run the MRI through an ensemble of LLMs and check to see if they converge among each other
5.
▲
by
skeptrune
3mo ago
Well put
6.
▲
by
skeptrune
3mo ago
This is fun
7.
▲
Declank – Remove AI Watermarks from Images
(declank.skeptrune.com)
3 points
by
skeptrune
3mo ago
|
0 comments
8.
▲
Show HN: Declank – Remove AI Watermarks from Images
(declank.skeptrune.com)
3 points
by
skeptrune
3mo ago
|
0 comments
9.
▲
Tokenmaxxing: One AI budget, four jobs
(mintlify.com)
2 points
by
skeptrune
4mo ago
|
0 comments
10.
▲
by
skeptrune
4mo ago
I really appreciate how "Jony Ive" this looks. Feels like they absolutely nailed the style. I personally feel like it looks like a disposable tech hardware product, but to each their own. I'm sure a lot of people will love it
11.
▲
by
skeptrune
4mo ago
I really thought the photos were real as i was reading. Wow
12.
▲
Audit Logs Wall of Shame
(audit-logs.tax)
2 points
by
skeptrune
5mo ago
|
0 comments
13.
▲
by
skeptrune
5mo ago
What a win it is for open source that qwen and kimi show up on this at all.
14.
▲
by
skeptrune
5mo ago
lmao, i love this
15.
▲
by
skeptrune
5mo ago
this is awesome. beyond happy to see it
16.
▲
by
skeptrune
6mo ago
Working on publishing those, but publishing benchmarks requires a lot of attention to detail so it will likely be a bit longer.
17.
▲
by
skeptrune
6mo ago
agreed. hopefully we can get there soon
18.
▲
by
skeptrune
6mo ago
Yea we did and actually use Daytona for another product, but it would have been too slow here.
19.
▲
by
skeptrune
6mo ago
yea chromadb is not the point. multiple data storage solutions work
20.
▲
by
skeptrune
6mo ago
agreed!
21.
▲
by
skeptrune
6mo ago
We would also be super interested to see that comparison. I agree that there isn't a specific reason why Chroma would be required to build something like this.
22.
▲
by
skeptrune
6mo ago
I agree that would have been the way to go given more time and resources. However, setting up a FUSE mount would have taken significantly longer and required additional infrastructure.
23.
▲
by
skeptrune
6mo ago
100% agree. However, if there were no resource tradeoffs, then a FUSE mount would probably be the way to go.
24.
▲
by
skeptrune
6mo ago
Modern OCR tooling is quite good. If the knowledge you are adding into your search database is able to be OCR'd then I think the approach we took here is able to be generalized.
25.
▲
by
skeptrune
6mo ago
Hmmm, the post is an attempt to explain that Mintlify migrated from embedding-retrieval->reranker->LLM to an agent loop with access to call POSIX tools as it desires. Perhaps we didn't provide enough detail?
26.
▲
by
skeptrune
6mo ago
Vector search has moved from a "complete solution" to just one tool among many which you should likely provide to an agent.
27.
▲
by
skeptrune
6mo ago
I think it's cool that LLMs can effectively do this kind of categorization on the fly at relatively large scale. When you give the LLM tools beyond just "search", it really is effectively cheating.
28.
▲
by
skeptrune
6mo ago
100% agree a FUSE mount would be the way to go given more time and resources. Putting Chroma behind a FUSE adapter was my initial thought when I was implementing this but it was way too slow. I think we would also need to optimize grep even
29.
▲
by
skeptrune
6mo ago
This is awesome! I'm happy someone made this exist.
30.
▲
The future of text layout is not CSS
(chenglou.me)
16 points
by
skeptrune
6mo ago
|
19 comments
More ›