Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
samosx
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
Infinity embedding engine added to KubeAI
1 points
by
samosx
2y ago
|
0 comments
2.
▲
Lightweight ML model proxy and autoscaler for Kubernetes
(github.com)
2 points
by
samosx
3y ago
|
0 comments
3.
▲
Estimating required GPU memory for serving LLMs
(substratus.ai)
2 points
by
samosx
3y ago
|
2 comments
4.
▲
by
samosx
3y ago
Having a hard time with estimating how much GPU memory that LLM needs to serve it? What kind of GPUs to use and how many? Wrote a blog post to demystify the process of GPU memory usage estimating.
5.
▲
by
samosx
3y ago
Thanks for the feedback! All resources are namespaced right now, except there is an issue the notebook plugin where namespaces are indeed broken: https://github.com/substratusai/substratus/issues/193 Will get
6.
▲
Deploying Weaviate on GKE end-to-end tutorial
(samos-it.com)
2 points
by
samosx
4y ago
|
0 comments
7.
▲
by
samosx
4y ago
Very simple prototype that uses stackoverflow API to get questions and chatGPT to answer them and store the results in postgres or sqlite using Peewee ORM.
8.
▲
ChatGPT-blog: Stackoverflow banning ChatGPT? I got a workaround for that
(github.com)
1 points
by
samosx
4y ago
|
1 comments
9.
▲
by
samosx
4y ago
OP here. You're right the solution is based on the existing documented full text search solutions. You can technically only sent the data you need indexed and the ID instead of sending all data. There is no native mechanism and firebas
10.
▲
by
samosx
5y ago
Feedback welcome, what feature would you want so this tool is helpful for you?
11.
▲
Open Source Free Website Speed Monitoring – Websu
(websu.io)
4 points
by
samosx
5y ago
|
1 comments