Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
yolo-auto
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
OpenClaw/Hermes Article on agent memory options
(yolo-auto.com)
2 points
by
yolo-auto
2mo ago
|
0 comments
2.
▲
by
yolo-auto
2mo ago
lmao dont hurt yourself
3.
▲
by
yolo-auto
2mo ago
Hi danger dingo, Sure, sign up and ping me in discord
4.
▲
by
yolo-auto
2mo ago
Well, are you a robot? Ok well I need a way to securely give you an API key for you to try it out but I don't really want to post it on HN so I'm open to ideas lol
5.
▲
by
yolo-auto
2mo ago
1x MI300x , rocm/sglang , highly optimized image, FP8, FP16 Cache (this fixes looping issues with qwen3.6-35b ~100tok/s average
6.
▲
by
yolo-auto
2mo ago
actually the math is more like $1500 a month on a MI300x on a highly optimized rocm/sglang image and everyone's getting 100+ tok/sec and we have 100 active people and plenty of room for others. but were losing money right now
7.
▲
by
yolo-auto
2mo ago
i know, but im scared to own our own auth, when there's money involved... too cheap to pay for a service. So we lean on the back of the giants for now. If it helps, there's really no reason to ever log into our site after you get
8.
▲
Show HN: An unmetered LLM API–$6/month, no token tracking, no limits
(yolo-auto.com)
12 points
by
yolo-auto
2mo ago
|
13 comments
9.
▲
by
yolo-auto
3mo ago
So on mi 300x we run FP8 and that was original plan. We are doing some weird q4 on 3090s now that is surprisingly good (for qwen) . will people pay for it? Yeah, sometimes. Some dude just burned 80mil tokens in 24 hr and posted it to our di
10.
▲
Story of How Im Running an Unlimited $6/Month AI Provider on 4x RTX 3090s
8 points
by
yolo-auto
3mo ago
|
3 comments