Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
ismaeel_bashir
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
8 ms
·
1.
▲
by
ismaeel_bashir
4mo ago
Thanks, the security point is valid, so let me be specific about how deployment works for us! There's no telemetry egress. Deployments are air-gapped and run in the customer's VPC, on their own hardware. We don't ship telemet
2.
▲
by
ismaeel_bashir
4mo ago
Yes :) This is actually a really cool feature of the platform. We ingest DCGM, CUPTI, and cgroups to give users granular telemetry of what exactly is going on in the hardware they allocated when running jobs on it. We also have profiler tha
3.
▲
by
ismaeel_bashir
4mo ago
Sure would be happy to :) I’ll send you an email, good luck with the book!
4.
▲
by
ismaeel_bashir
4mo ago
Nope :) the core model isn’t an LLM. It’s a custom architecture built from the ground up. We natively accept multimodal inputs such as source code, submission scripts and hardware topologies. The LLMs in the post are the baselines we beat.
5.
▲
by
ismaeel_bashir
4mo ago
Good point - people do set capacity aside, reserving it for later. But our utilisation measurements are from waste within a users allocation. It’s waste of what users are actually requesting and running, not from any reserved idle capacity.
6.
▲
Launch HN: Expanse (YC P26) – Unlock Wasted GPU Capacity
103 points
by
ismaeel_bashir
4mo ago
|
27 comments