Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
funfunfunction
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
by
funfunfunction
3mo ago
There's some benchmarks in the repo for AppWorld. Looks promising
2.
▲
by
funfunfunction
3mo ago
Cool project. A team at work was building something similar to internal use. I'm curious how this compares to just using Claude Code directly and giving it a dump of the agent traces? It seems like Claude could probably do some of the
3.
▲
by
funfunfunction
5mo ago
Hi all, we built a super easy way to train small language models from your production data - install a gateway to save your request/response data from a frontier provider, and use the traces to train a open source model that can be hos
4.
▲
Show HN: Project AELLA – Open LLMs for structuring 100M research papers
(aella.inference.net)
6 points
by
funfunfunction
10mo ago
|
2 comments
5.
▲
Hybrid-Attention models are the future for SLMs
(inference.net)
4 points
by
funfunfunction
11mo ago
|
0 comments
6.
▲
by
funfunfunction
11mo ago
We'll release the full data explorer soon, with more info. At the core of this project is a structured-extraction task using a custom Qwen 14B model, which we distilled from larger closed-source models. We needed a model we could run a
7.
▲
Show HN: Using LLMs and >1k 4090s to visualize 100k scientific research articles
(twitter.com)
5 points
by
funfunfunction
11mo ago
|
2 comments
8.
▲
by
funfunfunction
11mo ago
Creator of inference.net / schematron here. There is growing emphasis on efficiency as more companies adopt and scale with LLMs in their products. Developers might be fine paying GPT-5-Super-AGI-Thinking-Max prices to use the very be
9.
▲
Viral GPT wrappers are now training their own LLMs
(twitter.com)
8 points
by
funfunfunction
11mo ago
|
0 comments
10.
▲
by
funfunfunction
1y ago
OP here. I wanted a dead-simple way to quickly generate CLI commands without the overhead of Claude Code or Cursor, so I built it in an afternoon. The project uses some zsh magic to allow for quick editing of the model's response befor
11.
▲
UWU – generate CLI commands without leaving the terminal
(github.com)
16 points
by
funfunfunction
1y ago
|
2 comments
12.
▲
Show HN: UwU – Generate CLI commands inline with GPT-5
(github.com)
3 points
by
funfunfunction
1y ago
|
0 comments
13.
▲
How much energy does it take to produce an LLM token?
(energy.inference.net)
3 points
by
funfunfunction
1y ago
|
0 comments
14.
▲
by
funfunfunction
1y ago
This is a cheap marketing ploy for a GPU reseller with billboards on highway 101 into SF.
15.
▲
When to use model distillation in production
(inference.net)
1 points
by
funfunfunction
1y ago
|
0 comments
16.
▲
by
funfunfunction
1y ago
There are even companies starting to offer distillation as a service https://inference.net/explore/model-training
17.
▲
Show HN: Batch inference for large-scale synthetic data generation
(inference.net)
2 points
by
funfunfunction
2y ago
|
0 comments
18.
▲
Show HN: Costco for LLM Tokens
(inference.net)
6 points
by
funfunfunction
2y ago
|
0 comments
19.
▲
LLM Token Grants for Researchers
(inference.net)
2 points
by
funfunfunction
2y ago
|
0 comments
20.
▲
by
funfunfunction
2y ago
This is cool! I don’t see many people doing write ups on their tech stack as much any more. It’s nice to see the inside of a production-grade app like this. I’m curious, why command+r for the model? What benefits does it have over other SOT
21.
▲
by
funfunfunction
2y ago
It’s unlikely an individual would need this much capacity. Folks who need tokens at this level are apps with lots of users that don’t have their own GPUs. Think character.ai type apps. 10B tokens on together.ai is ~$2,000.
22.
▲
Inference.net: Wholesale LLM Tokens
(inference.net)
5 points
by
funfunfunction
2y ago
|
2 comments
23.
▲
by
funfunfunction
2y ago
Kuzco, Inc. ( https://kuzco.xyz ) | Full-stack SWEs and MLE | Full-time | San Francisco, CA We're building a serverless LLM inference network that makes use of underutilized capacity from GPU data centers. Our product is a sc
24.
▲
by
funfunfunction
2y ago
> generative AI is a product with no mass-market utility. > I am neither an engineer nor an economist. clearly.
25.
▲
by
funfunfunction
3y ago
Awesome project! I’m sure someone would be willing to buy this from you if the traffic is big enough. If you want to move on to other engineering projects this may be best path. If you’re interested in learning sales and marketing yourself
26.
▲
VC-backed AI startups are struggling. Indie devs are not
(twitter.com)
2 points
by
funfunfunction
3y ago
|
0 comments
27.
▲
Show HN: Chatbots for Technical Documentation
(usecontext.io)
2 points
by
funfunfunction
3y ago
|
1 comments
28.
▲
BabyAGI-ts: An NPM module to easily install, run, and play with BabyAGI locally
(github.com)
2 points
by
funfunfunction
3y ago
|
0 comments
29.
▲
by
funfunfunction
3y ago
People aren’t using ChatGPT because they can’t do it themselves, they’re using it to save time.
30.
▲
by
funfunfunction
3y ago
Hey if your project is public I would love to take a look if you don’t mind sharing a link
More ›