Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
brainless
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
brainless
6d ago
Yes they are great for learning at own pace, trying out new things. I have accepted two things that make me a happy engineer now: AGI is not here no matter what they say and LLMs are still very useful if one knows how to use them. They are
2.
▲
by
brainless
7d ago
More and more such experiments. I felt sad for a couple months when I realized that writing code will not be the same since. Now I am on the other side. LLMs are interesting in their own ways but as an engineer, this is a way to unlock a ne
3.
▲
by
brainless
7d ago
When we hyper focus on finding something, we find it all the time. The world of tiny LLMs is so interesting. It is unlocking novel ways to encode information. Why focus on the style of writing instead of the subject matter?
4.
▲
by
brainless
8d ago
This is a tiny LLM doing all the heavy-lifting. Any mention of the training process? I am obsessed with tiny LLMs and the do-one-thing-really-well approach that they seem to fit very well.
5.
▲
by
brainless
15d ago
I experiment a lot with local LLMs, particularly small ones like Qwen3.5 4B and 9B. I have build multiple experiments to make harnesses that use these models for code generation, planning, local search, etc. These are really good models but
6.
▲
by
brainless
17d ago
Somewhere around early 2000, I used to work part-time with a local video editing agency. Matrox video capture cards were the really high bar, other than Avid. At least that is what I remember. We had a couple of Matrox capture cards, expens
7.
▲
by
brainless
20d ago
I do not think Cloudflare was a less-than-peers optimized product when they launched. This is one of their blog posts which describes taking one aspect even further. I think Cloudflare became big only because they were so much more optimize
8.
▲
by
brainless
20d ago
I use Claude Code, Codex and opencode pretty much interchangeably. I am currently using Claude more this month because (stupidly) I paid for Max ($100) since I have a large client project. I generally use larger models to plan. All my gener
9.
▲
by
brainless
21d ago
I can understand that this is not the exact location where the floods actually happened. But I am sad and I am frustrated and I am angry. We had floods in Sikkim, which is practically a "stone throw" away from a world-level map po
10.
▲
by
brainless
21d ago
I am on the same boat, I live in a small Himalayan village. I have 2x5G based devices and one Wireless bridge (Ubiquiti LiteBeam M5) for a local broadband. Generally I get 50 Mbps, sometimes up to 100 Mbps
11.
▲
by
brainless
21d ago
HF is the default platform to find models. Not just ones published by big labs but also a lot of the distilled or fine-tuned versions, etc. If Nvidia buying HF makes it tough for all the diverse models on HF, then what are some alternatives
12.
▲
by
brainless
23d ago
I already do this. I live in a small village where there isn't even a wired Internet connection (wireless only). I work full-time with LLMs, on own product ideas and client projects (all LLM led). I started investing in farms, have 50
13.
▲
by
brainless
23d ago
Why is it important to train LLMs or even fine-tune them? LLMs have proven their point, costs are crashing, and there are more of them than most companies need. The real value is to unlock meaningful insights and directions from existing da
14.
▲
by
brainless
24d ago
I do not know where things will go but here are some things I have been feeling: - We have tons of languages (and libraries/frameworks on top) because we, humans, have too many preferences - Many of these preferences are also abstracti
15.
▲
How a Knowledge Graph Made Haiku as Accurate as Fable 5 [video]
(youtube.com)
3 points
by
brainless
1mo ago
|
0 comments
16.
▲
by
brainless
1mo ago
This is how my experiments go. And I am sure there are popular agents that do this. How I am trying is to create "Rust Engineer", "Typescript Engineer" or even "Rust Diesel Engineer". I have not tried fine-tuni
17.
▲
by
brainless
1mo ago
I am sorry I did not understand all of it. But, would this allow running large MoE LLMs on a local network with experts spread out over multiple cheaper GPUs (or even CPUs)? This would perhaps be more useful than over the Internet, within o
18.
▲
by
brainless
1mo ago
Good to see more harnesses coming out. I think the initial set of "features" that made into harnesses like tool calling, multi-turn chat, MCP, skills and so on can all be optimized. And then much more can be done on top. I am tryi
19.
▲
by
brainless
1mo ago
This is awesome. I will take some time to dig in. When I am not working for client(s), I focus entirely on tiny LLMs - I have specific approach to prompting, avoid multi-turn chat and build harness to fit the selected LLM as closely as poss
20.
▲
by
brainless
1mo ago
Any plans to support smaller models? I have a M4 Mac Mini with 16GB unified memory and an RTX 3060 (Laptop) with 6GB VRAM. My own product experiments all revolve around small models and harness around them. Happy to contribute.
21.
▲
by
brainless
1mo ago
I have a weird thought experiment: If you give a GPT-2/3 level LLM tools to search the internet - any document, can it build bigger, better LLMs? You may think this is not a good test because an older (or say a smaller) LLM can study f
22.
▲
by
brainless
1mo ago
Would it not be better to ask models to search the topic on the Internet and then answer? I do not understand why we expect small LLMs to answer from own knowledge.
23.
▲
by
brainless
2mo ago
I want small models to win and I am constantly experimenting with them. I have never tried fine-tuning and do not have that kind of budget. My approach is to remove some of the burden from models and bring into the agent. Tool calling is an
24.
▲
by
brainless
2mo ago
I do not know what/who Sybil is but yes, there is a lot of plumbing in order to make sure that host nodes are ZDR* compliant and much more. Zero visibility of source prompt. Also, this is the reason I want to start with trust based gro
25.
▲
by
brainless
2mo ago
I have been thinking of something on these lines but with much smaller models. The entire model has to fit on a single computer. Host owner would choose the model they prefer, perhaps because they already use it. Then it is more about util
26.
▲
by
brainless
2mo ago
This is good and I am saying this as someone using coding agents full on. I am a software engineer and I do use coding agents and I do believe that there can be spaces which do not encourage or allow projects built by generative AI. Detecti
27.
▲
by
brainless
2mo ago
It is all binary, all the way down. Code is text that passes the compiler's checks. Human language is text that has a really ambiguous compiler. And all text is still binary in the computer. If you think about it, the transformers arch
28.
▲
by
brainless
2mo ago
Yes, this is how I am building my agent as well. A chain of mostly deterministic steps for every incoming prompt. Run as many tools without help of LLMs, gather errors - feed to LLMs, then go back to deterministic steps as soon as possible.
29.
▲
by
brainless
2mo ago
Yes, I agree with this. I am not focusing on tests as much and I think that is a big mistake. Agents need to immediately understand something is off.
30.
▲
by
brainless
2mo ago
Not necessarily. I have been building client projects for the last few months only using coding agents. I use way more existing tools to handle a lot more of our digital footprint than I used to before: pdf, images, excel, ocr and many more
More ›