Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
markab21
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
markab21
2mo ago
We use it at a moderate scale, self-hosted on B300 hardware. It's great :D QA analysis of voice transcriptions. Napkin math: we operate at 2-5% of the cost of running on Equiv Frontier, though this changes near-weekly because pricing i
2.
▲
by
markab21
2mo ago
Yes - THIS! I can't even believe how exhausting it is to read. I'm not sure why or what changed in Fable. Did they do this writing-style output to give it more token compression during/for training or to prefer output for les
3.
▲
by
markab21
3mo ago
I'm mildly surprised that more people aren't using Nemo models for this reason. We've moved most of our processing to a combination of Nemo Ultra and Super, with some support for multi-model-specific tasks on Omni. The setup
4.
▲
by
markab21
6mo ago
I'll pipe in here as someone working on an agentic harness project using mastra as the harness. Nemotron3-super is, without question, my favorite model now for my agentic use cases. The closest model I would compare it to, in vibe and
5.
▲
by
markab21
6mo ago
I think the entire premise that the prompting is the surface area for optimizing the application is fundamentally the wrong framing, in the same way that in 1998 better cpam will save CGI. It's solving the wrong problems now, and the l
6.
▲
by
markab21
7mo ago
You just articulated why I struggle to personally connect with Gemini. It feels so unrelatable and exhausting to read its output. I prefer to read Opus/Deepseek/GLM over Gemini, Qwen and the open source GPT models. Maybe it is RLH
7.
▲
by
markab21
7mo ago
And I think you basically just described the OpenAI approach to building models and serving them.
8.
▲
by
markab21
7mo ago
Shaking fist at clouds!!
9.
▲
by
markab21
8mo ago
It's getting a lot easier to do this using sub-agents with tools in Claude. I have a fleet of Mastra agents (TypeScript). I use those agents inside my project as CLI tools to do repetitive tasks that gobble tokens such as scanning code
10.
▲
by
markab21
8mo ago
Basically looking for emergent behavior.
11.
▲
by
markab21
9mo ago
I love where you're going with this. In my experience it's not about a different persona, it's about constantly considering context that triggers, different activations enhance a different outcome. You can achieve the same th
12.
▲
by
markab21
1y ago
The skepticism is understandable given the trajectory of GPTs and custom instructions, but there's a meaningful technical difference here: the Apps SDK is built on the Model Context Protocol (MCP), which is an open specification rather
13.
▲
by
markab21
2y ago
It looks like it was built jointly with nvidia: https://huggingface.co/nvidia/Mistral-NeMo-12B-Instruct
14.
▲
by
markab21
3y ago
I bet you have been waiting years to pull that one out of your pocket. Well played sir! Nice shot man! :D
15.
▲
by
markab21
3y ago
I've found myself more and more using local models rather than ChatGPT; it was pretty trivial to set up Ollama+Ollama-WebUI, which is shockingly good. I'm so tired of arguing with ChatGPT (or what was Bard) to even get simple thin
16.
▲
by
markab21
3y ago
For Llama-based progress - Reddit - /r/LocalLlama has been my top source of info, although it's been getting a little more noisy lately. I also hang out on a few Discord servers: - Nous Research - TogetherAI / Firewor
17.
▲
by
markab21
3y ago
[flagged]
18.
▲
by
markab21
3y ago
Yeah, slow news day.
19.
▲
by
markab21
3y ago
Linux distributions, including Debian, offer a variety of desktop environments, each with its own design philosophy and user experience. If one environment doesn't suit your preferences, others might be more to your liking. It's w
20.
▲
by
markab21
3y ago
Debian and Ubuntu have similarities, but keep in mind - Ubuntu is derived from Debian, not the other way around. However, they differ in areas like release cycles, package management, and default configurations.
21.
▲
by
markab21
4y ago
I used it as a consultant on a development project to help me organize some of the milestones and design goals in some documentation. It wasn't that I didn't know the stuff, I do, but more helpful with quickly organizing and prese
22.
▲
by
markab21
5y ago
I'd be surprised if they don't have async mechanisms.
23.
▲
by
markab21
5y ago
Assume anything sent over a cellular network carrier via normal SMS can not only be retrieved, but intercepted.
24.
▲
by
markab21
6y ago
Why anyone would use Oracle for anything other than supporting legacy systems is beyond me.
25.
▲
by
markab21
7y ago
The claim is dead-right. I own a Tesla Model 3, my wife drives a BMW i3, my daughter has a leaf. The ONLY car we can effectively travel outside of the greater Tampa area without major headache is the Tesla. The ONLY car that I would try to
26.
▲
by
markab21
7y ago
Care to site/link these two studies?
27.
▲
by
markab21
7y ago
As someone who regularly pilots a piston-single, I'm looking forward to what hybrid or full electric can do for us little guys. The takeoff phase of flight is a high-risk situation in events such as fuel contamination which could prov
28.
▲
by
markab21
7y ago
> I'd like to get one but it doesn't fit my use case (namely long road trips in the summer). I own a P3D, facing the same issues of long road trips 1-2x a year I concluded that I'll just spend the few hundred dollars and r
29.
▲
by
markab21
8y ago
I've noticed my dog (black lab) doing similar behavior consistently in the mornings, also after she had her breakfast. She'll walk away from the kitchen and her dog bowl with her tail wagging and doing some kind of sneezing sniff
30.
▲
by
markab21
8y ago
Interesting. Do you have any thoughts on composite-based aircraft like a Cirrus and flutter? A few times I've had a Cirrus SR22 into a pretty steep descent with poor controller sequencing for an approach into busy terminal space and h
More ›