Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
bitexploder
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
bitexploder
3d ago
I was kinda kidding because parent called it "blazing" fast lol. I didn't think it was Rust. Blazing fast software is reserved as a descriptor for Rust programs :)
2.
▲
by
bitexploder
3d ago
It is entirely reasonable to run a model like DS Flash 4.1 locally. Not easily. But feasible. If you drop down 100b - 200b models they are far more feasible. DS Flash 4.1 is close to Opus 5. Beats Opus 4.8 in all of my little evals. You can
3.
▲
by
bitexploder
3d ago
You just have to Manhattan Project it. Fable will happily work on almost any piece in isolation. The hardest boundaries are domain specific terms you just can’t avoid. It can be tricky. I have found a way around it mostly. Or just use DeepS
4.
▲
by
bitexploder
3d ago
It is Rust now?
5.
▲
by
bitexploder
4d ago
That makes sense. I think I have approached becoming stuck. I reached a kind of extreme disassociated state once where I felt like I didn't exist and part of my brain started to panic. I meditate a lot. I reminded the part of me that w
6.
▲
by
bitexploder
5d ago
We have pretty strong indications of when and why things go wrong. Openness vs rigid thinking are an axis that are very strong correlative factors. A lot more nuance but there is a lot we do know about risk factors so it isn’t a complete qu
7.
▲
by
bitexploder
5d ago
“Inadvertently”.
8.
▲
by
bitexploder
5d ago
Someone was probably sad when they saw a compiler work for the first time after years of writing assembly. AI is intellectually different, progressing, and has no visible horizon, but in my life technology never did in computing. The point
9.
▲
by
bitexploder
6d ago
Hah, np, stego in general is really cool :)
10.
▲
by
bitexploder
6d ago
That is a better idea. Ingesting your corpus with a lot of traces that have semantic patterns. Semantic steganography that suffixes well to real math and science (and any) topics. <thinking> heh.
11.
▲
by
bitexploder
6d ago
Problem is how do you convince the model and training profess it matters. A one off canary is very unlikely to survive in the final model state.
12.
▲
by
bitexploder
7d ago
I am surprised at how well DSv4 flash does in the real world vs many benchmarks. You look at Flash 3.8 and it supposedly beats opus 5 and deepseek is far below.. but they were measuring efficiency, whatever that is… Something doesn’t add up
13.
▲
by
bitexploder
7d ago
I don't usually defend Apple products, but my AirPods Pro 2nd gen have been rock solid and taken insane abuse. I use them working on cars, runs, hiking, rain, shine, sauna. I drop them. They land in puddles. They smack concrete every c
14.
▲
by
bitexploder
7d ago
Thanks... I have been trying to figure out some things. Been doing my own evals. Flash 3.8 does burn a lot more tokens on high. Interesting how smart and not smart it is. For personal use almost impossible to justify the cost of 3.8 Flash c
15.
▲
by
bitexploder
7d ago
If you are okay with waiting use GLM 5.3 max. It costs more but still cheap. It is slow, but a very strong worker. Still dollars per day (at most) with heavy concurrent agent running. I load up planning and tasks in Opus or Sol, and just ha
16.
▲
by
bitexploder
7d ago
Which versions of flash and at what thinking levels? Which chinese flash models and at what thinking levels? What tasks? What completion rates? How was quality evaluated?
17.
▲
by
bitexploder
9d ago
Yep. It really isn’t nefarious or abusive. I keep mine slowed down to not scan often to be considerate. Even private trackers can require it now.
18.
▲
by
bitexploder
9d ago
For what it’s worth, I am very happy with Jellyfin and the *arr suite. It took a bit of agent prodding to get them all playing nicely together and bypassing Cloudflare CAPTCHAs. However, it's pretty sweet when you get it all working.
19.
▲
by
bitexploder
10d ago
It does seem like for new code that might help. There's some really good logic and wisdom in it, but it has to be applied very contextually to the exact problem you are trying to solve. If an agent is navigating a complex codebase, thi
20.
▲
by
bitexploder
10d ago
Sol is my current favorite model to interact with. So much less BS than Opus 5. Fable 5.1 is okay as is Fable 5 but it has Opus like tendencies. Sol is very good at following instructions and remembering them for a session.
21.
▲
by
bitexploder
11d ago
Amusingly, as an autonomous coding agent, I kind of like Opus 5. But I have to bound it on tasks or it just goes off the rails. But I'm bounded tasks, it is genuinely solid. It's kind of like the new Sonnet 5. Right now my favorit
22.
▲
by
bitexploder
12d ago
I still hold the line on interviewing. Maybe some don’t. I am sure it is true. My team us too small with too much responsibility to tolerate mediocrity to any real extent
23.
▲
by
bitexploder
12d ago
In modern America the answer to that question is often resoundingly yes. Not just hypothetical.
24.
▲
by
bitexploder
12d ago
I believe. I run it on my mac M5 pro at like 30t/s with some RAGs and let it work on stuff overnight and it's great. It isn't the same as the big models where things can be more unbounded, but if local models keep progressing
25.
▲
by
bitexploder
12d ago
I have a few attention and finish mechanisms in my prompts. I have been using it for a week and a half and with some prompt taming it is great. (I have early access to the models cause I work at the place that makes the model). None of my a
26.
▲
by
bitexploder
12d ago
Fable is okay, just slower, eats tokens and not any better at coding tasks. Maybe a little better, but not better enough. It's a lot faster to have a cheap and fast flash agent / sonnet do the implementation work with Fable taggin
27.
▲
by
bitexploder
13d ago
Opus 5 is a genuinely infuriating model. I hate it’s behavior.
28.
▲
by
bitexploder
13d ago
I feel like a lot happened this week and people are glazing how ridiculously strong Flash 3.8 is right now compared to Fable/Opus/Sol/Astra.
29.
▲
by
bitexploder
13d ago
I wish people could see how some of this reads. You are an “amateur” using a model 6-8 weeks behind? Really? Sigh.
30.
▲
by
bitexploder
13d ago
Does Claude not allow third party harnesses?
More ›