Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
xlayn
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
xlayn
6d ago
This is not a pro ai comment, I hope some bios bug bricks all the nvidia rtx gpus in the planet. With that said... ai generated music, the same can be said about software... but compilers are basically ai, there is no way a for loop on extr
2.
▲
by
xlayn
6d ago
Systems evolve around pros and cons, just like plants grow and bend and flex to capture the most sun we can put laws to say we fine 25% of their anual income and the next time the goverment takes becomes majoritary shareholder, and you will
3.
▲
by
xlayn
7d ago
unleashes unparallel, unrivaled, totally new, amazing, exciting new ways of interacting with the MOST POW3RFUL!!!!! then all the photos either what's on the inside it's repeated whats on the outside, or... and hold your breath...
4.
▲
by
xlayn
7d ago
The more advanced the effort to detect human vs AI the more advanced the AI will be to write just like human. I wonder if at some point it will become totally human like, and then what will be the reaction? will you accept it? will it be tu
5.
▲
by
xlayn
8d ago
Jevons Paradox doesn't say more radiologist will have more work and will be paid at the same price , it says that there will be more work for radiologists because now they are paid less and because of that people is able to afford mor
6.
▲
by
xlayn
9d ago
No they are not, try to change jobs and if you can't or takes a really long time you know this is some of the finest baloney. And no, it's not creating a lot of jobs, one one of the kind that will be destroyed by AI no, and that&#
7.
▲
by
xlayn
15d ago
Hey carloslfu, kudos from the other side of the internet, don't get down on people nitpicking everything here, experimenting and discovering is part of learning so keep going!, remember this is the place that said dropbox was dumb and
8.
▲
by
xlayn
17d ago
I pay a $100 subscription which feels infinite for me, even using only fable... 2 weeks ago I started getting notifications of running out of credit... to me it feels more like 30% than 17%... OHHHH and surprise, I'm in my 50% boosted&
9.
▲
by
xlayn
21d ago
For the impatient, I merged llama.cpp tentative branches to get it running here https://github.com/alainnothere/llama.cpp/tree/disk-cache-ev... , thing runs at 23.54 token/sec and my setup runs at high 30
10.
▲
by
xlayn
23d ago
This is the second post in the vibe of "you should not get angry at work, being angry is bad, and you want to be a professional bla bla bla", and what I see is some effort to start putting the blame into the worker of working cond
11.
▲
by
xlayn
25d ago
In animatrix there is this part of the video where the devised path forward is "the destruction of the sky" and you can see all this millitary and business people clapping... then the image turns to all that people as skeletons cl
12.
▲
by
xlayn
28d ago
Daniel, question I got the Qwen3.8-27B-UD-Q2_K_XL.gguf from https://huggingface.co/unsloth/Qwen3.8-27B-GGUF?show_file_in... and continue with my testing, but the model quickly felt into a loop of asking the same thing
13.
▲
by
xlayn
28d ago
I do use 2 amd gpus and I get high 40 for generation, 500 for pp and low 20/100 by the end of the context of 256k. llama-server --host 0.0.0.0 --port 8089 -m Qwen3.8-27B-UD-Q8_u.gguf --spec-type draft-mtp,ngram-mod --spec-draft-n-max 3
14.
▲
by
xlayn
28d ago
my bad, you are totally right, thanks!
15.
▲
by
xlayn
28d ago
Hey Unsloth, your gguf are the first ones I look for when I want to download a gguf model. Today I was trying in fact to see, what's the smallest Qwen3.8-27B that I could run and get good results, say restricting it to 16GB of ram.. so
16.
▲
by
xlayn
29d ago
In other news, automotive car company after spending billions on nvidia machinery discover that small company overseas deliver 90% of the puch for 0% of the price; anounces that will stop internal combustion engine development because it&#x
17.
▲
by
xlayn
29d ago
Remember, books ARE SO important and transmit so much information that they are used and destroyed to feed the machine so others cannot use them, let that sink in, then you will see the issue with libraries.
18.
▲
Qwen3.8-27B make medium the default effort level instead of xhigh
(github.com)
16 points
by
xlayn
29d ago
|
4 comments
19.
▲
by
xlayn
29d ago
Because it's so nice to see page after page after page of... but what if... let's consider... If you use a gguf, you can extract the template from the Qwen3.827B model using the following script https://github.com/
20.
▲
by
xlayn
29d ago
This is the correct answer, but let me add a corolarium, both github, whatever cursor launches and any other bigco does obeys to the same friking iron law: *MAKE MONEY/POWER AT ANY COST, MONEY/POWER IS THE OBJECTIVE*, so you will
21.
▲
Privibe – LLM CLI Local first+privacy focus+llama.cpp cache branch and Qwen3.x
(github.com)
1 points
by
xlayn
1mo ago
|
1 comments
22.
▲
by
xlayn
1mo ago
I've been working on this clone of mistral vibe, the idea is to have a cli that you can trust is not calling back home for any reason, I have done as much work as possible to remove any trace of phone back anywhere, hence the pri prefi
23.
▲
by
xlayn
1mo ago
this is my understanding, the default template keeps the thinking part but only for the last message, so the harness has to play along with the template and strip and add to keep the conversation matching what's there on the llama.cpp
24.
▲
by
xlayn
1mo ago
I have this branch of llama.cpp that among other things (like patching the template to not break the kv cache, and saving conversations to disk so you can resume quickly days after) also accept the reasoning effort flag here https:/&#
25.
▲
Privibe - LLM Cli Local first+privacy focus+llama.cpp cache branch + Qwen3.x
(github.com)
2 points
by
xlayn
1mo ago
|
1 comments
26.
▲
by
xlayn
1mo ago
I've been working on this clone of mistral vibe, the idea is to have a cli that you can trust is not calling back home for any reason, I have done as much work as possible to remove any trace of phone back anywhere, hence the pri prefi
27.
▲
by
xlayn
1mo ago
The unsloth Q8kxl https://huggingface.co/unsloth/Qwen3.8-27B-GGUF for some reason is looping and going crazy on the think part (I tried to search for an email to let the guys know but didn't find one)... I used t
28.
▲
by
xlayn
1mo ago
The file "Just loads" on llama.cpp, the Unsloth https://huggingface.co/unsloth/Qwen3.8-27B-GGUF is an MTP file, I see mostly the same speed on pp and generation. There has to be something wrong with those ben
29.
▲
by
xlayn
1mo ago
Absolutely agree, and yes AI is an slop-enabler tool.
30.
▲
by
xlayn
1mo ago
So if you take any of the options above, and as part of the instructions mention to be more concise, push to github, and then the reader is lazy and don't verify, is it slop? Hey option d, please create a website to sell stuff similar
More ›