Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
iamnotagenius
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
by
iamnotagenius
1y ago
православный is used in jargon for exactly that meaning. source: 45 years of native Russian speaking.
2.
▲
by
iamnotagenius
1y ago
They are different. Gemma 3 12b excels at natural languages but terrible at long context. Pixtral 12b is better at long context (not stellar), but worse at natural language.
3.
▲
by
iamnotagenius
1y ago
You can slightly improve output of non-thinking model if you add ad the end of prompt "output chain of though reasoning before outputting the result".
4.
▲
by
iamnotagenius
1y ago
We are told that writing must be pure, that it must come only from the sweat of the brow, the trembling hand, the solitary mind. That to use AI is to cheat, to dilute, to lessen the act of creation. But I say to you: Since when does it matt
5.
▲
by
iamnotagenius
1y ago
Whenever possible use local LLMs. You do need claude for everything.
6.
▲
by
iamnotagenius
1y ago
they all degrade well before 1M tokens.
7.
▲
by
iamnotagenius
1y ago
glm 4 (1 sec): To determine the cutoff frequency (fc ) for an RC circuit (since you've provided resistance R and capacitance C, but not inductance L), we can use the following formula: [.... calculation] So, the cutoff frequency is app
8.
▲
by
iamnotagenius
1y ago
Not quite true. Depends on number of KV heads. GLM4 32b at IQ4 quant and Q8 context can run full context with only 20GiB VRAM.
9.
▲
by
iamnotagenius
1y ago
Try to push your point to absurd you see why; hint - to analyze data pulled by tools you need knowledge already baked in. You have very limited context, you cannot just pull and pull data.
10.
▲
by
iamnotagenius
1y ago
I have stopped using cloud models half a year ago. The meager intelligence local models, ones I can run on my machine, is already giving me a great deal of productivity boost.
11.
▲
by
iamnotagenius
1y ago
> She thinks she's in some kind of Truman show that she calls "the game". Might be depersonalization. I had suffered from it in my twenties; everything feels fake, although you know it is not.
12.
▲
by
iamnotagenius
1y ago
Tiny, 4b or less models are designed for finetuning for some narrow tasks; this way can outperform large commercial models for a tiny fraction of price. Also great for code autocomplete. 7b-8b are great coding assistants if all you need is
13.
▲
by
iamnotagenius
1y ago
Fine-tuning is excellent way to reliably bake-in domain specific data into a model; there is a plenty of coding finetunes on Huggingface face, that outperforms foundation models on say coding, without significant loss in other domains.
14.
▲
by
iamnotagenius
1y ago
With all due respect to Deepseek, I would take their numbers with grain of salt, as they might as well be politically motivated.
15.
▲
by
iamnotagenius
1y ago
My 3060 idles sometimes at 19 watt, only sleep and wakeup of the machine helps.
16.
▲
by
iamnotagenius
1y ago
4060ti has abysmal bandwidth 288 Gb/sec which is a no go for llms.
17.
▲
by
iamnotagenius
1y ago
Mistral models though are not interesting as models. Context handling is weak, language is dry, coding mediocre; not sure why would anyone chose it over Chinese (Qwen, GLM, Deepseek) or American models (Gemma, Command A, Llama).
18.
▲
by
iamnotagenius
1y ago
Interesting, but not exactly practical for a local LLM user, as 4-bit is how LLM's are run locally.
19.
▲
by
iamnotagenius
2y ago
Well because you explicitly ask it to demonstrate the physics, it came out way too detailed, but point is that it adds details on its own to scenes, make more realistic, not that dry LLama 3.3 style.
20.
▲
by
iamnotagenius
2y ago
Here is the sentence : (She screamed, which echoed off the tile walls. “This is my life now,” she said to her reflection, which looked back at her with a mix of disgust and pity.) Looks good to me. try it on Lmarena.ai.
21.
▲
by
iamnotagenius
2y ago
though the dividends were not obvious to the lay people vs now. Which means that upcoming winter won't be as cold.
22.
▲
by
iamnotagenius
2y ago
Priuses were already popular then; the writing on the wall was already visible.
23.
▲
by
iamnotagenius
2y ago
I liked Grok 3 fiction writing style; catches lots of physics of mundane situations such as ringing echo in a closed bathroom we all know well; the prose feels very lively as the result. Kinda like R1 makes situations sharp with details, Gr
24.
▲
by
iamnotagenius
2y ago
Not quite true; LLMs are very expensive to run; bert or other tranformer specfically built for translation can be cheaper to run.
25.
▲
by
iamnotagenius
2y ago
precisely; however this time we will have tangible results from the ongoing AI summer; that would be generative art, and coding/writing/journalist assistants.
26.
▲
by
iamnotagenius
2y ago
[flagged]
27.
▲
by
iamnotagenius
2y ago
Fringe means that it attracts weirdos, demagogues etc, as things normally fringe for a reason, but even if they are not, people have so much reputational risk for publicly advertising these views, so that either suggests they are grifters,
28.
▲
by
iamnotagenius
2y ago
> I was expecting a lot more from (I think) the first politician following the Austrian school of economics. This is exactly why you've heard only demagoguery; Austrian school is fringe.
29.
▲
by
iamnotagenius
2y ago
[flagged]
30.
▲
by
iamnotagenius
2y ago
[flagged]
More ›