Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
Hfuffzehn
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
Hfuffzehn
2mo ago
While that might might be true nowadays, Anthropic has started to work very hard on fixing that. They just started with it not helping with software security.
2.
▲
by
Hfuffzehn
2mo ago
Agreed. Demoralization in e.g. Germany has reached insane levels. This is a problem of the direction things are going. Things are getting worse and that is causing the demoralization, not the absolute level of "how good things still ar
3.
▲
by
Hfuffzehn
3mo ago
I never felt more violated than when a coworker 15 years younger than me started to touch my work laptop screen. I mean, she was right, it turned out to be a touch screen, but really who does something like that?
4.
▲
by
Hfuffzehn
3mo ago
I think you are misapprehending the ruling. I someone would put the libellous statement into Google Translate in English and the LLM spits out the German translation would that be a crime from Google? My point is exactly that the court has
5.
▲
by
Hfuffzehn
3mo ago
The argument starts with search engine -> search terms -> search command Google accepted the argument up to that point. Then they argued that they are not responsible for the search result. The court argued that what they did was not
6.
▲
by
Hfuffzehn
3mo ago
https://the-decoder.com/wp-content/uploads/2026/06/26_O_869_... "Tatbestand Die Verfügungsklägerinnen begehren von der Verfügungsbeklagten die Unterlassung von Darstellungen KI-generierter Antworten
7.
▲
by
Hfuffzehn
3mo ago
Sure that could be the worst case. But then we are in the case law vs. continental law debate for lawyers. When I wrote "Gemini is not illegal" it would have been more correct to say that the court has not decided about Gemini but
8.
▲
by
Hfuffzehn
3mo ago
Yes, the monopoly is not relevant for the court. It is relevant for Google though, because they want to transfer it to another product. And the court is saying that whatever that new product is, Google is not allowed to mislead the public b
9.
▲
by
Hfuffzehn
3mo ago
If I get it correctly I like the ruling. So Google has established a product called Search. For that product rules have been established. Google has monopolized that product. Now Google is replacing that product with a new product. But they
10.
▲
by
Hfuffzehn
4mo ago
And as I am on holiday today I will try to help them out: GPT-5.4 mini Haiku 4.5 MAI-Code SWE-Bench Pro 54.4 % 35.2% 51.2% Terminal-Bench 2.0 60.0 % 41.6% 54.8% Source: https:/&#x
11.
▲
by
Hfuffzehn
4mo ago
So I guess the important link the marketing department forgot is this one: https://docs.github.com/en/copilot/reference/copilot-billing... Model Input Cached input Output MAI-Code-1-Flash $0.75 $0.075 $4.50 C
12.
▲
by
Hfuffzehn
4mo ago
This is a very good comment. But notice how even in software engineering there is still disagreement about these structural safeguards. So yes, we can say the LLM created bad code when it does not compile or fails prewritten tests. But expe
13.
▲
by
Hfuffzehn
4mo ago
I agree. But notice that you assume that there is a metric with which you can messure improvement. Which is fine if you are measuring against your personal taste. But it might be that the optimization target itself has a ceiling. If you
14.
▲
by
Hfuffzehn
4mo ago
https://docs.github.com/en/copilot/reference/copilot-billing... Model Input Cached input Output MAI-Code-1-Flash $0.75 $0.075 $4.50
15.
▲
by
Hfuffzehn
4mo ago
The first time I was impressed by AI coding was when I pointed it at some switch case monster code and told it to replace it with a strategy pattern. And it did just fine. So no matter what you think about vibe coding, using AI for these sl
16.
▲
by
Hfuffzehn
4mo ago
Agreed. Seems like this could have been a nice model if we would still be in the old GitHub Copilot free request/ premium multiplier mode. It could have been a good compromise to somehow reign in the costs for Microsoft. But with Copil
17.
▲
by
Hfuffzehn
4mo ago
That's really nice of them. That means Jensen can add another 30 times faster when comparing Rubin to Blackwell without having to actually do anything. Hopefully that means he won't have any problem to make another 150 billion in
18.
▲
by
Hfuffzehn
4mo ago
The interesting thing for me is that I do not feel like the writing of LLMs has improved very much lately stylistically. They have reached a "good" level some time ago but the newer models havn't brought such improvements tha
19.
▲
by
Hfuffzehn
4mo ago
The main insight here I think is that LLMs are great tools for iterative development and iterative problem solving in general. You can very effectivly iterate alone using the LLM as a mirror, rephrasing what you put in and adding a bit. You
20.
▲
by
Hfuffzehn
4mo ago
I haven't really deeply thought about frontend JS for many years. Back then the question we were looking at was whether it would be good idea to move away from SAP UI5. The alternatives back then where React, Angular and Vue. The concl
21.
▲
by
Hfuffzehn
4mo ago
This is really tickling the conspiracy theorist part of my brain. "Independent open-source project · not affiliated with DeepSeek" "Reasonix only targets DeepSeek because..." "Why DeepSeek only? Can I swap to Claude
22.
▲
by
Hfuffzehn
4mo ago
With DeepSeek making their price rebates permanent we now have some data what China values data access at. Western providers of the open weight models are 3 times or more as expensive as DeepSeek itself right now. Of course the data access
23.
▲
by
Hfuffzehn
4mo ago
I miss him too. Even though I had the experiences he discribes with Douglas Adams first before discovering Terry Pratchett.
24.
▲
by
Hfuffzehn
4mo ago
You have correctly identified that getting a "high-quality harness (ie preloaded instructions from md files, including custom skills)" is the (or at least a) hard part. Because you have to adjust the harness to your problem space
25.
▲
by
Hfuffzehn
4mo ago
Isn't blaming AI for that similar to blaming C for buffer overflows? More people are producing more code because of easier tools. Most code is bad. But that's not the tools fault. And in the end it is a problem of processes and cu
26.
▲
by
Hfuffzehn
5mo ago
Yes, but that market is not b2b, less commercialized, more end consumer focused and more bring your own key. That's why I find it interesting. Anthropic is not interested in building a moat there and OpenAI has given up on their announ
27.
▲
by
Hfuffzehn
5mo ago
Sure, but the best statistics about what models people are actually using when they can choose is probably from openrouter: https://openrouter.ai/apps/category/entertainment/roleplay
28.
▲
by
Hfuffzehn
5mo ago
From what I can gather Grok is not used for roleplay much. It is considered to inconsistant and crazy. People are mostly using GLM and Deepseek via API and Gemma4 and Mistral finetunes locally. It seems to me like the roleplay market is com
29.
▲
by
Hfuffzehn
5mo ago
Yes, I agree. I deliberatly formulated that channeling myself as the kid who actually found his drumming valuable but didn't have the money to buy (all) of it. Who was annoyed at society deciding I should not have it. So I still don&#x
30.
▲
by
Hfuffzehn
5mo ago
And at that moment societies might actually have to think deeply about the value copyright provides. Because having access to the condensed knowledge of humanity might be more valuable for society then having access to Lars Ulrich's sh
More ›