Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
m3h
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
m3h
21d ago
I'm having a tough time using Opus to produce blog-grade writing. Claude produces jargon heavy, market-y language that no human would ever write, much less want to read. This issue exists even when I ask it to write internal reports or
2.
▲
Ask HN: Opus 5 is unusable for writing, even internal use. Alternatives/fix?
9 points
by
m3h
22d ago
|
3 comments
3.
▲
by
m3h
26d ago
You're looking for Poteto's /unslop which cleans up sloppy AI writing. Someone made a Claude version of her skills: https://github.com/michael-denyer/pstack-claude
4.
▲
Kitesurf: Agent-first browser that runs in V8 isolates
(blog.cloudflare.com)
221 points
by
m3h
1mo ago
|
63 comments
5.
▲
by
m3h
2mo ago
Someone who until yesterday did not seem bothered by his technology being possibly used to bomb elementary girls school in another country seems to suddenly care about the repression of citizens in yet another country. No, we don't buy
6.
▲
by
m3h
2mo ago
Is there a specific list of changes they made to the system prompt? They're claiming they removed 80% of it. That's quite substantial. It would be good to know what the model knows to do by training and what we need to avoid over-
7.
▲
ChatGPT and Codex Weekly Users Cross 10M
(twitter.com)
3 points
by
m3h
2mo ago
|
0 comments
8.
▲
by
m3h
2mo ago
A while back, I saw a similar feature land in Codex (I'm using the VS Code plugin) but it got removed quickly. What is the chance that an LLM recommended this same idea to the Claude PM or lead? I see LLMs across different providers co
9.
▲
by
m3h
2mo ago
> Kimi K3 is Kimi’s most capable model to date, with 2.8 trillion parameters. This puts them on the top of the largest open models list: Kimi K3 2.8T DeepSeek-V4-Pro 1.6T (49B active) Kimi K2.6 ~1T (32B act
10.
▲
by
m3h
2mo ago
The author shared their experience building the first version in a month: https://themackabu.dev/blog/js-in-one-month And then the follow up few months later: https://themackabu.dev/blog/ant-part-t
11.
▲
by
m3h
2mo ago
We have an official pelican on a bicycle from the OpenAI livestream: https://imgshare.cc/mz9xwut3
12.
▲
by
m3h
2mo ago
The speed up numbers based on their testing: Codebase | TypeScript 6 | TypeScript 7 | Speedup ------------|--------------|--------------|-------- vscode | 125.7s | 10.6s | 11.9x sentry | 139.8s
13.
▲
by
m3h
2mo ago
When I reviewed the conversations affected by this issue, they did not always align with my feeling of "degraded output". Some were definitely below par, and I recall having to iterate on the generated code more than I wanted to.
14.
▲
by
m3h
2mo ago
Indeed, it looks like my work has suffered from the clustering issue as well: reasoning_output_tokens count percent ━━━━━━━━━━━━━━━━━━━━━━━━━ ━━━━━━━ ━━━━━━━━━ 0 873 28.5948 ─────────────────
15.
▲
by
m3h
3mo ago
Also, kudos to the Z.ai team for adding Linux support from day one.
16.
▲
by
m3h
3mo ago
Z.ai documents integrations with nearly all the popular CLI-based agents: https://docs.z.ai/devpack/tool/others If you're already used to your TUI coding agent, you don't need the desktop agent. Although
17.
▲
by
m3h
3mo ago
Correct. Albeit the nuance here is that a more capable model might solve problems more efficiently and faster, possibly saving you tokens. As with any new model, you won't know the real impact until you start using it for your workload
18.
▲
by
m3h
3mo ago
I didn't realize GPT 5.3 Codex was that good. OpenAI claims to have made their new Terra model as good as GPT 5.5, but with half the cost per intelligence. Hopefully, this will bring it closer to the price you're expecting (or eve
19.
▲
by
m3h
3mo ago
I think you should try an OpenAI model like GPT 5.5. It is better at following instructions and boundaries set during prompt. It feels like a more capable "agent assistant" than Claude models but without loss of intelligence. Most
20.
▲
GPT 5.5 uses Grug Brained talk during reasoning for 2x token efficiency
(youtube.com)
4 points
by
m3h
3mo ago
|
0 comments
21.
▲
by
m3h
3mo ago
Important to note: "Sonnet 5 is an upgrade to Sonnet 4.6, but it uses an updated tokenizer that changes how the model processes text to improve performance (this is similar to the tokenizer change we introduced with Claude Opus 4.7). T
22.
▲
by
m3h
3mo ago
Why is Claude Sonnet 5 allowed to be released but OpenAI Terra not? Are they not the same class of models?
23.
▲
Vibe Coding to Agentic Engineering: A Three-Phase Workflow with Claude Code
(apimatic.io)
2 points
by
m3h
3mo ago
|
0 comments
24.
▲
by
m3h
3mo ago
If GPT-5.6 preview is not available outside US government approved "trusted partners", I don't see how the General Available can be trusted later. Who knows what they will fix, block or change in the model between the preview
25.
▲
AI SDK 7 is available
(vercel.com)
2 points
by
m3h
3mo ago
|
0 comments
26.
▲
US Government Asks OpenAI to Stagger AI Model Release
(bloomberg.com)
2 points
by
m3h
3mo ago
|
1 comments
27.
▲
Vibe Coding to Agentic Engineering with Claude Code
(apimatic.io)
1 points
by
m3h
3mo ago
|
0 comments
28.
▲
by
m3h
3mo ago
Or we could simply hallucinate that the packages are there at the three houses. Hallucinations all the way down...
29.
▲
Codex usage grows after Fable nerf model release
(twitter.com)
2 points
by
m3h
3mo ago
|
0 comments
30.
▲
Replace your CI with a merge queue
(blog.exe.dev)
4 points
by
m3h
3mo ago
|
0 comments
More ›