Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
artursapek
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
Pair programming: still a good idea
(revise.io)
5 points
by
artursapek
1d ago
|
0 comments
2.
▲
by
artursapek
5d ago
I only use OpenRouter to benchmark models. For my production app, I have direct integrations with OpenAI, Anthropic, Google, xAI. Even just input caching is enough of a reason to do that, assuming you are trying to build the fastest and mos
3.
▲
Claude Public Artifacts
(google.com)
2 points
by
artursapek
16d ago
|
0 comments
4.
▲
Working on documents with your AI chat
(revise.io)
3 points
by
artursapek
1mo ago
|
0 comments
5.
▲
by
artursapek
1mo ago
Yes, it uses block IDs - documents are made of blocks like paragraphs and lists. Formatting survives because the agent works in HTML.
6.
▲
by
artursapek
1mo ago
Hi HN. I've been bootstrapping this project full-time for the last 12 months. Would love to get some feedback on the MCP integration! I think it's some of the best UX available for working on documents, with AI.
7.
▲
Show HN: A free DOCX editor with MCP server for editing
(revise.io)
14 points
by
artursapek
1mo ago
|
3 comments
8.
▲
by
artursapek
1mo ago
These prices are not real. They already said so.
9.
▲
by
artursapek
2mo ago
It’s an attempt to build a locomotion model. It uses a physics engine (Avian in Bevy) to animate a walking biped from first principles. An earlier version can be seen here https://x.com/tmuxvim/status/2081822899149
10.
▲
by
artursapek
2mo ago
WASD or just touch on mobile
11.
▲
Gait Game
(gait.game)
2 points
by
artursapek
2mo ago
|
4 comments
12.
▲
by
artursapek
2mo ago
It's definitely not cheaper than Sonnet on my benchmark, but it's cheaper than Fable and outperforms it. Which is big IMO. https://revise.io/errata-bench
13.
▲
ErrataBench
(revise.io)
1 points
by
artursapek
2mo ago
|
0 comments
14.
▲
by
artursapek
2mo ago
HN users are world champions are trivializing difficult things with snarky comments
15.
▲
Kimi K3 first open model in ErrataBench Top 10
(revise.io)
3 points
by
artursapek
2mo ago
|
0 comments
16.
▲
by
artursapek
2mo ago
Yep, I've been taking glycine and magnesium for years. I am not as consistent as I should be but it makes a big difference when I use them.
17.
▲
GPT 5.6 sets new record on proofreading benchmark
(twitter.com)
1 points
by
artursapek
2mo ago
|
0 comments
18.
▲
Mantissa, a distributed workload orchestration system
(mantissa.io)
4 points
by
artursapek
2mo ago
|
0 comments
19.
▲
by
artursapek
3mo ago
I run a proofreading benchmark that tests how well models can find and fix errors in English text. They get several passes in a simple agent loop. Sonnet 5 is definitely better than Sonnet 4.6, but inferior on both quality and cost to GLM 5
20.
▲
by
artursapek
3mo ago
haha yeah I've bet the last 12 months of my career on a .io
21.
▲
by
artursapek
3mo ago
The .ai TLD is some tiny island with a few thousand people
22.
▲
by
artursapek
3mo ago
Trivial to simulate basic keystrokes. But I don't think it's trivial to simulate the natural process of drafting something. There's no concrete heuristic or algorithm (yet) for judging these types of replays, but I'd be
23.
▲
"No, I swear I wrote this."
(revise.io)
17 points
by
artursapek
3mo ago
|
46 comments
24.
▲
by
artursapek
3mo ago
They claim extreme performance on ExploitBench, which Mythos was touted as being incredible at. https://x.com/OpenAI/status/2070555278576439306
25.
▲
Which LLM is the best proofreader?
(revise.io)
1 points
by
artursapek
3mo ago
|
0 comments
26.
▲
by
artursapek
3mo ago
Fable 5 beats GPT 5.5 in my proofreading benchmark. And it does so at approximately the same total cost; it used significantly fewer turns than 5.5 https://x.com/tmuxvim/status/2064452096800198930
27.
▲
by
artursapek
3mo ago
I would expect Apple to hedge their bet on Gemini and build everything so that the model can be swapped out in the future.
28.
▲
by
artursapek
3mo ago
I use Carplay all the time and I didn't even realize it has voice control. I just set things up on my phone and drive.
29.
▲
by
artursapek
3mo ago
I think it's fair to say that OpenAI has at least partially won the "consumer AI" segment.
30.
▲
by
artursapek
3mo ago
You’re not responding to anything the parent said.
More ›