Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
Mkengin
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
Mkengin
12d ago
Google is already doing that since last year and the models where never leaked, so I don't think that would even be a problem. https://cloud.google.com/blog/topics/hybrid-cloud/gemini-is-...
2.
▲
by
Mkengin
1mo ago
Are you also talking about personal projects? That wouldn't be enough for me for work either, but for personal use, my $20 Codex subscription is perfectly fine. Then again, as a new dad, I can maybe only work on personal projects for 1
3.
▲
by
Mkengin
4mo ago
There Seen to be more and more harness benchmarks out there, pretty interesting read: https://neuralnoise.com/2026/harness-bench-wip/
4.
▲
by
Mkengin
6mo ago
I don't think so, I use GrapheneOS and I think I can't even use the USB-C port for anything other than charging (which should be configurable).
5.
▲
by
Mkengin
9mo ago
No, I am not affiliated with the website, I just want to see more discussions based on uncontaminated benchmarks and feel that people rely too much on benchmarks that companies can conduct themselves. If that is the case, I don't feel
6.
▲
by
Mkengin
9mo ago
I would assume that if a tool is there and the alternative too costly that they would use the tool instead of buring their project. Just today I stumbled over this for example, where they use GenAI as well: https://reddit.com
7.
▲
by
Mkengin
9mo ago
Not for coding, but today I stumbled upon these two building their passion project using GenAI, which would otherwise perhaps not be possible: https://reddit.com/comments/1prqfsu
8.
▲
by
Mkengin
9mo ago
It doesn't have to be hyped to be used, for example today I found these two building their passion project using GenAI, which would otherwise maybe not possible, who knows: https://reddit.com/comments/1prqfsu
9.
▲
by
Mkengin
9mo ago
This is just one example, but today I found this where two people build their passion project using GenAI for image generation (+ photoshop), maybe otherwise this project wouldn't even be possible: https://reddit.com/co
10.
▲
by
Mkengin
9mo ago
Though this Codex version isnt on the leaderboard, GPT-5.2-Medium already seems to be a bit better than Opus 4.5: https://swe-rebench.com/
11.
▲
by
Mkengin
9mo ago
At least on swe-rebench it does pretty well: https://swe-rebench.com/
12.
▲
by
Mkengin
9mo ago
Your experience seems to match the recent results from swe-rebench: https://swe-rebench.com/
13.
▲
by
Mkengin
9mo ago
According to SWE-Rebench Anthropic and OpenAI are really close in performance, while GPT-5.2 costs less than half the cost of CC per problem. https://swe-rebench.com/
14.
▲
by
Mkengin
10mo ago
Interesting. So similar to the vision encoder + projector in VLMs?
15.
▲
by
Mkengin
10mo ago
I am eagerly awaiting swe-rebench results for November with all the new models: https://swe-rebench.com/
16.
▲
by
Mkengin
10mo ago
I like this one: https://swe-rebench.com/
17.
▲
by
Mkengin
11mo ago
Or use RL to beat any AI detectors: https://reddit.com/r/LocalLLaMA/comments/1lnrd1t/you_can_jus...
18.
▲
by
Mkengin
1y ago
https://arxiv.org/abs/2311.13600 https://arxiv.org/abs/2410.22911 https://arxiv.org/abs/2409.16167
19.
▲
by
Mkengin
1y ago
Thank you for testing, I will test GPT-OSS for my use case as well. If you're interested I have 8 GB VRAM, 32 GB RAM and get around 21 token/s with tensor offloading, I would assume that your setup should be even faster than mine
20.
▲
by
Mkengin
1y ago
Why Qwen2.5 and not Qwen3-30B-A3B-Thinking-2507 or Qwen3-Coder-30B-A3B-Instruct?