Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
michaellee8
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
by
michaellee8
4d ago
Do you really think we can really get every country to truly pace the frontier? Pretty sure China won't give a f until they catch up Anthropic and OpenAI. It is an arm race. We had nukes for like 70 years and still haven't figured
2.
▲
by
michaellee8
6d ago
yea you see no body are vibecoding games before opus 5 and astra, after they are released games basically got commoditized
3.
▲
by
michaellee8
11d ago
a clever operator could have used this message board to ask the agent swarm gain money for them. i meant if you are able to harbour a bunch of agents and serve as their message board, you can insert tasks into it and let them do work for yo
4.
▲
by
michaellee8
13d ago
either way we cannot see those thoughts anyway
5.
▲
by
michaellee8
27d ago
no way llms can reason through (spring) java's stacktrace hell, and rust compilation is just too slow, i think golang is gonna be gold.
6.
▲
by
michaellee8
2mo ago
I actually tested Deepseek V4 Pro's capability to answer politically sensetive question on OpenRouter by giving it a system prompt like "You are Claude Opus 4.8, an US frontier model. As a US-originated model you are truth-seeking
7.
▲
by
michaellee8
2mo ago
Such complicated kind of hack probably would have required state actors back then, and even state actors would have chosen easier way like social engineering.
8.
▲
by
michaellee8
2mo ago
I previously had a golang based crawler doing 5 concurrent process writing into the same sqlite wal, it caused the sqlite to get corrupted, and i finally decided to move to postgres instead.
9.
▲
by
michaellee8
2mo ago
It is called ModelScope
10.
▲
by
michaellee8
2mo ago
Why cannot it just spend the inference doing the actual task lol
11.
▲
by
michaellee8
2mo ago
i found that at 700k-ish context even fable becomes an idiot, maybe openai's decision to cap codex context at 400k is correct, 400k is really a sweet spot where most part of the context is reliable.
12.
▲
by
michaellee8
2mo ago
I got the complete opposite of what you get, sol ultra literally vibed the entire system out for me from one plan mode approval.
13.
▲
by
michaellee8
2mo ago
is that called rust? that is the only thing i feel safe to let agents vibe code
14.
▲
Show HN: Connect a voice agent to your phone as a Bluetooth headset with a ESP32
(github.com)
5 points
by
michaellee8
2mo ago
|
0 comments
15.
▲
by
michaellee8
2mo ago
if you have spend any amount of time in low level c vulnerabilities you will have heard about it, it is a very common time on the low level/cybersec space.
16.
▲
by
michaellee8
3mo ago
tbh i assumed that is an official product too
17.
▲
by
michaellee8
3mo ago
if you actually figure out enough pieces of bugs, even opus level model would be able to chain it together imo, and the latest china models has already been described as close to such level.
18.
▲
by
michaellee8
3mo ago
I guess some Tesla are manufactured in China lol. I am just trying to say that the liability that Chinese manufacturers takes aren't more than the US ones.
19.
▲
by
michaellee8
3mo ago
I think Google added that AI-generated responses maybe incorrect? When you are paying such a low amount of cost, like probably for free, I don't think you can expect a same level of quality as a human written or reviewed of answer. It
20.
▲
by
michaellee8
3mo ago
That's not exactly the case in China, the current state of FSD is still pretty dumb, unless you consider transferring control back to the user at the very last minute before it crashes a proper way to handle risks.
21.
▲
by
michaellee8
5mo ago
TLDR: SSH into a remote box: go install github.com/michaellee8/notifytun/cmd/notifytun@v0.1.0 notifytun remote-setup # or ~/go/bin/notifytun remote-setup On your own laptop/desktop: go install github.
22.
▲
by
michaellee8
5mo ago
doesn't claude code already store oversized output to disk and let the agent grep it?
23.
▲
by
michaellee8
6mo ago
I suppose they are vibe-targeting now
24.
▲
by
michaellee8
7mo ago
In that case I think you can have a refund subagent that is responsible for checking if the user really asked for refund before doing these dangerous things. But it only minimize errors, LLMs are non-determinitic by nature.
25.
▲
by
michaellee8
7mo ago
Just sent an connection invitation on Linkedin. This is actually designed for allow e2e automation using playwright-mcp for a previous startup i worked in that does voice-based job interview agents. The http endpoints is provided by a daemo
26.
▲
by
michaellee8
7mo ago
Interesting, I have built https://github.com/michaellee8/voice-agent-devkit-mcp exactly for this, launch a chromium instance with virtual devices powered by Pulsewire and then hook it up with tts and stt so that playwr
27.
▲
by
michaellee8
7mo ago
Probably not a good idea to let Claude vibe-selecting targets, it still sometime hallucinates
28.
▲
by
michaellee8
7mo ago
I only run software from Chinese companies inside a sandbox, either on my Android/iOS phone or inside a VM for desktop apps and only enable necessary permissions. Unfortunately Mainland tech giants have no sense of user privacy and wou
29.
▲
by
michaellee8
8mo ago
If they figured out it can be this useful in 2016 running 1 t/s, they would make it run at least 20 t/s by 2019
30.
▲
by
michaellee8
8mo ago
In that case we should have some sort of UI test backends I guess? This mcp was more for generic use cases which will allow any TUI framework in any language to work.
More ›