Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
agentdev001
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
31.
▲
by
agentdev001
28d ago
You're asking the wrong question here. There is: How exactly does Claude Code* not support AGENTS.md? How exactly does the Claude suite of Models not support AGENTS.md? The problem being pointed at in parent linked is referring to the
32.
▲
by
agentdev001
1mo ago
Yes, ohmypi is an opinionated set of features on the base pi harness.
33.
▲
by
agentdev001
1mo ago
Im having a hard time figuring out what the intended user interface is here. The screenshots im seeing makes this look to be an electron app, but the repo seems to be implying that this is a deployed web app. If this is just another librech
34.
▲
by
agentdev001
2mo ago
No offense intended by my initial reply, I understand how I may have came off in that way however. On the 'extreme' bit, this might certainly be along the lines of personal experience- but I find LinkedIn, HN, and Reddit to be amo
35.
▲
by
agentdev001
2mo ago
There is clear irony in the statement "... an absolutist cold-turkey-approach." ... "except some things like Hacker News, LinkedIn, YouTube." That sounds closer to the opposite end of the extreme.
36.
▲
by
agentdev001
2mo ago
Anecdotally, when I see coding agents preform this action- I see them using bash. IMO less tools is better, if the agent has a shell- so not having a dedicated cut/paste tool is good.
37.
▲
by
agentdev001
2mo ago
Then the agent runtime should be happening in a sandbox, where policy is enforced by a gateway external to it. Bound the agent's autonomy based on of the affects the agent's actions. Approvals should be made into a contract before
38.
▲
by
agentdev001
2mo ago
The way I try to illustrate this to my peers, in the context of automating with llms, is to "do as much of the deterministic work as possible before and after involving an agent". Tbf this is largely a restatement of your comment;
39.
▲
by
agentdev001
2mo ago
Gotcha, I feel like model or provider-specific installs would be a nice QoL improvement in that case. Presumably, part of this issue (beyond the ethos of minimalism) is the aim of shipping shipping an agnostic toolset. For myself, im openai
40.
▲
by
agentdev001
2mo ago
I'd like to understand what features you're referring to that are missing from base-install Pi CLI.
41.
▲
by
agentdev001
2mo ago
Obligatory yes, but only if you're subscription-based and not pay-per-token as enterprise users are.
42.
▲
by
agentdev001
2mo ago
Am I wrong to be somewhat peeved by the use of "RAG" in these contexts? I always read things like this, and wonder if instead the author should be saying "Semantic Retrieval" or something something Vector, etc. Retrieval
43.
▲
by
agentdev001
2mo ago
Nice work! Excited to try $YOUR_HARNESS out! Reading your comment reminded me; I actually did something quite similar at $MY_BETTER_STARTUP! My approach is slightly different, however, employing what I like to call State-Horizon-Aware-Rercu
44.
▲
by
agentdev001
3mo ago
Sounds like user error to me. Codex gives an llm a tool to allow it to use shell in the context of the host and user in which it is running. If a resource is sensitive, and accessible in that context, then the user is doing something wrong.
45.
▲
by
agentdev001
3mo ago
Ah, I wasn't aware things regressed there. Yea certainly workarounds n soft fork sorts of things definitely would work- but thats a bummer than things have changed. From watching Pr's and issues- seems like openai at least wants t
46.
▲
by
agentdev001
3mo ago
You can use the Codex harness with non-openai providers if you want.
47.
▲
by
agentdev001
3mo ago
I keep butting into the question of; why opencode, when you've got codex available? Codex is open source as well, and i can't seem to picture a situation where one would want Opencode over Codex. As far as I can tell, they tick th
48.
▲
by
agentdev001
3mo ago
The "Very good" I'm referring to is far better than only 99%. I can't offer solid stats off the top sadly, so you'll have to just take my word for it ;) I'll take the opportunity to note that if you're run
49.
▲
by
agentdev001
3mo ago
I find papers/articles which discuss solutions that rely heavily on a model in the middle unreadable, if the models used are not discussed. The data you need to get into context for a small model, vs a big boy frontier model, vs a fine
50.
▲
by
agentdev001
3mo ago
I have gone through this process and evaluated the results. Maybe you're referring to their comment as written, but going through what OC described + handholding leads to very good results in my experience.
51.
▲
by
agentdev001
3mo ago
Nvidia Openshell solves most of the hard problems I've run into while building stuff in this space. Observability is, for my purposes, solved by a given framework supporting OpenTelemetry. Guardrails is where I've gotten the most