Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
palguna26
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
by
palguna26
6d ago
At this point, i firmly believe all companies are building benchmaxed models, which perform well on older benchmarks but struggle on new one, terminal bench is the best example.
2.
▲
by
palguna26
2mo ago
I just used it yesterday, and overall it does a very clean job, i was given the architect title, which is really close to how i actually use coding agents, it also gives u feedbacks on where you were weak as well, so you should definitely g
3.
▲
by
palguna26
2mo ago
I personally use Codex(plus) a lot, with sol+ponytail, u get concise to the point answers rather than long explanation that openai models are known for, and till date im really satisfied with its coding performance. I also use opencode to t
4.
▲
by
palguna26
4mo ago
I believe customer support should actually be done by humans, as customers feel undervalued when they talk to scrappy voice agents, atleast make an effort to use good ones, so it tries to speak like humans
5.
▲
by
palguna26
4mo ago
Im building termyte, the runtime safety layer for ai agents.
6.
▲
by
palguna26
5mo ago
I just recently shifted to codex since i got frustrated with the token usage which did not allow me to get my work done. Here's my honest opinion i think with the right configs codex does a very good job, like cc is better in terms of
7.
▲
by
palguna26
5mo ago
I just shipped a causal memory system for AI agents and am now working on the mcp for claude code. It's open source u can check it out on: https://github.com/CausalOS/causalos-python