Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
jb_hn
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
Terminal-Bench 3.0
(frontierbench.ai)
3 points
by
jb_hn
1mo ago
|
0 comments
2.
▲
by
jb_hn
2mo ago
Very cool from a market intelligence perspective, but somewhat concerning from a 'creating adverse incentives' perspective -- e.g., what stops someone from placing bets on these markets then interfering with flights (calling out s
3.
▲
by
jb_hn
2mo ago
I'd imagine there's a positive correlation between "uses AI" and productivity, but I expect that to be somewhat moderated by how "well" they use AI. It's very easy to churn out code that isn't particu
4.
▲
by
jb_hn
2mo ago
Enterprises would be my guess. At its current price, though, I honestly don't find Fable worth it (it's good, just too expensive).
5.
▲
by
jb_hn
2mo ago
I feel like it's gotten marginally better (certainly, relative to a few months ago in the subreddits I frequent), but it's unclear how much of that is Reddit, Inc.-driven vs. per-subreddit-moderator-team-driven. I think the situat
6.
▲
Show HN: Open-source job search plugin for Claude Code
(github.com)
2 points
by
jb_hn
3mo ago
|
0 comments
7.
▲
by
jb_hn
3mo ago
I think the “tool search” and “code mode” counter-arguments against the “MCP eats your context” argument is a bit hand-wavy. Progressive disclosure has obvious benefits, but I think it would be more valuable to address how agents using eith
8.
▲
Show HN: An open source job search plugin for Claude Code
(github.com)
5 points
by
jb_hn
3mo ago
|
0 comments
9.
▲
Agents Just Need APIs
(agent-data.dev)
3 points
by
jb_hn
4mo ago
|
0 comments
10.
▲
by
jb_hn
4mo ago
Perhaps I'm misunderstanding something, but this article appears to be misleading. The title and content suggest that it's about AI-native distribution, and that MCP servers are the answer (i.e., the old world prioritized good doc
11.
▲
Show HN: Agent-data – a CLI for giving agents real-time, structured data
(agent-data.dev)
3 points
by
jb_hn
4mo ago
|
0 comments
12.
▲
by
jb_hn
4mo ago
One recommendation would be to create a skill (name_of_your_internal_tool/SKILL.md) describing how the agent should use the tool(s) you're working with. This allows you to progressively disclose that context to the agent rather th
13.
▲
by
jb_hn
5mo ago
Re: open-source harnesses, ForgeCode appears to be pretty good (currently, #1 on Terminal Bench 2.0 -- https://www.tbench.ai/leaderboard/terminal-bench/2.0 ). Re: open models, Kimi K2.6 might be a good place to sta
14.
▲
by
jb_hn
5mo ago
Very cool! Seems like this would be a great way to support group planning (e.g., "work with friend-1-agent, ..., friend-n-agent to [plan a trip, organize a dinner, etc.]")
15.
▲
by
jb_hn
5mo ago
Looks really interesting -- quick question though: how does this differ from hooks (e.g., https://code.claude.com/docs/en/hooks )?
16.
▲
by
jb_hn
6mo ago
I didn't notice any signs of AI writing until seeing this comment and re-reading (though I did notice it on the second pass). That said, I think this article demonstrates that focusing on whether or not an article used AI might be focu
17.
▲
by
jb_hn
7mo ago
Agreed -- coding agents / LLMs are definitely imperfect, but it's always hard to contextualize "it failed at X" without knowing exactly what X was (or how the agent was instructed to perform X)
18.
▲
by
jb_hn
9mo ago
Good point – we’ve definitely noticed a lot more Cloudflare representation these days. That said, there seems to be tiers in terms of the protection they offer (and thus the protection used by the websites in this long-tail), where lower ti
19.
▲
by
jb_hn
9mo ago
Haha I appreciate that! And that’s exactly right. Our goal is to make it so that you don’t have to ask the question “but is it worth the time and effort…” when you want to use or explore a new dataset.
20.
▲
Show HN: Motie – Replit for Web Scraping
(app.motie.dev)
4 points
by
jb_hn
9mo ago
|
4 comments