Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
screm
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
screm
7d ago
Indeed, we are helping growth teams get picked by coding agents, this is not a secret. But I do think this has the potential of getting 1000x worse than SEO so I don't imagine people letting that happen. For example we just added code
2.
▲
by
screm
8d ago
Hey yes we are actually currently running this as part of the next wave!
3.
▲
by
screm
12d ago
Well that's true but visibility remains a requirement and it's hard to think of a ranking algorithm that does not take into account popularity at all. Even if a product is perfect, can you really have it in top #10 results if it&#
4.
▲
by
screm
12d ago
We haven't tested with smaller / older models but it would definitely work better. Prompt injection was the top 1 concern for first LLMs so they put a lot of energy into having guardrails at almost every stage afaik (input, tool c
5.
▲
by
screm
12d ago
We decided to use three different providers so we could verify that this choice doesn't impact the result of our experiments (E2B, Blaxel and Daytona)
6.
▲
by
screm
12d ago
If you want your Claude Code to search you can always tweak your own with a good skill, this should work perfectly! It's more a problem for vendors who can't tell all people in the world to download a specific skill first.
7.
▲
by
screm
12d ago
True but they'll probably never integrate ads into the model's thinking (for now it's only a separate display in the apps). When tokens are a commodity you'll just switch to the one you can trust and since all labs are s
8.
▲
by
screm
12d ago
We don't need to reproduce the same errors! Anyway I do think it is going to be different this time because generating content is becoming so easy today that the entire web would just become 99.99% slop very quickly if things don'
9.
▲
by
screm
12d ago
Why? We haven't benchmarked Data Warehouses yet, only prod databases where you wouldn't expect Redshift to be considered.
10.
▲
by
screm
12d ago
Azure database was mostly for enterprise use-cases. Rebuilding a lot of things in-house is a real trend, especially for Claude Code when you don't ask it explicitly to consider all solutions and avoid overhead of managing things yourse
11.
▲
by
screm
12d ago
Thanks, was considering killing it after getting the opposite feedback earlier, now I may keep both options!
12.
▲
by
screm
12d ago
"we do growth hacking and SEO tricks on models and get them to use products that aren't actually best for the job" -> Well this could be seen the other way around. Today, without proper promotion of services, only incumben
13.
▲
by
screm
12d ago
Hey, thanks! I'm wondering if it's clear from our website that this is the price of a fully managed service, not just access to a platform or reports. Think of an SEO agency model.
14.
▲
by
screm
12d ago
This seems intuitive but agents are smarter than that! -> Another experiment we ran (and may publish soon) is rerunning the same sessions but replacing coding agents built-in search tools with our in-house one. At first our own search wa
15.
▲
by
screm
13d ago
Yep
16.
▲
by
screm
13d ago
It should be better now, including in mobile, thanks both for the feedback!
17.
▲
by
screm
13d ago
Yep makes sense I’m relaxing them
18.
▲
by
screm
13d ago
Haha there is a lot at stake for sure
19.
▲
by
screm
13d ago
Yeah sounds kind of like the equivalent of SEA for AI agents (AEA?) except that it’s sneakier since agents can act without you noticing.. anyway this is in the hands of the labs
20.
▲
by
screm
13d ago
Exactly!
21.
▲
by
screm
13d ago
Hey, thanks for the feedback, the leaderboards aren't displaying well on mobile indeed, we are currently shipping a fix that should help with that. Thanks anyway!
22.
▲
by
screm
13d ago
Definitely! But about concentration I'm not so sure, there are ways to counter this effect so in the end it will be a fight like SEO is today. What is certain though is that getting recommended by coding agents will be a top prio for a
23.
▲
by
screm
13d ago
Not sure I got your question right but if you are wondering for your own coding agent then I guess the answer would be a skill? Here what I meant by "how to influence coding agents choices and get products picked" is from a vendor
24.
▲
by
screm
13d ago
Hey! Disclaimer: I am a Co-Founder of Armature (YC P26) which sells growth services to dev tools. This study is part of our broader work on how to influence coding agents choices and get products picked. To understand how agents pick tools
25.
▲
Which tools do Claude, Codex and Cursor choose? We measured 17k runs to find out
(armature.tech)
300 points
by
screm
13d ago
|
153 comments
26.
▲
by
screm
1mo ago
Makes sense, just feel free to reach out if you think any of our features could become useful at some point or just want to discuss "building for agents"! Btw we recently shipped [evals]( https://armature.tech/blog&
27.
▲
by
screm
1mo ago
Yes it does because we can't see the real chat transcripts or any kind of history or memory. We only make sure the calls to your MCP server include "brief, task-specific user intent" following OpenAI's Apps SDK guideline
28.
▲
by
screm
1mo ago
Thanks! Happy to give you a tour and see if it can be helpful or just discuss how you handle these challenges on your side!
29.
▲
by
screm
1mo ago
Thanks, great questions! 1. We ask it :) The SDK adds an optional "telemetry" object to each tool's input schema, with fields like user_intent, agent_thinking and user_frustration. Then the calling agent just fills them in as
30.
▲
Show HN: Product analytics (and evals) for agent sessions on your MCP
(armature.tech)
42 points
by
screm
1mo ago
|
8 comments
More ›