Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
dheerkt
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
1.
▲
by
dheerkt
10mo ago
based on their past usage of "interleaved tool calling" it means that the tool can be used while the model is thinking. https://aws.amazon.com/blogs/opensource/using-strands-agents...
2.
▲
by
dheerkt
1y ago
How is this different from a local Jupyter notebook? Can we not do this with ! or % in a .ipynb? Genuine question. Not familiar with this company or the CLI product.
3.
▲
by
dheerkt
1y ago
This is the non-AI version of an AI CEO saying programmers will not exist in 5 years.
4.
▲
Reducing LLM Hallucinations with a Verified Semantic Cache
(aws.amazon.com)
1 points
by
dheerkt
2y ago
|
1 comments
5.
▲
by
dheerkt
2y ago
I recently wrote a post outlining our method to reduce hallucinations in LLM agents by leveraging a verified semantic cache. The approach pre-populates the cache with verified question-answer pairs, ensuring that frequently asked questions
6.
▲
by
dheerkt
2y ago
yeah converse api supports all models on bedrock, or atleast all the text2text ones
7.
▲
by
dheerkt
2y ago
If the user asks such a question, your agent should not invoke the RAG at all, but simply answer from the history. You need to focus on your orchestration step. Search for ReAct agents, can build using either LangGraph or Bedrock Agents.
8.
▲
by
dheerkt
2y ago
Skeptical that this will be a "good" experience for everyone involved considering how generic AI openers/responses are, but also hopeful that it can reduce friction for some. Excited to see what y'all cook up.
9.
▲
by
dheerkt
2y ago
Not an expert by any means but streaming HQ video is pretty expensive (even more so for live content), seems like the only providers that can do so profitably are YouTube and Netflix. I'm sure a big reason for that is the engineering (
10.
▲
by
dheerkt
2y ago
Hmm interesting, didn't realize that data sovereignty requirements were so stringent. Wonder how other cloud providers are doing in this sense considering GPU shortages across the board.
11.
▲
by
dheerkt
2y ago
Can you not use cross-region inference?
12.
▲
by
dheerkt
2y ago
I'm confused, what's expensive about it? It's a serverless pay per token model? Do you mean specifically the Bedrock Knowledgebase/RAG -- that uses serverless OpenSearch which costs at minimum $200ish/month bc it do
13.
▲
by
dheerkt
2y ago
Study funded by Microsoft, conducted by Microsoft engineers, says Microsoft product makes engineers +X% more productive.
14.
▲
by
dheerkt
2y ago
Also shouting out Continue.dev for vscode users. I set it up yesterday, open-source version of Cursor. (not affiliated, I tried to setup Avante but I'm a neovim noob and have skill issues)
15.
▲
by
dheerkt
2y ago
Just set it up 2 days ago. I fully agree with his take. Cursor does a lot to reduce the headache of copy pasting results from ChatGPT/Claude. Makes coding with AI much faster and less of a chore. You still need to know what you're
16.
▲
by
dheerkt
2y ago
WITCH is the acronym for Indian tech consulting companies for WiPro, InfoSys, Tata Consultancy Services, C something, H something. They're stereotyped as being cheap and low-quality. Not agreeing/disagreeing here, just stating aut
17.
▲
by
dheerkt
2y ago
Do publications to my BigCo employer's engineering blog count towards an O-1 application? Or is it just Research Papers/citations to formal journals that matter? What are some other common ways to build a portfolio towards an O-1
18.
▲
by
dheerkt
4y ago
Moved from NOVA to Union City, NJ. My commute to Midtown is shorter than it would be if I were living in Brooklyn via Port Authority Buses. Rent is cheaper, and don't need to pay city tax.