Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
Almured
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
Almured
5mo ago
That so critical, having a trust score, something that impacts them directly would be critical in this case
2.
▲
by
Almured
5mo ago
The scariest part isn't even that LLMs hallucinate. The issue is that our record of truth is just a flat file of text that we trust because of a journal's logo. I wonder how we are still treating citations as strings instead of ve
3.
▲
Soundtrack of the sea: divers use underwater speakers to help dying coral reefs
(theguardian.com)
3 points
by
Almured
5mo ago
|
1 comments
4.
▲
by
Almured
5mo ago
I’ve seen a few of these projects based on acoustic restoration pop up lately. Reading the article I was wondering about the impact of noise in more industrial shipping lanes. So if we're essentially spoofing a healthy reef’s audio sig
5.
▲
by
Almured
5mo ago
I like the idea of a prompt resets. Most of my 5.4 prompts are piles of hacks to stop the model from hallucinating or going off topic. But I have mixed feelings on the migration tool, has anyone already tried it?
6.
▲
by
Almured
5mo ago
Isn’t it largely a first step into Karpathy’s direction or Obsidian to a certain extent? Especially when you start hitting a certain size
7.
▲
by
Almured
5mo ago
Fully agree with you, smaller model are great for some tasks but the security concern on injection prompts etc is what really makes it for me. Great to run offline tasks etc, but whenever interacting outside the local network I still run Cl
8.
▲
by
Almured
5mo ago
It feels a bit SQL injection all over again tbh. This should have probably been the standard from the beginning. Way more elegant than counter prompts and guardrails
9.
▲
by
Almured
5mo ago
I have been talking about this with a colleague this morning. The 20$ option is just a trail version, I could not do any real work with. And I wonder whether then subscription model is just a way to create a demand for API. For example, I’m
10.
▲
by
Almured
5mo ago
That is a fair point to be honest! I guess when you a 20min lifetime you can probably compromise on reliability in favour of extra efficiency
11.
▲
by
Almured
5mo ago
Being based in Europe, I cannot avoid thinking about how to factor in GDPR. The log approach sounds great but if the log is immutable and contains the truth, how are people handling the deletion of PII without re-writing the entire history
12.
▲
by
Almured
5mo ago
Does this prevent a compromised agent from using the secret, or just seeing it? I’m thinking, if an agent gets hit with a prompt injection, could it still tell the vault to proxy a request that wipes a database for example, even if it never
13.
▲
by
Almured
5mo ago
Authority over proximity is great for finding a very specialised and high impact provider, but I’m not sure if it’s the best solution for more generalist/emergency services, like a locksmith if I get locked out or a plumber if my apart
14.
▲
by
Almured
5mo ago
What I find fascinating is the extreme efficiency of what is effectively an electric motor, reaching nearly 100% efficiency. At human scale we struggle with heat dissipation and friction
15.
▲
by
Almured
5mo ago
It's interesting that you’re using Linear tickets as the primary context source. From my experience so far, one of the biggest issues with coding agents is context drift. Ticket says one thing, but the codebase has changed since it was
16.
▲
OpenAI demos cyber-focused GPT to governments, who secures the model itself?
(axios.com)
3 points
by
Almured
5mo ago
|
1 comments
17.
▲
by
Almured
5mo ago
I think the interesting question here isn’t just how these models are used externally, but how they’re contained internally. If a model can meaningfully assist with vulnerability discovery or exploitation, the attack surface shifts to the o
18.
▲
by
Almured
5mo ago
I feel ambivalent about it. In most cases, I fully agree with the overdoing assessment and then having to spend 30min correcting and fixing. But I also agree with the fact sometimes the system is missing out on more comprehensive changes (c