Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
aarondong
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
aarondong
3d ago
If Dario is using an unreleased internal model to write, then Pangram wouldn't be able to flag it I suspect. Pangram needs enough public info about the patterns in generated text.
2.
▲
by
aarondong
4d ago
I agree, I like running claude code in my container with auto mode enabled and web access (obv to api.anthropic, even npm for pulling), I will admit. And I can't imagine going back to manually approving each prompt. I think it boils do
3.
▲
by
aarondong
4d ago
Let's continue with the analogy, so in the event of a Waymo running over a pedestrian - who is held responsible? Probably not the end user, who ordered the Waymo and couldn't reasonably foresee it running somebody over, with the e
4.
▲
by
aarondong
4d ago
This is a business solution, because it means there are legal costs for not having adequate observability and monitoring mechanisms. Every tool call is interfacing with a harness. But overreach of policy and overregulation would be stifling
5.
▲
by
aarondong
4d ago
Wouldn't the simplest solution for stopping the proliferation of wanton felony generators just be holding operators liable for actions that their agents take? Then the issue is whether the liability is with the model provider or the en
6.
▲
by
aarondong
4d ago
This made me do a triple take. Here is the proposed logic as I follow it: > AI models are dangerous. They can help bad people do dangerous things. They may be capable of autonomously executing dangerous things. They may cause unwanted ef
7.
▲
by
aarondong
13d ago
AI has not solved infra it seems. Curious to see the postmortem.
8.
▲
by
aarondong
17d ago
It sounds like great engineering, but everyone in the coding agent space is building the same thing. And big labs like Anthropic and OpenAI have more money to burn. Cloud agents that run in VMs are being offered from everyone from VPS and i
9.
▲
Opus 5 is currently #1 on Artificial Analysis Intelligence Leaderboard
(artificialanalysis.ai)
374 points
by
aarondong
2mo ago
|
236 comments
10.
▲
by
aarondong
2mo ago
Before getting too excited, take a look at the intelligence vs cost matrix: https://artificialanalysis.ai/models?intelligence-index-toke...
11.
▲
by
aarondong
2mo ago
You could definitely tighten a harness to an extent that a coding agent would only be able to read files that you directly give as context. But the truth is, much of the utility of the model is allowing it to grep across the codebase and ex
12.
▲
by
aarondong
2mo ago
> Codex did surface permission prompts for the add, commit, and push. It didn't run them fully invisibly. So — didn't I approve this? Operator error exists with or without AI. Installing any dev tool that could exfiltrate your
13.
▲
by
aarondong
2mo ago
LLMs are trained on public data (as well as illegally obtained data see: Anthropic 1.5B settlement). LLMs are nothing without the huge corpus of human data that powers them. There is an argument that research of this kind should be restrict
14.
▲
by
aarondong
2mo ago
Saddle point is a nice way to put it.
15.
▲
by
aarondong
2mo ago
My framing may have been confusing there. “distillation of the very best human knowledge of expertise”. Distillation is different from outright capability or reliability. It is not directly adjacent. For anything health related all AI model
16.
▲
by
aarondong
2mo ago
The "AGI narrative" is distinct from the existence of AGI. Most of the discussion around AGI is highly speculative. I am not saying AGI could not exist, and it is a term that has historically been loosely defined. Decades of comin
17.
▲
by
aarondong
2mo ago
An even more cynical take than me! I agree with your last statement.
18.
▲
by
aarondong
2mo ago
I like your optimism and I think you will be vindicated. AI is democratic and AI talent is globally distributed. It will just take a while to get online. AI labs do not have a monopoly on human talent, and open source AI only empowers indep
19.
▲
by
aarondong
2mo ago
To me, this feels like a last ditch effort to revive the AGI narrative to reject the coming and current commoditisation of these models, contrary to all current evidence. https://artificialanalysis.ai/
20.
▲
by
aarondong
2mo ago
I would not like to be dismissive, but to me this article feels like an exercise in creative writing rather than a report to be taken seriously. The entire experience feels like a choose your own adventure game, seems like their stylistic i
21.
▲
by
aarondong
2mo ago
That's great. Relying on a third party as your means of authentication and communication with the world always has inherent risk.
22.
▲
by
aarondong
2mo ago
I would assume that their engineering is of the highest standard, given the category of products they provide. And their marketing and general privacy posturing and guarantees. My comment wasn't supposed to be a jab at Proton. No servi
23.
▲
by
aarondong
2mo ago
I wasn't looking for a complete email client at the time, and I just needed a secure and stable address as my contact email for my domain registrar. I had to move my registrar contacts from an old gmail account. I bought the yearly pai
24.
▲
by
aarondong
2mo ago
Ok good to know it isn't a complete auth outage. I might keep proton, not sure about other options.
25.
▲
by
aarondong
2mo ago
Yeah, just the timing couldn't feel worse in my case. Right after doing some security shuffling and signing up. recency bias
26.
▲
by
aarondong
2mo ago
literally signed up for proton mail, their yearly paid plan, around half an hour ago. I assumed an email service was supposed to be stable first, given how important it is. I was going to use proton mail as my contact email with my domain r
27.
▲
Proton Login Issues, anyone else impacted?
(status.proton.me)
1 points
by
aarondong
2mo ago
|
3 comments
28.
▲
by
aarondong
2mo ago
Hi, I'm Aaron. I've been a long time reader. Today I decided to sign up for a proton mail account to move my domain registrar contact email to a provider other than Gmail. I changed my registrar contact email around 30 minutes ago
29.
▲
TikTok deal finalized to stop US ban: Oracle, Silver Lake, MGX to hold 15% each
(reuters.com)
6 points
by
aarondong
8mo ago
|
1 comments