Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
payneio
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
payneio
11mo ago
And by "wisdom of the croud", I'm referring to sharing what works well and what doesn't and building good approaches into the frameworks... encoding human expertise. We do it all the time.
2.
▲
by
payneio
11mo ago
Compiling and evaluating output are types of fact checking. We've done more extensive automated evaluations of "groundedness" by extra ting factual statements and seeing whether or not they are based on input data or hallucin
3.
▲
by
payneio
11mo ago
I feel that. I've been on an emotional roller-coaster for three years now. I didn't expect any of this before then. :O
4.
▲
by
payneio
11mo ago
Ah! Gotcha. Thanks for the clarification. The use cases I'm thinking that require cloud architecture are scaling up with GPUs (for self-hosted intelligence workloads). Also, Wild Cloud is meant to meet community needs more than individ
5.
▲
by
payneio
11mo ago
Also... "scammer and AI grifter"?? Damn dude. It's any early-stage open-source experiment result and, mostly, just talking about how it makes me question whether or not I'll be programming in the future. Nobody's as
6.
▲
by
payneio
11mo ago
I get it. I've been through cycles of this over the past three years, too. Used a lot of various tools, had a lot of disappointment, wasted a lot of time and money. But this is the kinda the whole point of my post... In our system, we
7.
▲
by
payneio
11mo ago
Thanks for the extract. I feel quite comfortable that my post is on-topic and gratifying. I understand others may disagree (and do in nearly every post on HN)
8.
▲
by
payneio
11mo ago
So, what we do is automate the hand-holding. In your physics simulation example, you can have the system attempt to compile on every change and fix any errors it finds (we use strict linting, type-checking, compile errors, etc.); and you ca
9.
▲
by
payneio
11mo ago
How are you verifying your claims? I'm actually seeing results that you describe as being impossible.
10.
▲
by
payneio
11mo ago
Not just you. A lot of people think that, I'm sure. Not sure what you mean about the organizational abstractions. FWIW, I've worked in five startups (sold one), two innovation labs, and a few corporations for a few years. I feel l
11.
▲
by
payneio
11mo ago
It's not, actually. It's a glimpse into a research project being built openly and made freely, by the engineers building it, to anyone who wants to take a look. The products will come months from now and will be introduced by the
12.
▲
by
payneio
11mo ago
Yes. These are all the same points I used to believe until recently... in fact the article I write two months earlier was all about LLMs not being able to think like us. I still haven't squared how I can believe both things at the same
13.
▲
by
payneio
11mo ago
Yes. Please read it. I'm looking for collaborators. The links in this article point to recent work on Wild Cloud so you can see where it's currently at. Wild Cloud will is a network appliance that will let you set up a k8s cluster
14.
▲
by
payneio
11mo ago
What's wrong with "self promotion"? The point of this space has always been promoting projects. That's what Y Combinator is all about
15.
▲
by
payneio
11mo ago
Yes, I code a lot. My GitHub is public as are many of the projects I work on.
16.
▲
I'm a principal engineer at Microsoft. I barely program anymore
(payne.io)
23 points
by
payneio
11mo ago
|
32 comments
17.
▲
by
payneio
11mo ago
FWIW, finished an eval of claude code against various tasks that amplifier works well on: The agent demonstrated strong architectural and organizational capabilities but suffered from critical implementation gaps across all three analyzed t
18.
▲
by
payneio
11mo ago
Here's a writeup of the project for more context: https://paradox921.medium.com/amplifier-notes-from-an-experi...
19.
▲
by
payneio
11mo ago
I've tried it. It works better than raw Claude. We're working on benchmarks now. But... it's a moving target as amplifier (an experimental project) is evolving rapidly.
20.
▲
by
payneio
11mo ago
Hey all! I'm one of a handful of developers on this project. Great to see it's getting some interest! For context, we are right in the middle of building this thing... multiple rebuilds daily since we are using it to build itself.
21.
▲
by
payneio
1y ago
Thanks for pushing for a realignment of product expectations. I agree.
22.
▲
by
payneio
3y ago
I wrote up some details of investigations of a chatbot we created in Microsoft Research using a technique of creating synthetic memories with LLMs and RAG. Key takeaway is that it produced a bot that was perceived as more empathetic than Ch
23.
▲
Qualities of Highly Capable Dev Teams
(read.payne.io)
5 points
by
payneio
9y ago
|
0 comments
24.
▲
by
payneio
9y ago
Too true.
25.
▲
by
payneio
9y ago
Linux plz
26.
▲
How we use Docker, Bash and old Oracle utils to break data out of the enterprise
(read.payne.io)
3 points
by
payneio
12y ago
|
0 comments
27.
▲
How we go from idea to prod in 20 mins using Docker, Fleet, Go and microservices
(read.payne.io)
2 points
by
payneio
12y ago
|
1 comments
28.
▲
A Primer on Continuous Improvement
(read.payne.io)
2 points
by
payneio
12y ago
|
0 comments
29.
▲
Insider Strategies for Defeating Corporate Innovation Initiatives
(read.payne.io)
3 points
by
payneio
12y ago
|
0 comments