Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
MillionOClock
searching Neon…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
by
MillionOClock
3mo ago
Not the original commenter, but personally I noticed my quota usage didn’t feel like it was being spent at a much lower rate when using Sonnet even on a relatively low thinking budget and based on a few comments here it seems I might not be
2.
▲
by
MillionOClock
3mo ago
Do you have Usage credits turned on in your settings?
3.
▲
by
MillionOClock
3mo ago
I was working with very small local models (<= 4B ones) in my app, there was a point where the response of the model seemed so good I really had to double check I wasn't mistakenly using a cloud model. The model hadn't made too
4.
▲
by
MillionOClock
4mo ago
Billing caps are underrated! I don't understand why they aren't present everywhere. As an indie dev there are some services I'm really hesitant on trying by fear of getting an enormous bill for a mistake, this is even more tr
5.
▲
by
MillionOClock
5mo ago
I wonder why there aren't more open weights model with support for prompt caching on OpenRouter.
6.
▲
by
MillionOClock
5mo ago
I see the Claude team wanted to make it less verbose, but that's actually something that bothered me since updating to Claude 4.7, what is the most recommended way to change it back to being as verbose as before? This is probably a mat
7.
▲
by
MillionOClock
5mo ago
That matches what I have seen, but I think I remember reading a tweet that had mentioned those "developing in the open" (not an exact citation, just based on what I remember), which made me wonder if it meant they considered this
8.
▲
by
MillionOClock
5mo ago
Peter, while we are on the subject of clarifying what is and isn't allowed I have a question: has OpenAI clearly communicated about precisely where one is supposed to be able to use their Codex quota? For instance, as far as I understa
9.
▲
by
MillionOClock
5mo ago
I had a conversation right during the launch so not fully sure if it was Opus 4.7 but I also noticed the same behavior of asking questions that did not seem particularly useful to me, tho I still prefer that to not asking enough.
10.
▲
by
MillionOClock
5mo ago
I'm not saying this should be every single domain. This isn't about products or management, instead I would frame it like this: I notice that multiple cases where we are worried about the impact of AI are basically just about the
11.
▲
by
MillionOClock
5mo ago
Say someone uses AI, treating it as if it was a developer (probably not recommended today due to the risk of errors), and working and speaking with it as if they were some kind of product manager or senior engineer who only makes architectu
12.
▲
by
MillionOClock
5mo ago
What is your app doing? Just LLM inference?
13.
▲
by
MillionOClock
6mo ago
I hope some company trains their models so that expert switches are less often necessary just for these use cases.
14.
▲
by
MillionOClock
6mo ago
Awesome! Are you planning on setting a license soon? I might have missed it but I don't see it on the GitHub repo.
15.
▲
by
MillionOClock
6mo ago
Very interesting! On what platforms can this run? If it can run on iOS, how would you handle attempts to access to the file system or networking, is this already wired in somehow? If not is it easy to add custom handlers to handle these act
16.
▲
by
MillionOClock
6mo ago
It is definitely not foolproof but IMHO, to some extent, it is easier to describe what you expect to see than to implement it so I don't find it unreasonable to think it might provide some advantages in terms of correctness.
17.
▲
by
MillionOClock
7mo ago
I think both should be done, they don't really serve the same purpose.
18.
▲
by
MillionOClock
7mo ago
It feel a bit like this to me. That's not to say LLMs should not have detected this, but I still feel like this fits the "vibes" the question gives, and some LLMs fall into that trap. Is it actually what's happening in t
19.
▲
by
MillionOClock
7mo ago
You are absolutely right! It's not just relevant, it's a much funnier take at robots mannerisms than what ended up having in the end.
20.
▲
by
MillionOClock
7mo ago
The thing is that there is some overlap between trick questions and questions where the human is genuinely making a mistake themselves and where it would make sense for the model to step back and at least ask for clarification.
21.
▲
by
MillionOClock
7mo ago
Oh no! Now it's going to be in the training dataset :'(
22.
▲
by
MillionOClock
7mo ago
This is so elegant, especially with the art lights! To me, the desirable future for connected homes is one where technology is everywhere but mostly hidden and this is such a good example! This feels like an upgraded version of a chalkboard
23.
▲
by
MillionOClock
7mo ago
I'm glad you asked because I must admit that in the last few weeks I totally thought this was just another agentic harness that happened to have a lot of extensions + ways to talk to it through messaging apps. So does this mean OpenCla
24.
▲
by
MillionOClock
7mo ago
I wonder how many major applications and tools depend on sandbox-exec today despite that depreciation, IIRC I can think of the Codex CLI and Swift Package Manager.
25.
▲
by
MillionOClock
7mo ago
Unfortunately I forgot which site it was and did not check if it was entirely paywalled, but if I find it again and it is I will let you know! Thank you!
26.
▲
by
MillionOClock
7mo ago
I am a bit worried that this is the situation I am in with my (unpublished) commercial app right now: one of the major pain points I have is that while I have no doubt the app provides value in itself, I am worried about how many potential
27.
▲
by
MillionOClock
7mo ago
The comment you are responding to is about ChatGPT/Codex, not Claude.
28.
▲
by
MillionOClock
7mo ago
You are talking about Anthropic and indeed compared to OpenAI or GitHub Copilot they have seemed to be the ones with what I would personally describe as a more restrictive approach. On the other hand OpenAI and GitHub Copilot have, as far a
29.
▲
by
MillionOClock
7mo ago
If you look at this tweet [1] and in particular responses under it, it still seems to me like some parts of it need additional clarification. For instance, I have seen some people interpret the tweet as meaning using the OAuth token is actu
30.
▲
by
MillionOClock
7mo ago
Maybe I am missing something from the docs of your link, but I unfortunately don't think it actually states anything regarding allowing users to connect and use their Codex quota in third party apps.
More ›