12 ms·
GPT-6 Astra on OpenRouter
- d2p 12d agoOdd that the tool call failure rate is so high (5%) for the OpenAI provider than Azure (0.2-0.5%).
- StopTheCringe 12d ago[flagged]
- Osama0456 12d ago[flagged]
- r_lee 12d agothat's literally what the link is for...
- kingstnap 12d agoIts also available finally to Pro users! Just took 24 hours.
- InsideOutSanta 12d agoThey gave out bankable resets for every day people on pro plans didn't get Astra. Given that, I wish they'd waited a few more days before activating it on my account :-D
- wincy 12d agoThey haven’t activated Astra for me yet, I have two resets now. I’ve been using the opportunity to test out how good 5.6 Sol is at computer use asking it to generate stuff in Blender which has been… interesting Edit: nevermind it JUST gave me a notification to use it!
- paxys 12d agoThat's a pretty genius internal incentive to move fast.
- embedding-shape 12d ago> They gave out bankable resets for every day people on pro plans didn't get Astra. Yeah, when I saw that Tweet I knew the person was saying it because they knew it'll be available within 24h.
- wincy 12d agoI must have been one of the last ones to get it because I managed to snag two bankable resets. Already reset once after having it clear up geometry in Blender, but it did a fantastic job!
- cromka 10d agoMeanwhile Fable with it's restrictions is still only available on 100 USD plan upwards.
- r_lee 12d agois Azure for this actually ZDR?
- jiggawatts 12d agoIt cracks me up that you have to ask in a public forum because the vendors purposefully obscure this critical information. The only reason most of my customers would use Azure Foundry instead of OpenAI directly is the ZDR assurance but it is so incredibly difficult to extract out of their model menu. There is no trivial way to block non-ZDR models either so every customer has to “vet” and individually approve models. If anyone from Microsoft is reading this: get your act together! You’re failing at the one thing people might want to pay you to do!
- r_lee 12d agofrom my understanding, ZDR for this is for vetted customers only + their whole "privacy preserving abuse monitoring" or whatever is active but again, seems like there's no word from Azure if this applies to them. very confusing.
- vatsachak 12d agoDamn I am so hyped
- deleted 12d ago[deleted]
- gavinray 12d agoI have GPT-6 access in Codex and OpenAI API now I'm a Business plan user with Cyber verification enabled, FWIW.
- embedding-shape 12d agoSame just got access literally this minute, Pro user here, no Cyber verification but have passed my ID over to them back in 2024 or something, maybe at the ChatGPT 3 API launch or something? Has there been anything published about if Astra uses different amount of usage from your subscription plan compared to Sol? Don't recall coming across that in the press releases.
- deleted 12d ago[deleted]
- deleted 12d ago[deleted]
- algoth1 12d agoJust got it. European plus user here. Only codex, no chatgpt
- cute_boi 12d agoI got it, but sadly no resets...
- noob0053 12d ago[dead]
- taywrobel 12d agoYou made an account just to post this?
- simonw 12d agoI posted this in the other Astra thread but it's just fallen off the homepage, so... Pelicans from Astra, plus 5.6 Sol, Terra, Luna for comparison: https://static.simonwillison.net/static/2026/gpt-6-and-5.6-pelicans.html https://static.simonwillison.net/static/2026/gpt-6-and-5.6-p... I think this is a genuinely interesting comparison grid. Astra may be more expensive, but if you have a budget of 10 cents for a Pelican Astra low gives you something SO much better than the other models. Astra uses less tokens overall too, for better results. Astra transcript here: https://tools.simonwillison.net/markdown-svg-renderer?url=https%3A%2F%2Fgist.github.com%2Fsimonw%2Ff789d2784fc6c5b870cc80f0b7cd9d01 https://tools.simonwillison.net/markdown-svg-renderer?url=ht...
- leoqa 12d ago[flagged]
- ComplexSystems 12d agoIt's absolutely related.
- samuelknight 12d agoHow are we supposed to know if Astra is frontier without the pelican?
- satvikpendem 12d agoIt's simonw. It's interesting to see their pelican benchmark, another comment by a different author elsewhere here shows some very good SVG generation too.
- Xunjin 12d agoDo you have other ideas of "combinations" for this kind of benchmark? I'm wondering if this is being trained on by the models today.
- r_lee 12d agoI have a feeling they've been doing that for a while now, even if unintentional, as it's such a well known benchmark
- XCSme 12d agoThat's some crazy SVG generation: https://aibenchy.com/compare/openai-gpt-6-astra-high/google-gemini-3-8-flash-high/anthropic-claude-fable-5-1-high/#showcase=df3e2eaf3a701c60 https://aibenchy.com/compare/openai-gpt-6-astra-high/google-... It took a while to test it, initially OpenRouter was giving Not Found errors for this model ID.
- embedding-shape 12d agoAt the bottom it says "Score 98.58", what measure is used for this score? It's kind of horrible, the perspective is all off (legs of the table makes that very obvious), the mouse/hamster has two mouths, a stub for a right paw, looks like left hand holds a melon on a stick or something, and there are pluses in the background for some reason. Not sure it'd call it "close to perfect" which the score seems to want to indicate.
- XCSme 12d agoYeah, that's confusing, the score is for the entire benchmark, not for SVG generation only. Good point about the mouths, I just noticed, lol Imo, it's still better than most models, I personally like the stylized perspective. You can view here all generations for all models: https://aibenchy.com/showcase/ https://aibenchy.com/showcase/
- XCSme 12d agoDo you prefer the fable one? It's more "correct" but looks a lot worse in my opinion: https://aibenchy.com/compare/openai-gpt-6-astra-high/google-gemini-3-8-flash-high/anthropic-claude-fable-5-1-high/#showcase=f36ffedba4331387 https://aibenchy.com/compare/openai-gpt-6-astra-high/google-...
- kulahan 12d agoNo, both of them look awful. I genuinely am unsure if this is a meme you're making that's going over my head...?
- 12d ago
- starik36 12d agoWhat is the actual utility of using this model on Azure? It's twice as expensive, according to the link. Do Azure offer something that simply hitting the OpenAI endpoint doesn't provide?
- itsjustkev 12d agoCompared to OpenAI flex? I'm pretty sure that is their batch processing endpoint, which is naturally cheaper.
- claiir 12d agoThey’re ZDR and the OAI ones aren’t
- 6thbit 12d agoInvoking simonw for pelicans pretty please.
- drivers99 12d agoThat's over here: https://news.ycombinator.com/item?id=49570643 https://news.ycombinator.com/item?id=49570643
- OpenGayEye_ 12d ago[flagged]
- vb-8448 12d agoPlayed in codex app a couple of hours today: it feels much faster than SOL, even if the TPS is half of it.
- WASDx 12d agoLikely because it uses fewer thinking tokens (that you don't see anyways).
- marsven_422 12d ago[dead]
- jaesonaras 12d agoAnyone had success using Astra as a Foundry model via Github Copilot? The error I get is that tooling is not available if reasoning has a value.
- sumedh 12d agoJust got access to it on Plus plan in Australia. 2 Banked resets as well.
- jjcm 12d agoIt's ability to handle non-90 degree cutouts and shapes for web dev is one of the best I've seen. The vision model on this is VERY capable. Here's an image design source of truth: https://image.non.io/78f4cd8b-2560-4643-9a51-96a89171f994.webp https://image.non.io/78f4cd8b-2560-4643-9a51-96a89171f994.we... And here's the page it build from it: https://image.non.io/e7d3a9e5-f9df-4fd8-b79f-1f90280f978f.webp https://image.non.io/e7d3a9e5-f9df-4fd8-b79f-1f90280f978f.we... Note the flowing svg lines, and how accurately it recreated them. Here's Opus 5 for comparison - you can really see how while Astra really recreated the flow that was in the original design, opus only got the general vibe: https://image.non.io/dfe13de0-4487-431f-8b69-544ff3030dac.webp https://image.non.io/dfe13de0-4487-431f-8b69-544ff3030dac.we... One thing I will say is you are paying for quality. That site build cost $24 - extremely non-trivial for a simple frontend.
- mydreamof 12d agoI don't get it. For me it seems Opus was more accurate in terms of for example this small building in the right down corner
- CapsAdmin 12d agoI would say opus was in some ways more accurate, but missed the higher level curvature feel of the site that astra picked up on. It sounds completely trivial and likely I'm wrong here, but could it be that opus saw the reference image squished? That might explain the sharper horizontal curvature
- jjcm 11d agoOpus 5 is quite good, and has been the SOTA for this test (by my own subjective comparison) for the last few months. One thing it's always struggled with though is recreating smooth SVG curves/cutouts. Another way to think about it is you could prompt Astra to put the building back in. You couldn't prompt Opus 5 to get the correct curvature of the line / cutout. That part has always been a huge struggle for models.
- copperx 12d ago> That site build cost $24 - extremely non-trivial for a simple frontend. I would say that $24 is trivial IF that's the final design. The truth is that the cost doesn't leave much room for error or experimentation.
- MisterMunchkin 12d ago$10/$50 is incredibly expensive compared to Chinese models which are cents. I think they’re really going to struggle selling these models long-term. My company is already massively cutting down on access because they’ve realised most people don’t actually produce any value using it. All the tokenmaxers have ruined it for the rest of us now that accounting have seen the costs.
- KptMarchewa 12d agothe only thing that matters is cost per task. Astra seems to be massively efficient.
- simianwords 12d agoNo it doesn't. Any source?
- gentlewater 12d agoNot really comparable IMO. Astra and Fable are not the every day workhorse you reach for to do basic tasks (unless your company has fuck you-money), they’re the tool you break out when you need the absolute strongest performance. There are plenty of tasks where finding and fixing one or two extra edge cases saves the business a lot of money, even if the cost is high. The best example would be scanning for vulnerabilities, if these models weren’t kneecapped in that area.
- ghosty141 12d agoWe have ChatGPT Pro at work and I usually use Terra medium/high and only bring out Sol High when the big or feature actually requires "thinking"/complex behavior. This has worked pretty well for me and it's very token efficient
- yurishimo 11d agoHow much code are you shipping in a day? I find I can pretty comfortable use Sol high most of the day and stay within the 5 hour limit. I’ve got too many meetings to allow me time to write code continuously for an entire day. I usually finish about one ticket a day and then review 1-3 tickets for my colleagues.
- friendlypenguin 12d agoI was really hoping for Astra to be less expensive then Opus...
- redox99 12d agoIt is (uses way less tokens)
- simianwords 12d agoNo it isn't cheaper, any source for task vs price comparison to Sol? Edit: GPT-6 Astra (low): 57 Intelligence Index, $7.70/M tokens GPT-5.6 Sol (high): 57 Intelligence Index, $3.08/M tokens So for the same measured intelligence, Sol costs only 40% as much — i.e. ~60% cheaper, while Astra is ~2.5× more expensive. Why is the burden of proof on me tho!?
- wickedsight 12d agoThis is the second comment I see where you write "no it isn't", without providing a source for your statement. Then you follow it by asking for a source. So is there a source you can provide to back up your statement?
- dgellow 12d agoIt’s supposed to work the other way, if you make a positive claim that it is cheaper, where is your proof?
- haaz 12d agoAstra is obviously better than sol. Look at the benchmarks, astra medium costs the same as sol X high whilst costing the same per task https://artificialanalysis.ai/models/releases/gpt-6-astra https://artificialanalysis.ai/models/releases/gpt-6-astra
- slopinthebag 12d agothey said opus, not sol, it's cheaper than opus by almost half at max. astra high is also 3x cheaper than opus max at basically the same intelligence. astra high is also about as expensive as sol max while being more intelligent. astra medium is cheaper than sol max while also being cheaper and roughly same intelligence. im going to replace my sol usage with astra high/medium i think caveat: benchmarks are really fuzzy with llms
- 1saadcodes 12d agoThe higher price seems less important if it actually gets the job done with fewer tokens. I'm still very worried that this will end up coming back to bite us, by becoming more expensive once they inevitably nerf it. Every major model provider does that now after all
- forrestthewoods 12d agoThrew $10 at this to help me prepare for my league’s fantasy auction this weekend. It spend $3.50 and then said “this action would cause you to go above your spending limit”. Then I threw $100 for a Codex Max sub and it included Astra and it did it for me. Sure seems like Astra is expensive AF.
- upcoming-sesame 12d agoAny tips on using Astra as orchestrator with Luna workers efficiently in codex?
- logged4upvoting 12d ago(This applies to Sol but probably works for Astra too) I've created with Sol a skill called Low Quota Mode that intends to reduce the use of tokens usages by the frontier (intelligent model) and delegate the use of bulk reading of docs/code and implementation to a sub-agent running Luna Max. Sol is asked to supervise, read the diffs and approves the commit/pr. The skill might need some iterations while you use it, for example at the end of a rough session you can ask Sol how did it went, which were the points of conflict with Luna and try to iron them little by little by editing the skill. Also in difficult tasks, ask to babysit the sub-agent model, I've seen it makes more effort into communication between frontier and sub-agent to guide the task with more care. So far it has reduced my tokens usage a lot (have not quantified but the quota lasts more).
- DrProtic 12d agoCould you shoot us the github gist?
- pointitkememe 12d ago[flagged]
- boxed 12d agoThe lemur is worse though. The hands don't follow physics for example. You claim this would be easier, but then the model is worse.
- thejassbrain 12d agogreat
- swe_dima 12d agoAccording to the metrics the "fast" mode is not any faster...
- NSUserDefaults 12d agoI misread the title as GPTA-6, wondering what sort of crazy crossover was happening.
- ebiester 12d agoFor those on plus, are you seeing astra limited to medium? Considering the rate limits that probably makes sense, but I'm wondering what different groups have access to.
- saidnooneever 12d agoin the middel of a coding session with 5.6 sol. astra popped up, swapped, asked review, it fixed a few really critical bugs immediately. pretty nice. one around some resource lifetimes in a rendering pipeline which would have been a nightmare to find manually. almost had the feelin it was watching its little brother fail and had to 'step in' for a moment :'). time to go play outside...
- killerstorm 12d agoI tried it with some humanities questions and with the default OR system prompt (no prompt?) it seems to be rather mild - lacking usual AI mannerisms. Kinda cool.
- cmrdporcupine 11d agoI'm appreciating Astra writes much more humanely somehow than Sol did, and certainly better than the Anthropic models. It's terse, like all GPT models by default, but the sentences feel less obscurantist. It's also more pro-active about problem solving.
- christophilus 12d agoTangentially related, but Astra generated some of the worst Odin code I’ve ever seen. Turns out AGI is indistinguishable from an Oracle subcontractor who hates tech and hates his job.
- frenchtoast8 12d agoI signed up for OpenRouter, loaded it up with $25, and after running a test prompt immediately had my account suspended. There’s no way to talk to support, emails go nowhere, and their Discord is swarmed with people who also are getting no responses to anything. I would recommend staying away from OpenRouter. No matter how good the service is, if anything does go wrong, you have no recourse and you lose every credit in your account. Ironically some of the few responses I actually saw in the Discord were doubling down on their “no refunds no matter what” policy.
- miyuru 12d agotake screenshots of the responses and do a chargeback.
- predkambrij 12d agoI did get a response when sending to support@openrouter.ai about a year ago (complaining about some information claims), not about my account.
- frenchtoast8 11d agoI emailed them a few days ago and the automated reply says to expect a week for a response. That wouldn’t be concerning except on Discord there are people begging for help after waiting multiple weeks.
- Gecko4072 12d agoThey were just bought for $7B and rely on Discord?
- bellowsgulch 12d agoGod, that's a step up from AI support, which is so sad to type out loud.
- deleted 12d ago[deleted]
- deleted 12d ago
- azinman2 12d agoI’m confused. I thought it was being only rolled out to select partners for now?
- kzrdude 12d agoThe names, I thought they were going to keep Sol, Terra, Luna for a while (while increasing versions). Are names like Sol and Astra really burned as one-offs? I think that's a waste of a good model name. Hope they keep them going with updates, as nicknames for the various model sizes on offer.
- kdnvk 11d agoAstra is a fourth tier above Sol, analogous to Fable and Opus.
- gertlabs 12d ago[dead]
- ellessarr 11d ago2× price only wins for review if it catches bugs the cheap model drops — nobody runs that test, everyone quotes the benchmark.
- theagenticleade 11d agoThis feels like a genuine step change in AI development. Inline with Dec 2025 release of Opus 4.6. Where do we go from here??
- sejje 11d agoFaster; cheaper.
- thenthenthen 11d agoAnyone tried it for pcb/circuit design yet?
- hermesrouter 11d ago[flagged]