7 ms·
h/t to DeepInfra for being the first 3rd party provider for it on OpenRouter (https://openrouter.ai/z-ai/glm-5.3?endpoint=b711bea7-3994-4935-873d-fb845bd40c00 h
by fra 19d ago
h/t to DeepInfra for being the first 3rd party provider for it on OpenRouter (https://openrouter.ai/z-ai/glm-5.3?endpoint=b711bea7-3994-4935-873d-fb845bd40c00 https://openrouter.ai/z-ai/glm-5.3?endpoint=b711bea7-3994-49...).
- andrewmunsell 19d agoIt's also now live on Ollama Cloud as of a couple minutes ago
- ljlolel 19d agoon my TrustedRouter: z-ai/glm-5.3: also Z.ai, Novita, Atlas Cloud, IO.NET
- johndough 19d agoI have seen you advertise your website a few times. I like the idea of not having to trust the router, so I took some time out of my day to critique your website: https://files.catbox.moe/v68cf7.png https://files.catbox.moe/v68cf7.png My visit to your website went like this: 1. Visit models page 2. Try to find GLM-5.3-Flash (which is among the ~5 models that 90% of people currently care about) 3. Give up scrolling (which would have taken OVER 50 SCROLLS!!!) and use Ctrl + F 4. Try to find input/output/cached price 5. Scroll all the way up to find out which column is what 6. Notice that output price is cut off 7. Notice that the scroll bar is over 100 scrolls further down the page 8. Use Shift + Wheel to scroll horizontally (most visitors probably won't know this trick) 9. Notice that cached price is missing 10. Conclude that this is probably not a serious offering and bounce There are probably more issues later on, but this is how far I got. I would suggest you to: - Deslopify all pages that a user may visit before conversion - List important models first (see OpenRouter rankings) - Move the most important information (model name/input/output/cached price) to the left - Disaggregate the prices per provider (maybe subtables per model? not sure) - Measure cache hit rate and compute effective price per provider (see OpenRouter) (- Optional: Fix the broken link on your HN profile page. Currently, the only way to get from this comment to your website is a search engine.)
- ljlolel 19d agothanks for the feedback, didn't realize people looked there instead of just asking their agents these days. I updated that page https://trustedrouter.com/models https://trustedrouter.com/models
- crossroadsguy 19d agoHow am I supposed to navigate around there? For example the pricing page is empty or is that how it was supposed to look? On models and providers pages there are lists but no way to filter or get any kind of meaningful info. Or is this a WIP/POC?
- ljlolel 19d agoupdated the pricing page link to the models page, and added search we are doing billions of tokens a day and thousands of users
- stavros 19d agoHave you guys been having a good experience with OpenRouter? I tried it out recently with Claude, and it cached no tokens, charging me $200 for one conversation of 11 messages.
- akie 19d agoI just checked because I was a bit paranoid, but I have a 96.6% cache hit rate for GPT-5.6 Luna and 96.8% for Opus 5.
- stavros 19d agoHm, thanks, it must have been some OpenWebUI bug, thank you.
- DefineOutside 19d agoI tried using deepseek v4 flash with OpenRouter. It switches between providers too eagerly which resets the cache. Then, each provider begins to rate limit me for providing so many uncached tokens, so it just keeps on switching providers. I'm paying for every token... why rate limit me? It was unusable compared to just using the official Deepseek provider which has a much better cache rate.
- eikenberry 19d agoYou really need to select your provider with Openrouter to get the best experience. https://openrouter.ai/docs/guides/routing/provider-selection https://openrouter.ai/docs/guides/routing/provider-selection
- creativeSlumber 19d agotheir cache hit rate is 67%. In comparison the provider with the highest hit rate is at 95%.
- matheusmoreira 19d agoDeepInfra has excellent terms of service too!