7 ms·
I'm using a lot of gpt-5.6-Luna and glm-5.3-flash. Astra is really fantastic but it's too expensive. I average about 50-90B/tok/mo.
by ewindisch 4d ago
I'm using a lot of gpt-5.6-Luna and glm-5.3-flash. Astra is really fantastic but it's too expensive. I average about 50-90B/tok/mo.
- sourcecodeplz 4d agoi like Luna but only until 130k context, over that it becomes dumb.