7 ms·
I found it far too expensive for Anthropic. Entire context of every conversation is sent each time you type anything. Switched to a local model running from Oll
by NetworkPerson 8mo ago
I found it far too expensive for Anthropic. Entire context of every conversation is sent each time you type anything. Switched to a local model running from Ollama. Not quite as smart as opus, but good enough for my needs.
- CjHuber 8mo agoDoes it not use prompt caching?