4 ms·
You could burn a few tokens by prompting the LLM to summarize the omit all of the "wrong" answers (where it overrode itself), after the user accepts the respons
by dpkirchner 24d ago
You could burn a few tokens by prompting the LLM to summarize the omit all of the "wrong" answers (where it overrode itself), after the user accepts the response. It'd probably reduce token use in the end, especially for long sessions.