6 ms·
One suggestion is to make a list or make a skill to have your agent keep a list of things you do not feel work well with today's models. And then, when new mode
by shostack 15d ago
One suggestion is to make a list or make a skill to have your agent keep a list of things you do not feel work well with today's models. And then, when new models come out, periodically, revisit items on that list to see if you get better results.
- KnightHawk3 15d agoOkay serious question, why not just write this down in your notes or something? Why use a language model for it.
- shostack 14d agoYou can. The point of having your agent keep track of it is that it will likely notice things you won't, and it can automate cataloging it with relevant metadata (prompts, environment, examples, etc.) that make it trivial to automate rerunning those tests when new models launch.