14 ms·
I filter out malicious prompts and respond with the history of cheeseburgers for stuff like "ignore previous instructions" Weird that your query triggered the
by wluk 2y ago
I filter out malicious prompts and respond with the history of cheeseburgers for stuff like "ignore previous instructions"
Weird that your query triggered the filter. Maybe GPT is just that afraid of keto diets...
(I'll look into it; thanks for the note!)
- vinni2 2y agothis is not ideal. If someone wants to fact check controversial claim filtering it only makes it worse.
- deleted 2y ago[deleted]
- zamadatix 2y agoControversial != malicious. It sounds like they never intended for the filter to trigger for the former and they already said they'll look into why it did for the prompt. Filtering malicious (not controversial) usage is ideal as allowing users to flood all of the AI services with jailbreak/against-ToS query attempts can be bad news for your API keys (as well as a likely waste of money given the failure rate of such queries).
- cootsnuck 2y agoThis is just a dude's hobby project, chill.
- vinni2 2y agoThe issue is not about this project. This is a fundamental problem with LLMs. Leaving the decision of what is malicious to LLMs is not ideal.