3 ms·
You can see the list of things the moderation endpoint scans for in the OpenAI documentation: https://platform.openai.com/docs/guides/moderation/overview https:
by netruk44 3y ago
You can see the list of things the moderation endpoint scans for in the OpenAI documentation: https://platform.openai.com/docs/guides/moderation/overview https://platform.openai.com/docs/guides/moderation/overview
I'm unsure of what the "GPT-4 powered moderation system" entails, though.
Conjecture: My unsubstantiated guess would be them prompting GPT-4 with something like "Is the following excerpt considered to be harmful or unsafe: {training data}" and then limiting the output to just a few words like "Yes", "No" and "It's unclear".
- MallocVoidstar 3y agoAlways funny when I see people talk about using LLMs for creative writing when both OpenAI and Anthropic believe that generating any amount of sex or violence is grounds for a ban.