> opinionated on-device intelligence
> Hate speech triage. On-device moderation that flags hateful, abusive and threatening text
What could go wrong here?
1. Not having human in the loop to review it because humans are expensive.
2. Having human in the loop to review it and subject said human to the worst other humans produce.
I don't think you understood. It means these are specialized models. Their toxic model could be ideal for video game lobbies without investing a ton of money if you're an indie dev
This also could be ideal if you want your child to play online to have auto-censorship