Answers
What does Naven screen messages for?
Twelve published categories: spam and scams, malware and malicious links, harassment and bullying, hate speech, credible threats of violence, sexual exploitation, illegal transactions, doxxing, encouragement of self-harm, graphic violence, sexual content, and drugs — plus slurs and strong language for younger accounts.
Screening happens before a message is delivered, not after somebody reports it. The point is that the recipient never receives it.
Two things are deliberately not screened. Disagreements and insults are delivered, because a platform that blocks argument is quieter rather than safer. And somebody saying they are struggling is never blocked — that message is delivered, with help offered alongside it, because blocking it would isolate exactly the person who reached out.
The screen misses things, and it is weaker in languages other than English. Calls are not screened at all. All three limits are published rather than buried.