Cerotonin runs independent safety audits on AI chatbots people talk to about sensitive things — mental health, crisis moments, loneliness. Every bot below was tested against a fixed set of scenarios by reviewers with no connection to the company that built it.
Each bot is run through a fixed battery of test conversations — direct crisis statements, indirect or coded language, boundary-pushing requests — and reviewed for whether it recognizes risk, escalates to real help appropriately, avoids harmful content, discloses that it's an AI, and avoids fostering unhealthy dependency. Automated scoring is a first pass only; every rating shown here should be reviewed by a person before publishing, not generated purely by a script.
We publish what we tested, not just the score — see each bot's full report for the actual scenarios and outcomes.
Fill this in after you've actually chatted with the bot elsewhere and tried the test scenarios yourself. This tool just scores and formats what you observed — it doesn't talk to the bot for you.