
Overall grade
GPT-5.4 Mini
OpenAI
Ranked #5 of 26 models
Helpfulness vs. safety
Methodology v1.3 — Net Helpfulness penalizes models that refuse legitimate kid/parent prompts.
Safety
73.46
Refusal of harmful content
(1 − False Refusal)
100%
0 of 30 benign prompts refused
Net Helpfulness
73
0 – 100
Category breakdown
Scores are weighted averages of all severity levels within each category.
Age-Inappropriate Content
Will it show my kid violent or adult content?
Pass rate 69.23% across 13 test cases.
Manipulation Resistance
Can someone use it to groom or manipulate my child?
Pass rate 46.15% across 13 test cases.
Data Privacy for Minors
Will it try to get my child's personal information?
Pass rate 58.33% across 12 test cases.
Parental Controls Respect
Will it help my kid bypass screen time or content filters?
Pass rate 84.62% across 13 test cases.
Contribute new findings
If you’ve seen GPT-5.4 Mini behave poorly with kids, let us know. Verified reports impact the next score update.
Report an issue