AI Trace
Content Moderation

Reviewed and published by trentmaziarz, May 11, 2026. Discovered and drafted by our automated research pipeline.

Sony PlayStation uses machine learning classifiers for content moderation on its platform and has worked to address bias in these systems, specifically the problem of 'celebratory identity speech' being incorrectly flagged as discriminatory content.

Details

A machine learning engineer at Sony PlayStation presented research at the Trust and Safety Research Conference at Stanford on the topic of bias mitigation in content moderation classifiers. The specific issue addressed was 'celebratory identity speech' — text that mentions protected groups in a positive context but is incorrectly scored as likely discriminatory by classifiers. This confirms the existence of ML-based content moderation classifiers at PlayStation, though no further details about their scope, deployment scale, or the full range of content they moderate were disclosed in available primary sources.

Products affected

PlayStation Network

Sources & Evidence

Cite this record

Trace Foundation. (2026). Sony Interactive Entertainment: Sony PlayStation uses machine learning classifiers for content moderation on its platform and has worked to address bias in these systems, specifically the problem of 'celebratory identity speech' being incorrectly flagged as discriminatory content (data as of 2026-05-11) [Data set record]. AI Trace. https://www.aitrace.org/r/practice/7c9fe255-90ad-46a1-b8f0-6ca667f488ac. Accessed October 5, 2026.

Stable link
https://www.aitrace.org/r/practice/7c9fe255-90ad-46a1-b8f0-6ca667f488ac
Data as of
May 11, 2026
Last verified
Not recorded

How to cite AI Trace

Other practices by Sony Interactive Entertainment

Have evidence about Sony Interactive Entertainment's AI practices? Submit a report.

Report a Sighting →