Harmful Content Detection for an LLM
Published 3 October 2026
ML System Design — Harmful Content Detection for an LLM¶
The ML design round asked me to design a system for detecting harmful content in an LLM-based product. The discussion was broader than a traditional social-media content-moderation problem. It appeared to combine classic integrity concerns with problems that arise specifically when the content is generated or processed by an LLM.