A toxicity scorer is an automated tool or machine learning model that analyzes a given text and outputs a numerical score reflecting the degree or likelihood of harmful, offensive, or abusive language. Within natural language processing and content moderation, these systems are trained to identify attributes such as hate speech, harassment, profanity, insults, and identity attacks. They are widely used to flag inappropriate material in online communities, clean training datasets for artificial intelligence, and measure the effectiveness of safety interventions and detoxification techniques in generative language models.