
Salesforce is expanding its Einstein Trust Layer with a new toxicity detection system. It monitors and filters out harmful or inappropriate content in customer conversations. This feature is becoming extremely relevant in today’s time when more and more companies are adopting AI-generated customer interactions.
The feature addresses a key risk associated with AI deployments: large language models (LLMs), while powerful, can still produce offensive or unsafe responses. That’s why Salesforce’s latest solution is designed to catch those moments before they escalate into real-world consequences like customer complaints, brand damage, or legal trouble.