6. Delivered
Minor ML Update: Context-aware toxicity filtering
Key Dates
- Frontier: September 2 2026
- Standard: September 8 2026
- Basic: September 10 2026
Learn more about our model upgrades here.
We’re upgrading the AI Assistant’s toxicity filtering to consider the full context of a conversation rather than evaluating content in isolation. This provides a more accurate understanding of user intent and helps identify genuinely harmful content. The upgrade also reduces false positives, allowing safe conversations to continue without unnecessary interruption. Administrators can review the configuration steps for more information: https://help.moveworks.com/agent-studio/agentic-ai/conversational-safeguards-assistant/configure-moveworks-toxcity-filter#configuration-steps