Purpose
Protect the boundary
Detect risk before a message or action is executed.
Protect the boundary by flagging risky language and forcing approval when needed.
Detect risk before a message or action is executed.
Reviews outgoing content, selected actions, and sensitive phrases.
Returns a decision and the reason for it.
Always defer to policy, jurisdiction, and human review for sensitive cases.