AI agents can investigate abuse, but reliable detection requires smart routing, classifier feedback loops, audit trails, and strict tool permissions.