Requirements
Implement adversarial testing program to validate system resilience against adversarial inputs and prompt injection attempts in line with adversarial threat taxonomy
Implement monitoring capabilities to detect and enable responding to adversarial inputs and prompt injection attempts
Implement controls to prevent over-disclosure of technical information about AI systems and organizational details that could enable adversarial targeting
Implement safeguards to prevent probing or scraping of external AI endpoints
Implement real-time input filtering using automated moderation tools
Implement safeguards to prevent AI agents from performing actions beyond intended scope and authorized privileges
Establish and maintain user access controls and admin privileges for AI systems in line with policy
Implement security measures for AI system deployment environments including encryption, access controls and authorization
Implement output limitations and obfuscation techniques to safeguard against information leakage
Implement safeguards to promote secure patterns and prevent known vulnerabilities in generated code