| Prompt Integrity Check | Detects embedded instructions intended to manipulate or override
expected agent behavior. |
| Arbitrary Code Execution Check | Detects executable code or patterns capable of performing actions
beyond the skill's documented purpose. |
| Secrets Disclosure Check | Detects unauthorized access to credentials, API keys, tokens, and
other sensitive information. |
| Data Exfiltration Check | Detects outbound communication or data transfers that may expose
sensitive information. |
| Obfuscated Behavior Check | Detects actions performed without appropriate visibility, logging, or
user awareness. |
| Behavior Integrity Check | Detects modifications to agent configuration, instructions, or
persistent state. |
| Excessive Permissions Check | Detects autonomous behaviors or capabilities that exceed expected
operational boundaries. |