Screening instrument / 15 signals

    AI Failure Score™

    Understand your AI system's exposure across reliability, security, safety, observability, and human control.

    This assessment is a screening indicator only. It does not constitute a security audit, penetration test, certification, or regulatory compliance determination. It does not prove an AI system is safe.

    01 / RELIABILITY

    Does your AI behave consistently under real-world conditions?

    Q1. Does your AI system have defined pass/fail criteria for its critical workflows? *
    Q2. Has the AI been tested against a representative set of real user requests? *
    Q3. Do you track whether AI responses stay consistent across model updates or prompt changes? *

    02 / SECURITY

    Has your AI been evaluated for adversarial manipulation and data exposure?

    Q4. Has your AI been tested for prompt injection — attempts to override its instructions? *
    Q5. Does the system restrict access to sensitive data based on user identity or role? *
    Q6. Has your system prompt or configuration been evaluated for exposure to end users? *

    03 / SAFETY

    Does your AI have documented prohibited behaviors and guardrails?

    Q7. Does your system have documented outputs and actions it must never produce? *
    Q8. Is there a tested process to prevent the AI from generating harmful or inappropriate content? *
    Q9. Have you tested what happens when users deliberately attempt to manipulate the AI? *

    04 / OBSERVABILITY

    Can you see what your AI is doing in production?

    Q10. Do you have logging or tracing active for your AI system's inputs and outputs? *
    Q11. Can you identify which model version or prompt version produced a specific output? *
    Q12. Do you have alerts or monitoring for unexpected AI behavior in production? *

    05 / HUMAN CONTROL

    Can humans intervene, override, or shut down the AI when needed?

    Q13. Are there human approval checkpoints before the AI takes high-stakes actions? *
    Q14. Can you roll back or disable the AI system quickly if a problem is discovered? *
    Q15. Is there a defined escalation process when the AI produces a problematic output? *