Red teaming that deliberately stress-tests the system: Which activity best fits this specific goal?
Red teaming proactively stress-tests the system with adversarial and edge-case inputs to expose failure modes before harm occurs.
The question
A safety-critical medical-triage AI is in production, and the governance team wants a periodic activity that proactively probes the system with adversarial inputs and unexpected scenarios to surface failure modes before real patients are affected. Which activity best fits this specific goal?
Preparing for AIGP? Take the free 5-min readiness quiz →
- A financial audit that verifies the accuracy of the accounting entries associated with operating and maintaining the triage system in production.Plausible but wrong: a financial audit checks accounting records, not the model's resilience to adversarial or unexpected inputs.
- Threat modeling that maps potential attack paths and abuse cases against the system's architecture to prioritize where defenses are needed.Plausible but wrong: threat modeling is a valuable analytical exercise, but it maps risks on paper rather than actively probing the live system with inputs.
- Red teaming that deliberately stress-tests the system with adversarial and edge-case inputs to surface failure modes before they cause harm. ✓Correct: red teaming proactively probes a system with adversarial and edge-case inputs to reveal failure modes, matching the stated goal precisely.
- User acceptance testing that confirms clinicians find the interface usable and that outputs are presented in the expected on-screen format.Plausible but wrong: acceptance testing checks usability and presentation, not adversarial resilience or hidden failure modes.
The trap
Confusing threat modeling (mapping risks analytically) with red teaming (actively probing the system). How to remember it
Red teaming proactively stress-tests the system with adversarial and edge-case inputs to expose failure modes before harm occurs.
How many of these would you get right?
One of 1581 AIGP questions on Certsqill. Take a free five-minute check and see your score per domain — not one number, but which section to open tonight.
Test your AIGP readiness — freeMore Understanding How to Govern AI Development questions
- Treat the input shift as data drift and trigger scheduled: Which maintenance action does this pattern most →
- Log and document each incident with root cause: Which action best satisfies both aims? →
- Model or data drift: Which contributing factor best describes this cause? →
- All 426 Understanding How to Govern AI Development questions →
Part of the Certsqill AIGP question bank · Understanding How to Govern AI Development ·
Every answer, right and wrong, comes with its own explanation.