Supervised learning plus reinforcement learning | AIGP
7-day money-back guarantee — full refund within 7 days of purchase if you've completed under 20% of the questions. See pricing →
Certifications Tools Flashcards Career Paths Exam Guides Blog Pricing For Teams About

Language

✓ EnglishDeutschEspañolFrançaisPortuguês
Check readiness — free →

Supervised learning plus reinforcement learning: What classification is supported?

AIGP Understanding the Foundations of AI Governance Hard

Labels indicate supervised learning, rewards indicate reinforcement learning, and pilot success does not establish broad trustworthiness.

The question

A retailer tests two demand-planning systems. System A learns from historical sales records paired with known demand quantities. System B repeatedly proposes inventory actions and receives reward feedback based on stockouts and excess inventory. Both systems perform well in a pilot, but neither pilot establishes fairness or reliability across all stores. What classification is supported?

Preparing for AIGP? Take the free 5-min readiness quiz →

  1. Classify System A as unsupervised and System B as supervised learning
    Known demand quantities are labels for System A, while reward feedback identifies System B with reinforcement learning.
  2. Supervised learning plus reinforcement learning
    System A uses labeled demand quantities, and System B uses reward feedback. The pilot does not support broad fairness or reliability claims across stores.
  3. Classify both systems as supervised learning from business data
    System B is described as learning through reward feedback, not labeled target examples; business data alone does not establish supervised learning.
  4. Treat pilot performance as proof of fairness and reliability across stores
    Good pilot performance is evidence about the tested setting, not proof of fairness or reliability across all stores.
The trap
Separate how a system learns from what its evaluation actually demonstrates.

How to remember it

Labels indicate supervised learning, rewards indicate reinforcement learning, and pilot success does not establish broad trustworthiness.

How many of these would you get right?

One of 1581 AIGP questions on Certsqill. Take a free five-minute check and see your score per domain — not one number, but which section to open tonight.

Test your AIGP readiness — free

More Understanding the Foundations of AI Governance questions

Part of the Certsqill AIGP question bank · Understanding the Foundations of AI Governance · Every answer, right and wrong, comes with its own explanation.