AWS practice questions: 1100 questions with full explanations
- Questions on the exam
- about 85 — vendor indicates, no fixed count published
- Time allowed
- 170 minutes format →
1100 practice questions for AWS Certified Machine Learning Engineer – Associate MLA-C02, grouped by exam domain. Every question below shows all four options, which one is correct, and why each of the other three is not — the wrong answers are where most candidates lose marks.
Not sure where you stand? Take the free 5-min AWS readiness check →
AWS certification: requirements, cost and exam format → · AWS exam format → ·
Questions by domain
- 1: Data Preparation for ML and AI — 308 questions →
- 2: ML Model and Foundation Model (FM) Development — 264 questions →
- 3: Deployment and Orchestration of ML and AI Workflows — 264 questions →
- 4: Operating, Monitoring, and Securing ML and AI Solutions — 264 questions →
Sample questions
Store encrypted Parquet objects in S3 and scan only: Which approach best meets these requirements?
- Store encrypted Parquet files on EFS for shared training access.EFS supports shared filesystem access, but the stated durable analytical object workload is better suited to S3.
- Store encrypted CSV files in S3 and scan every field during training.CSV is row-oriented, and scanning every field prevents efficient column pruning for selective analytical queries.
- Store encrypted Parquet objects in S3 and scan only required columns. ✓S3 provides durable object storage, Parquet supports column pruning, and server-side encryption protects the objects.
- Store encrypted Parquet files on EBS volumes attached to training instances.EBS is attached block storage and is less suitable than S3 for a durable, shared, historical object dataset.
All 308 1: Data Preparation for ML and AI questions →
Supervised fine-tuning on the labeled multilingual cases: What should it evaluate next?
- The previously tested few-shot prompt without new examples.Repeating an unchanged failed prompt does not introduce the task adaptation requested by the scenario.
- Retrieval of the same unchanged category definitions.The stem states that this approach was already evaluated and failed the required consistency threshold.
- A larger output-token allowance for each classification.A higher output allowance does not directly teach the desired category boundaries or terminology.
- Supervised fine-tuning on the labeled multilingual cases. ✓Labeled task examples can adapt classification behavior after the stated prompting and retrieval approaches failed validation.
All 264 2: ML Model and Foundation Model (FM) Development questions →
Use a multi-container inference pipeline: Which strategy should it choose?
- Deploy one real-time classifier and duplicate preprocessing logic in every client application.Client-side preprocessing does not provide one managed ordered path and can create inconsistent transformations.
- Use a multi-container inference pipeline. ✓A multi-container inference pipeline sequences compatible preprocessing and inference containers within one managed endpoint.
- Connect separate asynchronous preprocessing and classification endpoints through a queued workflow.Separate asynchronous stages do not satisfy the required synchronous combined response.
- Place preprocessing and classification models in a multi-model endpoint for sequential execution.Multi-model endpoints select among models; they do not inherently sequence preprocessing and classification containers.
All 264 3: Deployment and Orchestration of ML and AI Workflows questions →
Publish separate loop and truncation metrics: What monitoring design should be implemented?
- Publish separate loop and truncation metrics, with request-correlated logs for investigation. ✓Separate metrics support independent thresholds, while correlated request logs preserve diagnostic context.
- Use model-quality drift monitoring to identify tool loops before instrumenting tool execution metrics and request logs.Model-quality metrics do not directly measure orchestration loops or truncation, and telemetry is needed to investigate those events.
- Alarm on total requests and review individual assistant requests only after users report failures.Total request volume does not distinguish loops from truncations and delays detection until users notice failures.
- Store tool responses with request identifiers but combine loop and truncation counts into one daily metric.Identifiers help tracing, but one combined metric prevents separate thresholds and obscures which failure is increasing.
All 264 4: Operating, Monitoring, and Securing ML and AI Solutions questions →
Distribute objects across additional prefixes and adjust: What should the engineer do first?
- Convert all objects to EFS files without changing the ingestion readers.Changing storage services alone does not diagnose the concentrated request pattern or guarantee compatible reader behavior.
- Distribute objects across additional prefixes and adjust the reader concurrency gradually. ✓Spreading requests across prefixes and tuning concurrency addresses concentrated request load while validating ingestion behavior.
- Increase the model endpoint instance count before investigating ingestion metrics.Endpoint capacity does not resolve S3 prefix throttling occurring before data reaches the model.
- Increase worker memory because throttling responses indicate insufficient processing memory.Throttling responses identify service request pressure, not a worker-memory shortage, especially with normal CPU and object sizes.
All 308 1: Data Preparation for ML and AI questions →
Train a supervised image classifier with an explanation: Which approach best meets these requirements?
- Use unsupervised anomaly detection.Anomaly detection finds unusual examples but does not directly classify the predefined defect categories represented by the labels.
- Use retrieval over product manuals for image decisions.Retrieval supplies document context but does not provide reliable category-specific visual defect classification or rejection evidence.
- Use a generative image model to describe defects.Image generation or description is unnecessary for fixed-label classification and can add cost and latency.
- Train a supervised image classifier with an explanation method such as saliency maps for its predefined defect labels. ✓Supervised classification uses the labeled categories, while an explanation method supports inspection decisions and economical inference.
All 264 2: ML Model and Foundation Model (FM) Development questions →
Use an asynchronous inference endpoint with configured: Which strategy is most appropriate?
- Use a real-time endpoint and wait synchronously.Synchronous real-time inference keeps the client waiting and is unsuitable for delayed completion.
- Use serverless inference with a substantially increased client timeout for each document.A longer timeout does not provide the queued request and deferred-result pattern required here.
- Use an asynchronous inference endpoint with configured output handling and notifications. ✓Asynchronous inference queues supported long-running or large requests and delivers results through configured output handling.
- Submit each document as a separate batch transform job and notify users after completion.Batch transform is intended for offline datasets, not individual delayed requests requiring endpoint-style submission.
All 264 3: Deployment and Orchestration of ML and AI Workflows questions →
Compare recent feature distributions with a training: Which monitoring approach best detects potentially harmf
- Compare current device and region counts with operational capacity thresholds each day.Capacity thresholds can reveal infrastructure pressure, but they do not compare production feature distributions with the training population.
- Compare recent feature distributions with a training baseline using scheduled data-drift monitoring. ✓Scheduled baseline comparisons can detect changes in device and region distributions before ground-truth outcomes are available.
- Run an A/B experiment to determine whether production feature proportions match the original training distribution before changing the model.A/B testing compares model variants or user experiences; it is not the appropriate mechanism for comparing production inputs with a training baseline.
- Wait for labels and measure recommendation accuracy after outcomes are reported.Waiting for labels delays detection and measures model quality rather than current input-distribution change.
All 264 4: Operating, Monitoring, and Securing ML and AI Solutions questions →
AWS exam: the facts
How many questions are on the AWS exam?
Around 85. The vendor does not publish a fixed count for AWS, so this is the figure it indicates rather than a guaranteed number.
How long is the AWS exam?
170 minutes. Across 85 questions that is about 120 seconds per question.
What topics does the AWS exam cover?
4 domains: 1: Data Preparation for ML and AI, 2: ML Model and Foundation Model (FM) Development, 3: Deployment and Orchestration of ML and AI Workflows, 4: Operating, Monitoring, and Securing ML and AI Solutions. Weights: 1: Data Preparation for ML and AI 28%, 2: ML Model and Foundation Model (FM) Development 24%, 3: Deployment and Orchestration of ML and AI Workflows 24%, 4: Operating, Monitoring, and Securing ML and AI Solutions 24%.
How many AWS practice questions does Certsqill have?
1100, spread across 4 exam domains. Every one shows all options, which is correct, and why each of the others is not.
Would you pass AWS today?
Five minutes, and you get a score per domain — not one number, but which section to open tonight.
Test your AWS readiness — free