Which ML Methodology for Customer Tiers with Unlabeled Data?
A company has petabytes of unlabeled customer data to use for an advertisement campaign. The company wants to classify its customers into tiers to advertise and promote the company's products. Which methodology should the company use to meet these requirements?
Community Votes
100% of anonymous learners picked answer B. Votes are pick records left by other test-takers — they are not the verified answer.
Community Insight
This question tests the core difference between supervised and unsupervised learning, and the trap is assuming 'classification' always requires labeled training data.
On the AIF-C01 exam, classifying customers into tiers using petabytes of unlabeled data points to unsupervised learning. Community consensus is that clustering techniques group unlabeled customer data without predefined labels.
A. Supervised learning is the most common incorrect choice because classification tasks are often associated with supervised algorithms, but supervised learning requires labeled data, which is absent here.
Community Discussion (8 comments)
Comments & Corrections
No comments yet — spotted an error or have a note? Share it below.
Expert Analysis
Why the Answer Is Correct
Unsupervised learning is designed for unlabeled datasets, discovering hidden patterns and groupings without predefined labels. As commenters noted, the keyword 'unlabeled data' directly points to unsupervised learning, and clustering techniques like K-Means or hierarchical clustering can classify customers into tiers. Comments 1 and 2 highlight that grouping similar customers into clusters perfectly meets the advertisement campaign requirement.Why the Other Options Are Wrong
A supervised learning requires labeled input-output pairs, but the data has no labels, so it cannot be used directly. C reinforcement learning uses rewards and penalties in an environment, not suitable for static customer data. D RLHF is a specialized fine-tuning method for language models, not for customer tier classification. Comments show unanimous support for B, reinforcing that other options conflict with the unlabeled-data constraint.Community Comment Notes
The community vote is 100% for B, with many comments emphasizing the 'unlabeled data' keyword. Comment 7 specifically says 'Keyword - Unlabeled Data' as the deciding factor. Comments 1 and 2 provide detailed explanations that clustering is the appropriate unsupervised technique for this scenario.Official Reference
Exam Strategy
On exam day, scan for key data terms like 'unlabeled' or 'labeled.' If the question says unlabeled data and the goal is grouping, pick unsupervised learning immediately; then use clustering as your justification in review.
Related Analysis
Practice All AIF-C01 Questions
Access 100 questions with complete answers and detailed explanations.
View Full AIF-C01 Practice Test →