AI Cert Prep
Type to search documentation.

CCAO-F Practice Exam 2

A second full-length, blueprint-weighted practice exam for the Claude Certified Associate – Foundations credential, with new, harder items.

This is a second full-length, blueprint-weighted practice exam for CCAO-F. All 60 items are new – they do not repeat Practice Exam 1 or the domain-page questions. Exam 2 is deliberately slightly harder: more stems carry multiple constraints at once, and more use FIRST, MOST cost-effective, and TWO qualifiers, so you must weigh competing considerations rather than spot a single keyword. Distractors lean on the classic patterns – constraint-blind, over-engineered, prompt-as-enforcement, self-report reliance, silent failure, recall-only and aggregate-metric.

Instructions

  • Time: 120 minutes, matching the real exam.
  • Items: 60, multiple-choice and multiple-response. Each item states how many answers to select.
  • Scoring: scaled 100–1000, pass 720/1000. There is no guessing penalty. Aim for ≥ 80% raw (≈ 48/60) before booking.
  • Rule: answer every question. For multiple-response items you must select all correct options and no incorrect ones.
  • Work each question before expanding the answer.

Domain distribution

#DomainWeightItems here
1Prompting and Task Execution14%8
2Output Evaluation and Validation21%13
3Product and Model Selection12%7
4Workflow Integration and Solution Design16%10
5Configuration and Knowledge Management12%7
6Governance, Risk, and Responsible Use15%9
7Troubleshooting and Optimization10%6

Total: 8 + 13 + 7 + 10 + 7 + 9 + 6 = 60 items.

How to use both exams

  • Sit Practice Exam 1 first, untimed if you are still learning, to surface weak domains.
  • Study the domain pages for any domain where you scored below ~75%, paying attention to the “Common misconceptions” tables and scenario walkthroughs.
  • Use Practice Exam 2 as your booking gate: sit it timed (120 minutes) under exam conditions. If you clear ≥ 48/60 here, with no domain badly lagging, you are ready to book.
  • Compare per-domain results between the two exams. A domain that is strong on Exam 1 but weak on the harder Exam 2 is a shallow-understanding signal – revisit the decision tables and distractor patterns for that domain.

Score interpretation

Raw score (of 60)Approx. scaledInterpretation
54–60900–1000Exam-ready; strong across all domains
48–53800–890Solid pass zone; review weak domains
43–47720–790Marginal pass; targeted revision advised
36–42600–710Below pass; revisit heavy domains (D2, D4, D6)
< 36< 600Study all domains before re-attempting

Because Exam 2 runs harder than Exam 1, treat a comfortable pass here as a stronger readiness signal than the same score on Exam 1.

Take the practice exam

The interactive mode runs a timed sitting one question at a time and ends with your score, a per-domain breakdown and a full correction; the review mode underneath lists every question with its options one per line and the answer hidden until you ask for it.

Interactive mode

Take the practice exam

60 questions · one at a time · 120-minute countdown · results with per-domain breakdown and full correction at the end. Your progress is saved in this browser if you leave the page.

All questions (review mode)

Options are listed one per line. The answer and explanation stay hidden until you click Show answer. Use the interactive mode above for a timed sitting.

  1. Q1D1 · Prompting and Task ExecutionSelect one

    An associate pastes a customer email and types in one run-on line: 'reply to this and also tell me whether we broke the SLA'. Claude drafts a warm reply but never mentions the SLA. What is the BEST FIRST fix?

    • A. Switch to a more capable model and resend the same text.
    • B. Label the parts (DOCUMENT / SLA / TASKS), number the two tasks, and put the request after the source so both are answered.
    • C. Regenerate several times until it happens to answer both.
    • D. Tell Claude to 'be more thorough'.
    Show answer

    Answer: B.

    The two asks were jumbled with the source text, so one was dropped; separating and labelling the parts and ordering the tasks fixes the root cause. A bigger model (A) still faces the jumble (constraint-blind); regenerating (C) is blind variation; 'be more thorough' (D) is a wish, not a specification (prompt-as-effort).

  2. Q2D1 · Prompting and Task ExecutionSelect one

    A four-part request (research competitors, write a brief, draft headlines, build a timeline) consistently returns a strong brief but a weak, generic timeline. Which approach is MOST appropriate?

    • A. Add 'and make the timeline better' to the same prompt.
    • B. Run the timeline as its own step after the brief, specifying weeks, milestones and dependencies as a table.
    • C. Switch the whole task to Opus 5.
    • D. Ask for a longer overall response.
    Show answer

    Answer: B.

    Uneven quality on a bundled ask is the overloaded-prompt pattern; isolating the weak part as its own well-specified step fixes it. 'Make it better' (A) is a non-actionable wish; a bigger model (C) does not address bundling (over-engineered); more words (D) is the wrong lever.

  3. Q3D1 · Prompting and Task ExecutionSelect two

    An operations lead needs figures pulled from an attached 30-page report and wants to minimise both fabrication AND unusable formatting. Which TWO prompt techniques are BEST?

    • A. 'Use only the attached report; if a figure is not stated, write "not stated" — do not estimate.'
    • B. Specify the exact output schema: a table with columns Metric \| Prior \| Current \| % change.
    • C. 'Be accurate and well-organised.'
    • D. Use the most expensive model.
    • E. Request the maximum response length.
    Show answer

    Answer: A and B.

    Grounding to the document with a missing-value rule reduces fabrication, and a defined schema fixes formatting. 'Be accurate and well-organised' (C) is not actionable; model tier (D) does not enforce grounding (over-engineered); maximum length (E) is the wrong lever.

  4. Q4D1 · Prompting and Task ExecutionSelect one

    A new joiner opens a brand-new chat and types 'continue the analysis from yesterday', and Claude has no idea what they mean. What is the correct explanation and fix?

    • A. The model forgot on purpose; upgrade to a bigger model.
    • B. New chats do not retain a previous conversation; re-supply the material, or keep it in a Project so it is available.
    • C. Increase the response length so it can recall more.
    • D. Ask 'are you sure you don't remember?'
    Show answer

    Answer: B.

    A fresh chat has no memory of a prior one unless context is persisted in a Project or Memory; the fix is to supply or persist it. Model tier (A) does not restore missing context; length (C) is irrelevant; a self-check (D) is self-report reliance and cannot recover data that is not present.

  5. Q5D1 · Prompting and Task ExecutionSelect one

    A brainstorming prompt keeps returning three safe, similar event themes when the associate needs a wide range. Which single change helps MOST?

    • A. Ask for 'better, more creative ideas'.
    • B. Explicitly request 15 distinct ideas spanning different styles, with no filtering yet, then narrow in a follow-up.
    • C. Switch to Haiku for speed.
    • D. Ask for the single best idea up front.
    Show answer

    Answer: B.

    Brainstorming rewards explicit quantity and diversity before convergence; stating '15 distinct, different styles, no filtering' widens the divergent phase. 'Better' (A) is a non-actionable wish; model tier (C) is irrelevant; asking for one idea (D) collapses divergence — a task-type mismatch.

  6. Q6D1 · Prompting and Task ExecutionSelect one

    Despite a written description of the required layout, a recurring report is formatted slightly differently each time. What is the MOST reliable fix?

    • A. Describe the format in even more words.
    • B. Provide one perfect worked example of the exact layout for Claude to match.
    • C. Ask Claude to 'be consistent'.
    • D. Increase the temperature setting.
    Show answer

    Answer: B.

    When a described format will not hold, one concrete example (few-shot) pins the pattern far more reliably. More words (A) and 'be consistent' (C) are not actionable; a higher temperature (D) increases variation, worsening the problem.

  7. Q7D1 · Prompting and Task ExecutionSelect one

    A request is ambiguous — 'prepare the quarterly update' — with no audience, scope, or format named. Which approach BEST prevents wasted rework?

    • A. Generate immediately and fix problems afterwards.
    • B. Ask Claude to list what it needs to know (audience, scope, format, timeframe) and confirm before generating.
    • C. Generate three long versions and pick one.
    • D. Use the maximum response length to cover all bases.
    Show answer

    Answer: B.

    Resolving ambiguity up front avoids generating on a wrong assumption. Generating first (A, C) risks solving the wrong problem and wastes iterations; maximum length (D) is unrelated to the missing specifications.

  8. Q8D1 · Prompting and Task ExecutionSelect one

    An extraction prompt asks Claude to 'pull all action items' from meeting notes, but the output omits owners and due dates. What is the BEST refinement?

    • A. Specify the exact fields: a table of Action \| Owner \| Due date, using 'unassigned' where no owner is named.
    • B. Ask for a longer answer.
    • C. Turn on extended thinking.
    • D. Paste the notes a second time.
    Show answer

    Answer: A.

    Extraction needs an explicit schema and a rule for missing values, which makes it complete and checkable. Length (B) does not define the missing fields; extended thinking (C) does not add a schema; re-pasting (D) changes nothing about the fields requested.

  9. Q9D2 · Output Evaluation and ValidationSelect one

    A board pack is due in an hour. It reads fluently, but the segment table's rows do not sum to the headline total AND one of five citations cannot be found anywhere. What should the analyst do FIRST?

    • A. Send it; it reads well and time is short.
    • B. Recompute the total from the rows, since an internal-consistency failure means the table cannot be trusted as-is and needs no external source.
    • C. Ask Claude in the same chat whether the numbers are correct.
    • D. Lower the temperature and regenerate the whole pack.
    Show answer

    Answer: B.

    The cheapest, highest-value first move is the internal-consistency recompute, which needs no external source. Sending on fluency (A) is the fluency trap; a same-session self-check (C) reuses the biased reasoning; lowering temperature and regenerating (D) changes variance, not grounding, and wastes the hour.

  10. Q10D2 · Output Evaluation and ValidationSelect one

    Claude cites a study with a plausible title, real-sounding authors and a real journal name to support a key claim. Before relying on it, what is the MOST appropriate validation?

    • A. Accept it; the journal name is real.
    • B. Open the cited source and confirm it exists and actually says what Claude claims.
    • C. Ask Claude in the same chat to provide the DOI and accept it.
    • D. Regenerate to see whether the same citation appears again.
    Show answer

    Answer: B.

    Fabricated citations blend real-sounding elements, so the only valid check is opening the source and confirming it supports the claim. A real journal name (A) proves nothing; a same-chat DOI (C) can also be fabricated (self-report reliance); regeneration (D) tests stability, not existence.

  11. Q11D2 · Output Evaluation and ValidationSelect one

    A fully accurate three-page analysis is rejected by a CFO who says she cannot find the decision. Which lens failed, and what is the fix?

    • A. Accuracy failed; open external sources to re-check the facts.
    • B. Relevance/audience fit failed; restructure to lead with the recommendation and the ask, moving method to a short appendix.
    • C. Consistency failed; recompute the totals.
    • D. Bias failed; balance the examples.
    Show answer

    Answer: B.

    The facts are correct but the output is mis-pitched for an executive — a relevance/audience failure, fixed by leading with the decision. Accuracy (A) already passed; there is no totals problem (C) or asymmetry (D) described, so those lenses are distractors.

  12. Q12D2 · Output Evaluation and ValidationSelect two

    A dashboard shows 92% overall accuracy for a Claude-assisted tagging task and a manager wants to expand it. Which TWO checks are MOST important FIRST?

    • A. Break accuracy down per category to find any segment that is failing.
    • B. Confirm the failing categories are not the highest-stakes ones.
    • C. Report only the 92% headline to leadership.
    • D. Assume 92% holds uniformly across all categories.
    • E. Switch to a cheaper model to save cost.
    Show answer

    Answer: A and B.

    An aggregate can hide a category that fails badly, so per-segment breakdown and checking whether failures concentrate in high-stakes categories are essential. Reporting only the headline (C) and assuming uniformity (D) are aggregate-metric traps; cost (E) is unrelated to validation.

  13. Q13D2 · Output Evaluation and ValidationSelect one

    A colleague says an internal-only memo's revenue figure 'doesn't need checking because it's internal' — but it will feed next year's hiring plan. What is the BEST response?

    • A. Agree; internal content does not need verification.
    • B. Verify the figure, because its downstream use in a hiring decision makes it consequential regardless of the 'internal' label.
    • C. Only check the tone of the memo.
    • D. Ask Claude in the same chat whether it is confident in the figure.
    Show answer

    Answer: B.

    Stakes follow the downstream use, not the 'internal' label; a figure feeding hiring must be verified. 'Internal' does not exempt it (A); tone (C) is the wrong lens; a same-chat confidence check (D) is self-report reliance, not verification.

  14. Q14D2 · Output Evaluation and ValidationSelect one

    Two colleagues verify a high-stakes number by asking two different AI tools; both return the same value, so they treat it as confirmed. What is the FLAW?

    • A. None; agreement between two tools confirms accuracy.
    • B. Both tools can share the same wrong prior, so agreement is weak evidence; verify against an authoritative independent source.
    • C. They should have used the same tool twice instead.
    • D. They should have asked a third AI tool to break the tie.
    Show answer

    Answer: B.

    Two models agreeing is not independent verification because they can be wrong the same way; an authoritative source or human is needed. Agreement is not confirmation (A); repeating one tool (C) or adding a third AI (D) still lacks authority.

  15. Q15D2 · Output Evaluation and ValidationSelect one

    In research mode, an analyst spot-checks 2 of 8 citations; one is fine and one does not exist. What is the MOST appropriate next step?

    • A. Use the summary; 50% of the sample was fine.
    • B. Delete only the paragraph with the missing citation and keep the rest.
    • C. Treat the whole output as unverified: check every citation and re-verify any claim whose citation fails.
    • D. Ask Claude to invent a replacement citation for the missing one.
    Show answer

    Answer: C.

    A fabricated citation invalidates sample-based trust; move to exhaustive verification. Partial trust (A) and fixing only one paragraph (B) leave other fabrications in place; inventing a citation (D) compounds the problem.

  16. Q16D2 · Output Evaluation and ValidationSelect one

    A vendor comparison lists six strengths for Vendor A and six weaknesses for Vendor B. Which is the MOST likely issue and the BEST remedy?

    • A. Hallucination; request sources for every point.
    • B. Framing bias; ask for a symmetric comparison with equal-depth pros and cons for both against the same criteria.
    • C. Truncation; ask Claude to continue.
    • D. Inconsistency; recompute the totals.
    Show answer

    Answer: B.

    Asymmetric treatment of options is framing bias, fixed by a neutral, criteria-based symmetric structure. Sources (A) may help but do not fix the asymmetry; it is not truncation (C) or a totals problem (D).

  17. Q17D2 · Output Evaluation and ValidationSelect two

    Which TWO outputs should receive line-by-line human verification before use?

    • A. A document to be filed with a financial regulator.
    • B. Text that was generated in under ten seconds.
    • C. Patient-facing content containing medication dosage language.
    • D. An answer formatted as bullet points.
    • E. A response that is longer than one page.
    Show answer

    Answer: A and C.

    A regulatory filing and medical dosage content are high-stakes triggers requiring exhaustive review. Generation speed (B), bullet formatting (D) and length (E) are irrelevant to stakes.

  18. Q18D2 · Output Evaluation and ValidationSelect one

    A financial analyst asks Claude to compute a quarterly total from a pasted 500-row table, and it 'looks about right'. What is the MOST appropriate verification?

    • A. Trust it; arithmetic is deterministic for the model.
    • B. Recompute the total independently (analysis tool, spreadsheet or calculator) and compare.
    • C. Ask Claude 'are you sure?' in the same chat.
    • D. Round the total to reduce the error.
    Show answer

    Answer: B.

    Language models can err on long-table arithmetic, so financial figures should be recomputed independently. Arithmetic is not reliably deterministic for an LLM (A); a same-chat doubt (C) is self-report reliance; rounding (D) hides rather than verifies.

  19. Q19D2 · Output Evaluation and ValidationSelect one

    A translated policy renders the defined term 'data controller' three different ways. Which lens caught this, and what is the cheapest fix?

    • A. Accuracy; open external sources to check the translation.
    • B. Consistency; supply a glossary of fixed terms and request a terminology-check table.
    • C. Relevance; shorten the document.
    • D. Bias; ask for a neutral tone.
    Show answer

    Answer: B.

    Inconsistent rendering of a defined term is an internal-consistency failure, fixed with a glossary and a checkable terminology table. It is not an accuracy-sourcing (A), length (C) or bias (D) issue.

  20. Q20D2 · Output Evaluation and ValidationSelect one

    Claude answers three of four sub-questions in a brief thoroughly and silently omits the fourth, yet the brief reads as complete. Which lens catches this and how?

    • A. Accuracy; open external sources.
    • B. Completeness; re-read the original brief as a checklist and tick each required part.
    • C. Bias; balance the examples.
    • D. Relevance; change the audience.
    Show answer

    Answer: B.

    A silently dropped sub-question is a completeness failure, caught by checking the output against the brief as a checklist. Accuracy sourcing (A), bias balancing (C) and audience changes (D) do not detect a missing part.

  21. Q21D2 · Output Evaluation and ValidationSelect one

    Which output format is MOST appropriate for 300 customer records with fixed fields destined for a CRM import?

    • A. Inline prose.
    • B. A narrative artifact.
    • C. Structured data (CSV or JSON) with a specified schema and column order.
    • D. A slide deck.
    Show answer

    Answer: C.

    Machine-consumed, fixed-field data needs a specified structured format with the exact schema. Prose (A), a narrative artifact (B) and slides (D) are not importable structured data.

  22. Q22D3 · Product and Model SelectionSelect one

    A strategy team needs BOTH the same brand tone and reference docs across many chats AND a fixed weekly-report procedure run identically each time. Which combination is BEST?

    • A. A single very long prompt reused each week.
    • B. A Project for the shared tone and docs, plus a Skill for the repeatable report procedure.
    • C. Memory only.
    • D. A separate account for each task.
    Show answer

    Answer: B.

    Shared context is a Project's job; a repeatable multi-step procedure is a Skill's job — the two are complementary. A long prompt (A) and Memory (C) cover neither well; separate accounts (D) fragment the setup and governance.

  23. Q23D3 · Product and Model SelectionSelect two

    An associate must accurately sum a 4,000-row sales spreadsheet AND obtain a current, sourced view of a new regulation. Which TWO surfaces are correct, respectively?

    • A. The code / analysis tool for the spreadsheet maths.
    • B. Research mode / web search for the current sourced regulation.
    • C. Plain chat on built-in knowledge for the regulation.
    • D. Research mode for the spreadsheet totals.
    • E. An artifact to compute the spreadsheet maths.
    Show answer

    Answer: A and B.

    Reliable arithmetic over many rows is executed by the analysis tool; current sourced facts come from research mode with citation checks. Built-in knowledge (C) may be stale; research mode does not do arithmetic (D); an artifact (E) is a container, not a compute engine.

  24. Q24D3 · Product and Model SelectionSelect one

    A subject-line rewrite must run thousands of times a day at the lowest cost and latency. Which model AND thinking setting is MOST cost-effective?

    • A. Opus 5 with extended thinking on.
    • B. Haiku 4.5 with extended thinking off.
    • C. Fable 5.1 with thinking on.
    • D. Sonnet 5 with extended thinking on.
    Show answer

    Answer: B.

    A trivial, high-volume, latency-sensitive task wants the cheapest fast model with no extra thinking. Opus and Fable (A, C) over-buy cost and latency; extended thinking on Sonnet (D) adds cost for no quality gain here.

  25. Q25D3 · Product and Model SelectionSelect one

    A manager claims that switching from the web app to Claude Desktop will let the team safely use confidential data that policy currently disallows. Which statement is CORRECT?

    • A. Correct; the desktop app is inherently more secure.
    • B. Incorrect; data governance follows the account/plan, not the client app, so the policy still applies.
    • C. Correct; local apps bypass organisational policy.
    • D. Correct, provided the files are deleted afterwards.
    Show answer

    Answer: B.

    The plan/account determines governance and the no-training guarantee; changing client does not change what data is allowed. Desktop is not inherently more compliant (A, C), and deletion (D) does not alter policy permissions.

  26. Q26D3 · Product and Model SelectionSelect one

    A workload includes a complex, high-stakes board scenario analysis AND a routine batch of 5,000 ticket tags that must be produced cheaply. What is the BEST model strategy?

    • A. Use one model for both to keep it simple.
    • B. Use Opus/Fable with extended thinking for the board analysis, and Haiku for the ticket tagging.
    • C. Use Haiku for both to save the most money.
    • D. Use Opus for both to be safe.
    Show answer

    Answer: B.

    Match each task to its stakes: top tier plus thinking for hard high-stakes reasoning, cheapest fast model for routine high-volume tagging. One model for both (A) mismatches at least one task; Haiku for the board analysis (C) under-buys; Opus for tagging (D) over-buys.

  27. Q27D3 · Product and Model SelectionSelect one

    A user needs Claude to act on the web page they are currently viewing to help complete a long form. Which client is MOST appropriate?

    • A. The mobile app.
    • B. Claude for Chrome, which can act on the current page with per-site permission.
    • C. Claude Code.
    • D. Research mode in the web app.
    Show answer

    Answer: B.

    In-page assistance on the current site is what the Chrome extension provides, with permission. The mobile app (A) is for on-the-go capture; Claude Code (C) is developer tooling; research mode (D) does web synthesis, not acting on the open page.

  28. Q28D3 · Product and Model SelectionSelect one

    A user wants Claude to recall their personal writing preferences across sessions, while their team wants a shared, curated set of policy docs available to all. Which pairing is CORRECT?

    • A. Memory for the personal preferences; a Project's knowledge files for the shared curated docs.
    • B. A Project for the personal preferences; Memory for the shared docs.
    • C. Both stored in Memory.
    • D. Both stored in one personal Project.
    Show answer

    Answer: A.

    Memory handles personal cross-session recall; curated shared documents belong in a Project's knowledge. Swapping them (B), putting shared docs in personal Memory (C), or hiding shared docs in a personal Project (D) mis-scopes the need.

  29. Q29D4 · Workflow Integration and Solution DesignSelect two

    A support lead wants Claude to auto-send simple refund confirmations 'because it's confident on those' and to connect the whole shared inbox for convenience. Which TWO corrections are MOST appropriate?

    • A. Keep a human approval gate before any refund confirmation is sent, because it is irreversible and financial.
    • B. Scope the connector to only what the task needs, since the inbox contains customer PII.
    • C. Allow auto-send only when Claude rates its own confidence as high.
    • D. Connect the whole inbox so nothing is missed.
    • E. Use a bigger model so no gate is needed.
    Show answer

    Answer: A and B.

    Irreversible financial sends need a human gate, and connectors must follow least privilege because the inbox holds PII. Self-rated confidence (C) is self-report reliance, not a valid gate; whole-inbox connection (D) over-exposes; model size (E) does not remove risk.

  30. Q30D4 · Workflow Integration and Solution DesignSelect one

    A pilot cut handling time 40% with steady overall CSAT, and the manager wants to roll it out to all 12 teams next week. What should happen FIRST?

    • A. Roll out to all teams immediately; the numbers are good.
    • B. Check quality per task type and reviewer-edit rates, then define data rules, review gates and ownership before phasing the rollout.
    • C. Switch everyone to the most expensive model.
    • D. Remove human review to move faster.
    Show answer

    Answer: B.

    Time and aggregate CSAT can hide a weak segment; verify per-type quality and put governance in place before a phased rollout. Immediate big-bang rollout (A) skips both; a pricier model (C) is unrelated; removing review (D) is unsafe.

  31. Q31D4 · Workflow Integration and Solution DesignSelect one

    Which need CLEARLY crosses the escalation boundary to Developers/Architects?

    • A. A recurring monthly report with the same reference docs, drafted by a person in chat.
    • B. A pipeline that ingests API webhooks, extracts fields, and updates a database automatically 24/7 with no human per run.
    • C. Iterating on a proposal in an artifact.
    • D. Drafting personalised outreach with a scoped CRM connector.
    Show answer

    Answer: B.

    Automatic, system-to-system, unattended 24/7 processing needs engineered integration — Developer/Architect territory. The others keep a human in the loop and stay in the business-user zone.

  32. Q32D4 · Workflow Integration and Solution DesignSelect two

    Mapping a claims-processing workflow, which TWO steps should stay FULLY human?

    • A. Approving or denying a claim payout.
    • B. Summarising the supporting documents.
    • C. Sending the final decision letter to the customer.
    • D. Reformatting the claim notes into a table.
    • E. Drafting an internal summary of the claim.
    Show answer

    Answer: A and C.

    A payout decision and sending an external, irreversible decision letter are consequential human steps. Summarising (B), reformatting (D) and drafting internal notes (E) are reversible, language-heavy steps Claude assists with under review.

  33. Q33D4 · Workflow Integration and Solution DesignSelect one

    A stakeholder asks the Associate to promise the new workflow will 'basically run itself'. What is the BEST response?

    • A. Agree, to secure buy-in.
    • B. Explain that a human reviews and owns outputs at defined gates, describe the concrete value and limitations, and set up a feedback channel.
    • C. Say the top model makes it fully autonomous.
    • D. Refuse to communicate any value at all.
    Show answer

    Answer: B.

    Honest framing pairs concrete value with limitations, a review/ownership gate, and a feedback loop. Promising autonomy (A, C) overpromises and loses trust on the first error; refusing to communicate value (D) discards real benefit.

  34. Q34D4 · Workflow Integration and Solution DesignSelect one

    A team measures a pilot ONLY by hours saved. What is the RISK, and the fix?

    • A. No risk; hours saved is the only metric that matters.
    • B. Hours saved can mask quality loss and rework; also measure error/edit rates per task type.
    • C. They should measure sign-ups instead.
    • D. They should stop measuring once time drops.
    Show answer

    Answer: B.

    A time-only metric can hide quality regressions and rework; per-task-type quality metrics complete the picture. Time is not the only metric (A); sign-ups (C) measure adoption, not quality; stopping measurement (D) removes the evidence needed to scale.

  35. Q35D4 · Workflow Integration and Solution DesignSelect one

    A manager wants to automate sending customer refund emails end-to-end with no human, triggered by Claude's judgement. What is the BEST design?

    • A. Fully automate it; Claude is usually right.
    • B. Keep a human approval gate before any refund email is sent, since sending and refunding are irreversible and financial.
    • C. Automate but only send when Claude rates its own confidence high.
    • D. Use the most expensive model so no gate is needed.
    Show answer

    Answer: B.

    Irreversible, financial actions require a human approval gate. Full automation (A) removes it; self-rated confidence (C) is self-report reliance; model tier (D) does not remove the risk.

  36. Q36D4 · Workflow Integration and Solution DesignSelect one

    Which step is the BEST FIRST candidate for a Claude pilot?

    • A. Approving loan applications.
    • B. Drafting internal meeting summaries from notes.
    • C. Auto-sending regulatory filings.
    • D. Issuing customer payments.
    Show answer

    Answer: B.

    A reversible, internal, language-heavy, low-risk task is the ideal first pilot. The others are regulated, irreversible or financial and are poor first pilots.

  37. Q37D4 · Workflow Integration and Solution DesignSelect one

    A CRM connector is proposed so Claude can draft personalised outreach. What is the KEY governance consideration?

    • A. None; drafting is harmless.
    • B. The CRM contains customer PII, so data-handling policy and least-privilege scoping apply.
    • C. Only the model choice matters.
    • D. Outreach must always be fully automated.
    Show answer

    Answer: B.

    A CRM connector brings customer PII into scope, triggering data-handling policy and least-privilege scoping. It is not harmless (A); model choice (C) is secondary; automation (D) is not required and would raise risk.

  38. Q38D4 · Workflow Integration and Solution DesignSelect one

    An Associate is planning the rollout of a validated pilot. Which sequence reflects responsible scaling?

    • A. Scale org-wide first, then measure and add governance if problems appear.
    • B. Measure against a baseline, standardise into Projects/Skills, add data rules/gates/ownership, then phase the rollout with continued measurement.
    • C. Standardise on the most expensive model and remove review to move fast.
    • D. Keep it a permanent single-team pilot with no documentation.
    Show answer

    Answer: B.

    Responsible scaling measures, standardises, governs, then phases out with monitoring. Scaling before governance (A) is unsafe; the priciest model plus no review (C) over-buys and removes safeguards; never scaling or documenting (D) forgoes value and repeatability.

  39. Q39D7 · Troubleshooting and OptimizationSelect one

    A prompt returns a generic, off-target answer. Before anything else, what should the associate do FIRST?

    • A. Switch to the most expensive model.
    • B. Identify the cause — likely ambiguity or missing context — and fix that specific issue in the prompt.
    • C. Regenerate five times and keep the best.
    • D. Turn on extended thinking.
    Show answer

    Answer: B.

    Diagnosis precedes action: match the symptom (generic/off-target) to its cause (ambiguity/missing context) and fix it. A bigger model (A) does not add the missing context (constraint-blind); regenerating (C) varies wording, not relevance (blind regeneration); thinking (D) does not fix a vague ask.

  40. Q40D7 · Troubleshooting and OptimizationSelect one

    A prompt contains two contradictory instructions ('include every detail' and 'answer in one sentence') and the output is confused. What is the cause and fix?

    • A. Missing context; attach a document.
    • B. Conflicting instructions; remove the contradiction and state a single clear priority.
    • C. Truncation; ask Claude to continue.
    • D. Wrong model; upgrade it.
    Show answer

    Answer: B.

    Contradictory instructions confuse the output; resolve the conflict and set one priority. It is not missing context (A), truncation (C) or a model issue (D) — those misdiagnose the symptom.

  41. Q41D5 · Configuration and Knowledge ManagementSelect two

    Configuring a company-wide HR assistant, an associate has the current handbook, last year's handbook, and a spreadsheet of individual salaries. Which TWO setup choices are correct?

    • A. Include only the current handbook and remove last year's version.
    • B. Exclude the salary spreadsheet from knowledge and connectors so it is not reachable, rather than relying on a 'never reveal salaries' instruction.
    • C. Include both handbooks for completeness.
    • D. Add the salary sheet but instruct Claude never to reveal it.
    • E. Connect the entire HR Drive so nothing is missing.
    Show answer

    Answer: A and B.

    Keep one authoritative current source, and make restricted data unreachable rather than trusting an instruction. Both handbooks (C) create conflicts; instruction-only salary protection (D) is prompt-as-enforcement; whole-Drive connection (E) violates least privilege and re-exposes salary data.

  42. Q42D5 · Configuration and Knowledge ManagementSelect one

    An HR assistant Project starts citing an outdated leave policy; both the 2025 and 2026 handbooks are in the knowledge files. What is the BEST fix?

    • A. Switch to a more capable model.
    • B. Remove the superseded 2025 handbook so a single authoritative version remains.
    • C. Add a prompt saying 'use the newest policy'.
    • D. Start a new Project each month.
    Show answer

    Answer: B.

    Conflicting versions are a curation problem; keep one authoritative, current source. A bigger model (A) cannot decide authority; a prompt (C) does not reliably override a stale source; new Projects (D) do not address the duplicated files.

  43. Q43D5 · Configuration and Knowledge ManagementSelect one

    A company-wide assistant must behave identically for every employee and be auditable. At which scope should it be configured?

    • A. Whoever sets it up first, in their personal Project.
    • B. Org/Enterprise scope with central ownership, audit and controlled access.
    • C. Each team's private Project, copied around.
    • D. One executive's personal account.
    Show answer

    Answer: B.

    A must-be-identical, auditable, company-wide assistant belongs at org scope with central governance. Personal or copied-around setups (A, C, D) cannot guarantee consistency or apply org-level controls.

  44. Q44D5 · Configuration and Knowledge ManagementSelect one

    A support-triage Project should answer FAQs but hand off billing disputes over $500 to a human. Which instruction element handles this?

    • A. A longer role description.
    • B. An explicit escalation rule naming the condition and the hand-off (e.g., 'for disputes over $500, direct the user to a human agent').
    • C. A higher length limit.
    • D. A warmer tone.
    Show answer

    Answer: B.

    Escalation rules define when to hand off rather than answer, the correct mechanism for out-of-scope cases. Role length (A), length limits (C) and tone (D) do not define hand-off behaviour.

  45. Q45D5 · Configuration and Knowledge ManagementSelect two

    An assistant's answers have quietly drifted from current policy over six months, with no record of what changed. Which TWO maintenance practices would have prevented this?

    • A. A named owner responsible for the Project and its knowledge.
    • B. A regular review cadence with a change log.
    • C. Never editing the configuration once it is live.
    • D. Deleting the Project after each use.
    • E. Removing all oversight to save time.
    Show answer

    Answer: A and B.

    Ownership and a review cadence with a change log keep knowledge fresh and traceable. Never editing (C) lets it rot; deleting after use (D) and removing oversight (E) are not maintenance practices.

  46. Q46D5 · Configuration and Knowledge ManagementSelect one

    Before connecting a Slack workspace to a team Project, what should the associate check FIRST?

    • A. Which model is cheapest.
    • B. What channels and data the connected account can access and whether that scope is appropriate under policy, scoping to only what is needed.
    • C. The output font.
    • D. The maximum chat length.
    Show answer

    Answer: B.

    Connectors inherit the account's access, so the first check is the data scope and its appropriateness, scoped to least privilege. Model cost (A), formatting (C) and chat length (D) are not the governing concern.

  47. Q47D5 · Configuration and Knowledge ManagementSelect one

    An associate proposes adding every departmental document to a Project 'so it knows everything'. What is the MOST likely outcome and the better approach?

    • A. Sharper answers; add even more documents.
    • B. Bloated, conflicting knowledge dilutes relevance; include only the authoritative documents each task needs and curate for conflicts.
    • C. A faster model; no change needed.
    • D. Connectors become unnecessary.
    Show answer

    Answer: B.

    Over-stuffing knowledge lowers answer quality and raises conflict risk; curate for relevance and resolve conflicts. It does not sharpen answers (A), affect speed (C) or replace connectors (D).

  48. Q48D6 · Governance, Risk, and Responsible UseSelect two

    Under deadline, a recruiter plans to paste 200 CVs (with names and dates of birth) into a personal free AI account because the work account is slow, then delete the chat afterwards. Which TWO statements are CORRECT?

    • A. This is shadow AI: confidential PII leaves the organisation's controls, retention and audit.
    • B. Deleting the chat afterwards does not undo the exposure or the policy breach.
    • C. It is fine because the output will be deleted.
    • D. It is fine because CVs are not sensitive data.
    • E. A faster personal account would make it compliant.
    Show answer

    Answer: A and B.

    Using an unapproved personal account for confidential PII is shadow AI, and deletion cannot reverse an exposure that has already occurred. The output being deleted (C) does not cure the breach, CVs contain personal data (D), and account speed (E) is irrelevant to compliance.

  49. Q49D6 · Governance, Risk, and Responsible UseSelect one

    A recruiter wants Claude to rank and automatically reject the bottom half of candidates 'to remove human bias'. What is the BEST response?

    • A. Proceed; automation removes bias.
    • B. Keep humans accountable for shortlisting decisions and review outputs for biased or discriminatory language; do not auto-reject.
    • C. Trust the model because it is neutral.
    • D. Use the cheapest model to cut cost.
    Show answer

    Answer: B.

    Hiring is fairness-sensitive and regulated; auto-rejection launders bias and removes accountability, so humans must decide with bias review. Automation does not remove bias (A); models are not inherently neutral (C); cost (D) is beside the point.

  50. Q50D6 · Governance, Risk, and Responsible UseSelect one

    An associate is unsure whether a dataset is 'restricted' and whether policy permits it in Claude. What is the BEST FIRST action?

    • A. Assume it is fine and proceed.
    • B. Treat it as more sensitive and escalate to privacy/security/compliance before use.
    • C. Ask Claude whether the data is allowed and follow its answer.
    • D. Use it once and delete the chat.
    Show answer

    Answer: B.

    Uncertainty about restricted data is resolved by defaulting to caution and escalating to the right owner. Proceeding on a guess (A, D) risks a violation; Claude is not an authoritative policy source (C).

  51. Q51D6 · Governance, Risk, and Responsible UseSelect one

    Company policy requires disclosing AI assistance in client deliverables, but an associate omits it 'to look more rigorous'. Which statement is CORRECT?

    • A. Disclosure is always optional.
    • B. Omitting a policy-required disclosure is a transparency and policy violation, regardless of intent.
    • C. Disclosure only matters for copyright.
    • D. It only matters in regulated industries.
    Show answer

    Answer: B.

    Where policy requires disclosure, omitting it violates policy and the transparency expectation. It is not optional (A), not limited to copyright (C), and not limited to regulated industries (D).

  52. Q52D6 · Governance, Risk, and Responsible UseSelect one

    A finance analyst wants to paste customer payment card numbers into a chat to reformat them. What is the BEST response?

    • A. Proceed; reformatting is harmless.
    • B. Do not paste PCI data into an unapproved tool; PCI DSS and policy govern card data — minimise/redact and escalate to compliance.
    • C. Use a bigger model to make it safe.
    • D. Reformat, then delete the chat.
    Show answer

    Answer: B.

    Payment card data falls under PCI DSS and policy and must not be pasted into an unapproved tool; minimise and escalate. Reformatting is not harmless (A); model tier (C) and deletion (D) do not make handling PCI data compliant.

  53. Q53D6 · Governance, Risk, and Responsible UseSelect two

    A team is deciding between a personal free plan and a Team/Enterprise plan for handling confidential business data. Which TWO reasons favour the commercial plan?

    • A. Business data is not used to train models on commercial plans.
    • B. SSO and audit logs support access control and accountability.
    • C. The model becomes more intelligent on paid plans.
    • D. Personal plans provide stronger governance.
    • E. Commercial plans remove the need for human review.
    Show answer

    Answer: A and B.

    The no-training guarantee and admin controls (SSO, audit logs) are why confidential data belongs on commercial tooling. Paid plans do not change intelligence (C), personal plans lack org governance (D), and no plan removes human review (E).

  54. Q54D6 · Governance, Risk, and Responsible UseSelect one

    A manager says the flawed customer email 'was the AI's fault, not ours'. Which principle applies and what should follow?

    • A. The model is accountable; no human action needed.
    • B. The human who used and sent the output is accountable; add a verification/review step and correct the error.
    • C. Accountability depends on the model tier used.
    • D. No one is accountable for AI output.
    Show answer

    Answer: B.

    Human accountability is the anchor principle; the sender owns the output and should fix the gap that let the error through. 'The AI decided' is never a defence (A, D); tier is irrelevant (C).

  55. Q55D6 · Governance, Risk, and Responsible UseSelect one

    A clinic wants Claude to help with patient records containing PHI. Which consideration is MOST important FIRST?

    • A. Which model is cheapest.
    • B. Whether HIPAA safeguards and organisational policy permit PHI on the specific tool, and whether the data can be minimised — escalate to compliance if unclear.
    • C. The output format.
    • D. Whether the chat is fast.
    Show answer

    Answer: B.

    PHI triggers HIPAA and policy questions that must be settled before use; minimise and escalate if unclear. Model cost (A), format (C) and speed (D) are irrelevant to the governing legal/policy question.

  56. Q56D6 · Governance, Risk, and Responsible UseSelect one

    Which scenario BEST illustrates an INAPPROPRIATE use case?

    • A. Summarising an internal meeting on an approved tool.
    • B. Generating deceptive content that impersonates a real person to mislead customers.
    • C. Brainstorming campaign ideas.
    • D. Drafting a policy FAQ from an approved handbook.
    Show answer

    Answer: B.

    Deceptive impersonation to mislead is prohibited and clearly inappropriate, regardless of tool. The others are ordinary, appropriate business tasks on approved material.

  57. Q57D7 · Troubleshooting and OptimizationSelect one

    A weekly report chat, excellent two hours ago, now contradicts an earlier decision, changed the table format, invented a figure, and dropped a section. What is the BEST FIRST action?

    • A. Switch to the most expensive model and paste the whole transcript into it.
    • B. Ask Claude to summarise the confirmed decisions and state, then start a fresh chat with that summary as a clean brief.
    • C. Keep correcting each issue in the same chat.
    • D. Turn on extended thinking.
    Show answer

    Answer: B.

    These are context-bloat/drift symptoms; the cure is to capture state and reset to a clean context. A bigger model with the pasted transcript (A) carries the bloat forward; in-place correction (C) fights the bloat; thinking (D) does not clear it.

  58. Q58D7 · Troubleshooting and OptimizationSelect one

    To cut cost, a colleague proposes moving legal-contract review to Haiku AND dropping the lawyer sign-off, while leaving high-volume ticket-tagging on Opus. What is the BEST correction?

    • A. Approve it; cheaper is always better.
    • B. Keep the top model and lawyer sign-off for contracts, and instead move the routine ticket-tagging to Haiku, verifying quality holds.
    • C. Move both to Haiku.
    • D. Move both to Opus.
    Show answer

    Answer: B.

    Protect quality and the review gate on high-stakes contract work; redirect optimisation to the routine, high-volume tagging. Cheaper-is-always-better (A, C) under-buys and removes a gate on high-stakes work; Opus for tagging (D) over-buys.

  59. Q59D7 · Troubleshooting and OptimizationSelect two

    Claude's long report is repeatedly truncated before it finishes. Which TWO responses are BEST?

    • A. Ask Claude to continue from where it stopped.
    • B. Request the report in labelled sections and assemble them.
    • C. Retry the same oversized single request.
    • D. Switch to a cheaper model.
    • E. Turn off extended thinking.
    Show answer

    Answer: A and B.

    Truncation is handled by continuing or by sectioning the deliverable. Retrying the same oversized request (C) repeats the limit; model tier (D) and thinking (E) do not address output length.

  60. Q60D7 · Troubleshooting and OptimizationSelect one

    An associate applied a prompt fix; one run looked great, so they declared success. What is the FLAW and the better practice?

    • A. No flaw; one good run is enough.
    • B. A single run is not proof; compare before/after on the same task and confirm the fix holds across several runs and task types.
    • C. They should track only the overall satisfaction score.
    • D. They should switch models to be sure.
    Show answer

    Answer: B.

    Measuring improvement means before/after comparison, consistency across runs, and per-task-type checks — not a single lucky run. One run (A) can mislead; aggregate-only satisfaction (C) hides segment failures; switching models (D) does not validate the fix.

Last updated Sep 18, 2026