Exam guide·8 min read·3 October 2026

Claude Exam Task Statements: CCAR-F Domain Guide

Master all 30 claude exam task statements across the five CCAR-F domains. Learn domain weights, key anti-patterns, and scenario-study strategy for exam day.

By Solomon Udoh · AI Architect & Certification Lead

Claude Exam Task Statements: CCAR-F Domain Guide

The 30 claude exam task statements map every CCAR-F question to a specific, testable skill. Learning to read them is the fastest route to focused preparation and the surest way to avoid wasted study time.

The Claude Certified Architect, Foundations exam (CCAR-F) organises its 60 scenario-based items around five domains and 30 task statements. Each question anchors to at least one task statement; high-weight domains produce proportionally more questions per sitting. Before diving into practice exams, download the official task statement list from Anthropic's certification portal to confirm the version you are studying.

What are the CCAR-F task statements?

Task statements are the atomic skill descriptors that make up each exam domain. They describe what a Claude Partner Architect must be able to do, not merely recall. A task statement in Domain 1 might require you to identify the correct fix when an agentic loop anti-pattern causes premature termination. One in Domain 2 might ask you to diagnose a tool misrouting problem given a description that is too broad to guide tool selection.

Because every item is scenario-based, task statements read as judgment tasks: "evaluate", "diagnose", "select", "design". Memorising definitions will not pass the exam. You need to apply each task statement to a realistic situation under time pressure, with plausible wrong answers that reflect common architectural mistakes.

How are the five domains and their task statements weighted?

The CCAR-F has five domains, each containing a cluster of task statements. Domain weights determine roughly how many of the 60 questions will test skills from each cluster.

DomainTitleWeightApprox. questions
1Agentic Architecture & Orchestration27%~16
2Tool Design & MCP Integration18%~11
3Claude Code Configuration & Workflows20%~12
4Prompt Engineering & Structured Output20%~12
5Context Management & Reliability15%~9

Domain 1 at 27% is the single largest domain, making its task statements the highest-return study target. Domain 5 at 15% is the smallest, but its task statements on context reliability appear inside multi-domain scenarios, so gaps there compound. Anthropic does not publish the raw-to-scaled score conversion, so the approximate question counts are illustrations based on 60 items times stated percentages; your actual sitting draws 4 scenarios at random from a bank of 6.

Which task statement areas carry the most weight in Domain 1?

Agentic Architecture & Orchestration at 27% is the dominant domain. Its task statements cover:

  • Designing the agentic loop around stop_reason so the model continues, terminates, or escalates correctly rather than hitting an arbitrary cap
  • Choosing between hub-and-spoke architecture and flat multi-agent topologies based on coordination complexity
  • Diagnosing narrow decomposition failure when a coordinator splits a task into pieces that cannot reassemble coherently
  • Selecting parallel subagent spawning versus sequential execution when latency and error isolation requirements differ
  • Enforcing constraints via hooks versus prompts using the hooks vs. prompts decision framework

A common exam pattern: you are given a multi-agent research pipeline where the coordinator spawns three subagents, one fails silently, and the synthesis step produces a confident but incomplete answer. The task statement asks you to identify the correct architectural fix. The answer is almost never "add a retry loop" as a first move; it is to surface structured error metadata so the failure is visible before synthesis begins.

What does the exam test in Domain 2 task statements?

Domain 2 covers Tool Design & MCP Integration at 18%. Its task statements focus on:

  • Writing tool descriptions that function as a selection mechanism, guiding Claude to the right tool without ambiguity
  • Deciding when to split a broad tool into narrower variants rather than patching the description
  • Configuring tool_choice appropriately versus fixing the underlying tool design
  • Scoping MCP servers to the right level in the MCP scoping hierarchy
  • Handling the four error categories correctly via the isError flag

The exam will not ask you to recite tool_choice parameter values from memory. It will present a scenario where the model consistently routes to the wrong tool and ask which fix to apply first. The task statement tests whether you reach for the low-effort, high-leverage solution before reaching for configuration.

Here is the pattern of a tool description the exam expects you to flag as broken:

text
Tool name: search
Description: Searches things and returns results.

And the corrected form the task statements reward:

text
Tool name: web_search
Description: Queries the public web for current information not in the model's
training data. Use when the user asks about events after the knowledge cutoff,
real-time prices, or live URLs. Do NOT use for internal document retrieval.

The difference is not stylistic. A vague description produces systematic misrouting at scale; a scoped description with explicit exclusions narrows the model's tool selection to the intended use case.

How do Domain 3 task statements differ from the others?

Claude Code Configuration & Workflows at 20% has a practical, operational flavour. Its task statements ask you to:

  • Navigate the three-level configuration hierarchy (user, project, and system) and identify which layer takes precedence when they conflict
  • Choose between CLAUDE.md, .claude/rules/ path-scoped files, skills, hooks, and permissions for a given guidance scenario
  • Design hooks for deterministic enforcement where prompt-based instructions would be insufficient or unreliable
  • Set up CI/CD pipelines using Claude Code in headless mode for automated workflows

These task statements lean toward system-level thinking. When the exam presents an organisation-wide compliance requirement, such as preventing secrets in commit messages, the correct answer is a hook, not a system prompt instruction. The task statement is testing whether you understand the reliability gap between advisory text and enforced behaviour. An instruction in a prompt can be overridden or forgotten; a hook runs deterministically on every relevant event.

What makes Domain 4 task statements particularly testable?

Domain 4 covers Prompt Engineering & Structured Output at 20%. Its task statements include:

  • Constructing few-shot examples that close the gap between intended and actual model behaviour for ambiguous edge cases
  • Designing JSON schemas to prevent fabricated fields in extraction workflows
  • Adding a validation loop before downstream consumption of structured output
  • Choosing between goal-based and step-based prompts based on task complexity and the degree of procedural constraint required

The exam consistently rewards deterministic solutions over probabilistic ones when stakes are high, proportionate fixes, and root-cause tracing.

Anthropic , CCAR-F Exam Guide

This principle runs through Domain 4 task statements directly. A structured output scenario typically involves an extraction pipeline where Claude occasionally fabricates a field the schema does not define. The task statement rewards candidates who add schema validation before the pipeline's next stage rather than prompting Claude to "be more careful", because the latter is probabilistic and the former is enforceable.

python
# Exam-preferred pattern: validate before consuming structured output
import json
import jsonschema
def extract_and_validate(raw_output: str, schema: dict) -> dict:
parsed = json.loads(raw_output)
jsonschema.validate(parsed, schema) # raises ValidationError on fabricated fields
return parsed

The exam will not ask you to write this function from memory. It will present a pipeline description and ask which architectural change most reliably prevents fabricated-field errors reaching downstream consumers.

How do Domain 5 task statements appear across multi-domain scenarios?

Context Management & Reliability at 15% is the smallest domain, but its task statements appear inside scenarios that are nominally classified under Domain 1 or 4. When a long-running agent degrades in quality after many turns, that is a context management failure even if the scenario is labelled as an agentic architecture question.

The key task statements in this domain cover:

  • Recognising context degradation in extended sessions and distinguishing it from model capability limits
  • Choosing between summary injection, fresh session start, and session forking based on current task state and the cost of lost context
  • Avoiding the stale context problem when resuming long investigations after interruption

Candidates who treat context as a first-class resource rather than an implementation detail consistently perform better on multi-domain scenarios that mix Domain 1 and Domain 5 task statements.

How should you map your study plan to task statements?

A task-statement-first study plan runs as follows:

  1. Download the official CCAR-F exam guide from the Anthropic certification portal.
  2. For each of the 30 task statements, write a one-sentence description of what a correct answer looks like.
  3. Find at least one anti-pattern for each task statement (what wrong answers look like and why they fail).
  4. Work through scenario-based practice questions and tag each question to the task statement it tests.
  5. Use your domain-level score report to identify which task statement clusters still need work.

Our concept library at /concepts covers 174 atomic concepts mapped to all five CCAR-F domains and 30 task statements. Each concept page cross-references the task statements it supports, so a low score in Domain 2 leads you directly to the specific concepts underlying those task statements rather than to a broad domain review.

The platform's adaptive engine uses Bayesian Knowledge Tracing with a 0.90 mastery threshold. It will not mark a task statement cluster as ready until your answer pattern is consistent across multiple question variants, not just lucky on one attempt.

The CCAR-F passing score is 720 on a 100 to 1000 scale. Anthropic does not publish the raw-to-scaled conversion, so focus on task statement mastery rather than trying to count minimum correct answers.

What anti-patterns appear most often across task statement questions?

The exam's most testable anti-patterns cluster around five failure modes that cut across multiple domains:

Anti-patternDomain(s)Why the exam tests it
Arbitrary iteration caps1Masks real loop termination logic; not a deterministic fix
Natural-language-only enforcement1, 3Unreliable for compliance; hooks are the correct answer
Overstuffed tool sets2Causes systematic misrouting; splitting is the fix
Schema-free extraction4Produces fabricated fields at scale in production pipelines
Stale context on session resume5Degrades reliability in long-running agentic tasks

Recognising these anti-patterns as wrong answers is a skill the task statements test directly. Scenario questions are written so that the anti-pattern option sounds reasonable; the task statement tests whether you know why it fails, not just that it does.

Frequently asked questions

What are the 30 claude exam task statements on the CCAR-F?
The 30 task statements are the atomic skill descriptors distributed across five domains in the CCAR-F exam guide. They describe what a Claude Certified Architect must be able to do, using verbs like evaluate, diagnose, and design. The complete list is in Anthropic's official CCAR-F exam guide, available from the certification portal. AI Skill Certs maps all 30 to its 174-concept library.
How many task statements are in each CCAR-F domain?
Anthropic does not publish a per-domain task statement count, only the five domain weights and the total of 30 task statements. Based on the weighting, Domain 1 (27%) and the 20% domains (3 and 4) carry the most task statements, but the exam guide is the authoritative source for the exact distribution per domain.
Do I need to memorise task statements word-for-word to pass?
No. The CCAR-F is scenario-based, not recall-based. Each item tests whether you can apply a task statement to a realistic situation. You need to understand what each task statement asks you to judge or design, not reproduce its wording. The passing score is 720 on a 100 to 1000 scale, rewarding applied judgment over memorisation.
Which domain's task statements should I study first?
Domain 1, Agentic Architecture & Orchestration, at 27% of the exam is the highest-return starting point. Its task statements on agentic loop design, subagent orchestration, and hook-versus-prompt enforcement appear in the most questions and also surface inside multi-domain scenarios tagged to other domains.
How do task statements differ from exam domains on the CCAR-F?
Domains are the five broad topic areas with published percentage weights. Task statements are the specific, verb-led skills within each domain, totalling 30 across the exam. A domain is a container; a task statement is the measurable unit the exam question actually tests. Study plans built around task statements are more precise than domain-level study plans.
Can I see a sample question mapped to a specific task statement?
Anthropic's official exam guide includes sample scenario questions mapped to domains, and the score report after your sitting shows percent-correct by domain. AI Skill Certs practice exams (60 items, scored 100 to 1000 with 720 as the pass bar) link each question to the relevant concept in the 174-item library, which is itself mapped to the 30 task statements.

People also ask

What are claude exam task statements used for?
Claude exam task statements define exactly which skills each CCAR-F question tests. Anthropic uses them to build scenario items that assess practical judgment, not recall. Candidates use them as a study scaffold: mapping practice questions to task statements reveals which skill clusters need more work before the exam.
How many domains does the Claude Certified Architect exam have?
The CCAR-F has five domains: Agentic Architecture & Orchestration (27%), Tool Design & MCP Integration (18%), Claude Code Configuration & Workflows (20%), Prompt Engineering & Structured Output (20%), and Context Management & Reliability (15%). Together they contain 30 task statements across 60 scenario-based items.
Are claude exam task statements published by Anthropic?
Yes. Anthropic publishes the full list of 30 task statements in the official CCAR-F exam guide, available from the Claude Partner Network certification portal. The guide also includes domain weights, exam format details, and sample questions. It is the authoritative source; third-party summaries should be cross-checked against it.
What is the passing score for the Claude certification exam?
The CCAR-F passing score is 720 on a scaled range of 100 to 1000. The score report shows pass or fail, the scaled score, and percent-correct by domain. Anthropic does not publish the raw-to-scaled conversion, so do not attempt to calculate a minimum question count as a pass target.
How is the CCAR-F exam structured around scenarios?
Each CCAR-F sitting draws 4 scenarios at random from a bank of 6. Each scenario generates multiple items, and every item is scenario-based, testing practical judgment against the 30 task statements. The format rewards candidates who can diagnose architectural problems and select proportionate fixes, not those who recall definitions.

About the author

Solomon Udoh

AI Architect & Certification Lead

Solomon Udoh is an AI Architect who designs and ships production agent systems on the Claude API and Claude Code. He built AI Skill Certs' adaptive engine and authored its 174-concept knowledge graph, mapping every Claude Certified Architect - Foundations objective to hands-on, exam-aligned practice.

  • Designs production multi-agent systems on the Claude API and Agent SDK
  • Author of the AI Skill Certs knowledge graph (174 mapped exam concepts)
  • Builds with MCP, Claude Code, structured outputs, and agentic loops daily
  • Reviews every concept page against the official Anthropic exam guide

You might also like

Ready to put it into practice?

Study every exam concept with an adaptive tutor.

Start studying