End-to-end Claude solution design: architecture patterns, RAG and integration, evaluation, governance and stakeholder delivery, for senior architects who own production systems.
Free sample questions
No account needed. Every question explains why every answer is right or wrong, just like the full bank.
lock_openFree sampleIntegrationmedium
A hospital network is piloting a discharge-summary agent that reads a patient's notes, medications and results and drafts a summary for the attending clinician to edit. The agent was built on the hospital's shared clinical tool gateway and inherited every tool on it, including place_medication_order and cancel_appointment, which the summary role never calls. Clinicians have said they will abandon the pilot if it adds interruptions to their workflow, and go-live is in three weeks. The security lead asks for a recommendation on the two write tools. What should the architect recommend?
- ARemove both write tools from the agent's configuration so its tool set covers only the read operations the summary role usescheck_circle Correct
- BKeep both tools but require the clinician to approve each order or cancellation in a confirmation dialog before it runs
- CKeep both tools and log every invocation to the security monitoring platform, with an alert on any call from the agent
- DAdd a system-prompt instruction forbidding medication orders and cancellations, and verify it with a red-team test suite
When an agent holds a capability its role never uses, removing that capability beats confirming, logging or instructing against its use. Least privilege is a structural control: a tool that is not in the agent's configuration cannot be invoked by any prompt, error or injected instruction. Confirmations, logs and prompt rules all leave the capability reachable and either add friction or act after the harm. Because the summary role is read-only, removing the write tools costs nothing in function while removing the risk entirely.
Why A is correct: Correct. The summary role never places orders or cancels appointments, so removing those tools eliminates the risk outright rather than detecting or gating it. It adds no clinician interruptions, shrinks the attack surface a manipulated prompt could reach, and is a configuration change that fits the three-week timeline.
Why B is wrong: A confirmation step feels safe because a human sees every write before it lands. It is wrong because it is a compensating control on a capability the role does not need, it adds exactly the workflow interruptions clinicians said would end the pilot, and approval fatigue tends to turn confirmations into reflexive clicks.
Why C is wrong: Logging with alerting is tempting because it adds no clinician friction and gives the security team visibility. It is wrong because it is a detective control: an erroneous medication order or cancelled appointment has already happened by the time the alert fires, and the capability itself remains.
Why D is wrong: An instruction plus red-team testing looks rigorous and costs nothing at runtime. It is wrong because a prompt instruction is a soft control that injected content or an unusual input can override, and passing a test suite does not prove the model will refuse every case when the tool is still available to call.
lock_openFree sampleSolution Design & Architecturemedium
A retail bank's head of complaints asks for a customer-facing chatbot. Discovery shows that 18 per cent of complaints breach the regulator's five-business-day acknowledgement deadline, and case timing shows most of the delay sits in staff manually sorting each complaint into one of 12 regulatory categories before it can be routed. The team has eight weeks and must show the sponsor a measurable improvement in that breach rate. What should the architect recommend?
- ABuild the customer-facing chatbot as requested, measured by how many complaints customers submit through it rather than by email each week
- BFine-tune a model on five years of past complaints so it learns the 12 categories, and begin routing only once that training has finished
- CDeploy an agent that investigates each complaint, decides the outcome and sends the final response letter to the customer without staff review
- DUse Claude to classify each new complaint into the 12 categories and route it, measured by breach rate and by accuracy on a labelled samplecheck_circle Correct
Separate the outcome a sponsor needs from the mechanism they proposed, and aim the solution at the measured bottleneck with a matching success metric. The business problem is the acknowledgement breach rate, and the evidence places the delay in manual categorisation. Framing the solution as classification plus routing attacks that cause directly, fits an LLM's strength with unstructured text, and lets the team prove success with the breach rate and a labelled accuracy sample within the deadline. A chatbot answers the proposed mechanism rather than the problem.
Why A is wrong: It is tempting because it delivers exactly what the sponsor asked for. It is wrong because the measured delay sits in internal categorisation, not in how complaints arrive, so a new intake channel leaves the breach rate untouched and its success metric says nothing about the original problem.
Why B is wrong: It is tempting because historical labelled complaints look like ideal training data. It is wrong because fine-tuning comes before any prompt-based classification has been tried, adds data preparation and evaluation work that threatens the eight-week deadline, and may not beat a well-prompted model on a 12-category task.
Why C is wrong: It is tempting because it promises to clear the whole backlog at once. It is wrong because it is far larger than the stated problem, removes human judgement from a high-impact regulated decision, and the acknowledgement deadline only requires faster categorisation and routing.
Why D is correct: This targets the step the timing data identified as the bottleneck, uses a language model for a text classification task it suits, and ties success to the breach rate the sponsor cares about plus an accuracy check that catches misrouting.
lock_openFree sampleClaude Models, Prompting & Context Engineeringmedium
A software company runs two Claude-based review paths. An editor plug-in suggests fixes as developers type and must return its first token within 1.5 seconds. A nightly job reviews every merged pull request that touches concurrency code, and engineers read its findings the next morning. On 200 held-out concurrency pull requests, enabling extended thinking on the nightly reviewer raised the share of seeded defects found from 58 to 81 per cent and added about 40 seconds per review, and the engineering director has approved the higher cost for that job. Where should extended thinking be enabled?
- AOn both paths, so that developers and the nightly job receive reasoning of the same depth.
- BOn the editor plug-in only, because interactive suggestions reach developers the soonest.
- COn the nightly concurrency review only, leaving the editor plug-in to answer without it.check_circle Correct
- DOn neither path, moving the nightly job to the fastest tier to offset the cost of the reviews.
Enable extended thinking where measured reasoning gains matter and latency is slack, and keep it off paths bound by a tight interactive latency budget. Extended thinking lets the model spend additional tokens reasoning before it produces the final answer, which improves results on multi-step problems such as concurrency defects but lengthens time to the answer and raises token spend. The decision is therefore made per path: an offline job with an approved budget and a measured quality gain is where the trade pays, while an interactive path with a 1.5 second first-token budget is where it does not.
Why A is wrong: Consistency across tools is an appealing principle, and the nightly gain suggests the plug-in might improve too. It is wrong because extended thinking adds generation time before the answer, which the 1.5 second first-token budget on the editor path cannot absorb, and the plug-in's quick fix suggestions were never shown to need deeper reasoning.
Why B is wrong: It is tempting to put the best reasoning where users see it first. It is wrong on both counts: the plug-in has the tight latency budget that extended thinking would break, and the measured 23-point gain in defects found belongs to the nightly concurrency review, which this option leaves without it.
Why C is correct: This is correct because extended thinking earns its latency where a task needs multi-step reasoning and the consumer is not waiting: the nightly review has no one reading it until morning, a measured gain from 58 to 81 per cent, and approved cost. The editor plug-in keeps its 1.5 second budget by answering directly.
Why D is wrong: Cost reduction is a reasonable instinct for a job that runs on every merged pull request. It is wrong because the director has already approved the higher cost, and dropping both extended thinking and model capability on hard concurrency reasoning gives up the measured improvement in defect detection that the job exists to deliver.
More free CCAR-P practice questions, every answer explainedFrequently asked questions
- How many questions are on the Claude Certified Architect Professional exam?
- The Claude Certified Architect - Professional (CCAR-P) exam has 63 questions and runs for 120 minutes. The format is proctored multiple-choice and multiple-response; each item states how many responses to select.
- What score do I need to pass Claude Certified Architect Professional?
- The pass mark is 720 / 1000. Examworthy gives you a per-domain readiness score so you can see which domains are holding you back before you book.
- How much does the Claude Certified Architect Professional exam cost?
- The exam costs 175 USD to sit. Practising on Examworthy is free to start, and every answer is explained, right and wrong.
- Is there a Claude Certified Architect Professional practice exam?
- Yes. Examworthy's exam mode runs a timed Claude Certified Architect Professional practice exam (mock) paced to match the real exam, scored per domain so you can see exactly where you stand. Timed mocks are free with an account.
- How does Examworthy help me prepare for Claude Certified Architect Professional?
- Every practice question explains why the right answer is right and why each wrong one is wrong, mapped to the official blueprint domains. You learn the reasoning, not just the letter.
- Is Examworthy affiliated with Anthropic?
- No. Examworthy is not affiliated with or endorsed by Anthropic. Our questions are original, blueprint-aligned practice material; we never reproduce live exam items.
Examworthy is not affiliated with or endorsed by Anthropic. All questions are original, blueprint-aligned practice material. We never reproduce live exam items. CCAR-P and related marks belong to their respective owners.