Summer Certification Special Limited Time 70% Discount Offer - Ends in 0d 00h 00m 00s - Coupon code: force70

Anthropic Claude Certified Architect - Professional CCAR-P Question # 24 Topic 3 Discussion

Anthropic Claude Certified Architect - Professional CCAR-P Question # 24 Topic 3 Discussion

CCAR-P Exam Topic 3 Question 24 Discussion:
Question #: 24
Topic #: 3

A revenue projection assistant has missed its monthly cost target by 38 percent. Profiling shows three contributors: a 6,000-token policy preamble repeated on every call (45 percent of cost), retrieval of historical sales chunks averaging 3,000 tokens per call (30 percent), and inference on a flagship-tier model (25 percent). Stakeholders require that projection accuracy remain unchanged.

Which two optimizations should you sequence first to reduce cost without affecting accuracy? (Select two.)

Each correct answer presents part of the solution.


A.

Reduce the number of historical sales chunks retrieved across each query run.


B.

Truncate the policy preamble to remove non-essential clauses from the prompt.


C.

Enable prompt caching on the static policy preamble across the recurring calls.


D.

Switch the workload to a smaller, faster Claude model tier across all queries.


E.

Cache common retrieved sales chunks accessed across many of the daily queries.


Get Premium CCAR-P Questions

Contribute your Thoughts:


Chosen Answer:
This is a voting comment (?). It is better to Upvote an existing comment if you don't have anything to add.