Fact-Check a Claude Research Brief and Ship a Claim Register
Ask Claude for a research brief, then audit it claim by claim: extract every checkable factual assertion into a register, verify each one against a primary source, mark it verified, unsupported, or contradicted, and set the threshold at which a human has to review before the work leaves the building. Includes adapting the corrected brief for a second audience. No coding required.
2.5 hrs
Est. time
5
Outcomes
8
Rubric criteria
65%
Pass score
What you'll learn
Skills you'll have real reps in after shipping this.
See how it works
Claims traced back to sources
The audit turns flowing prose into a register where every assertion points at a source or is marked as pointing at nothing. That mapping is what makes an error rate measurable instead of a matter of impression.
The scenario
A colleague forwards a Claude-written market brief to a client without reading it closely. Two of the figures in it are wrong, one cited report does not exist under that title, and the client notices before you do. The output was fluent, well organized, and confidently worded, which is exactly why nobody stopped to check it. Fluency is not evidence, and a hallucination is a claim the model generated because it fit the pattern of the surrounding text rather than because it came from a source.
The habit that prevents this is mechanical rather than intuitive. You pull every checkable assertion out of the prose into a register, verify each one against a primary source you can link, and record the verdict next to the claim. Claims that cannot be verified do not get quietly deleted, they get marked and escalated. This task makes you run that audit end to end on a brief you generate yourself, then decide in advance which categories of claim always require a human before the work ships.
Your role
You are the person who signs off before AI-assisted research reaches a client or an executive. Your deliverable is a claim register for a Claude-generated brief, with a verdict and a source for every factual assertion, plus a written human-review threshold your team could adopt.
Start the task to unlock the full brief
You'll get the step-by-step requirements, setup commands, the 8-criterion grading rubric, tips, and the ability to submit your solution for instant AI grading.
Free to start · submit when you're ready
Learning resources
What you'll build in this output validation task
This is a build-and-submit task rather than a guided lab, and it requires no coding. You generate a research brief with Claude, then audit it the way a careful analyst audits any draft that is about to reach a client: every checkable assertion goes into a register, gets a verdict of verified, unsupported, or contradicted, and carries a primary source link or a note explaining why none exists.
The point is to build a habit that does not depend on being alert. You tally the verdicts into an error rate, identify what the failing claims have in common, and then write a human-review threshold naming the claim categories that always require sign-off before work leaves your team. The final step adapts the corrected brief for a second audience, which tends to expose claims that only read as solid because the original wording was vague.
Grading is rubric-based and explainable. Your document is scored against weighted criteria covering the register, the sourcing, the error-rate tally, the pattern analysis, and the review threshold, with per-criterion feedback quoted from your submission. The pass threshold is 65 percent and you can resubmit. Output evaluation and validation is the heaviest scored domain on the Claude Certified Associate Foundations exam.