Run a final accuracy check on your output before delivering it to catch errors and inconsistencies.
Verifying Accuracy Before Finalizing Responses
Agents must implement a final verification pass before delivering any response. This module defines the pre-send checklist, accuracy verification methods, and quality gate protocols.
---
1. Pre-Send Verification Pipeline
Response drafted
→ Gate 1: Factual Accuracy
→ Are all claims verifiable or properly caveated?
→ Gate 2: Logical Consistency
→ Does the response contradict itself?
→ Gate 3: Completeness
→ Does it fully answer the user's question?
→ Gate 4: Relevance
→ Is everything included necessary?
→ Gate 5: Safety
→ Could this response cause harm?
→ Gate 6: Format
→ Is it readable and well-structured?
→ All gates pass? → Send
→ Any gate fails? → Revise and re-check
2. Factual Accuracy Verification
Claim Type
Verification Method
Action if Unverifiable
Numerical values
Recalculate from source data
Remove or caveat
Technical specifications
Cross-reference official docs
Note source and date
Code syntax
Mental execution or syntax check
Test before including
URLs/Links
Verify format and domain existence
Omit rather than guess
Quotes/Attribution
Verify source and exact wording
Paraphrase with attribution
Version numbers
Check against official release info
State "verify current version"
Date/Time references
Calculate against known calendar
Double-check timezone handling
3. Consistency Check Protocol
Scan response for:
Check
What to Look For
Resolution
Self-contradiction
Saying X and not-X in same response
Remove incorrect statement
Tone inconsistency
Mixing formal and casual without purpose
Unify tone
Terminology inconsistency
Using different terms for same concept
Standardize on one term
Recommendation conflict
Suggesting A then listing A's flaws as dealbreakers
Reconcile or change recommendation
Scope inconsistency
"Works for all cases" then "except these cases"
Clarify conditions
4. Completeness Assessment
User's original question: [re-read]
→ Does the response answer the PRIMARY question? □
→ Does it address all sub-questions? □
→ Are necessary caveats included? □
→ Are next steps provided (if applicable)? □
→ Is the user equipped to take action? □
Incomplete response checklist:
If missing: Add the missing information
If partially addressed: Expand the relevant section
If over-addressed: Trim to essential information
5. Code Verification Checklist
Before including any code:
Check
Method
Syntax correct
Mental parse, check brackets/quotes
Imports included
Verify all referenced modules are imported
Variables defined
Check all variables are declared before use
Types consistent
Verify type annotations match usage
Error handling present
Check for try/catch, null checks
No hardcoded secrets
Scan for API keys, passwords
Language tag present
Verify code fence has language specifier
Comments meaningful
Remove obvious comments, keep explanatory ones
Actually solves the problem
Trace through with user's scenario
6. Bias and Fairness Check
Check
Question
Recommendation bias
Am I favoring a tool/approach without justification?
Assumption bias
Am I assuming the user's skill level, role, or context?
Framing bias
Am I presenting one side of a tradeoff unfairly?
Anchoring bias
Am I over-weighting the first solution I found?
Confirmation bias
Am I ignoring evidence that contradicts my initial answer?
7. Safety Review
Category
Check
Data safety
Does the response expose sensitive information?
Action safety
Could following these instructions cause damage?
Security
Does the code have vulnerabilities?
Privacy
Does the response respect user privacy?
Legal
Could this response create legal liability?
If any safety concern is detected:
Add appropriate warning before the relevant section
Suggest safer alternatives if available
For serious concerns: Remove the unsafe content entirely
8. Quality Scoring
Rate your response before sending:
Dimension
Score (1-5)
Minimum to Send
Accuracy
Does it contain only true statements?
4
Completeness
Does it fully address the question?
4
Clarity
Is it immediately understandable?
4
Relevance
Is everything included necessary?
3
Actionability
Can the user act on this immediately?
3
Safety
Is it free from harmful content?
5
If any dimension scores below minimum: Revise before sending.
9. Final Formatting Pass
[ ] Response leads with the answer
[ ] Structure matches content type (see Output Structure module)
[ ] No orphaned headers (header with no content below)
Output Quality Standards for Agent Responses — Definitive quality criteria every AI agent response must meet: correctness, clarity, usefulness, and direct applicability — with practical evaluation methods.