Capture the exact claim
Write down the precise capability, workflow or performance promise before judging it.
Methodology
VetAI Trust evaluates claims by separating what is asserted, what is supported by vendor evidence, what has independent support and what has been independently reproduced.
The evidence ladder
A lack of independent verification is an evidence status, not an accusation. It means readers should know what is known and what still needs testing.
Core flow
VetAI Trust tries to keep the evaluative chain legible. Different policy pages explain the rules around each step; the methodology page explains the assessment logic itself.
Write down the precise capability, workflow or performance promise before judging it.
Review papers, validation reports, primary sources and disclosed limitations.
Benchmark only when an independent test answers a real public-interest question.
Readers should see what is supported, what is unresolved and what changed over time.
Evaluation workflow
The workflow is intentionally inspectable. Readers should be able to see how a conclusion was reached and where uncertainty remains.
Identify the exact wording, context, product scope and practical implication.
Collect public materials, papers, validation studies and any vendor-disclosed evidence.
Ask whether the evidence actually supports the claim, not merely a neighboring or narrower claim.
Use relevant workflow, expert, specialized veterinary AI and frontier-model baselines where appropriate.
Design independent testing only when review alone cannot answer the public-interest question.
Publish findings with methodology, caveats, conflicts and unanswered questions.
Claims must be specific enough to evaluate. Broad phrases are translated into testable questions before review.
Press language, abstracts and marketing summaries are not substitutes for study design and underlying methods.
A test result is only meaningful when readers know what it was compared against and why that baseline was chosen.
Policy links
This page explains how evidence is assessed. It deliberately avoids duplicating the policy pages that govern independence, benchmark access, privacy and corrections.