Reliability and validity are different questions. Getting the same score twice shows the instrument is steady. It does not prove the score matches what a trained human reader would conclude.
Truverai has not run a blind expert agreement study, so agreement with expert judgment is not claimed anywhere on the site.
Truverai scores are reproducible within a published lock: engine v4.7, prompt hash 1a6c25c7e25ddbf8, model google/gemini-3-flash-preview at temperature 0, and a median of three runs. That is reproducibility within a fixed configuration, not a claim that a different language model would return the same number.
Model independence is not claimed either. Different language models read the same passage differently, which is why the configuration is fixed and published rather than swapped for whichever model is newest.
This page needs JavaScript for the full interactive view. The summary above is the same information in plain HTML.