Section 186 · Appendices, Tools, Templates, and Reference
Appendix: Glossary of AI Testing Terms
A shared vocabulary makes AI quality work easier to teach, debate, and improve.
glossary terms
What to do
- Define runnable checks that exercise glossary terms.
- Set acceptable outcomes and blocker failures for glossary terms before running the evaluation.
- Run representative cases for glossary terms and preserve the failures that would change the decision.
Evidence to preserve
- Preserve the inputs, versions, configurations, raw outcomes, and results for glossary terms needed to reproduce work on Appendix: Glossary of AI Testing Terms.
- Report results for glossary terms by relevant slice, separate blocker failures from averages, state uncertainty and blind spots, and connect the result to a release decision.
Expert note
In a real release review, treat the glossary as a living artifact. Update it when the organization invents new failure categories, metrics, release gates, or governance concepts.
Continue the conversation
Apply this to your context.
Save your product context once, then open a focused conversation that combines it with this concept.
Cite this page
Jason Arbon. "Appendix: Glossary of AI Testing Terms." Testing AI Knowledge Edition, section 186.
https://jarbon.ai/testing-ai/knowledge/ch186-glossary-terms.html