Representative applications
Three ways to examine Generative AI and Large Language Models
The examples consider searching internal knowledge with source citations, drafting documents from approved facts and templates, and assisting code or document review with human verification; none is presented as client evidence.
01Searching internal knowledge with source citations
Evaluation for searching internal knowledge with source citations would examine grounded task correctness while applying this control: Require grounding, citations, or refusal for factual knowledge tasks
02Drafting documents from approved facts and templates
Evaluation for drafting documents from approved facts and templates would examine citation validity and evidence coverage while applying this control: Constrain tools, schemas, permissions, and consequential actions
03Assisting code or document review with human verification
Evaluation for assisting code or document review with human verification would examine safety, refusal, and policy compliance while applying this control: Test prompt injection, data leakage, misuse, and unsafe outputs