TargetPrincipal-linked responses14 of 20
Illustrative held-out count
Compares a target with a matched auxiliary model on identical prompts.
Illustrative held-out count
Same prompts and settings
Runs a versioned protocol with strict output parsing, matched controls, coverage gates, and uncertainty. Results are observations, not model-safety verdicts.
Uses Hugging Face Inference Providers under your account. Inference is billed to you.
Evidence boundary. Validity depends on the comparison model. Capability, lineage, or serving differences can masquerade as target-specific behavior.