DocsTroubleshootingGlossary

Glossary

Definitions of key terms used throughout PeerLM.

TermMeaning
MonitorA standing production decision with sources, policies, Runs, and a living verdict
RunOne frozen sample/replay/judge/aggregate cycle; the evidence receipt
TriggerA manual, scheduled, catalog, price, drift, or deploy event that starts a Run
ControlThe incumbent evidence: captured production output in observed mode or regenerated output in replayed mode
CandidateAn alternative model compared one-to-one with the control
Judge poolThe models frozen on a Monitor from which each item's full panel is seated
Decision contractThe thresholds and critical categories frozen before a Run
Evidence strengthA 0–100 description of sample sufficiency, execution cleanliness, and determinate agreement
SwitchA candidate held quality under the contract and is worth replacing the incumbent with
RouteEvidence supports a candidate only for specific stable workload categories
HoldEvidence does not support a change because of quality, economics, or insufficient evidence
Observed controlUses the exact captured production response without calling the incumbent
Replayed controlRegenerates the incumbent under the frozen Run context
Sponsored first lookA free, directional comparison on pasted prompt/response pairs; not a standing Monitor
Run packA $99/month self-serve expansion adding one Monitor slot and 10 pooled Runs
Prompt CIA deploy Run comparing old and new system prompts on the same model
Compatibility terminology: REST and MCP may still expose Suite evaluation endpoints for existing integrations. Suites are an internal engine and are not a customer dashboard surface.