Advanced Evidence Profile (AEP)¶
Status: Optional · informative · offline analysis only
UHBS version: 4.5.1 · AEP schema version: 1.0.0
Evaluation scope: Laboratory / sandbox only — not real-world production testing
Lab evaluation only
UHBS and AEP are for isolated lab and sandbox evaluation of decoys. Do not point harnesses, collectors, or AEP trial collection at production systems, customer environments, or unauthorized targets. The optional UHQS > 80 “production baseline” is an internal gate after lab grading, not permission to test in the real world.
UHQS remains the normative implementation-quality and safety grade. The optional Advanced Evidence Profile adds controlled lab evidence about adversarial perception, engagement, distinguishability, and cost under declared experimental conditions. AEP does not change UHQS.
Academic credit: AEP’s design vocabulary draws on cited research (Zhu 2019; Collins et al. 2024; Ersok et al. 2022; Li et al. 2020). See Research foundations & credits — citing those works does not imply their authors endorse UHBS.
flowchart LR
coreRun["Core UHBS lab"] --> scorecard["UHQS scorecard (unchanged)"]
trialEvidence["Local controlled trial evidence"] --> offlineAep["Offline AEP analyzer"]
experimentManifest["Experiment manifest"] --> offlineAep
scorecard --> addendum["Advanced Evidence Addendum"]
offlineAep --> addendum
What AEP is¶
AEP is a lab controlled-experiment layer that supplements a valid UHBS scorecard with evidence about:
- whether a decoy changes attacker/agent behavior relative to a matched lab reference
- engagement (dwell, exchanges, capability expenditure)
- distinguishability across fingerprint layers (with TPR/FPR)
- defender utility deltas under an explicit utility model (VoD)
What AEP is not¶
- Not a replacement for Modules A–F
- Not part of UHQS, δ_C, weights, or letter grade
- Not a certification or “proven against all attackers” claim
- Not real-world / production testing — lab and sandbox only
- Not permission to test production systems, customer networks, or unauthorized targets
- Not an attack launcher —
uhbs aeponly reads local files - Attestations require
sandbox_onlyandno_production_assets
Who should use it¶
Honeypot developers, academic researchers, red teams, deception engineers, and evaluators comparing LLM / ICS / tarpit behavior in sandboxed lab studies.
Decision guide¶
| Goal | Use |
|---|---|
| Implementation quality / release gate | Ordinary UHBS scorecard (Modules A–F, UHQS) |
| Controlled decoy-versus-reference evidence | Add AEP |
| Organization or campaign maturity | CDMM / Engage — related frameworks |
Status vocabulary¶
AEP reports valid | inconclusive | control_failed — not pass/fail grades.
Optional SLM evaluator (alpha)¶
For labs that want help drafting AEP trial JSONL with a small/local model (or a deterministic mock) before offline analyze:
- SLM evaluator (alpha) — what it is for, activation checklist,
providers (
mock/recorded/ loopbackopenai_compatible) - Default: off. Install does not enable it; you must edit
aep-slm.yaml. - Does not change UHQS. Not exposed via AI-host MCP.
- Published: https://uhbs.github.io/uhbs-standard/mkdocs/advanced-evidence/slm-alpha/
Most users should start with ordinary AEP tutorials and skip SLM until they need it.
Start here¶
- Beginner tutorial —
uhbs aep example beginnerthen analyze - CLI reference —
pip install 'uhbs[aep]' - Methodology · Metrics · Runbook
- Research foundations
- Optional: SLM evaluator (alpha) — off by default; edit
aep-slm.yamlto unlock (does not change UHQS)