Runs the chained identity + campaign clustering pipeline against all seven fixtures via from_synthetic / from_synthetic_identity adapters and ratchets every YAML floor to 1.0 — the production clusterer (and the reference clusterers used in the per-fixture tests) all score perfectly across ARI / homogeneity / completeness / singleton_recall on each fixture. Three substrate fixes surfaced by the ratchet: - Tuning: shared_infra now Jaccards payload+C2 only; decky_set moved into cohort_weight to prevent fleet-scarcity false-merges (F1's shared_wordlist failure mode). Tier weight raised to 1.0 so shared payload+C2 alone crosses threshold (F5's intended pass). - Adapter: from_synthetic_identity now reads SyntheticSession started_at + duration_s for session_windows and per-decky timestamps (the production-row adapter still uses start_ts/end_ts when available). - Fixture data: paused_campaign.yaml's JA3 collided exactly with vpn_hopping.yaml's (same TLS extension list). The collision fused two unrelated campaigns under the chained identity layer in the noise_floor composite. Made paused's JA3 distinct. Also wires Campaign / CampaignsResponse into models/__init__.py's __all__ that was missed in the schema commit.
25 lines
810 B
YAML
25 lines
810 B
YAML
# Bounds for fixture 7 (slow_burn).
|
|
#
|
|
# Ground truth at campaign-level: 1 campaign of 3 observation rows
|
|
# (one per operational window — recon, exploit, action). A correct
|
|
# algorithm scores 1.0 across every metric on this fixture.
|
|
#
|
|
# Completeness is the load-bearing metric: a clusterer that lets
|
|
# multi-week silence fragment the campaign tanks completeness (the
|
|
# one true class is split across the operational windows). The
|
|
# adversarial recency_decay_clusterer demonstrates this and the
|
|
# bound below rejects it.
|
|
#
|
|
# Campaign-level fixture only — the three DSL actors model the
|
|
# operator's three operational windows by design.
|
|
#
|
|
# Bounds are loose at v1; tighten as the algorithm matures.
|
|
adjusted_rand_index:
|
|
min: 1.0
|
|
homogeneity:
|
|
min: 1.0
|
|
completeness:
|
|
min: 1.0
|
|
singleton_recall:
|
|
min: 1.0
|