give segmentation QA the same threshold-independent view as detection

Detection QA reports a precision/recall curve, average precision and a
calibration sweep; segmentation QA reported a single operating point. Both rank
their outputs by confidence, so the same view applies, and the asymmetry meant
the two panels answered different questions about comparable runs — an
inconsistency introduced when detection gained the curve.

Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
This commit is contained in:
Jens
2026-08-22 20:47:30 +02:00
co-authored by Claude Opus 5
parent ff4a15aa74
commit 1a1a9af6e7
6 changed files with 37 additions and 1 deletions
+1
View File
@@ -223,5 +223,6 @@ def compare_segmentation_run_with_reference(
iou_threshold=payload.iou_threshold,
class_name=payload.class_name,
min_confidence=payload.min_confidence,
calibration_thresholds=payload.calibration_thresholds,
)
)