COGNIT-Guard: Calibrated Standalone Direct-Decision Guardrails with Heterogeneous CPU-NPU Confidence Cascading under Explicit Latency and False-Positive Constraints
COGNIT-Guard cascades a CPU gatekeeper to a 322M model for low-latency prompt safety screening.
COGNIT-Guard couples a validation-calibrated CPU fast gatekeeper with confidence-gated escalation to Laya-322M, a 322M bidirectional direct-decision safety model, under an asymmetric false-positive penalty. On the unseen DUCS-Bench split (N=607) it reports 98.85% accuracy, 0.42% benign false-positive rate (1/238), 1.12% ECE, and a 0.0104 Brier score. On Huawei Ascend 910C NPUs, pure NPU inference averages 21.77 ms, while the deployed CPU-NPU cascade averages 41.63 ms at 99.23% accuracy and 0.00% FPR. Experience replay restores SafetyBench-ZH out-of-domain accuracy to 64.10-65.05%.
51