ZeroHour

Search: “adversarial attacks”

13,254 items

MarkSec: Capability-Aware Evaluation of Adversarial Attacks Against LLM Watermarks

MarkSec unifies evaluation of stealing, scrubbing, and spoofing attacks against LLM watermarks with quality-constrained success metrics under shared reporting protocols.

MarkSec is a framework unifying analysis of stealing, scrubbing, and spoofing attacks against LLM watermarks under shared detector calibration, metric definitions, and reporting protocols. It introduces a quality-constrained attack success metric that jointly assesses attack effectiveness and text quality. Experiments across representative watermark families, attacks, LLMs, and datasets show that attacks strongest by watermark removal alone can fall behind general rewriting when success requires acceptable text quality, and stealing-based scrubbers often underperform the best general-scrubbing baselines.

arXiv cs.CR · 2d agoResearch

A GAN-Based Framework for Robust DDoS Attack Detection

WGAN-GP-generated adversarial DDoS traffic augments training data, improving detection resilience against evasion attempts.

Researchers built a DDoS detection framework combining Random Forests, deep neural ensembles, and Transformer-based models trained on CICDDoS2019 with synthetic adversarial flows generated by a Wasserstein GAN with gradient penalty. Hybrid datasets of benign, malicious, and generated traffic taught models more generalizable decision boundaries. Experiments showed improved accuracy and resilience against unseen adversarial traffic, validated on real-world generated flows.

arXiv cs.CR · 1d agoResearch