Laidlaw Research Poster- Guo Yixuan
This project investigates how canary-based attribution can still identify AI agents when their inputs are paraphrased. We find that adding a small amount of canary redundancy improves detection far more than tightening thresholds, making redundancy the key design lever for robust attribution.