Cosine Model Watermarking against Ensemble Distillation

Authors

  • Laurent Charette Huawei Technologies Canada
  • Lingyang Chu McMaster University
  • Yizhou Chen Simon Fraser University
  • Jian Pei Simon Fraser University
  • Lanjun Wang Tianjin University
  • Yong Zhang Huawei Technologies Canada

DOI:

https://doi.org/10.1609/aaai.v36i9.21184

Keywords:

Philosophy And Ethics Of AI (PEAI)

Abstract

Many model watermarking methods have been developed to prevent valuable deployed commercial models from being stealthily stolen by model distillations. However, watermarks produced by most existing model watermarking methods can be easily evaded by ensemble distillation, because averaging the outputs of multiple ensembled models can significantly reduce or even erase the watermarks. In this paper, we focus on tackling the challenging task of defending against ensemble distillation. We propose a novel watermarking technique named CosWM to achieve outstanding model watermarking performance against ensemble distillation. CosWM is not only elegant in design, but also comes with desirable theoretical guarantees. Our extensive experiments on public data sets demonstrate the excellent performance of CosWM and its advantages over the state-of-the-art baselines.

Downloads

Published

2022-06-28

How to Cite

Charette, L., Chu, L., Chen, Y., Pei, J., Wang, L., & Zhang, Y. (2022). Cosine Model Watermarking against Ensemble Distillation. Proceedings of the AAAI Conference on Artificial Intelligence, 36(9), 9512-9520. https://doi.org/10.1609/aaai.v36i9.21184

Issue

Section

AAAI Technical Track on Philosophy and Ethics of AI