ICLR Poster Copyright-Protected Language Generation via Adaptive Model Fusion

Poster

Copyright-Protected Language Generation via Adaptive Model Fusion

Javier Abad · Konstantin Donhauser · Francesco Pinto · Fanny Yang

Hall 3 + Hall 2B #537

[ Abstract ] [ Project Page ]

Thu 24 Apr 7 p.m. PDT — 9:30 p.m. PDT

Oral presentation: Oral Session 4E
Fri 25 Apr 12:30 a.m. PDT — 2 a.m. PDT

Abstract:

The risk of language models reproducing copyrighted material from their training data has led to the development of various protective measures. Among these, inference-time strategies that impose constraints via post-processing have shown promise in addressing the complexities of copyright regulation. However, they often incur prohibitive computational costs or suffer from performance trade-offs. To overcome these limitations, we introduce Copyright-Protecting Model Fusion (CP-Fuse), a novel approach that combines models trained on disjoint sets of copyrighted material during inference. In particular, CP-Fuse adaptively aggregates the model outputs to minimize the reproduction of copyrighted content, adhering to a crucial balancing property to prevent the regurgitation of memorized data. Through extensive experiments, we show that CP-Fuse significantly reduces the reproduction of protected material without compromising the quality of text and code generation. Moreover, its post-hoc nature allows seamless integration with other protective measures, further enhancing copyright safeguards. Lastly, we show that CP-Fuse is robust against common techniques for extracting training data.

Live content is unavailable. Log in and register to view live content