DuFFin: A Dual-Level Fingerprinting Framework for LLMs IP Protection

Yuliang Yan; Haochun Tang; Shuo Yan; Enyan Dai

DuFFin: A Dual-Level Fingerprinting Framework for LLMs IP Protection

Yuliang Yan, Haochun Tang, Shuo Yan, Enyan Dai

TL;DR

This work tackles the problem of protecting LLM intellectual property under black-box access. It introduces DuFFin, a dual-level fingerprinting framework that combines Trigger-DuFFin (trigger-prompt fingerprints) and Knowledge-DuFFin (domain knowledge fingerprints) to non-invasively verify ownership of pirated models. Fingerprints are extracted via a secret-key–driven process and merged with a weighted distance to decide ownership, enabling verification even when models are fine-tuned, quantized, or RLHF-aligned. Experiments on multiple protected models and unseen variants show high IP-ROC scores (often >0.95) and strong generalization, demonstrating DuFFin’s practical potential for LLM IP protection.

Abstract

Large language models (LLMs) are considered valuable Intellectual Properties (IP) for legitimate owners due to the enormous computational cost of training. It is crucial to protect the IP of LLMs from malicious stealing or unauthorized deployment. Despite existing efforts in watermarking and fingerprinting LLMs, these methods either impact the text generation process or are limited in white-box access to the suspect model, making them impractical. Hence, we propose DuFFin, a novel $\textbf{Du}$al-Level $\textbf{Fin}$gerprinting $\textbf{F}$ramework for black-box setting ownership verification. DuFFin extracts the trigger pattern and the knowledge-level fingerprints to identify the source of a suspect model. We conduct experiments on a variety of models collected from the open-source website, including four popular base models as protected LLMs and their fine-tuning, quantization, and safety alignment versions, which are released by large companies, start-ups, and individual users. Results show that our method can accurately verify the copyright of the base protected LLM on their model variants, achieving the IP-ROC metric greater than 0.95. Our code is available at https://github.com/yuliangyan0807/llm-fingerprint.

DuFFin: A Dual-Level Fingerprinting Framework for LLMs IP Protection

TL;DR

Abstract

DuFFin: A Dual-Level Fingerprinting Framework for LLMs IP Protection

TL;DR

Abstract

Paper Structure

Table of Contents

Figures (5)