Sampling-based watermark distilled Llama 2 7B using the KGW \(k=0, \gamma=0.25, \delta=1\) watermarking strategy in the paper On the Learnability of Watermarks for Language Models.
The following hyperparameters were used during training:
Sampling-based watermark distilled Llama 2 7B using the KGW \(k=0, \gamma=0.25, \delta=1\) watermarking strategy in the paper On the Learnability of Watermarks for Language Models.
The following hyperparameters were used during training: