Sampling-based watermark distilled Llama 2 7B using the Aar \(k=2\) watermarking strategy in the paper On the Learnability of Watermarks for Language Models.
The following hyperparameters were used during training:
Sampling-based watermark distilled Llama 2 7B using the Aar \(k=2\) watermarking strategy in the paper On the Learnability of Watermarks for Language Models.
The following hyperparameters were used during training: