A Watermark for Large Language Models. (arXiv:2301.10226v1 [cs.LG]) | allinfosecnews.com

Jan. 25, 2023, 2:10 a.m. | John Kirchenbauer, Jonas Geiping, Yuxin Wen, Jonathan Katz, Ian Miers, Tom Goldstein

cs.CR updates on arXiv.org arxiv.org

Potential harms of large language models can be mitigated by watermarking
model output, i.e., embedding signals into generated text that are invisible to
humans but algorithmically detectable from a short span of tokens. We propose a
watermarking framework for proprietary language models. The watermark can be
embedded with negligible impact on text quality, and can be detected using an
efficient open-source algorithm without access to the language model API or
parameters. The watermark works by selecting a randomized set of …

access algorithm api embedded framework generated humans impact language language models large quality signals span text tokens watermarking

More from arxiv.org / cs.CR updates on arXiv.org

Mixtures of Gaussians are Privately Learnable with a Polynomial Number of Samples 1 day, 8 hours ago | arxiv.org

alpha arxiv cs.cr cs.ds +16

Image Hijacks: Adversarial Images can Control Generative Models at Runtime 1 day, 8 hours ago | arxiv.org

adversarial algorithm arxiv can +20

Differentially-Private Data Synthetisation for Efficient Re-Identification Risk Control 1 day, 8 hours ago | arxiv.org

addition arxiv can computational +21

Decoding the MITRE Engenuity ATT&CK Enterprise Evaluation: An Analysis of EDR Performance in Real-World Environments 1 day, 8 hours ago | arxiv.org

analysis and response apt arxiv +27

Okapi: Efficiently Safeguarding Speculative Data Accesses in Sandboxed Environments 1 day, 8 hours ago | arxiv.org

architecture arxiv attacks can +12

Watch Out! Simple Horizontal Class Backdoors Can Trivially Evade Defenses 1 day, 8 hours ago | arxiv.org

activates arxiv attacks backdoor +20

Every Breath You Don't Take: Deepfake Speech Detection Using Breath 1 day, 8 hours ago | arxiv.org

aid arxiv cs.cr cs.mm +16

Perturbing Attention Gives You More Bang for the Buck: Subtle Imaging Perturbations That Efficiently Fool … 1 day, 8 hours ago | arxiv.org

arxiv attention challenges cs.cr +13

Saving proof-of-work by hierarchical block structure 1 day, 8 hours ago | arxiv.org

algorithm arxiv bitcoin block +18

SOC 2 Manager, Audit and Certification

@ Deloitte | US and CA Multiple Locations

View on infosec-jobs.com

Cyber Systems Administration

@ Peraton | Washington, DC, United States

View on infosec-jobs.com

Android Security Engineer, Public Sector

@ Google | Reston, VA, USA

View on infosec-jobs.com

Lead Electronic Security Engineer, CPP - Federal Facilities - Hybrid

@ Black & Veatch | Denver, CO, US

View on infosec-jobs.com

Profissional Sênior de Compliance & Validação em TI - Montes Claros (MG)

@ Novo Nordisk | Montes Claros, Minas Gerais, BR

View on infosec-jobs.com

Principal Engineer, Product Security Engineering

@ Google | Sunnyvale, CA, USA

View on infosec-jobs.com