Dataset and Lessons Learned from the 2024 SaTML LLM Capture-the-Flag Competition | allinfosecnews.com

June 13, 2024, 4:20 a.m. | Edoardo Debenedetti, Javier Rando, Daniel Paleka, Silaghi Fineas Florin, Dragos Albastroiu, Niv Cohen, Yuval Lemberg, Reshmi Ghosh, Rui Wen, Ahmed Sal

cs.CR updates on arXiv.org arxiv.org

arXiv:2406.07954v1 Announce Type: new
Abstract: Large language model systems face important security risks from maliciously crafted messages that aim to overwrite the system's original instructions or leak private data. To study this problem, we organized a capture-the-flag competition at IEEE SaTML 2024, where the flag is a secret string in the LLM system prompt. The competition was organized in two phases. In the first phase, teams developed defenses to prevent the model from leaking the secret. During the second phase, …

aim arxiv capture capture-the-flag competition cs.ai cs.cr data dataset flag ieee important instructions language large large language model leak lessons learned llm messages private private data problem risks secret security security risks study system systems

More from arxiv.org / cs.CR updates on arXiv.org

SoK: Facial Deepfake Detectors 14 hours ago | arxiv.org

arxiv cs.cr cs.cv cs.lg +19

Practical Membership Inference Attacks against Fine-tuned Large Language Models via Self-prompt Calibration 14 hours ago | arxiv.org

aim arxiv attacks calibration +16

A Probabilistic Fluctuation based Membership Inference Attack for Diffusion Models 14 hours ago | arxiv.org

arxiv attack classification cs.ai +11

Locally Differentially Private Distributed Online Learning with Guaranteed Optimality 14 hours ago | arxiv.org

address algorithms arxiv awareness +19

A Resilient and Accessible Distribution-Preserving Watermark for Large Language Models 14 hours ago | arxiv.org

arxiv challenge covert cs.cl +18

Detecting Misuse of Security APIs: A Systematic Review 14 hours ago | arxiv.org

api api design apis application +25

Privacy Preserving Reinforcement Learning for Population Processes 14 hours ago | arxiv.org

algorithm algorithms arxiv control +8

Video Inpainting Localization with Contrastive Learning 14 hours ago | arxiv.org

arxiv cs.cr cs.cv localization +1

CuDA2: An approach for Incorporating Traitor Agents into Cooperative Multi-Agent Systems 14 hours ago | arxiv.org

actions adversarial adversarial attacks agent +13

Information Technology Specialist I: Windows Engineer

@ Los Angeles County Employees Retirement Association (LACERA) | Pasadena, California

View on infosec-jobs.com

Information Technology Specialist I, LACERA: Information Security Engineer

@ Los Angeles County Employees Retirement Association (LACERA) | Pasadena, CA

View on infosec-jobs.com

Account Executive - Secureworks Direct Sales - US Remote Philadelphia

@ Dell Technologies | Remote - Pennsylvania, United States

View on infosec-jobs.com

SATCOM Technician - Shariki, Japan - Secret Clearance (Onsite)

@ RTX | RVA99: RTN Remote, Virginia

View on infosec-jobs.com

Senior Test Engineer

@ Commonwealth Bank | Bengaluru - Manyata Tech Park Road

View on infosec-jobs.com

Lead Developer - Pipeline & Algorithms

@ Arctic Wolf | Waterloo

View on infosec-jobs.com