'Many-Shot Jailbreaking' Defeats Gen AI Security Guardrails | allinfosecnews.com

April 4, 2024, 5:19 p.m. |

BankInfoSecurity.com RSS Syndication www.bankinfosecurity.com

'Fictitious Dialogue' About Harmful Content Subverts Defenses, Researchers Find
After testing safety features built into generative artificial intelligence tools developed by the likes of Anthropic, OpenAI and Google DeepMind, researchers have discovered that a technique called "many-shot jailbreaking" can be used to defeat safety guardrails and obtain prohibited content.

ai security anthropic artificial artificial intelligence called can defenses features gen gen ai generative generative artificial intelligence google google deepmind guardrails intelligence jailbreaking openai researchers safety security testing tools

More from www.bankinfosecurity.com / BankInfoSecurity.com RSS Syndication

US and Allies Issue Cyber Alert on Threats to OT Systems an hour ago | www.bankinfosecurity.com

alert allies america attacks +26

Verizon DBIR: Cyber Defenders Are Facing Exploit Fatigue 2 hours ago | www.bankinfosecurity.com

critical critical vulnerabilities cyber cyberattacks +17

Corelight Gets $150M to Expand Detection, Improve Workflows 2 hours ago | www.bankinfosecurity.com

and response capabilities corelight detection +19

Lawmakers Grill UnitedHealth CEO on Change Healthcare Attack 3 hours ago | www.bankinfosecurity.com

attack ceo change change healthcare +17

GitLab Hackers Use 'Forgot Your Password' to Hijack Accounts 3 hours ago | www.bankinfosecurity.com

accounts agency attacker cisa +19

Correlating Cyber Investments With Business Outcomes 9 hours ago | www.bankinfosecurity.com

business ceo cisos cyber +13

Qantas Airways Says App Showed Customers Each Other's Data 11 hours ago | www.bankinfosecurity.com

airline airways app breach +11

Verizon Breach Report: Vulnerability Hacks Tripled in 2023 19 hours ago | www.bankinfosecurity.com

alex author breach data +23

How Personal Branding Can Elevate Your Tech Career 1 day, 2 hours ago | www.bankinfosecurity.com

brand branding can candidates +8

Azure DevSecOps Cloud Engineer II

@ Prudent Technology | McLean, VA, USA

View on infosec-jobs.com

Security Engineer III - Python, AWS

@ JPMorgan Chase & Co. | Bengaluru, Karnataka, India

View on infosec-jobs.com

SOC Analyst (Threat Hunter)

@ NCS | Singapore, Singapore

View on infosec-jobs.com

Managed Services Information Security Manager

@ NTT DATA | Sydney, Australia

View on infosec-jobs.com

Senior Security Engineer (Remote)

@ Mattermost | United Kingdom

View on infosec-jobs.com

Penetration Tester (Part Time & Remote)

@ TestPros | United States - Remote

View on infosec-jobs.com