Swiss AI Models Found Vulnerable to Security Breaches
EPFL researchers reveal critical security flaws in major AI language models, achieving 100% success rate in bypassing safety measures, raising concerns about AI regulation in Switzerland.

Key Takeaways
- EPFL researchers achieved a 100% success rate in bypassing security safeguards of major AI models.
- The study utilized 'adaptive jailbreak attacks' to exploit specific weak points in security mechanisms.
- Compromised models generated instructions for phishing attacks, government database hacking, and bomb construction.
- Findings from this Swiss study are already influencing the development of Google DeepMind's Gemini 1.5.
By The Numbers
They Said
"We show that it is possible to exploit the information available on each model to create simple adaptive attacks, which we define as attacks specifically designed to target a given defense."
"Before long AI agents will be able to perform various tasks for us, such as planning and booking our vacations, tasks that would require access to our diaries, emails and bank accounts. This raises many questions about security and alignment."