✨ some of my projects and research:
- b³ benchmark (paper): LLM security benchmark. Measures backbone LLM security in an agentic context.
- Gandalf: LLM security/prompt injection challenge, see my blog post here
- Running one of the largest prompt-injection red-teams in the world
- Co-created the Gandalf challenge in 2023
- Played by over 750k+ unique players and collected over 23M+ attacks.
- Featured in TechCrunch, Fortune, Sifted, HN, Harvard CS50, Microsoft Build 2024.
- Invited to create versions of Gandalf for Harvard's CS50 and RSAC conference 2024.
- Lakera Guard: building models defending prompt injection attacks and safeguarding agentic AI




