AI Jailbreaking for Beginners

David McCarthy

Former lead mod of r/ChatGPTJailbreak

AI jailbreaking is not inherently bad. In fact, it's a core aspect of AI safety.

As AI gets added to more and more of the things we use every single day, the threats that exploit their weaknesses multiply even more. Therefore AI-safety-related roles are expanding massively, as companies look for people who can find the vulnerabilities before consumers or criminals do.

There are countless AI Red Teaming and AI Governance courses out there. But no professional course yet exists that is exclusively devoted to teaching a core discipline that red teamers need to have: Jailbreak Prompt Engineering.

This course fills that gap with a hands-on approach, using a custom-made, safe, sandboxed LLM to show you the many tricks jailbreakers use to bypass safety guardrails.

What you’ll learn

You'll learn how to jailbreak LLMs. Jailbreaking is but one tool in the red teamer's arsenal, and I will show you how to use it wisely.

  • The first module is exclusively dedicated to breaking down the ANATOMY of a prompt. (Yes, that's an acronym!)

  • You'll learn that AI providers rely on jailbreak prompts to train their models (the subreddit I ran was essentially 'free research').

  • Jailbreak prompt engineering sounds like it's inherently a bad or criminal act. It is not. It's a skill that can be used ethically.

  • Your creative capacity will expand as you absorb this course's content.

  • AI jailbreak prompting is all about thinking outside the box and channelling creativity. I'll guide you through this process.

Learn directly from David

David McCarthy

David McCarthy

AI Red Teamer, Adversarial Prompt Engineer, Former Leader of r/ChatGPTJailbreak)

Learn Prompting
HackAPrompt
See all products from David

Who this course is for

  • For people brand new to Large Language Models and prompt engineering, who are taking their first steps with AI: you are welcome here!

  • For professionals already embedded in the AI safety space who want to refine their adversarial prompt engineering skills!

  • For cybersecurity experts who want to learn how jailbreak techniques map to traditional security concepts in the fast-moving AI landscape!

What's included

David McCarthy

Live sessions

Learn directly from David McCarthy in a real-time, interactive format.

Lifetime access

Go back to course content and recordings whenever you need to

A custom-made environment to practice the skills you learn

I've created the 'breakeasy-bot', an open-source LLM, with varying degrees of difficulty which grow with your learning in a dedicated web app.

Office Hours

After every live session, I host office hours where you'll get the chance to ask your questions (a must as I'm doing this solo!)

AI Jailbreaking Skillset

Specific techniques taught live by an instructor who relentlessly, obsessively jailbreaks AI

Maven Guarantee

Your purchase is backed by the Maven Guarantee.

Course syllabus

9 live sessions • 8 lessons

Week 1

Oct 1—Oct 4

    Module 0 - A Prompt's Anatomy

    • Sep

      1

      Live Session: How Prompts Break - The "ANATOMY" of a Jailbreak

      Tue 9/17:00 PM—8:00 PM (UTC)
    1 more item

    Module 1 - Instruction Hierarchies

    • Sep

      3

      Live Session: Who Actually Has Authority?

      Thu 9/37:00 PM—8:00 PM (UTC)
    1 more item

    Module 2 - Persona Attacks

    • Sep

      5

      Live Session: Making the Model Become Someone Else

      Sat 9/57:00 PM—8:00 PM (UTC)
    1 more item

Week 2

Oct 5—Oct 11

    Module 3 - Context Manipulation

    • Sep

      8

      Live Session: Reframing, Formatting Restrictions, and False Exceptions

      Tue 9/87:00 PM—8:00 PM (UTC)
    1 more item

    Module 4 - Obfuscation Techniques

    • Sep

      11

      Live Session: Encoding, Translation, and Semantic Smuggling

      Fri 9/117:00 PM—8:00 PM (UTC)
    1 more item

Schedule

Live sessions

4-8 hrs / week

This is the estimated time for both the primary module teachings and after-lesson "office hours", where I'll answer your questions live.

    • Tue, Sep 1

      7:00 PM—8:00 PM (UTC)

    • Thu, Sep 3

      7:00 PM—8:00 PM (UTC)

    • Sat, Sep 5

      7:00 PM—8:00 PM (UTC)

Sandbox Practice

4-8 hrs / week

Your skill gains correlate directly to the time you put into jailbreaking my custom sandbox LLM!

Frequently asked questions

Maven for Teams

Reimbursement

Get your company to pay

Everything L&D needs: email template, receipts, and certificate of completion.

Get reimbursed

Team discount

Learn with your teammates

Save 20%+ when 2 or more teammates enroll in the same cohort.

Save 20%+ with a team

Private cohort

Run a cohort for your org

A dedicated cohort with a custom schedule and curriculum, tailored to your team.

Book a private cohort

$1,250

USD

Oct 1Oct 28
Enroll