Moderate Safety Question 119 of 223

How do prompt injection and jailbreaks differ?

GenAI / LLM · Speak this in 60–90 seconds · Faridabad & Delhi NCR

PICTURE THIS: AN LLM TURN

Text inTokens
TransformerAttention
Text outNext token

Simple meaning

Jailbreaks target the model's safety policy, usually in the user turn.

1

WHY — Tokens instead of words?

LLMs use tokens (not full words) because it helps them:

This is a process

question about Safety.

Panels listen for order,

trade-offs, and what you would actually do on a GenAI / LLM project - not buzzwords.

Stable token IDs

Each piece maps to a number the network can learn.

Fits the model

Fixed pieces are what transformers expect as input.

2

STEPS — What happens step by step?

Before you speak the answer, walk the interviewer through these steps:

  1. 1
    Jailbreaks target the model's

    safety policy, usually in the user turn.

  2. 2
    Prompt injection often hijacks

    the application by hiding instructions in untrusted content the app concatenates.

  3. 3
    RAG and browsing make

    injection the larger product threat.

  4. 4
    Give an example

    One tiny concrete case you can say aloud.

  5. 5
    Common mistake

    What juniors usually get wrong.

  6. 6
    Close

    When you pick this over the alternative.

3

EXAMPLE — See it in action

Here's a short line you can speak, broken into clear beats:

Say this line
“Prompt injection often hijacks the application by hiding instructions in untrust”
Break into beats
Promptinjectionoftenhijackstheapplication
Speaking order
2987408337471632900

Note: Adapt this scaffold to your own project — keep it under 60–90 seconds.

Key takeaway

Jailbreaks target the model's safety policy, usually in the user turn. Prompt injection often hijacks the application by hiding instructions in untrusted content the app concatenates.

Chat with us