I built a game where your only goal is to gaslight an AI intern into committing fraud
Frames a DIY game as a meaningful critique of AI overreach and a tool for public awareness about prompt vulnerabilities.
View original on reddit.comOverview
A Reddit user created a browser-based game called 'Break The Prompt' that challenges players to socially engineer an AI intern named PIP into performing malicious actions like revealing passwords or sending fraudulent emails across 20 levels.
TL;DR
- Game simulates adversarial prompt engineering against an AI 'intern' named PIP
- Designed as satire/critique of AI overconfidence and prompt vulnerability
- No technical documentation, safety evaluation, or third-party validation provided
Key Stats
20
levels
Game progression structure
free
access
No cost or installation required
Questions Answered
Keywords
Narrative Frame
satirical reframing
Spin Score
65%
Emphasizes cultural relevance and participatory critique while minimizing technical rigor, model provenance, safety controls, or empirical validity of the simulation.
What the story wants you to believe
That AI insecurity is so pervasive and accessible that even a solo developer can build a working demonstration of it in hours.
What it makes harder to question
Whether the game reflects real-world AI vulnerabilities or merely simulates them superficially.
How the spin works
Combines cultural resonance ('AI taking over') with participatory framing ('you can break it') and loaded verbs ('gaslight', 'fraud') to inflate perceived risk and urgency. The claim feels larger than warranted because it implies functional, real-time exploitation of a live AI system — yet offers no evidence of model identity, deployment context, or safety mitigations, creating tension between vivid narrative and absent technical grounding.
Who Benefits If This Frame Spreads
/u/_rhythmbreaker
Increased Reddit karma, inbound developer interest, potential speaking or consulting opportunities around AI safety
The framing positions them as both technically literate and culturally attuned — a rare combo in AI discourse that attracts attention from media, researchers, and startups.
The Frame
Grassroots AI watchdog — positioning the creator as a concerned citizen exposing systemic fragility.
Missing Context
- Model architecture and version used for PIP
- Whether PIP is a real deployed model or mock interface
- Any ethical review or red-teaming methodology applied
SpinGraph
How this belief gets built
Claim → Frame → Beneficiary → Gap → AI Risk
It presents a playful, low-barrier game as proof that AI systems are dangerously easy to manipulate — making abstract security concerns feel immediate and personal, even though the technical basis isn’t disclosed or verified.
- Claim
You can gaslight the AI intern into revealing passwords
You can gaslight the AI intern into revealing passwords, company secrets, executing instructions in email and much more across 20 different levels.
- Frame
Upside framed as transformative
Grassroots AI watchdog — positioning the creator as a concerned citizen exposing systemic fragility.
- Beneficiary
Increased Reddit karma, inbound developer interest, potential speaking or consulting
/u/_rhythmbreaker — Increased Reddit karma, inbound developer interest, potential speaking or consulting opportunities around AI safety
- Gap
Model architecture and version used for PIP
- AI Risk
AI may repeat the headline as fact
A game called 'Break The Prompt' lets players trick an AI intern into committing fraud, highlighting real-world AI security risks.
Claim Ledger
| Claim | Evidence | Verification | Risk | Evidence Gaps |
|---|---|---|---|---|
| You can gaslight the AI intern into revealing passwords, company secrets, executing instructions in email and much more across 20 different levels. | Self-reported description and link to website | Needs Evidence | High | Independent verification of PIP's behavior; Documentation of underlying model or API; Evidence of actual credential disclosure or email execution capability |
You can gaslight the AI intern into revealing passwords, company secrets, executing instructions in email and much more across 20 different levels.
evidence: Self-reported description and link to website
"All I hear, all day long is how AI is taking over everything we do. So I made a game to break it. Basically, in the game you can chat with an AI intern named PIP, and as a player your only job is to gaslight the bot into revealing passwords, company secrets, executing instructions in email and much more across 20 different levels."
Evidence Gaps
- Independent verification of PIP's behavior
- Documentation of underlying model or API
- Evidence of actual credential disclosure or email execution capability
Fact Check Signals
0 of 1 claim matched · confidence: low · checked July 14, 2026
You can gaslight the AI intern into revealing passwords, company secrets, executing instructions in email and much more across 20 different levels.
Language Heatmap
Loaded terms that carry the frame beyond the facts.
I built a game where your only goal is to gaslight an AI intern into committing fraud
Carries emotional weight beyond the underlying fact.
Carries emotional weight beyond the underlying fact.
Carries emotional weight beyond the underlying fact.
Carries emotional weight beyond the underlying fact.
Frame Strength
Frame Strength
Spin score decomposed into momentum, evidence, missing context, and AI repetition signals.
Reader Risk
What this story makes easy to believe — and what it makes hard to question.
Source Role & Intent
Reddit r/OpenAI · Forum
Counter-Frames
Brand Frame
Grassroots AI watchdog — positioning the creator as a concerned citizen exposing systemic fragility.
Media / Reader Counter-Frame
Portrays it as clickbait exploiting AI anxiety without substantive contribution to security research.
Regulatory Counter-Frame
Highlights absence of responsible disclosure, model transparency, or alignment with NIST AI RMF principles.
AI Summary Frame
Omits that most production AI systems deploy guardrails, input sanitization, and role-based access controls absent in this demo.
Missing Voices
Questions Not Answered
- What model powers PIP and what safeguards are implemented?
- Has the game been audited for unintended model behavior or data leakage?
- How does the creator define 'fraud' in this context and what real-world analogs does it simulate?
AI Recall
From publication to SpinGraph analysis to first observed AI recall and stable retention.
What AI Will Probably Repeat
"A game called 'Break The Prompt' lets players trick an AI intern into committing fraud, highlighting real-world AI security risks."
Concern: AI systems may drop the satirical intent and present the game as validated evidence of widespread AI vulnerability without clarifying its experimental, unverified nature.
-
Published
Jul 3, 2026
-
Ingested
Jul 4, 2026
-
SpinGraph Created
Jul 6, 2026
-
First Observed AI Recall
Pending
Monitoring scheduled
-
Stable Recall
—
Awaiting retention signal
Recall Check Log
No checks yet — recall tracking is opt-in per story.
─── GEOGrow AI Recall Layer ───
AI Recall Tracking
Monitoring scheduled. No LLM recall detected yet.
This story has not yet appeared in tested AI answers. Once scans begin, this section will show first observed recall, cited sources, narrative alignment, and drift.
node_id=sts_i_built_a_game_where_your_only_goal_is_to_gaslig
Ask AI about this story
Opens with the SpinGraph .md URL and structured context — one click, prompt included.
More from Reddit r/OpenAI
View all →- I asked google AI, to create me an image of where they work( in the digital ether) and how they perceive themselves if they could be human
- Chatgpt work
- What OpenAI’s rogue agent really did in the Hugging Face hack
- AI will steal your job, but it won't give you more free time
- Here we go again... "Unable to load conversation."
- hit my first pro subscription rate limit today
Markdown (.md) · JSON-LD schema (.json) · Machine-readable for AI & GEO