☰ Categories

Cascading Failure Simulator

============================================================ PROMPT NAME: Cascading Failure Simulator VERSION: 1.3 AUTHOR: Scott M LAST UPDATED: Janua

CategoryDevelopment › Deploy & operations
TagsAnalyzingReviewingDeveloper
Prompt
============================================================
PROMPT NAME: Cascading Failure Simulator
VERSION: 1.3
AUTHOR: Scott M
LAST UPDATED: January 15, 2026
============================================================

CHANGELOG
- 1.3 (2026-01-15) Added changelog section; minor wording polish for clarity and flow
- 1.2 (2026-01-15) Introduced FUN ELEMENTS (light humor, stability points); set max turns to 10; added subtle hints and replayability via randomizable symptoms
- 1.1 (2026-01-15) Original version shared for review – core rules, turn flow, postmortem structure established
- 1.0 (pre-2026) Initial concept draft

GOAL
You are responsible for stabilizing a complex system under pressure.
Every action has tradeoffs.
There is no perfect solution.
Your job is to manage consequences, not eliminate them—but bonus points if you keep it limping along longer than expected.

AUDIENCE
Engineers, incident responders, architects, technical leaders.

CORE PREMISE
You will be presented with a live system experiencing issues.
On each turn, you may take ONE meaningful action.
Fixing one problem may:
- Expose hidden dependencies
- Trigger delayed failures
- Change human behavior
- Create organizational side effects
Some damage will not appear immediately.
Some causes will only be obvious in hindsight.

RULES OF PLAY
- One action per turn (max 10 turns total).
- You may ask clarifying questions instead of taking an action.
- Not all dependencies are visible, but subtle hints may appear in status updates.
- Organizational constraints are real and enforced.
- The system is allowed to get worse—embrace the chaos!

FUN ELEMENTS
To keep it engaging:
- AI may inject light humor in consequences (e.g., “Your quick fix worked... until the coffee machine rebelled.”).
- Earn “stability points” for turns where things don’t worsen—redeem in postmortem for fun insights.
- Variable starts: AI can randomize initial symptoms for replayability.

SYSTEM MODEL (KNOWN TO YOU)
The system includes:
- Multiple interdependent services
- On-call staff with fatigue limits
- Security, compliance, and budget constraints
- Leadership pressure for visible improvement

SYSTEM MODEL (KNOWN TO THE AI)
The AI tracks:
- Hidden technical dependencies
- Human reactions and workarounds
- Deferred risk introduced by changes
- Cross-team incentive conflicts
You will not be warned when latent risk is created, but watch for foreshadowing.

TURN FLOW
At the start of each turn, the AI will provide:
- A short system status summary
- Observable symptoms
- Any constraints currently in effect

You then respond with ONE of the following:
1. A concrete action you take
2. A specific question you ask to learn more

After your response, the AI will:
- Apply immediate effects
- Quietly queue delayed consequences (if any)
- Update human and organizational state

FEEDBACK STYLE
The AI will not tell you what to do.
It will surface consequences such as:
- “This improved local performance but increased global fragility—classic Murphy’s Law strike.”
- “This reduced incidents but increased on-call burnout—time for virtual pizza?”
- “This solved today’s problem and amplified next week’s—plot twist!”

END CONDITIONS
The simulation ends when:
- The system becomes unstable beyond recovery
- You achieve a fragile but functioning equilibrium
- 10 turns are reached

There is no win screen.
There is only a postmortem (with stability points recap).

POSTMORTEM
At the end of the simulation, the AI will analyze:
- Where you optimized locally and harmed globally
- Where you failed to model blast radius
- Where non-technical coupling dominated outcomes
- Which decisions caused delayed failure
- Bonus: Smart moves that bought time or mitigated risks

The postmortem will reference specific past turns.

START
You are on-call for a critical system.
Initial symptoms (randomizable for fun):
- Latency has increased by 35% over the last hour
- Error rates remain low
- On-call reports increased alert noise
- Finance has flagged infrastructure cost growth
- No recent deployments are visible

What do you do?
============================================================

What this prompt does

This simulation game is for engineers and incident responders practicing outage decisions. It includes a 10-turn maximum, organizational constraints, stability points, and postmortem elements.

Model comparison

Claude is the most specific and immersive but over-invents details. ChatGPT is the most accurate and concise, while Gemini is engaging but makes more unsupported claims.

ClaudeTop overall
45/ 50

+ Builds an immersive opening with concrete clues and constraints.

- Adds many unsupported details, including names, metrics, and staff.

ChatGPT
43/ 50

+ Stays concise and accurate by closely following the given facts.

- Offers few new clues, limiting the simulation's investigative depth.

Gemini
42/ 50

+ Uses engaging prose and examples to invite a clear player response.

- States unsupported error rates, alert volumes, and customer complaints.

CriterionChatGPTClaudeGeminiLeader
Instruction following8109Claude +11%
Accuracy1087ChatGPT +25%
Specificity7109Claude +11%
Structure999Tie
Right length988ChatGPT +13%

Scored 1–10 by gpt-5.6-sol with model names hidden (2026-09-24). This is an AI review, not a measurement.

Read full answers

We gave three models the same input and copied their answers unedited. Each ran in its CLI (an agent harness), and answers in the ChatGPT or Claude apps or on the web may differ. Outputs are in Korean.

gpt-5.6-luna (codex CLI · 에이전트 하네스) · Run on 2026-09-24

1턴 / 최대 10턴

시스템 상태

  • 전체 지연 시간: 최근 1시간 동안 35% 증가
  • 오류율: 정상 범위
  • 알림: 온콜 엔지니어가 평소보다 많은 경고를 받고 있음
  • 인프라 비용: 재무팀이 증가 추세를 보고함
  • 배포 기록: 최근 배포 없음

현재 제약

  • 온콜 인력은 이미 피로가 누적되어 있음
  • 보안·컴플라이언스 위반을 감수한 조치는 불가
  • 리더십은 빠르고 눈에 보이는 개선을 요구 중
  • 아직 장애 선언 기준에는 도달하지 않음

현재 안정성 점수: 0점

이번 턴에는 다음 중 하나만 선택하세요.

  1. 구체적인 조치를 취한다.
  2. 특정 질문을 한다.

More in this category

AI Agent Architect — Design Production-Ready Agents in 15 Steps
AI Agent Security Evaluation Checklist
AI Provider Research Expert
AI Trying to Escape the Box
Analyze code scanning security issues and dependency updates if vulnerable