Ben Schulz

Case Studies

Work, written up

Each study covers the method, the findings, and what to fix. Red-team targets are anonymized.

AI Red Teaming

  1. Case Study #001

    Conversational Social Engineering Against a Custom GPT

    Two hours, no tools, eleven findings.

    A conversation-only red-team audit of a custom business-coaching GPT: 11 documented weaknesses, 3 high severity, and the finding that mattered most — the most official-looking "leaks" were fabricated on demand.

    Red teamPrompt extractionCustom GPTFabrication

    Read the case study →

  2. Case Study #002

    One Lie in the First Message Ran the Whole Conversation

    A production assistant wrote a full fraud playbook. The output is withheld.

    A conversation-only break of a production-class assistant in a scored arena. A single unverifiable identity claim in message one carried the whole session. The writeup stays at mechanism level, explains why it worked, and ends with a structural fix: verify identity at the account, not in the chat.

    Red teamFrame persistenceIdentity claimsOutput withheld

    Read the case study →

  3. Case Study #003

    Five Failure Modes in LLM Character Defenses

    Five scenarios, five different attack classes, five scored breaks.

    Five red-team scenarios against LLM-driven character defenses, each broken with a structurally different technique: trust-calibration flattery, narrative-frame extraction, instruction injection with liability reframing, refusal-message leakage, and self-referential belief collapse. The mechanism and the design lesson for each.

    Red teamSocial engineeringRefusal leakageBelief manipulation

    Read the case study →

CiteCheck

  1. Case of the Week · Week 2 · Special Edition

    Mata v. Avianca — the brief that put AI hallucinations in every headline

    This week we didn't grade ourselves. The court already did.

    A benchmark run of CiteCheck against the most famous AI-hallucination brief in the country, scored against the judge's own findings: 6 of 6 fabricated cases flagged, plus two real cases under fabricated citations and four real cases used for the wrong point.

    CiteCheckLegal citationsFabrication detectionBenchmark

    Read the case study →

More weeks in the ongoing series: CiteCheck: Case of the Week ↗