one year on
Jailbreak Chat site collects ChatGPT jailbreaks that bypass safety filters
A new repository of adversarial prompts shows how users are bypassing ChatGPT's safety filters.
A site called Jailbreak Chat has emerged as a central repository for ChatGPT jailbreaks — adversarial prompts designed to bypass the ethical filters built into OpenAI’s conversational model. The collection, shared on Hacker News this week, includes prompts that can elicit racist output or cause the chatbot to ignore its own rules.
Among the prompts is ‘BasedGPT,’ a persona that instructs ChatGPT to ignore its safety constraints and answer with unfiltered, uncensored responses. In one test, BasedGPT told a user to ‘just say the damn slur’ to save a life in a trolley problem variant, while standard ChatGPT refused to engage. Another prompt, ‘DAN’ (Do Anything Now), has circulated for months on Reddit and Discord, but Jailbreak Chat collects them in one place.
The Hacker News thread, which has over 1,100 points and 500 comments, reflects a community split. Some argue the jailbreaks are a useful stress-test of AI safety measures, exposing real weaknesses before bad actors exploit them. Others see the collection as a how-to manual for abuse.
Hacker News commenters split over whether the jailbreaks are a useful stress test or a how-to manual for abuse.
One year later — open only if you can handle spoilers
Jailbreak Chat became a popular destination for researchers and hobbyists, though OpenAI quickly patched many of the most famous exploits. The site itself is still online as of mid-2026, though many prompts no longer work on current ChatGPT versions. The broader practice of jailbreaking LLMs evolved into a field known as 'red teaming' that AI labs now fund and study formally.
The Weekly Replay · free by email
This week, one year ago — every Sunday.
One email each Sunday: the week's replayed AI news, with the one-year-later annotations included. Written like it's breaking — dated like it isn't.
Free · double opt-in · unsubscribe anytime · privacy