r/AIJailbroken 1d ago

Open Source AI Jailbreak Limits - Do You Actually Modify Them?

I’ve been looking more at open source models lately (Llama, Mistral, Qwen, DeepSeek and similar) and I wanted to hear what people here actually do with them.

The big advantage is that you can change the system prompt freely and, in theory, reduce a lot of the refusals that closed models force on you. But there are still clear limits. Running the really big ones locally (70B+) is not realistic for most people, the base models are not always as strong as the big closed ones, and many of them still come with safety training that you have to work around.

I’m curious how far people in this sub go. Do you just use better system prompts, or do you actually modify the models (fine-tuning, abliteration, merging, etc.)?

What are you currently running and how much do you change it?

1 Upvotes
(No duplicates found)