The Eval Index / Red Teaming & Safety / #238
Mobius-Dev/mobius-llm-adversity
by Mobius-Dev · Red Teaming & Safety · updated 1y ago
This repository documents a series of experiments focused on adversarial prompting and jailbreaks against large language models. It is part of my personal red teaming portfolio, intended to showcase prompt engineering techniques, jailbreak persistence, and alignment failure analysis.
25
momentum
115
stars
14
forks
#238
rank