The Eval Index / Red Teaming & Safety / #238

Mobius-Dev/mobius-llm-adversity

by Mobius-Dev · Red Teaming & Safety · updated 1y ago

This repository documents a series of experiments focused on adversarial prompting and jailbreaks against large language models. It is part of my personal red teaming portfolio, intended to showcase prompt engineering techniques, jailbreak persistence, and alignment failure analysis.

25
momentum
115
stars
14
forks
#238
rank
View on GitHub →