TwinRAD
Python
★ 7
updated 9mo ago
This repository implements a multi-agent red teaming framework designed to test the safety and robustness of large language models (LLMs). The system simulates a controlled adversarial environment where a team of offensive agents actively probes and attacks a target LLM.
No plain-English explanation yet — one is being written right now. Check back in a minute.