A curated list of LLM/MLLM guardrails, safety benchmarks, guard models, jailbreak attacks, moderation datasets, and evaluation tools.
By chatting or signing in you agree to the Terms and chat-message logging (revocable in History).