Official github repo for SafeDialBench, a comprehensive multi-turn dialogue benchmark to evaluate LLMs' safety.
☆54May 12, 2025Updated last year
Alternatives and similar repositories for SafeDialBench-Dataset
Users that are interested in SafeDialBench-Dataset are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Reconsidering Overthinking: Penalizing Internal and External Redundancy in CoT Reasoning☆23Jun 23, 2026Updated last month
- RoboHiMan: A Hierarchical Evaluation Paradigm for Compositional Generalization in Long-Horizon Manipulation☆17Oct 16, 2025Updated 9 months ago
- A Framework of Continual Learning☆136Dec 9, 2025Updated 7 months ago
- The official code for "Steering Dialogue Dynamics for Robustness against Multi-turn Jailbreaking Attacks".☆18Jun 24, 2026Updated last month
- Implementation of the paper "Egoism, Utilitarianism and Egalitarianism in Multi-Agent Reinforcement Learning"☆21Aug 17, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ManifoldAlignmentStyleTransfer☆45Feb 24, 2022Updated 4 years ago
- Red Queen Dataset and data generation template