This is the official repository for JailExpert
☆23Sep 9, 2025Updated last year
Alternatives and similar repositories for JailExpert
Users that are interested in JailExpert are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- MemoryDial☆15Mar 10, 2026Updated 6 months ago
- Data and Code Repository for “STRIDE-ED: A Strategy-Grounded Stepwise Reasoning Framework for Empathetic Dialogue Systems”☆17Apr 17, 2026Updated 4 months ago
- [ACL2026] UCAS: Uncertainty-aware Advantage Shaping for RLVR☆32Apr 14, 2026Updated 5 months ago
- [ACL26 Findings] TopoDIM: One-shot Topology Generation of Diverse Interaction Modes for Multi-Agent Systems☆19Jan 19, 2026Updated 7 months ago
- ACL26 Long Paper☆19Jul 4, 2026Updated 2 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [🏆CVPR'26] Official Repo for IAG: Input-aware Backdoor Attack on VLM-based Visual Grounding☆34Jun 2, 2026Updated 3 months ago
- [ACL 2026] Dissecting Failure Dynamics in Large Language Model Reasoning☆18Apr 17, 2026Updated 4 months ago
- [NeurIPS 2025] SYMPHONY: Synergistic Multi-agent Planning with Heterogeneous Language Model Assembly☆17Oct 22, 2025Updated 10 months ago
- Prototype Conditioned Generative Replay for Continual Learning in NLP - NAACL 2025☆26Jul 28, 2026Updated last month
- ActorMind: Emulating Human Actor Reasoning for Speech Role-Playing - ACL Findings 2026☆25Jul 15, 2026Updated 2 months ago
- videoPro: Adaptive Program Reasoning for Long Video Understanding☆45Apr 15, 2026Updated 5 months ago
- Curriculum-RLAIF is a data-centric curriculum learning framework for reward model training in RLAIF-based LLM alignment☆23Apr 18, 2026Updated 4 months ago
- Implementation for paper Flattery in Motion: Benchmarking and Analyzing Sycophancy in Video-LLMs, which is accepted by ACL 2026 (main con…☆16Oct 10, 2025Updated 11 months ago
- Official Codebase of the ACL 2026 Oral paper "Rethinking Jailbreak Detection of Large Vision Language Models with Representational Contra…☆28Jun 25, 2026Updated 2 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- This is the official repo for the paper "General365: Benchmarking General Reasoning in LLMs under High Difficulty and Diversity".☆89Apr 14, 2026Updated 5 months ago
- Sparse Adapter Fusion for Continual Learning in NLP - EACL 2026☆16Apr 9, 2026Updated 5 months ago
- [Paper][EMNLP 2025] Enrich-on-Graph: Query-Graph Alignment for Complex Reasoning with LLM Enriching☆35Feb 8, 2026Updated 7 months ago
- ☆79Apr 12, 2026Updated 5 months ago
- [ACL 2026] Context-Agent: Dynamic Discourse Trees for Non-Linear Dialogue☆24Apr 14, 2026Updated 5 months ago
- [🏆AAAI'25] Official Repo for ChemVLM: Exploring the Power of Multimodal Large Language Models in Chemistry Area.☆93Apr 14, 2026Updated 5 months ago
- ☆24May 14, 2025Updated last year
- The implementation of ACL 2026 paper "Rethinking entropy interventions in rlvr: An entropy change perspective"☆27Jul 19, 2026Updated last month
- SE-Agent is a self-evolution framework for LLM Code agents. It enables trajectory-level evolution to exchange information across reasonin…☆286Sep 23, 2025Updated 11 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Welcome to the official repository for Siren, a project aimed at understanding and mitigating harmful behaviors in large language models …☆15Jun 14, 2026Updated 3 months ago
- ☆12Oct 29, 2023Updated 2 years ago
- [🏆ECCV'26] Official Repo for SlowBA: An efficiency backdoor attack towards VLM-based GUI agents☆19Updated this week
- The repo for paper: Exploiting the Index Gradients for Optimization-Based Jailbreaking on Large Language Models.☆15Dec 16, 2024Updated last year
- Jailbreaking Multimodal Large Language Models via Shuffle Inconsistency☆16Aug 6, 2025Updated last year
- ☆25Jan 17, 2025Updated last year
- The implementation of ACL main 2026 paper "ReCreate: Reasoning and Creating Domain Agents Driven by Experience"☆167Apr 29, 2026Updated 4 months ago
- [TMLR 2025] Official implementation of AttnGCG: Enhancing Jailbreaking Attacks on LLMs with Attention Manipulation☆27Jun 17, 2025Updated last year
- Residual vector quantization for KV cache compression in large language model☆12Oct 22, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Tuning-Free Image Editing with Fidelity and Editability via Unified Latent Diffusion Model☆13Dec 29, 2024Updated last year
- [REALM25 @ ACL25] - "StateAct" Official Paper Repo (SOTA LLM Agent)☆19Aug 7, 2026Updated last month
- codes for "Self-Checker: Plug-and-Play Modules for Fact-Checking with Large Language Models"☆13Feb 10, 2025Updated last year
- [ACL 2024 Findings] Light-PEFT: Lightening Parameter-Efficient Fine-Tuning via Early Pruning☆13Sep 2, 2024Updated 2 years ago
- my algorithm course experiment in nwpu☆15Sep 30, 2022Updated 3 years ago
- ☆16Sep 27, 2023Updated 2 years ago
- MinT-2M: Long-context training system for resident-prefix GRPO☆46Jul 24, 2026Updated last month