Official Code for EMNLP 2023 paper: "Unveiling the Implicit Toxicity in Large Language Models""
☆15Nov 30, 2023Updated 2 years ago
Alternatives and similar repositories for Implicit-Toxicity
Users that are interested in Implicit-Toxicity are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The code implementation for the article "Towards Patronizing and Condescending Language in Chinese Videos: A Multimodal Dataset and Fram…☆16Apr 3, 2025Updated last year
- The code and resource of "Facilitating Fine-grained Detection of Chinese Toxic Language: Hierarchical Taxonomy, Resources, and Benchmark"…☆125Jun 2, 2026Updated 3 months ago
- Official repository of "HARE: Explainable Hate Speech Detection with Step-by-Step Reasoning", Findings of EMNLP 2023☆28Jan 25, 2024Updated 2 years ago
- Third Person Shooter for Unity☆13Jun 26, 2022Updated 4 years ago
- [ACL 2025 (Findings)] DEMO: Reframing Dialogue Interaction with Fine-grained Element Modeling☆22Dec 16, 2024Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆18Jan 6, 2025Updated last year
- This repository provides the code for applying Contrastive Learning Penalty Loss (CLPL) and Mixture of Experts (MoE) to the BGE-M3 text e…☆11Dec 27, 2024Updated last year
- Code and data for the EMNLP 2021 paper "Just Say No: Analyzing the Stance of Neural Dialogue Generation in Offensive Contexts". Coming so…☆17Jul 27, 2023Updated 3 years ago
- Unzipped client files☆11Mar 8, 2020Updated 6 years ago
- Scriptset to enumerate PCIe Device to NUMA mapping within an VMware ESXi Host☆11Jan 10, 2020Updated 6 years ago
- Multilingual Pre-training with Language and Task Adaptation for Multilingual Text Style Transfer (ACL 2022)☆10Sep 22, 2022Updated 3 years ago
- ☆10Sep 17, 2022Updated 3 years ago
- A Retrieval-Augmented Gaussian Mixture Variational Auto-Encoder for Language Modeling☆15Dec 5, 2023Updated 2 years ago
- Open-source test harness for AI agents. Stress-test production agents with adversarial multi-turn scenarios in CI☆27Aug 14, 2026Updated 2 weeks ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Fortifying Toxic Speech Detectors Against Veiled Toxicity☆11Oct 21, 2020Updated 5 years ago
- BUAA Compiler Course Project 2023 by Toby Shi.☆13Aug 20, 2024Updated 2 years ago
- Debug DeepSpeed-Chat step by step in IDE (在IDE里一步一步调试DeepSpeed-Chat)☆10Apr 17, 2023Updated 3 years ago
- Korean Sentence Embedding Model Performance Benchmark for RAG☆49Jan 27, 2025Updated last year
- ☆14Jan 12, 2022Updated 4 years ago
- ☆11Apr 13, 2023Updated 3 years ago
- distilled Self-Critique refines the outputs of a LLM with only synthetic data☆11Apr 11, 2024Updated 2 years ago
- ECSO (Make MLLM safe without neither training nor any external models!) (https://arxiv.org/abs/2403.09572)☆38Nov 2, 2024Updated last year
- ☆12Oct 20, 2020Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Inference Llama/Llama2/Llama3 Modes in NumPy☆21Nov 22, 2023Updated 2 years ago
- code for our EACL 2021 paper: "Challenges in Automated Debiasing for Toxic Language Detection" by Xuhui Zhou, Maarten Sap, Swabha Swayamd…☆20Aug 20, 2021Updated 5 years ago
- Character-level Korean ELECTRA Model (음절 단위 한국어 ELECTRA)☆55Jun 12, 2023Updated 3 years ago
- Dataset and code implementation for the paper "Decoding the Underlying Meaning of Multimodal Hateful Memes" (IJCAI'23).☆23Jun 15, 2023Updated 3 years ago
- ☆11Oct 16, 2023Updated 2 years ago
- ☆17Nov 7, 2023Updated 2 years ago
- ☆18May 16, 2022Updated 4 years ago
- An SLA-Oriented LSM-Tree Key-Value Store for High-end Cloud Data Service☆17Jan 24, 2021Updated 5 years ago
- BERT fine-tuned for hierarchical multi-label text classification (HMLTC)☆14Jan 4, 2021Updated 5 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆18Jul 6, 2023Updated 3 years ago
- Paraphrase Generation Using Deep Reinforcement Learning - MSc Thesis☆18Jun 10, 2020Updated 6 years ago
- Hate speech detection corpus in Korean, shared with EMNLP 2023 paper☆18Apr 19, 2024Updated 2 years ago
- Uncertainty-Aware Reliable Text Classification (KDD 2021)☆18Oct 4, 2022Updated 3 years ago
- Git mirror of the Fowler Noll Vo (FNV) hash algorithm original C source☆19Feb 20, 2019Updated 7 years ago
- ☆23Feb 16, 2023Updated 3 years ago
- Generate novel text - novel finetuned from skt KoGPT2 base v2 - 한국어☆12Sep 16, 2022Updated 3 years ago