Code for our WOAH@ACL 2021 Paper on Data Integration for Toxic Comment Classification: Making More Than 40 Datasets Easily Accessible in One Unified Format
☆30Nov 25, 2021Updated 4 years ago
Alternatives and similar repositories for toxic-comment-collection
Users that are interested in toxic-comment-collection are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- MetricEval: A framework that conceptualizes and operationalizes four main components of metric evaluation, in terms of reliability and va…☆12Nov 6, 2023Updated 2 years ago
- [ACL 2023] Counterspeeches up my sleeve! Intent Distribution Learning and Persistent Fusion for Intent-Conditioned Counterspeech Generati…☆10Sep 23, 2023Updated 2 years ago
- Can we use explanations to improve hate speech models? Our paper accepted at AAAI 2021 tries to explore that question.☆251Jun 12, 2023Updated 3 years ago
- Explaining neural decisions contrastively to alternative decisions.☆24Mar 18, 2021Updated 5 years ago
- Hugging Face and Pyserini interoperability☆20May 18, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A repo to keep all resources about interpretability in NLP organised and up to date☆13Nov 22, 2020Updated 5 years ago
- ☆24Apr 2, 2024Updated 2 years ago
- A Dataset and Results for Classifying Emotions Across Languages☆10Jun 20, 2021Updated 5 years ago
- ☆19Feb 22, 2024Updated 2 years ago
- Code of Robust Lottery Tickets for Pre-trained Language Models (ACL2022)☆20Jul 18, 2022Updated 4 years ago
- ☆23Jun 12, 2023Updated 3 years ago
- Ongoing research training transformer language models at scale, including: BERT & GPT-2☆20Oct 23, 2023Updated 2 years ago
- Code for "Goodtriever: Toxicity Mitigation with Retrieval-augmented Language Models"☆25May 30, 2024Updated 2 years ago
- 2020 Summer Olympics medals per million people☆12Aug 8, 2021Updated 5 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Stanford Internet Observatory publications☆14Dec 2, 2021Updated 4 years ago
- Add Darkmode in Shiny App☆10Aug 23, 2022Updated 3 years ago
- Official repository for our EACL 2023 paper "LongEval: Guidelines for Human Evaluation of Faithfulness in Long-form Summarization" (https…☆45Aug 10, 2024Updated 2 years ago
- Code for FACTOID dataset paper in LREC 2022☆18Dec 19, 2022Updated 3 years ago
- ☆20Dec 16, 2020Updated 5 years ago
- The LM Contamination Index is a manually created database of contamination evidences for LMs.☆82Apr 11, 2024Updated 2 years ago
- ☆13Sep 13, 2018Updated 7 years ago
- The official repository of "Improving Large Language Models via Fine-grained Reinforcement Learning with Minimum Editing Constraint"☆39Jan 12, 2024Updated 2 years ago
- "Why do I feel offended?" - Korean Dataset for Offensive Language Identification (EACL2023 Findings)☆15May 14, 2023Updated 3 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Addressing common clinical biases in medical language models☆17Jul 27, 2024Updated 2 years ago
- Notebooks, slides and dataset of the CorrelAid Machine Learning Winter School☆12Jul 13, 2022Updated 4 years ago
- CounterGeDi is a pipeline that aims at controlling the counter speech generated to make it emotional, polite and detoxified. Paper accept…☆11Jul 19, 2022Updated 4 years ago
- ☆10Sep 17, 2022Updated 3 years ago
- Release of the ConditionalQA dataset☆21Nov 2, 2021Updated 4 years ago
- MRP - Multilevel + BART = BARP☆10Jan 1, 2023Updated 3 years ago
- Arabic Dialect Identification on AOC data.☆24Mar 2, 2019Updated 7 years ago
- This repository contains papers and resources pertaining to Hate speech research.☆44May 30, 2021Updated 5 years ago
- DeClarE: Debunking Fake News and False Claims using Evidence-Aware Deep Learning☆23Aug 23, 2023Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Towards cross-lingual distributed representations without parallel text trained with adversarial autoencoders☆22Aug 11, 2016Updated 10 years ago
- Fortifying Toxic Speech Detectors Against Veiled Toxicity☆11Oct 21, 2020Updated 5 years ago
- Fine-tuned transformers for protest event detection.☆11Mar 9, 2021Updated 5 years ago
- Workshop Materials "Advanced Bayesian Statistical Modeling in R and Stan "☆12Nov 23, 2023Updated 2 years ago
- Debug DeepSpeed-Chat step by step in IDE (在IDE里一步一步调试DeepSpeed-Chat)☆10Apr 17, 2023Updated 3 years ago
- TGLS: Unsupervised Text Generation by Learning from Search☆25Jan 5, 2021Updated 5 years ago
- Official repository of "Distort, Distract, Decode: Instruction-Tuned Model Can Refine its Response from Noisy Instructions", ICLR 2024 Sp…☆21Mar 7, 2024Updated 2 years ago