Few-shot Learning with Auxiliary Data
☆31Dec 8, 2023Updated 2 years ago
Alternatives and similar repositories for FLAD
Users that are interested in FLAD are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Implementation of "Visualize Before You Write: Imagination-Guided Open-Ended Text Generation".☆17Feb 3, 2023Updated 3 years ago
- Data and code for APPDIA: A Discourse-aware Transformer-based Style Transfer Model for Offensive Social Media Conversations (COLING 2022)…☆13Sep 8, 2022Updated 3 years ago
- Official repository for MATES: Model-Aware Data Selection for Efficient Pretraining with Data Influence Models [NeurIPS 2024]☆80Nov 14, 2024Updated last year
- ☆14Apr 22, 2024Updated 2 years ago
- Emotion-Aware Dialogue Response Generation by Multi-Task Learning☆13Jan 22, 2022Updated 4 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- My Implementation for the paper EDA: Easy Data Augmentation Techniques for Boosting Performance on Text Classification Tasks using Tensor…☆12Mar 18, 2022Updated 4 years ago
- Official Code Repository for [AutoScale📈: Scale-Aware Data Mixing for Pre-Training LLMs] Published as a conference paper at **COLM 2025*…☆14Aug 8, 2025Updated last year
- ☆15Oct 4, 2024Updated last year
- ☆20Jun 27, 2026Updated last month
- An implementation of online data mixing for the Pile dataset, based on the GPT-NeoX library.☆14Jan 9, 2024Updated 2 years ago
- Fairer Preferences Elicit Improved Human-Aligned Large Language Model Judgments (Zhou et al., EMNLP 2024)☆14Oct 3, 2024Updated last year
- Code for the ACL 2023 paper: "Rethinking the Role of Scale for In-Context Learning: An Interpretability-based Case Study at 66 Billion Sc…☆36Sep 16, 2023Updated 2 years ago
- [SIGIR24] Pre-training with Bag-of-Word Prediction for Dense Passage Retrieval☆18Feb 29, 2024Updated 2 years ago
- Implementation of Gradient Information Optimization (GIO) for effective and scalable training data selection☆14Jun 22, 2023Updated 3 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Data set for LREC 2020 paper "I Feel Offended, Don't Be Abusive!"☆19Sep 23, 2023Updated 2 years ago
- SCT: An Efficient Self-Supervised Cross-View Training For Sentence Embedding (TACL)☆16Jul 27, 2024Updated 2 years ago
- Offical code of the paper Large Language Models Are Implicitly Topic Models: Explaining and Finding Good Demonstrations for In-Context Le…☆76Mar 20, 2024Updated 2 years ago
- A simple Python implementation of forward-forward NN training by G. Hinton from NeurIPS 2022☆21Dec 2, 2022Updated 3 years ago
- A library implementing the kernels for and experiments using extrinsic gauge equivariant vector field Gaussian Processes☆26Oct 28, 2021Updated 4 years ago
- Automatic prompt optimization framework for multi-step agent tasks.☆37Nov 12, 2024Updated last year
- This repo contains all the codes for SEScore implementation☆15Mar 3, 2025Updated last year
- Code base for the EMNLP 2021 Findings paper: Cartography Active Learning☆14Jun 3, 2025Updated last year
- Auxiliary tasks for task-oriented dialogue systems. Published in ICNLSP'22 and indexed in the ACL Anthology.☆17Feb 27, 2023Updated 3 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- The offcial repository for 'CharacterBERT and Self-Teaching for Improving the Robustness of Dense Retrievers on Queries with Typos', SIGI…☆16May 4, 2022Updated 4 years ago
- Thank you BART! Rewarding Pre-Trained Models Improves Formality Style Transfer (ACL 2021)☆30Oct 25, 2022Updated 3 years ago
- Extracting Cultural Commonsense Knowledge at Scale (WWW 2023)☆11Feb 15, 2024Updated 2 years ago
- Knowledge Graph-augmented NMT☆11Sep 20, 2021Updated 4 years ago
- Python wrapper for making MySQL queries easier☆10Mar 13, 2023Updated 3 years ago
- Joint Selection for Large-Scale Pre-Training Data via Policy Gradient-based Mask Learning☆21Jan 4, 2026Updated 7 months ago
- Code for our EMNLP-2023 paper: "Active Instruction Tuning: Improving Cross-Task Generalization by Training on Prompt Sensitive Tasks"☆26Nov 16, 2023Updated 2 years ago
- Targeted Data Generation with Large Language Models☆19Jun 25, 2024Updated 2 years ago
- Implementation of Implicit Graphon Neural Representation☆13Sep 1, 2023Updated 2 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Meta in-context learning for protein fitness prediction☆21Feb 7, 2025Updated last year
- Governance of the Commons Simulation (GovSim)☆81Jan 19, 2025Updated last year
- Mining Discriminative Components with Random Forests☆17Jun 13, 2016Updated 10 years ago
- [ICML 2024] Selecting High-Quality Data for Training Language Models☆204Dec 8, 2025Updated 8 months ago
- ☆17Mar 3, 2025Updated last year
- data collator for UL2 and U-PaLM☆29Aug 20, 2023Updated 3 years ago
- Simple and scalable tools for data-driven pretraining data selection.☆30Jun 9, 2025Updated last year