chenllliang/Gradient-Vaccine

Readme badge preview -

If you own this repo, copy the snippet below and add it to your README.md

[![RelatedRepos](https://img.shields.io/badge/related-repos-yellow)](https://relatedrepos.com/gh/chenllliang/Gradient-Vaccine)

chenllliang / Gradient-Vaccine

(Unofficial) Implementation of ICLR 2021 paper "Gradient Vaccine: Investigating and Improving Multi-task Optimization in Massively Multilingual Models"

☆14

Alternatives and similar repositories for Gradient-Vaccine

Users that are interested in Gradient-Vaccine are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.

Sorting:

M3-IT / YING-VLM
View on GitHub
Vision Large Language Models trained on M3IT instruction tuning dataset
☆17Aug 16, 2023Updated 2 years ago
dqxiu / KAssess
View on GitHub
☆14Oct 28, 2023Updated 2 years ago
Yifan-Song793 / InfoCL
View on GitHub
Findings of EMNLP 2023: InfoCL: Alleviating Catastrophic Forgetting in Continual Text Classification from An Information Theoretic Perspe…
☆14Aug 13, 2024Updated last year
pkunlp-icler / MLS
View on GitHub
Source code of our paper "Focus on the Target’s Vocabulary: Masked Label Smoothing for Machine Translation" @ ACL 2022
☆13Apr 13, 2022Updated 4 years ago
TobiasLee / VEC
View on GitHub
Visual and Embodied Concepts evaluation benchmark
☆21Oct 10, 2023Updated 2 years ago
Virtual machines for every use case on DigitalOcean • Ad
Get dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
Wangpeiyi9979 / ACA
View on GitHub
EMNLP2022: Learning Robust Representations for Continual Relation Extraction via Adversarial Class Augmentation
☆15Oct 19, 2022Updated 3 years ago
Zce1112zslx / IKE
View on GitHub
☆41Nov 30, 2023Updated 2 years ago
dqxiu / CaliNet
View on GitHub
☆32Oct 17, 2022Updated 3 years ago
lancopku / clip-openness
View on GitHub
[ACL 2023] Delving into the Openness of CLIP
☆24Jan 11, 2023Updated 3 years ago
yilinyang7 / fairseq_multi_fix
View on GitHub
Code and Data release for "Improving Multilingual Translation by Representation and Gradient Regularization" (Yang et al. EMNLP 2021), an…
☆13Aug 12, 2024Updated last year
chenllliang / MLS
View on GitHub
Source code of our paper "Focus on the Target’s Vocabulary: Masked Label Smoothing for Machine Translation" @ACL-2022
☆18May 19, 2022Updated 4 years ago
RunxinXu / Make-Information-Extraction-Great-Again
View on GitHub
An (incomplete) overview of information extraction
☆43Apr 28, 2022Updated 4 years ago
Zce1112zslx / ChID_baseline
View on GitHub
计算语言学22-23学年秋季学期课程大作业baseline实现
☆38Dec 8, 2022Updated 3 years ago
Wangpeiyi9979 / ESD
View on GitHub
Code for NAACL2022 Long Paper "An Enhanced Span-based Decomposition Method for Few-Shot Sequence Labeling"
☆27Nov 9, 2022Updated 3 years ago
Deploy on Railway without the complexity - Free Credits Offer • Ad
Connect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
S1s-Z / SCL-RAI
View on GitHub
[COLING'22] Code for "SCL-RAI: Span-based Contrastive Learning with Retrieval Augmented Inference for Unlabeled Entity Problem in NER"
☆45Sep 10, 2024Updated last year
pkunlp-icler / IKE
View on GitHub
☆25Feb 27, 2023Updated 3 years ago
pkunlp-icler / Two-Stage-CAMRP
View on GitHub
Source code for paper "A Two-Stage Method for Chinese AMR Parsing" @ CAMRP-2022 & CCL-2022
☆24Aug 29, 2024Updated last year
WeiminXiong / RationaleCL
View on GitHub
Rationale-enhanced language models are better continual relation learners (EMNLP 2023 Main Conference)
☆12Oct 11, 2023Updated 2 years ago
S1s-Z / SANTA
View on GitHub
[ACL'23] Code for "SANTA: Separate Strategies for Inaccurate and Incomplete Annotation Noise in Distantly-Supervised Named Entity Recogni…
☆39Apr 28, 2025Updated last year
lancopku / DCKD
View on GitHub
Code and data for Distributional Correlation–Aware Knowledge Distillation for Stock Trading Volume Prediction (ECML-PKDD 22)
☆16Sep 6, 2022Updated 3 years ago
kkk-an / UltraIF
View on GitHub
Code of EMNLP 2025 paper 'UltraIF: Advancing Instruction Following from the Wild'.
☆21Apr 3, 2025Updated last year
zkcpku / HiT-hierarchy-transformer
View on GitHub
code for "Implant Global and Local Hierarchy Information to Sequence based Code Representation Models"
☆12Dec 13, 2024Updated last year
KbsdJames / MATH-Minos
View on GitHub
The implementation of paper "LLM Critics Help Catch Bugs in Mathematics: Towards a Better Mathematical Verifier with Natural Language Fee…
☆38Jul 25, 2024Updated 2 years ago
Managed Database hosting by DigitalOcean • Ad
PostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
lancopku / MUKI
View on GitHub
[Findings of EMNLP22] From Mimicking to Integrating: Knowledge Integration for Pre-Trained Language Models
☆19Mar 16, 2023Updated 3 years ago
F2-Song / ICDPO
View on GitHub
The official implementation of "ICDPO: Effectively Borrowing Alignment Capability of Others via In-context Direct Preference Optimization…
☆16Feb 15, 2024Updated 2 years ago
tongmeihan1995 / DocEE
View on GitHub
DocEE: A Large-Scale and Fine-grained Benchmark for Document-level Event Extraction
☆42Apr 19, 2023Updated 3 years ago
Nesvilab / MSBooster
View on GitHub
MSBooster allows users to add deep learning-based features to .pin files before Percolator PSM rescoring
☆22Jul 18, 2026Updated last week
WeiChengTseng / Pytorch-PCGrad
View on GitHub
Pytorch reimplementation for "Gradient Surgery for Multi-Task Learning"
☆403Jun 22, 2021Updated 5 years ago
lscpku / VITATECS
View on GitHub
☆18Jul 10, 2024Updated 2 years ago
cisnlp / mPLM-Sim
View on GitHub
mPLM-Sim: Better Cross-Lingual Similarity and Transfer in Multilingual Pretrained Language Models
☆11Jan 19, 2024Updated 2 years ago
alon-albalak / online-data-mixing
View on GitHub
An implementation of online data mixing for the Pile dataset, based on the GPT-NeoX library.
☆14Jan 9, 2024Updated 2 years ago
zxytim / arithmetic-encoding-compression
View on GitHub
☆11Apr 3, 2023Updated 3 years ago
AI Agents on DigitalOcean Gradient AI Platform • Ad
Build production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
zhuzilin / vllm-group
View on GitHub
☆12Nov 5, 2024Updated last year
gagneurlab / spectralis
View on GitHub
☆17Oct 1, 2024Updated last year
chenllliang / CTDNN
View on GitHub
MMM 2021: Crossed-Time Delay Neural Network for Speaker Recognition
☆11Dec 4, 2021Updated 4 years ago
llyx97 / sparse-and-robust-PLM
View on GitHub
[NeurIPS 2022] "A Win-win Deal: Towards Sparse and Robust Pre-trained Language Models", Yuanxin Liu, Fandong Meng, Zheng Lin, Jiangnan Li…
☆21Jan 9, 2024Updated 2 years ago
RunxinXu / ChildTuning
View on GitHub
Source code for our EMNLP'21 paper 《Raise a Child in Large Language Model: Towards Effective and Generalizable Fine-tuning》
☆61Nov 6, 2021Updated 4 years ago
IBM / docAMR
View on GitHub
code for document level AMR representation and evaluation
☆46Sep 17, 2025Updated 10 months ago
zhang-xuan1314 / MTL-BERT
View on GitHub
☆23Oct 6, 2022Updated 3 years ago