BLOOM+1: Adapting BLOOM model to support a new unseen language
☆74Mar 2, 2024Updated 2 years ago
Alternatives and similar repositories for multilingual-modeling
Users that are interested in multilingual-modeling are comparing it to the libraries listed below
Sorting:
- ☆13Aug 23, 2024Updated last year
- ☆21Dec 5, 2022Updated 3 years ago
- German Alpaca Dataset (Cleaned + Translated)☆26Apr 6, 2023Updated 2 years ago
- Official code for the paper Improving Language Plasticity via Pretraining with Active Forgetting, NeurIPS 2023☆21Mar 12, 2026Updated last week
- A library for parameter-efficient and composable transfer learning for NLP with sparse fine-tunings.☆75Aug 9, 2024Updated last year
- ☆21Feb 13, 2023Updated 3 years ago
- Codebase, data and models for hallucination of pruned models☆16Jan 11, 2025Updated last year
- Can LLMs generate code-mixed sentences through zero-shot prompting?☆11Apr 18, 2023Updated 2 years ago
- EMNLP 2021 - Frustratingly Simple Pretraining Alternatives to Masked Language Modeling☆34Nov 21, 2021Updated 4 years ago
- Jojajovai Guarani-Spanish Parallel Corpus☆19Jul 5, 2022Updated 3 years ago
- Code and dataset for the EMNLP 2021 Finding paper "Can NLI Models Verify QA Systems’ Predictions?"☆25Jul 21, 2023Updated 2 years ago
- Repository for sharing the data in the Tamasheq language, one of the target languages for the low-resource speech translation track at IW…☆18Nov 30, 2022Updated 3 years ago
- ☆11Aug 12, 2020Updated 5 years ago
- Glot500: Scaling Multilingual Corpora and Language Models to 500 Languages -- ACL 2023☆106Apr 20, 2024Updated last year
- ☆11Oct 3, 2021Updated 4 years ago
- ☆93Feb 13, 2024Updated 2 years ago
- [ACL 2023]: Training Trajectories of Language Models Across Scales https://arxiv.org/pdf/2212.09803.pdf☆25Nov 14, 2023Updated 2 years ago
- ☆18Nov 25, 2022Updated 3 years ago
- Code and Data release for "Improving Multilingual Translation by Representation and Gradient Regularization" (Yang et al. EMNLP 2021), an…☆13Aug 12, 2024Updated last year
- ☆44Sep 16, 2020Updated 5 years ago
- CCQA A New Web-Scale Question Answering Dataset for Model Pre-Training☆32Jul 20, 2022Updated 3 years ago
- Adversarial Test Dataset for Korean Multi-turn Response Selection☆34Dec 16, 2021Updated 4 years ago
- ☆16Aug 20, 2020Updated 5 years ago
- Evaluation results for Machine Translation within the BigScience project☆11May 15, 2023Updated 2 years ago
- State-of-the-art LLM-based translation models.☆582Apr 9, 2025Updated 11 months ago
- Language models scale reliably with over-training and on downstream tasks☆100Apr 2, 2024Updated last year
- Web UI for docsQA. Main branch: https://jina-docqa-ui.netlify.app/☆20Aug 29, 2022Updated 3 years ago
- ☆11Jun 23, 2022Updated 3 years ago
- Code for evaluating uncertainty estimation methods for Transformer-based architectures in natural language understanding tasks.☆44Aug 16, 2021Updated 4 years ago
- GPT-jax based on the official huggingface library☆13Jun 22, 2021Updated 4 years ago
- ☆16May 14, 2024Updated last year
- This is project for korean auto spacing☆12Aug 3, 2020Updated 5 years ago
- exBERT on Transformers🤗☆10Jun 14, 2021Updated 4 years ago
- A Multilingual Replicable Instruction-Following Model☆97Jun 11, 2023Updated 2 years ago
- ☆14Oct 6, 2025Updated 5 months ago
- Anh - LAION's multilingual assistant datasets and models☆27Apr 5, 2023Updated 2 years ago
- Do Multilingual Language Models Think Better in English?☆42Aug 3, 2023Updated 2 years ago
- Curriculum training☆22Jun 25, 2025Updated 8 months ago
- Train 🤗transformers with DeepSpeed: ZeRO-2, ZeRO-3☆23May 20, 2021Updated 4 years ago