Training and evaluation code for the paper "Headless Language Models: Learning without Predicting with Contrastive Weight Tying" (https://arxiv.org/abs/2309.08351)
☆29Apr 17, 2024Updated 2 years ago
Alternatives and similar repositories for headless-lm
Users that are interested in headless-lm are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A Scheduler for Batched LLM Inference☆19Oct 5, 2025Updated 10 months ago
- Python source code for EMNLP 2021 Findings paper: "Subword Mapping and Anchoring Across Languages".☆13Sep 17, 2021Updated 4 years ago
- 🚀🤗 A collection of templates for Hugging Face Spaces☆35Oct 9, 2023Updated 2 years ago
- ☆16Jun 14, 2024Updated 2 years ago
- [ICASSP 2025] AnCoGen: Analysis, Control and Generation of Speech with a Masked Autoencoder☆14Mar 11, 2025Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆16May 14, 2024Updated 2 years ago
- MUX-PLMs: Pretraining LMs with Data Multiplexing☆15Jan 29, 2023Updated 3 years ago
- mPLM-Sim: Better Cross-Lingual Similarity and Transfer in Multilingual Pretrained Language Models☆11Jan 19, 2024Updated 2 years ago
- Evaluate your models with A/B test experiments☆14Jan 5, 2023Updated 3 years ago
- GPT2 Byte Pair Encoding implementation in Golang☆25Jul 9, 2025Updated last year
- A Multilingual Keyboard Layout-Based Typo Generator☆17Nov 23, 2025Updated 8 months ago
- Named Entity (NER) annotations of the Hebrew Treebank (Haaretz newspaper) corpus, including: morpheme and token level NER labels, nested …☆11Dec 27, 2021Updated 4 years ago
- ☆10Oct 15, 2019Updated 6 years ago
- Library for fast text representation and classification. Fix compatibility with numpy 2☆16Nov 21, 2024Updated last year
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Parity-Aware Byte-Pair Encoding: Improving Cross-lingual Fairness in Tokenization [ACL 2026]☆20Apr 18, 2026Updated 3 months ago
- [WAVC 2024] Official implementation of the paper: Semantic Generative Augmentations for Few-shot Counting☆13May 1, 2024Updated 2 years ago
- T-Projection is a method to perform high-quality Annotation Projection of Sequence Labeling datasets.☆13Nov 21, 2023Updated 2 years ago
- A Python database interface for eXist-db☆15Aug 1, 2026Updated last week
- ☆11Oct 3, 2021Updated 4 years ago
- [ICML 2026] Improving GPT via a simple normalization strategy☆15May 22, 2026Updated 2 months ago
- Linear Attention for Efficient Bidirectional Sequence Modeling☆18May 13, 2025Updated last year
- ☆12Mar 15, 2024Updated 2 years ago
- Goldfish: Monolingual language models for 350 languages.☆28Mar 4, 2026Updated 5 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Resources related to EMNLP 2021 paper "FAME: Feature-Based Adversarial Meta-Embeddings for Robust Input Representations"☆13Dec 14, 2021Updated 4 years ago
- DPO, but faster 🚀☆52Dec 6, 2024Updated last year
- A framework for adversarial attacks against token classification models☆33Nov 6, 2021Updated 4 years ago
- An opinionated NLP research template☆10Aug 29, 2024Updated last year
- ☆20Apr 26, 2026Updated 3 months ago
- (ACL-IJCNLP 2021) Convolutions and Self-Attention: Re-interpreting Relative Positions in Pre-trained Language Models.☆21Jul 13, 2022Updated 4 years ago
- [NeurIPS 2024] 🕸 GlotCC Dataset and Pipline☆21Apr 6, 2025Updated last year
- Exploring Few-Shot Adaptation of Language Models with Tables☆25Aug 22, 2022Updated 3 years ago
- torch_remat fine-grained activation checkpointing API☆15Updated this week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Repo for training MLMs, CLMs, or T5-type models on the OLM pretraining data, but it should work with any hugging face text dataset.☆98Feb 9, 2023Updated 3 years ago
- A neural network that jointly part-of-speech tags and lemmatizes sentences, boosting accuracy for morphologically-rich languages (Czech, …☆34Apr 5, 2019Updated 7 years ago
- ☆19Jan 13, 2025Updated last year
- Lowering PyTorch's Memory Consumption for Selective Differentiation☆12Aug 29, 2024Updated last year
- Air and bone conduction speech☆27Nov 26, 2022Updated 3 years ago
- ☆19Feb 18, 2026Updated 5 months ago
- T-scan: an analysis tool for dutch texts to assess the complexity of the text, based on original work by Rogier Kraf☆19May 28, 2025Updated last year