Tool for converting LLMs from uni-directional to bi-directional by removing causal mask for tasks like classification and sentence embeddings. Compatible with ๐ค transformers.
โ66Dec 12, 2024Updated last year
Alternatives and similar repositories for BiLLM
Users that are interested in BiLLM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A Simple but Powerful SOTA NER Model | Official Code For Label Supervised LLaMA Finetuningโ152Mar 17, 2024Updated 2 years ago
- โ20Apr 8, 2025Updated last year
- [SIGIR 2025] The official repo for "Scaling Sparse and Dense Retrieval in Decoder-Only LLMs"โ22Mar 31, 2025Updated last year
- code for piccolo embedding model from SenseTimeโ145May 21, 2024Updated 2 years ago
- Leveraging passage embeddings for efficient listwise reranking with large language models.โ51Dec 7, 2024Updated last year
- Managed Kubernetes at scale on DigitalOcean โข AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Improving Text Embedding of Language Models Using Contrastive Fine-tuningโ64Aug 2, 2024Updated 2 years ago
- Train and Infer Powerful Sentence Embeddings with AnglE | ๐ฅ SOTA on STS and MTEB Leaderboardโ574Mar 22, 2026Updated 5 months ago
- ROUTE: Robust Multitask Tuning and Collaboration for Text-to-SQL (ICLR 2025 Pytorch Code)โ16May 15, 2025Updated last year
- Code for EMNLP 2023 paper: DALE: Generative Data Augmentation for Low-Resource Legal NLPโ11Oct 27, 2023Updated 2 years ago
- Official repository for paper "ReasonIR Training Retrievers for Reasoning Tasks".โ230Jul 2, 2026Updated 2 months ago
- Cascade bert+word vec and one layer FLAT, trained by adversarial FGM and Stochastic Weight Averagingโ23Nov 4, 2021Updated 4 years ago
- This repository helps you evaluate your models on the FreshStack benchmark!โ34Dec 9, 2025Updated 9 months ago
- A package for fine tuning of pretrained NLP transformers using Semi Supervised Learningโ14Oct 27, 2021Updated 4 years ago
- Easy modernBERT fine-tuning and multi-task learningโ66Mar 13, 2026Updated 6 months ago
- Managed Kubernetes at scale on DigitalOcean โข AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Cocktail: A Comprehensive Information Retrieval Benchmark with LLM-Generated Documents Integrationโ15Jun 4, 2024Updated 2 years ago
- Efficient Pre-training of Masked Language Model via Concept-based Curriculum Maskingโ13Feb 5, 2023Updated 3 years ago
- Source code for SummaReranker (ACL 2022)โ24Jan 7, 2024Updated 2 years ago
- [Findings of ACL'2023] Improving Contrastive Learning of Sentence Embeddings from AI Feedbackโ40Aug 14, 2023Updated 3 years ago
- A PyTorch implementation of Proxy Anchor Loss based on CVPR 2020 paper "Proxy Anchor Loss for Deep Metric Learning"โ11Jan 16, 2021Updated 5 years ago
- โ26Oct 20, 2022Updated 3 years ago
- Starbucks: Improved Training for 2D Matryoshka Embeddingsโ25Jun 30, 2025Updated last year
- โ13Feb 17, 2025Updated last year
- Chance-corrected Agreement Coefficientsโ17Aug 22, 2026Updated 3 weeks ago
- Wordpress hosting with auto-scaling - Free Trial Offer โข AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [ACL 2023] Few-shot Reranking for Multi-hop QA via Language Model Promptingโ27Oct 19, 2025Updated 11 months ago
- Tevatron - Unified Document Retrieval Toolkit across Scale, Language, and Modality. Demo in SIGIR 2023, SIGIR 2025.โ750Jul 18, 2026Updated 2 months ago
- Streamlit UI to remove duplicate or near duplicate imagesโ12Mar 25, 2023Updated 3 years ago
- Code for KaLM-Embedding modelsโ119Jun 30, 2025Updated last year
- Official Code For TDEER: An Efficient Translating Decoding Schema for Joint Extraction of Entities and Relations (EMNLP 2021)โ41Jul 27, 2024Updated 2 years ago
- The training codes of Jasper-Token-Compression-600Mโ22Nov 19, 2025Updated 10 months ago
- CCKS 2020: ้ขๅไธญๆ็ญๆๆฌ็ๅฎไฝ้พๆไปปๅกโ43Mar 27, 2021Updated 5 years ago
- A collection of scripts for retrieving, storing, and querying SureChEMBL data.โ42Jul 11, 2024Updated 2 years ago
- Unified Learned Sparse Retrieval Frameworkโ67May 13, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient โข AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Code for embedding and retrieval research.โ16Oct 24, 2023Updated 2 years ago
- Code for ACL paper "Zero-Shot Text Classification via Self-Supervised Tuning"โ29Sep 25, 2023Updated 2 years ago
- LLM for NERโ83Jul 29, 2024Updated 2 years ago
- GPU-accelerated algorithm for subsampling datasets while preserving diversityโ27Jan 12, 2024Updated 2 years ago
- โ13Mar 22, 2023Updated 3 years ago
- โ53Sep 11, 2024Updated 2 years ago
- [ACL 2025] AIR-Bench: Automated Heterogeneous Information Retrieval Benchmarkโ167Mar 29, 2026Updated 5 months ago