Token-level adaptation of LoRA matrices for downstream task generalization.
☆15Apr 14, 2024Updated 2 years ago
Alternatives and similar repositories for LoRA-TLE
Users that are interested in LoRA-TLE are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Mixture of Expert (MoE) techniques for enhancing LLM performance through expert-driven prompt mapping and adapter combinations.☆11Feb 11, 2024Updated 2 years ago
- Comprehensive analysis of difference in performance of QLora, Lora, and Full Finetunes.☆83Sep 10, 2023Updated 2 years ago
- Script for processing OpenAI's PRM800K process supervision dataset into an Alpaca-style instruction-response format☆27Jul 12, 2023Updated 3 years ago
- [SIGIR'24] The official implementation code of MOELoRA.☆193Jul 22, 2024Updated 2 years ago
- Repository for SoMeLVLM: A Large Vision Language Model for Social Media Processing☆14Oct 9, 2025Updated 10 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆13Jun 18, 2024Updated 2 years ago
- LREC-COLING 2024: DiffusionABSA: Let’s Rectify Step by Step: Improving Aspect-based Sentiment Analysis with Diffusion Models☆24Oct 6, 2024Updated last year
- A Data Source for Reasoning Embodied Agents☆20Sep 18, 2023Updated 2 years ago
- Image Diffusion block merging technique applied to transformers based Language Models.☆55May 8, 2023Updated 3 years ago
- A video scene detection algorithm is designed to detect a variety of different scenes within a video. There is a very simple definition f…☆10Jan 4, 2022Updated 4 years ago
- Zeus LLM Trainer is a rewrite of Stanford Alpaca aiming to be the trainer for all Large Language Models☆69Aug 27, 2023Updated 2 years ago
- fastNLP reimplementation of the paper "A Novel Cascade Binary Tagging Framework for Relational Triple Extraction"☆11Dec 11, 2020Updated 5 years ago
- ☆21Oct 6, 2023Updated 2 years ago
- QLoRA: Efficient Finetuning of Quantized LLMs☆11Jun 1, 2023Updated 3 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Sample app for a Python API using FastAPI and neomodel☆12Jul 1, 2024Updated 2 years ago
- a pipeline for using api calls to agnostically convert unstructured data into structured training data☆32Sep 22, 2024Updated last year
- Multi-Domain Expert Learning☆66Jan 23, 2024Updated 2 years ago
- Python code which creates a semantic search bot over any available corpus☆17May 22, 2023Updated 3 years ago
- Code repository for the c-BTM paper☆109Sep 26, 2023Updated 2 years ago
- Repository containing code for the paper "Learning to Learn to Disambiguate: Meta-Learning for Few-Shot Word Sense Disambiguation", publi…☆12Nov 12, 2020Updated 5 years ago
- A curated reading list of research in Adaptive Computation, Inference-Time Computation & Mixture of Experts (MoE).☆164Jan 1, 2025Updated last year
- ☆11Oct 3, 2021Updated 4 years ago
- T2NER: Transformers based Transfer Learning Framework for Named Entity Recognition (EACL 2021)☆11Sep 24, 2022Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆23Dec 22, 2023Updated 2 years ago
- Train transformer language models with reinforcement learning.☆20Dec 26, 2023Updated 2 years ago
- Code for LAMOL: LAnguage MOdeling for Lifelong Language Learning☆95Aug 28, 2020Updated 5 years ago
- 🚀 A curated list of awesome Desktop Extensions (DXT) and MCP servers for Claude Desktop. Discover, share, and contribute to the growing …☆21Jul 1, 2025Updated last year
- Source code for our "D-REPTILE" paper at EACL 2021☆13Jan 19, 2021Updated 5 years ago
- ☆415Nov 2, 2023Updated 2 years ago
- The official implementation of the EMNLP 2023 paper "Paraphrase Types for Generation and Detection"☆12Oct 20, 2024Updated last year
- Must-read Papers on Large Language Model (LLM) Continual Learning☆150Nov 14, 2023Updated 2 years ago
- Merge Transformers language models by use of gradient parameters.☆215Aug 8, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- An MCP server that interfaces with OpenAI, Google, and Anthropic's APIs to give Claude Code "coworkers" to help it on difficult problems.☆17Apr 3, 2025Updated last year
- This repo is our code and dataset for paper De-biasing Distantly Supervised Named Entity Recognition via Causal Intervention.☆13Sep 2, 2021Updated 4 years ago
- ☆27Jul 29, 2025Updated last year
- ☆16Oct 24, 2021Updated 4 years ago
- ☆79Apr 29, 2024Updated 2 years ago
- A benchmark to evaluate search-augmented LLMs☆17Aug 28, 2025Updated 11 months ago
- ☆17Jan 30, 2024Updated 2 years ago