Multi-Domain Expert Learning
☆66Jan 23, 2024Updated 2 years ago
Alternatives and similar repositories for MDEL
Users that are interested in MDEL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Exploring finetuning public checkpoints on filter 8K sequences on Pile☆116Mar 22, 2023Updated 3 years ago
- ☆20Jul 12, 2023Updated 3 years ago
- A library for squeakily cleaning and filtering language datasets.☆50Jul 10, 2023Updated 3 years ago
- A repository of projects and datasets under active development by Alignment Lab AI☆22Dec 22, 2023Updated 2 years ago
- Image Diffusion block merging technique applied to transformers based Language Models.☆55May 8, 2023Updated 3 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Token-level adaptation of LoRA matrices for downstream task generalization.☆15Apr 14, 2024Updated 2 years ago
- Code repository for the c-BTM paper☆109Sep 26, 2023Updated 2 years ago
- An OpenAI API compatible LLM inference server based on ExLlamaV2.☆24Feb 9, 2024Updated 2 years ago
- A curated list of Natural Language Generation papers, tutorials, and blogs.☆12Dec 13, 2018Updated 7 years ago
- ☆21Oct 6, 2023Updated 2 years ago
- ☆14Jul 13, 2025Updated last year
- See https://github.com/cuda-mode/triton-index/ instead!☆11May 8, 2024Updated 2 years ago
- Ongoing research training transformer language models at scale, including: BERT & GPT-2☆20Oct 23, 2023Updated 2 years ago
- Utilities for Training Very Large Models☆58Sep 25, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- This repository contains code for cleaning your training data of benchmark data to help combat data snooping.☆28Apr 21, 2023Updated 3 years ago
- ☆30Sep 28, 2023Updated 2 years ago
- ☆33Jul 31, 2024Updated 2 years ago
- Awesome-RL-Reasoning☆17Jul 23, 2026Updated 2 weeks ago
- Solution for the Foursquare - Location Matching competition☆14Jul 8, 2022Updated 4 years ago
- Code and data to explore neural scaling laws of xLSTM and Transformer models.☆23Apr 8, 2026Updated 4 months ago
- Lecture of Practical Parallel Computing☆35Jun 1, 2026Updated 2 months ago
- ☆28Aug 30, 2023Updated 2 years ago
- Checkpointable dataset utilities for foundation model training☆32Jan 29, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆63Sep 23, 2024Updated last year
- Open sourced backend for Martian's LLM Inference Provider Leaderboard☆21Aug 13, 2024Updated last year
- ☆10Sep 7, 2020Updated 5 years ago
- ☆79Apr 29, 2024Updated 2 years ago
- Full finetuning of large language models without large memory requirements☆92Sep 22, 2025Updated 10 months ago
- Spherical Merge Pytorch/HF format Language Models with minimal feature loss.☆153Sep 10, 2023Updated 2 years ago
- Swarming algorithms like PSO, Ant Colony, Sakana, and more in PyTorch 😊☆151Updated this week
- An experiment to see if chatgpt can improve the output of the stanford alpaca dataset☆12Mar 29, 2023Updated 3 years ago
- Codes written for some competitions☆13Dec 8, 2016Updated 9 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- direct preference optimization with only 1 model copy :)☆14Oct 2, 2023Updated 2 years ago
- ☆13Jul 30, 2026Updated last week
- Patch for MPT-7B which allows using and training a LoRA☆57May 20, 2023Updated 3 years ago
- Using NLP techniques to summarize prompts for program synthesis☆17Sep 26, 2023Updated 2 years ago
- Retro styled terminal shell☆26May 8, 2024Updated 2 years ago
- Official implementation of Paper "System-Aware 4-Bit KV-Cache Quantization for Real-World LLM Serving"☆30Apr 17, 2026Updated 3 months ago
- Some simple scripts that I use day-to-day when working with LLMs and Huggingface Hub☆161Sep 26, 2023Updated 2 years ago