Fine-tuning large language models with huggingface transformers and deepspeed
☆31Dec 11, 2023Updated 2 years ago
Alternatives and similar repositories for llmft
Users that are interested in llmft are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code for paper "Factual Confidence of LLMs: on Reliability and Robustness of Current Estimators"☆17Dec 4, 2024Updated last year
- Code for T-MARS data filtering☆35Aug 23, 2023Updated 3 years ago
- Code for "Preference Tuning For Toxicity Mitigation Generalizes Across Languages." Paper accepted at Findings of EMNLP 2024☆18Mar 25, 2025Updated last year
- [Kauf & Ivanova, ACL 2023] A Better Way to Do Masked Language Model Scoring☆13Dec 1, 2023Updated 2 years ago
- Efficient Scaling laws and collaborative pretraining.☆24Jul 19, 2026Updated 2 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆13Jul 8, 2023Updated 3 years ago
- [CVPR 2021] ColorRL: Reinforced Coloring for End-to-End Instance Segmentation In Computer Vision and Pattern Recognition Conference☆11Nov 19, 2025Updated 10 months ago
- Code for the examples presented in the talk "Training a Llama in your backyard: fine-tuning very large models on consumer hardware" given…☆15Oct 16, 2023Updated 2 years ago
- Synthetic pretraining data by rephrasing the web☆35Jun 5, 2026Updated 3 months ago
- Repository for reproducing `Model-Based Robust Deep Learning`☆17Jan 22, 2021Updated 5 years ago
- ☆30Jun 19, 2023Updated 3 years ago
- ☆19Jul 4, 2025Updated last year
- Generate a dataset to finetune a LLM to generate Cypher code from questions given in natural language (English).☆15May 24, 2024Updated 2 years ago
- Data Valuation on In-Context Examples (ACL23)☆24Jan 12, 2025Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆35Jan 29, 2023Updated 3 years ago
- Placeholder repository☆15Mar 16, 2022Updated 4 years ago
- Merging, linking and placing compounds by stitching bound compounds together like a reanimated corpse☆12Feb 22, 2024Updated 2 years ago
- Evaluating the Factuality of Large Language Models using Large-Scale Knowledge Graphs☆35Sep 3, 2024Updated 2 years ago
- Code for ECML-PKDD 2022 Paper --- CMG: A Class-Mixed Generation Approach to Out-of-Distribution Detection☆13Oct 12, 2022Updated 3 years ago
- A compact high-signal benchmark for evaluating frontier agents☆39Aug 3, 2026Updated last month
- Official code for Cross-Domain Policy Adaptation by Capturing Representation Mismatch (ICML 2024)☆15Aug 15, 2025Updated last year
- Open source code for EigenGame.☆35May 15, 2023Updated 3 years ago
- Code for the paper "Pretraining task diversity and the emergence of non-Bayesian in-context learning for regression"☆27Jun 28, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Research code for the paper "How Good is Your Tokenizer? On the Monolingual Performance of Multilingual Language Models"☆28Oct 3, 2021Updated 4 years ago
- ☆26May 30, 2023Updated 3 years ago
- ☆14May 29, 2024Updated 2 years ago
- ☆19Oct 2, 2023Updated 2 years ago
- Accelerating Transfer Learning with Robust Neural Nets☆11Oct 2, 2020Updated 5 years ago
- 元培学院地下室预约系统☆14Jan 24, 2022Updated 4 years ago
- Word embeddings from PPMI-weighted and dirichlet-smoothed co-occurrence matrices☆10Aug 3, 2020Updated 6 years ago
- Minimal RLHF implementation built on top of minGPT.☆32Jul 4, 2024Updated 2 years ago
- [EMNLP 2022] Adapting a Language Model While Preserving its General Knowledge☆21Feb 12, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Reinforcement Learning inside a 3D soccer simulation☆36Sep 15, 2024Updated 2 years ago
- Reproducible code for Augmentation paper☆17Jan 23, 2019Updated 7 years ago
- ☆18Nov 7, 2022Updated 3 years ago
- Repo to reproduce the First-Explore paper results☆39May 6, 2026Updated 4 months ago
- ACL24☆11Jun 7, 2024Updated 2 years ago
- Algebraic value editing in pretrained language models☆71Nov 1, 2023Updated 2 years ago
- Official PyTorch Implementation of "CoSSL: Co-Learning of Representation and Classifier for Imbalanced Semi-Supervised Learning" (CVPR 20…☆52Sep 21, 2022Updated 3 years ago