llama fine-tuning with lora
☆140May 8, 2024Updated 2 years ago
Alternatives and similar repositories for llama-lora-fine-tuning
Users that are interested in llama-lora-fine-tuning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- llama2 finetuning with deepspeed and lora☆176Jul 28, 2023Updated 2 years ago
- A full pipeline to finetune Vicuna LLM with LoRA and RLHF on consumer hardware. Implementation of RLHF (Reinforcement Learning with Human…☆220May 20, 2024Updated 2 years ago
- Original PyTorch Implementation for the EMNLP 2023 Paper "Beyond Detection: A Defend-and-Summarize Strategy for Robust and Interpretable …☆16Dec 14, 2023Updated 2 years ago
- ☆19Dec 12, 2023Updated 2 years ago
- MASSW is a comprehensive text dataset on Multi-Aspect Summarization of Scientific Workflows. MASSW includes more than 152,000 peer-review…☆22May 16, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- 基于prompt的中文文本分类。☆55May 6, 2023Updated 3 years ago
- Tuning LLMs with no tears💦; Sample Design Engineering (SDE) for more efficient downstream-tuning.☆1,015Apr 27, 2024Updated 2 years ago
- ☆71Jul 11, 2024Updated 2 years ago
- Finetune LLaMA-7B with Chinese instruction datasets☆136May 8, 2023Updated 3 years ago
- ☆10Jul 8, 2021Updated 5 years ago
- [AAAI 2024] GMMFormer: Gaussian-Mixture-Model Based Transformer for Efficient Partially Relevant Video Retrieval☆21May 10, 2024Updated 2 years ago
- AI-WordCards is an innovative project that leverages the power of GPT, StableDiffusion, and DALL-E3 to create educational and engaging wo…☆11May 16, 2024Updated 2 years ago
- GNNs-Pytorch-Geometric-Examples☆10Jun 14, 2020Updated 6 years ago
- VSCode code generation Extension based on SalesForce CodeGen☆32Jan 27, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- APIBench is a benchmark for evaluating the performance of API recommendation approaches released in the paper "Revisiting, Benchmarking a…☆66Apr 3, 2023Updated 3 years ago
- [ICLR 2026🔥] SHIELD: Suppressing Hallucinations In LVLM Encoders via Bias and Vulnerability Defense☆17Mar 24, 2026Updated 3 months ago
- Code for the paper "Rethinking Stealthiness of Backdoor Attack against NLP Models" (ACL-IJCNLP 2021)☆24Dec 9, 2021Updated 4 years ago
- Agentic translation using reflection workflow, refactored and sugared.☆11Sep 25, 2024Updated last year
- Easy and Efficient Finetuning LLMs. (Supported LLama, LLama2, LLama3, Qwen, Baichuan, GLM , Falcon) 大模型高效量化训练+部署.☆620Jan 24, 2025Updated last year
- Methods for removing learned data from neural nets and evaluation of those methods☆38Nov 26, 2020Updated 5 years ago
- Instruct-tune LLaMA on consumer hardware☆18,909Jul 29, 2024Updated last year
- Llama2 chinese finetuning☆38Aug 2, 2023Updated 2 years ago
- A virtual caregiver system that extracts the expression of mental and physical health states through dialogue-based human-computer intera…☆15Jan 29, 2023Updated 3 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆15May 28, 2024Updated 2 years ago
- Code used for analysis and visualiation of ocean model data during my postdoc☆12Mar 1, 2023Updated 3 years ago
- ☆15Feb 26, 2025Updated last year
- onebot v11 adapter in plugin☆10Mar 5, 2023Updated 3 years ago
- Code for the paper "Multi-Field Adaptive Retrieval," a research project on a semi-structured document retrieval☆18Feb 13, 2026Updated 5 months ago
- ☆14Jun 3, 2023Updated 3 years ago
- Recommender systems with large language models (Paper list)☆65Nov 20, 2023Updated 2 years ago
- tiny HTTP request lib base on fetch api☆14Jun 26, 2020Updated 6 years ago
- ☆15May 16, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- LLaMA-TRL: Fine-tuning LLaMA with PPO and LoRA☆240Aug 17, 2025Updated 11 months ago
- Code & data for EMNLP 2020 paper "MOCHA: A Dataset for Training and Evaluating Reading Comprehension Metrics".☆16May 3, 2022Updated 4 years ago
- Text Style Transfer: A Review☆13Jun 1, 2019Updated 7 years ago
- [ICML 2024] Matrix Variational Masked Autoencoder (M-MAE) for ICML paper "Information Flow in Self-Supervised Learning" (https://arxiv.or…☆15Sep 17, 2024Updated last year
- ☆17May 31, 2023Updated 3 years ago
- A simple script for extracting plain text from arxiv dataset: https://www.kaggle.com/Cornell-University/arxiv☆15Dec 7, 2020Updated 5 years ago
- ☆18Jun 25, 2026Updated 3 weeks ago