Collection of papers using LLaMA as backbone model
☆50Apr 6, 2025Updated last year
Alternatives and similar repositories for LLaMA-Paper-List
Users that are interested in LLaMA-Paper-List are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆14Aug 16, 2023Updated 3 years ago
- The offical code of "Parameter-Efficient Learning for Text-to-Speech Accent Adaptation"☆12Aug 29, 2023Updated 3 years ago
- ☆19Jan 21, 2022Updated 4 years ago
- ☆20Sep 2, 2021Updated 5 years ago
- [PACT'24] GraNNDis. A fast and unified distributed graph neural network (GNN) training framework for both full-batch (full-graph) and min…☆10Aug 13, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Hands-On Reinforcement Learning with TensorFlow & TRFL☆14Jan 18, 2021Updated 5 years ago
- Information and Materials for the Deep Learning Course☆30Jun 16, 2022Updated 4 years ago
- Code for "Towards Revealing the Mystery behind Chain of Thought: a Theoretical Perspective"☆21Jul 16, 2023Updated 3 years ago
- [ICLR 2022] Training L_inf-dist-net with faster acceleration and better training strategies☆22Mar 16, 2022Updated 4 years ago
- The complete tips and notes from Duolingo Hebrew course in one file☆13Mar 1, 2019Updated 7 years ago
- [NeurIPS 2021] Fast Certified Robust Training with Short Warmup☆25Jun 7, 2025Updated last year
- ☆16May 1, 2026Updated 4 months ago
- ☆12Sep 23, 2023Updated 2 years ago
- COBS: COmprehensive Building Simulator☆16Jun 23, 2022Updated 4 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Self-supervised learning course at CS HSE☆12May 24, 2023Updated 3 years ago
- [HPCA'24] Smart-Infinity: Fast Large Language Model Training using Near-Storage Processing on a Real System☆52Jul 21, 2025Updated last year
- NCSU CSC-326 Course Page☆12Dec 5, 2018Updated 7 years ago
- Code for the ACL 2022 (Long paper): "New Intent Discovery with Pre-training and Contrastive Learning".☆14Jul 18, 2022Updated 4 years ago
- [NeurIPS'23] Binary Classification with Confidence Difference☆10May 13, 2024Updated 2 years ago
- Using multiple sensor modalities to improve exploration for robotic manipulation tasks with sparse rewards☆10Sep 17, 2019Updated 6 years ago
- A simple and efficient llama3 local service deployment solution that supports real-time streaming response and is optimized for common Ch…☆13Jul 31, 2024Updated 2 years ago
- ☆10Sep 27, 2022Updated 3 years ago
- Code and data for the paper: On the Reliability of Psychological Scales on Large Language Models☆31Dec 15, 2025Updated 8 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ODYSSEUS/EduCOSMOS☆10Aug 12, 2020Updated 6 years ago
- ☆10Jun 21, 2021Updated 5 years ago
- [KDD'22] Partial Label Learning with Discrimination Augmentation☆10May 21, 2024Updated 2 years ago
- Learning from Indirect Observations☆11Jul 16, 2021Updated 5 years ago
- The dataset repo of "CLCIFAR: CIFAR-Derived Benchmark Datasets with Human Annotated Complementary Labels" paper☆17May 11, 2026Updated 3 months ago
- Official repository of paper [FALQON: Accelerating LoRA Fine-tuning with Low-Bit Floating-Point Arithmetic, NeurIPS 2025]☆21Dec 2, 2025Updated 9 months ago
- [ICLR 2025] "Noisy Test-Time Adaptation in Vision-Language Models"☆14Feb 22, 2025Updated last year
- Assignments for CS294-112.☆17Jul 13, 2018Updated 8 years ago
- Official code for ICML 2024 paper on Persona In-Context Learning (PICLe)☆28Jun 27, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ICLR 2024] This is the official implementation for the paper: "Beyond imitation: Leveraging fine-grained quality signals for alignment"☆10May 5, 2024Updated 2 years ago
- Serial Contrastive Knowledge Distillation for Continual Few-shot Relation Extraction, Findings of ACL 2023☆14May 12, 2023Updated 3 years ago
- Репозиторий курса "Практические аспекты обучения больших языковых моделей", ВМК МГУ, осень, 2024☆20Dec 24, 2024Updated last year
- [CVPR 2025] An Implementation of the paper "Pre-Instruction Data Selection for Visual Instruction Tuning"☆17Jun 9, 2025Updated last year
- It's All In the Teacher: Zero-Shot Quantization Brought Closer to the Teacher [CVPR 2022 Oral]☆29Sep 15, 2022Updated 3 years ago
- Smoothing video traffic to make it a friendlier internet neighbor☆14Apr 23, 2024Updated 2 years ago
- [ICML 2024] Code release for "On the Emergence of Cross-Task Linearity in Pretraining-Finetuning Paradigm"☆11Feb 20, 2025Updated last year