Collection of papers using LLaMA as backbone model
☆50Apr 6, 2025Updated last year
Alternatives and similar repositories for LLaMA-Paper-List
Users that are interested in LLaMA-Paper-List are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Chang Gung University Computer Science / Artificial Intelligence learning material☆27Sep 5, 2024Updated last year
- Knowledge-Infused Topic Model (KITM) for empathetic dialogue generation, integrating topic modeling with commonsense reasoning from COMET…☆22Nov 30, 2025Updated 8 months ago
- The offical code of "Parameter-Efficient Learning for Text-to-Speech Accent Adaptation"☆12Aug 29, 2023Updated 2 years ago
- Code for SRMRL☆19Sep 5, 2021Updated 4 years ago
- ☆20Sep 2, 2021Updated 4 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Code for Submission titled "Guidance Learning for Multi-Domain DIalogue Management"☆23Feb 2, 2022Updated 4 years ago
- finetune llama2 with traditional chinese dataset☆39Aug 8, 2023Updated 3 years ago
- Distilling Task-Specific Knowledge from BERT into Simple Neural Networks.☆15Aug 28, 2020Updated 5 years ago
- Refined dataset for Stanford Sentiment Treebank used in Yoon Kim (2014).☆12Apr 1, 2018Updated 8 years ago
- [PACT'24] GraNNDis. A fast and unified distributed graph neural network (GNN) training framework for both full-batch (full-graph) and min…☆10Aug 13, 2024Updated 2 years ago
- ☆15Sep 10, 2019Updated 6 years ago
- 交大電信所機器學習作業(簡仁宗老師)☆13Jan 20, 2017Updated 9 years ago
- [ICLR 2022] Training L_inf-dist-net with faster acceleration and better training strategies☆22Mar 16, 2022Updated 4 years ago
- ☆16May 1, 2026Updated 3 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [USENIX Security 2025] Topic-FlipRAG: Topic-Orientated Adversarial Opinion Manipulation Attacks to Retrieval-Augmented Generation Models☆17Jun 21, 2025Updated last year
- [HPCA'24] Smart-Infinity: Fast Large Language Model Training using Near-Storage Processing on a Real System☆52Jul 21, 2025Updated last year
- Code for the ACL 2022 (Long paper): "New Intent Discovery with Pre-training and Contrastive Learning".☆14Jul 18, 2022Updated 4 years ago
- [NeurIPS'23] Binary Classification with Confidence Difference☆10May 13, 2024Updated 2 years ago
- A simple and efficient llama3 local service deployment solution that supports real-time streaming response and is optimized for common Ch…☆13Jul 31, 2024Updated 2 years ago
- ☆10Sep 27, 2022Updated 3 years ago
- Code and data for the paper: On the Reliability of Psychological Scales on Large Language Models☆31Dec 15, 2025Updated 7 months ago
- ODYSSEUS/EduCOSMOS☆10Aug 12, 2020Updated 6 years ago
- ☆10Jun 21, 2021Updated 5 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- [KDD'22] Partial Label Learning with Discrimination Augmentation☆10May 21, 2024Updated 2 years ago
- Learning from Indirect Observations☆11Jul 16, 2021Updated 5 years ago
- Qualifying Exam Preparing☆18May 7, 2025Updated last year
- Official repository of paper [FALQON: Accelerating LoRA Fine-tuning with Low-Bit Floating-Point Arithmetic, NeurIPS 2025]☆21Dec 2, 2025Updated 8 months ago
- A benchmark on visual perception in text strings for both LLMs and MLLMs.☆16Apr 7, 2026Updated 4 months ago
- Here, we compare Q(\sigma) learning presented by Sutton and Barto in [1] to Tree-Backup, n-step Expected Sarsa, and n-step Sarsa.☆15Feb 17, 2017Updated 9 years ago
- [ICLR 2026] | MMSU: A Massive Multi-task Spoken Language Understanding and Reasoning Benchmark☆17Feb 12, 2026Updated 6 months ago
- Serial Contrastive Knowledge Distillation for Continual Few-shot Relation Extraction, Findings of ACL 2023☆14May 12, 2023Updated 3 years ago
- Code for Semi-crowdsourced Clustering with Deep Generative Models☆12Dec 9, 2022Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [NeurIPS 2022] "Adversarial Training with Complementary Labels: On the Benefit of Gradually Informative Attacks"☆13Nov 11, 2022Updated 3 years ago
- It's All In the Teacher: Zero-Shot Quantization Brought Closer to the Teacher [CVPR 2022 Oral]☆29Sep 15, 2022Updated 3 years ago
- Smoothing video traffic to make it a friendlier internet neighbor☆14Apr 23, 2024Updated 2 years ago
- [ICML 2024] Code release for "On the Emergence of Cross-Task Linearity in Pretraining-Finetuning Paradigm"☆11Feb 20, 2025Updated last year
- Repository for the DPP'23 course☆11May 2, 2024Updated 2 years ago
- A Transformer Framework Based Couplet Task☆23Oct 29, 2023Updated 2 years ago
- A browser extension that uses ChatGPT to group your browser tabs into logical categories.☆23Sep 26, 2024Updated last year