awesome LLM papers! 🚀 🚀 🚀
☆48Jul 3, 2025Updated last year
Alternatives and similar repositories for Awesome-Awesome-LLM
Users that are interested in Awesome-Awesome-LLM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 🌟 The ultimate meta-collection of 200+ awesome LLM repositories☆33Updated this week
- ☆17Jun 10, 2025Updated last year
- When Reasoning Meets Its Laws☆38Jan 2, 2026Updated 7 months ago
- GAMER: Generative Augmentation and Multi-Level Behavior Modeling for Sequential Recommendation☆32Jul 29, 2026Updated last week
- Transformer from Scratch in PyTorch☆18Mar 26, 2022Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆14Apr 1, 2024Updated 2 years ago
- ☆17Nov 20, 2024Updated last year
- Tiny-R2: A hybrid architecture integrating SWA, CSA, HCA, mHC, and DSMoE under the DeepSeek V4 design paradigm, enabling single-GPU OPD p…☆46May 30, 2026Updated 2 months ago
- ☆16Jun 15, 2026Updated last month
- Hello-RAG 从零实现RAG及其各种高级用法☆20Jul 5, 2025Updated last year
- Project of ACL 2025 "UAlign: Leveraging Uncertainty Estimations for Factuality Alignment on Large Language Models"☆15Mar 25, 2025Updated last year
- This is the official code of the paper "SupFusion: Supervised LiDAR-Camera Fusion for 3D Object Detection"☆17Aug 23, 2023Updated 2 years ago
- Code for Multi-Aspect Cross-modal Quantization for Generative Recommendation. (AAAI 2026 Oral)☆46Dec 9, 2025Updated 8 months ago
- 电子科技大学编译原理实验:设计程序,能够对特定Pascal代码进行语法分析和词法分析。☆15Apr 23, 2019Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- "Parallel Test-Time Scaling for Latent Reasoning Models"☆24Apr 12, 2026Updated 3 months ago
- ☆16Feb 6, 2025Updated last year
- SMART introduces a novel test-time framework where Small Language Models (SLMs) reason step-by-step, and Large Language Models (LLMs) pro…☆12Jul 9, 2025Updated last year
- 电子科技大学2020年《人工智能》课程的平时作业和实验☆14Nov 26, 2022Updated 3 years ago
- Official implementation of the paper "Understanding Language Prior of LVLMs by Contrasting Chain-of-Embedding"☆18Sep 30, 2025Updated 10 months ago
- [AAAI 2025] Neural-Symbolic Collaborative Distillation: Advancing Small Language Models for Complex Reasoning Tasks☆12Jun 19, 2025Updated last year
- adong's skills☆28May 10, 2026Updated 3 months ago
- ☆18Mar 15, 2026Updated 4 months ago
- KnowLA: Enhancing Parameter-efficient Finetuning with Knowledgeable Adaptation, NAACL 2024☆16Jul 29, 2024Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆14Sep 2, 2024Updated last year
- DRACO: Byzantine-resilient Distributed Training via Redundant Gradients☆23Dec 9, 2018Updated 7 years ago
- ☆19Apr 12, 2026Updated 3 months ago
- Source codes for the paper "Personalized Dynamic Music Emotion Recognition with Dual-Scale Attention-Based Meta-Learning" (PDMER) which p…☆14Mar 24, 2025Updated last year
- v1: Learning to Point Visual Tokens for Multimodal Grounded Reasoning☆21Updated this week
- [KDD'25] Code of "Bridging Textual-Collaborative Gap through Semantic Codes for Sequential Recommendation".☆18Jul 3, 2026Updated last month
- The official code for paper "Token-Level Collaborative Alignment for LLM-based Generative Recommendation"☆19Jun 9, 2026Updated 2 months ago
- 实现CS336的作业1,并从头开始构建一个transformer模型。Implement CS336's job 1 and build a transformer from scratch☆15Aug 1, 2025Updated last year
- Code repository for the ICML 2026 Oral paper "Characterizing, Evaluating, and Optimizing Complex Reasoning".☆18Jun 21, 2026Updated last month
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆14Jul 6, 2023Updated 3 years ago
- [KDD 2026] "Breaking Information Cocoons: A Hyperbolic Graph-LLM Framework for Exploration and Exploitation in Recommender Systems"☆16Jan 29, 2025Updated last year
- ☆22Dec 18, 2025Updated 7 months ago
- LightRFT (Light Reinforcement Fine-Tuning) is an advanced reinforcement learning fine-tuning framework designed for Large Language Models…☆19Jan 12, 2026Updated 6 months ago
- [EMNLP 2025] Reasoning-to-Defend: Safety-Aware Reasoning Can Defend Large Language Models from Jailbreaking☆12Aug 22, 2025Updated 11 months ago
- UESTC-Operating System Experiment(电子科技大学 操作系统 实验)一、进程与资源管理实验 二、虚拟内存综合实验☆13Jun 18, 2019Updated 7 years ago
- ☆94May 4, 2026Updated 3 months ago