Full stack LLM (Pre-training/finetuning, PPO(RLHF), Inference, Quant, etc.)
☆31Feb 21, 2025Updated last year
Alternatives and similar repositories for FullLLM
Users that are interested in FullLLM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Simplest AlphaZero Implementation☆25Nov 6, 2024Updated last year
- ☆10Nov 1, 2021Updated 4 years ago
- Official codes of KDD'24 paper "HiFGL: A Hierarchical Framework for Cross-silo Cross-device Federated Graph Learning"☆10Sep 4, 2024Updated 2 years ago
- ☆14May 13, 2025Updated last year
- Python SDK for DBD Products☆15Feb 13, 2026Updated 7 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [CVPR'25] Attention IoU: Examining Biases in CelebA using Attention Maps☆13Mar 26, 2025Updated last year
- 轻量化的实例分割模型☆14Apr 12, 2020Updated 6 years ago
- [ACL 2023] Solving Math Word Problems via Cooperative Reasoning induced Language Models (LLMs + MCTS + Self-Improvement)☆51Dec 15, 2023Updated 2 years ago
- Jina Embedding Models on AWS SageMaker☆15Jun 4, 2026Updated 3 months ago
- ☆11May 9, 2023Updated 3 years ago
- Papers on fairness☆12Oct 20, 2020Updated 5 years ago
- ☆90Sep 15, 2026Updated last week
- ☆15Jan 14, 2026Updated 8 months ago
- ☆47Nov 8, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- NeurIPS 2024: SciFIBench: Benchmarking Large Multimodal Models for Scientific Figure Interpretation☆13May 24, 2025Updated last year
- ☆11Jun 15, 2019Updated 7 years ago
- 使用Few-Shot方法来做文本分类任务,基于THUCNews数据☆10Jun 4, 2020Updated 6 years ago
- Official implementation of the paper "ALTER: Augmentation for Large-Table-Based Reasoning"☆15Aug 26, 2024Updated 2 years ago
- ☆14Oct 11, 2023Updated 2 years ago
- ☆17Jun 23, 2020Updated 6 years ago
- ☆20Apr 9, 2025Updated last year
- Code for the EACL 2024 paper: "Small Language Models Improve Giants by Rewriting Their Outputs"☆12Apr 20, 2024Updated 2 years ago
- ChartSum is a large scale benchmark for automatic chart to text summarization☆11Jul 20, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Implementation of Kronecker Attention in Pytorch☆20Sep 12, 2020Updated 6 years ago
- ☆14Mar 9, 2023Updated 3 years ago
- Estimate the Deterministic Input, Noisy "And" Gate (DINA) cognitive diagnostic model parameters using the Gibbs sampler described by Culp…☆16Sep 28, 2025Updated 11 months ago
- ☆135Jul 8, 2024Updated 2 years ago
- Train I3D on NTU-RGB+D dataset in keras☆11Feb 5, 2019Updated 7 years ago
- ☆16Sep 8, 2025Updated last year
- HalluciDoctor: Mitigating Hallucinatory Toxicity in Visual Instruction Data (Accepted by CVPR 2024)☆52Jul 16, 2024Updated 2 years ago
- ☆13Jun 18, 2019Updated 7 years ago
- ☆13Oct 19, 2022Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Experiments generating text with state-of-the-art deep-learning models (GPT2, Transformer XL, ...)☆13May 15, 2019Updated 7 years ago
- A huge dataset for Document Visual Question Answering☆24Jul 29, 2024Updated 2 years ago
- ☆27Jan 29, 2026Updated 7 months ago
- AI-TP: Attention-based Interaction-aware Trajectory Prediction for Autonomous Driving☆24Mar 7, 2024Updated 2 years ago
- Official code for the paper: Scaling Transformers for Discriminative Recommendation via Generative Pretraining☆32Sep 1, 2025Updated last year
- ☆18Mar 28, 2021Updated 5 years ago
- This repository contains reference implementation for multi-LLM ToM paper (accepted to EMNLP 2023), Theory of Mind for Multi-Agent Collab…☆20Jun 11, 2024Updated 2 years ago