mini project for nanorllm
☆62Mar 31, 2026Updated 5 months ago
Alternatives and similar repositories for nanorllm
Users that are interested in nanorllm are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This is the code of a agentic rag method with dynamic workflow.☆15Jan 22, 2026Updated 7 months ago
- This is the implementation repository of our SOSP'24 paper: Aceso: Achieving Efficient Fault Tolerance in Memory-Disaggregated Key-Value …☆24Oct 20, 2024Updated last year
- Implementation of the logging layer of our SOSP '23 paper Halfmoon☆11Jul 28, 2023Updated 3 years ago
- APRIL: Active Partial Rollouts in Reinforcement Learning to Tame Long-tail Generation. A system-level optimization for scalable LLM tra…☆63Oct 11, 2025Updated 10 months ago
- An LLM training framework built from the ground up, featuring a custom BumbleBee architecture and end-to-end support for multiple open-so…☆101Aug 24, 2026Updated last week
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- [ICDCS 2023] Evaluation and Optimization of Gradient Compression for Distributed Deep Learning☆10Apr 28, 2023Updated 3 years ago
- This repository contains the TLA+ specification of the ownership and the reliable commit protocols for transactions in Zeus work that app…☆20Jun 12, 2022Updated 4 years ago
- [ICML 2026] Less Is More: Training-Free Sparse Attention with Global Locality for Efficient Reasoning☆35Sep 12, 2025Updated 11 months ago
- 华科东五实验室Latex Beamer PPT模板☆22May 3, 2018Updated 8 years ago
- [ICML 2025] AdaDecode: Accelerating LLM Decoding with Adaptive Layer Parallelism☆20Jul 14, 2025Updated last year
- A local cache for the Anna KVS☆24Jun 2, 2020Updated 6 years ago
- ☆51Apr 15, 2026Updated 4 months ago
- [AFK] Hardware router in Chisel (THU Network Joint Lab 2020)☆14Oct 8, 2020Updated 5 years ago
- Pipeline-Parallel Lecture: Simplest Dualpipe Implementation.☆31Sep 17, 2025Updated 11 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- 这是一个从零开始构建的强化学习人类反馈(RLHF)学习代码库,实现了 PPO、GRPO、GSPO 以及相关的策略优化算法,并提供了清晰、可复现的训练流程。由于文档是由latex文件转译过来,如果md文件渲染异常,请用VScode的md插件打开☆89Dec 19, 2025Updated 8 months ago
- Fixing GRPO training collapse in long-horizon multi-tool agents. A lightweight PRM-Lite + LATA joint approach achieves +37% over vanilla …☆215Jun 27, 2026Updated 2 months ago
- The code for the paper "Adversarial Decomposition of Text Representation", NAACL 2019☆29Dec 8, 2022Updated 3 years ago
- Claw-R1: Empowering OpenClaw with Advanced Agentic RL.☆195Aug 22, 2026Updated last week
- An LLM Agent Framework for Automated 3D Cutscene Generation.☆18Jun 29, 2026Updated 2 months ago
- Democratizing Reinforcement Learning for LLMs☆5,808Aug 24, 2026Updated last week
- Website about the project WorkflowHub☆13Aug 21, 2026Updated last week
- A MVP implementation of distributed query engine cut from datafusion-ballista codebase for learning purpose.☆12Jan 10, 2025Updated last year
- Reading seminar in Harvard Cloud Networking and Systems Group☆16Aug 29, 2022Updated 4 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆1,309May 20, 2026Updated 3 months ago
- Code and Data for PeerQA: A Scientific Question Answering Dataset from Peer Reviews, NAACL 2025 https://aclanthology.org/2025.naacl-long.…☆16Jun 1, 2026Updated 2 months ago
- ☆13Jun 1, 2022Updated 4 years ago
- ☆11Nov 4, 2022Updated 3 years ago
- ☆17May 10, 2024Updated 2 years ago
- A project implementing various agentic RL based on the Slime post-training framework☆527Apr 11, 2026Updated 4 months ago
- Lithops application examples☆11Dec 19, 2024Updated last year
- ☆14May 21, 2024Updated 2 years ago
- SC'25 UltraAttn: Efficiently Parallelizing Attention through Hierarchical Context-Tiling☆16Aug 14, 2025Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆25Jul 7, 2024Updated 2 years ago
- ☆16Mar 24, 2026Updated 5 months ago
- Tiny-FSDP, a minimalistic re-implementation of the PyTorch FSDP☆111Aug 20, 2025Updated last year
- ☆22Dec 14, 2024Updated last year
- ☆30Oct 3, 2022Updated 3 years ago
- DeDe (OSDI '25): an optimization framework for large-scale resource allocation☆15May 18, 2026Updated 3 months ago
- codes for Efficient Test-Time Scaling via Self-Calibration☆21Sep 13, 2025Updated 11 months ago