训练营训练方向项目
☆30Jan 28, 2026Updated 6 months ago
Alternatives and similar repositories for TinyInfiniTrain
Users that are interested in TinyInfiniTrain are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Let's Learn AI SYStem☆52Jul 13, 2026Updated 3 weeks ago
- InfiniTensor 大模型与人工智能系统训练营 CUDA 方向作业与项目系统☆56Updated this week
- ☆47Updated this week
- 实验:rust 实现 llama2 推理☆17Feb 23, 2024Updated 2 years ago
- ☆17Jul 17, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- InfiniTensor is a high-performance inference engine tailored for GPUs and AI accelerators. Its design focuses on effective deployment and…☆377Updated this week
- [ICDCS 2023] Evaluation and Optimization of Gradient Compression for Distributed Deep Learning☆10Apr 28, 2023Updated 3 years ago
- ☆187Updated this week
- [AFK] Hardware router in Chisel (THU Network Joint Lab 2020)☆14Oct 8, 2020Updated 5 years ago
- 电子科技大学的计算机组成原理实验课,并在此基础上实现了汇编编译器☆10Jun 14, 2022Updated 4 years ago
- ☆12Jan 19, 2022Updated 4 years ago
- 算子库☆17Jul 9, 2025Updated last year
- A domain-specific language (DSL) based on Triton but providing higher-level abstractions.☆331Updated this week
- Reading seminar in Harvard Cloud Networking and Systems Group☆16Aug 29, 2022Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- XDMA PCIe to DDR4 and GPIO and BRAM for the Innova-2 Flex XCKU15P FPGA☆22Mar 7, 2024Updated 2 years ago
- ☆17May 10, 2024Updated 2 years ago
- ☆20Dec 24, 2023Updated 2 years ago
- 根据计算机组成与原理的课程设计要求编写的 cpu 模拟器,可以读取特定的汇编指令集文件,并以执行一条微指令为最小单位进行单步执行和全部执行。☆13Dec 26, 2019Updated 6 years ago
- SC'25 UltraAttn: Efficiently Parallelizing Attention through Hierarchical Context-Tiling☆16Aug 14, 2025Updated 11 months ago
- ☆16Mar 24, 2026Updated 4 months ago
- DeDe (OSDI '25): an optimization framework for large-scale resource allocation☆15May 18, 2026Updated 2 months ago
- ☆28Updated this week
- ☆20Jun 1, 2026Updated 2 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Autonomous CUDA kernel optimization agent with structured task specs and per-config scoring☆17Jun 17, 2026Updated last month
- 我的一生一芯项目☆16Dec 14, 2021Updated 4 years ago
- Deferred Continuous Batching in Resource-Efficient Large Language Model Serving (EuroMLSys 2024)☆19May 28, 2024Updated 2 years ago
- Tiny-DeepSpeed, a minimalistic re-implementation of the DeepSpeed library☆52Aug 20, 2025Updated 11 months ago
- [NSDI'26] PolyRL is a reinforcement learning framework for LLM that harvest spot instances on the cloud to reduce cost.☆19Mar 30, 2026Updated 4 months ago
- [Accepted to SOSP 2026] Fast Deterministic LLM Inference☆24Jul 24, 2026Updated 2 weeks ago
- Official implementation of CrossPipe: Towards Optimal Pipeline Schedules for Cross-Datacenter Training (ATC '25), built on top of Megatro…☆17Jul 6, 2025Updated last year
- A lightweight post-training framework for LLMs and VLMs. 51 algorithms, 38 verified models. Scales with DeepSpeed, vLLM, and Ray.☆19Updated this week
- An Efficient Design of Intelligent Network Data Plane☆64Mar 16, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Arya: Arbitrary Graph Pattern Mining with Decomposition-based Sampling☆18Sep 27, 2023Updated 2 years ago
- Artifact of Chimera☆18May 6, 2025Updated last year
- Code repository for the SOSP'25 paper DCP: Addressing Input Dynamism In Long-Context Training via Dynamic Context Parallelism.☆22Nov 28, 2025Updated 8 months ago
- ☆14Aug 18, 2024Updated last year
- This is the implementation repository of our SOSP'24 paper: Aceso: Achieving Efficient Fault Tolerance in Memory-Disaggregated Key-Value …☆24Oct 20, 2024Updated last year
- BytePS examples (Vision, NLP, GAN, etc)☆19Nov 24, 2022Updated 3 years ago
- 国科大一生一芯第二期: RISCV-64 五级流水线CPU☆19Apr 17, 2021Updated 5 years ago