Throughput-oriented multi-turn inference engine for KernelBench [ICML '25]
☆24May 27, 2025Updated last year
Alternatives and similar repositories for caesar
Users that are interested in caesar are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆68Jul 14, 2025Updated last year
- Ship correct and fast LLM kernels to PyTorch☆155Jan 14, 2026Updated 8 months ago
- ☆28Mar 31, 2026Updated 6 months ago
- Horizontal Fusion☆24Jan 7, 2022Updated 4 years ago
- TritonBench: Benchmarking Large Language Model Capabilities for Generating Triton Operators☆143Jun 14, 2025Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Document Automation Reference Kit☆16Jun 27, 2024Updated 2 years ago
- KernelBench: Can LLMs Write GPU Kernels? - Benchmark + Toolkit with Torch -> CUDA (+ more DSLs)☆1,283Mar 24, 2026Updated 6 months ago
- This project includes code for using the AsyncWebRL and WebGym frameworks to train web agent models.☆49Jun 9, 2026Updated 4 months ago
- CUDA-L1: Improving CUDA Optimization via Contrastive Reinforcement Learning☆324Nov 3, 2025Updated 11 months ago
- Use the tokenizer in parallel to achieve superior acceleration☆20Mar 21, 2024Updated 2 years ago
- Code for the paper "Function-Space Learning Rates"☆23Jun 3, 2025Updated last year
- ☆107Nov 22, 2025Updated 10 months ago
- ☆10Oct 8, 2021Updated 5 years ago
- Winner 🏆 (Agent-only) MLSys 2026 - FlashInfer AI Kernel Generation Contest for the DeepSeek Sparse Attention (DSA) track with an average…☆313Sep 12, 2026Updated 3 weeks ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- A very hacky set of functions for getting plotly to do what I want when doing mech interp research, designed to be compatible with PyTorc…☆15Jun 16, 2023Updated 3 years ago
- [ICML2024 Spotlight] Fine-Tuning Pre-trained Large Language Models Sparsely☆24Jun 26, 2024Updated 2 years ago
- 基于PyTorch GPT-2的针对各种数据并行pretrain的研究代码.☆11Dec 16, 2022Updated 3 years ago
- Automated High-Performance GPU Kernel Generation☆139Jun 1, 2026Updated 4 months ago
- ☆14Oct 7, 2023Updated 3 years ago
- Vortex: Programmable Sparse Attention for Agents as Algorithm Designers☆69Aug 22, 2026Updated last month
- Personal solutions to the Triton Puzzles☆22Jul 18, 2024Updated 2 years ago
- Moatless Testbeds allows you to create isolated testbed environments in a Kubernetes cluster where you can apply code changes through git…☆14Apr 9, 2025Updated last year
- ☆17Jun 8, 2019Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official implementation of TBA for async LLM post-training.☆32Nov 5, 2025Updated 11 months ago
- ☆18Feb 4, 2026Updated 8 months ago
- Agent-native Seedance 2.0 short-film studio: cli for AI, canvas for human☆16Sep 8, 2026Updated last month
- Building the Virtuous Cycle for AI-driven LLM Systems☆289Sep 25, 2026Updated 2 weeks ago
- TREC QA dataset for question answering cleaned for usage in Question Answering☆14Aug 26, 2019Updated 7 years ago
- Odysseus: Playground of LLM Sequence Parallelism☆83Jun 17, 2024Updated 2 years ago
- Fine-Tuning Pre-trained Transformers into Decaying Fast Weights☆20Oct 9, 2022Updated 4 years ago
- 🔥 LLM-powered GPU kernel synthesis: Train models to convert PyTorch ops into optimized Triton kernels via SFT+RL. Multi-turn compilation…☆151Nov 10, 2025Updated 10 months ago
- [Archived] Superseded by humanfia/humanize — workflow-native terminal coding agent, a fork of can1357/oh-my-pi☆76Updated this week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- PerFlow-AI is a programmable performance analysis, modeling, prediction tool for AI system.☆33May 12, 2026Updated 4 months ago
- Implements the Messenger environment and EMMA model.☆25Jun 14, 2023Updated 3 years ago
- ☆22May 30, 2026Updated 4 months ago
- Recursive Bayesian Networks☆11May 11, 2025Updated last year
- Enhancing the convergence speed by 2x and improving the training success of Physics-Informed Neural Networks (PINNs).☆14Oct 14, 2024Updated last year
- Silly twitter torch implementations.☆48Oct 14, 2022Updated 3 years ago
- ☆14May 12, 2025Updated last year