Can AI Agents Build Bespoke Systems?
β98Sep 19, 2026Updated this week
Alternatives and similar repositories for vibesys
Users that are interested in vibesys are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- β20Dec 4, 2025Updated 9 months ago
- [MLSys 26] π₯ Solution for Gated Delta Net Track of MLSys 26 Flash infer competitionβ36May 22, 2026Updated 3 months ago
- An open toolkit and public dataset hub for collecting, sanitizing, analyzing, and visualizing coding agent traces.β124Aug 22, 2026Updated 3 weeks ago
- A Streaming-Native Serving Engine for TTS/STS Modelsβ79Jun 20, 2026Updated 2 months ago
- [ASPLOS' 26] TetriServe: Efficiently Serving Mixed DiT Workloadsβ19Mar 12, 2026Updated 6 months ago
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- SC'25 UltraAttn: Efficiently Parallelizing Attention through Hierarchical Context-Tilingβ16Aug 14, 2025Updated last year
- A high-performance, universal serving framework for any-to-any models.β88Updated this week
- A language-modelβpowered compressor for natural language textβ53Oct 23, 2025Updated 10 months ago
- β37Aug 7, 2025Updated last year
- Vortex: Programmable Sparse Attention for Agents as Algorithm Designersβ68Aug 22, 2026Updated 3 weeks ago
- β16May 27, 2026Updated 3 months ago
- β46Oct 15, 2025Updated 11 months ago
- AI-Driven Research Systems (ADRS)β154Dec 17, 2025Updated 9 months ago
- [MLSys 2026] AccelOpt: Self-improving Agents for AI Accelerator Kernel Optimizationβ74Jul 28, 2026Updated last month
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [EuroSys'25] Mist: Efficient Distributed Training of Large Language Models via Memory-Parallelism Co-Optimizationβ24Apr 13, 2026Updated 5 months ago
- β17Jun 15, 2026Updated 3 months ago
- Agent-friendly GPU profile-query CLIβ122Aug 21, 2026Updated 3 weeks ago
- Artifact Evaluation for SpecFS [FAST'26]β34Dec 28, 2025Updated 8 months ago
- An agent for CUDA compute-communication kernel co-designβ36May 7, 2026Updated 4 months ago
- High Performance Sorting Based Distributed memory K-mer counterβ15Dec 8, 2025Updated 9 months ago
- Dataflow-Oriented Reinforcement Learning for (Multi-)Agentic LLMsβ103Jul 26, 2026Updated last month
- Code repository for the SOSP'25 paper DCP: Addressing Input Dynamism In Long-Context Training via Dynamic Context Parallelism.β22Nov 28, 2025Updated 9 months ago
- A throughput-oriented high-performance serving framework for LLMsβ976Mar 29, 2026Updated 5 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer β’ AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Kernel Design Agents (KDA) is a agent-centric workflow to write high-performance CUDA Kernels.β1,049Updated this week
- β457Aug 26, 2026Updated 3 weeks ago
- ζ΅ζ±ε€§ε¦ζδ½η³»η»θ―Ύη¨δ»εΊοΌζζ‘£β18Jan 14, 2026Updated 8 months ago
- Translate C macros and condtional compilation to Rustβ27Aug 2, 2026Updated last month
- TurboServe: Serving Streaming Video Generation Efficiently and Economicallyβ140Jul 12, 2026Updated 2 months ago
- vLLM Daily Summarization of Merged PRsβ54Updated this week
- PaperHelper: Knowledge-Based LLM QA Paper Reading Assistant with Reliable Referencesβ22Jun 13, 2024Updated 2 years ago
- β51Jul 27, 2026Updated last month
- β25Jan 18, 2026Updated 8 months ago
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A benchmarking framework for on-device AIβ20Aug 20, 2026Updated 3 weeks ago
- Foundry materializes CUDA graphs along with its execution context to disk to support fast cold start of serving engines.β61Updated this week
- Winner π (Agent-only) MLSys 2026 - FlashInfer AI Kernel Generation Contest for the DeepSeek Sparse Attention (DSA) track with an averageβ¦β307Sep 12, 2026Updated last week
- β284Sep 5, 2026Updated 2 weeks ago
- [ICLR'25] Fast Inference of MoE Models with CPU-GPU Orchestrationβ268Nov 18, 2024Updated last year
- mKernel: fast multi-node, multi-GPU fused kernelsβ277Updated this week
- β232Aug 26, 2026Updated 3 weeks ago