Can AI Agents Build Bespoke Systems?
β84Jul 20, 2026Updated this week
Alternatives and similar repositories for vibesys
Users that are interested in vibesys are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- β20Dec 4, 2025Updated 7 months ago
- [MLSys 26] π₯ Solution for Gated Delta Net Track of MLSys 26 Flash infer competitionβ35May 22, 2026Updated last month
- An open toolkit and public dataset hub for collecting, sanitizing, analyzing, and visualizing coding agent traces.β49Jul 2, 2026Updated 2 weeks ago
- A Streaming-Native Serving Engine for TTS/STS Modelsβ74Jun 20, 2026Updated last month
- [ASPLOS' 26] TetriServe: Efficiently Serving Mixed DiT Workloadsβ17Mar 12, 2026Updated 4 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- SC'25 UltraAttn: Efficiently Parallelizing Attention through Hierarchical Context-Tilingβ16Aug 14, 2025Updated 11 months ago
- A high-performance, universal serving framework for any-to-any models.β56Updated this week
- A language-modelβpowered compressor for natural language textβ53Oct 23, 2025Updated 8 months ago
- β37Aug 7, 2025Updated 11 months ago
- Vortex: Programmable Sparse Attention for Agents as Algorithm Designersβ68Jun 24, 2026Updated 3 weeks ago
- β16May 27, 2026Updated last month
- β45Oct 15, 2025Updated 9 months ago
- AI-Driven Research Systems (ADRS)β145Dec 17, 2025Updated 7 months ago
- [MLSys 2026] AccelOpt: Self-improving Agents for AI Accelerator Kernel Optimizationβ57Jun 18, 2026Updated last month
- Virtual machines for every use case on DigitalOcean β’ AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- [EuroSys'25] Mist: Efficient Distributed Training of Large Language Models via Memory-Parallelism Co-Optimizationβ24Apr 13, 2026Updated 3 months ago
- TurboServe: Serving Streaming Video Generation Efficiently and Economicallyβ33Jul 12, 2026Updated last week
- β16Jun 15, 2026Updated last month
- Agent-friendly GPU profile-query CLIβ104Jun 22, 2026Updated 3 weeks ago
- Artifact Evaluation for SpecFS [FAST'26]β33Dec 28, 2025Updated 6 months ago
- An agent for CUDA compute-communication kernel co-designβ35May 7, 2026Updated 2 months ago
- High Performance Sorting Based Distributed memory K-mer counterβ15Dec 8, 2025Updated 7 months ago
- Dataflow-Oriented Reinforcement Learning for (Multi-)Agentic LLMsβ95Updated this week
- Code repository for the SOSP'25 paper DCP: Addressing Input Dynamism In Long-Context Training via Dynamic Context Parallelism.β21Nov 28, 2025Updated 7 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A throughput-oriented high-performance serving framework for LLMsβ968Mar 29, 2026Updated 3 months ago
- β310Jun 9, 2026Updated last month
- Implementation from scratch in C of the Multi-head latent attention used in the Deepseek-v3 technical paper.β18Jan 15, 2025Updated last year
- Winner π (Agent-only) MLSys 2026 - FlashInfer AI Kernel Generation Contest for the DeepSeek Sparse Attention (DSA) track with an averageβ¦β148Jun 10, 2026Updated last month
- ζ΅ζ±ε€§ε¦ζδ½η³»η»θ―Ύη¨δ»εΊοΌζζ‘£β18Jan 14, 2026Updated 6 months ago
- β754Jun 2, 2026Updated last month
- Translate C macros and condtional compilation to Rustβ26Jun 23, 2026Updated 3 weeks ago
- vLLM Daily Summarization of Merged PRsβ51Updated this week
- PaperHelper: Knowledge-Based LLM QA Paper Reading Assistant with Reliable Referencesβ22Jun 13, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- β41Jul 1, 2026Updated 2 weeks ago
- β23Jan 18, 2026Updated 6 months ago
- A benchmarking framework for on-device AIβ19Jun 7, 2026Updated last month
- Foundry materializes CUDA graphs along with its execution context to disk to support fast cold start of serving engines.β45Jul 8, 2026Updated last week
- β231Updated this week
- mKernel: fast multi-node, multi-GPU fused kernelsβ251Jun 21, 2026Updated 3 weeks ago
- [ICLR'25] Fast Inference of MoE Models with CPU-GPU Orchestrationβ267Nov 18, 2024Updated last year