Can AI Agents Build Bespoke Systems?
β89Aug 8, 2026Updated this week
Alternatives and similar repositories for vibesys
Users that are interested in vibesys are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- β20Dec 4, 2025Updated 8 months ago
- [MLSys 26] π₯ Solution for Gated Delta Net Track of MLSys 26 Flash infer competitionβ36May 22, 2026Updated 2 months ago
- An open toolkit and public dataset hub for collecting, sanitizing, analyzing, and visualizing coding agent traces.β84Jul 24, 2026Updated 2 weeks ago
- A Streaming-Native Serving Engine for TTS/STS Modelsβ74Jun 20, 2026Updated last month
- [ASPLOS' 26] TetriServe: Efficiently Serving Mixed DiT Workloadsβ17Mar 12, 2026Updated 4 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- SC'25 UltraAttn: Efficiently Parallelizing Attention through Hierarchical Context-Tilingβ16Aug 14, 2025Updated 11 months ago
- A high-performance, universal serving framework for any-to-any models.β67Updated this week
- A language-modelβpowered compressor for natural language textβ53Oct 23, 2025Updated 9 months ago
- β37Aug 7, 2025Updated last year
- Vortex: Programmable Sparse Attention for Agents as Algorithm Designersβ68Updated this week
- β16May 27, 2026Updated 2 months ago
- β46Oct 15, 2025Updated 9 months ago
- AI-Driven Research Systems (ADRS)β147Dec 17, 2025Updated 7 months ago
- [MLSys 2026] AccelOpt: Self-improving Agents for AI Accelerator Kernel Optimizationβ57Jul 28, 2026Updated last week
- Managed Database hosting by DigitalOcean β’ AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- TurboServe: Serving Streaming Video Generation Efficiently and Economicallyβ40Jul 12, 2026Updated 3 weeks ago
- [EuroSys'25] Mist: Efficient Distributed Training of Large Language Models via Memory-Parallelism Co-Optimizationβ24Apr 13, 2026Updated 3 months ago
- β16Jun 15, 2026Updated last month
- Agent-friendly GPU profile-query CLIβ108Updated this week
- Artifact Evaluation for SpecFS [FAST'26]β34Dec 28, 2025Updated 7 months ago
- An agent for CUDA compute-communication kernel co-designβ36May 7, 2026Updated 3 months ago
- High Performance Sorting Based Distributed memory K-mer counterβ15Dec 8, 2025Updated 8 months ago
- Dataflow-Oriented Reinforcement Learning for (Multi-)Agentic LLMsβ99Jul 26, 2026Updated 2 weeks ago
- Code repository for the SOSP'25 paper DCP: Addressing Input Dynamism In Long-Context Training via Dynamic Context Parallelism.β22Nov 28, 2025Updated 8 months ago
- Managed Kubernetes at scale on DigitalOcean β’ AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- A throughput-oriented high-performance serving framework for LLMsβ974Mar 29, 2026Updated 4 months ago
- β364Jun 9, 2026Updated 2 months ago
- Winner π (Agent-only) MLSys 2026 - FlashInfer AI Kernel Generation Contest for the DeepSeek Sparse Attention (DSA) track with an averageβ¦β151Jun 10, 2026Updated last month
- β813Jun 2, 2026Updated 2 months ago
- ζ΅ζ±ε€§ε¦ζδ½η³»η»θ―Ύη¨δ»εΊοΌζζ‘£β18Jan 14, 2026Updated 6 months ago
- Translate C macros and condtional compilation to Rustβ27Aug 2, 2026Updated last week
- vLLM Daily Summarization of Merged PRsβ53Updated this week
- PaperHelper: Knowledge-Based LLM QA Paper Reading Assistant with Reliable Referencesβ22Jun 13, 2024Updated 2 years ago
- β47Jul 27, 2026Updated 2 weeks ago
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- β24Jan 18, 2026Updated 6 months ago
- A benchmarking framework for on-device AIβ20Jul 30, 2026Updated last week
- Foundry materializes CUDA graphs along with its execution context to disk to support fast cold start of serving engines.β52Jul 8, 2026Updated last month
- β242Jul 29, 2026Updated last week
- [ICLR'25] Fast Inference of MoE Models with CPU-GPU Orchestrationβ266Nov 18, 2024Updated last year
- mKernel: fast multi-node, multi-GPU fused kernelsβ263Aug 3, 2026Updated last week
- β177May 24, 2026Updated 2 months ago