Can AI Agents Build Bespoke Systems?
β92Aug 29, 2026Updated this week
Alternatives and similar repositories for vibesys
Users that are interested in vibesys are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- β20Dec 4, 2025Updated 8 months ago
- [MLSys 26] π₯ Solution for Gated Delta Net Track of MLSys 26 Flash infer competitionβ36May 22, 2026Updated 3 months ago
- An open toolkit and public dataset hub for collecting, sanitizing, analyzing, and visualizing coding agent traces.β104Aug 22, 2026Updated last week
- A Streaming-Native Serving Engine for TTS/STS Modelsβ79Jun 20, 2026Updated 2 months ago
- [ASPLOS' 26] TetriServe: Efficiently Serving Mixed DiT Workloadsβ18Mar 12, 2026Updated 5 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- SC'25 UltraAttn: Efficiently Parallelizing Attention through Hierarchical Context-Tilingβ16Aug 14, 2025Updated last year
- A high-performance, universal serving framework for any-to-any models.β80Updated this week
- A language-modelβpowered compressor for natural language textβ53Oct 23, 2025Updated 10 months ago
- β37Aug 7, 2025Updated last year
- Vortex: Programmable Sparse Attention for Agents as Algorithm Designersβ68Aug 22, 2026Updated last week
- β16May 27, 2026Updated 3 months ago
- β46Oct 15, 2025Updated 10 months ago
- AI-Driven Research Systems (ADRS)β152Dec 17, 2025Updated 8 months ago
- [MLSys 2026] AccelOpt: Self-improving Agents for AI Accelerator Kernel Optimizationβ61Jul 28, 2026Updated last month
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- TurboServe: Serving Streaming Video Generation Efficiently and Economicallyβ44Jul 12, 2026Updated last month
- [EuroSys'25] Mist: Efficient Distributed Training of Large Language Models via Memory-Parallelism Co-Optimizationβ24Apr 13, 2026Updated 4 months ago
- β17Jun 15, 2026Updated 2 months ago
- Agent-friendly GPU profile-query CLIβ119Aug 21, 2026Updated last week
- Artifact Evaluation for SpecFS [FAST'26]β34Dec 28, 2025Updated 8 months ago
- An agent for CUDA compute-communication kernel co-designβ36May 7, 2026Updated 3 months ago
- High Performance Sorting Based Distributed memory K-mer counterβ15Dec 8, 2025Updated 8 months ago
- Dataflow-Oriented Reinforcement Learning for (Multi-)Agentic LLMsβ102Jul 26, 2026Updated last month
- Code repository for the SOSP'25 paper DCP: Addressing Input Dynamism In Long-Context Training via Dynamic Context Parallelism.β22Nov 28, 2025Updated 9 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A throughput-oriented high-performance serving framework for LLMsβ974Mar 29, 2026Updated 5 months ago
- β413Updated this week
- Winner π (Agent-only) MLSys 2026 - FlashInfer AI Kernel Generation Contest for the DeepSeek Sparse Attention (DSA) track with an averageβ¦β157Aug 21, 2026Updated last week
- β882Jun 2, 2026Updated 2 months ago
- ζ΅ζ±ε€§ε¦ζδ½η³»η»θ―Ύη¨δ»εΊοΌζζ‘£β18Jan 14, 2026Updated 7 months ago
- Translate C macros and condtional compilation to Rustβ27Aug 2, 2026Updated 3 weeks ago
- vLLM Daily Summarization of Merged PRsβ53Updated this week
- PaperHelper: Knowledge-Based LLM QA Paper Reading Assistant with Reliable Referencesβ22Jun 13, 2024Updated 2 years ago
- β50Jul 27, 2026Updated last month
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- β25Jan 18, 2026Updated 7 months ago
- A benchmarking framework for on-device AIβ20Aug 20, 2026Updated last week
- Foundry materializes CUDA graphs along with its execution context to disk to support fast cold start of serving engines.β57Updated this week
- β255Updated this week
- [ICLR'25] Fast Inference of MoE Models with CPU-GPU Orchestrationβ267Nov 18, 2024Updated last year
- mKernel: fast multi-node, multi-GPU fused kernelsβ270Updated this week
- β207Updated this week