A high-performance RDMA distributed file system for fast LLM Inference and GPU Training.
☆47Feb 20, 2026Updated 6 months ago
Alternatives and similar repositories for blackbird
Users that are interested in blackbird are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Automatically Analyze your Model Traces☆47Mar 16, 2026Updated 5 months ago
- VUA stands for 'VAST Undivided Attention'. It's a global KVCache storage solution optimizing LLM time to first token (TTFT) and GPU utili…☆38Mar 12, 2026Updated 5 months ago
- Highly Performant, Modular, Memory Safe and Production-ready Inference, Ingestion and Indexing built in Rust 🦀☆1,306Aug 12, 2026Updated 2 weeks ago
- Persistent Memory Test Suite☆14Apr 29, 2020Updated 6 years ago
- Prototyp MegaScale-Infer: Serving Mixture-of-Experts at Scale with Disaggregated Expert Parallelism☆34Apr 4, 2025Updated last year
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- This is the official Python version of CoreInfer: Accelerating Large Language Model Inference with Semantics-Inspired Adaptive Sparse Act…☆19Oct 25, 2024Updated last year
- ☆16Jun 4, 2024Updated 2 years ago
- Open-source transpiler for CUDA Tile (13.1) migration☆19Dec 9, 2025Updated 8 months ago
- Algorithms for approximate attention in LLMs☆22Apr 14, 2025Updated last year
- The official implementation of paper: SimLayerKV: A Simple Framework for Layer-Level KV Cache Reduction.☆55Oct 18, 2024Updated last year
- ☆82Updated this week
- Aggressive decode optimizations for Qwen3-0.6B on RTX 5090☆60Feb 25, 2026Updated 6 months ago
- Take tiles from an mbtiles file locally (eg vector tiles) and extracts them to ZXY structure in an S3 bucket or locally.☆11May 19, 2020Updated 6 years ago
- A prefill & decode disaggregated LLM serving framework with shared GPU memory and fine-grained compute isolation.☆131Dec 25, 2025Updated 8 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ReactJs app boilerplate using esbuild☆10Mar 22, 2024Updated 2 years ago
- ☆16Aug 20, 2024Updated 2 years ago
- Provides a tool to test CXL with Kernel and Qemu setup☆22Apr 10, 2026Updated 4 months ago
- Docker container for a jupyter notebook with a GeoTrellis environment.☆10Aug 15, 2017Updated 9 years ago
- ☆19Updated this week
- code poems☆10Feb 23, 2022Updated 4 years ago
- ☆11Mar 26, 2015Updated 11 years ago
- ☆20Mar 11, 2025Updated last year
- ☆13Jan 31, 2015Updated 11 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Decode nvme trace events☆18Jan 8, 2021Updated 5 years ago
- Docker configuration for CartoDB☆11Oct 24, 2016Updated 9 years ago
- GeoZarr extension for OpenLayers☆12Jun 27, 2024Updated 2 years ago
- GPU-accelerated LLM Training Simulator☆22Jun 26, 2025Updated last year
- Analysis library to create data views from queries☆17Jan 5, 2023Updated 3 years ago
- Call-graph generator.☆23Jun 13, 2018Updated 8 years ago
- FLA but cuTile☆27Apr 17, 2026Updated 4 months ago
- Source code of paper ''KVSharer: Efficient Inference via Layer-Wise Dissimilar KV Cache Sharing''☆31Oct 24, 2024Updated last year
- Some Docker images to play with.☆13Jun 16, 2016Updated 10 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆28Jul 29, 2025Updated last year
- Model explanation provides the ability to interpret the effect of the predictors on the composition of an individual score.☆13Jan 21, 2021Updated 5 years ago
- The Auditree common fetchers, checks and harvest reports library.☆22Nov 9, 2023Updated 2 years ago
- This is a repo listing papers/blogs/news related to CXL. Let's take the leap to Next-Gen memory system with the awesome CXL☆21Jul 10, 2024Updated 2 years ago
- clustering algorithm implementation☆13May 13, 2026Updated 3 months ago
- A Collection for Distributed Reinforcement Learning Papers☆18Sep 24, 2025Updated 11 months ago
- ☆18Dec 9, 2025Updated 8 months ago