LLM-powered Python
☆16Apr 8, 2026Updated 3 months ago
Alternatives and similar repositories for Nerif
Users that are interested in Nerif are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- PyTorch library for cost-effective, fast and easy serving of MoE models.☆321Updated this week
- 合肥工业大学 校园网自动登陆脚本☆11Mar 17, 2017Updated 9 years ago
- AI model training on heterogeneous, geo-distributed resources☆46Nov 24, 2025Updated 8 months ago
- ☆29Aug 14, 2024Updated last year
- 中国科学技术大学计算机学院课程资源备份,最新的请查看--->☆17Feb 23, 2019Updated 7 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Triton to TVM transpiler.☆24Oct 14, 2024Updated last year
- ☆26Sep 1, 2020Updated 5 years ago
- Rodinia benchmark☆24Jul 5, 2024Updated 2 years ago
- Serverless LLM Serving for Everyone.☆693May 4, 2026Updated 2 months ago
- This is my personal website☆15Jun 5, 2026Updated last month
- Clang-based translator for OP2☆12Jul 17, 2022Updated 4 years ago
- 合工大系列软件(刷试题库,刷网选,抢课,刷慕课,刷NIEI等)☆23Nov 1, 2017Updated 8 years ago
- Fantasy Ptrace☆23Mar 14, 2018Updated 8 years ago
- Optimize tensor program fast with Felix, a gradient descent autotuner.☆33Mar 5, 2026Updated 4 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Standalone Web IDE☆39Dec 31, 2015Updated 10 years ago
- DeepGEMM: clean and efficient FP8 GEMM kernels with fine-grained scaling☆32Updated this week
- ☆27Oct 1, 2025Updated 9 months ago
- byocc.cc☆22Jun 5, 2026Updated last month
- FlashSparse significantly reduces the computation redundancy for unstructured sparsity (for SpMM and SDDMM) on Tensor Cores through a Swa…☆39Oct 5, 2025Updated 9 months ago
- WaferLLM: Large Language Model Inference at Wafer Scale☆112Jun 12, 2026Updated last month
- A lightweight, user-friendly data-plane for LLM training.☆40Sep 10, 2025Updated 10 months ago
- RL-Scope: Cross-Stack Profiling for Deep Reinforcement Learning Workloads☆48Apr 7, 2021Updated 5 years ago
- Codebase for "Uni[MASK]: Unified Inference in Sequential Decision Problems"☆57Jul 3, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- A browser extension that allows users to search bookmarked tabs and browsing history within the browser.☆11Feb 24, 2024Updated 2 years ago
- An efficient concurrent graph processing system☆46Oct 27, 2021Updated 4 years ago
- build frida or hluda in docker☆13Dec 30, 2021Updated 4 years ago
- ☆26Aug 11, 2023Updated 2 years ago
- Demo app with Loguru logging, async middleware to generate X-request-Id. Works with Gunicorn or Uvicorn, and is safe to use with async/th…☆10Feb 2, 2022Updated 4 years ago
- The Web Metadata Extraction Toolkit is designed to streamline the process of extracting, cleaning, and analyzing metadata from websites. …☆18Jul 8, 2024Updated 2 years ago
- From Dataset Labeling, Entity Extraction to production Knowledge Graph Deployment: The Power of NLP and LLMs Combined.☆12Jun 19, 2026Updated last month
- System for automated integration of deep learning backends.☆47Aug 15, 2022Updated 3 years ago
- NAACL '24 (Best Demo Paper RunnerUp) / MlSys @ NeurIPS '23 - RedCoast: A Lightweight Tool to Automate Distributed Training and Inference☆69Dec 9, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Lecture notes of Probability Theory.☆51Jun 20, 2018Updated 8 years ago
- A tool to plot the memory and cpu usage.☆14May 9, 2024Updated 2 years ago
- vLLM Router☆56Mar 11, 2024Updated 2 years ago
- 🪵 Troncos - Collection of logging, tracing and profiling tools☆13Updated this week
- ☆11Mar 2, 2023Updated 3 years ago
- Pie: Programmable LLM Serving☆188Updated this week
- GPU Code optimizer for stencil computations. Refer to our IPDPS'19 paper for more details☆25Sep 27, 2019Updated 6 years ago