Open-source calculator for LLM system requirements.
☆183Dec 18, 2024Updated last year
Alternatives and similar repositories for LLM-Tools
Users that are interested in LLM-Tools are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A simple program scheduler for your code on different devices.☆12Mar 8, 2026Updated 6 months ago
- A calculator to estimate the memory footprint, capacity, and latency on VMware Private AI with NVIDIA.☆40Aug 5, 2025Updated last year
- ☆101Apr 2, 2025Updated last year
- torch.compile artifacts for common deep learning models, can be used as a learning resource for torch.compile☆19Dec 22, 2023Updated 2 years ago
- best llms in russian☆63May 23, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ACM SoCC 2019, "Coupling Decentralized Key-Value Stores with Erasure Coding"☆15May 22, 2021Updated 5 years ago
- MMLU eval for RU/EN☆16Jul 31, 2023Updated 3 years ago
- Beyond KV Caching: Shared Attention for Efficient LLMs☆20Jul 19, 2024Updated 2 years ago
- Project Mitosis Introduction☆19Nov 13, 2022Updated 3 years ago
- ☆23May 29, 2023Updated 3 years ago
- MERA (Multimodal Evaluation for Russian-language Architectures) is a new open benchmark for the Russian language for evaluating fundament…☆63Oct 7, 2024Updated last year
- Evolutionary generation of efficient GPU kernels☆35Jul 8, 2026Updated 2 months ago
- The source code for running LLMs on the AAAR-1.0 benchmark.☆20Apr 5, 2025Updated last year
- Codes for paper "Stylized Story Generation with Style-Guided Planning"☆12May 9, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆22Oct 30, 2024Updated last year
- Effective LLM Alignment Toolkit☆154Jun 25, 2025Updated last year
- InfiniStore: an elastic serverless cloud storage system (VLDB'23)☆24May 5, 2023Updated 3 years ago
- Calculate token/s & GPU memory requirement for any LLM. Supports llama.cpp/ggml/bnb/QLoRA quantization☆1,401Dec 3, 2024Updated last year
- Codebase for character-centric story understanding☆14Jan 20, 2022Updated 4 years ago
- Estimate Your LLM's Token Toll Across Various Platforms and Configurations☆39Nov 9, 2025Updated 10 months ago
- Yet Another RussNet☆15Oct 21, 2021Updated 4 years ago
- LLM Inference analyzer for different hardware platforms☆125Jul 30, 2026Updated last month
- A toy implementation about Program Dependence Graph using LLVM☆13Sep 27, 2023Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Russian paraphrasers. Generate paraphrases with mt5, gpt2, etc.☆57May 27, 2023Updated 3 years ago
- ☆13Jun 4, 2025Updated last year
- Efficient and easy multi-instance LLM serving☆562Mar 12, 2026Updated 6 months ago
- CHAE: Fine-Grained Controllable Story Generation with Characters, Actions and Emotions☆11Jan 31, 2023Updated 3 years ago
- Materials that help you to navigate in the world of NLP☆21Sep 10, 2025Updated last year
- 📚 LaTeX templates and tools for creating beautiful, structured documents 📝☆14Oct 24, 2025Updated 10 months ago
- Code for ACL2023 paper: Pre-Training to Learn in Context☆106Jul 26, 2024Updated 2 years ago
- DeepSeek-V3/R1 inference performance simulator☆196Mar 27, 2025Updated last year
- Source Code for Partial Interference☆10Dec 17, 2022Updated 3 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- An Agile Chisel-Based SoC Design Framework☆25Dec 29, 2021Updated 4 years ago
- (Obsoleted) A speech signal processing library designed for CVE of Rocaloid Project.☆15Dec 7, 2013Updated 12 years ago
- A Distributed Analysis and Benchmarking Framework for Apache OpenWhisk Serverless Platform☆12Dec 11, 2018Updated 7 years ago
- Optimizing Review Generation Through Prompt Generation☆17Apr 15, 2024Updated 2 years ago
- Official Code for NAACL 2022 paper: "Persona-Guided Planning for Controlling the Protagonist's Persona in Story Generation"☆16Sep 1, 2022Updated 4 years ago
- ☆14Jul 13, 2025Updated last year
- Very concise example of integrated gradients (a method to reveal areas of attention in input images)☆10Jun 17, 2019Updated 7 years ago