Fast hierarchical embedding cache for recommenders
☆22Jul 5, 2026Updated last month
Alternatives and similar repositories for nv-embedding-cache
Users that are interested in nv-embedding-cache are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆17Jul 2, 2012Updated 14 years ago
- ☆15Jan 29, 2025Updated last year
- ☆57Oct 17, 2023Updated 2 years ago
- HierarchicalKV is a part of NVIDIA Merlin and provides hierarchical key-value storage to meet RecSys requirements. The key capability of…☆208May 22, 2026Updated 2 months ago
- ☆12May 12, 2023Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- CUDA Embedding Lookup Kernel Library☆49Jun 26, 2026Updated last month
- Tree-sitter organization info☆17Mar 4, 2026Updated 5 months ago
- Parallelizing Google's PageRank algorithm in C++ with CUDA framework on GPU. Conducted some experiments, tried new ideas. Final Course Pr…☆10May 20, 2021Updated 5 years ago
- Persistent Memory Development Kit☆18Updated this week
- benchmark for linux server☆13Nov 6, 2016Updated 9 years ago
- An MCP server that connects to your React Native application debugger☆32Mar 28, 2025Updated last year
- Python package for collecting ACS and geospatial data from the Census API☆24Feb 19, 2025Updated last year
- Telegram-style circular reveal theme transitions for React Native (Expo Module). iOS: CAShapeLayer mask animation. Android: PixelCopy + P…☆31Aug 1, 2026Updated last week
- Examples for Recommenders - easy to train and deploy on accelerated infrastructure.☆296Updated this week
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆18Apr 8, 2022Updated 4 years ago
- Simple HTTP serving for PyTorch 🚀☆10Oct 15, 2020Updated 5 years ago
- Repository containing pruned models and related information☆38Mar 17, 2021Updated 5 years ago
- pytorch code examples for measuring the performance of collective communication calls in AI workloads☆21Sep 18, 2025Updated 10 months ago
- Easily create a mirror of crates.io (crate downloads only, not the website)☆13Aug 26, 2018Updated 7 years ago
- ucas hpc course code☆15May 24, 2023Updated 3 years ago
- A curated list of Awesome Knative resources.☆21Feb 3, 2021Updated 5 years ago
- hipDF - GPU DataFrame Library☆20Jul 31, 2026Updated last week
- a collection of Tools, Command, Tips used by me☆17Jul 26, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Web Desktop App Using Integrated Extjs 4.2 and Node.js☆14Apr 6, 2016Updated 10 years ago
- A hopefully fast symbol table (string <=> integer sequence number)☆17Apr 14, 2026Updated 3 months ago
- K8s DRA driver for AMD GPUs☆28Updated this week
- A very simple tool to rewrite parameters such as attributes and constants for OPs in ONNX models. Simple Attribute and Constant Modifier …☆15Feb 6, 2026Updated 6 months ago
- ☆16Dec 21, 2019Updated 6 years ago
- Promise-based wrapper for worker_threads☆18Jul 12, 2023Updated 3 years ago
- Papers and Codes about Deep Metric Learning/Deep Embedding☆35Feb 17, 2020Updated 6 years ago
- Matrix-Vector Multiplication Using Shared and Coalesced Memory Access☆16Apr 9, 2013Updated 13 years ago
- 本人的文章和笔记,充当懒人 Blog 来用。☆12May 31, 2026Updated 2 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- A global file-based mutex lock using Google Cloud Storage☆29Jun 29, 2018Updated 8 years ago
- My portfolio built with Next.js, Tailwind, Giscus, Umami, Upstash, MDX, Content Collections, Bun, and Vercel.☆24Updated this week
- Implementation for Trained Ternary Network.☆108Jan 13, 2017Updated 9 years ago
- A New Format for SIMD-accelerated SpMV☆22Apr 4, 2022Updated 4 years ago
- 演示 vllm 对中文大语言模型的神奇效果☆31Nov 4, 2023Updated 2 years ago
- A SPI Master IP written in verilog which is then used to output characters entered on a keypad to a serial LCD screen☆20Dec 5, 2014Updated 11 years ago
- A Synchronization-Free Algorithm for Parallel Sparse Triangular Solves (SpTRSV)☆23Feb 14, 2020Updated 6 years ago