Because it's there.
☆16Sep 22, 2024Updated last year
Alternatives and similar repositories for cuda-substrings
Users that are interested in cuda-substrings are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Modal LLM LLama.cpp based model deployment as part of series of Model as a Service (MaaS)☆17Mar 23, 2026Updated 3 months ago
- Server bots for Poe☆18Nov 17, 2025Updated 8 months ago
- A collection of optimized ComfyUI-based cloud inference endpoints, built on ComfyDeploy and Modal☆16Nov 5, 2024Updated last year
- A webhook that integrates the W&B model registry with Modal Labs☆15Dec 24, 2023Updated 2 years ago
- Orchestrate Modal and OpenAI workloads with Dagster☆13Dec 11, 2024Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Privacy-preserving k-means clustering on data owned by multiple parties☆14May 10, 2016Updated 10 years ago
- various experiments for scaling inference time compute with small reasoning models☆17Jan 16, 2025Updated last year
- Exploration into the Scaling Value Iteration Networks paper, from Schmidhuber's group☆37Sep 23, 2024Updated last year
- run dreambooth training on modal☆22Nov 21, 2022Updated 3 years ago
- Voice data <= 10 mins can also be used to train a good VC model!☆12Dec 5, 2023Updated 2 years ago
- Seldon Core Operator for Kubernetes☆13Nov 5, 2019Updated 6 years ago
- A collection of tools for your LLMs that run on Modal☆25Feb 28, 2025Updated last year
- This repository contains code for the paper "Better Estimation of the KL Divergence Between Language Models"☆19May 30, 2025Updated last year
- on-chain LLMs for the Internet Computer☆17Jan 14, 2026Updated 6 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Exercises & Notes from the LinkedIn Learning courses about Behavioral, Creational & Structural Design Patterns by Bethan Palmer☆18Jun 7, 2025Updated last year
- A tool for benchmarking LLMs on Modal☆56Aug 29, 2025Updated 10 months ago
- ☆18Apr 23, 2025Updated last year
- Deploy a FastHTML app in just a few lines of simple python code on Modal's serverless infra.☆26Aug 19, 2024Updated last year
- ☆21Apr 14, 2026Updated 3 months ago
- 🧹 FileRefine is a Rust-based CLI tool that renames files in a directory to remove unwanted or problematic characters from filenames.☆22Feb 1, 2026Updated 5 months ago
- A cog implementation of Nvidia's Triton server☆18Oct 23, 2024Updated last year
- Embedded systems starter for IOT practices☆12Jul 11, 2022Updated 4 years ago
- Terraform Azure Verified Resource Module for Cognitive Service☆14Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A Swift package for interacting with selenium and undetected-chromedriver through python by using PythonKit.☆13Jun 21, 2025Updated last year
- A Survey on Causal Generative Modeling (TMLR 2024)☆15Dec 18, 2024Updated last year
- A simple react app with a canvas to draw an image. Real time image-to-image inference using Stable Diffusion XL Turbo and Modal☆25Mar 8, 2024Updated 2 years ago
- Code for creating weather forecast data cubes with a Modal serverless pipeline☆25Oct 27, 2025Updated 8 months ago
- ☆21Updated this week
- ValTown MCP Server - Execute ValTown functions from AI assistants☆15Aug 12, 2025Updated 11 months ago
- Eden is building autonomous creative agents.☆30Updated this week
- ☆14Feb 24, 2023Updated 3 years ago
- Integrating SSE with NVIDIA Triton Inference Server using a Python backend and Zephyr model. There is very less documentation how to use …☆10May 29, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Triton implementation of GPT/LLAMA☆22Aug 28, 2024Updated last year
- High performance JSON middleware for Swift projects☆13Jul 4, 2026Updated 2 weeks ago
- Well documented examples of running distributed training jobs on Modal☆29Updated this week
- Simple dockerized Rust/Axum based HTTP server for demo purposes.☆19Updated this week
- A speedy, flexible router for Rust.☆13Jul 8, 2026Updated last week
- ☆30Apr 6, 2026Updated 3 months ago
- Elixir: Train a Large Language Model on a Small GPU Cluster☆16Jun 8, 2023Updated 3 years ago