☆158Mar 9, 2023Updated 3 years ago
Alternatives and similar repositories for fast_gpt2
Users that are interested in fast_gpt2 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ANE accelerated embedding models!☆20Dec 11, 2024Updated last year
- ☆15Oct 24, 2023Updated 2 years ago
- Standalone commandline CLI tool for compiling Triton kernels☆20Sep 13, 2024Updated 2 years ago
- ☆568Feb 5, 2024Updated 2 years ago
- A Demo server serving Bert through ONNX with GPU written in Rust with <3☆42Jul 30, 2021Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Bleeding edge low level Rust binding for GGML☆18Jun 26, 2024Updated 2 years ago
- Directly Connecting Python to LLMs via Strongly-Typed Functions, Dataclasses, Interfaces & Generic Types☆401Mar 3, 2025Updated last year
- ☆28Aug 10, 2023Updated 3 years ago
- ☆24May 23, 2023Updated 3 years ago
- ☆133Nov 24, 2023Updated 2 years ago
- Find out why your CoreML model isn't running on the Neural Engine!☆30Jun 18, 2024Updated 2 years ago
- LLaMa 7b with CUDA acceleration implemented in rust. Minimal GPU memory needed!☆112Jul 27, 2023Updated 3 years ago
- Implementation of Token Shift GPT - An autoregressive model that solely relies on shifting the sequence space for mixing☆49Jan 27, 2022Updated 4 years ago
- Your one stop CLI for ONNX model analysis.☆48Nov 13, 2022Updated 3 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- See https://youtube-dl.org/☆10Oct 24, 2020Updated 5 years ago
- An example of using Torch rust bindings to serve trained machine learning models via Actix Web☆17Aug 15, 2021Updated 5 years ago
- lightweight composable proxy made for chatgpt plugins☆34May 24, 2023Updated 3 years ago
- Barebones Rust EVM Implementation☆12Feb 9, 2022Updated 4 years ago
- Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the PaLM architecture. Basically ChatGPT but with PaLM☆7,857Sep 20, 2026Updated 2 weeks ago
- Simulate VRGDAs in your browser. Perfect for prototyping and use with tabletop games.☆29Oct 6, 2022Updated 4 years ago
- ☆12Jan 30, 2023Updated 3 years ago
- ☆131Apr 23, 2023Updated 3 years ago
- Using short models to classify long texts☆21Mar 8, 2023Updated 3 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Exporting youtube videos using whisper☆17Sep 27, 2022Updated 4 years ago
- ☆13Jul 12, 2024Updated 2 years ago
- A helper that sends logs over TCP, for programs without stdout, such as TUIs.☆24Feb 14, 2026Updated 7 months ago
- ☆15Sep 15, 2023Updated 3 years ago
- ☆72Mar 15, 2024Updated 2 years ago
- Run Reth+Prometheus+Grafana locally via docker-compose☆71Jun 27, 2023Updated 3 years ago
- A single-binary, GPU-accelerated LLM server (HTTP and WebSocket API) written in Rust☆79Jan 14, 2024Updated 2 years ago
- a GGUF file parser☆21Aug 31, 2026Updated last month
- ☆37May 31, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Slow burner for protocol assets.☆33Mar 21, 2023Updated 3 years ago
- Command Line Interface for Hugging Face Inference Endpoints☆65Apr 10, 2024Updated 2 years ago
- Adapter / facade for language models (OpenAI, Anthropic, Cohere, local transformers, etc)☆20Sep 21, 2023Updated 3 years ago
- Kernl lets you run PyTorch transformer models several times faster on GPU with a single line of code, and is designed to be easily hackab…☆1,584Jan 28, 2026Updated 8 months ago
- ☆55Mar 7, 2023Updated 3 years ago
- GPU accelerated client-side embeddings for vector search, RAG etc.☆64Dec 4, 2023Updated 2 years ago
- Evaluation tool for LLM QA chains☆1,104May 10, 2023Updated 3 years ago