☆53Jan 18, 2024Updated 2 years ago
Alternatives and similar repositories for barrel-rec-pytorch
Users that are interested in barrel-rec-pytorch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆19Dec 4, 2025Updated 9 months ago
- Parallel Associative Scan for Language Models☆18Jan 8, 2024Updated 2 years ago
- Efficient PScan implementation in PyTorch☆17Jan 2, 2024Updated 2 years ago
- Official implementation for "Let Offline RL Flow: Training Conservative Agents in the Latent Space of Normalizing Flows", NeurIPS 2022, O…☆12Jan 31, 2023Updated 3 years ago
- Trace LLM calls (and others) and visualize them in WandB, as interactive SVG or using a streaming local webapp☆14Feb 18, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- NanoGPT-speedrunning for the poor T4 enjoyers☆72Apr 22, 2025Updated last year
- ☆14May 4, 2026Updated 4 months ago
- Improving Neural Text Generation with Reinforcement Learning☆23Jan 13, 2021Updated 5 years ago
- Simplex Random Feature attention, in PyTorch☆76Oct 10, 2023Updated 2 years ago
- Pytorch Implementation of Residual Multiplicative Filter Networks, NeurIPS 2022☆22Nov 17, 2022Updated 3 years ago
- Official implementation for "Q-Ensemble for Offline RL: Don't Scale the Ensemble, Scale the Batch Size", NeurIPS 2022, Offline RL Worksho…☆21Feb 27, 2023Updated 3 years ago
- ☆13Aug 7, 2021Updated 5 years ago
- Simple replication of [ColBERT-v1](https://arxiv.org/abs/2004.12832).☆83Mar 18, 2024Updated 2 years ago
- Simple GRPO scripts and configurations.☆59Feb 6, 2025Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Official implementation of the paper "Linear Transformers with Learnable Kernel Functions are Better In-Context Models"☆168Jan 16, 2025Updated last year
- A collection of various custom nodes for ComfyUI (Work in progress)☆15Jun 9, 2025Updated last year
- utilities to facilitate working with codebases that don't ascribe to normal package management paradigms, e.g. ML research code that can …☆13Nov 26, 2022Updated 3 years ago
- ☆40Jan 5, 2024Updated 2 years ago
- minGPT in JAX☆49Jan 10, 2022Updated 4 years ago
- ESGD-M is a stochastic non-convex second order optimizer, suitable for training deep learning models, for PyTorch.☆57Sep 18, 2022Updated 3 years ago
- ☆34May 14, 2025Updated last year
- supporting pytorch FSDP for optimizers☆84Dec 8, 2024Updated last year
- Fast approximate inference on a single GPU with sparsity aware offloading☆38Jan 4, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Minimal (400 LOC) implementation Maximum (multi-node, FSDP) GPT training☆132Apr 17, 2024Updated 2 years ago
- Solve puzzles. Learn CUDA.☆63Dec 13, 2023Updated 2 years ago
- Author implementation of Monte Carlo Augmented Actor Critic in PyTorch☆18Oct 24, 2022Updated 3 years ago
- Optimizing Causal LMs through GRPO with weighted reward functions and automated hyperparameter tuning using Optuna☆60Oct 18, 2025Updated 10 months ago
- A discrete sequential VAE☆42Apr 22, 2020Updated 6 years ago
- Annotated version of the Mamba paper☆502Feb 27, 2024Updated 2 years ago
- 🚀💼☆18Updated this week
- Примеры пропозалов для подачи заявки в Open.TLab☆27Dec 15, 2022Updated 3 years ago
- Diffusion models in PyTorch☆134Aug 28, 2026Updated last week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Accelerated First Order Parallel Associative Scan☆199Jan 7, 2026Updated 8 months ago
- Adaptive Subgoal Search☆20Apr 3, 2023Updated 3 years ago
- ACCO: An optimization algorithm for sharded distributed LLM training.☆13May 22, 2025Updated last year
- ☆33Sep 22, 2025Updated 11 months ago
- ☆16May 8, 2024Updated 2 years ago
- DashAttention: Differentiable and Adaptive Sparse Hierarchical Attention☆23May 25, 2026Updated 3 months ago
- A minimal home grid world environment to evaluate language understanding in interactive agents.☆24Sep 6, 2023Updated 3 years ago