Simple and efficient RevNet-Library for PyTorch with XLA and DeepSpeed support and parameter offload
☆132Aug 6, 2022Updated 4 years ago
Alternatives and similar repositories for revlib
Users that are interested in revlib are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A case study of efficient training of large language models using commodity hardware.☆67Aug 4, 2022Updated 4 years ago
- HomebrewNLP in JAX flavour for maintable TPU-Training☆50Jan 20, 2024Updated 2 years ago
- Memory-efficient transformer. Work in progress.☆19Sep 17, 2022Updated 3 years ago
- Drop-in replacement for any ResNet with a significantly reduced memory footprint and better representation capabilities☆206Apr 24, 2024Updated 2 years ago
- Contrastive Language-Image Pretraining☆147Sep 6, 2022Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A scalable Dreamer implementation in JAX☆10May 22, 2022Updated 4 years ago
- PyTorch Framework for Developing Memory Efficient Deep Invertible Networks☆257Mar 12, 2026Updated 4 months ago
- Implementation of the 😇 Attention layer from the paper, Scaling Local Self-Attention For Parameter Efficient Visual Backbones☆199Mar 24, 2021Updated 5 years ago
- Transformers with doubly stochastic attention☆55Sep 14, 2022Updated 3 years ago
- Training a model similar to OpenAI DALL-E with volunteers from all over the Internet using hivemind and dalle-pytorch (NeurIPS 2021 demo)☆27May 29, 2023Updated 3 years ago
- The Intermediate Goal of the project is to train a GPT like architecture to learn to summarise reddit posts from human preferences, as th…☆12Jul 14, 2021Updated 5 years ago
- ☆33Mar 1, 2023Updated 3 years ago
- Fast Discounted Cumulative Sums in PyTorch☆99Aug 28, 2021Updated 4 years ago
- ESGD-M is a stochastic non-convex second order optimizer, suitable for training deep learning models, for PyTorch.☆57Sep 18, 2022Updated 3 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Implementation of Token Shift GPT - An autoregressive model that solely relies on shifting the sequence space for mixing☆49Jan 27, 2022Updated 4 years ago
- ☆15Sep 15, 2022Updated 3 years ago
- ☆27May 9, 2022Updated 4 years ago
- Implicit MLE: Backpropagating Through Discrete Exponential Family Distributions☆261Oct 29, 2023Updated 2 years ago
- ☆15Feb 28, 2022Updated 4 years ago
- ☆774Jan 27, 2024Updated 2 years ago
- An attempt at the implementation of Glom, Geoffrey Hinton's new idea that integrates concepts from neural fields, top-down-bottom-up proc…☆197Mar 27, 2021Updated 5 years ago
- An open source implementation of CLIP.☆33Nov 7, 2022Updated 3 years ago
- CLASP - Contrastive Language-Aminoacid Sequence Pretraining☆142Sep 17, 2021Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆14Dec 28, 2021Updated 4 years ago
- Stores the MHub models dockerfiles and scripts.☆11Apr 30, 2026Updated 3 months ago
- To be a next-generation DL-based phenotype prediction from genome mutations.☆19May 17, 2021Updated 5 years ago
- ☆65Nov 4, 2021Updated 4 years ago
- Swarm training framework using Haiku + JAX + Ray for layer parallel transformer language models on unreliable, heterogeneous nodes☆241May 12, 2023Updated 3 years ago
- ☆161Jun 13, 2022Updated 4 years ago
- ☆32Sep 24, 2019Updated 6 years ago
- This Is a Multiple format Video Steganography project.☆10Feb 4, 2024Updated 2 years ago
- ARC Community Project☆23Aug 2, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Library for 8-bit optimizers and quantization routines.☆779Aug 18, 2022Updated 3 years ago
- This extension contains only a module with some tools to install PyTorch inside Slicer, using the best possible version.☆38Jul 30, 2026Updated last week
- Minimal A2C/A3C example of an LSTM-based meta-learner.☆13Feb 2, 2021Updated 5 years ago
- clip retrieval benchmark☆17May 4, 2022Updated 4 years ago
- ☆39Oct 3, 2022Updated 3 years ago
- ☆18Oct 3, 2024Updated last year
- A dashboard for exploring timm learning rate schedulers☆20Nov 22, 2024Updated last year