Contrastive Reinforcement Learning
☆68Sep 15, 2026Updated last week
Alternatives and similar repositories for contrastive-rl
Users that are interested in contrastive-rl are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Explorations into NEAT and some of its derivative research☆42Sep 7, 2026Updated 2 weeks ago
- Causal Attention with Lookahead Keys☆28Sep 26, 2025Updated 11 months ago
- Embedding and readout for simple multi-categorical and gaussian continuous☆20Jul 5, 2026Updated 2 months ago
- Explorations into the proposed Streaming Deep Reinforcement Learning, from University of Alberta☆32Sep 6, 2026Updated 2 weeks ago
- Implementation of various evolutionary algorithms, starting with evolutionary strategies☆51Sep 4, 2026Updated 2 weeks ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Implementation of ResFit, Residual Off-Policy RL for Finetuning Behavior Cloning Policies☆17Sep 29, 2025Updated 11 months ago
- Some utility functions to help myself (and perhaps others) go faster with ML/AI work☆54Updated this week
- Training framework for Large Behavioral Models☆28Sep 17, 2025Updated last year
- LocoFormer - Generalist Locomotion via Long-Context Adaptation☆124Sep 11, 2026Updated last week
- Implementation of Recurrent Independent Mechanisms in Pytorch☆27Apr 6, 2026Updated 5 months ago
- Implementation of Dex1B: Learning with 1B Demonstrations for Dexterous Manipulation, from Ye et al of UCSD☆30Jul 27, 2026Updated last month
- Implementation of rectified flow and some of its followup research / improvements in Pytorch☆486Sep 12, 2026Updated last week
- Implementation of AlphaGenome, Deepmind's updated genomic attention model☆101Mar 25, 2026Updated 5 months ago
- Efficiently discovering algorithms via LLMs with evolutionary search and reinforcement learning.☆17Apr 22, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- An implementation of PPO in Pytorch☆126Updated this week
- Unofficial implementation of Hippoformer, Integrating Hippocampus-inspired Spatial Memory with Transformers☆55Aug 4, 2026Updated last month
- Implementation of the MetaController proposed in "Emergent temporal abstractions in autoregressive models enable hierarchical reinforceme…☆106Aug 14, 2026Updated last month
- Unofficial implementation of Tiny Recursive Model (TRM), improvement to HRM from Sapient AI, by Alexia Jolicoeur-Martineau☆192Dec 23, 2025Updated 9 months ago
- Implementation of Fast Weight Attention☆35Updated this week
- Attempt to make multiple residual streams from Bytedance's Hyper-Connections paper accessible to the public☆188May 13, 2026Updated 4 months ago
- Implementation of RL-100, Performant Robotic Manipulation with Real-World Reinforcement Learning☆66Nov 26, 2025Updated 9 months ago
- Implementation and explorations into Blackbox Gradient Sensing (BGS), an evolutionary strategies approach proposed in a Google Deepmind p…☆20Apr 17, 2026Updated 5 months ago
- Efficiently discovering algorithms via LLMs with evolutionary search and reinforcement learning.☆144Nov 18, 2025Updated 10 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Implementation of ViLLA-X, Enhancing Latent Action Modeling in Vision-Language-Action Models☆23Aug 27, 2025Updated last year
- Explorations into the proposed SDFT, Self-Distillation Enables Continual Learning, from Shenfeld et al. of MIT☆33Feb 6, 2026Updated 7 months ago
- Implementation of Soft Actor Critic and some of its improvements in Pytorch☆72Updated this week
- Exploration into the Scaling Value Iteration Networks paper, from Schmidhuber's group☆37Sep 23, 2024Updated last year
- Explorations into some of the approaches advocated by Yann LeCun, and just a more wholistic architecture (JEPA) in general☆126Updated this week
- Implementation of the new SOTA for model based RL, from the paper "Improving Transformer World Models for Data-Efficient RL", in Pytorch☆156May 2, 2025Updated last year
- Synchronized Curriculum Learning for RL Agents☆124Jul 31, 2026Updated last month
- Implementation of Amplify, Actionless Motion Priors for Robot Learning from Videos☆30Sep 26, 2025Updated 11 months ago
- Implementation of the proposed DeepCrossAttention by Heddes et al at Google research, in Pytorch☆107Apr 3, 2026Updated 5 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [ICML 2026] Official Code for Rectified LpJEPA: Joint-Embedding Predictive Architectures with Sparse and Maximum-Entropy Representations☆85Feb 15, 2026Updated 7 months ago
- Testing various improvements to Ranger21 for 2022☆19Nov 6, 2024Updated last year
- BM-MAE: Multimodal Masked Autoencoder Pre-training for 3D MRI-based Brain Tumor Analysis with Missing Modalities☆34Aug 24, 2025Updated last year
- Official repository for the paper "Random Shuffle Transformer for Image Restoration".☆17Jan 9, 2024Updated 2 years ago
- This repository collects lecture slides, assignments (CAs), code notebooks, reports, and reference papers used in the "Deep Generative Mo…☆21Feb 14, 2026Updated 7 months ago
- 33B Chinese LLM, DPO QLORA, 100K context, AirLLM 70B inference with single 4GB GPU☆14May 5, 2024Updated 2 years ago
- vLLM adapter for a TGIS-compatible gRPC server.☆58Updated this week