Implementation and explorations into PopuLoRA, Co-Evolving LLM Populations for Reasoning Self-Play
☆18Sep 29, 2026Updated this week
Alternatives and similar repositories for populora
Users that are interested in populora are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Implementation and explorations into DiscoRL, Discovering state-of-the-art reinforcement learning algorithms, David Silver's last work at…☆22Jun 13, 2026Updated 3 months ago
- Implementation of Recurrent Independent Mechanisms in Pytorch☆27Apr 6, 2026Updated 5 months ago
- JAX compilation of RDDL description files, and a differentiable planner in JAX.☆18Jun 28, 2026Updated 3 months ago
- Code for Fast-weight Product Key Memory (FwPKM)☆26Mar 18, 2026Updated 6 months ago
- Learning diverse options through the Laplacian representation.☆23Jan 5, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Implementation of a transformer for reinforcement learning using `x-transformers`☆72Sep 25, 2025Updated last year
- A framework for evaluating LLMs in Atari games☆15Apr 21, 2025Updated last year
- Behavior Injection: Preparing Language Models for Reinforcement Learning (NeurIPS 2025)☆17Jul 1, 2025Updated last year
- Implementation of UltraMem, improved Product Key Memory design, from Bytedance AI labs☆28Nov 4, 2025Updated 11 months ago
- Embedding and readout for simple multi-categorical and gaussian continuous☆20Jul 5, 2026Updated 3 months ago
- [NeurIPS 2023] Official code release accompanying the paper "NetHack is Hard to Hack" (Piterbarg, Pinto, Fergus)☆14Oct 30, 2023Updated 2 years ago
- Causal Attention with Lookahead Keys☆28Sep 26, 2025Updated last year
- Implementation of 2-simplicial attention proposed by Clift et al. (2019) and the recent attempt to make practical in Fast and Simplex, Ro…☆49Sep 2, 2025Updated last year
- Code for "Meta Learning Backpropagation And Improving It" @ NeurIPS 2021 https://arxiv.org/abs/2012.14905☆33Jan 9, 2022Updated 4 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Implementation of the model architecture for SRT-H☆30Updated this week
- Implementation of Poly-attention, a higher-order self-attention proposed by Chakrabarti et al. of Columbia☆55Aug 19, 2026Updated last month
- ☆11Aug 7, 2024Updated 2 years ago
- A technical exploration comparing standard deep RL (PPO) against biologically plausible learning rules on a custom Pong environment. Ever…☆29May 19, 2026Updated 4 months ago
- [CHIL 2024] Interpretation of Intracardiac Electrograms Through Textual Representations☆12Sep 4, 2024Updated 2 years ago
- Efficiently discovering algorithms via LLMs with evolutionary search and reinforcement learning.☆17Apr 22, 2025Updated last year
- ☆28Jul 14, 2024Updated 2 years ago
- ☆25Dec 11, 2024Updated last year
- The implementation of "The Kanerva Machine" with Pytorch and Pyro☆12Jun 14, 2018Updated 8 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Property-based testing for OCaml, built on Hypothesis☆24Updated this week
- Dataset for the paper: "A multi-task semi-supervised framework for Text2Graph & Graph2Text"☆25Feb 19, 2022Updated 4 years ago
- Evaluating Durability: Benchmark Insights into Multimodal Watermarking☆12Jun 7, 2024Updated 2 years ago
- Some utility functions to help myself (and perhaps others) go faster with ML/AI work☆54Updated this week
- Explorations into some of the approaches advocated by Yann LeCun, and just a more wholistic architecture (JEPA) in general☆126Sep 21, 2026Updated 2 weeks ago
- o lua m'lua☆12May 27, 2024Updated 2 years ago
- PyTorch implementation for the Deep Symbolic Simplification Without Human Knowledge☆14Feb 25, 2021Updated 5 years ago
- Symbol-Equivariant Recurrent Reasoning Model☆18Mar 4, 2026Updated 7 months ago
- Simple and Ideal Circuit Simulation☆13Dec 4, 2017Updated 8 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Convert GitHub PRs into Harbor tasks☆89Jul 13, 2026Updated 2 months ago
- ☆10Feb 16, 2025Updated last year
- Official repository for the paper "Automating Continual Learning"☆21Jun 11, 2025Updated last year
- ☆15Feb 8, 2023Updated 3 years ago
- ☆12Nov 21, 2023Updated 2 years ago
- Portfolio Optimization Book☆13Dec 16, 2024Updated last year
- Implementation of the Extended Exposure Fusion (EEF) proposed in [WACV '20 paper reference]☆12Dec 21, 2019Updated 6 years ago