Implementation of UltraMem, improved Product Key Memory design, from Bytedance AI labs
☆28Nov 4, 2025Updated 8 months ago
Alternatives and similar repositories for ultra-mem
Users that are interested in ultra-mem are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Implementation of Strassen attention, from Kozachinskiy et al. of National Center of AI in Chile☆41Jul 8, 2025Updated last year
- Implementation of 2-simplicial attention proposed by Clift et al. (2019) and the recent attempt to make practical in Fast and Simplex, Ro…☆49Sep 2, 2025Updated 10 months ago
- Implementation of Recurrent Independent Mechanisms in Pytorch☆27Apr 6, 2026Updated 3 months ago
- Unofficial implementation of Hippoformer, Integrating Hippocampus-inspired Spatial Memory with Transformers☆53Apr 28, 2026Updated 2 months ago
- Implementation and explorations into PopuLoRA, Co-Evolving LLM Populations for Reasoning Self-Play☆15Jun 3, 2026Updated last month
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Causal Attention with Lookahead Keys☆28Sep 26, 2025Updated 10 months ago
- Implementation of various evolutionary algorithms, starting with evolutionary strategies☆51May 10, 2026Updated 2 months ago
- Unofficial implementation of Tiny Recursive Model (TRM), improvement to HRM from Sapient AI, by Alexia Jolicoeur-Martineau☆191Dec 23, 2025Updated 7 months ago
- Official Code Repositiry for "RaDeR: Reasoning-aware Dense Retrieval Models" accepted at Main Conference EMNLP 2025☆18Jun 23, 2025Updated last year
- Run ops on Apple ANE in NPU register with pure python on M1 Asahi Linux. No Espresso, No CoreML, no metal, no .mlmodels file, no .hwx fil…☆16Jun 28, 2026Updated 3 weeks ago
- a fast and lightweight distributed background task processing framework with seamless scheduling.☆15Mar 30, 2026Updated 3 months ago
- Implementation and explorations into DiscoRL, Discovering state-of-the-art reinforcement learning algorithms, David Silver's last work at…☆21Jun 13, 2026Updated last month
- Implementation of the fast weight product key memory from Sakana AI☆19Apr 1, 2026Updated 3 months ago
- Repository for paper: Contexts are Never Long Enough: Structured Reasoning for Scalable Question Answering over Long Document Sets☆27Apr 27, 2026Updated 3 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Implementation of Poly-attention, a higher-order self-attention proposed by Chakrabarti et al. of Columbia☆52Updated this week
- Pytorch implementation of Evolutionary Policy Optimization, from Wang et al. of the Robotics Institute at Carnegie Mellon University☆110May 18, 2026Updated 2 months ago
- ☆20Sep 16, 2025Updated 10 months ago
- Implementation of the new SOTA for model based RL, from the paper "Improving Transformer World Models for Data-Efficient RL", in Pytorch☆155May 2, 2025Updated last year
- a simple API to use CUPTI☆10Aug 19, 2025Updated 11 months ago
- Implementation of Multiscreen proposed by Ken Nakanishi for "Screening is Enough"☆18May 13, 2026Updated 2 months ago
- ☆16Feb 9, 2026Updated 5 months ago
- ☆24May 21, 2026Updated 2 months ago
- NICE: Neurogenesis Inspired Contextual Encoding for Replay-free Class Incremental Learning☆29Jul 28, 2024Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- JAX implementation of configurable LLM distillation training☆24Nov 15, 2025Updated 8 months ago
- Implementation of the proposed Adam-atan2 from Google Deepmind in Pytorch☆143Jul 17, 2026Updated last week
- Jax Codebase for Evolutionary Strategies at the Hyperscale☆350Feb 27, 2026Updated 4 months ago
- Learning to Skip the Middle Layers of Transformers☆17Aug 7, 2025Updated 11 months ago
- Efficient implementation (and explorations) into polar coordinate positional embedding (PoPE) - from Gopalakrishnan et al. under Schmidhu…☆71Jun 21, 2026Updated last month
- Efficiently discovering algorithms via LLMs with evolutionary search and reinforcement learning.☆17Apr 22, 2025Updated last year
- ☆192Oct 31, 2025Updated 8 months ago
- Marketplace ML experiment - training without backprop☆28Sep 9, 2025Updated 10 months ago
- ☆24Dec 11, 2024Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Recovery-Bench is a benchmark for evaluating the capability of LLM agents to recover from mistakes☆27Jun 17, 2026Updated last month
- ☆13Oct 14, 2024Updated last year
- A transformer that executes a one-instruction Turing-complete computer — two approaches: hand-coded weights (no training) and learned fro…☆41Mar 3, 2026Updated 4 months ago
- ☆13Jun 7, 2023Updated 3 years ago
- Triton‑style kernel toolkit for MLX plus a small upstream incubator: prototype, benchmark, and upstream fusions for Apple Silicon☆47Mar 31, 2026Updated 3 months ago
- Implementation of Danijar's latest iteration for his Dreamer line of work☆209Jul 4, 2026Updated 3 weeks ago
- Explorations into NEAT and some of its derivative research☆41Jul 6, 2026Updated 3 weeks ago