The simplest, fastest repository for training/finetuning medium-sized GPTs.
☆38Dec 3, 2023Updated 2 years ago
Alternatives and similar repositories for nanoGPT-jax
Users that are interested in nanoGPT-jax are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Jax/Flax rewrite of Karpathy's nanoGPT☆66Feb 15, 2023Updated 3 years ago
- A videogame made with PyGame turned into an Open AI Gym Learning Environment for Reinforcement Learning agents.☆14Jan 3, 2023Updated 3 years ago
- Blazingly fast implementation of the Datasaurus paper. Same Stats, Different Graphs.☆19Mar 22, 2026Updated 5 months ago
- On the Feasibility of Cross-Task Transfer with Model-Based Reinforcement Learning☆16Apr 30, 2023Updated 3 years ago
- ☆15Jan 16, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A simple but instructive implementation of DP, TP, FSDP, FSDP+TP using pytorch distributed primitives☆24Apr 12, 2026Updated 4 months ago
- Distributed pretraining of large language models (LLMs) on cloud TPU slices, with Jax and Equinox.☆27Sep 29, 2024Updated last year
- ☆19Apr 16, 2022Updated 4 years ago
- Decision Transformer JAX - Reproduction of 'Decision Transformer: Reinforcement Learning via Sequence Modeling' in JAX and Haiku☆13Aug 14, 2024Updated 2 years ago
- Actor-Sharer-Learner training framework for off-policy DRL algorithms☆22Dec 29, 2024Updated last year
- minGPT in JAX☆49Jan 10, 2022Updated 4 years ago
- Prioritized Experience Replay implementation with proportional prioritization☆90Jul 18, 2023Updated 3 years ago
- Adaptation of DQN, DDQN and COMA for multi-agent Gym environments☆10Oct 3, 2023Updated 2 years ago
- ☆23Jun 8, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code for reproducing the results from the paper Avoiding Side Effects in Complex Environments☆12Jun 3, 2021Updated 5 years ago
- This is a port of Mistral-7B model in JAX☆34Jul 1, 2024Updated 2 years ago
- Platform to run interactive Reinforcement Learning agents in a Minecraft Server☆57Apr 21, 2026Updated 4 months ago
- An implementation of MuZero in JAX.☆58Nov 8, 2022Updated 3 years ago
- ☆14Jun 11, 2025Updated last year
- Scripts for performing experiments on semantic networks generated by a machine learning model.☆12Apr 7, 2019Updated 7 years ago
- [NeurIPS'22] Official Repository for Characterizing Datapoints via Second-Split Forgetting☆16Aug 11, 2023Updated 3 years ago
- MLJ Interface for ScikitLearn.jl☆14May 22, 2024Updated 2 years ago
- Fast Flexible Replay Buffer Library (Mirror repository of https://gitlab.com/ymd_h/cpprb)☆77Dec 14, 2024Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Train very large language models in Jax.☆208Oct 21, 2023Updated 2 years ago
- Codebase for a Marimba playing robot☆16Nov 6, 2024Updated last year
- Repo for materials for coordinating work on improving Julia's function documentation☆10Jul 30, 2022Updated 4 years ago
- A MATLAB function library containing encoders, decoders and weight enumerators for Reed-Muller codes.☆13Aug 19, 2023Updated 3 years ago
- ☆13Jan 16, 2019Updated 7 years ago
- MI and Formal Verification of NNs on Algorithmic tasks!☆18Mar 18, 2024Updated 2 years ago
- Simple single file implementations of Reinforcement Learning algorithms in Julia☆24Feb 15, 2025Updated last year
- Learning bisimulation metrics for control, particularly suited to sparse reward settings☆11Feb 28, 2023Updated 3 years ago
- ☆26Jun 14, 2022Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A codebase dedicated to exploring multimodal learning approaches by integrating images of host galaxies of supernovae and their correspon…☆17Jan 15, 2025Updated last year
- PlaNet: Learning Latent Dynamics for Planning from Pixels☆10Feb 13, 2020Updated 6 years ago
- Rainbow DQN implementation accompanying the paper "Fast and Data-Efficient Training of Rainbow" which reaches 205.7 median HNS after 10M …☆46Dec 11, 2021Updated 4 years ago
- An implementation of a Brownian motion using ClojureScript with re-frame and Highcharts☆11Feb 8, 2019Updated 7 years ago
- Brax + Pufferlib + CARBS for gpu-accelerated robotics RL☆12Jun 12, 2025Updated last year
- clip mp4 for the most viral moments with Deepseek☆14Jul 15, 2025Updated last year
- Soft Actor-Critic implementation with SOTA model-free extension (REDQ) and SOTA model-based extension (MBPO).☆15Feb 21, 2021Updated 5 years ago