The simplest, fastest repository for training/finetuning medium-sized GPTs.
☆38Dec 3, 2023Updated 2 years ago
Alternatives and similar repositories for nanoGPT-jax
Users that are interested in nanoGPT-jax are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Jax/Flax rewrite of Karpathy's nanoGPT☆66Feb 15, 2023Updated 3 years ago
- A videogame made with PyGame turned into an Open AI Gym Learning Environment for Reinforcement Learning agents.☆14Jan 3, 2023Updated 3 years ago
- Blazingly fast implementation of the Datasaurus paper. Same Stats, Different Graphs.☆19Mar 22, 2026Updated 4 months ago
- On the Feasibility of Cross-Task Transfer with Model-Based Reinforcement Learning☆16Apr 30, 2023Updated 3 years ago
- ☆14Jan 16, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆19Apr 16, 2022Updated 4 years ago
- Decision Transformer JAX - Reproduction of 'Decision Transformer: Reinforcement Learning via Sequence Modeling' in JAX and Haiku☆13Aug 14, 2024Updated last year
- Actor-Sharer-Learner training framework for off-policy DRL algorithms☆22Dec 29, 2024Updated last year
- minGPT in JAX☆49Jan 10, 2022Updated 4 years ago
- Prioritized Experience Replay implementation with proportional prioritization☆91Jul 18, 2023Updated 3 years ago
- Adaptation of DQN, DDQN and COMA for multi-agent Gym environments☆10Oct 3, 2023Updated 2 years ago
- ☆13Dec 21, 2018Updated 7 years ago
- ☆23Jun 8, 2021Updated 5 years ago
- Code for reproducing the results from the paper Avoiding Side Effects in Complex Environments☆12Jun 3, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This is a port of Mistral-7B model in JAX☆34Jul 1, 2024Updated 2 years ago
- Platform to run interactive Reinforcement Learning agents in a Minecraft Server☆57Apr 21, 2026Updated 3 months ago
- An implementation of MuZero in JAX.☆58Nov 8, 2022Updated 3 years ago
- Example code snipped to visualize a neural network fitting a surface to random points in space☆12Dec 26, 2021Updated 4 years ago
- Minimal but scalable implementation of large language models in JAX☆34Nov 28, 2025Updated 8 months ago
- ☆23Aug 19, 2022Updated 3 years ago
- Integrates Imbue's Cost Aware pareto-Region Bayesian Search (CARBS) with Weights and Biases (WanDB)☆12Mar 17, 2025Updated last year
- ☆14Jun 11, 2025Updated last year
- Code accompanying the paper "Disparate Impact in Differential Privacy from Gradient Misalignment".☆11Apr 4, 2023Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- coloring terminal text with intensities (used for plotting probability, entropy with tokens)☆12Oct 11, 2024Updated last year
- ☆25Jan 2, 2019Updated 7 years ago
- Fast Flexible Replay Buffer Library (Mirror repository of https://gitlab.com/ymd_h/cpprb)☆77Dec 14, 2024Updated last year
- Train very large language models in Jax.☆208Oct 21, 2023Updated 2 years ago
- ☆143Mar 31, 2023Updated 3 years ago
- A MATLAB function library containing encoders, decoders and weight enumerators for Reed-Muller codes.☆13Aug 19, 2023Updated 2 years ago
- ☆13Jan 16, 2019Updated 7 years ago
- Simple single file implementations of Reinforcement Learning algorithms in Julia☆24Feb 15, 2025Updated last year
- Learning Off-Policy with Online Planning [CoRL 2021 Best Paper Finalist]☆42Aug 27, 2022Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- ☆26Jun 14, 2022Updated 4 years ago
- A codebase dedicated to exploring multimodal learning approaches by integrating images of host galaxies of supernovae and their correspon…☆17Jan 15, 2025Updated last year
- Rainbow DQN implementation accompanying the paper "Fast and Data-Efficient Training of Rainbow" which reaches 205.7 median HNS after 10M …☆44Dec 11, 2021Updated 4 years ago
- An implementation of a Brownian motion using ClojureScript with re-frame and Highcharts☆11Feb 8, 2019Updated 7 years ago
- ☆13Sep 13, 2015Updated 10 years ago
- Brax + Pufferlib + CARBS for gpu-accelerated robotics RL☆12Jun 12, 2025Updated last year
- CUDA integration for the NNlib API☆14Feb 5, 2023Updated 3 years ago