Alpha-Zero Connect Four NN trained via self play
☆27Jun 30, 2026Updated 2 months ago
Alternatives and similar repositories for c4a0
Users that are interested in c4a0 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A minimal home grid world environment to evaluate language understanding in interactive agents.☆24Sep 6, 2023Updated 2 years ago
- manipulating cointegrated pairs to achieve a market-neutral strategy that outperforms indices☆11Jan 12, 2021Updated 5 years ago
- ☆25May 23, 2025Updated last year
- A number of agents (PPO, MuZero) with a Perceiver-based NN architecture that can be trained to achieve goals in nethack/minihack environm…☆44Sep 19, 2022Updated 3 years ago
- Action Value Gradient Algorithm☆30May 18, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Implementation in the framework of my bachelor thesis: Generative Modelling using Capsule Generative Adversarial Networks☆12Feb 20, 2026Updated 6 months ago
- coloring terminal text with intensities (used for plotting probability, entropy with tokens)☆12Oct 11, 2024Updated last year
- ML-Constructive is a deep learning based constructive heuristic for the Traveling Salesman Problem.☆11Feb 10, 2024Updated 2 years ago
- Scratchpad/Chain-of-Thought Prompts☆12Jun 6, 2022Updated 4 years ago
- Code to reproduce key results accompanying "SAEs (usually) Transfer Between Base and Chat Models"☆13Jul 18, 2024Updated 2 years ago
- ☆14Mar 2, 2025Updated last year
- ☆11Dec 10, 2020Updated 5 years ago
- Source code for the paper "Positional Attention: Expressivity and Learnability of Algorithmic Computation"☆14May 26, 2025Updated last year
- Official codebase for Improving Computational Efficiency in Visual Reinforcement Learning via Stored Embeddings.☆21Mar 5, 2021Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Groq-powered MAD: The first work to explore Multi-Agent Debate with Large Language Models :D☆12Jul 5, 2024Updated 2 years ago
- ☆45Apr 30, 2018Updated 8 years ago
- ApertureDB Python Client☆13Updated this week
- Scala Reactive Programming, published by Packt☆10Jan 30, 2023Updated 3 years ago
- This is the code which powers the Twitter Bot https://twitter.com/RGB_Colours☆15Apr 14, 2017Updated 9 years ago
- Experiments Notebook of "Understanding the Skill Gap in Recurrent Language Models: The Role of the Gather-and-Aggregate Mechanism"☆17Apr 30, 2025Updated last year
- Replayable Nova-Seed → MARK → Mini Sovereign → AGI Jobs → Evidence Docket → vNext loop☆16Updated this week
- Text preprocessing package for use in NLP tasks https://pypi.org/project/textcl/☆12Aug 9, 2024Updated 2 years ago
- Minimal implementation of TokenFormer for inference and learning☆13Nov 6, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- I can haz planetz?☆12Jun 12, 2020Updated 6 years ago
- ☆19Dec 4, 2025Updated 8 months ago
- Schedule free optimiser implemented in JAX using Optimistix☆14May 29, 2024Updated 2 years ago
- Parallel Associative Scan for Language Models☆18Jan 8, 2024Updated 2 years ago
- ☆65Jun 12, 2025Updated last year
- ☆13Jul 12, 2024Updated 2 years ago
- A simulator of Michelson interferometer.☆13Nov 23, 2020Updated 5 years ago
- ☆11Feb 4, 2022Updated 4 years ago
- Simulating a 2D Hovering SpaceX Grasshopper with a Thrust Vector Control) engine.☆13Dec 28, 2015Updated 10 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Hybrid Deep Sequential Modeling for Social Text-Driven Stock Prediction-Dataset☆22Aug 19, 2018Updated 8 years ago
- ☆18Apr 19, 2024Updated 2 years ago
- Code for the paper Don't Pay Attention☆60Sep 25, 2025Updated 11 months ago
- ☆16Jul 16, 2024Updated 2 years ago
- This repo contains the code for the reinforcement learning course project https://github.com/cuhkrlcourse☆12May 24, 2020Updated 6 years ago
- Portuguese translation of the GLUE benchmark and Scitail dataset☆33Jun 27, 2022Updated 4 years ago
- Grokking on modular arithmetic in less than 150 epochs in MLX☆15Oct 24, 2024Updated last year