LLM RL envs done right, plus some training code
☆17Feb 1, 2026Updated 5 months ago
Alternatives and similar repositories for gyllm
Users that are interested in gyllm are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Jido implementation of Managed Agents on Phoenix☆23Updated this week
- Ludic – an LLM-RL library for the era of experience☆67Jan 9, 2026Updated 6 months ago
- ☆12Aug 8, 2023Updated 2 years ago
- Code for 'Answer Matching Outperforms Multiple Choice for Language Model Evaluation' paper☆18Jul 4, 2025Updated last year
- Apertium tools☆20May 27, 2021Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- An archive of learning resources assembled by current Exun members and alumni.☆15Jun 23, 2026Updated 3 weeks ago
- Reinforcement learning with Equinox☆21Mar 4, 2025Updated last year
- An unbounded n-gram language model on Tiny Shakespeare☆22Jan 21, 2026Updated 5 months ago
- ☆10Nov 6, 2024Updated last year
- A lightweight computational physics framework, based on the organization of turboWAVE. Implements a "Simulation, PhysicsModule, ComputeTo…☆12Jul 13, 2026Updated last week
- ☆24Sep 24, 2024Updated last year
- ☆20May 23, 2025Updated last year
- A Python DSL for Apple Metal GPU compute☆25Jun 20, 2026Updated last month
- Agents, and RL environment, for optimizing GPU kernels on AMD ROCm using LLM agents. Benchmarks LLM serving workloads end-to-end, profile…☆71Updated this week
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- ☆12Dec 30, 2020Updated 5 years ago
- ☆13Apr 16, 2025Updated last year
- JAX bindings for the flash-attention3 kernels☆23Jan 2, 2026Updated 6 months ago
- ☆24Dec 17, 2025Updated 7 months ago
- Free Draft-and-Verification: Toward Lossless Parallel Decoding for Diffusion Large Language Models☆23May 19, 2026Updated 2 months ago
- ☆27Jun 16, 2026Updated last month
- ☆12Apr 29, 2022Updated 4 years ago
- Transformer LIbrary Docker Stacks☆14Feb 4, 2022Updated 4 years ago
- "We must know. We shall know." - David Hilbert☆21Sep 8, 2025Updated 10 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Repo of "Seeing the Whole Elephant: A Benchmark for Failure Attribution in LLM-based Multi-Agent Systems" (ACL 2026)☆16Apr 27, 2026Updated 2 months ago
- Code and Configs for Asynchronous RLHF: Faster and More Efficient RL for Language Models☆68Mar 5, 2026Updated 4 months ago
- CoEvolve: Training LLM Agents via Agent-Data Mutual Evolution☆19Apr 27, 2026Updated 2 months ago
- Tensorflow 2.0 implementation of STAR RNN☆10Jun 7, 2020Updated 6 years ago
- Experiments in protein folding through language modeling☆10Dec 10, 2021Updated 4 years ago
- Unofficially Implements https://arxiv.org/abs/2112.05682 to get Linear Memory Cost on Attention for PyTorch☆12Jan 16, 2022Updated 4 years ago
- A framework for implementing equivariant DL☆10May 25, 2021Updated 5 years ago
- A Codex skill (via CLI) which runs codex in parallel to reflect on previous conversations to brainstorm some skills you could add to your…☆19Jan 23, 2026Updated 5 months ago
- RWKV6 in native pytorch and triton:)☆11Aug 4, 2024Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Hugging Face and Pyserini interoperability☆20May 18, 2023Updated 3 years ago
- Github Pages template for academic personal websites, forked from mmistakes/minimal-mistakes☆11Jul 10, 2024Updated 2 years ago
- Fast and generic implementation of the minhash algorithm☆14Jul 8, 2025Updated last year
- SkillHack: A Benchmark for Skill Transfer in Open-Ended Reinforcement Learning☆17Oct 23, 2022Updated 3 years ago
- density growing clustering☆10Dec 2, 2021Updated 4 years ago
- Hausa-NMT: Empirical Study of Neural Machine translation for English-Hausa-English☆17Oct 20, 2020Updated 5 years ago
- Released Apertium language modules☆40May 27, 2021Updated 5 years ago