π± A little course on Reinforcement Learning Environments for evaluating and training Language Models
β226May 27, 2026Updated 4 months ago
Alternatives and similar repositories for llm-rl-environments-lil-course
Users that are interested in llm-rl-environments-lil-course are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Using Llama2 with Haystack, the NLP/LLM framework.β16Jul 21, 2023Updated 3 years ago
- Hugging Face's take home challenge for post-training internships, now open for you to try!β51Apr 23, 2026Updated 5 months ago
- Carla worldsim envrionment for decision-making evaluation and RLβ19Apr 22, 2026Updated 5 months ago
- A Python SDK to interact with deepset Cloudβ16Aug 17, 2026Updated last month
- A course on context engineering with code agents.β100Sep 18, 2026Updated last week
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A repo of resource for the GPU Mode talk on OpenEnv.β16Jan 14, 2026Updated 8 months ago
- Use `outlines` generators with Haystack.β14Sep 21, 2026Updated last week
- β24Mar 28, 2023Updated 3 years ago
- β25Jan 28, 2026Updated 8 months ago
- Lab Cookbookβ42Aug 5, 2026Updated last month
- a whirlwind tour to deep learning and deep learning systemsβ83Updated this week
- A full model customization repo for Liquid Foundation Modelsβ215Updated this week
- π€ A GitHub bot that summarizes your actions throughout a day and lists all your achievements. Built with Haystack and OpenAIβ16Apr 14, 2023Updated 3 years ago
- Examples, end-2-end tutorials and apps built using Liquid AI Foundational Models (LFM) and the LEAP SDKβ2,513Sep 18, 2026Updated last week
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Agentic RL Training at Scaleβ2,092Updated this week
- Example code using the DSPy framework.β20May 30, 2024Updated 2 years ago
- A benchmarking harness for coding agents.β17Updated this week
- Our library for RL environments + evalsβ4,652Updated this week
- FineEnvs β RL Environments 101: building and scaling RL environments in the age of LLMsβ262Updated this week
- minimalistic AI library that resembles HF's transformersβ13Dec 31, 2024Updated last year
- Train LLM on Hugging Face infraβ73Updated this week
- Website for Haystack, the open source LLM frameworkβ24Updated this week
- j1-micro (1.7B) & j1-nano (600M) are absurdly tiny but mighty reward models.β108Jul 19, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- a set of scripts to easily convert all training data from huggingface into alpaca instruct or sharegpt format, which should allow for easβ¦β20Mar 14, 2025Updated last year
- A PyTorch implementation of the GPT-OSS-20B architecture. All components are coded from scratch: RoPE with YaRN, RMSNorm, SwiGLU with claβ¦β237Dec 2, 2025Updated 9 months ago
- The best ChatGPT that $100 can buy.β59Updated this week
- Open source framework for evaluating AI Agentsβ32Feb 24, 2026Updated 7 months ago
- π Unstructured Data Connectors for Haystack 2.0β18Sep 21, 2023Updated 3 years ago
- β10Dec 3, 2024Updated last year
- Semantic Memory for AI Data Agentsβ31Sep 11, 2026Updated 2 weeks ago
- β10Oct 24, 2024Updated last year
- A tiny 1.3M parameter model that plays DOOM, outperforming LLMs up to 92,000x its size.β27May 11, 2026Updated 4 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Project code for training LLMs to write better unit tests + codeβ22May 19, 2025Updated last year
- An educational distributed training and inference library for neural nets using local computingβ85Jun 10, 2026Updated 3 months ago
- FrontierSWE is an ultra long-horizon coding agent benchmark that tests implementation, performance eng and ML researchβ230Aug 13, 2026Updated last month
- β38May 4, 2026Updated 4 months ago
- Analyze coinbase orderbook in real-time in Python with Bytewaxβ11Apr 23, 2024Updated 2 years ago
- β19Apr 22, 2026Updated 5 months ago
- An interface library for RL post training with environments.β2,619Updated this week