π± A little course on Reinforcement Learning Environments for evaluating and training Language Models
β221May 27, 2026Updated 2 months ago
Alternatives and similar repositories for llm-rl-environments-lil-course
Users that are interested in llm-rl-environments-lil-course are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Using Llama2 with Haystack, the NLP/LLM framework.β16Jul 21, 2023Updated 3 years ago
- Hugging Face's take home challenge for post-training internships, now open for you to try!β50Apr 23, 2026Updated 3 months ago
- A course on context engineering with code agents.β81May 26, 2026Updated 2 months ago
- A Python SDK to interact with deepset Cloudβ16Aug 10, 2026Updated last week
- A repo of resource for the GPU Mode talk on OpenEnv.β16Jan 14, 2026Updated 7 months ago
- End-to-end encrypted cloud storage - Proton Drive β’ AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Use `outlines` generators with Haystack.β14Updated this week
- β25Mar 28, 2023Updated 3 years ago
- β25Jan 28, 2026Updated 6 months ago
- Lab Cookbookβ42Aug 5, 2026Updated 2 weeks ago
- a whirlwind tour to deep learning and deep learning systemsβ82Updated this week
- A full model customization repo for Liquid Foundation Modelsβ197Updated this week
- β38Jul 8, 2025Updated last year
- π€ A GitHub bot that summarizes your actions throughout a day and lists all your achievements. Built with Haystack and OpenAIβ16Apr 14, 2023Updated 3 years ago
- Examples, end-2-end tutorials and apps built using Liquid AI Foundational Models (LFM) and the LEAP SDKβ2,404Updated this week
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Example code using the DSPy framework.β20May 30, 2024Updated 2 years ago
- Building and Scaling RL environments in the age of LLMsβ178Aug 6, 2026Updated last week
- A benchmarking harness for coding agents.β16Jul 31, 2026Updated 2 weeks ago
- Agentic RL Training at Scaleβ1,943Updated this week
- Our library for RL environments + evalsβ4,521Updated this week
- Website for Haystack, the open source LLM frameworkβ23Aug 11, 2026Updated last week
- minimalistic AI library that resembles HF's transformersβ13Dec 31, 2024Updated last year
- Environments by the Prime Intellect Research Teamβ119Updated this week
- Train LLM on Hugging Face infraβ72May 26, 2026Updated 2 months ago
- Simple, predictable pricing with DigitalOcean hosting β’ AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- j1-micro (1.7B) & j1-nano (600M) are absurdly tiny but mighty reward models.β106Jul 19, 2025Updated last year
- Simple Question Answering system, based on data crawled from Twin Peaks Wiki. It is built using π Haystack, an awesome open-source frameβ¦β11Jun 22, 2023Updated 3 years ago
- β169Jun 18, 2026Updated 2 months ago
- a set of scripts to easily convert all training data from huggingface into alpaca instruct or sharegpt format, which should allow for easβ¦β20Mar 14, 2025Updated last year
- A PyTorch implementation of the GPT-OSS-20B architecture. All components are coded from scratch: RoPE with YaRN, RMSNorm, SwiGLU with claβ¦β238Dec 2, 2025Updated 8 months ago
- The best ChatGPT that $100 can buy.β58Updated this week
- Open source framework for evaluating AI Agentsβ32Feb 24, 2026Updated 5 months ago
- π Unstructured Data Connectors for Haystack 2.0β18Sep 21, 2023Updated 2 years ago
- β10Dec 3, 2024Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Semantic Memory for AI Data Agentsβ30Updated this week
- β10Oct 24, 2024Updated last year
- SRE Agent β An AI-powered MCP server for production incident triage. Takes natural-language symptom reports, plans structured investigatiβ¦β36Mar 6, 2026Updated 5 months ago
- A tiny 1.3M parameter model that plays DOOM, outperforming LLMs up to 92,000x its size.β26May 11, 2026Updated 3 months ago
- Project code for training LLMs to write better unit tests + codeβ22May 19, 2025Updated last year
- An educational distributed training and inference library for neural nets using local computingβ75Jun 10, 2026Updated 2 months ago
- FrontierSWE is an ultra long-horizon coding agent benchmark that tests implementation, performance eng and ML researchβ214Updated this week