☆52Feb 20, 2026Updated 5 months ago
Alternatives and similar repositories for trl-tuto
Users that are interested in trl-tuto are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This repository contains all code examples for my TensorFlow World talk about "Advanced model deployments with TensorFlow Serving"☆17Dec 8, 2022Updated 3 years ago
- Train your own SOTA deductive reasoning model☆111Mar 6, 2025Updated last year
- ☆15Jun 2, 2025Updated last year
- Notebooks from DS3 course on practical optimization☆15Jan 5, 2021Updated 5 years ago
- Layerwise Batch Entropy Regularization☆24Aug 3, 2022Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A FastAPI server that turns markdown prompt files into API endpoints with minimal configuration.☆15Sep 5, 2025Updated 10 months ago
- ☆12May 27, 2024Updated 2 years ago
- ☆12Mar 18, 2024Updated 2 years ago
- ☆25Jan 28, 2026Updated 6 months ago
- ☆17Mar 24, 2026Updated 4 months ago
- Simple UI for debugging correlations of text embeddings☆315May 28, 2025Updated last year
- MLflow is Open source platform for the machine learning lifecycle so here you can learn MLflow End to End Example with Prediction.☆13Jun 14, 2022Updated 4 years ago
- Code for "Counterfactual Token Generation in Large Language Models", Arxiv 2024.☆34Nov 7, 2024Updated last year
- Karpathy's llama2.c transpiled to MLX for Apple Silicon☆14Dec 28, 2023Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- A starter kit for evaluating benchmarks on the 🤗 Hub☆18Apr 8, 2026Updated 3 months ago
- Advanced Data Science with IBM Specialization☆12Aug 9, 2021Updated 4 years ago
- Inference Llama 2 in one file of pure C☆14Jul 24, 2023Updated 3 years ago
- ☆170Jun 3, 2024Updated 2 years ago
- SNoRe: Scalable Unsupervised Learning of Symbolic Node Representations☆11Sep 26, 2023Updated 2 years ago
- Code on IART: Intent-aware Response Ranking with Transformers in Information-seeking Conversation Systems (WWW 2020)☆11Apr 18, 2021Updated 5 years ago
- A complete waste of time☆15Dec 11, 2022Updated 3 years ago
- Open-source Python implementation of a Claude-Code-like agent (Gemini CLI, Codex, Copilot...). Usable in CLI, GUI, or as a python library…☆24Updated this week
- ☆10Jul 18, 2022Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Official GitHub repository of the lecture "Multimodal Deep Learning for Recommendation", at the 2024 ACM RecSys Summer School☆12Oct 12, 2024Updated last year
- Lightblue LLM Eval Framework: tengu, elyza100, ja-mtbench, rakuda☆19Apr 29, 2026Updated 3 months ago
- A minimal re-implementation of orthogonal fine-tuning (OFT), a diffusion method, for LLMs. Based on nanoGPT and minLoRA.☆14Nov 17, 2023Updated 2 years ago
- Simple model for sentence compression (a.k.a Baseline in Klerke et al., NAACL 2016)☆10Dec 16, 2018Updated 7 years ago
- Ludic – an LLM-RL library for the era of experience☆67Jan 9, 2026Updated 6 months ago
- Train LLM on Hugging Face infra☆72May 26, 2026Updated 2 months ago
- An R implementation of some of the data science methods for wind energy (DSWE) applications.☆11Feb 6, 2024Updated 2 years ago
- AI for a cure, a combination of Latent-GAN and VAE-JTNN to create 100% valid drug like molecules☆10Mar 16, 2020Updated 6 years ago
- A collection of fine-tuning notebooks!☆32Oct 5, 2023Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Python library to use Pleias-RAG models☆72Jul 1, 2026Updated 3 weeks ago
- A framework for orchestrating AI agents using a mermaid graph☆76May 16, 2024Updated 2 years ago
- Distributed Reinforcement Learning for LLM Fine-Tuning with multi-GPU utilization☆22Mar 12, 2025Updated last year
- The CODWOE shared task invites you to compare two types of semantic descriptions: dictionary glosses and word embedding representations. …☆12Jul 13, 2022Updated 4 years ago
- ☆13Jul 19, 2021Updated 5 years ago
- ☆10Oct 24, 2024Updated last year
- Tiny evaluation of leading LLMs on competitive programming problems☆14Apr 10, 2026Updated 3 months ago