fast trainer for educational purposes
☆27Jul 24, 2026Updated last month
Alternatives and similar repositories for mini_trainer
Users that are interested in mini_trainer are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- An algorithm-focused interface for common llm training, continual learning, and reinforcement learning techniques☆92Updated this week
- Synthetic Data Generation Toolkit for LLMs☆157Updated this week
- ☆22Jun 5, 2025Updated last year
- A repo for open research on building large reasoning models☆153Jul 3, 2026Updated last month
- A pytorch implementation of Deep Functional Map (FMNet).☆15May 6, 2020Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 🤖ConvRe🤯: An Investigation of LLMs’ Inefficacy in Understanding Converse Relations (EMNLP 2023)☆24Oct 10, 2023Updated 2 years ago
- A MBTI test on Large Language Model like GPT-3.☆28May 2, 2022Updated 4 years ago
- Efficient multi-prompt evaluation of LLMs☆33Dec 6, 2024Updated last year
- Implementation of ICML 2023 paper: Future-conditioned Unsupervised Pretraining for Decision Transformer☆29Jul 25, 2023Updated 3 years ago
- Group-conditional DRO to alleviate spurious correlations☆15Jul 15, 2021Updated 5 years ago
- Gecko Architecture☆18Jan 13, 2026Updated 7 months ago
- [KDD 2023] code for "Test accuracy vs. generalization gap: model selection in NLP without accessing training or testing data" https://arx…☆12Oct 17, 2022Updated 3 years ago
- ☆58Jun 1, 2026Updated 2 months ago
- [ICML'24] TroVE: Inducing Verifiable and Efficient Toolboxes for Solving Programmatic Tasks☆33Sep 20, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [TPAMI] "Symbolic Visual Reinforcement Learning: A Scalable Framework with Object-Level Abstraction and Differentiable Expression Search"…☆18Jan 4, 2023Updated 3 years ago
- The official implementation of the paper "Self-Updatable Large Language Models by Integrating Context into Model Parameters"☆16May 18, 2025Updated last year
- Pretraining with Natural and Synthetic Data for Few-shot Table-based Question Answering☆31Dec 2, 2022Updated 3 years ago
- Scale digital agent rollouts without pain.☆36Jun 18, 2026Updated 2 months ago
- ☆18Jan 17, 2024Updated 2 years ago
- Collaborative inference of latent diffusion via hivemind☆12May 29, 2023Updated 3 years ago
- Evaluating LLMs with fewer examples☆185Jul 4, 2026Updated last month
- ☆14Feb 1, 2024Updated 2 years ago
- Generate sentences from a probabilistic context-free grammar.☆17Nov 8, 2024Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Blog post☆17Feb 16, 2024Updated 2 years ago
- Word Embeddings for Low Resource Languages: The Case of Buryat☆10Mar 12, 2025Updated last year
- ☆12Dec 13, 2022Updated 3 years ago
- Collections of RLxLM experiments using minimal codes☆14Feb 17, 2025Updated last year
- ☆18Mar 25, 2021Updated 5 years ago
- Repository for getting started with the OfficeQA Benchmark.☆184Aug 6, 2026Updated 3 weeks ago
- ☆52Mar 17, 2025Updated last year
- Some improvements on Adam☆28Nov 5, 2020Updated 5 years ago
- ☆13Feb 25, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Russian dialog datasets parsers and crawlers.☆15Sep 6, 2021Updated 4 years ago
- ☆55Aug 25, 2023Updated 3 years ago
- [NeurIPS 2021] code for "Taxonomizing local versus global structure in neural network loss landscapes" https://arxiv.org/abs/2107.11228☆20Jan 7, 2022Updated 4 years ago
- Benchmark structured generation libraries☆31Oct 25, 2024Updated last year
- The official repository for the paper "From Zero to Hero: Examining the Power of Symbolic Tasks in Instruction Tuning".☆64Apr 18, 2023Updated 3 years ago
- ☆16Jun 26, 2026Updated 2 months ago
- Unofficial baselines for ManiSkill, including RL and BC algorithms.☆22Jun 6, 2024Updated 2 years ago