j1-micro (1.7B) & j1-nano (600M) are absurdly tiny but mighty reward models.
☆108Jul 19, 2025Updated last year
Alternatives and similar repositories for j1-micro
Users that are interested in j1-micro are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Inference-time scaling for LLMs-as-a-judge.☆348Nov 5, 2025Updated 10 months ago
- a single interface around speech-to-speech foundation models☆28Jun 27, 2025Updated last year
- ☆39Aug 4, 2025Updated last year
- A framework for optimizing DSPy programs with RL☆339Jan 12, 2026Updated 8 months ago
- Approximating the joint distribution of language models via MCTS☆22Nov 3, 2024Updated last year
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Entropy Based Sampling and Parallel CoT Decoding☆17Oct 9, 2024Updated last year
- nyc is so back☆21Jun 27, 2025Updated last year
- ⚖️ Awesome LLM Judges ⚖️☆203Apr 28, 2025Updated last year
- High-Performance Engine for Multi-Vector Search☆282Sep 10, 2026Updated 2 weeks ago
- look how they massacred my boy☆63Oct 16, 2024Updated last year
- Exploring Applications of GRPO☆251Aug 25, 2025Updated last year
- smolLM with Entropix sampler on pytorch☆148Oct 31, 2024Updated last year
- ☆32Sep 4, 2025Updated last year
- Storing long contexts in tiny caches with self-study☆337Mar 23, 2026Updated 6 months ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Is In-Context Learning Sufficient for Instruction Following in LLMs? [ICLR 2025]☆34Jan 23, 2025Updated last year
- Our library for RL environments + evals☆4,651Updated this week
- An introduction to LLM Sampling☆80Dec 15, 2024Updated last year
- ☆139Mar 20, 2025Updated last year
- ☆19Mar 3, 2025Updated last year
- ☆29Jan 14, 2025Updated last year
- rl from zero pretrain, can it be done? yes.☆296Sep 28, 2025Updated 11 months ago
- A collection of Compound Retrieval Systems implemented with DSPy and Weaviate.☆100Jun 1, 2026Updated 3 months ago
- Mine-tuning is a methodology for synchronizing human and AI attention.☆21Jun 16, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆12Dec 30, 2020Updated 5 years ago
- Small, simple agent task environments for training and evaluation☆20Nov 1, 2024Updated last year
- This repository helps you evaluate your models on the FreshStack benchmark!☆34Dec 9, 2025Updated 9 months ago
- Python SDK for Modaic☆28Updated this week
- Official Code For Dual Grained Quantization: Efficient Fine-Grained Quantization for LLM☆14Dec 27, 2023Updated 2 years ago
- ☆15Dec 12, 2024Updated last year
- Easiest way to give context to LLMs; Attachments has the ambition to be the general funnel for any files to be transformed into images+te…☆369Jun 10, 2026Updated 3 months ago
- Late Interaction Models Training & Retrieval☆895Jul 23, 2026Updated 2 months ago
- Agent Engineering course files☆71Jul 12, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official code repository for the paper "Internal Activation as the Polar Star for Steering Unsafe LLM Behavior"☆15May 31, 2026Updated 3 months ago
- ☆36Updated this week
- ☆59Aug 19, 2025Updated last year
- ☆69May 23, 2025Updated last year
- ☆53Dec 2, 2025Updated 9 months ago
- Red-Teaming Language Models with DSPy☆275Feb 13, 2025Updated last year
- KernelBench v2: Can LLMs Write GPU Kernels? - Benchmark with Torch -> Triton (and more!) problems☆26Jul 4, 2025Updated last year