☆33Oct 2, 2025Updated 9 months ago
Alternatives and similar repositories for muCode
Users that are interested in muCode are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆23Jun 2, 2026Updated last month
- Official repo for "Binary Retrieval-augmented Reward Mitigates Hallucinations"☆15Nov 13, 2025Updated 8 months ago
- ☆17Feb 22, 2025Updated last year
- ☆13Apr 17, 2025Updated last year
- ☆20Jun 16, 2026Updated last month
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆17Jul 31, 2025Updated 11 months ago
- Code of "Regularized Best-of-N Sampling with Minimum Bayes Risk Objective for Language Model Alignment" (2025).☆14Apr 4, 2025Updated last year
- From Accuracy to Robustness: A Study of Rule- and Model-based Verifiers in Mathematical Reasoning.☆24Oct 7, 2025Updated 9 months ago
- Fetches, extracts, and parses data from the arxiv bucket on Amazon S3☆22Jul 12, 2019Updated 7 years ago
- Repository for the paper Stream of Search: Learning to Search in Language☆154Feb 3, 2025Updated last year
- Q-Probe: A Lightweight Approach to Reward Maximization for Language Models☆40Jun 10, 2024Updated 2 years ago
- B-STAR: Monitoring and Balancing Exploration and Exploitation in Self-Taught Reasoners☆86May 21, 2025Updated last year
- Source Code for <Target-Side Data Augmentation for Sequence Generation>☆12Oct 6, 2021Updated 4 years ago
- Recipes to train the self-rewarding reasoning LLMs.☆231Mar 2, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Implementation for the paper "Fictitious Synthetic Data Can Improve LLM Factuality via Prerequisite Learning"☆11Jan 10, 2025Updated last year
- ☆19Jun 3, 2024Updated 2 years ago
- 2020 THU-CST 软件工程大作业 企业固定资产管理系统 后端☆11Jul 17, 2021Updated 5 years ago
- ☆16Dec 10, 2025Updated 7 months ago
- The Dataset and Official Implementation for <Discursive Socratic Questioning: Evaluating the Faithfulness of Language Models’ Understandi…☆18Aug 7, 2024Updated last year
- ☆19Nov 12, 2024Updated last year
- [ICML 2025] Teaching Language Models to Critique via Reinforcement Learning☆127May 6, 2025Updated last year
- [ASE 2025] CoSIL: Issue Localization via Iteritive Code Graph Searching☆23May 31, 2026Updated last month
- Both Text and Images Leaked! A Systematic Analysis of Data Contamination in Multimodal LLM | EMNLP 2025 Findings☆18Oct 17, 2025Updated 9 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Under construction☆13Jan 15, 2025Updated last year
- ☆341Jun 5, 2025Updated last year
- ☆13Jul 8, 2023Updated 3 years ago
- [TMLR] Process Reward Models That Think☆89Nov 29, 2025Updated 7 months ago
- ☆32Oct 8, 2025Updated 9 months ago
- A LLM Multi-Agent Framework toward Ultra Large-Scale Code Generation and Optimization☆18Dec 22, 2024Updated last year
- The official repository for SkyLadder: Better and Faster Pretraining via Context Window Scheduling☆43Dec 29, 2025Updated 6 months ago
- MarketGPT: Developing a Pre-trained transformer (GPT) for Modeling Financial Time Series☆19Sep 5, 2025Updated 10 months ago
- [ICML 2025] M-STAR (Multimodal Self-Evolving TrAining for Reasoning) Project. Diving into Self-Evolving Training for Multimodal Reasoning☆75Jul 13, 2025Updated last year
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆14Jun 3, 2025Updated last year
- Code for 'Contrastive Multi-Document Question Generation'☆11Oct 16, 2022Updated 3 years ago
- Marketplace ML experiment - training without backprop☆28Sep 9, 2025Updated 10 months ago
- ☆15Jan 27, 2025Updated last year
- ☆26Jul 26, 2025Updated 11 months ago
- Code for "Language Models Can Learn from Verbal Feedback Without Scalar Rewards"☆65Jan 5, 2026Updated 6 months ago
- Collection of LLM completions for reasoning-gym task datasets☆31Jul 4, 2025Updated last year