Simple framework for training and evaluating math reasoning agents using local models, GRPO and vLLM.
☆41Jun 22, 2026Updated 3 weeks ago
Alternatives and similar repositories for DeepMath
Users that are interested in DeepMath are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Mixture of Experts from scratch☆14Apr 12, 2024Updated 2 years ago
- ☆21Aug 19, 2025Updated 11 months ago
- Improving Math reasoning through Direct Preference Optimization with Verifiable Pairs☆21Mar 20, 2025Updated last year
- ☆23Jul 5, 2024Updated 2 years ago
- ☆22Jul 10, 2026Updated last week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ICLR'24] Official code for "C-TPT: Calibrated Test-Time Prompt Tuning for Vision-Language Models via Text Feature Dispersion"☆23Jun 9, 2024Updated 2 years ago
- Official implementation of the paper "Next Embedding Prediction Makes World Models Stronger"☆37Mar 5, 2026Updated 4 months ago
- This is the Placeholder for Llama. Starting with Llama 3☆11May 20, 2024Updated 2 years ago
- [BMVC 2025 🔥] CalibPrompt is the first framework that enhances Med-VLM calibration during prompt tuning.☆16Jul 13, 2026Updated last week
- Jointly Optimizing Large Language Models for Reasoning and Self-Refinement☆15Apr 22, 2026Updated 2 months ago
- It's for SK Networks AI 1☆15Oct 24, 2024Updated last year
- ☆16Sep 25, 2025Updated 9 months ago
- Repo for collaboration on OSS agentic code search☆70Apr 29, 2026Updated 2 months ago
- [ACL 2026 Main] Official Repo for Paper "Which Reasoning Trajectories Teach Students to Reason Better? A Simple Metric of Informative Ali…☆17Jul 1, 2026Updated 2 weeks ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Advancing Small and Medium-sized Code Agents.☆17May 29, 2026Updated last month
- Model for multi-turn tool calling of bash functions☆22Jan 26, 2026Updated 5 months ago
- AI Agent for voice-driven math visualization☆34Apr 17, 2026Updated 3 months ago
- Codes and data for AAAI-24 paper "Advancing Spatial Reasoning in Large Language Models: An In-depth Evaluation and Enhancement Using the …☆14Apr 23, 2024Updated 2 years ago
- ☆17Jun 3, 2025Updated last year
- (ACL2025 Findings) Official code for the paper "STeCa: Step-level Trajectory Calibration for LLM Agent Learning"☆29Mar 2, 2026Updated 4 months ago
- ☆100Jun 23, 2025Updated last year
- ☆10May 9, 2016Updated 10 years ago
- ☆18Jul 31, 2025Updated 11 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆22Dec 24, 2025Updated 6 months ago
- ☆19Mar 26, 2026Updated 3 months ago
- Implementation of a GPT-4o like Multimodal from Scratch using Python☆78Apr 4, 2025Updated last year
- Training tiny models to prove hard theorems☆81Mar 5, 2026Updated 4 months ago
- ☆18Jan 1, 2026Updated 6 months ago
- 🔥🔥🔥 [NeurIPS2025] MM-Agent: LLM as Agents for Real-world Mathematical Modeling Problem☆608May 22, 2026Updated last month
- Implementation of the model: "Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models" in PyTorch☆28Jul 13, 2026Updated last week
- Link Prediction with Signed Latent Factors in Signed Social Networks (SIGKDD 2019)☆15Oct 24, 2021Updated 4 years ago
- NeurIPS 2025: Discriminative Constrained Optimization for Reinforcing Large Reasoning Models☆53Mar 14, 2026Updated 4 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Local AI code review assistant built on dspy-go.☆29Apr 22, 2026Updated 2 months ago
- Official Implementation of ISR-DPO:Aligning Large Multimodal Models for Videos by Iterative Self-Retrospective DPO (AAAI'25)☆23Nov 25, 2025Updated 7 months ago
- Official Implementation of Papar CM2☆25Apr 21, 2026Updated 3 months ago
- Video-CoM: Interactive Video Reasoning via Chain of Manipulations☆22Jun 17, 2026Updated last month
- The code of paper "Nonlinear Hybrid Planning with Deep Net Learned Transition Models and Mixed-Integer Linear Programming." published on …☆10Apr 27, 2018Updated 8 years ago
- BigCodeBench: Benchmarking Code Generation with Diverse Function Calls and Complex Instructions☆26Aug 8, 2024Updated last year
- [CVPR 2022] Official code for the paper: "A Stitch in Time Saves Nine: A Train-Time Regularizing Loss for Improved Neural Network Calibra…☆33Nov 9, 2022Updated 3 years ago