☆35Mar 18, 2026Updated 4 months ago
Alternatives and similar repositories for RationaleRM
Users that are interested in RationaleRM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ACL 2026] This is the repo of Data Darwinism.☆26Apr 16, 2026Updated 3 months ago
- Archer2.0 evolves from its predecessor by introducing ASPO, which overcomes fundamental PPO-Clip limitations to prevent premature converg…☆31Oct 10, 2025Updated 9 months ago
- Your efficient and accurate answer verification system for RL training.☆42Jun 23, 2025Updated last year
- Co-evolving policy actors and experience extractors for efficient experience-driven agent RL☆51May 12, 2026Updated 2 months ago
- ☆25Oct 23, 2025Updated 9 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- (CVPR 2026) Sampling Algorithm for paper "Ani3DHuman: Photorealistic 3D Human Animation with Self-guided Stochastic Sampling"☆23Jun 3, 2026Updated 2 months ago
- [ICML 2025] M-STAR (Multimodal Self-Evolving TrAining for Reasoning) Project. Diving into Self-Evolving Training for Multimodal Reasoning☆75Jul 13, 2025Updated last year
- ☆17Feb 4, 2026Updated 6 months ago
- Inverse Constitutional AI [ICLR 2025]: compressing pairwise preference data into a short constitution of principles.☆42May 6, 2026Updated 3 months ago
- When and How Much to Imagine: Adaptive Test-Time Scaling with World Models for Visual Spatial Reasoning☆18Jun 2, 2026Updated 2 months ago
- ☆35Feb 12, 2026Updated 5 months ago
- A Comprehensive survey on business use cases of AI that help them thrive in the digital economy☆13Oct 7, 2020Updated 5 years ago
- OccuBench: Evaluating AI Agents on Real-World Professional Tasks via Language World Models☆21Apr 14, 2026Updated 3 months ago
- Skill-Targeted Adaptive Training☆25Mar 12, 2026Updated 4 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [NeurIPS 2025] RL Tango: Reinforcing Generator and Verifier Together for Language Reasoning☆59Oct 23, 2025Updated 9 months ago
- Can We Predict Before Executing Machine Learning Agents?☆21Jul 7, 2026Updated 3 weeks ago
- 🏆 Ambassador Paper for Innovative Use of NLP for Building Educational Applications 2023: Is ChatGPT a Good Teacher Coach? Measuring Zero…☆14Jul 21, 2024Updated 2 years ago
- NeurIPS'24 - LLM Safety Landscape☆40Oct 21, 2025Updated 9 months ago
- Official code for Guiding Language Model Math Reasoning with Planning Tokens☆19Feb 29, 2024Updated 2 years ago
- Second Renaissance website 🌄☆11Jul 29, 2026Updated last week
- ProAct is a framework designed to enable Large Language Model (LLM) agents to perform accurate, multi-turn lookahead reasoning in interac…☆18Feb 11, 2026Updated 5 months ago
- [NO LONGER MAINTAINED, SUPERSEDED BY https://github.com/trueagi-io/pln-experimental and https://github.com/trueagi-io/PLN]. Probabilisti…☆16Sep 20, 2025Updated 10 months ago
- Official implementation of "Reward Prediction with Factorized World States"☆20Mar 11, 2026Updated 4 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Jointly Optimizing Large Language Models for Reasoning and Self-Refinement☆15Apr 22, 2026Updated 3 months ago
- Geometric Problem Solving Integrating FormalGeo Symbolic System and Hypergraph Neural Network.☆16Sep 23, 2025Updated 10 months ago
- [ICML 2025] Closed-Loop Long-Horizon Robotic Planning via Equilibrium Sequence Modeling☆13May 5, 2025Updated last year
- [ICML 2023] "Data Efficient Neural Scaling Law via Model Reusing" by Peihao Wang, Rameswar Panda, Zhangyang Wang☆14Jan 4, 2024Updated 2 years ago
- ☆78Apr 26, 2026Updated 3 months ago
- This is the github to open source benchmark AdvancedIF, see LAMA L1387358RCRO☆37Nov 26, 2025Updated 8 months ago
- A Julia package for differentiating through expectations with Monte-Carlo estimates☆16Nov 25, 2024Updated last year
- WideSearch: Benchmarking Agentic Broad Info-Seeking☆148Oct 9, 2025Updated 9 months ago
- ☆11Jul 30, 2024Updated 2 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- A gem for converting between hiragana, katakana, and romaji alphabets for the Japanese language☆15Jan 18, 2023Updated 3 years ago
- This is a AUTOSAR documents specific retriever based on LLM and RAG.☆16Nov 12, 2024Updated last year
- kNN-TL: k-Nearest-Neighbor Transfer Learning for Low-Resource Neural Machine Translation (ACL2023)☆11Jul 26, 2023Updated 3 years ago
- Minsk in VB☆11May 10, 2022Updated 4 years ago
- In this model I have created a basic AI chatbot Interface with External plugin abilities; with visual basic An Interface AI_Contracts en…☆10May 2, 2021Updated 5 years ago
- Dockerfile for AmigaOS Cross-Compiler Toolchain☆11Mar 12, 2018Updated 8 years ago
- The code repo for the paper "Differentiable Evolutionary Reinforcement Learning"☆18Jan 6, 2026Updated 7 months ago