Official Repository of "[ICLR26] TRAPO: A Semi-Supervised Reinforcement Learning Framework for Boosting LLM Reasoning"
☆29Feb 6, 2026Updated 8 months ago
Alternatives and similar repositories for TRAPO
Users that are interested in TRAPO are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ACL 2025] RealHiTBench: A Comprehensive Realistic Hierarchical Table Benchmark for Evaluating LLM-Based Table Analysis☆27Aug 8, 2025Updated last year
- Code for CVPR 2024 paper: Positive-Unlabeled Learning by Latent Group-Aware Meta Disambiguation☆22May 19, 2024Updated 2 years ago
- [ICLR 2025 Spotlight] Realistic Evaluation of Deep Partial-Label Learning Algorithms☆15Feb 2, 2025Updated last year
- [EMNLP-2025] R1-Zero on ANY TASK☆33Nov 9, 2025Updated 11 months ago
- Beyond Myopia: Learning from Positive and Unlabeled Data through Holistic Predictive Trends [NeurIPS 2023]☆10Jan 28, 2024Updated 2 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Source code for NeurIPS 2022 paper SoLar☆30Dec 20, 2023Updated 2 years ago
- Official code of "ALIM: Adjusting Label Importance Mechanism for Noisy Partial Label Learning"☆22Sep 25, 2023Updated 3 years ago
- [CVPR 2026] Official Implementation of DC-Merge: Improving Model Merging with Directional Consistency☆20Apr 29, 2026Updated 5 months ago
- To Trust Or Not To Trust Your Vision-Language Model's Prediction☆15May 30, 2025Updated last year
- ☆33Aug 21, 2025Updated last year
- Scaling Test-time Training for LLM Reasoning☆37Apr 14, 2026Updated 5 months ago
- ☆29Jul 16, 2024Updated 2 years ago
- Official code repository for the paper "ToMAP: Training Opponent-Aware LLM Persuaders with Theory of Mind"☆27Sep 25, 2025Updated last year
- [TOIS 2024] Target-constrained Bidirectional Planning for Generation of Target-oriented Proactive Dialogue☆14Oct 18, 2025Updated 11 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official implementation of Tabular Transfer Learning via Prompting LLMs (COLM 2024).☆13Aug 6, 2024Updated 2 years ago
- The official implementation of MaskGRPO: Consolidating Reinforcement Learning for Multimodal Discrete Diffusion Models. (ICLR 2026, arxiv…☆19Jan 27, 2026Updated 8 months ago
- The official implementation of the paper "A Dual-Space Framework for General Knowledge Distillation of Large Language Models".☆17Jan 4, 2026Updated 9 months ago
- ABC: Achieving Better Control of Multimodal Embeddings using VLMs [TMLR2025]☆20Aug 21, 2025Updated last year
- ☆14Mar 4, 2024Updated 2 years ago
- ICCV 2023 - AdaptGuard: Defending Against Universal Attacks for Model Adaptation☆11Dec 23, 2023Updated 2 years ago
- ☆11Jul 3, 2024Updated 2 years ago
- Exploring prompt tuning with pseudolabels for multiple modalities, learning settings, and training strategies.☆50Nov 8, 2024Updated last year
- Official implementation for SPA: A Graph Spectral Alignment Perspective for Domain Adaptation (NeurIPS 2023)☆18Dec 21, 2023Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆15Jan 14, 2026Updated 8 months ago
- ☆14Oct 31, 2022Updated 3 years ago
- ☆54Jun 3, 2026Updated 4 months ago
- Generate custom text files for dataloader within UDA methods☆14May 24, 2023Updated 3 years ago
- ☆51May 13, 2024Updated 2 years ago
- [NeurIPS 2025] TTRL: Test-Time Reinforcement Learning☆1,126Apr 15, 2026Updated 5 months ago
- [IJCAI 2023] ProMix: Combating Label Noise via Maximizing Clean Sample Utility☆80May 10, 2024Updated 2 years ago
- [ICLR 2024] Towards Elminating Hard Label Constraints in Gradient Inverision Attacks☆14Feb 6, 2024Updated 2 years ago
- Official implementation of "Automated Generation of Challenging Multiple-Choice Questions for Vision Language Model Evaluation" (CVPR 202…☆39May 26, 2025Updated last year
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- code released for our TIP 2021 paper "Adversarial Domain Adaptation with Prototype-based Normalized Output Conditioner"☆15May 24, 2023Updated 3 years ago
- [CVPR'25] Attention IoU: Examining Biases in CelebA using Attention Maps☆13Mar 26, 2025Updated last year
- Code and results accompanying our paper titled CHiLS: Zero-Shot Image Classification with Hierarchical Label Sets☆60Jun 4, 2023Updated 3 years ago
- ☆15Aug 5, 2024Updated 2 years ago
- Entropy-Driven GRPO with Guided Error Correction for Advantage Diversity☆22Aug 28, 2025Updated last year
- Official Repository of "Learning what reinforcement learning can't"☆89Dec 30, 2025Updated 9 months ago
- Official implementation of Latent-GRPO: reinforcement learning for vocabulary-space latent reasoning.☆21Aug 31, 2026Updated last month