[ICLR 2026] LongRLVR: Long-Context Reinforcement Learning Requires Verifiable Context Rewards.
☆19Mar 16, 2026Updated 5 months ago
Alternatives and similar repositories for LongRLVR
Users that are interested in LongRLVR are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official implementation of TDC.☆15Jul 22, 2025Updated last year
- Fine-Tuning Pre-trained Transformers into Decaying Fast Weights☆20Oct 9, 2022Updated 3 years ago
- Official implementation of Self-Taught Agentic Long Context Understanding (ACL 2025).☆14Sep 22, 2025Updated 11 months ago
- Gemini evaluations☆26Feb 19, 2026Updated 6 months ago
- [ICML 2025 Spotlight] RAPID: Long-Context Inference with Retrieval-Augmented Speculative Decoding☆22Mar 2, 2025Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Official implementation of MATPO: Multi-Agent Tool-Integrated Policy Optimization.☆83Oct 31, 2025Updated 9 months ago
- This repository includes the implementation and results of the paper "ChatGPT is fun, but it is not funny! Humor is still challenging Lar…☆13Jul 13, 2023Updated 3 years ago
- ☆11Aug 10, 2024Updated 2 years ago
- [ACL26 Findings] LiCoMemory: Lightweight and Cognitive Agentic Memory for Efficient Long-Term Reasoning☆50Jan 6, 2026Updated 7 months ago
- The official implementation of LIFT: Improving Long Context Understanding of Large Language Models through Long Input Fine-Tuning☆15Mar 14, 2025Updated last year
- ☆14Feb 28, 2024Updated 2 years ago
- [ACL 2026] VGPO: Visually-Guided Policy Optimization for Multimodal Reasoning☆33Apr 14, 2026Updated 4 months ago
- Code and data for EMNLP2019 Paper "Uncover the Ground-Truth Relations in Distant Supervision: A Neural Expectation-Maximization Framework…☆10May 24, 2020Updated 6 years ago
- SDPG: Self-Distilled Policy Gradient☆54Jun 15, 2026Updated 2 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆14Feb 26, 2024Updated 2 years ago
- ☆17Jul 31, 2025Updated last year
- Code for paper "ProgGen: Generating Named Entity Recognition Datasets Step-by-step with Self-Reflexive Large Language Models"☆17Mar 29, 2024Updated 2 years ago
- ☆16Aug 11, 2025Updated last year
- Reproducing R1 for Code with Reliable Rewards☆13Apr 9, 2025Updated last year
- Search Self-Play: Pushing the Frontier of Agent Capability without Supervision☆26Dec 30, 2025Updated 7 months ago
- knrm文本相似度☆10Aug 1, 2020Updated 6 years ago
- A comprehensive and efficient long-context model evaluation framework☆31Feb 25, 2026Updated 6 months ago
- HiPRAG (Hierarchical Process Rewards for Efficient Agentic Retrieval Augmented Generation) is a reinforcement learning method designed fo…☆26Oct 10, 2025Updated 10 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- All-in-one benchmarking platform for evaluating LLM.☆15Nov 12, 2025Updated 9 months ago
- code for "Fine-grained Entity Typing via Label Reasoning" EMNLP2021☆13May 27, 2022Updated 4 years ago
- code for paper Hierarchical Retrieval-Augmented Generation Model with Rethink for Multi-hop Question Answering☆14Aug 13, 2024Updated 2 years ago
- ☆15Aug 12, 2022Updated 4 years ago
- ☆25Jan 22, 2024Updated 2 years ago
- Code for paper: Weakly- and Semi-supervised Evidence Extraction☆15Apr 12, 2021Updated 5 years ago
- Agent-Omit: Training Efficient LLM Agents for Adaptive Thought and Observation Omission via Reinforcement Learning☆33May 11, 2026Updated 3 months ago
- ☆15Jan 9, 2018Updated 8 years ago
- Official code for "Flatten Graphs as Sequences: Transformers are scalable graph generators" (NeurIPS 2025)☆18Oct 17, 2025Updated 10 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Image Classification Tutorial: ConvNext--> 98.8% on CIFAR10 + 92.4% on CIFAR100; ResNet18 -- 95.6% on CIFAR10 + 79.1% on CIFAR100☆15Jun 2, 2025Updated last year
- Codebase for multilingual neural machine translation☆13Nov 24, 2022Updated 3 years ago
- ☆17Apr 18, 2024Updated 2 years ago
- Official repo for "PAPO: Perception-Aware Policy Optimization for Multimodal Reasoning"☆158Feb 4, 2026Updated 6 months ago
- Official implementation of 'RiskPO: Risk-based Policy Optimization via Verifiable Reward for LLM Post-Training', accepted by ICLR 2026☆19Oct 15, 2025Updated 10 months ago
- ACM Multimedia 2023 (Oral) - RTQ: Rethinking Video-language Understanding Based on Image-text Model☆15Apr 7, 2026Updated 4 months ago
- Make math learning simpler, starting with Nano Math plus , too!☆45Jul 16, 2026Updated last month