ICML 2026: "ProRL: Effective Reinforcement Learning for Proactive Recommendation via Rectified Policy Gradient Estimation"
☆47Jun 12, 2026Updated 2 months ago
Alternatives and similar repositories for ProRL
Users that are interested in ProRL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆47May 22, 2026Updated 3 months ago
- ☆26May 17, 2026Updated 3 months ago
- ☆16Aug 27, 2025Updated last year
- A multimodal context reasoning approach that introduce the multi-view semantic alignment information via prefix tuning.☆15Sep 14, 2023Updated 2 years ago
- The pytorch implementation of the SAFE model presented in NAACL-Findings-2022☆17Mar 10, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ACL 2024] Masked Thought: Simply Masking Partial Reasoning Steps Can Improve Mathematical Reasoning Learning of Language Models☆27Jul 9, 2024Updated 2 years ago
- ☆17Dec 29, 2025Updated 8 months ago
- 电子科技大学高级计算机视觉课程的作业代码☆13Sep 5, 2020Updated 5 years ago
- EvoTest: Evolutionary Test-Time Learning for Self-Improving Agentic Systems (ICLR'26)☆28Nov 3, 2025Updated 9 months ago
- This is an official repository of paper "Refining Action Segmentation with Hierarchical Video Representations", which is accepted as a re…☆17Oct 11, 2021Updated 4 years ago
- ☆19Aug 29, 2024Updated 2 years ago
- 2022微信大数据挑战赛_rank12☆18Aug 18, 2022Updated 4 years ago
- SIGIR 2024, Session-based recommendation, Co-occurrence patterns of ID, Fine-grained preferences of Modality, Disentanglement learning.☆23Sep 5, 2024Updated last year
- Searching a High Performance Feature Extractor for Text Recognition Network. TPAMI 2022☆13Nov 25, 2022Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- A Network Integration Approach for Drug-Target Interaction Prediction☆13Apr 5, 2025Updated last year
- Graph neural network for predicting energy of known and hypothetical crystal structures☆10Jan 26, 2022Updated 4 years ago
- [NeurIPS 2023] "Combating Bilateral Edge Noise for Robust Link Prediction"☆11Nov 3, 2023Updated 2 years ago
- Data-efficient Fine-tuning for LLM-based Recommendation (SIGIR'24)☆40Feb 21, 2025Updated last year
- this is based on the paper Chain-of-Retrieval Augmented Generation☆15Mar 29, 2025Updated last year
- Automatically installs and configures XFCE, XRDP and variables for a one-script setup☆14Apr 14, 2021Updated 5 years ago
- DDI-LLM:💊 Drug-Drug Interaction Prediction: Experimenting With Large Language-Based Drug Information Embedding For Multi-View Representa…☆14Sep 15, 2023Updated 2 years ago
- ☆10Apr 11, 2022Updated 4 years ago
- Klear-Reasoner: Advancing Reasoning Capability via Gradient-Preserving Clipping Policy Optimization☆82Dec 25, 2025Updated 8 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- The official implementation of ACL'24 paper: Synergistic Interplay between Search and Large Language Models for Information Retrieval.☆36Jun 6, 2024Updated 2 years ago
- Experiments codes for WSDM '24 paper "MultiFS: Automated Multi-Scenario Feature Selection in Deep Recommender Systems"☆11May 31, 2024Updated 2 years ago
- Code for paper "Towards Open-World Recommendation with Knowledge Augmentation from Large Language Models"☆112Nov 14, 2024Updated last year
- Code for 'Diff-MSR: A Diffusion Model Enhanced Paradigm for Cold-Start Multi-Scenario Recommendation' accepted to WSDM 2024☆15Aug 1, 2025Updated last year
- [IJCAI 2023] Graph Propagation Transformer for Graph Representation Learning