[ACL2026] UCAS: Uncertainty-aware Advantage Shaping for RLVR
☆32Apr 14, 2026Updated 5 months ago
Alternatives and similar repositories for UCAS
Users that are interested in UCAS are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- MemoryDial☆15Mar 10, 2026Updated 6 months ago
- Data and Code Repository for “STRIDE-ED: A Strategy-Grounded Stepwise Reasoning Framework for Empathetic Dialogue Systems”☆17Apr 17, 2026Updated 4 months ago
- [ACL26 Findings] TopoDIM: One-shot Topology Generation of Diverse Interaction Modes for Multi-Agent Systems☆19Jan 19, 2026Updated 7 months ago
- [ACL 2026] Dissecting Failure Dynamics in Large Language Model Reasoning☆18Apr 17, 2026Updated 4 months ago
- This is the official repository for JailExpert☆23Sep 9, 2025Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- [🏆CVPR'26] Official Repo for IAG: Input-aware Backdoor Attack on VLM-based Visual Grounding☆34Jun 2, 2026Updated 3 months ago
- ActorMind: Emulating Human Actor Reasoning for Speech Role-Playing - ACL Findings 2026☆25Jul 15, 2026Updated 2 months ago
- [NeurIPS 2025] SYMPHONY: Synergistic Multi-agent Planning with Heterogeneous Language Model Assembly☆17Oct 22, 2025Updated 10 months ago
- Implementation for paper Flattery in Motion: Benchmarking and Analyzing Sycophancy in Video-LLMs, which is accepted by ACL 2026 (main con…☆16Oct 10, 2025Updated 11 months ago
- Curriculum-RLAIF is a data-centric curriculum learning framework for reward model training in RLAIF-based LLM alignment☆23Apr 18, 2026Updated 4 months ago
- Official Codebase of the ACL 2026 Oral paper "Rethinking Jailbreak Detection of Large Vision Language Models with Representational Contra…☆28Jun 25, 2026Updated 2 months ago
- This is the official repo for the paper "General365: Benchmarking General Reasoning in LLMs under High Difficulty and Diversity".☆89Apr 14, 2026Updated 5 months ago
- This is the official repo for the paper "AMO-Bench: Large Language Models Still Struggle in High School Math Competitions".☆172Feb 6, 2026Updated 7 months ago
- [Paper][EMNLP 2025] Enrich-on-Graph: Query-Graph Alignment for Complex Reasoning with LLM Enriching☆35Feb 8, 2026Updated 7 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ACL 2026] Context-Agent: Dynamic Discourse Trees for Non-Linear Dialogue☆24Apr 14, 2026Updated 5 months ago
- [🏆AAAI'25] Official Repo for ChemVLM: Exploring the Power of Multimodal Large Language Models in Chemistry Area.☆93Apr 14, 2026Updated 5 months ago
- [AAAI2024] Debiasing Multimodal Sarcasm Detection with Contrastive Learning☆17Jan 5, 2024Updated 2 years ago
- The implementation of ACL 2026 paper "Rethinking entropy interventions in rlvr: An entropy change perspective"☆27Jul 19, 2026Updated last month
- The implementation of ACL main 2026 paper "ReCreate: Reasoning and Creating Domain Agents Driven by Experience"☆167Apr 29, 2026Updated 4 months ago
- Residual vector quantization for KV cache compression in large language model☆12Oct 22, 2024Updated last year
- ☆13May 31, 2023Updated 3 years ago
- 本项目综合运用d3、echarts来完成可视化工作,实现了对nba两场比赛的可视化数据分析,包括球员运动轨迹、个人数据、传球次数以及得分位置等多种可交互式图表。通过可视化方法,我们能够进一步深入分析球队的具体情况,便于制定更佳的战术。☆15Dec 19, 2022Updated 3 years ago
- Tuning-Free Image Editing with Fidelity and Editability via Unified Latent Diffusion Model☆13Dec 29, 2024Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆14Feb 14, 2024Updated 2 years ago
- MinT-2M: Long-context training system for resident-prefix GRPO☆46Jul 24, 2026Updated last month
- [NIPS2025] A decentralized, RAG-enhanced multi-agent framework for LLMs with dynamic task routing and agent evolution.☆68Oct 2, 2025Updated 11 months ago
- Project page and release repository for PhysEditWorld, a dataset toward physics-editable world models.☆18Jul 2, 2026Updated 2 months ago
- ☆26Jun 3, 2026Updated 3 months ago
- Official repo for PlanViz: Evaluating Planning-Oriented Image Generation and Editing for Computer-Use Tasks☆17Feb 17, 2026Updated 6 months ago
- An official implementation of Random Policy Valuation is Enough for LLM Reasoning with Verifiable Rewards☆36Oct 3, 2025Updated 11 months ago
- The official repository for Trust-Region Adaptive Policy Optimization (TRAPO) – a novel hybrid framework designed to enhance large langua…☆16Mar 2, 2026Updated 6 months ago
- EMNLP'2024: Knowledge Verification to Nip Hallucination in the Bud☆23Mar 10, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Prompt Tuning based Adapter for Vision-Language Model Adaption☆16Sep 1, 2023Updated 3 years ago
- New API official website documentation☆25Nov 30, 2025Updated 9 months ago
- Official implementation of 'MeanSE: Efficient Generative Speech Enhancement with Mean Flows'☆22Oct 11, 2025Updated 11 months ago
- [WWW'26 Oral] ColorBench: a graph-structured benchmark for complex, long-horizon tasks in mobile GUI agents.☆16Apr 13, 2026Updated 5 months ago
- Efficient Expert Pruning for Sparse Mixture-of-Experts Language Models: Enhancing Performance and Reducing Inference Costs☆25Nov 11, 2025Updated 10 months ago
- Source code of NAACL 2025 Findings "Scaling Up Membership Inference: When and How Attacks Succeed on Large Language Models"☆16Dec 16, 2025Updated 9 months ago
- Official implementation of the paper "Robust and Resource-Efficient Data-Free Knowledge Distillation by Generative Pseudo Replay" (AAAI-2…☆18May 5, 2022Updated 4 years ago