[ACL 2026] Dissecting Failure Dynamics in Large Language Model Reasoning
☆19Apr 17, 2026Updated 5 months ago
Alternatives and similar repositories for GUARD
Users that are interested in GUARD are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ACL2026] UCAS: Uncertainty-aware Advantage Shaping for RLVR☆31Apr 14, 2026Updated 5 months ago
- Data and Code Repository for “STRIDE-ED: A Strategy-Grounded Stepwise Reasoning Framework for Empathetic Dialogue Systems”☆17Apr 17, 2026Updated 5 months ago
- ActorMind: Emulating Human Actor Reasoning for Speech Role-Playing - ACL Findings 2026☆26Jul 15, 2026Updated 2 months ago
- [NeurIPS 2025] SYMPHONY: Synergistic Multi-agent Planning with Heterogeneous Language Model Assembly☆17Oct 22, 2025Updated 11 months ago
- Implementation for paper Flattery in Motion: Benchmarking and Analyzing Sycophancy in Video-LLMs, which is accepted by ACL 2026 (main con…☆16Oct 10, 2025Updated 11 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Official Codebase of the ACL 2026 Oral paper "Rethinking Jailbreak Detection of Large Vision Language Models with Representational Contra…☆28Jun 25, 2026Updated 3 months ago
- This is the official repository for JailExpert☆22Sep 9, 2025Updated last year
- Curriculum-RLAIF is a data-centric curriculum learning framework for reward model training in RLAIF-based LLM alignment☆23Apr 18, 2026Updated 5 months ago
- Prototype Conditioned Generative Replay for Continual Learning in NLP - NAACL 2025☆25Jul 28, 2026Updated 2 months ago
- videoPro: Adaptive Program Reasoning for Long Video Understanding☆46Apr 15, 2026Updated 5 months ago
- ACL26 Long Paper☆19Jul 4, 2026Updated 2 months ago
- Sparse Adapter Fusion for Continual Learning in NLP - EACL 2026☆16Apr 9, 2026Updated 5 months ago
- This is the official repo for the paper "General365: Benchmarking General Reasoning in LLMs under High Difficulty and Diversity".☆87Apr 14, 2026Updated 5 months ago
- [Paper][EMNLP 2025] Enrich-on-Graph: Query-Graph Alignment for Complex Reasoning with LLM Enriching☆35Feb 8, 2026Updated 7 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [🏆AAAI'25] Official Repo for ChemVLM: Exploring the Power of Multimodal Large Language Models in Chemistry Area.☆93Apr 14, 2026Updated 5 months ago
- [ACL 2026] Context-Agent: Dynamic Discourse Trees for Non-Linear Dialogue☆24Apr 14, 2026Updated 5 months ago
- The implementation of ACL 2026 paper "Rethinking entropy interventions in rlvr: An entropy change perspective"☆27Jul 19, 2026Updated 2 months ago
- SE-Agent is a self-evolution framework for LLM Code agents. It enables trajectory-level evolution to exchange information across reasonin…☆289Sep 23, 2025Updated last year
- MinT-2M: Long-context training system for resident-prefix GRPO☆47Jul 24, 2026Updated 2 months ago
- [ICLR 2026] Thinking on the Fly: Test-Time Reasoning Enhancement via Latent Thought Policy Optimization☆32Mar 6, 2026Updated 6 months ago
- 本项目综合运用d3、echarts来完成可视化工作,实现了对nba两场比赛的可视化数据分析,包括球员运动轨迹、个人数据、传球次数以及得分位置等多种可交互式图表。通过可视化方法,我们能够进一步深入分析球队的具体情况,便于制定更佳的战术。☆15Dec 19, 2022Updated 3 years ago
- Tuning-Free Image Editing with Fidelity and Editability via Unified Latent Diffusion Model☆13Dec 29, 2024Updated last year
- [REALM25 @ ACL25] - "StateAct" Official Paper Repo (SOTA LLM Agent)☆19Aug 7, 2026Updated last month
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆12Sep 8, 2023Updated 3 years ago
- [ACM MM25] LongWriter-V: Enabling Ultra-Long and High-Fidelity Generation in Vision-Language Models☆25Mar 29, 2025Updated last year
- my algorithm course experiment in nwpu☆15Sep 30, 2022Updated 4 years ago
- DataMosaic: Explainable and Verifiable Document-Based Data Analytics☆20Jun 30, 2025Updated last year
- [AAAI2024] Debiasing Multimodal Sarcasm Detection with Contrastive Learning☆17Jan 5, 2024Updated 2 years ago
- The official implementation for "Mitigating Overthinking in Large Reasoning Models via Manifold Steering"☆15May 29, 2025Updated last year
- IntentVCNet: Bridging Spatio-Temporal Gaps for Intention-Oriented Controllable Video Captioning☆19Aug 16, 2025Updated last year
- Project page and release repository for PhysEditWorld, a dataset toward physics-editable world models.☆18Jul 2, 2026Updated 3 months ago
- ☆26Jun 3, 2026Updated 4 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Official repo for PlanViz: Evaluating Planning-Oriented Image Generation and Editing for Computer-Use Tasks☆17Feb 17, 2026Updated 7 months ago
- [ICLR 2026 Oral] RAIN-Merging☆16Mar 9, 2026Updated 6 months ago
- [CVPR'25] Attention IoU: Examining Biases in CelebA using Attention Maps☆13Mar 26, 2025Updated last year
- Smart Crop Try, 努力仿照 “全能扫描王”☆16Feb 23, 2021Updated 5 years ago
- An official implementation of Random Policy Valuation is Enough for LLM Reasoning with Verifiable Rewards☆36Oct 3, 2025Updated last year
- [NeurIPS 2026] AgentForesight: Online Auditing for Early Failure Prediction in Multi-Agent Systems☆19May 12, 2026Updated 4 months ago
- The official repository of paper "Learn to Memorize: Optimizing LLM-based Agents with Adaptive Memory Framework".☆18Sep 22, 2025Updated last year