Official repository for paper "DeepCritic: Deliberate Critique with Large Language Models"
☆41Jun 24, 2025Updated last year
Alternatives and similar repositories for DeepCritic
Users that are interested in DeepCritic are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ACL 2025] Cautious Next Token Prediction☆16Jul 24, 2025Updated 11 months ago
- [ICLR 2025] Code&Data for the paper "Super(ficial)-alignment: Strong Models May Deceive Weak Models in Weak-to-Strong Generalization"☆15Jun 21, 2024Updated 2 years ago
- (ICLR 2025) AgentRefine: Enhancing Agent Generalization through Refinement Tuning☆20Nov 22, 2025Updated 8 months ago
- Official PyTorch implementation for "Effective and Efficient Masked Image Generation Models"☆35Apr 8, 2025Updated last year
- [ICLR'26] RM-R1: Unleashing the Reasoning Potential of Reward Models☆167Jun 26, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Reproducing R1 for Code with Reliable Rewards☆13Apr 9, 2025Updated last year
- This repository contains the code for the paper “Neuro-Symbolic Query Compiler”, accepted to the Findings of ACL 2025.☆17Oct 20, 2025Updated 9 months ago
- ☆20Nov 3, 2024Updated last year
- Implementation of Prompt-to-Prompt Image Editing with Cross Attention Control☆16Apr 5, 2023Updated 3 years ago
- TACO: TFBS-Aware Cis-Regulatory Element Optimization☆23Aug 1, 2025Updated 11 months ago
- ☆65Mar 30, 2026Updated 3 months ago
- [ICCV 2023 Workshop] The Official Implementation of The First Prize Solution for RVOS Competition☆14Jan 1, 2024Updated 2 years ago
- Official PyTorch implementation for "Diffusion Models and Semi-Supervised Learners Benefit Mutually with Few Labels"☆96Jan 17, 2024Updated 2 years ago
- "Good scientific writing is not a matter of life and death; it is much more serious than that."☆14Apr 29, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆15Sep 24, 2023Updated 2 years ago
- [Preprint] Why is the State of Neural Network Pruning so Confusing? On the Fairness, Comparison Setup, and Trainability in Network Prunin…☆41Sep 9, 2025Updated 10 months ago
- UFT: Unifying Supervised and Reinforcement Fine-Tuning☆31Jun 30, 2025Updated last year
- We release our code and data for SEAS in this repository.☆21Dec 23, 2024Updated last year
- SynthRL: Scaling Visual Reasoning with Verifiable Data Synthesis☆70Jul 24, 2025Updated 11 months ago
- ☆16May 18, 2026Updated 2 months ago
- [TKDE 2024, CIKM 2022] SLA²P: Self-supervised Anomaly Detection with Adversarial Perturbation.☆39Dec 26, 2024Updated last year
- [ICLR 2025] A Comprehensive Framework for Developing and Evaluating Multimodal Role-Playing Agents☆101Feb 2, 2026Updated 5 months ago
- [ICLR 2026] Do Not Let Low-Probability Tokens Over-Dominate in RL for LLMs☆46May 20, 2025Updated last year
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- ☆16Mar 6, 2025Updated last year
- ☆20Jan 18, 2026Updated 6 months ago
- Scratchpad/Chain-of-Thought Prompts☆12Jun 6, 2022Updated 4 years ago
- ☆16Mar 22, 2025Updated last year
- ☆20Nov 27, 2022Updated 3 years ago
- ☆50Aug 5, 2025Updated 11 months ago
- HierSearch: A Hierarchical Enterprise Deep Search Framework Integrating Local and Web Searches☆40Oct 9, 2025Updated 9 months ago
- Code for the paper "RAP: Robustness-Aware Perturbations for Defending against Backdoor Attacks on NLP Models" (EMNLP 2021)☆25Oct 21, 2021Updated 4 years ago
- The code and data for the paper JiuZhang3.0☆49May 26, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [EMNLP 2024] A Multi-level Hallucination Diagnostic Benchmark for Tool-Augmented Large Language Models.☆22Sep 23, 2024Updated last year
- RL with Experience Replay☆59Jul 27, 2025Updated 11 months ago
- Code for ACL2024 paper - Adversarial Preference Optimization (APO).☆54Jun 3, 2024Updated 2 years ago
- Anti exploration in offline reinforcement learning☆11May 17, 2021Updated 5 years ago
- Code for the paper "Rethinking Stealthiness of Backdoor Attack against NLP Models" (ACL-IJCNLP 2021)☆24Dec 9, 2021Updated 4 years ago
- The official repository of paper "Pass@k Training for Adaptively Balancing Exploration and Exploitation of Large Reasoning Models''☆113Aug 15, 2025Updated 11 months ago
- Code & data for ICLR 2024 spotlight paper: 🍯MUSTARD: Mastering Uniform Synthesis of Theorem and Proof Data☆43May 29, 2024Updated 2 years ago