[ICML 2026] An Evaluation Suite for Chain-of-Thought Controllability
☆53Mar 10, 2026Updated 5 months ago
Alternatives and similar repositories for CoTControl
Users that are interested in CoTControl are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code used for "Training Agents to Self-Report Misbehavior"☆18Feb 27, 2026Updated 6 months ago
- Codebase for the work “Why Does Self-Distillation (Sometimes) Degrade the Reasoning Capability of LLMs?”☆75Apr 14, 2026Updated 4 months ago
- Generalizing from SIMPLE to HARD Visual Reasoning: Can We Mitigate Modality Imbalance in VLMs?☆19Jun 3, 2025Updated last year
- Open-source RL Framework with Online Teacher-Student Distillation☆22Mar 5, 2026Updated 5 months ago
- ☆44Feb 18, 2026Updated 6 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆19Apr 28, 2026Updated 4 months ago
- A curated list of resources on on-policy distillation☆25Apr 13, 2026Updated 4 months ago
- ☆31Jul 1, 2026Updated 2 months ago
- Evaluating Agent Safety in Realistic, High-Risk Simulations☆34Aug 10, 2026Updated 3 weeks ago
- ☆18Mar 10, 2026Updated 5 months ago
- [ICML 2026] Official implementation for paper "Unsafer in Many Turns: Benchmarking and Defending Multi-Turn Safety Risks in Tool-Using Ag…☆35Jul 31, 2026Updated last month
- [ICLR26] Safety Mirage: How Spurious Correlations Undermine VLM Safety Fine-Tuning and Can Be Mitigated by Machine Unlearning☆21Apr 16, 2026Updated 4 months ago
- ☆26Jan 5, 2026Updated 7 months ago
- ☆47Sep 30, 2025Updated 11 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [ICLR 2026] Seeing Through Words: Controlling Visual Retrieval Quality with Language Models☆28Mar 19, 2026Updated 5 months ago
- Official code for the paper: "Multi-User Large Language Model Agents"☆34Updated this week
- This is a collection of awesome papers I have read (carefully or roughly) in the fields of computer vision, machine learning, pattern rec…☆31Aug 8, 2024Updated 2 years ago
- Repository for the paper: "TiC-LM: A Web-Scale Benchmark for Time-Continual LLM Pretraining" ACL Oral 2025☆24Apr 19, 2026Updated 4 months ago
- [CVPR 2026] Machine Mental Imagery: Empower Multimodal Reasoning with Latent Visual Tokens☆297Aug 2, 2025Updated last year
- ☆67May 21, 2025Updated last year
- [ACL 2025] Data and Code for Paper VLSBench: Unveiling Visual Leakage in Multimodal Safety☆62Jul 21, 2025Updated last year
- ☆17Updated this week
- ☆40May 7, 2026Updated 3 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Topology Distillation for Recommender System (KDD'21)☆13Sep 2, 2021Updated 5 years ago
- ☆21Apr 3, 2026Updated 5 months ago
- Experiments Notebook of "Understanding the Skill Gap in Recurrent Language Models: The Role of the Gather-and-Aggregate Mechanism"☆17Apr 30, 2025Updated last year
- Jailbreak Evo☆23Jun 2, 2025Updated last year
- A unified framework for vision-language environments with Gymnasium-compatible interface☆38Mar 17, 2026Updated 5 months ago
- Official implementation of Tabular Transfer Learning via Prompting LLMs (COLM 2024).☆13Aug 6, 2024Updated 2 years ago
- [CVPR Findings 2026] HoliSafe: Holistic Safety Benchmarking and Modeling for Vision-Language Model☆17Mar 8, 2026Updated 5 months ago
- ☆41Aug 11, 2026Updated 3 weeks ago
- [ICML 2026] Set Diffusion: Interpolating Token Orderings between Autoregression and Diffusion for Fast and Flexible Decoding☆26Aug 9, 2026Updated 3 weeks ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Repository for "Training Language Models To Explain Their Own Computations"☆37Jul 7, 2026Updated last month
- [ICLR 2025] Understanding and Enhancing Safety Mechanisms of LLMs via Safety-Specific Neuron☆36Apr 30, 2025Updated last year
- The official implementation of the paper "AgentLAB: Benchmarking LLM Agents against Long-Horizon Attacks"☆30Jun 1, 2026Updated 3 months ago
- ☆36Jun 13, 2025Updated last year
- [ACL 2026] Render-of-Thought: Rendering Textual Chain-of-Thought as Images for Visual Latent Reasoning☆94Jan 22, 2026Updated 7 months ago
- SkillJect: Automating Stealthy Skill-Based Prompt Injection for Coding Agents with Trace-Driven Closed-Loop Refinement☆77Jun 11, 2026Updated 2 months ago
- ☆180Aug 15, 2025Updated last year