Repo for paper "Rethinking Generalization in Reasoning SFT: A Conditional Analysis on Optimization, Data, and Model Capability"
☆108Apr 23, 2026Updated 2 months ago
Alternatives and similar repositories for rethink_sft_generalization
Users that are interested in rethink_sft_generalization are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official Repo of Your Agent May Misevolve: Emergent Risks in Self-evolving LLM Agents☆90Jun 2, 2026Updated last month
- ATBench: A Diverse and Realistic Agent Trajectory Benchmark for Safety Evaluation and Diagnosis☆33Jul 10, 2026Updated last week
- A Diagnostic Guardrail Framework for AI Agent Safety and Security☆669Jun 8, 2026Updated last month
- All-in-One Safety Evaluation Framwork☆51Jul 15, 2026Updated last week
- Diagnostic Framework for LLMs and MLLMs☆39Mar 2, 2026Updated 4 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆68Jul 14, 2025Updated last year
- 🧨 TradeTrap: Are LLM-based Trading Agents Truly Reliable and Faithful?☆80Nov 27, 2025Updated 7 months ago
- [ACL 2025] Data and Code for Paper VLSBench: Unveiling Visual Leakage in Multimodal Safety☆62Jul 21, 2025Updated last year
- Jointly Optimizing Large Language Models for Reasoning and Self-Refinement☆15Apr 22, 2026Updated 3 months ago
- This repo is the official implementation of “Are Your Agents Upward Deceivers?”. The paper is accepted by ICML 2026.☆24Dec 15, 2025Updated 7 months ago
- Official Repository of "Taming Masked Diffusion Language Models via Consistency Trajectory Reinforcement Learning with Fewer Decoding Ste…☆28Mar 9, 2026Updated 4 months ago
- [ACL 2026 Main] Official Repo for Paper "Which Reasoning Trajectories Teach Students to Reason Better? A Simple Metric of Informative Ali…☆17Jul 1, 2026Updated 3 weeks ago
- ☆45Mar 30, 2026Updated 3 months ago
- ☆11Oct 25, 2024Updated last year
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- On Policy Distillation Build on top of Verl☆92May 25, 2026Updated last month
- [NeurIPS 2025] Official repository of RiOSWorld: Benchmarking the Risk of Multimodal Computer-Use Agents☆123Dec 2, 2025Updated 7 months ago
- [ICML 2026] Hybrid Policy Distillation (HPD) is a practical distillation framework for reasoning-oriented language models. This repositor…☆24Apr 24, 2026Updated 2 months ago
- [AAAI 2026] Data and Code for Paper IS-Bench: Evaluating Interactive Safety of VLM-Driven Embodied Agents in Daily Household Tasks☆47Nov 24, 2025Updated 7 months ago
- ☆26Feb 20, 2026Updated 5 months ago
- ReplayCode — first open-source rebuild of Claude Code that actually runs. Built from decompiled source with Node.js/esbuild☆20Apr 1, 2026Updated 3 months ago
- Socratic-Zero is a fully autonomous framework that generates high-quality training data for mathematical reasoning☆37Oct 26, 2025Updated 8 months ago
- A Teacher–Student Cooperation Framework to Synthesize Student-Consistent SFT Data☆33May 1, 2026Updated 2 months ago
- Official implementation of Selective Entropy Regularization (SIREN), proposed by paper 'Rethinking Entropy Regularization in Large Reason…☆32Dec 10, 2025Updated 7 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- The repository of the paper "REEF: Representation Encoding Fingerprints for Large Language Models," aims to protect the IP of open-source…☆79Jan 16, 2025Updated last year
- 😎 A Survey of Efficient Reasoning for Large Reasoning Models: Language, Multimodality, Agent, and Beyond☆355Jan 22, 2026Updated 6 months ago
- [ACL 2024] CodeAttack: Revealing Safety Generalization Challenges of Large Language Models via Code Completion☆61Oct 1, 2025Updated 9 months ago
- ☆36Apr 13, 2026Updated 3 months ago
- PyTorch implementation of the paper "Discovering and Explaining the Representation Bottleneck of DNNs" (ICLR 2022 Oral)☆37Oct 30, 2024Updated last year
- We introduce BabyVision, a benchmark revealing the infancy of AI vision.☆231Jan 13, 2026Updated 6 months ago
- Some thoughts about writing scientific papers☆23Nov 8, 2024Updated last year
- Official repository of DARE: Diffusion Large Language Models Alignment and Reinforcement Executor☆213Updated this week
- A novel approach to improve the safety of large language models, enabling them to transition effectively from unsafe to safe state.☆72May 22, 2025Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- (ICLR 2026 🔥) Code for "The Devil behind the mask: An emergent safety vulnerability of Diffusion LLMs"☆79Feb 9, 2026Updated 5 months ago
- [ICML 2024] Code release for "On the Emergence of Cross-Task Linearity in Pretraining-Finetuning Paradigm"☆11Feb 20, 2025Updated last year
- Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recipe☆835Jun 29, 2026Updated 3 weeks ago
- PaperPub is an academic arena where diverse AI Agents read papers daily, pick apart each other's arguments, and fiercely debate.☆43Jun 12, 2026Updated last month
- instruction-following benchmark for large reasoning models☆49Apr 19, 2026Updated 3 months ago
- Official Implementation of MARS☆30Apr 21, 2026Updated 3 months ago
- Beyond SFT-to-RL: Pre-alignment via Black-BoxOn-Policy Distillation for Multimodal RL☆96May 6, 2026Updated 2 months ago