Repo for paper "Rethinking Generalization in Reasoning SFT: A Conditional Analysis on Optimization, Data, and Model Capability"
☆109Aug 15, 2026Updated last month
Alternatives and similar repositories for rethink_sft_generalization
Users that are interested in rethink_sft_generalization are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ATBench: A Diverse and Realistic Agent Trajectory Benchmark for Safety Evaluation and Diagnosis☆40Updated this week
- A Diagnostic Guardrail Framework for AI Agent Safety and Security☆700Jun 8, 2026Updated 3 months ago
- All-in-One Safety Evaluation Framwork☆55Aug 12, 2026Updated last month
- Diagnostic Framework for LLMs and MLLMs☆39Mar 2, 2026Updated 6 months ago
- ☆68Jul 14, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 🧨 TradeTrap: Are LLM-based Trading Agents Truly Reliable and Faithful?☆83Nov 27, 2025Updated 9 months ago
- [ACL 2025] Data and Code for Paper VLSBench: Unveiling Visual Leakage in Multimodal Safety☆62Jul 21, 2025Updated last year
- Jointly Optimizing Large Language Models for Reasoning and Self-Refinement☆14Sep 14, 2026Updated last week
- This repo is the official implementation of “Are Your Agents Upward Deceivers?”. The paper is accepted by ICML 2026.☆24Dec 15, 2025Updated 9 months ago
- Official Repository of "Taming Masked Diffusion Language Models via Consistency Trajectory Reinforcement Learning with Fewer Decoding Ste…☆28Mar 9, 2026Updated 6 months ago
- [ACL 2026 Main] Official Repo for Paper "Which Reasoning Trajectories Teach Students to Reason Better? A Simple Metric of Informative Ali…☆16Jul 1, 2026Updated 2 months ago
- ☆45Mar 30, 2026Updated 5 months ago
- ☆11Oct 25, 2024Updated last year
- Trace origins, shared sources, and contamination risk☆29May 27, 2026Updated 3 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- On Policy Distillation Build on top of Verl☆100Sep 3, 2026Updated 2 weeks ago
- The official code of FineRMoE.☆23Mar 17, 2026Updated 6 months ago
- [NeurIPS 2025] Official repository of RiOSWorld: Benchmarking the Risk of Multimodal Computer-Use Agents☆124Dec 2, 2025Updated 9 months ago
- [ICML 2026] Hybrid Policy Distillation (HPD) is a practical distillation framework for reasoning-oriented language models. This repositor…☆25Apr 24, 2026Updated 4 months ago
- [AAAI 2026] Data and Code for Paper IS-Bench: Evaluating Interactive Safety of VLM-Driven Embodied Agents in Daily Household Tasks☆52Nov 24, 2025Updated 9 months ago
- ReplayCode — first open-source rebuild of Claude Code that actually runs. Built from decompiled source with Node.js/esbuild☆20Apr 1, 2026Updated 5 months ago
- Socratic-Zero is a fully autonomous framework that generates high-quality training data for mathematical reasoning☆37Oct 26, 2025Updated 10 months ago
- ☆28Feb 20, 2026Updated 7 months ago
- A Teacher–Student Cooperation Framework to Synthesize Student-Consistent SFT Data☆37May 1, 2026Updated 4 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official implementation of Selective Entropy Regularization (SIREN), proposed by paper 'Rethinking Entropy Regularization in Large Reason…☆32Dec 10, 2025Updated 9 months ago
- The repository of the paper "REEF: Representation Encoding Fingerprints for Large Language Models," aims to protect the IP of open-source…☆80Jan 16, 2025Updated last year
- JoinAI是一个开源仓库,专注于算法工程能力的培养,包括工程和数学原理的整理☆11Apr 20, 2025Updated last year
- 😎 A Survey of Efficient Reasoning for Large Reasoning Models: Language, Multimodality, Agent, and Beyond☆360Jan 22, 2026Updated 8 months ago
- [ACL 2024] CodeAttack: Revealing Safety Generalization Challenges of Large Language Models via Code Completion☆60Oct 1, 2025Updated 11 months ago
- ☆37Apr 13, 2026Updated 5 months ago
- PyTorch implementation of the paper "Discovering and Explaining the Representation Bottleneck of DNNs" (ICLR 2022 Oral)☆37Oct 30, 2024Updated last year
- We introduce BabyVision, a benchmark revealing the infancy of AI vision.☆254Jan 13, 2026Updated 8 months ago
- Some thoughts about writing scientific papers☆23Nov 8, 2024Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Official repository of DARE: Diffusion Large Language Models Alignment and Reinforcement Executor☆215Jul 30, 2026Updated last month
- A novel approach to improve the safety of large language models, enabling them to transition effectively from unsafe to safe state.☆72May 22, 2025Updated last year
- (ICLR 2026 🔥) Code for "The Devil behind the mask: An emergent safety vulnerability of Diffusion LLMs"☆79Feb 9, 2026Updated 7 months ago
- [ICML 2024] Code release for "On the Emergence of Cross-Task Linearity in Pretraining-Finetuning Paradigm"☆11Feb 20, 2025Updated last year
- ☆30May 22, 2024Updated 2 years ago
- Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recipe☆1,022Aug 20, 2026Updated last month
- instruction-following benchmark for large reasoning models☆49Apr 19, 2026Updated 5 months ago