☆26Feb 12, 2026Updated 6 months ago
Alternatives and similar repositories for DataChef
Users that are interested in DataChef are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ACL2025 Findings] Official code for MIG: Automatic Data Selection for Instruction Tuning by Maximizing Information Gain in Semantic Spac…☆29Aug 30, 2025Updated 11 months ago
- arXiv 2024 | ZIP: entropy-law data selection for efficient LLM alignment.☆28Jun 10, 2026Updated 2 months ago
- CoEvolve: Training LLM Agents via Agent-Data Mutual Evolution☆23Apr 27, 2026Updated 3 months ago
- ☆16Aug 11, 2025Updated last year
- ROUTE: Robust Multitask Tuning and Collaboration for Text-to-SQL (ICLR 2025 Pytorch Code)☆16May 15, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Training codebase for K2-V2☆22Dec 17, 2025Updated 8 months ago
- ☆18Sep 5, 2024Updated last year
- AcademiClaw: When Students Set Challenges for AI Agents — a bilingual benchmark of 80 university student-sourced academic tasks.☆18Jun 26, 2026Updated 2 months ago
- The original Shared Recurrent Memory Transformer implementation☆36Jul 11, 2025Updated last year
- ☆22Jan 2, 2026Updated 7 months ago
- [ACM MM25] LongWriter-V: Enabling Ultra-Long and High-Fidelity Generation in Vision-Language Models☆24Mar 29, 2025Updated last year
- [ICLR'26] ssToken: Self-modulated and Semantic-aware Token Selection for LLM Fine-tuning☆16Feb 28, 2026Updated 5 months ago
- ☆28May 27, 2025Updated last year
- ☆15Apr 25, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Multilingual and Multiculture Benchmark and LLM☆43Jul 30, 2026Updated 3 weeks ago
- [ICML'26] Learning Query-Aware Budget-Tier Routing for Runtime Agent Memory☆23Jun 10, 2026Updated 2 months ago
- ☆15Apr 11, 2024Updated 2 years ago
- ☆23Jun 16, 2026Updated 2 months ago
- Mixture-of-Basis-Experts for Compressing MoE-based LLMs☆37Dec 24, 2025Updated 8 months ago
- Self Evolving Large Multimodal Models with Continuous Rewards☆27Jun 9, 2026Updated 2 months ago
- ☆16Jul 23, 2024Updated 2 years ago
- Code for the paper "Stack Attention: Improving the Ability of Transformers to Model Hierarchical Patterns"☆18Mar 15, 2024Updated 2 years ago
- The benchmark and datasets of the ICML 2024 paper "VisionGraph: Leveraging Large Multimodal Models for Graph Theory Problems in Visual C …☆17May 27, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- A Teacher–Student Cooperation Framework to Synthesize Student-Consistent SFT Data☆37May 1, 2026Updated 3 months ago
- ☆29Sep 4, 2025Updated 11 months ago
- ☆15Jan 24, 2025Updated last year
- An automated data pipeline scaling RL to pretraining levels☆76Jun 2, 2026Updated 2 months ago
- ☆22Jun 23, 2026Updated 2 months ago
- Official repository for MiniAppBench. Contains the complete pipeline and codebase for LLM-powered interactive HTML generation and agentic…☆24Mar 9, 2026Updated 5 months ago
- Official implementation of ECCV24 paper: POA☆24Aug 8, 2024Updated 2 years ago
- WikiVideo: Article Generation from Multiple Videos☆15Nov 14, 2025Updated 9 months ago
- How Instruction and Reasoning Data shape Post-Training: Data Quality through the Lens of Layer-wise Gradients☆21Jun 17, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- The official paper for EnvFactory: Scaling Tool-Use Agents via Executable Environments Synthesis and Robust RL.☆94Jun 5, 2026Updated 2 months ago
- HiPRAG (Hierarchical Process Rewards for Efficient Agentic Retrieval Augmented Generation) is a reinforcement learning method designed fo…☆26Oct 10, 2025Updated 10 months ago
- ☆32Feb 11, 2026Updated 6 months ago
- GFT: From Imitation to Reward Fine-Tuning with Unbiased Group Advantages and Dynamic Coefficient Rectification☆36Jun 10, 2026Updated 2 months ago
- Official implementation of the KDD'26 paper "ManCAR: Manifold-Constrained Latent Reasoning with Adaptive Test-Time Computation for Sequen…☆23May 28, 2026Updated 2 months ago
- Unsupervised Domain Adaptation on Graphs☆15Apr 6, 2022Updated 4 years ago
- Official Repo for DAC-RL: Training LLMs for Divide-and-Conquer Reasoning Elevates Test-Time Scalability☆16Feb 26, 2026Updated 6 months ago