This is the official repository of the paper "BalanceSFT: Improving LLM Function Calling with Balanced Training Signals and Data Hardness"
☆57Jul 2, 2026Updated 2 months ago
Alternatives and similar repositories for BalanceSFT
Users that are interested in BalanceSFT are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This is the official repository of the paper Exploring Superior Function Calls via Reinforcement Learning.☆33Aug 11, 2025Updated last year
- Laos_System provides a configurable end-to-end pipeline that converts clinical speech/text notes into structured JSON documents for: admi…☆22Jan 7, 2026Updated 7 months ago
- Agentic Learning Powered by AWorld☆126Jun 18, 2026Updated 2 months ago
- ☆22Apr 22, 2026Updated 4 months ago
- [ICLR'26] R-HORIZON: How Far Can Your Large Reasoning Model Really Go in Breadth and Depth?☆26May 9, 2026Updated 3 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Search, understand, reproduce, and improve an idea with ease☆1,229Updated this week
- Our paper is titled "NUS-IDS at FinCausal 2021: Dependency Tree in Graph Neural Networks for better Cause-Effect Span Detection".☆13Feb 11, 2022Updated 4 years ago
- Heterogenous, Task- and Domain-Specific Benchmark for Unsupervised Sentence Embeddings used in the TSDAE paper: https://arxiv.org/abs/210…☆30Jan 4, 2022Updated 4 years ago
- Official repository for MiniAppBench. Contains the complete pipeline and codebase for LLM-powered interactive HTML generation and agentic…☆24Mar 9, 2026Updated 5 months ago
- ☆16Sep 6, 2023Updated 2 years ago
- ABench is an evolving open-source benchmark suite designed to rigorously evaluate and enhance Large Language Models (LLMs) on complex cro…☆29Jul 30, 2026Updated last month
- AI电力能耗预测☆12Aug 13, 2018Updated 8 years ago
- ☆16Nov 14, 2018Updated 7 years ago
- ☆11Oct 25, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆27Jun 10, 2025Updated last year
- IE-Critic-R1: Advancing the Explanatory Measurement of Text-Driven Image Editing for Human Perception Alignment☆20Nov 26, 2025Updated 9 months ago
- ☆14Jul 31, 2025Updated last year
- Code and data for paper "SpatialWorld: Benchmarking Interactive Spatial Reasoning of Multimodal Agents in Real-World Tasks".☆55Jun 18, 2026Updated 2 months ago
- Code for paper "Towards Efficient Pareto Set Approximation via Weight-Ensembling Mixture of Experts"☆11Jul 30, 2026Updated last month
- The code implementation of MuScleLoRA (Accepted in ACL 2024)☆11Dec 1, 2024Updated last year
- [ICLR2026] Video-GPT via Next Clip Diffusion.☆45Jun 2, 2025Updated last year
- ☆20Nov 22, 2025Updated 9 months ago
- ☆43Jun 12, 2023Updated 3 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- This is a collective repository for all 3D and 4D Object Generation papers☆20May 22, 2026Updated 3 months ago
- ☆25Sep 4, 2025Updated last year
- This repository includes code and materials for the paper "Efficient PRM Training Data Synthesis via Formal Verification" (ACL 2026 Findi…☆19Apr 7, 2026Updated 4 months ago
- [CVPR 2026 Oral] SenCache: Accelerating Diffusion Model Inference via Sensitivity-Aware Caching☆24Jun 5, 2026Updated 2 months ago
- Qwen-WisdomVast is a large model trained on 1 million high-quality Chinese multi-turn SFT data, 200,000 English multi-turn SFT data, and …☆17Apr 12, 2024Updated 2 years ago
- Resources and paper list for 'Scaling Environments for Agents'. This repository accompanies our survey on how environments contribute to …☆74Jan 28, 2026Updated 7 months ago
- ☆230Jun 2, 2025Updated last year
- inductive reasoning benchmark with subregular hierarchy for string-to-string transformation☆20Jun 27, 2025Updated last year
- ☆11Jun 15, 2019Updated 7 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- A Benchmark for Fine-Grained Relational Memory Discrimination in Long-Horizon AI Agents☆22Jun 9, 2026Updated 2 months ago
- ☆17Feb 26, 2024Updated 2 years ago
- Official implementation of TBA for async LLM post-training.☆32Nov 5, 2025Updated 9 months ago
- MrlX: A Multi-Agent Reinforcement Learning Framework☆221Jan 19, 2026Updated 7 months ago
- ☆17Aug 1, 2025Updated last year
- [NeurIPS 2025] Bag of Tricks for Inference-time Computation of LLM Reasoning☆16Sep 20, 2025Updated 11 months ago
- Implementation of NAACL'19 Strong and Simple Baselines for Multimodal Utterance Embeddings☆10Jun 4, 2019Updated 7 years ago