NexusRaven-13B, a new SOTA Open-Source LLM for function calling. This repo contains everything for reproducing our evaluation on NexusRaven-13B and baselines.
☆325Sep 29, 2023Updated 3 years ago
Alternatives and similar repositories for NexusRaven
Users that are interested in NexusRaven are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆418Feb 13, 2024Updated 2 years ago
- Nexusflow function call, tool use, and agent benchmarks.☆28Dec 13, 2024Updated last year
- Chat language model that can use tools and interpret the results☆1,595Jun 30, 2026Updated 3 months ago
- [COLING 2025] ToolEyes: Fine-Grained Evaluation for Tool Learning Capabilities of Large Language Models in Real-world Scenarios☆74May 13, 2025Updated last year
- [ICML 2024] LLMCompiler: An LLM Compiler for Parallel Function Calling☆1,892Jul 10, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ICLR'24 spotlight] An open platform for training, serving, and evaluating large language model for tool learning.☆5,751May 21, 2025Updated last year
- Gorilla: Training and Evaluating LLMs for Function Calls (Tool Calls)☆13,042Apr 13, 2026Updated 5 months ago
- xLAM: A Family of Large Action Models to Empower AI Agent Systems☆636Jun 2, 2026Updated 4 months ago
- ☆35Feb 8, 2024Updated 2 years ago
- Official repository for LongChat and LongEval☆535May 24, 2024Updated 2 years ago
- FireAct: Toward Language Agent Fine-tuning☆297Oct 22, 2023Updated 2 years ago
- AgentTuning: Enabling Generalized Agent Abilities for LLMs☆1,504Oct 31, 2023Updated 2 years ago
- The code implementation of MAGDi: Structured Distillation of Multi-Agent Interaction Graphs Improves Reasoning in Smaller Language Models…☆41Feb 5, 2024Updated 2 years ago
- Syntax Error-Free and Generalizable Tool Use for LLMs via Finite-State Decoding☆31Jan 28, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Official Repo for ICLR 2024 paper MINT: Evaluating LLMs in Multi-turn Interaction with Tools and Language Feedback by Xingyao Wang*, Ziha…☆144Jun 4, 2024Updated 2 years ago
- [ICLR 2024] Lemur: Open Foundation Models for Language Agents☆555Oct 28, 2023Updated 2 years ago
- Exploring limitations of LLM-as-a-judge☆20Aug 17, 2024Updated 2 years ago
- ☆15Nov 22, 2023Updated 2 years ago
- LLMs build upon Evol Insturct: WizardLM, WizardCoder, WizardMath☆9,478Jun 7, 2025Updated last year
- Large-scale, Informative, and Diverse Multi-round Chat Data (and Models)☆2,904Mar 13, 2024Updated 2 years ago
- A Comprehensive Benchmark to Evaluate LLMs as Agents (ICLR'24)☆3,765Feb 8, 2026Updated 8 months ago
- Xwin-LM: Powerful, Stable, and Reproducible LLM Alignment☆1,037May 31, 2024Updated 2 years ago
- Tools for merging pretrained large language models.☆7,392Sep 12, 2026Updated 3 weeks ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A cog implementation of Nvidia's Triton server☆18Oct 23, 2024Updated last year
- ☆123Jun 6, 2024Updated 2 years ago
- [COLM 2024] OpenAgents: An Open Platform for Language Agents in the Wild☆4,863Nov 18, 2024Updated last year
- Open Source Projects from Pallas Lab☆21Oct 10, 2021Updated 5 years ago
- Code for the paper "Rethinking Benchmark and Contamination for Language Models with Rephrased Samples"☆327Dec 20, 2023Updated 2 years ago
- [EMNLP 2024] RoTBench: A Multi-Level Benchmark for Evaluating the Robustness of Large Language Models in Tool Learning☆15May 13, 2025Updated last year
- ☆21Apr 29, 2024Updated 2 years ago
- Build Hierarchical Autonomous Agents through Config. Collaborative Growth of Specialized Agents.☆329Nov 27, 2023Updated 2 years ago
- ToRA is a series of Tool-integrated Reasoning LLM Agents designed to solve challenging mathematical reasoning problems by interacting wit…☆1,125Feb 22, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [NAACL 2025] Representing Rule-based Chatbots with Transformers☆23Feb 9, 2025Updated last year
- Code and data for "Lumos: Learning Agents with Unified Data, Modular Design, and Open-Source LLMs"☆477Mar 19, 2024Updated 2 years ago
- S-LoRA: Serving Thousands of Concurrent LoRA Adapters☆1,923Jan 21, 2024Updated 2 years ago
- Pre-training code for CrystalCoder 7B LLM☆61May 10, 2024Updated 2 years ago
- ToolBench, an evaluation suite for LLM tool manipulation capabilities.☆183Jul 27, 2026Updated 2 months ago
- ☆105Dec 6, 2024Updated last year
- [ACL2024] T-Eval: Evaluating Tool Utilization Capability of Large Language Models Step by Step☆313Apr 3, 2024Updated 2 years ago