Official Repo for CRMArena and CRMArena-Pro
☆142Jul 22, 2026Updated last week
Alternatives and similar repositories for CRMArena
Users that are interested in CRMArena are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official implementation of "CONCRETE: Improving Cross-lingual Fact Checking with Cross-lingual Retrieval" (COLING'22)☆15Oct 13, 2022Updated 3 years ago
- Official implementation of the ACL 2023 paper: "Zero-shot Faithful Factual Error Correction"☆17Aug 14, 2023Updated 2 years ago
- [NeurIPS'25] Beyond Accuracy: Dissecting Mathematical Reasoning for LLMs Under Reinforcement Learning☆16Dec 12, 2025Updated 7 months ago
- ☆396Jul 23, 2025Updated last year
- ☆102Jul 16, 2026Updated last week
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆28Nov 19, 2025Updated 8 months ago
- Code and data for the ACL 2024 Findings paper "Do LVLMs Understand Charts? Analyzing and Correcting Factual Errors in Chart Captioning"☆27Jun 5, 2024Updated 2 years ago
- A toolkit for dialogue system evaluation via crowdsourcing☆18Apr 25, 2023Updated 3 years ago
- UQ: Assessing Language Models on Unsolved Questions☆30Aug 26, 2025Updated 11 months ago
- ☆13Sep 11, 2024Updated last year
- Lab Cookbook☆38Updated this week
- this is a repository that gives the power of mixture of workflows a concept inspired by the mixture of agents.☆13Aug 19, 2024Updated last year
- [ICLR 2025] DSBench: How Far are Data Science Agents from Becoming Data Science Experts?☆125Aug 17, 2025Updated 11 months ago
- A monolithic index that supports worst-case optimal joins (WCOJ) by providing all collation orders in a single redundancy eliminating dat…☆18Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆13Oct 20, 2022Updated 3 years ago
- ☆15Nov 23, 2023Updated 2 years ago
- Structured Prediction with Deep Value Networks (PyTorch implementation)☆13Jul 25, 2024Updated 2 years ago
- ☆27May 9, 2022Updated 4 years ago
- Memory-Bounded GPU Acceleration for Vector Search☆33Dec 29, 2025Updated 7 months ago
- ☆33Jun 2, 2026Updated last month
- ☆28May 15, 2024Updated 2 years ago
- The goal of this experiment is to take articles and certain metadata and group them by topic.☆11Apr 14, 2016Updated 10 years ago
- The Stanford Word Substitution (Swords) Benchmark☆33Mar 24, 2022Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- This repository contains the codes of "A Lip Sync Expert Is All You Need for Speech to Lip Generation In the Wild", published at ACM Mult…☆15Nov 11, 2022Updated 3 years ago
- A framework for comprehensive diagnosis and optimization of agents using simulated, realistic synthetic interactions☆1,253Jul 14, 2026Updated 2 weeks ago
- ☆10Oct 22, 2024Updated last year
- ☆18Sep 15, 2025Updated 10 months ago
- The raw UserRL repo under construction☆114Jun 2, 2026Updated last month
- [ICLR 2026] RPG: KL-Regularized Policy Gradient (https://arxiv.org/abs/2505.17508)☆76Jun 29, 2026Updated last month
- AuditNLG: Auditing Generative AI Language Modeling for Trustworthiness☆103Jun 2, 2026Updated last month
- ☆20Mar 12, 2025Updated last year
- xLAM: A Family of Large Action Models to Empower AI Agent Systems☆636Jun 2, 2026Updated last month
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- repo to hold my collection of NotebookLM Infographics and Slides☆30Mar 28, 2026Updated 4 months ago
- A tool for calling (and calling out to) large language models.☆16Aug 13, 2024Updated last year
- Intro to using DSPy with Kuzu to enrich the data within the Nobel Laureate mentorship network☆16Sep 16, 2025Updated 10 months ago
- Work-in-progress unofficial asynchronous API wrapper for Whatnot API.☆14Apr 18, 2024Updated 2 years ago
- ☆28May 19, 2025Updated last year
- ☆24Mar 23, 2026Updated 4 months ago
- R1-Searcher++: Incentivizing the Dynamic Knowledge Acquisition of LLMs via Reinforcement Learning☆82May 25, 2025Updated last year