Official Repo for CRMArena and CRMArena-Pro
☆146Jul 22, 2026Updated last month
Alternatives and similar repositories for CRMArena
Users that are interested in CRMArena are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Holistic Coverage and Faithfulness Evaluation of Large Vision-Language Models (ACL-Findings 2024)☆16Apr 23, 2024Updated 2 years ago
- Official implementation of the ACL 2023 paper: "Zero-shot Faithful Factual Error Correction"☆17Aug 14, 2023Updated 3 years ago
- [NeurIPS'25] Beyond Accuracy: Dissecting Mathematical Reasoning for LLMs Under Reinforcement Learning☆16Dec 12, 2025Updated 8 months ago
- ☆415Jul 23, 2025Updated last year
- ☆106Jul 16, 2026Updated last month
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Implementation of AdaCQR(COLING 2025)☆15Dec 30, 2024Updated last year
- Code and data for the ACL 2024 Findings paper "Do LVLMs Understand Charts? Analyzing and Correcting Factual Errors in Chart Captioning"☆27Jun 5, 2024Updated 2 years ago
- A toolkit for dialogue system evaluation via crowdsourcing☆18Apr 25, 2023Updated 3 years ago
- UQ: Assessing Language Models on Unsolved Questions☆30Aug 26, 2025Updated last year
- ☆13Sep 11, 2024Updated last year
- Lab Cookbook☆42Aug 5, 2026Updated last month
- this is a repository that gives the power of mixture of workflows a concept inspired by the mixture of agents.☆13Aug 19, 2024Updated 2 years ago
- [ICLR 2025] DSBench: How Far are Data Science Agents from Becoming Data Science Experts?☆128Aug 17, 2025Updated last year
- A monolithic index that supports worst-case optimal joins (WCOJ) by providing all collation orders in a single redundancy eliminating dat…☆18Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆13Oct 20, 2022Updated 3 years ago
- ☆16Nov 23, 2023Updated 2 years ago
- CAR-bench☆41Aug 27, 2026Updated last week
- Structured Prediction with Deep Value Networks (PyTorch implementation)☆13Jul 25, 2024Updated 2 years ago
- Memory-Bounded GPU Acceleration for Vector Search☆33Dec 29, 2025Updated 8 months ago
- The official repo for the code and data of paper SMART☆43Feb 20, 2025Updated last year
- Production-grade embedding generation, for any length of text, for transformer models.☆23Aug 10, 2026Updated 3 weeks ago
- ☆33Jun 2, 2026Updated 3 months ago
- ☆28May 15, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- A framework for comprehensive diagnosis and optimization of agents using simulated, realistic synthetic interactions☆1,257Jul 14, 2026Updated last month
- ☆10Oct 22, 2024Updated last year
- ☆18Sep 15, 2025Updated 11 months ago
- The raw UserRL repo under construction☆119Jun 2, 2026Updated 3 months ago
- [ICLR 2026] RPG: KL-Regularized Policy Gradient (https://arxiv.org/abs/2505.17508)☆76Jun 29, 2026Updated 2 months ago
- ☆16Nov 4, 2021Updated 4 years ago
- AuditNLG: Auditing Generative AI Language Modeling for Trustworthiness☆103Jun 2, 2026Updated 3 months ago
- ☆20Mar 12, 2025Updated last year
- [EMNLP26 Findings] Official repository for DataChef: Cooking Up Optimal Data Recipes for LLM Adaptation via Reinforcement Learning☆27Feb 12, 2026Updated 6 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- the implementation of "Entity Resolution via Hierarchical Graph Attention Network"☆24Aug 29, 2023Updated 3 years ago
- xLAM: A Family of Large Action Models to Empower AI Agent Systems☆638Jun 2, 2026Updated 3 months ago
- Convert CVXPY expressions to PyTorch expressions☆18Jul 8, 2025Updated last year
- ☆23Jan 10, 2025Updated last year
- ☆18Feb 7, 2021Updated 5 years ago
- ☆24Mar 23, 2026Updated 5 months ago
- R1-Searcher++: Incentivizing the Dynamic Knowledge Acquisition of LLMs via Reinforcement Learning☆82May 25, 2025Updated last year