A multi-language code evaluation tool.
☆28Jan 26, 2024Updated 2 years ago
Alternatives and similar repositories for code-evaluator
Users that are interested in code-evaluator are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [NeurIPS 2021] "Adversarial GLUE: A Multi-Task Benchmark for Robustness Evaluation of Language Models" by Boxin Wang*, Chejian Xu*, Shuoh…☆13Apr 3, 2023Updated 3 years ago
- ROUGE for multilingual Summarization☆25Oct 11, 2021Updated 4 years ago
- MaXM is a suite of test-only benchmarks for multilingual visual question answering in 7 languages: English (en), French (fr), Hindi (hi),…☆13Jan 16, 2024Updated 2 years ago
- ☆17Apr 7, 2025Updated last year
- [ACL 2024 Findings] MathBench: A Comprehensive Multi-Level Difficulty Mathematics Evaluation Dataset☆116May 22, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆29Apr 8, 2025Updated last year
- 儿童故事常识推理与寓意理解评测(Commonsense Reasoning and Moral Understanding Evaluation in Children's Stories,CRMU)☆18Oct 22, 2024Updated last year
- Code for ACL 2022 long paper: Can Prompt Probe Pretrained Language Models? Understanding the Invisible Risks from a Causal View☆10May 17, 2022Updated 4 years ago
- The Chia Network Nebula Graph database Importer☆11Jan 16, 2023Updated 3 years ago
- ☆13Aug 27, 2021Updated 4 years ago
- ☆19Jul 5, 2024Updated 2 years ago
- ☆14Oct 11, 2023Updated 2 years ago
- ☆16Oct 3, 2024Updated last year
- [ECCV2022] Dense Siamese Network for Dense Unsupervised Learning☆29Jul 21, 2022Updated 4 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Running inference on the ZeroSCROLLS benchmark☆22Apr 18, 2024Updated 2 years ago
- NeurIPS 2025☆15Feb 4, 2026Updated 6 months ago
- Nebula docker image for development☆16Apr 1, 2026Updated 4 months ago
- Open-source evaluation toolkit of large vision-language models (LVLMs), support ~100 VLMs, 30+ benchmarks☆15Feb 17, 2025Updated last year
- Official code of our work, VCSR: Mutable CSR Graph Format Using Vertex-Centric Packed Memory Array [CCGrid 2022].☆14Jun 30, 2022Updated 4 years ago
- Evaluating LLMs' multi-round chatting capability via assessing conversations generated by two LLM instances.☆163May 22, 2025Updated last year
- The official implementation of "Ada-LEval: Evaluating long-context LLMs with length-adaptable benchmarks"☆56May 22, 2025Updated last year
- The Math23k dataset for downloading☆22Apr 16, 2022Updated 4 years ago
- Group R-CNN for Point-based Weakly Semi-supervised Object Detection (CVPR2022)☆136May 24, 2023Updated 3 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- OpenClaw memory plugin backed by a self-hosted Mem0 REST API☆17May 21, 2026Updated 2 months ago
- 利用 docker 快速建立 pgadmin4、Linux install pgadmin4☆11Apr 7, 2024Updated 2 years ago
- Minimal example of a OpenAI chat clone written in Streamlit with SOTA features.☆25Jul 11, 2023Updated 3 years ago
- The specification of the LDBC Financial Benchmark☆19Jan 9, 2026Updated 7 months ago
- Articles to share☆14Mar 26, 2023Updated 3 years ago
- Demo springboot + elastic search + kibana☆11Apr 3, 2017Updated 9 years ago
- An example to implement a new backbone with OpenMMLab framework.☆27Mar 22, 2022Updated 4 years ago
- Packed Memory Array☆17May 14, 2014Updated 12 years ago
- ☆24Oct 8, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Code and data for the paper "Understanding Hidden Context in Preference Learning: Consequences for RLHF"☆27Aug 21, 2024Updated last year
- From Accuracy to Robustness: A Study of Rule- and Model-based Verifiers in Mathematical Reasoning.☆24Oct 7, 2025Updated 10 months ago
- 2021 ~ present. NLP 관련 공부 기록☆20Feb 13, 2026Updated 5 months ago
- Source code for SummaReranker (ACL 2022)☆24Jan 7, 2024Updated 2 years ago
- Direct port of the Bullet physics engine to JavaScript using Emscripten☆12Apr 1, 2025Updated last year
- ☆20Mar 22, 2024Updated 2 years ago
- word2vec implementation (for skip-gram and cbow) and simple application of word2vec in sentiment analysis☆21Jan 25, 2019Updated 7 years ago