Source Code of Arena Leaderboard Methodology
☆120Aug 4, 2026Updated last month
Alternatives and similar repositories for arena-rank
Users that are interested in arena-rank are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Your agent is powerful but it doesn't know you. VibeLens visualizes agent sessions, personalizes your agents, provides dashboard analytic…☆20Updated this week
- ☆65May 13, 2025Updated last year
- [ICML 2026] Official implementation for paper "Unsafer in Many Turns: Benchmarking and Defending Multi-Turn Safety Risks in Tool-Using Ag…☆37Jul 31, 2026Updated last month
- A framework for evaluating the effectiveness of chain-of-thought reasoning in language models.☆19Feb 6, 2025Updated last year
- PhD/MBA-level human-annotated rubrics dataset across Physics, Chemistry, Finance and Consulting☆34Oct 30, 2025Updated 10 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Prompt-to-Leaderboard☆273May 9, 2025Updated last year
- [NeurIPS'25 Spotlight] MJ-VIDEO: Fine-Grained Benchmarking and Rewarding Video Preferences in Video Generation☆20Feb 23, 2025Updated last year
- anp-examples☆18Jul 16, 2025Updated last year
- Temporal and Causal Reasoning (dataset)☆10Apr 19, 2022Updated 4 years ago
- ☆38Jan 31, 2023Updated 3 years ago
- Developer project for getting basic API integrations working in under 5 minutes☆11May 22, 2026Updated 3 months ago
- Learning adapter weights from task descriptions☆21Nov 12, 2023Updated 2 years ago
- Documentation Builder for docs.openclaw.ai☆24Updated this week
- ☆17Oct 11, 2023Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Process-global proxy routing for Node.js☆32Updated this week
- an implementation of paper"Retentive Network: A Successor to Transformer for Large Language Models" https://arxiv.org/pdf/2307.08621.pdf☆11Jul 25, 2023Updated 3 years ago
- Official repository for Decentralized Arena via Collective LLM Intelligence☆18May 19, 2025Updated last year
- ☆23May 25, 2023Updated 3 years ago
- ☆27Aug 16, 2025Updated last year
- Skills for Knowledge Catalog☆33Updated this week
- ☆19Mar 16, 2025Updated last year
- Full-stack Amazon India-inspired e-commerce clone built with Next.js, TypeScript, Prisma, MongoDB, and Stripe.☆22Sep 2, 2026Updated 2 weeks ago
- Guide: from fragile multi-agent app to prod ready with orra - code and resources.☆14Mar 24, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆17Nov 16, 2020Updated 5 years ago
- Infer Ring is an iOS and macOS app that facilitates cross-device LLM inference using MLX☆24May 19, 2026Updated 4 months ago
- Fine-tunes a student LLM using teacher feedback for improved reasoning and answer quality. Implements GRPO with teacher-provided evaluati…☆54May 7, 2025Updated last year
- 3rd-place solution to the ARC-AGI-3 Preview Challenge — Explore It Till You Solve It☆22Jan 7, 2026Updated 8 months ago
- TakeNote is a Markdown note taking app for macOS and iOS with helpful AI features.☆20Jun 14, 2026Updated 3 months ago
- Polkadot Protocol Conformance Tests☆10May 22, 2023Updated 3 years ago
- 裁判文书数据☆11Oct 26, 2020Updated 5 years ago
- Find context neurons in Pythia models.☆13Jun 13, 2023Updated 3 years ago
- Parser-Based Interactive Fiction Platform developed in Typescript with a Custom Modeling Language☆16Updated this week
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Training and data processing code for Saiga☆56Jan 2, 2026Updated 8 months ago
- ☆17Sep 1, 2026Updated 2 weeks ago
- a Python library that uses Reinforcement Learning (RL) to train LLMs.☆43Jul 12, 2026Updated 2 months ago
- ☆18Jan 13, 2025Updated last year
- Prompt based agentic developer primitives☆35Jun 17, 2025Updated last year
- ☆42Apr 14, 2026Updated 5 months ago
- Enjoy puzzle-solving directly in your browser.☆33Apr 18, 2025Updated last year