Ko-Arena-Hard-Auto: An automatic LLM benchmark for Korean
☆22Apr 23, 2025Updated last year
Alternatives and similar repositories for ko-arena-hard-auto
Users that are interested in ko-arena-hard-auto are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A simple JSON parser specifically designed to handle malformed JSON output from Large Language Models (LLMs) like GPT, Claude, and others…☆27Jun 20, 2025Updated last year
- 한글 텍스트 임베딩 모델 리더보드☆97Oct 22, 2024Updated last year
- BERT score for text generation☆12Jan 15, 2025Updated last year
- This repository aims to develop CoT Steering based on CoT without Prompting. It focuses on enhancing the model’s latent reasoning capabil…☆116Jun 25, 2025Updated last year
- Gunmo-emo-classification: 한국어 감정 다중 분류 모델 제작법☆28Dec 12, 2023Updated 2 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- ☆64Jul 21, 2025Updated last year
- ☆14Jan 31, 2025Updated last year
- Flask App Which detects 15 variety of plants [Pepper , Potato , Tomato ]☆11Aug 27, 2020Updated 5 years ago
- nanoRLHF: from-scratch journey into how LLMs and RLHF really work.☆195Jul 13, 2026Updated 2 weeks ago
- Run and manage local Codex workflows from trusted Discord channels.☆19Apr 27, 2026Updated 3 months ago
- The most modern LLM evaluation toolkit☆70Apr 30, 2026Updated 3 months ago
- prototype of plant-disease-detector☆10Apr 21, 2021Updated 5 years ago
- ☆10Dec 19, 2023Updated 2 years ago
- Korean Translation Benchmark, LLM-as-a-judge☆23Oct 23, 2025Updated 9 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- huggingface에 있는 한국어 데이터 세트☆37Oct 10, 2024Updated last year
- A lightweight adjustment tool for smoothing token probabilities in the Qwen models to encourage balanced multilingual generation.☆106Jul 9, 2025Updated last year
- ☆29Nov 10, 2024Updated last year
- Repository for "Scaling Evaluation-time Compute with Reasoning Models as Process Evaluators"☆12Mar 25, 2025Updated last year
- 🎹 Instruct.KR 2025 Summer Meetup: 오픈소스 LLM, vLLM으로 Production까지 🎹☆23Aug 2, 2025Updated 11 months ago
- KoCommonGEN v2: A Benchmark for Navigating Korean Commonsense Reasoning Challenges in Large Language Models☆25Aug 24, 2024Updated last year
- 《GPT-4, ChatGPT, 라마인덱스, 랭체인을 활용한 인공지능 프로그래밍》 예제 코드☆10Jan 16, 2024Updated 2 years ago
- Korean-MTEB☆100May 12, 2026Updated 2 months ago
- Korean Sentence Embedding Model Performance Benchmark for RAG☆49Jan 27, 2025Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Using decision tree and random forest models, predict the winner of an NBA regular season game☆15Jun 7, 2018Updated 8 years ago
- 랭체인 & 랭그래프로 AI 에이전트 개발하기 소스 코드☆12Mar 3, 2025Updated last year
- Train GEMMA on TPU/GPU! (Codebase for training Gemma-Ko Series)☆50Mar 2, 2024Updated 2 years ago
- ☆38Oct 4, 2023Updated 2 years ago
- This is a hands-on for ML beginners to perform SimCSE step-by-step. Implemented both supervised SimCSE and unsupervisied SimCSE, and dist…☆22Oct 6, 2023Updated 2 years ago
- ☆12Dec 20, 2024Updated last year
- Forked repo from https://github.com/EleutherAI/lm-evaluation-harness/commit/1f66adc☆81Feb 28, 2024Updated 2 years ago
- An unofficial implementation of SOLAR-10.7B model and the newly proposed interlocked-DUS(iDUS) implementation and experiment details.☆14Mar 20, 2024Updated 2 years ago
- 1-Click is all you need.☆63Apr 29, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆20Jul 24, 2024Updated 2 years ago
- hwplib 패키지 python에서 쉽게 사용 할수 있게 만든 github repo 입니다.☆54Mar 29, 2025Updated last year
- Official codes for NAACL 2025 paper "LLMs Are Biased Towards Output Formats! Systematically Evaluating and Mitigating Output Format Bias …☆11Nov 25, 2025Updated 8 months ago
- Pretraining and finetuning for visual instruction following with Mixture of Experts☆15Jan 30, 2024Updated 2 years ago
- Official onboarding skill for HWPX document automation with AI agents.☆19Updated this week
- ☆40Mar 9, 2026Updated 4 months ago
- HieraPlan - Hierarchical Task Planner for llm agents☆17Mar 20, 2025Updated last year