Ko-Arena-Hard-Auto: An automatic LLM benchmark for Korean
☆22Apr 23, 2025Updated last year
Alternatives and similar repositories for ko-arena-hard-auto
Users that are interested in ko-arena-hard-auto are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A simple JSON parser specifically designed to handle malformed JSON output from Large Language Models (LLMs) like GPT, Claude, and others…☆27Jun 20, 2025Updated last year
- 한글 텍스트 임베딩 모델 리더보드☆97Oct 22, 2024Updated last year
- BERT score for text generation☆12Jan 15, 2025Updated last year
- This repository aims to develop CoT Steering based on CoT without Prompting. It focuses on enhancing the model’s latent reasoning capabil…☆116Jun 25, 2025Updated last year
- Gunmo-emo-classification: 한국어 감정 다중 분류 모델 제작법☆28Dec 12, 2023Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆64Jul 21, 2025Updated last year
- ☆14Jan 31, 2025Updated last year
- Run and manage local Codex workflows from trusted Discord channels.☆20Apr 27, 2026Updated 3 months ago
- prototype of plant-disease-detector☆10Apr 21, 2021Updated 5 years ago
- ☆10Dec 19, 2023Updated 2 years ago
- Korean Translation Benchmark, LLM-as-a-judge☆23Oct 23, 2025Updated 9 months ago
- ☆29Nov 10, 2024Updated last year
- huggingface에 있는 한국어 데이터 세트☆37Oct 10, 2024Updated last year
- A lightweight adjustment tool for smoothing token probabilities in the Qwen models to encourage balanced multilingual generation.☆107Jul 9, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Repository for "Scaling Evaluation-time Compute with Reasoning Models as Process Evaluators"☆12Mar 25, 2025Updated last year
- 🎹 Instruct.KR 2025 Summer Meetup: 오픈소스 LLM, vLLM으로 Production까지 🎹☆23Aug 2, 2025Updated last year
- KoCommonGEN v2: A Benchmark for Navigating Korean Commonsense Reasoning Challenges in Large Language Models☆25Aug 24, 2024Updated last year
- Korean-MTEB☆103May 12, 2026Updated 3 months ago
- Korean Sentence Embedding Model Performance Benchmark for RAG☆49Jan 27, 2025Updated last year
- 인공지능에 대한 배경지식이 없어도 LLM을 학습시켜 나만의 GPT를 만들 수 있는 오픈소스 솔루션 | 🏆 2023 공개SW 개발자대회 장려상☆11Nov 9, 2023Updated 2 years ago
- ☆10Aug 13, 2023Updated 3 years ago
- Miscellaneous codes and writings for MLOps☆16Apr 8, 2026Updated 4 months ago
- 랭체인 & 랭그래프로 AI 에이전트 개발하기 소스 코드☆12Mar 3, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Train GEMMA on TPU/GPU! (Codebase for training Gemma-Ko Series)☆50Mar 2, 2024Updated 2 years ago
- ☆38Oct 4, 2023Updated 2 years ago
- Parses, Analyzes and Predicts for the Korean Baseball League☆17Dec 8, 2022Updated 3 years ago
- This is a hands-on for ML beginners to perform SimCSE step-by-step. Implemented both supervised SimCSE and unsupervisied SimCSE, and dist…☆22Oct 6, 2023Updated 2 years ago
- ☆12Dec 20, 2024Updated last year
- An unofficial implementation of SOLAR-10.7B model and the newly proposed interlocked-DUS(iDUS) implementation and experiment details.☆14Mar 20, 2024Updated 2 years ago
- 1-Click is all you need.☆63Apr 29, 2024Updated 2 years ago
- hwplib 패키지 python에서 쉽게 사용 할수 있게 만든 github repo 입니다.☆54Mar 29, 2025Updated last year
- Official codes for NAACL 2025 paper "LLMs Are Biased Towards Output Formats! Systematically Evaluating and Mitigating Output Format Bias …☆11Nov 25, 2025Updated 8 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Pretraining and finetuning for visual instruction following with Mixture of Experts☆15Jan 30, 2024Updated 2 years ago
- 자체 구축한 한국어 평가 데이터셋을 이용한 한국어 모델 평가☆31May 31, 2024Updated 2 years ago
- ☆40Mar 9, 2026Updated 5 months ago
- HieraPlan - Hierarchical Task Planner for llm agents☆17Mar 20, 2025Updated last year
- The source code of the game I made for the HuggingFace game jam☆16Jul 25, 2023Updated 3 years ago
- 행정안전부에서 마련한 〈디지털 정부서비스 UI/UX 가이드라인〉을 준수하는 것을 목표로 삼고 있는 크로스-프레임워크 컴포넌트 라이브러리이다.☆11Apr 24, 2024Updated 2 years ago
- This is a Korean OCR Python code using the Pororo library☆87May 24, 2023Updated 3 years ago