LLM-as-a-judge using G-eval Scratch
☆15Oct 12, 2025Updated 9 months ago
Alternatives and similar repositories for LLM-as-a-judge-using-G-eval
Users that are interested in LLM-as-a-judge-using-G-eval are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Python based Vectorizing Framework☆22Updated this week
- Make running benchmark simple yet maintainable, again. Now only supports Korean-based cross-encoder.☆35Dec 2, 2025Updated 8 months ago
- 🎹 Instruct.KR 2025 Summer Meetup: 오픈소스 LLM, vLLM으로 Production까지 🎹☆23Aug 2, 2025Updated last year
- Andrej karpathy이 만든 Obsidian-RAG☆24Apr 8, 2026Updated 4 months ago
- This repository aims to develop CoT Steering based on CoT without Prompting. It focuses on enhancing the model’s latent reasoning capabil…☆116Jun 25, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- View and manage Claude Code tasks and memory in a floating Hammerspoon window with live updates.☆32Jun 25, 2026Updated last month
- ☆48Updated this week
- Liner LLM Meetup archive☆70Mar 27, 2024Updated 2 years ago
- A loader that lets you try running LLMs built for WebGPU.☆29Dec 20, 2023Updated 2 years ago
- Combining ontology and knowledge graph for an ultimate GraphRAG system.☆58Updated this week
- The list of NLP paper and news I've checked. There might be short description of them (abstract) in Korean.☆38Updated this week
- ☆64Jul 21, 2025Updated last year
- Official repository for KoMT-Bench built by LG AI Research☆73Aug 8, 2024Updated 2 years ago
- ☆21Updated this week
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- generate synthetic data for LLM fine-tuning in arbitrary situations within systematic way☆22Mar 18, 2024Updated 2 years ago
- PERM GaussianKG☆10Nov 24, 2021Updated 4 years ago
- ☆119Oct 13, 2025Updated 9 months ago
- Calibrating LLM Confidence by Probing Perturbed Representation Stability☆19Jul 5, 2025Updated last year
- SKT A.X LLM 4.0☆158Feb 11, 2026Updated 5 months ago
- 밑바닥부터 손수 만들어 보면서 개념을 익히는 핸즈온 Deep Agents 튜토리얼☆135Nov 29, 2025Updated 8 months ago
- KoTAN: Korean Translation and Augmentation with fine-tuned NLLB☆23Jan 4, 2024Updated 2 years ago
- 한국어 벤치마크 평가 코드 통합본(?)☆21Nov 15, 2024Updated last year
- HSEB: Hybrid Search Engine Benchmark☆21Oct 5, 2025Updated 10 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- 42dot LLM consists of a pre-trained language model, 42dot LLM-PLM, and a fine-tuned model, 42dot LLM-SFT, which is trained to respond to …☆134Mar 7, 2024Updated 2 years ago
- Kor-IR: Korean Information Retrieval Benchmark☆17Jul 3, 2024Updated 2 years ago
- BERT score for text generation☆12Jan 15, 2025Updated last year
- A local-LLM based Graph RAG agent using FalkorDB☆21Dec 24, 2025Updated 7 months ago
- NSMC, KorSTS ... fine-tunings☆18Feb 23, 2022Updated 4 years ago
- Performs benchmarking on two Korean datasets with minimal time and effort.☆47Updated this week
- KoCommonGEN v2: A Benchmark for Navigating Korean Commonsense Reasoning Challenges in Large Language Models☆25Aug 24, 2024Updated last year
- hwplib 패키지 python에서 쉽게 사용 할수 있게 만든 github repo 입니다.☆54Mar 29, 2025Updated last year
- fine-tuning tutorial☆20May 30, 2026Updated 2 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Kaggle에서 진행하는 경진대회의 코드를 올려둔 공간입니다.☆34Feb 1, 2024Updated 2 years ago
- Fine-tuned KoGPT2 chatbot demo with translated PersonaChat (ongoing)☆13Apr 17, 2022Updated 4 years ago
- A collection of Python agent samples built with the Google Agent Development Kit (ADK), demonstrating integrations with services like B…☆21May 8, 2026Updated 3 months ago
- Forked repo from https://github.com/EleutherAI/lm-evaluation-harness/commit/1f66adc☆81Feb 28, 2024Updated 2 years ago
- ☆122Feb 25, 2026Updated 5 months ago
- Bias, Hate classification with KoELECTRA 👿☆27Jun 12, 2023Updated 3 years ago
- ☆54Apr 28, 2024Updated 2 years ago