A hackable, simple, and reseach-friendly GRPO Training Framework with high speed weight synchronization in a multinode environment.
☆38Aug 27, 2025Updated last year
Alternatives and similar repositories for RLYX
Users that are interested in RLYX are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Open Source + Multilingual MLLM + Fine-tuning + Distillation + More efficient models and learning + ?☆19Jan 31, 2025Updated last year
- 한국어 LLM 리더보드 및 모델 성능/안전성 관리☆22Sep 26, 2023Updated 2 years ago
- StrategyQA 데이터 세트 번역☆22Apr 12, 2024Updated 2 years ago
- Performs benchmarking on two Korean datasets with minimal time and effort.☆48Aug 6, 2026Updated 3 weeks ago
- 심층강화학습 책 https://hiddenbeginner.github.io/Deep-Reinforcement-Learnings☆11May 10, 2024Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Open-source RL Framework with Online Teacher-Student Distillation☆22Mar 5, 2026Updated 5 months ago
- ☆21May 4, 2026Updated 3 months ago
- OCR Engine☆17Dec 31, 2021Updated 4 years ago
- Triton kernels for Flux☆23Jul 7, 2025Updated last year
- The most modern LLM evaluation toolkit☆70Apr 30, 2026Updated 4 months ago
- ☆15Oct 10, 2023Updated 2 years ago
- BERTScore for Korean☆81Feb 22, 2024Updated 2 years ago
- Korean Multi-task Instruction Tuning☆156Dec 20, 2023Updated 2 years ago
- Experiments Notebook of "Understanding the Skill Gap in Recurrent Language Models: The Role of the Gather-and-Aggregate Mechanism"☆17Apr 30, 2025Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Repository for "Scaling Evaluation-time Compute with Reasoning Models as Process Evaluators"☆12Mar 25, 2025Updated last year
- Korean Math Word Problems☆59Jan 14, 2022Updated 4 years ago
- [KO-Platy🥮] Korean-Open-platypus를 활용하여 llama-2-ko를 fine-tuning한 KO-platypus model☆73Aug 24, 2025Updated last year
- Submission archive for the MS MARCO passage ranking leaderboard☆13Apr 21, 2023Updated 3 years ago
- ☆13Jan 12, 2023Updated 3 years ago
- Official Codebase of "A Closer Look at Weakly-Supervised Audio-Visual Source Localization" (NeurIPS 2022)☆22Dec 6, 2022Updated 3 years ago
- Benchmark in Korean Context☆139Sep 26, 2023Updated 2 years ago
- 한국어 노이즈 생성을 위한 라이브러리입니다.☆23May 18, 2023Updated 3 years ago
- KoRean based ELECTRA pre-trained models (KR-ELECTRA) for Tensorflow and PyTorch☆15Feb 13, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆14May 8, 2023Updated 3 years ago
- ☆15Apr 6, 2026Updated 4 months ago
- #인권코퍼스☆31Oct 6, 2023Updated 2 years ago
- 날짜, 장소, 사람, 기관, 시간☆23Jan 10, 2023Updated 3 years ago
- ☆19Oct 24, 2023Updated 2 years ago
- the benchmark for finance☆11Jul 4, 2023Updated 3 years ago
- KoLLaVA: Korean Large Language-and-Vision Assistant (feat.LLaVA)☆295Sep 20, 2024Updated last year
- ☆21Sep 6, 2021Updated 4 years ago
- 한국어 언어 모델 학습을 위한 프로젝트(Flax, Pytorch with Huggingface Accelerate)☆32Sep 13, 2023Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- 한국어 언어모델 다분야 사고력 벤치마크☆208Oct 17, 2024Updated last year
- Optimized primitives for collective multi-GPU communication☆11May 8, 2024Updated 2 years ago
- generate synthetic data for LLM fine-tuning in arbitrary situations within systematic way☆22Mar 18, 2024Updated 2 years ago
- 자연어 처리 기반 [한글 서술형 수학문제 데이터셋] 공개 저장소입니다.☆14Jun 12, 2023Updated 3 years ago
- ☆122Apr 21, 2023Updated 3 years ago
- Context Modeling with Speaker's Pre-trained Memory Tracking for Emotion Recognition in Conversation (NAACL 2022)☆63Mar 17, 2023Updated 3 years ago
- [ACL'26] Official Code for "ReflexiCoder: Teaching Large Language Models to Self-Reflect on Generated Code and Self-Correct It via Reinfo…☆23Apr 24, 2026Updated 4 months ago