Script for processing OpenAI's PRM800K process supervision dataset into an Alpaca-style instruction-response format
☆27Jul 12, 2023Updated 3 years ago
Alternatives and similar repositories for prm800k-denorm
Users that are interested in prm800k-denorm are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Token-level adaptation of LoRA matrices for downstream task generalization.☆15Apr 14, 2024Updated 2 years ago
- Mixture of Expert (MoE) techniques for enhancing LLM performance through expert-driven prompt mapping and adapter combinations.☆11Feb 11, 2024Updated 2 years ago
- Repository for Skill Set Optimization☆14Jul 26, 2024Updated 2 years ago
- This repository contains the code for the paper The Open Proof Corpus: Building a Large-Scale, Human-Validated Dataset of LLM-Generated P…☆18Aug 4, 2025Updated last year
- ☆15Sep 7, 2022Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Eh, simple and works.☆27Dec 9, 2023Updated 2 years ago
- ☆75Sep 5, 2023Updated 2 years ago
- Your efficient and accurate answer verification system for RL training.☆42Jun 23, 2025Updated last year
- 🏆 Ambassador Paper for Innovative Use of NLP for Building Educational Applications 2023: Is ChatGPT a Good Teacher Coach? Measuring Zero…☆14Jul 21, 2024Updated 2 years ago
- Official Implementation of NeurIPS'23 Paper "Cross-Episodic Curriculum for Transformer Agents"☆32Oct 12, 2023Updated 2 years ago
- ☆26Dec 20, 2023Updated 2 years ago
- Get Telemetry Data from YAMCS in OpenMCT☆10Sep 15, 2017Updated 8 years ago
- Official github repo for the paper "Compression Represents Intelligence Linearly" [COLM 2024]☆152Sep 20, 2024Updated last year
- Datasets collection and preprocessings framework for NLP extreme multitask learning☆197Jul 9, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- LLM plugin for models hosted by Anyscale Endpoints☆35Apr 22, 2024Updated 2 years ago
- Run evaluation on LLMs using human-eval benchmark☆431Sep 12, 2023Updated 2 years ago
- Teaching Models to Express Their Uncertainty in Words☆39May 26, 2022Updated 4 years ago
- Comprehensive analysis of difference in performance of QLora, Lora, and Full Finetunes.☆83Sep 10, 2023Updated 2 years ago
- Debian packaging for NNCP [archived], moved to https://salsa.debian.org/go-team/packages/nncp☆14Feb 18, 2023Updated 3 years ago
- ☆10Feb 12, 2024Updated 2 years ago
- kNN-TL: k-Nearest-Neighbor Transfer Learning for Low-Resource Neural Machine Translation (ACL2023)☆11Jul 26, 2023Updated 3 years ago
- Official code for the paper: DRA-GRPO: Exploring Diversity-Aware Reward Adjustment for R1-Zero-Like Training of Large Language Models☆24Jan 6, 2026Updated 7 months ago
- (ICML 2024) Alphazero-like Tree-Search can guide large language model decoding and training☆287May 26, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Code for the paper LeanReasoner: Boosting Complex Logical Reasoning with Lean: https://arxiv.org/pdf/2403.13312.pdf☆27May 25, 2024Updated 2 years ago
- BERT score for text generation☆12Jan 15, 2025Updated last year
- Formal verification of Rust code with AI-assisted specification and proof.☆15Updated this week
- Learning to code TensorFlow☆10Jan 14, 2018Updated 8 years ago
- Code of ACM MM 2023 Paper: A Symbolic Characters Aware Model for Solving Geometry Problems☆16Dec 27, 2023Updated 2 years ago
- Code for Semi-crowdsourced Clustering with Deep Generative Models☆12Dec 9, 2022Updated 3 years ago
- A cog model for the all-mpnet-base-v2 sentence-transformers embedding model.☆15Jan 3, 2024Updated 2 years ago
- Code for "Critique Fine-Tuning: Learning to Critique is More Effective than Learning to Imitate" [COLM 2025]☆182Jul 8, 2025Updated last year
- Code for Contrastive Preference Learning (CPL)☆184Nov 22, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- This is MPE-pytorch, fix some bugs.☆11Apr 26, 2020Updated 6 years ago
- Using multiple LLMs for ensemble Forecasting☆16Jan 17, 2024Updated 2 years ago
- InsTag: A Tool for Data Analysis in LLM Supervised Fine-tuning☆289Aug 20, 2023Updated 2 years ago
- Train a SmolLM-style llm on fineweb-edu in JAX/Flax with an assortment of optimizers.☆19Jul 24, 2025Updated last year
- Repo for Anonymous purpose, pls don't distribute☆10Oct 2, 2024Updated last year
- PennyLane/PyTorch implementation of Quantum agents in the Gym: a variational quantum algorithm for deep Q-learning (Skolik et al., 2021)☆40Mar 15, 2023Updated 3 years ago
- ☆16Jul 9, 2025Updated last year