The code of paper "Toward Optimal LLM Alignments Using Two-Player Games".
☆17Jun 20, 2024Updated last year
Alternatives and similar repositories for gpo
Users that are interested in gpo are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆13Feb 11, 2021Updated 5 years ago
- ☆21Oct 15, 2022Updated 3 years ago
- [ICLR 2022] Boosting Randomized Smoothing with Variance Reduced Classifiers☆11Mar 29, 2022Updated 4 years ago
- ☆12Dec 9, 2020Updated 5 years ago
- How Robust are Randomized Smoothing based Defenses to Data Poisoning? (CVPR 2021)☆14Jul 16, 2021Updated 4 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- SMART introduces a novel test-time framework where Small Language Models (SLMs) reason step-by-step, and Large Language Models (LLMs) pro…☆12Jul 9, 2025Updated 11 months ago
- 👀 VITRina: VIsual Token Representations☆11Jun 15, 2023Updated 3 years ago
- ☆13Jun 4, 2024Updated 2 years ago
- Code and Dataset release of "Carpe Diem: On the Evaluation of World Knowledge in Lifelong Language Models" (NAACL 2024)☆10Oct 16, 2024Updated last year
- ☆16Jul 17, 2022Updated 3 years ago
- zero-vocab or low-vocab embeddings☆18Jul 17, 2022Updated 3 years ago
- AutoLR: Layer-wise Pruning and Auto-tuning of Learning Rates in Fine-tuning of Deep Networks☆17Jan 27, 2021Updated 5 years ago
- Implicit Differentiable Optimal Control (IDOC) with JAX☆12May 11, 2022Updated 4 years ago
- VC-FB and MC-FB algorithms from "Zero-Shot Reinforcement Learning from Low Quality Data" (NeurIPS 2024)☆29Jan 14, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ACL 2024] ValueBench: Towards Comprehensively Evaluating Value Orientations and Understanding of Large Language Models☆27Jan 11, 2025Updated last year
- Simple Conversational Data Augmentation for Semi-supervised Abstractive Conversation Summarization☆10Mar 7, 2022Updated 4 years ago
- The repo for using the model https://huggingface.co/thu-coai/Attacker-v0.1☆13Apr 23, 2025Updated last year
- [NeurIPS 2025 Spotlight] Implementation of "KLASS: KL-Guided Fast Inference in Masked Diffusion Models"☆32Jan 3, 2026Updated 5 months ago
- ☆13Feb 2, 2023Updated 3 years ago
- ☆13Apr 24, 2022Updated 4 years ago
- Code for EMNLP 2023 long paper: An Iteratively Parallel Generation Method with the Pre-Filling Strategy for Document-level Event Extracti…☆19Feb 2, 2025Updated last year
- Repo for the paper "Bounding Training Data Reconstruction in Private (Deep) Learning".☆12Jun 16, 2023Updated 3 years ago
- [ICLR 2024]: Is Self-Repair a Silver Bullet for Code Generation?☆15May 2, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Fault-Tolerant Offline Multi-Agent Path Planning (AAAI-23)☆15Jul 20, 2023Updated 2 years ago
- ☆10Jun 4, 2024Updated 2 years ago
- Code for ACL 2022 long paper: Can Prompt Probe Pretrained Language Models? Understanding the Invisible Risks from a Causal View☆10May 17, 2022Updated 4 years ago
- ☆13Nov 7, 2023Updated 2 years ago
- ☆14Apr 19, 2022Updated 4 years ago
- Microsoft question-answering dataset☆10Jun 16, 2023Updated 3 years ago
- ☆18Feb 7, 2021Updated 5 years ago
- 海康威视工业相机sdk的ros驱动包。☆13Jul 4, 2021Updated 4 years ago
- Kernel Playground - A playground to run large scale experiments on the Linux Kernel☆21Nov 8, 2025Updated 7 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [ICSE 2023] Differentiable interpretation and failure-inducing input generation for neural network numerical bugs.☆13Jan 5, 2024Updated 2 years ago
- ☆12Dec 4, 2022Updated 3 years ago
- ☆15Apr 4, 2023Updated 3 years ago
- Official implementation: Large Language Models are Interpretable Learners - Google☆13Jun 29, 2024Updated last year
- ☆18May 14, 2025Updated last year
- A C++ hierarchical task network (HTN) planner library that interprets a custom domain language.☆20Dec 9, 2023Updated 2 years ago
- ☆13Nov 25, 2019Updated 6 years ago