☆138May 12, 2026Updated 3 months ago
Alternatives and similar repositories for OpenNovelty
Users that are interested in OpenNovelty are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ACL 2026 - Muse: Towards Reproducible Long-Form Song Generation with Fine-Grained Style Control☆120Apr 11, 2026Updated 4 months ago
- Reasoning or Memorization? Unreliable Results of Reinforcement Learning Due to Data Contamination.☆21Jul 18, 2025Updated last year
- Codes for the paper "BAPO: Stabilizing Off-Policy Reinforcement Learning for LLMs via Balanced Policy Optimization with Adaptive Clipping…☆95Jan 29, 2026Updated 6 months ago
- [AAAI 2024] LLMEval Phase II dataset — professional domain evaluation across 12 academic disciplines☆71May 21, 2026Updated 2 months ago
- VLA^2: Empowering Vision-Language-Action Models with an Agentic Framework for Unseen Concept Manipulation☆33Nov 3, 2025Updated 9 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Adapter-X: A Novel General Parameter-Efficient Fine-Tuning Framework for Vision☆11Jul 22, 2024Updated 2 years ago
- CycleResearcher: Improving Automated Research via Automated Review☆400Mar 5, 2026Updated 5 months ago
- We propose Reinforcement Learning from Community Feedback (RLCF), a training paradigm that uses large-scale community signals as supervis…☆427Jul 22, 2026Updated 3 weeks ago
- [AAAI 2024] LLMEval Phase I dataset — 17 categories, 453 questions, 2186 annotators for Chinese LLM evaluation☆114May 21, 2026Updated 2 months ago
- Official repo for "SE-Bench: Benchmarking Self-Evolution with Knowledge Internalization"☆28Mar 24, 2026Updated 4 months ago
- The Github repo for our survey paper: "Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large…☆154Apr 15, 2026Updated 3 months ago
- ☆116Dec 5, 2025Updated 8 months ago
- Predictable MDP Abstraction for Unsupervised Model-Based RL (ICML 2023)☆33Feb 6, 2023Updated 3 years ago
- [ACL 2026] Dissecting Failure Dynamics in Large Language Model Reasoning☆18Apr 17, 2026Updated 3 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Real-Time Fluid Simulation using WebGPU☆21Updated this week
- ☆15Apr 17, 2026Updated 3 months ago
- Code for Evolving Language Models without Labels: Majority Drives Selection, Novelty Promotes Variation (EVOL-RL).☆51Mar 31, 2026Updated 4 months ago
- [ICLR 2026] The implementation of paper "AlphaSteer: Learning Refusal Steering with Principled Null-Space Constraint"☆62Nov 20, 2025Updated 8 months ago
- Data and Code Repository for “STRIDE-ED: A Strategy-Grounded Stepwise Reasoning Framework for Empathetic Dialogue Systems”☆17Apr 17, 2026Updated 3 months ago
- Official repository for AutoRule: Reasoning Chain-of-thought Extracted Rule-based Rewards Improve Preference Learning☆17Jul 24, 2025Updated last year
- MOSS-Audio-Tokenizer is a Causal Transformer-based audio tokenizer built on the CAT architecture. Trained on 3M hours of diverse audio, i…☆250Jun 16, 2026Updated last month
- Pre-trained, Scalable, High-performance Reward Models via Policy Discriminative Learning.☆168Sep 23, 2025Updated 10 months ago
- [ICML 2026] Hybrid Policy Distillation (HPD) is a practical distillation framework for reasoning-oriented language models. This repositor…☆24Apr 24, 2026Updated 3 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- SaTML'23 paper "Backdoor Attacks on Time Series: A Generative Approach" by Yujing Jiang, Xingjun Ma, Sarah Monazam Erfani, and James Bail…☆21Feb 5, 2023Updated 3 years ago
- DELLA-Merging: Reducing Interference in Model Merging through Magnitude-Based Sampling☆38Jul 12, 2024Updated 2 years ago
- Curriculum-RLAIF is a data-centric curriculum learning framework for reward model training in RLAIF-based LLM alignment☆23Apr 18, 2026Updated 3 months ago
- Implementation of the ICML 2024 paper "Training Large Language Models for Reasoning through Reverse Curriculum Reinforcement Learning" pr…☆117Feb 9, 2024Updated 2 years ago
- A mass-spring system simulator that animates realistic hanging, pinned, falling, colliding, and folding cloth behaviors.☆10May 18, 2021Updated 5 years ago
- [ECCV 2026] Implementation of "CollectionLoRA: Collecting 50 Effects in 1 LoRA via Multi-Teacher On-Policy Distillation"☆29Jun 25, 2026Updated last month
- CL-bench: A Benchmark for Context Learning☆576May 12, 2026Updated 3 months ago
- CVPR2026 (main)☆16Mar 30, 2026Updated 4 months ago
- ☆25Jan 29, 2026Updated 6 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- HTML Agent based on NexAU☆16Nov 20, 2025Updated 8 months ago
- Idea2Paper Offical Demo☆1,386Mar 24, 2026Updated 4 months ago
- Source code for the Refined Policy Distillation paper.☆21Jul 18, 2025Updated last year
- A curated list of awesome resources about reward construction for AI agents. This repository covers cutting-edge research, and practical …☆61Sep 1, 2025Updated 11 months ago
- Use the tokenizer in parallel to achieve superior acceleration☆20Mar 21, 2024Updated 2 years ago
- Short RL☆19Apr 16, 2026Updated 3 months ago
- [AAAI26 oral] CronusVLA: Towards Efficient and Robust Manipulation via Multi-Frame Vision-Language-Action Modeling☆114Jan 11, 2026Updated 7 months ago