☆141May 12, 2026Updated 4 months ago
Alternatives and similar repositories for OpenNovelty
Users that are interested in OpenNovelty are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ACL 2026 - Muse: Towards Reproducible Long-Form Song Generation with Fine-Grained Style Control☆125Apr 11, 2026Updated 5 months ago
- Reasoning or Memorization? Unreliable Results of Reinforcement Learning Due to Data Contamination.☆22Jul 18, 2025Updated last year
- Codes for the paper "BAPO: Stabilizing Off-Policy Reinforcement Learning for LLMs via Balanced Policy Optimization with Adaptive Clipping…☆96Jan 29, 2026Updated 7 months ago
- VLA^2: Empowering Vision-Language-Action Models with an Agentic Framework for Unseen Concept Manipulation☆33Nov 3, 2025Updated 10 months ago
- CycleResearcher: Improving Automated Research via Automated Review☆402Mar 5, 2026Updated 6 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- We propose Reinforcement Learning from Community Feedback (RLCF), a training paradigm that uses large-scale community signals as supervis…☆432Updated this week
- [AAAI 2024] LLMEval Phase I dataset — 17 categories, 453 questions, 2186 annotators for Chinese LLM evaluation☆114May 21, 2026Updated 4 months ago
- this is for the ACM MM paper---Backdoor Attack on Crowd Counting☆17Jul 10, 2022Updated 4 years ago
- We introduce 'Thinking with Video', a new paradigm leveraging video generation for multimodal reasoning. Our VideoThinkBench shows that S…☆320Aug 23, 2026Updated last month
- ☆75Mar 7, 2024Updated 2 years ago
- ☆204Jan 28, 2026Updated 7 months ago
- In this work, we investigate the compositionality of large language models (LLMs) in mathematical reasoning. Specifically, we construct a…☆59Mar 15, 2025Updated last year
- ☆116Dec 5, 2025Updated 9 months ago
- Predictable MDP Abstraction for Unsupervised Model-Based RL (ICML 2023)☆32Feb 6, 2023Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Submission Guide + Discussion Board for AI Singapore Global Challenge for Safe and Secure LLMs (Track 1A).☆16Jul 4, 2024Updated 2 years ago
- [ACL 2026] Dissecting Failure Dynamics in Large Language Model Reasoning☆19Apr 17, 2026Updated 5 months ago
- ☆15Apr 17, 2026Updated 5 months ago
- Code for Evolving Language Models without Labels: Majority Drives Selection, Novelty Promotes Variation (EVOL-RL).☆52Mar 31, 2026Updated 5 months ago
- [ICLR 2026] The implementation of paper "AlphaSteer: Learning Refusal Steering with Principled Null-Space Constraint"☆65Nov 20, 2025Updated 10 months ago
- Data and Code Repository for “STRIDE-ED: A Strategy-Grounded Stepwise Reasoning Framework for Empathetic Dialogue Systems”☆17Apr 17, 2026Updated 5 months ago
- Real-Time Fluid Simulation using WebGPU☆23Aug 14, 2026Updated last month
- Official repository for AutoRule: Reasoning Chain-of-thought Extracted Rule-based Rewards Improve Preference Learning☆17Jul 24, 2025Updated last year
- A 1.6B causal Transformer audio tokenizer with streaming, variable bitrates, and semantic alignment across speech, sound, and music☆258Jun 16, 2026Updated 3 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Pre-trained, Scalable, High-performance Reward Models via Policy Discriminative Learning.☆168Sep 23, 2025Updated last year
- [ICML 2026] Hybrid Policy Distillation (HPD) is a practical distillation framework for reasoning-oriented language models. This repositor…☆25Apr 24, 2026Updated 5 months ago
- SaTML'23 paper "Backdoor Attacks on Time Series: A Generative Approach" by Yujing Jiang, Xingjun Ma, Sarah Monazam Erfani, and James Bail…☆21Feb 5, 2023Updated 3 years ago
- a Video Quality Analysis Toolkit☆14May 16, 2025Updated last year
- DELLA-Merging: Reducing Interference in Model Merging through Magnitude-Based Sampling☆38Jul 12, 2024Updated 2 years ago
- Curriculum-RLAIF is a data-centric curriculum learning framework for reward model training in RLAIF-based LLM alignment☆23Apr 18, 2026Updated 5 months ago
- Implementation of the ICML 2024 paper "Training Large Language Models for Reasoning through Reverse Curriculum Reinforcement Learning" pr…☆117Feb 9, 2024Updated 2 years ago
- ☆12Mar 1, 2023Updated 3 years ago
- CL-bench: A Benchmark for Context Learning☆583May 12, 2026Updated 4 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- CVPR2026 (main)☆17Mar 30, 2026Updated 5 months ago
- ☆35Mar 18, 2026Updated 6 months ago
- ☆25Jan 29, 2026Updated 7 months ago
- HTML Agent based on NexAU☆16Nov 20, 2025Updated 10 months ago
- Source code for the Refined Policy Distillation paper.☆23Jul 18, 2025Updated last year
- Use the tokenizer in parallel to achieve superior acceleration☆20Mar 21, 2024Updated 2 years ago
- [AAAI26 oral] CronusVLA: Towards Efficient and Robust Manipulation via Multi-Frame Vision-Language-Action Modeling☆116Jan 11, 2026Updated 8 months ago