☆135May 12, 2026Updated 2 months ago
Alternatives and similar repositories for OpenNovelty
Users that are interested in OpenNovelty are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Reasoning or Memorization? Unreliable Results of Reinforcement Learning Due to Data Contamination.☆21Jul 18, 2025Updated last year
- Codes for the paper "BAPO: Stabilizing Off-Policy Reinforcement Learning for LLMs via Balanced Policy Optimization with Adaptive Clipping…☆94Jan 29, 2026Updated 5 months ago
- [AAAI 2024] LLMEval Phase II dataset — professional domain evaluation across 12 academic disciplines☆71May 21, 2026Updated 2 months ago
- VLA^2: Empowering Vision-Language-Action Models with an Agentic Framework for Unseen Concept Manipulation☆31Nov 3, 2025Updated 8 months ago
- Adapter-X: A Novel General Parameter-Efficient Fine-Tuning Framework for Vision☆11Jul 22, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- CycleResearcher: Improving Automated Research via Automated Review☆399Mar 5, 2026Updated 4 months ago
- We propose Reinforcement Learning from Community Feedback (RLCF), a training paradigm that uses large-scale community signals as supervis…☆425Updated this week
- ☆11Jan 25, 2022Updated 4 years ago
- [AAAI 2024] LLMEval Phase I dataset — 17 categories, 453 questions, 2186 annotators for Chinese LLM evaluation☆114May 21, 2026Updated 2 months ago
- [ICCV 2025] Preacher: Paper-to-Video Agentic System☆50Sep 1, 2025Updated 10 months ago
- this is for the ACM MM paper---Backdoor Attack on Crowd Counting☆17Jul 10, 2022Updated 4 years ago
- We introduce 'Thinking with Video', a new paradigm leveraging video generation for multimodal reasoning. Our VideoThinkBench shows that S…☆315Jun 21, 2026Updated last month
- Official repo for "SE-Bench: Benchmarking Self-Evolution with Knowledge Internalization"☆28Mar 24, 2026Updated 4 months ago
- The Github repo for our survey paper: "Locate, Steer, and Improve: A Practical Survey of Actionable Mechanistic Interpretability in Large…☆150Apr 15, 2026Updated 3 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆15Aug 23, 2025Updated 11 months ago
- ☆116Dec 5, 2025Updated 7 months ago
- Predictable MDP Abstraction for Unsupervised Model-Based RL (ICML 2023)☆33Feb 6, 2023Updated 3 years ago
- The official implementation of "ML-Agent: Reinforcing LLM Agents for Autonomous Machine Learning Engineering"☆71Jun 21, 2025Updated last year
- ☆15Apr 17, 2026Updated 3 months ago
- [ICLR 2026] The implementation of paper "AlphaSteer: Learning Refusal Steering with Principled Null-Space Constraint"☆61Nov 20, 2025Updated 8 months ago
- The official implementation of ManiAgent☆32Jan 4, 2026Updated 6 months ago
- Repository of the paper ''CritiQ: Mining Data Quality Criteria from Human Preferences". Code for CritiQ Flow & Training CritiQ Scorer.☆22Dec 11, 2025Updated 7 months ago
- Official repository for AutoRule: Reasoning Chain-of-thought Extracted Rule-based Rewards Improve Preference Learning☆17Jul 24, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- MOSS-Audio-Tokenizer is a Causal Transformer-based audio tokenizer built on the CAT architecture. Trained on 3M hours of diverse audio, i…☆248Jun 16, 2026Updated last month
- Pre-trained, Scalable, High-performance Reward Models via Policy Discriminative Learning.☆166Sep 23, 2025Updated 10 months ago
- [ICML 2026] Hybrid Policy Distillation (HPD) is a practical distillation framework for reasoning-oriented language models. This repositor…☆24Apr 24, 2026Updated 3 months ago
- SaTML'23 paper "Backdoor Attacks on Time Series: A Generative Approach" by Yujing Jiang, Xingjun Ma, Sarah Monazam Erfani, and James Bail…☆21Feb 5, 2023Updated 3 years ago
- A Docker-first, non-preemptive multi-agent coordination runtime☆16Jul 9, 2026Updated 2 weeks ago
- Curriculum-RLAIF is a data-centric curriculum learning framework for reward model training in RLAIF-based LLM alignment☆23Apr 18, 2026Updated 3 months ago
- Implementation of the ICML 2024 paper "Training Large Language Models for Reasoning through Reverse Curriculum Reinforcement Learning" pr…☆116Feb 9, 2024Updated 2 years ago
- A mass-spring system simulator that animates realistic hanging, pinned, falling, colliding, and folding cloth behaviors.☆10May 18, 2021Updated 5 years ago
- CL-bench: A Benchmark for Context Learning☆573May 12, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- CVPR2026 (main)☆16Mar 30, 2026Updated 3 months ago
- ☆34Mar 18, 2026Updated 4 months ago
- ☆25Jan 29, 2026Updated 5 months ago
- HTML Agent based on NexAU☆16Nov 20, 2025Updated 8 months ago
- Idea2Paper Offical Demo☆1,376Mar 24, 2026Updated 4 months ago
- Source code for the Refined Policy Distillation paper.☆22Jul 18, 2025Updated last year
- Use the tokenizer in parallel to achieve superior acceleration☆20Mar 21, 2024Updated 2 years ago