Skill-RM: Unifying Heterogeneous Evaluation Criteria via Agent Skill
☆25Jun 8, 2026Updated 3 months ago
Alternatives and similar repositories for Skill-RM
Users that are interested in Skill-RM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆23Jun 16, 2026Updated 3 months ago
- ☆18Feb 14, 2026Updated 7 months ago
- [NeurIPS'25 Spotlight] Fine-grained List-wise Alignment for Generative Medication Recommendation☆18May 19, 2026Updated 4 months ago
- ☆17Mar 9, 2026Updated 6 months ago
- A small, runnable reference implementation that accompanies Loop Engineering: A Practitioner's Guide. Each module maps to a chapter so yo…☆20Jul 30, 2026Updated 2 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- KERL: Knowledge-Enhanced Personalized Recipe Recommendation using Large Language Models☆15Jun 18, 2025Updated last year
- ☆87May 8, 2026Updated 4 months ago
- The official repo of ICML2026 Paper: Adversarial Latent Embedding Repair for LLM Continual Learning☆20Aug 25, 2026Updated last month
- [EMNLP 2023] ReLM: Leveraging Language Models for Enhanced Chemical Reaction Prediction.☆23Jan 28, 2024Updated 2 years ago
- ☆14Mar 18, 2025Updated last year
- This repo contains evaluation code for the paper "AV-Odyssey: Can Your Multimodal LLMs Really Understand Audio-Visual Information?"☆31Dec 23, 2024Updated last year
- The official repo of NeurIPS2025 Paper: High-Performance Arithmetic Circuit Optimization via Differentiable Architecture Search☆19Oct 19, 2025Updated 11 months ago
- [AAAI 2026] The official code for ``LoopLLM: Transferable Energy-Latency Attacks in LLMs via Repetitive Generation''☆17Mar 20, 2026Updated 6 months ago
- For ACL25 paper "WAFFLE: Multi-Modal Model for Automated Front-End Development" - by Shanchao Liang and Nan Jiang and Shangshu Qian and L…☆12May 28, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- PolicyShiftGuard: Benchmarking and Improving Policy-Adaptive Image Guardrails☆22Jul 8, 2026Updated 2 months ago
- Code repository for the ICML 2026 paper "Reasoning over Boundaries: Enhancing Specification Alignment via Test-time Deliberation".☆24Jun 14, 2026Updated 3 months ago
- ☆22Mar 11, 2026Updated 6 months ago
- LLMPerf is a library for validating and benchmarking LLMs☆11Aug 13, 2024Updated 2 years ago
- CLIPO: Contrastive Learning in Policy Optimization Generalizes RLVR☆22Apr 7, 2026Updated 5 months ago
- [ACL 2025 Findings] Understanding the Repeat Curse in Large Language Models from a Feature Perspective☆22Jun 13, 2025Updated last year
- Universal OKF-based memory system for Hermes agent - structured, persistent, agent-readable knowledge storage.☆35Jun 25, 2026Updated 3 months ago
- Synthetic pretraining data by rephrasing the web☆35Jun 5, 2026Updated 3 months ago
- [ICLR 2026] Official repo for "FrameThinker: Learning to Think with Long Videos via Multi-Turn Frame Spotlighting"☆56Oct 9, 2025Updated 11 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Code for Unsupervised multi-granular Chinese word segmentation and term discovery via graph partition [JBI]☆16Jan 28, 2022Updated 4 years ago
- Open Rubric System: Scaling Reinforcement Learning with Pairwise Adaptive Rubric☆29Mar 5, 2026Updated 6 months ago
- Code and data for "An Accurate Unsupervised Method for Joint Entity Alignment and Dangling Entity Detection".☆15Mar 26, 2022Updated 4 years ago
- CODA: Coordinating the Cerebrum and Cerebellum for a Dual-Brain Computer Use Agent with Decoupled Reinforcement Learning☆37Aug 28, 2025Updated last year
- Yet another Bloomfilter implementation in Python, compatible with Java's Guava library☆12Aug 10, 2024Updated 2 years ago
- official code for "BoostStep: Boosting mathematical capability of Large Language Models via improved single-step reasoning"☆37Jan 21, 2025Updated last year
- Official Implementation for *PaCo-RL: Advancing Reinforcement Learning for Consistent Image Generation with Pairwise Reward Modeling*☆43Dec 13, 2025Updated 9 months ago
- simple trainer for musicgen/audiocraft☆15Jul 14, 2023Updated 3 years ago
- ☆26May 30, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- This is the official repository of Daily-Omni: Towards Audio-Visual Reasoning with Temporal Alignment across Modalities☆48Jul 26, 2026Updated 2 months ago
- ☆31May 15, 2026Updated 4 months ago
- Code and dataset release for "PACS: A Dataset for Physical Audiovisual CommonSense Reasoning" (ECCV 2022)☆18Dec 20, 2022Updated 3 years ago
- Look Back to Reason Forward: Revisitable Memory for Long-Context LLM Agents☆45Apr 13, 2026Updated 5 months ago
- [Tech Report] Expanded Hyper-Connections☆68Jul 21, 2026Updated 2 months ago
- ☆36Jun 18, 2026Updated 3 months ago
- ☆40Aug 26, 2025Updated last year