☆31Apr 30, 2026Updated 5 months ago
Alternatives and similar repositories for skills-coach
Users that are interested in skills-coach are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆72Feb 4, 2024Updated 2 years ago
- [EMNLP 2025] Official code for the paper "SafeKey: Amplifying Aha-Moment Insights for Safety Reasoning"☆16May 12, 2026Updated 4 months ago
- [COLM'26] SkillLearnBench is the first benchmark for evaluating continual learning methods that automatically generate agent skills.☆83Updated this week
- Scaling Test-time Training for LLM Reasoning☆37Apr 14, 2026Updated 5 months ago
- ☆37May 11, 2026Updated 4 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Agent Skill Evaluation and Evolution: Frameworks and Benchmarks☆30Jul 15, 2026Updated 2 months ago
- ☆17Feb 4, 2025Updated last year
- ☆13Jan 25, 2026Updated 8 months ago
- ☆17Aug 1, 2025Updated last year
- ConfAgents: A Conformal-Guided Multi-Agent Framework for Cost-Efficient Medical Diagnosis☆15Jul 22, 2026Updated 2 months ago
- [TMLR 2025 & ICLR 2025 DeLTa] Official Implementation of Design Editing for Offline Model-based Optimization 🧬 🤖☆10Apr 17, 2025Updated last year
- SkillDAG benchmark reproduction repository☆54Jul 30, 2026Updated 2 months ago
- video_attack; Efficient Sparse Attacks on Videos using Reinforcement Learning☆15Oct 25, 2021Updated 4 years ago
- ☆48Apr 8, 2026Updated 5 months ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- MarsRL: Advancing Multi-Agent Reasoning System via Reinforcement Learning with Agentic Pipeline Parallelism☆18Nov 18, 2025Updated 10 months ago
- ☆12Oct 24, 2022Updated 3 years ago
- action recognition; video classification; LRCN; I3D☆16Aug 9, 2021Updated 5 years ago
- AgentsCourt: Building Judicial Decision-Making Agents with Court Debate Simulation and Legal Knowledge Augmentation (EMNLP 2024 Findings)☆19Dec 30, 2024Updated last year
- The official implementation of MaskGRPO: Consolidating Reinforcement Learning for Multimodal Discrete Diffusion Models. (ICLR 2026, arxiv…☆19Jan 27, 2026Updated 8 months ago
- ☆24Mar 8, 2024Updated 2 years ago
- CoEvoSkills: Self-Evolving Agent Skills via Co-Evolutionary Verification — COLM 2026☆72Sep 23, 2026Updated last week
- The official repository for Multi3WOZ: A Multilingual, Multi-Domain, Multi-Parallel Dataset for Training and Evaluating Culturally Adapte…☆17Jan 15, 2024Updated 2 years ago
- Benchmark self-evolving Agent upon realistic large-scale file workspaces☆78Aug 27, 2026Updated last month
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Official Repo for DAC-RL: Training LLMs for Divide-and-Conquer Reasoning Elevates Test-Time Scalability☆38Sep 20, 2026Updated last week
- ☆69Jul 1, 2026Updated 3 months ago
- ☆14Nov 2, 2022Updated 3 years ago
- Gaussian Membership Inference Privacy (NeurIPS 2023)☆12Jul 27, 2024Updated 2 years ago
- Baselines for Model-Based Optimization installation fixes and compatible with newer AMPERE+ GPUs (e.g. 3090)☆11Apr 30, 2023Updated 3 years ago
- [NeurIPS 2026] GenEnv: Difficulty-Aligned Co-Evolution Between LLM Agents and Environment Simulators☆67Sep 25, 2026Updated last week
- A Self-Evolving Framework Through Executable Subagent Accumulation and Reuse☆64Jul 10, 2026Updated 2 months ago
- ☆17Jun 25, 2025Updated last year
- an efficient method for detecting adversarial image examples☆19Jun 3, 2018Updated 8 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆25Dec 13, 2024Updated last year
- Code with CliqueFlowmer model for Optimal Computational Materials Discovery☆17Sep 6, 2026Updated 3 weeks ago
- ☆19Mar 31, 2024Updated 2 years ago
- Open-source Environment toolkit of claw-like agents, support task/harness generation and evaluation☆62May 7, 2026Updated 4 months ago
- DIVE: Scaling Diversity in Agentic Task Synthesis for Generalizable Tool Use☆30Mar 13, 2026Updated 6 months ago
- [TMLR 24'] TacoGFN: Target Conditioned GFlowNet for Structure-based Drug Design☆21May 31, 2026Updated 4 months ago
- Official repository of paper "Context-DPO: Aligning Language Models for Context-Faithfulness"☆23Feb 17, 2025Updated last year