☆47Apr 8, 2026Updated 3 months ago
Alternatives and similar repositories for Skill-Usage
Users that are interested in Skill-Usage are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆23Dec 17, 2024Updated last year
- ☆41May 12, 2026Updated 2 months ago
- [ACL'25] Code for "Aligning Large Language Models to Follow Instructions and Hallucinate Less via Effective Data Filtering"☆21Jul 23, 2025Updated last year
- Skill retrieval benchmark dataset and evaluation code.☆19May 8, 2026Updated 2 months ago
- ☆17Jan 27, 2026Updated 6 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- The implementation for SIGIR 2026: Learning to Retrieve from Agent Trajectories.☆56Jul 14, 2026Updated 2 weeks ago
- Benchmark self-evolving Agent upon realistic large-scale file workspaces☆49Updated this week
- Companion code to https://arxiv.org/abs/2402.15491☆22Sep 18, 2025Updated 10 months ago
- Memory Transfer Learning: How Memories are Transferred Across Domains in Coding Agents☆31Apr 16, 2026Updated 3 months ago
- ☆15Feb 21, 2024Updated 2 years ago
- ☆31Jun 2, 2026Updated 2 months ago
- Benchmark Test-Time Scaling of General LLM Agents☆21Apr 14, 2026Updated 3 months ago
- ☆22Jan 29, 2026Updated 6 months ago
- GenEnv: Difficulty-Aligned Co-Evolution Between LLM Agents and Environment Simulators☆62Dec 23, 2025Updated 7 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- A repository of OpenDecoder framework: Open Large Language Model Decoding to Incorporate Document Quality in RAG (WWW 2026)☆26Jan 27, 2026Updated 6 months ago
- ToolBridge: An Open-Source Dataset to Equip LLMs with External Tool Capabilities☆14Feb 11, 2025Updated last year
- Mobile GUI Agents under Real-world Threats: Are We There Yet?☆18May 18, 2026Updated 2 months ago
- Extended Inductive Reasoning for Personalized Preference Inference from Behavioral Signals☆11Jan 8, 2026Updated 6 months ago
- COS-PLAY: Co-Evolving LLM Decision and Skill Bank Agents for Long-Horizon Game Play☆29Jul 11, 2026Updated 3 weeks ago
- ☆30Apr 30, 2026Updated 3 months ago
- Models, data, and codes for the paper: MetaAligner: Towards Generalizable Multi-Objective Alignment of Language Models☆24Sep 26, 2024Updated last year
- The official implementation of Preference Data Reward-Augmentation.☆18May 1, 2025Updated last year
- ☆63Jun 2, 2026Updated 2 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Fast Memorization of Prompt Improves Context Awareness of Large Language Models (Findings of EMNLP 2024)☆22Oct 22, 2024Updated last year
- SkillX: Automatically Constructing Skill Knowledge Bases for Agents☆270Jul 5, 2026Updated 3 weeks ago
- ☆10Jun 15, 2024Updated 2 years ago
- Skill-Inject: Measuring Agent Vulnerability to Skill File Attacks☆88Updated this week
- PISmith: Reinforcement Learning-based Red Teaming for Prompt Injection Defenses☆22Jul 17, 2026Updated 2 weeks ago
- The implementation for ACL 2026: MatchTIR: Fine-Grained Supervision for Tool-Integrated Reasoning via Bipartite Matching.☆20Jul 27, 2026Updated last week
- OccuBench: Evaluating AI Agents on Real-World Professional Tasks via Language World Models☆21Apr 14, 2026Updated 3 months ago
- [NAACL 2025 Main] Official implementation of "From Allies to Adversaries: Manipulating LLM Tool Scheduling through Adversarial Injection"…☆22Jun 11, 2025Updated last year
- ☆17Aug 1, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official Implementation of "DeCoRe: Decoding by Contrasting Retrieval Heads to Mitigate Hallucination"☆30Dec 18, 2024Updated last year
- [AACL2025] Code for paper Chain-of-Query: Unleashing the Power of LLMs in SQL-Aided Table Understanding via Multi-Agent Collaboration☆21Jan 19, 2026Updated 6 months ago
- Model-based Hindsight Experience Replay☆10Jun 8, 2022Updated 4 years ago
- ☆18Apr 1, 2025Updated last year
- ☆14Oct 17, 2024Updated last year
- ☆59Jul 1, 2026Updated last month
- ☆18Mar 16, 2026Updated 4 months ago