☆32Feb 11, 2026Updated 7 months ago
Alternatives and similar repositories for AgentSkiller
Users that are interested in AgentSkiller are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- An automated pipeline that leverages LLM's meta-learning capability to iteratively design and refine red-teaming systems without human in…☆31May 24, 2026Updated 4 months ago
- Official implementation of the ΔBelief-RL method.☆31Feb 28, 2026Updated 7 months ago
- The official implementation of "EnvScaler: Scaling Tool-Interactive Environments for LLM Agent via Programmatic Synthesis".☆199Sep 3, 2026Updated last month
- All-in-one repository for Fine-tuning & Pretraining (Large) Language Models☆15Mar 8, 2023Updated 3 years ago
- ☆13Aug 12, 2026Updated last month
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [EMNLP26 Findings] Official repository for DataChef: Cooking Up Optimal Data Recipes for LLM Adaptation via Reinforcement Learning☆27Feb 12, 2026Updated 7 months ago
- ☆181May 13, 2026Updated 4 months ago
- Repository for "Scaling Evaluation-time Compute with Reasoning Models as Process Evaluators"☆12Mar 25, 2025Updated last year
- ☆39Jun 13, 2026Updated 3 months ago
- Code for EACL 26 Findings paper "I-MCTS: Enhancing Agentic AutoML via Introspective Monte Carlo Tree Search"☆13Jan 28, 2026Updated 8 months ago
- [NeurIPS 2025] Official PyTorch implementation of paper "Learning Grouped Lattice Vector Quantizers for Low-Bit LLM Compression".☆16Oct 24, 2025Updated 11 months ago
- The official implementation of the paper "AgentLAB: Benchmarking LLM Agents against Long-Horizon Attacks"☆42Jun 1, 2026Updated 4 months ago
- Code and Dataset release of "Carpe Diem: On the Evaluation of World Knowledge in Lifelong Language Models" (NAACL 2024)☆10Oct 16, 2024Updated last year
- mechanical stage☆12Aug 6, 2023Updated 3 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Codes for Merging Large Language Models☆37Aug 7, 2024Updated 2 years ago
- [ACL 2026 Oral] From Word to World: Can Large Language Models be Implicit Text-based World Models?☆78Apr 13, 2026Updated 5 months ago
- Agent Skills Enable a New Class of Realistic and Trivially Simple Prompt Injections☆23Jul 2, 2026Updated 3 months ago
- SuperOptiX: Full Stack Agentic AI Framework☆27Oct 1, 2026Updated last week
- ☆21Apr 21, 2026Updated 5 months ago
- Code for our paper titled "Lens: Rethinking Multilingual Enhancement for Large Language Models"☆12Oct 15, 2024Updated last year
- Metrics for evaluating biological sequence design☆16Sep 1, 2026Updated last month
- ☆24Dec 18, 2025Updated 9 months ago
- Verifying the optimization phases of the GraalVM compiler☆15Sep 28, 2026Updated last week
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [ACL 2026 Main] MCP-Flow: Facilitating LLM Agents to Master Real-World, Diverse and Scaling MCP Tools.☆26Apr 8, 2026Updated 6 months ago
- offical implementation of Jailbreak-R1☆16Jul 16, 2025Updated last year
- Agent World Model: Infinity Synthetic Environments for Agentic Reinforcement Learning☆467May 28, 2026Updated 4 months ago
- Open source codebase for PRBench☆20Jan 15, 2026Updated 8 months ago
- ☆28Jan 29, 2026Updated 8 months ago
- Russian Drug Reaction Corpus (RuDReC)☆13Dec 29, 2020Updated 5 years ago
- ☆19Jan 29, 2026Updated 8 months ago
- Vstream - Video Analytics pipeline with Hardware based accelerations (dev - stage)☆10Feb 2, 2024Updated 2 years ago
- Sutracli is an AI-powered code manager for coding agents. It spawns agents for multiple projects, connects repos through cross-indexing, …☆29Nov 7, 2025Updated 11 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Code and data release for FEABench: Evaluating Language Models on Multiphysics Reasoning Ability. [MATH-AI workshop, NeurIPS 2024]☆14May 7, 2025Updated last year
- MCP Atlas☆158Updated this week
- [ACL 2026] DR-Arena: an Automated Evaluation Framework for Deep Research Agents☆18Jul 8, 2026Updated 3 months ago
- The official implementation of NOSA (EMNLP 2026 main)☆19Sep 28, 2026Updated last week
- [ICLR 2026] The official implementation of the paper “Anchored Supervised Fine-Tuning”☆50Sep 8, 2026Updated last month
- LongSeeker: Elastic Context Orchestration for Long-Horizon Search Agents☆28May 29, 2026Updated 4 months ago
- This is a framework for using large language models to improve ASR recognition accuracy. You need to provide the recognized text and tag …☆19Jun 5, 2025Updated last year