Moonshot - A simple and modular tool to evaluate and red-team any LLM application.
☆352Jun 10, 2026Updated 3 months ago
Alternatives and similar repositories for moonshot
Users that are interested in moonshot are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Contains all assets to run with Moonshot Library (Connectors, Datasets and Metrics)☆45Feb 5, 2026Updated 7 months ago
- AI Verify☆96Mar 23, 2026Updated 5 months ago
- This repository stems from our paper, “Cataloguing LLM Evaluations ”, and serves as a living, collaborative catalogue of LLM evaluation fr…☆23Nov 16, 2023Updated 2 years ago
- ☆10Jan 14, 2025Updated last year
- Improving transparency of large language models' reasoning☆15Nov 25, 2025Updated 9 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Papers about red teaming LLMs and Multimodal models.☆179Jul 14, 2026Updated 2 months ago
- LLM evaluation.☆16Nov 7, 2023Updated 2 years ago
- Parallel Universal Dependencies.☆15May 6, 2026Updated 4 months ago
- ☆25Nov 27, 2023Updated 2 years ago
- This repository provides an export of various AWS demonstrations an instructor may leverage to deliver a course about running Amazon EKS …☆14Mar 2, 2023Updated 3 years ago
- Finetune Code for OpenThaiGPT 0.1.0-beta☆12Nov 11, 2023Updated 2 years ago
- ☆16Jun 15, 2024Updated 2 years ago
- Code and Data for Evaluating the Evaluators☆16Aug 20, 2025Updated last year
- South-East Asia Large Language Models☆425Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆31Apr 16, 2024Updated 2 years ago
- Independent robustness evaluation of Improving Alignment and Robustness with Short Circuiting☆18Apr 15, 2025Updated last year
- S-Eval: Towards Automated and Comprehensive Safety Evaluation for Large Language Models☆121Feb 13, 2026Updated 7 months ago
- LLM red teaming datasets from the paper 'Student-Teacher Prompting for Red Teaming to Improve Guardrails' for the ART of Safety Workshop …☆26Oct 12, 2023Updated 2 years ago
- 🤖🛡️🔍🔒🔑 Tiny package designed to support red teams and penetration testers in exploiting large language model AI solutions.☆26May 16, 2024Updated 2 years ago
- ☆16Aug 27, 2026Updated 3 weeks ago
- A minimal yet unstoppable blueprint for multi-agent AI—anchored by the rare, far-reaching “Multi-Agent AI DAO” (2017 Prior Art)—empowerin…☆38Jan 11, 2025Updated last year
- ☆98Nov 9, 2024Updated last year
- A summary of NSO Group/Circles documents, research and media clippings.☆12Apr 13, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Effective sampling methods within TensorFlow input functions.☆10Mar 24, 2023Updated 3 years ago
- Inspect: A framework for large language model evaluations☆2,806Updated this week
- ☆31Jul 14, 2023Updated 3 years ago
- Java library to tokenize Thai text into a list of TCCs☆22May 30, 2017Updated 9 years ago
- HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal☆1,049Aug 16, 2024Updated 2 years ago
- AI risk ontology☆26Aug 1, 2025Updated last year
- ☆10Mar 13, 2023Updated 3 years ago
- Repo containing documentation and explanation for CSET's harm taxonomy of incidents from AIID.☆21Jun 21, 2024Updated 2 years ago
- Bundle of security analysis scripts for keras tensorflow models☆16Apr 15, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Adding guardrails to large language models.☆7,433Updated this week
- The data and code for paper: "SOPBench: Evaluating Language Agents at Following Standard Operating Procedures and Constraints"☆19Updated this week
- ☆50Aug 3, 2024Updated 2 years ago
- The Dataset and Official Implementation for <Discursive Socratic Questioning: Evaluating the Faithfulness of Language Models’ Understandi…☆19Aug 7, 2024Updated 2 years ago
- ☆25Jun 2, 2026Updated 3 months ago
- Official PyTorch implementation of RACRO (https://www.arxiv.org/abs/2506.04559)☆19Jul 1, 2025Updated last year
- this is based on the paper Chain-of-Retrieval Augmented Generation☆15Mar 29, 2025Updated last year