Fiddler Auditor is a tool to evaluate language models.
β196Mar 11, 2024Updated 2 years ago
Alternatives and similar repositories for fiddler-auditor
Users that are interested in fiddler-auditor are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- β20Updated this week
- π LangKit: An open-source toolkit for monitoring Large Language Models (LLMs). π Extracts signals from prompts & responses, ensuring saβ¦β996Nov 22, 2024Updated last year
- Sample notebooks and prompts for LLM evaluationβ176Nov 2, 2025Updated 10 months ago
- Coffee Chat Voice Assistant is a voice-driven ordering system powered by Azure OpenAI GPT-4o Realtime API, simulating the experience of oβ¦β32Jun 22, 2026Updated 2 months ago
- python jupyter notebook tutorialsβ14Apr 14, 2026Updated 4 months ago
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A tool for evaluating LLMsβ428Mar 15, 2026Updated 5 months ago
- Building the laion5B paperβ35May 6, 2022Updated 4 years ago
- Just a bunch of benchmark logs for different LLMsβ130Jul 28, 2024Updated 2 years ago
- Evaluating LLMs with CommonGen-Liteβ95Mar 21, 2024Updated 2 years ago
- Sample project to get started with dbt-power-user vscode extension using dev-containerβ12Apr 5, 2024Updated 2 years ago
- An end to end ML project. Using MLflow for experiment tracking and model registry. Prefect for workflow orchestration. S3 for artifacts sβ¦β12Sep 11, 2022Updated 3 years ago
- Deepchecks: Tests for Continuous Validation of ML Models & Data. Deepchecks is a holistic open-source solution for all of your AI & ML vaβ¦β4,050Dec 28, 2025Updated 8 months ago
- This is a small lib for making simple multi criteria decisionsβ11Jul 20, 2019Updated 7 years ago
- π’ Open-Source Evaluation & Testing library for LLM Agentsβ5,801Updated this week
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Build, evaluate, understand, and fix LLM-based appsβ490Jan 16, 2024Updated 2 years ago
- Oxygen is a Robot Framework tool that empowers the user to convert the results of any testing tool or framework to Robot Framework's repoβ¦β26Jun 26, 2024Updated 2 years ago
- Repository for NPHardEval, a quantified-dynamic benchmark of LLMsβ66Mar 26, 2024Updated 2 years ago
- zero-vocab or low-vocab embeddingsβ18Jul 17, 2022Updated 4 years ago
- Proof of concept implementation of a cyber threat intelligence and incident handling platformβ11Feb 10, 2023Updated 3 years ago
- Python SDK for running evaluations on LLM generated responsesβ301Jun 6, 2025Updated last year
- A library for red-teaming LLM applications with LLMs.β28Oct 11, 2024Updated last year
- A Library for Scaling Mixed-Integer Optimization-Based Machine Learning.β11Jun 24, 2024Updated 2 years ago
- β10Oct 11, 2022Updated 3 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer β’ AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- LLM Prompt Injection Detectorβ1,521Aug 7, 2024Updated 2 years ago
- One scan for AI risk. `opena2a review` checks an AI project across credentials, shadow agents, MCP servers, and dependencies, returns a sβ¦β20Updated this week
- Apple OAuth2 Provider for Laravel Socialiteβ10May 8, 2020Updated 6 years ago
- Evidently is ββan open-source ML and LLM observability framework. Evaluate, test, and monitor any AI-powered system or data pipeline. Froβ¦β7,878Updated this week
- Simulate SDTM datasets in SAS.β12Jan 19, 2018Updated 8 years ago
- This sample shows how to use a Cosmos DB Trigger in Azure Functions Triggers (C# or Python) to automatically generate embeddings on data β¦β21Mar 11, 2025Updated last year
- Code for the papers "Induction of Subgoal Automata for Reinforcement Learning" (AAAI-20) and "Induction and Exploitation of Subgoal Automβ¦β14Aug 15, 2023Updated 3 years ago
- NeMo Guardrails is an open-source toolkit for easily adding programmable guardrails to LLM-based conversational systems.β7,048Updated this week
- This project involves using llamaindex Multi Agents concierge system and Qdrant vector database to customize the RAG application with useβ¦β57Aug 20, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- β28Aug 1, 2024Updated 2 years ago
- This repository contains the Julia code for the paper "Competitive Gradient Descent"β25Dec 18, 2019Updated 6 years ago
- New ways of breaking app-integrated LLMsβ2,133Jul 17, 2025Updated last year
- Metrics to evaluate the quality of responses of your Retrieval Augmented Generation (RAG) applications.β327Jul 10, 2025Updated last year
- Deepmark AI enables a unique testing environment for language models (LLM) assessment on task-specific metrics and on your own data so yoβ¦β104Nov 24, 2023Updated 2 years ago
- This project provides a solution to AWS customers for reporting on what tags exists, the resources they are applied to, and what resourceβ¦β25Feb 28, 2024Updated 2 years ago
- ToolBench, an evaluation suite for LLM tool manipulation capabilities.β182Jul 27, 2026Updated last month