In-Context Explainability 360 toolkit
☆72Mar 9, 2026Updated 5 months ago
Alternatives and similar repositories for ICX360
Users that are interested in ICX360 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The AI Steerability 360 toolkit is an extensible library for general purpose steering of LLMs.☆106Aug 7, 2026Updated 2 weeks ago
- Long-form factuality assessor for large language models☆38Updated this week
- Demo setups for ai-atlas-nexus☆16Jul 15, 2026Updated last month
- Uncertainty Quantification 360 (UQ360) is an extensible open-source toolkit that can help you estimate, communicate and use uncertainty i…☆269Sep 17, 2025Updated 11 months ago
- A library of components to help agent builders boost their agent performance (tool-calling, instruction following, policy, etc.)☆119Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- AI risk ontology☆26Aug 1, 2025Updated last year
- Code and data for the paper "LatentLens: Revealing Highly Interpretable Visual Tokens in LLMs"☆50Mar 31, 2026Updated 4 months ago
- Python code for perturbation-based saliency map☆12Jul 16, 2018Updated 8 years ago
- ☆16Jul 25, 2024Updated 2 years ago
- ☆13Oct 8, 2019Updated 6 years ago
- "TIGERScore: Towards Building Explainable Metric for All Text Generation Tasks" [TMLR 2024]☆32Dec 21, 2024Updated last year
- Attribute statements generated by LLMs to preceding tokens using attention weights.☆28Apr 22, 2025Updated last year
- Python labs demonstrating various techniques for reducing hallucinations in apps using large language models☆33Jun 27, 2026Updated last month
- A framework for designing, executing and analysing experiment campaigns☆62Updated this week
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- 🦄 Unitxt is a Python library for enterprise-grade evaluation of AI performance, offering the world's largest catalog of tools and data …☆217May 27, 2026Updated 2 months ago
- Code for "On Measuring Faithfulness of Natural Language Explanations"☆23Jul 14, 2026Updated last month
- DL Backtrace is a new explainablity technique for deep learning models that works for any modality and model type.☆27May 13, 2026Updated 3 months ago
- Mellea is a library for writing generative programs.☆1,798Updated this week
- ☆10Updated this week
- Exploring limitations of LLM-as-a-judge☆20Aug 17, 2024Updated 2 years ago
- QuoteSum is a textual QA dataset containing Semi-Extractive Multi-source Question Answering (SEMQA) examples written by humans, based on …☆13Mar 25, 2024Updated 2 years ago
- Bootstrap hypothesis testing Python Package. Bootstrapping is a simple method to compute statistics over your custom metrics, using only …☆14Aug 24, 2021Updated 5 years ago
- Prompt Declaration Language (PDL) is a declarative prompt programming language.☆310Updated this week
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆12Mar 5, 2025Updated last year
- learning AI from scratch☆14Feb 17, 2024Updated 2 years ago
- [SIGIR '22] Code for our SIGIR 2022 accepted paper : P3 Ranker: Mitigating the Gaps between Pre-training and Ranking Fine-tuning with Pr…☆18Sep 24, 2023Updated 2 years ago
- VertMetric: An abstractive summarization evaluation package. VERT stands for Versatile Evaluation of Reduced Texts.☆12Dec 20, 2018Updated 7 years ago
- Corpus exploration platform using advanced tools such as interactive summarization and multi document coreference resolution☆12Jun 15, 2023Updated 3 years ago
- 🪝PISCES - Precise In-Parameter Suppression for Concept EraSure in Large Language Models☆14Jun 28, 2026Updated last month
- HealthFC: Verifying Health Claims with Evidence-Based Medical Fact-Checking☆14Apr 11, 2025Updated last year
- Probing for Labeled Dependency Trees (ACL 2022) + Sorting LMs by Structure (NAACL 2022)☆10Jun 11, 2024Updated 2 years ago
- ☆23Sep 2, 2025Updated 11 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- The predecessor of CiteLab.☆18Feb 3, 2026Updated 6 months ago
- Self improving agents through iterations☆104Updated this week
- Interpretability and explainability of data and machine learning models☆1,795Aug 8, 2026Updated 2 weeks ago
- ☆19Feb 14, 2024Updated 2 years ago
- A Benchmark for Reasoning-Driven Retrieval in Medicine☆18Apr 12, 2026Updated 4 months ago
- CUGA is an open-source generalist agent harness for the enterprise, supporting complex task execution on web and APIs, OpenAPI/MCP integr…☆869Updated this week
- OccuBench: Evaluating AI Agents on Real-World Professional Tasks via Language World Models☆21Apr 14, 2026Updated 4 months ago