Audit any agent decision across its past, present, and future, on one typed graph.
☆27Sep 8, 2026Updated last month
Alternatives and similar repositories for auditable
Users that are interested in auditable are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆16Aug 30, 2025Updated last year
- AlphaDiana: A System for Evaluating Agentic Reasoning☆22Aug 12, 2026Updated last month
- LLM Benchmark☆45May 24, 2025Updated last year
- Source code for StageNet: Stage-Aware Neural Networks for Health Risk Prediction☆33May 7, 2025Updated last year
- NeurIPS 2025 Poster☆24Oct 17, 2025Updated 11 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- NeurIPS 2025 Poster☆26Feb 4, 2025Updated last year
- Code and data for EMNLP 2023 paper "Grounding Visual Illusions in Language: Do Vision-Language Models Perceive Illusions Like Humans?"☆15Jan 25, 2024Updated 2 years ago
- [NeurIPS 2024] "Can Language Models Perform Robust Reasoning in Chain-of-thought Prompting with Noisy Rationales?"☆41Jul 18, 2025Updated last year
- A curated list of awesome human-centered AI resources.☆47Apr 14, 2022Updated 4 years ago
- [NeurIPS 2024] HonestLLM: Toward an Honest and Helpful Large Language Model☆29Jun 10, 2025Updated last year
- [NeurIPS 2024, spotlight] Scaling Out-of-Distribution Detection for Multiple Modalities☆71Dec 3, 2025Updated 10 months ago
- Code for the multi-agent computer use project.☆25Jul 3, 2026Updated 3 months ago
- Code for Negation Neglect☆18May 22, 2026Updated 4 months ago
- [ICML 2025] "From Debate to Equilibrium: Belief-Driven Multi-Agent LLM Reasoning via Bayesian Nash Equilibrium"☆41Nov 23, 2025Updated 10 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆12Aug 8, 2023Updated 3 years ago
- Uncertainty-Aware Reliable Text Classification (KDD 2021)☆18Oct 4, 2022Updated 4 years ago
- ☆18Aug 19, 2024Updated 2 years ago
- ACL 2025: Search-based multilingual LLM evaluation that discovers English-correct, target-language-wrong failures; code and 6,713-pair da…☆42Sep 21, 2026Updated 2 weeks ago
- DynAuditClaw — A security audit skill that dynamically discovers your OpenClaw agent's real configuration, designs targeted attack scenar…☆15Apr 6, 2026Updated 6 months ago
- One config to rule all your AI agents: portable (every project, every session), effective (curated writing, routing, skills), and safer (…☆248Updated this week
- [ICML 2025] "From Passive to Active Reasoning: Can Large Language Models Ask the Right Questions under Incomplete Information?"☆48Oct 8, 2025Updated last year
- [arXiv:2510.06261] "AlphaApollo: A System for Deep Agentic Reasoning"☆51Aug 21, 2026Updated last month
- A simple CLI command that initialises a Kedro project from an existing Python package☆11Aug 23, 2024Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Base Kafka Producer, consumer, flask api and PySpark Structured streaming Job☆11Oct 20, 2021Updated 4 years ago
- ☆18Nov 27, 2017Updated 8 years ago
- Code for paper "AutoAudit: Mining Accounting and Time-Evolving Graphs" (Big Data 2020)☆19Aug 23, 2023Updated 3 years ago
- [Cell Patterns 2024] A comprehensive benchmark for COVID-19 predictive modeling using electronic health records in intensive care☆32May 24, 2026Updated 4 months ago
- ☆13Sep 5, 2024Updated 2 years ago
- Supplementary material for KDD 2018 workshop "DCSO: Dynamic Combination of Detector Scores for Outlier Ensembles"☆21Sep 8, 2026Updated last month
- The code for paper "MemGym: a Long-Horizon Memory Environment for LLM Agents".☆23Jun 2, 2026Updated 4 months ago
- Open-source EU AI Act compliance scanner. 51 checks across Articles 9-12, 14-15. Drop-in trust layers for LangChain, CrewAI, AutoGen, Ope…☆23Oct 1, 2026Updated last week
- Code for TKDE paper: Patient Health Representation Learning via Correlational Sparse Prior of Medical Features.☆11Jan 5, 2023Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [ICLR 26] The official code repository for the paper "Mirage or Method? How Model–Task Alignment Induces Divergent RL Conclusions".☆19Feb 9, 2026Updated 8 months ago
- [CVPR 2026 DataCV Workshop] 4KLSDB: A Large-Scale Native-4K Dataset and Benchmark for Image Restoration and Generation.☆36May 28, 2026Updated 4 months ago
- Sum of handwritten digits using Deep Sets, implemented in PyTorch.☆22Apr 17, 2019Updated 7 years ago
- A benchmark for evaluating the efficiency of LLM-generated code☆17Apr 17, 2025Updated last year
- [ICLR 2026] "Landscape of Thoughts: Visualizing the Reasoning Process of Large Language Models"☆63May 21, 2026Updated 4 months ago
- Supplementary material for paper "SynC: A Unified Framework for Generating Synthetic Population with Gaussian Copula"☆21Nov 9, 2020Updated 5 years ago
- pytorch-retain☆13Nov 20, 2017Updated 8 years ago