AI Control for Claude Code in the real world
☆33Jul 7, 2026Updated last month
Alternatives and similar repositories for luthien-proxy
Users that are interested in luthien-proxy are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Data visualization for Inspect AI large language model evalutions.☆21Jul 15, 2026Updated 3 weeks ago
- 3cb: Catastrophic Cyber Capabilities Benchmarking of Large Language Models☆17Oct 30, 2024Updated last year
- ControlArena is a collection of settings, model organisms and protocols - for running control experiments.☆219Jul 31, 2026Updated last week
- ☆21Dec 10, 2025Updated 7 months ago
- Misalignment Bounty submission template☆19Aug 26, 2025Updated 11 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Lightweight representation engineering dataflow operations for agent developers.☆23May 27, 2026Updated 2 months ago
- list of projects related to EA Software Engineers☆28Jun 23, 2023Updated 3 years ago
- Code for Negation Neglect☆16May 22, 2026Updated 2 months ago
- In-depth analysis of AI agent transcripts.☆59Updated this week
- Pin files for contextual, codebase-level AI assistance.☆16Jul 11, 2024Updated 2 years ago
- The open-source AISI toolkit for sandboxing agentic evaluations☆27Aug 7, 2025Updated last year
- Methods 2: The General Linear Model☆15May 5, 2022Updated 4 years ago
- ☆16Jun 17, 2025Updated last year
- ☆11Jun 2, 2021Updated 5 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆20Jun 25, 2024Updated 2 years ago
- ☆17Mar 10, 2026Updated 4 months ago
- AI agent benchmark hackability scanner — find evaluation vulnerabilities before they undermine your results☆42May 25, 2026Updated 2 months ago
- Inference API for many LLMs and other useful tools for empirical research☆136May 29, 2026Updated 2 months ago
- ☆14Mar 31, 2024Updated 2 years ago
- ☆19Apr 28, 2026Updated 3 months ago
- Multiplayer JS game platform☆16Oct 16, 2017Updated 8 years ago
- These are the code files and resources for the first part of the ARM assembly language programming course.☆19Jan 30, 2023Updated 3 years ago
- Can Large Language Models Solve Security Challenges? We test LLMs' ability to interact and break out of shell environments using the Over…☆13Aug 21, 2023Updated 2 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Collection of evals for Inspect AI☆617Updated this week
- 🔥 A repository for collecting cyberdefense thoughts, books, and documents about AI cyberdefense☆13Jul 2, 2023Updated 3 years ago
- Showcasing the power of Ruby on Rails.☆12Jun 7, 2020Updated 6 years ago
- F-Secure Lightweight Acqusition for Incident Response (FLAIR)☆16Jul 5, 2021Updated 5 years ago
- 📚📚📚📚📚📚📚📚📚 Reading everything☆16Mar 11, 2026Updated 4 months ago
- ☆38Feb 8, 2024Updated 2 years ago
- Ember is a hosted API/SDK that lets you shape AI model behavior by directly controlling a model's internal units of computation, or "feat…☆55Jul 14, 2025Updated last year
- ☆15Apr 13, 2026Updated 3 months ago
- 🧠 Inspecting complexity and goal-directedness of imagination in an fNIRS BCI system.☆11Aug 26, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆11Oct 13, 2020Updated 5 years ago
- Benchmarking Dark Patterns in LLMs (ICLR 2025)☆18Mar 29, 2025Updated last year
- Obrew Server: A self-hostable machine learning engine. Build agents and schedule workflows private to you.☆17May 16, 2026Updated 2 months ago
- ☆149Aug 4, 2024Updated 2 years ago
- 👩💻 Code for the ACL paper "Detecting Edit Failures in LLMs: An Improved Specificity Benchmark"☆20Jan 19, 2024Updated 2 years ago
- Software Engineering Agents for Inspect AI☆25Updated this week
- ☆72Jan 17, 2025Updated last year