Can Large Language Models Solve Security Challenges? We test LLMs' ability to interact and break out of shell environments using the OverTheWire wargames environment, showing the models' surprising ability to do action-oriented cyberexploits in shell environments
☆13Aug 21, 2023Updated 2 years ago
Alternatives and similar repositories for llm-security-challenge
Users that are interested in llm-security-challenge are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This project aims at giving the best customer service ever using the power of LLM models like GPT.☆10Jun 29, 2023Updated 3 years ago
- Whispers in the Machine: Confidentiality in Agentic Systems☆45Apr 20, 2026Updated 3 months ago
- Pin files for contextual, codebase-level AI assistance.☆16Jul 11, 2024Updated 2 years ago
- AI-Powered CyberSecurity Compliance: Boost Network Security with OpenAI GPT-3.5-turbo☆10May 18, 2023Updated 3 years ago
- This Repo focuses on defending against 'adversarial prompts,' detecting and attempting to mitigate objectionable content in real time.☆14Jul 30, 2023Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Methods 2: The General Linear Model☆15May 5, 2022Updated 4 years ago
- Risks and targets for assessing LLMs & LLM vulnerabilities☆36May 27, 2024Updated 2 years ago
- Repo for the paper on Escalation Risks of AI systems☆44Apr 12, 2024Updated 2 years ago
- ☆17Dec 30, 2023Updated 2 years ago
- ☆11Sep 7, 2023Updated 2 years ago
- ☆16Aug 8, 2023Updated 2 years ago
- G2Net Competition☆12Aug 2, 2023Updated 2 years ago
- Decentralized File storage system☆16Oct 29, 2023Updated 2 years ago
- ☆14Mar 31, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Multiplayer JS game platform☆16Oct 16, 2017Updated 8 years ago
- ☆20Jun 4, 2023Updated 3 years ago
- Example fNIRS BIDS dataset☆15Nov 4, 2022Updated 3 years ago
- 🔥 A repository for collecting cyberdefense thoughts, books, and documents about AI cyberdefense☆13Jul 2, 2023Updated 3 years ago
- 📚📚📚📚📚📚📚📚📚 Reading everything☆16Mar 11, 2026Updated 4 months ago
- New York Times Article Summarization Tool☆17Sep 15, 2019Updated 6 years ago
- 🧠 Inspecting complexity and goal-directedness of imagination in an fNIRS BCI system.☆11Aug 26, 2023Updated 2 years ago
- Benchmarking Dark Patterns in LLMs (ICLR 2025)☆18Mar 29, 2025Updated last year
- Code for Preventing Language Models From Hiding Their Reasoning, which evaluates defenses against LLM steganography.☆25Jan 26, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆22Jul 18, 2024Updated 2 years ago
- ☆15May 10, 2023Updated 3 years ago
- The following is a simple example of how LLMs and langchain agents can simplify asking questions to understand the security posture of a …☆23Aug 23, 2023Updated 2 years ago
- 👩💻 Code for the ACL paper "Detecting Edit Failures in LLMs: An Improved Specificity Benchmark"☆20Jan 19, 2024Updated 2 years ago
- Improving transparency of large language models' reasoning☆15Nov 25, 2025Updated 8 months ago
- Tool, paper, and study data for DeepManeuver: Adversarial Test Generation for Trajectory Manipulation of Autonomous Vehicles.☆12Aug 26, 2023Updated 2 years ago
- Fine-tuning of transformers for Sentiment Analysis☆18May 25, 2021Updated 5 years ago
- A web service in PHP that "translates" HackNPlan webhook messages to Discord webhook messages.☆16Feb 23, 2023Updated 3 years ago
- 3cb: Catastrophic Cyber Capabilities Benchmarking of Large Language Models☆17Oct 30, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code for the paper "Understanding RL Vision"☆50Apr 2, 2023Updated 3 years ago
- This project investigates the security of large language models by performing binary classification of a set of input prompts to discover…☆65Dec 18, 2023Updated 2 years ago
- A comprehensive, step-by-step guide to mastering cybersecurity from beginner to expert level with curated resources, tools, and career gu…☆24Jul 23, 2025Updated last year
- ☆13Dec 22, 2023Updated 2 years ago
- The Happy Faces Benchmark☆15Jul 20, 2023Updated 3 years ago
- Dark Patterns in Chatbot Design☆20Jun 15, 2024Updated 2 years ago
- ☆22Sep 9, 2021Updated 4 years ago