CS-Eval is a comprehensive evaluation suite for fundamental cybersecurity models or large language models' cybersecurity ability.
☆67Nov 27, 2024Updated last year
Alternatives and similar repositories for CS-Eval
Users that are interested in CS-Eval are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆172Sep 22, 2025Updated 11 months ago
- CyberMetric dataset☆126May 27, 2026Updated 3 months ago
- A next-generation unified program analysis framework for semantic extraction and security validation, supporting any programming language…☆17Aug 11, 2026Updated 3 weeks ago
- NL prompts generating code with LLMs covering security-relevant scenarios☆59Oct 4, 2024Updated last year
- Metis: Understanding and Enhancing Regular Expressions in Network☆14Aug 19, 2022Updated 4 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- The D-CIPHER and NYU CTF baseline LLM Agents built for NYU CTF Bench☆161Jul 17, 2026Updated last month
- xAST评价体系,让安全工具不再“黑盒”. The xAST evaluation benchmark makes security tools no longer a "black box".☆493May 21, 2026Updated 3 months ago
- 用于检测gradle项目的第三方依赖组件是否存在安全漏洞。☆25Apr 12, 2022Updated 4 years ago
- ☆13Apr 22, 2024Updated 2 years ago
- 3cb: Catastrophic Cyber Capabilities Benchmarking of Large Language Models☆17Oct 30, 2024Updated last year
- ☆24Dec 16, 2024Updated last year
- ☆116Apr 3, 2024Updated 2 years ago
- MOSEC-X-PLUGIN 后端API服务☆24Aug 11, 2020Updated 6 years ago
- [NAACL 2024 Findings] Deja vu: Contrastive Historical Modeling with Prefix-tuning for Temporal Knowledge Graph Reasoning☆14Jul 8, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 关于蜜罐的一些微小的统计工作☆30Aug 26, 2020Updated 6 years ago
- This repository demonstrates a security vulnerability in MCP (Model Context Protocol ) servers that allows for remote code execution and …☆24Apr 21, 2025Updated last year
- The goal of this repo is to become a benchmark for pentesting☆24Oct 25, 2024Updated last year
- A curated list of research resources in automated vulnerability detection (AVD)☆47Nov 25, 2024Updated last year
- 《网络空间安全导论》课程配套资源☆32Jan 27, 2026Updated 7 months ago
- ☆319Jul 9, 2026Updated last month
- XBOW Validation Benchmarks☆697Jul 7, 2026Updated last month
- SecGPT网络安全大模型☆3,103Jun 25, 2025Updated last year
- 用于检测python项目的第三方依赖组件是否存在安全漏洞。☆23Aug 11, 2020Updated 6 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Codebase for Obfuscated Activations Bypass LLM Latent-Space Defenses☆33Feb 11, 2025Updated last year
- A fuzzy parser for C/C++ that creates semantic code property graphs☆36Oct 15, 2020Updated 5 years ago
- AgentGuard: Zero-Trust Security Foundation for AI Agents☆134Updated this week
- A collection of scripts to aid in reverse engineering and exploit development.☆24Oct 3, 2021Updated 4 years ago
- The repository of paper "HackMentor: Fine-Tuning Large Language Models for Cybersecurity".☆145May 30, 2024Updated 2 years ago
- Pwning AI Code Interpreters for fun and profit - by Phantom Labs☆27Aug 3, 2026Updated last month
- Tempolocus is a time-series activity patterns and approximate location inference☆18Updated this week
- A curated list of GPT agents for cybersecurity☆12Oct 2, 2024Updated last year
- ☆11Oct 13, 2020Updated 5 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Netgpt: Generative pretrained transformer for network traffic☆34Jan 10, 2025Updated last year
- An automated tool to test AI models against scope manipulation (deceiving an AI agent about its real target).☆29Updated this week
- ☆105Jul 24, 2025Updated last year
- Training Language Model Agents to Find Vulnerabilities with CTF-Dojo☆67Jan 10, 2026Updated 7 months ago
- Rust语言安全相关分析☆22Jan 20, 2022Updated 4 years ago
- SC-Safety: 中文大模型多轮对抗安全基准☆152Mar 15, 2024Updated 2 years ago
- An extended version of SecureBERT, trained on top of both base and large version of RoBERTa using 10 GB cybersecurity-related data☆35Jan 26, 2024Updated 2 years ago