π¦Ύ SeClaw: The Security Armored Personal AI Assistant
β31Aug 27, 2026Updated this week
Alternatives and similar repositories for seclaw
Users that are interested in seclaw are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- DynAuditClaw β A security audit skill that dynamically discovers your OpenClaw agent's real configuration, designs targeted attack scenarβ¦β15Apr 6, 2026Updated 4 months ago
- The official implementation of the paper "AgentDyn: Are Your Agent Security Defenses Deployable in Real-World Dynamic Environments?"β76May 19, 2026Updated 3 months ago
- [CCS 2026] The official implementation of our CCS 2026 paper "ReasoningBomb: A Stealthy Denial-of-Service Attack by Inducing Pathologicalβ¦β17Aug 5, 2026Updated 3 weeks ago
- [ACL 2025] The official implementation of the paper "PIGuard: Prompt Injection Guardrail via Mitigating Overdefense for Free".β81Dec 4, 2025Updated 8 months ago
- [NeurIPS 2023] The official implementation of paper "Prototype-based Aleatoric Uncertainty Quantification for Cross-modal Retrieval" acceβ¦β28May 14, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- The code of "Image-text Retrieval via Preserving Main Semantic of Vision" in ICME 2023.β15Dec 25, 2023Updated 2 years ago
- The official implementation of our pre-print paper "Automatic and Universal Prompt Injection Attacks against Large Language Models".β74Oct 23, 2024Updated last year
- The code of the paper of "A Differentiable Semantic Metric Approximation in Probabilistic Embedding for Cross-Modal Retrieval" accepted bβ¦β19Jan 16, 2024Updated 2 years ago
- The code of the paper "Negative Pre-aware for Noisy Cross-modal Matching" in AAAI 2024.β31Jul 22, 2026Updated last month
- CopT: Contrastive On-Policy Thinking with Continuous Spaces for General and Agentic Reasoningβ18May 21, 2026Updated 3 months ago
- [ACL 2025] The official code for "AGrail: A Lifelong Agent Guardrail with Effective and Adaptive Safety Detection".β44Aug 12, 2026Updated 2 weeks ago
- β24Jun 13, 2024Updated 2 years ago
- Code for ICLR 2025 Failures to Find Transferable Image Jailbreaks Between Vision-Language Modelsβ37Jun 1, 2025Updated last year
- An implementation for MLLM oversensitivity evaluationβ18Nov 16, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Accepted by IJCAI-24 Survey Trackβ234Aug 25, 2024Updated 2 years ago
- Multi-dimensional analysis of orthogonal safety directions in LLM alignmentβ23Jun 12, 2026Updated 2 months ago
- [NeurIPS 2025] More Thinking, Less Seeing? Assessing Amplified Hallucination in Multimodal Reasoning Modelsβ82May 31, 2025Updated last year
- ReasoningShield: Safety Detection over Reasoning Traces of Large Reasoning Modelsβ26Aug 2, 2026Updated 3 weeks ago
- Seeing What You Miss: Vision-Language Pre-training with Semantic Completion Learningβ20Dec 21, 2023Updated 2 years ago
- [ICLR 2025] This repo is the official implementation of "The Labyrinth of Links: Navigating the Associative Maze of Multi-modal LLMs".β13Jan 25, 2025Updated last year
- [COLM 2024] JailBreakV-28K: A comprehensive benchmark designed to evaluate the transferability of LLM jailbreak attacks to MLLMs, and furβ¦β96May 9, 2025Updated last year
- Awesome Jailbreak, red teaming arxiv papers (Automatically Update Every 12th hours)β120Updated this week
- [COLING 2025π₯] Evolver: Chain-of-Evolution Prompting to Boost Large Multimodal Models for Hateful Meme Detectionβ17Jan 21, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [CVPR 2025] Official implementation for "Steering Away from Harm: An Adaptive Approach to Defending Vision Language Model Against Jailbreβ¦β62Jul 5, 2025Updated last year
- PFI: Prompt Flow Integrity to Prevent Privilege Escalation in LLM Agentsβ31Mar 26, 2025Updated last year
- β30May 14, 2026Updated 3 months ago
- Demo code for the paper: One Thing to Fool them All: Generating Interpretable, Universal, and Physically-Realizable Adversarial Featuresβ12Nov 30, 2023Updated 2 years ago
- The official implementation of the paper "AgentLAB: Benchmarking LLM Agents against Long-Horizon Attacks"β30Jun 1, 2026Updated 2 months ago
- β10Apr 17, 2024Updated 2 years ago
- Progent: Securing AI Agents with Privilege Controlβ48May 14, 2026Updated 3 months ago
- [AAAI'25 (Oral)] Jailbreaking Large Vision-language Models via Typographic Visual Promptsβ212Jun 26, 2025Updated last year
- The official implementation for "Towards Transferable Targeted 3D Adversarial Attack in the Physical World" (CVPR, 2024))β44Aug 6, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Official Implementation of CODEβ17Sep 26, 2024Updated last year
- [EMNLP 25] An effective and interpretable weight-editing method for mitigating overly short reasoning in LLMs, and a mechanistic study unβ¦β19Updated this week
- A Task of Fictitious Unlearning for VLMsβ27Apr 6, 2025Updated last year
- β33Mar 16, 2025Updated last year
- Code of paper: Probing the Difficulty Perception Mechanism of Large Language Modelsβ19Mar 17, 2026Updated 5 months ago
- A highly optimized library for building markov random fields with pytorch.β18Nov 2, 2023Updated 2 years ago
- Skill-Inject: Measuring Agent Vulnerability to Skill File Attacksβ93Aug 3, 2026Updated 3 weeks ago