Code for Chain-of-Thought Hijacking
☆28Nov 10, 2025Updated 9 months ago
Alternatives and similar repositories for Hijacking
Users that are interested in Hijacking are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆17Mar 10, 2026Updated 5 months ago
- ☆12Jun 5, 2024Updated 2 years ago
- Master PDF Summarization with Google Bard☆13Feb 29, 2024Updated 2 years ago
- Code and Data for Paper "Controlling Styles in Neural Machine Translation with Activation Prompt" (ACL 2023 Findings)☆16Dec 20, 2022Updated 3 years ago
- An automated pipeline that leverages LLM's meta-learning capability to iteratively design and refine red-teaming systems without human in…☆30May 24, 2026Updated 2 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆22Aug 6, 2023Updated 3 years ago
- ☆40Jul 9, 2025Updated last year
- Vultron is a protocol for Coordinated Vulnerability Disclosure☆21Updated this week
- Code to the paper: The Geometry of Refusal in Large Language Models: Concept Cones and Representational Independence☆35Jul 31, 2025Updated last year
- ☆19Jun 13, 2023Updated 3 years ago
- ☆16Apr 15, 2024Updated 2 years ago
- Code used for "Training Agents to Self-Report Misbehavior"☆18Feb 27, 2026Updated 5 months ago
- Geographical Information Science Tutorial and Practice course in UCAS☆13Dec 12, 2021Updated 4 years ago
- ☆16Nov 14, 2025Updated 9 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official implementation of ICLR 2026 paper "LUMINA: Detecting Hallucinations in RAG System with Context–Knowledge Signals"☆18Jan 31, 2026Updated 6 months ago
- Sequence-Level Mixed Sample Data Augmentation☆22Mar 7, 2021Updated 5 years ago
- ☆15May 13, 2026Updated 3 months ago
- The official repository for MaCh3☆13Updated this week
- Implementation of SLIM, a framework of dynamics skill lifecycle management for agentic reinforcement learning☆22May 12, 2026Updated 3 months ago
- C# Z80 assembler and IDE☆13Feb 10, 2014Updated 12 years ago
- Cybersecurity demo for Neo4j's Connections: Graphs in Cybersecurity 2021☆14May 17, 2022Updated 4 years ago
- Threat Detection Rules (Snort/Sigma/Yara)☆14Jan 23, 2024Updated 2 years ago
- Using LVGL for create UI and control more ESP with ESP NOW protocol.☆14May 1, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Multi-dimensional analysis of orthogonal safety directions in LLM alignment☆23Jun 12, 2026Updated 2 months ago
- RAG Hallucination Detecting By LRP.☆12Mar 31, 2025Updated last year
- DevSecOps for the AI-era CI/CD pipeline. Catches the bugs your AI coding assistant introduces — before they reach production.☆20Jul 14, 2026Updated last month
- Training scenarios for cyber ranges☆15Apr 24, 2020Updated 6 years ago
- An example of how to implement a simple traffic light example, on an Arduino, implemented in 4 ways, a switch case, a table of function p…☆11Jul 11, 2020Updated 6 years ago
- Training hybrid models for dummies.☆31Nov 1, 2025Updated 9 months ago
- Snooze - Log aggregation and alerting☆10Jul 29, 2026Updated 2 weeks ago
- Analyse Social Network of co-authors in DBLP website (https://dblp.uni-trier.de) using NetworkX.☆13May 27, 2020Updated 6 years ago
- An Autonomous Curriculum Reinforcement Learning framework that steers agents to continually learn in specific environments with zero huma…☆42Jun 7, 2026Updated 2 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Our research proposes a novel MoGU framework that improves LLMs' safety while preserving their usability.☆18Jan 14, 2025Updated last year
- ☆22Jul 28, 2026Updated 2 weeks ago
- ☆28Mar 4, 2025Updated last year
- ☆12Jun 11, 2021Updated 5 years ago
- [EMNLP 2025] Reasoning-to-Defend: Safety-Aware Reasoning Can Defend Large Language Models from Jailbreaking☆12Aug 22, 2025Updated 11 months ago
- ☆45Feb 11, 2025Updated last year
- ASIDE: Architectural Separation of Instructions and Data in Language Models [ICLR 2026]☆17Jun 10, 2026Updated 2 months ago