[FCS'24] LVLM Safety paper
☆19Jan 4, 2025Updated last year
Alternatives and similar repositories for LVLM-Safety
Users that are interested in LVLM-Safety are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [preprint] sparsity☆23Jul 26, 2026Updated 3 weeks ago
- VAEGAN, I Love u☆16Aug 15, 2023Updated 3 years ago
- The code for paper "MemGym: a Long-Horizon Memory Environment for LLM Agents".☆21Jun 2, 2026Updated 2 months ago
- [ICML'25] Our study systematically investigates massive values in LLMs' attention mechanisms. First, we observe massive values are concen…☆87Jun 20, 2025Updated last year
- Code for ACL 2024 findings paper "wav2vec-S: Adapting Pre-trained Speech Models for Streaming"☆13Apr 21, 2026Updated 3 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Demo code for the paper: One Thing to Fool them All: Generating Interpretable, Universal, and Physically-Realizable Adversarial Features☆12Nov 30, 2023Updated 2 years ago
- ☆13Jan 9, 2024Updated 2 years ago
- ☆19Mar 25, 2025Updated last year
- Agentic Risk Standard is a settlement-layer standard for trustworthy transactions with AI Agent☆34Mar 29, 2026Updated 4 months ago
- ☆18Nov 16, 2021Updated 4 years ago
- Code for EMNLP 2022 main conference paper "Information-Transport-based Policy for Simultaneous Translation"☆13Nov 3, 2022Updated 3 years ago
- Code and dataset for the paper: "Can Editing LLMs Inject Harm?" [AAAI'26]☆21Dec 26, 2025Updated 7 months ago
- ☆19Jun 3, 2024Updated 2 years ago
- [ACL 2024] CodeAttack: Revealing Safety Generalization Challenges of Large Language Models via Code Completion☆60Oct 1, 2025Updated 10 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆16Mar 22, 2025Updated last year
- The code for ICLR2025 paper "SLMRec: Empowering Small Language Models for Sequential Recommendation".☆52Jun 16, 2025Updated last year
- ☆10Jun 30, 2026Updated last month
- Your finetuned model's back to its original safety standards faster than you can say "SafetyLock"!☆11Oct 16, 2024Updated last year
- ☆22Oct 12, 2024Updated last year
- [ACL 2024] An easily extensible framework for simultaneous, text-to-text neural machine translation (SimulMT) for LLMs.☆18Apr 21, 2025Updated last year
- A novel approach to improve the safety of large language models, enabling them to transition effectively from unsafe to safe state.☆72May 22, 2025Updated last year
- Official code for Guiding Language Model Math Reasoning with Planning Tokens☆19Feb 29, 2024Updated 2 years ago
- ☆11Jan 19, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [ICLR 2025] Code&Data for the paper "Super(ficial)-alignment: Strong Models May Deceive Weak Models in Weak-to-Strong Generalization"☆15Jun 21, 2024Updated 2 years ago
- A toolkit for testing and improving named entity recognition [ESEC/FSE'23]☆11Aug 31, 2023Updated 2 years ago
- This is a project based on machine learning and deep learning method for playing Gobang by controlling mechanical arm(利用机械臂下五子棋)☆13Apr 16, 2023Updated 3 years ago
- Tasks for describing differences between text distributions.☆17Aug 9, 2024Updated 2 years ago
- ☆13Sep 12, 2024Updated last year
- EmojiCrypt: Prompt Encryption for Secure Communication with Large Language Models☆26Feb 21, 2024Updated 2 years ago
- CausalGym: Benchmarking causal interpretability methods on linguistic tasks☆54Nov 30, 2024Updated last year
- ☆119Jan 23, 2026Updated 6 months ago
- Papers about Explainable AI (Deep Learning-based)☆29Nov 14, 2025Updated 9 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- implementation of paper "Large Language Models are In-Context Semantic Reasoners rather than Symbolic Reasoners"☆20Aug 17, 2023Updated 3 years ago
- Demonstration Agents for AIOS☆19Dec 25, 2024Updated last year
- Source code of paper "Systematic Assessment of Factual Knowledge in Large Language Models" - EMNLP Findings 2023☆18Mar 17, 2026Updated 5 months ago
- Collection of Reverse Engineering in Large Model☆35Jan 8, 2025Updated last year
- The implementation for our paper, "Improving Simultaneous Machine Translation with Monolingual Data," accepted to AAAI 2023. 🎉☆12Jul 19, 2023Updated 3 years ago
- A large-scale dataset composed of high-quality synthetic images aimed at evaluating social biases in LVLMs☆16Apr 7, 2026Updated 4 months ago
- A lightweight tool for detecting bugs on Graph Database Management Systems☆16Jan 9, 2024Updated 2 years ago