☆34Nov 16, 2025Updated 8 months ago
Alternatives and similar repositories for latentqa
Users that are interested in latentqa are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆20Nov 28, 2024Updated last year
- ☆12Nov 13, 2024Updated last year
- This is the boilerplate for django project. There are so many settings configurations☆10Nov 7, 2025Updated 8 months ago
- ☆33Nov 28, 2024Updated last year
- [ICLR 2025] This repository contains the code to reproduce the results from our paper From Sparse Dependence to Sparse Attention: Unveili…☆12Mar 7, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Improving Steering Vectors by Targeting Sparse Autoencoder Features☆29Nov 20, 2024Updated last year
- [ICLR 25] A novel framework for building intrinsically interpretable LLMs with human-understandable concepts to ensure safety, reliabilit…☆33Feb 5, 2026Updated 5 months ago
- ☆25Mar 30, 2026Updated 3 months ago
- NestJS project template, configured with prisma and ejs☆12Dec 1, 2024Updated last year
- ☆18Aug 19, 2024Updated last year
- Evaluate interpretability methods on localizing and disentangling concepts in LLMs.☆58Oct 30, 2025Updated 8 months ago
- ☆15Jan 20, 2026Updated 6 months ago
- [ICLR 2025] "Rethinking LLM Unlearning Objectives: A Gradient Perspective and Go Beyond"☆16Feb 27, 2025Updated last year
- Localization of Knowledge in Text-to-Image Models☆11Oct 8, 2024Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- code for EMNLP 2024 paper: Neuron-Level Knowledge Attribution in Large Language Models☆52Nov 17, 2024Updated last year
- This repository contains the code and data for the paper "SelfIE: Self-Interpretation of Large Language Model Embeddings" by Haozhe Chen,…☆58Dec 9, 2024Updated last year
- Code repository for "Eliciting Secret Knowledge from Language Models"☆24Mar 30, 2026Updated 3 months ago
- ☆114Aug 8, 2024Updated last year
- Improving Your Model Ranking on Chatbot Arena by Vote Rigging (ICML 2025)☆27Feb 25, 2025Updated last year
- [ICLR 2022] Official Code Repository for "TRGP: TRUST REGION GRADIENT PROJECTION FOR CONTINUAL LEARNING"☆22Oct 5, 2022Updated 3 years ago
- ☆25Jul 8, 2026Updated 2 weeks ago
- ☆16Mar 13, 2025Updated last year
- ☆19Mar 5, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆178May 1, 2026Updated 2 months ago
- Competition of Mechanisms: Tracing How Language Models Handle Facts and Counterfactuals; ACL 2024☆13May 24, 2024Updated 2 years ago
- Official codebase for "Analyzing the Generalization and Reliability of Steering Vectors"☆22Dec 14, 2024Updated last year
- Progetto per la prova finale di Ingegneria del Software 2023-2024 al Politecnico di Milano☆10Oct 19, 2024Updated last year
- ☆260Nov 22, 2024Updated last year
- 免费梯子,免费VPN,真正免费的的VPN,shadowsocks,v2rey,官网地址www.dragonvpn.cc☆13Sep 4, 2024Updated last year
- ☆25Jan 28, 2025Updated last year
- A curated list of resources for activation engineering☆140Oct 2, 2025Updated 9 months ago
- Universal Neurons in GPT2 Language Models☆30May 28, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [ICML25] Official repo for "Towards LLM Unlearning Resilient to Relearning Attacks: A Sharpness-Aware Minimization Perspective and Beyond…☆24Sep 27, 2025Updated 9 months ago
- ☆15Oct 7, 2024Updated last year
- Stanford NLP Python library for benchmarking the utility of LLM interpretability methods☆210Mar 12, 2026Updated 4 months ago
- ☆41Dec 19, 2024Updated last year
- Concreteness☆20Nov 22, 2022Updated 3 years ago
- ☆43Nov 16, 2021Updated 4 years ago
- Multi-dimensional analysis of orthogonal safety directions in LLM alignment☆23Jun 12, 2026Updated last month