π Fast, accurate & memory efficient LLM distillation via tokens, classes and samples selection
β20Feb 3, 2026Updated 6 months ago
Alternatives and similar repositories for SE-KD3x
Users that are interested in SE-KD3x are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This repository contains the code for applying One-Token Approximation to a pretrained language model using subword-level tokenization.β12May 7, 2020Updated 6 years ago
- Hyper-networks for Unified Visual Representation (HUVR) use implicit neural representation to bridge the gap between understanding and geβ¦β32Jan 23, 2026Updated 6 months ago
- [ECCV 26'] Official codebase for the paper LaViTβ35Jul 30, 2026Updated last week
- β14Dec 12, 2024Updated last year
- An arbitrage bot is a smart contract connected to an external automation script that controls its operation.β2,595Updated this week
- Deploy open-source AI quickly and easily - Special Bonus Offer β’ AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- For further understanding the wide array of emotions embedded in human speech, we are introducing an emotional speech corpus. In contrastβ¦β11Oct 29, 2018Updated 7 years ago
- [EMNLP 2024 Main] Code for the paper "Dissecting Fine-Tuning Unlearning in Large Language Models"β14Oct 10, 2024Updated last year
- Pack of LLMs: Model Fusion at Test-Time via Perplexity Optimizationβ15Apr 25, 2024Updated 2 years ago
- MSP project: Latent Space Factorisation and Manipulation via Matrix Subspace Projection (ICML2020)β14Dec 4, 2021Updated 4 years ago
- Data and code for "Understanding Linearity of Cross-Lingual Word Embedding Mappings" (TMLR 2022)β12Jun 8, 2022Updated 4 years ago
- Localization of Knowledge in Text-to-Image Modelsβ11Oct 8, 2024Updated last year
- A Jensen-Shannon Divergence Driven Mechanistic Study of Context Attribution in Retrieval-Augmented Generationβ15Aug 28, 2025Updated 11 months ago
- β17May 5, 2024Updated 2 years ago
- Code for AttriBoT from "AttriBoT: A Bag of Tricks for Efficiently Approximating Leave-One-Out Context Attribution"β15Apr 21, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Implementation code for ACL2024οΌAdvancing Parameter Efficiency in Fine-tuning via Representation Editingβ15Apr 20, 2024Updated 2 years ago
- 2020 PTA History Test Questionsβ15Feb 7, 2023Updated 3 years ago
- β14Jul 28, 2025Updated last year
- A PyTorch implementation of alpha-GANβ15Jul 27, 2017Updated 9 years ago
- β19Sep 1, 2025Updated 11 months ago
- β15Jan 20, 2026Updated 6 months ago
- [EMNLP 2025] Layer-Aware Representation Filtering: Purifying Finetuning Data to Preserve LLM Safety Alignmentβ16Jul 22, 2025Updated last year
- β21Mar 25, 2023Updated 3 years ago
- Documentation for Enron data.β17Nov 23, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code for "Cross-Lingual Word Embedding Refinement by β1 Norm Optimisation" (NAACL 2021)β17Jun 16, 2022Updated 4 years ago
- Projects and Exercises for Udacity Data Analyst Nanodegree (2014-2015)β11May 29, 2015Updated 11 years ago
- β16Jul 1, 2024Updated 2 years ago
- [ICLR'24] Official code for "C-TPT: Calibrated Test-Time Prompt Tuning for Vision-Language Models via Text Feature Dispersion"β23Jun 9, 2024Updated 2 years ago
- Official code for ICML 2024 paper, "Connecting the Dots: Collaborative Fine-tuning for Black-Box Vision-Language Models"β19Jun 12, 2024Updated 2 years ago
- Code for our paper "Decomposing The Dark Matter of Sparse Autoencoders"β23Feb 6, 2025Updated last year
- ζ°ε θΊηζΆζοΌθͺθ±θΏεδΏ‘ζ―ζ±ζ»β15Mar 13, 2021Updated 5 years ago
- This repository presents the original implementation of Pretraining Data Detection for Large Language Models: A Divergence-based Calibratβ¦β23May 21, 2025Updated last year
- collab-dev - Collaboration Metrics for Code Reviewsβ23May 12, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Pretraining codes for ECG Semantic Integrator (ESI): A Foundation ECG Model Pretrained with LLM-Enhanced Cardiological Textβ20Feb 16, 2026Updated 5 months ago
- Solutions of Computer Systems: A Programmer's Perspective (3rd Edition) by Randal E. Bryant, David R. O'Hallaronβ17Sep 13, 2018Updated 7 years ago
- breadth-first search in parallelβ18May 30, 2013Updated 13 years ago
- Code repo for the model organisms and convergent directions of EM papers.β74Sep 22, 2025Updated 10 months ago
- Different hyperparameter optimization methods to get best performance for your Machine Learning Modelsβ19Oct 4, 2020Updated 5 years ago
- β21Sep 17, 2022Updated 3 years ago
- [ICLR 2025 Spotlight] Code release for "Sharpness-Aware Minimization Efficiently Selects Flatter Minima Late In Training"β19Feb 20, 2025Updated last year