[ICLR'25] Code for KaSA, an official implementation of "KaSA: Knowledge-Aware Singular-Value Adaptation of Large Language Models"
☆22Jan 16, 2025Updated last year
Alternatives and similar repositories for KaSA
Users that are interested in KaSA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [arXiv'26] Official Code for "WebGen-R1: Incentivizing Large Language Models to Generate Functional and Aesthetic Websites with Reinforce…☆21Jul 20, 2026Updated 2 months ago
- [ACL'26] Official Code for "ReflexiCoder: Teaching Large Language Models to Self-Reflect on Generated Code and Self-Correct It via Reinfo…☆23Apr 24, 2026Updated 5 months ago
- [ACL 2024 Findings] Light-PEFT: Lightening Parameter-Efficient Fine-Tuning via Early Pruning☆13Sep 2, 2024Updated 2 years ago
- CorDA: Context-Oriented Decomposition Adaptation of Large Language Models for task-aware parameter-efficient fine-tuning(NeurIPS 2024)☆60Jan 13, 2025Updated last year
- Implementation of paper "Approximate Nearest Neighbor Negative Contrastive Learning for Dense Text Retrieval"☆17Jan 10, 2022Updated 4 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Replication package for ISSTA2023 paper - Towards Efficient Fine-tuning of Pre-trained Code Models: An Experimental Study and Beyond☆23Apr 9, 2023Updated 3 years ago
- ☆20Mar 6, 2023Updated 3 years ago
- ☆28Jun 9, 2024Updated 2 years ago
- [CIKM 2023] This is the official source code of "TrendGCN: Enhancing the Robustness via Adversarial Learning and Joint Spatial-Temporal E…☆51Aug 11, 2023Updated 3 years ago
- generate synthetic data for LLM fine-tuning in arbitrary situations within systematic way☆22Mar 18, 2024Updated 2 years ago
- DCPO: Dynamic Adaptive Clipping for RL☆50Apr 1, 2026Updated 5 months ago
- [NAACL 2025] MiLoRA: Harnessing Minor Singular Components for Parameter-Efficient LLM Finetuning☆23May 31, 2025Updated last year
- [ACL 2025] Official code for ''Learning to Reason from Feedback at Test-Time''.☆13May 16, 2025Updated last year
- [NeurIPS'24 Oral] HydraLoRA: An Asymmetric LoRA Architecture for Efficient Fine-Tuning☆243Dec 3, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Facing Changes: Continual Entity Alignment for Growing Knowledge Graphs, ISWC 2022☆24Jul 27, 2022Updated 4 years ago
- ☆12Oct 10, 2021Updated 4 years ago
- This repository shows various ways of deploying a vision model (TensorFlow) from 🤗 Transformers.☆30Aug 22, 2022Updated 4 years ago
- Awesome-Low-Rank-Adaptation☆134Oct 13, 2024Updated last year
- (ECCV 2026): Official code for Flash-BoN: Instant Drafts for Inference-Time Scaling in Diffusion Models☆23Jul 9, 2026Updated 2 months ago
- ☆22Feb 5, 2024Updated 2 years ago
- ☆19Jan 3, 2025Updated last year
- https://blog.tensorflow.org/2021/12/continuous-adaptation-for-machine.html☆30Dec 10, 2021Updated 4 years ago
- ☆10Apr 15, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code for "Routing Manifold Alignment Improves Generalization of Mixture-of-Experts LLMs"☆19Nov 6, 2025Updated 10 months ago
- Demonstration of the Model Training as a CI/CD System in Vertex AI☆29Oct 26, 2021Updated 4 years ago
- ☆17Jun 10, 2025Updated last year
- ICML-2025 (Spotlight) "From Weight-Based to State-Based Fine-Tuning: Further Memory Reduction on LoRA with Parallel Control"☆15May 7, 2026Updated 4 months ago
- [ICML2025] Make LoRA Great Again: Boosting LoRA with Adaptive Singular Values and Mixture-of-Experts Optimization Alignment☆151Nov 9, 2025Updated 10 months ago
- Contests based Dataset for Code Generation☆13Dec 11, 2022Updated 3 years ago
- PELA: Learning Parameter-Efficient Models with Low-Rank Approximation [CVPR 2024]☆19Apr 14, 2024Updated 2 years ago
- ☆36Aug 23, 2023Updated 3 years ago
- A video question answering dataset that focuses on the dynamics properties of objects (velocity, acceleration) and their collisions withi…☆20Apr 23, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Experiments with reasoning models, training techniques, papers☆30Updated this week
- Code for studying the super weight in LLM☆124Dec 3, 2024Updated last year
- hllama is a library which aims to provide a set of utility tools for large language models.☆10Apr 16, 2024Updated 2 years ago
- showing various ways to serve Keras based stable diffusion☆112Feb 28, 2023Updated 3 years ago
- Application for searching images from natural language queries☆47Dec 10, 2021Updated 4 years ago
- ☆32Jun 6, 2025Updated last year
- [TIFS'22]Smart App Attack: Hacking Deep Learning Models in Android Apps☆16Jun 8, 2022Updated 4 years ago