[ICLR'25] Code for KaSA, an official implementation of "KaSA: Knowledge-Aware Singular-Value Adaptation of Large Language Models"
☆22Jan 16, 2025Updated last year
Alternatives and similar repositories for KaSA
Users that are interested in KaSA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [arXiv'26] Official Code for "WebGen-R1: Incentivizing Large Language Models to Generate Functional and Aesthetic Websites with Reinforce…☆20Jul 20, 2026Updated last week
- [ACL'26] Official Code for "ReflexiCoder: Teaching Large Language Models to Self-Reflect on Generated Code and Self-Correct It via Reinfo…☆22Apr 24, 2026Updated 3 months ago
- [ACL 2024 Findings] Light-PEFT: Lightening Parameter-Efficient Fine-Tuning via Early Pruning☆13Sep 2, 2024Updated last year
- Implementation of paper "Approximate Nearest Neighbor Negative Contrastive Learning for Dense Text Retrieval"☆17Jan 10, 2022Updated 4 years ago
- ☆18Apr 19, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆19Nov 16, 2023Updated 2 years ago
- ☆126Jul 6, 2024Updated 2 years ago
- ☆28Jun 9, 2024Updated 2 years ago
- Presents an optimized Apache Beam pipeline for generating sentence embeddings (runnable on Cloud Dataflow).☆20Mar 7, 2022Updated 4 years ago
- [CIKM 2023] This is the official source code of "TrendGCN: Enhancing the Robustness via Adversarial Learning and Joint Spatial-Temporal E…☆51Aug 11, 2023Updated 2 years ago
- generate synthetic data for LLM fine-tuning in arbitrary situations within systematic way☆22Mar 18, 2024Updated 2 years ago
- DCPO: Dynamic Adaptive Clipping for RL☆49Apr 1, 2026Updated 3 months ago
- [NAACL 2025] MiLoRA: Harnessing Minor Singular Components for Parameter-Efficient LLM Finetuning☆21May 31, 2025Updated last year
- [ACL 2025] Official code for ''Learning to Reason from Feedback at Test-Time''.☆13May 16, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [NeurIPS'24 Oral] HydraLoRA: An Asymmetric LoRA Architecture for Efficient Fine-Tuning☆239Dec 3, 2024Updated last year
- (ECCV 2026): Official code for Flash-BoN: Instant Drafts for Inference-Time Scaling in Diffusion Models☆18Jul 9, 2026Updated 2 weeks ago
- This repository shows various ways of deploying a vision model (TensorFlow) from 🤗 Transformers.☆30Aug 22, 2022Updated 3 years ago
- Implementations of Influential Recommender System☆12Oct 29, 2024Updated last year
- Run TFLITE models on the web☆13Jan 2, 2022Updated 4 years ago
- ☆19Jan 3, 2025Updated last year
- ☆17Jun 9, 2024Updated 2 years ago
- https://blog.tensorflow.org/2021/12/continuous-adaptation-for-machine.html☆30Dec 10, 2021Updated 4 years ago
- Code for "Routing Manifold Alignment Improves Generalization of Mixture-of-Experts LLMs"☆19Nov 6, 2025Updated 8 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆17Jun 10, 2025Updated last year
- ICML-2025 (Spotlight) "From Weight-Based to State-Based Fine-Tuning: Further Memory Reduction on LoRA with Parallel Control"☆15May 7, 2026Updated 2 months ago
- Data and code for "Chain-of-Thought in Neural Code Generation: From and For Lightweight Language Models", which accepted in TSE.☆15Jul 3, 2024Updated 2 years ago
- [ICML2025] Make LoRA Great Again: Boosting LoRA with Adaptive Singular Values and Mixture-of-Experts Optimization Alignment☆147Nov 9, 2025Updated 8 months ago
- Contests based Dataset for Code Generation☆13Dec 11, 2022Updated 3 years ago
- PELA: Learning Parameter-Efficient Models with Low-Rank Approximation [CVPR 2024]☆19Apr 14, 2024Updated 2 years ago
- ☆36Aug 23, 2023Updated 2 years ago
- A video question answering dataset that focuses on the dynamics properties of objects (velocity, acceleration) and their collisions withi…☆20Apr 23, 2025Updated last year
- This repository provides a framework to serve LLM(Large Language Model) based applications such as Chatbot.☆18Apr 20, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Code for studying the super weight in LLM☆124Dec 3, 2024Updated last year
- This project breathes life into video characters by using AI to describe their personality and then chat with you as them.☆48Mar 12, 2024Updated 2 years ago
- hllama is a library which aims to provide a set of utility tools for large language models.☆10Apr 16, 2024Updated 2 years ago
- A curated list of awesome multi-modal recommendation.☆10Mar 16, 2022Updated 4 years ago
- showing various ways to serve Keras based stable diffusion☆111Feb 28, 2023Updated 3 years ago
- Application for searching images from natural language queries☆47Dec 10, 2021Updated 4 years ago
- ☆21Apr 16, 2024Updated 2 years ago