[AAAI-2025] The offical code for SiTo (Similarity-based Token Pruning for Stable Diffusion Models)
☆47Jun 2, 2025Updated last year
Alternatives and similar repositories for SiTo
Users that are interested in SiTo are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICLR2025] Accelerating Diffusion Transformers with Token-wise Feature Caching☆224Mar 14, 2025Updated last year
- 📚 Collection of token-level model compression resources.☆203Sep 3, 2025Updated last year
- [ICME 2024 Oral] DARA: Domain- and Relation-aware Adapters Make Parameter-efficient Tuning for Visual Grounding☆22Feb 26, 2025Updated last year
- Official PyTorch code for ICLR 2025 paper "Gnothi Seauton: Empowering Faithful Self-Interpretability in Black-Box Models"☆24Mar 4, 2025Updated last year
- Official Implementation of Paper FOLDER (ICCV2025) and Turbo (ECCV2024)☆15Jun 27, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Aiming to integrate most existing feature caching-based diffusion acceleration schemes into a unified framework.☆112Oct 23, 2025Updated 11 months ago
- [ICCV2025] From Reusing to Forecasting: Accelerating Diffusion Models with TaylorSeers☆416Mar 2, 2026Updated 6 months ago
- (ICCV2025) EEdit⚡: Rethinking the Spatial and Temporal Redundancy for Efficient Image Editing☆63Sep 17, 2025Updated last year
- This is the open-source code for TokenCarve.☆25Jan 23, 2026Updated 8 months ago
- Multi-Stage Vision Token Dropping: Towards Efficient Multimodal Large Language Model☆38Jan 8, 2025Updated last year
- Source code for SWIFT, an efficient reward model.☆21Jan 13, 2026Updated 8 months ago
- [ICML 2026]☆18Jul 4, 2026Updated 2 months ago
- Based on BrainTransformers, BrainGPTForCausalLM is a Large Language Model (LLM) implemented using Spiking Neural Networks (SNN). We are e…☆39Oct 22, 2024Updated last year
- [EMNLP 2025 main 🔥] Code for "Stop Looking for Important Tokens in Multimodal Language Models: Duplication Matters More"☆122Oct 12, 2025Updated 11 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Official Implementation for paper "Pretraining A Large Language Model using Distributed GPUs: A Memory-Efficient Decentralized Paradigm"☆24May 8, 2026Updated 4 months ago
- ☆26Sep 4, 2025Updated last year
- SAM4SS: Tailoring SAM and SAM2 for Semantic Segmentation☆11Jul 31, 2024Updated 2 years ago
- Use Hermes Agent as the control plane for local coding agents like Codex, Kimi Code, Claude Code, OpenCode, and Gemini CLI.☆39Jul 22, 2026Updated 2 months ago
- [ICASSP 2024] VGDiffZero: Text-to-image Diffusion Models Can Be Zero-shot Visual Grounders☆17Feb 11, 2025Updated last year
- DisCa: Accelerating Video Diffusion Transformers with Distillation-Compatible Learnable Feature Caching☆25Apr 15, 2026Updated 5 months ago
- Memorize-and-Generate: Towards Long-Term Consistency in Real-Time Video Generation☆19Mar 20, 2026Updated 6 months ago
- NeurIPS 2025☆17Sep 24, 2025Updated last year
- An Open-Source Processor for Accelerating Spiking Neural Network☆14Sep 30, 2022Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Official code for paper: [CLS] Attention is All You Need for Training-Free Visual Token Pruning: Make VLM Inference Faster.☆116Jun 29, 2025Updated last year
- Code for our ICCV 2025 paper "Adaptive Caching for Faster Video Generation with Diffusion Transformers"☆173Nov 5, 2024Updated last year
- 东北大学2019届[2015级]本科毕设Latex模版☆11Jun 16, 2019Updated 7 years ago
- RISC-V-based many-core neuromorphic architecture☆18Aug 1, 2026Updated last month
- ☆18Apr 5, 2024Updated 2 years ago
- Paper writing guide for Zhuang Liu Lab @ Princeton University☆37Jun 24, 2026Updated 3 months ago
- Communication-Efficient Diffusion Denoising Parallelization via Reuse-then-Predict Mechanism (NIPS'25)☆17Oct 6, 2025Updated 11 months ago
- Official Code Implementation for 'A Simple Early Exiting Framework for Accelerated Sampling in Diffusion Models'☆20Jul 24, 2024Updated 2 years ago
- [ICML 2024] Unveiling and Harnessing Hidden Attention Sinks: Enhancing Large Language Models without Training through Attention Calibrati…☆45Jun 30, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [NeurIPS 2025] Official code for paper: Beyond Attention or Similarity: Maximizing Conditional Diversity for Token Pruning in MLLMs.☆109Sep 20, 2025Updated last year
- Official Implementation of "Semantics-Consistent Feature Search for Self-Supervised Visual Representation Learning" in AAAI2024.☆13Feb 28, 2024Updated 2 years ago
- ☆192Jan 14, 2025Updated last year
- Code for Panoramic Semantic Segmentation☆16Apr 26, 2024Updated 2 years ago
- This is the repository of the Paper GlowGAN☆17Oct 22, 2023Updated 2 years ago
- Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation☆15Aug 11, 2025Updated last year
- [NeurIPS'25] KVCOMM: Online Cross-context KV-cache Communication for Efficient LLM-based Multi-agent Systems☆18Nov 1, 2025Updated 10 months ago