[ICLR 2025 Spotlight] Official Implementation for ToST (Token Statistics Transformer)
☆135Feb 25, 2025Updated last year
Alternatives and similar repositories for ToST
Users that are interested in ToST are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The official repo for LIFT: Language-Image Alignment with Fixed Text Encoders☆43Jun 10, 2025Updated last year
- Official repository of Polarity-aware Linear Attention for Vision Transformers (ICLR 2025)☆92Aug 12, 2026Updated last month
- Flash-Linear-Attention models beyond language☆21Aug 28, 2025Updated last year
- This repository includes the official implementation our paper "Scaling White-Box Transformers for Vision"☆48Jun 3, 2024Updated 2 years ago
- Official PyTorch implementation of The Linear Attention Resurrection in Vision Transformer☆15Sep 7, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [𝗜𝗖𝗠𝗟 𝟮𝟬𝟮𝟲] Dispersion loss counteracts embedding condensation and improves generalization in small language models☆21May 21, 2026Updated 4 months ago
- [ICML 2025] Fourier Position Embedding: Enhancing Attention’s Periodic Extension for Length Generalization☆120Jun 2, 2025Updated last year
- [ECCV2022] Gumbel Optimised Loss for Long Tailed Instance Segmentation.☆18Nov 24, 2022Updated 3 years ago
- [NeurIPS 2024] Offical PyTorch implementation of All-in-One Image Coding for Joint Human-Machine Vision with Multi-Path Aggregation☆23Jul 8, 2025Updated last year
- Official implementation of the paper "Scalable Image Coding for Humans and Machines Using Feature Fusion Network".☆16Jan 8, 2025Updated last year
- Code for APLA: A Simple Adaptation Method for Vision Transformers☆16Apr 3, 2025Updated last year
- WeConvene: Learned Image Compression with Wavelet-Domain Convolution and Entropy Model☆47Aug 15, 2025Updated last year
- ☆93Aug 18, 2024Updated 2 years ago
- Here is the resources and code for the LotteryCodec.☆28Nov 3, 2025Updated 10 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [NeurIPS 2024] Image Understanding Makes for A Good Tokenizer for Image Generation☆21Dec 17, 2024Updated last year
- the code of GRFormer: Grouped Residual Self-Attention for Lightweight Single Image Super-Resolution☆26May 16, 2024Updated 2 years ago
- Seeing from Another Perspective: Evaluating Multi-View Understanding in MLLMs☆71Mar 22, 2026Updated 6 months ago
- [NeurIPS25 Spotlight] Official Implementation for CBSA (Contract-and-Broadcast Self-Attention)☆36Apr 3, 2026Updated 5 months ago
- All Points Matter: Entropy-Regularized Distribution Alignment for Weakly-supervised 3D Segmentation (NeurIPS 2023)☆32Nov 3, 2023Updated 2 years ago
- [ICML 2025] SparseLoRA: Accelerating LLM Fine-Tuning with Contextual Sparsity☆79Mar 10, 2026Updated 6 months ago
- Code for the paper "Interpreting and Improving Diffusion Models from an Optimization Perspective", appearing in ICML 2024☆15Sep 30, 2024Updated last year
- [ICCV2025 highlight]Rectifying Magnitude Neglect in Linear Attention☆64Jul 24, 2025Updated last year
- A Tight-fisted Optimizer (Tiger), implemented in PyTorch.☆12Jun 26, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆15Apr 6, 2023Updated 3 years ago
- ☆15Dec 3, 2023Updated 2 years ago
- UniGSC is a highly modular and extensible framework for compressing static and dynamic Gaussian Splats, supporting both video and point c…☆20Dec 18, 2025Updated 9 months ago
- Stick-breaking attention☆64Jul 1, 2025Updated last year
- Official implementation of "ImagineFSL: Self-Supervised Pretraining Matters on Imagined Base Set for VLM-based Few-shot Learning" [CVPR 2…☆30Sep 1, 2025Updated last year
- [CVPR 2025 Highlight] Meta LoRA / MetaPEFT: Meta-Learning Hyperparameters for Parameter-Efficient Fine-Tuning (LoRA, Adapter, Prompt Tuni…☆19Sep 12, 2026Updated last week
- [EMNLP 2024] Quantize LLM to extremely low-bit, and finetune the quantized LLMs☆17Jul 18, 2024Updated 2 years ago
- Unofficial implementation of Tensorial Radiance Fields (Chen & Xu ‘22)☆38Feb 17, 2026Updated 7 months ago
- A high-efficiency text embedding and reranking model based on RWKV architecture.☆22Sep 12, 2026Updated last week
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆39Oct 16, 2024Updated last year
- Official pytorch implementation for ESCNet:Edge-Semantic Collaborative Network for Camouflaged Object Detection☆24Jul 3, 2026Updated 2 months ago
- [AAAI 2026] Empowering DINO Representations for Underwater Instance Segmentation via Aligner and Prompter☆57Feb 3, 2026Updated 7 months ago
- [NeurIPS 2025 Spotlight] TPA: Tensor ProducT ATTenTion Transformer (https://arxiv.org/abs/2501.06425)☆463Sep 4, 2026Updated 2 weeks ago
- A generalized framework for subspace tuning methods in parameter efficient fine-tuning.☆182Jan 29, 2026Updated 7 months ago
- ☆21Apr 3, 2025Updated last year
- [NeurIPS2024] Causal Context Adjustment Loss for Learned Image Compression☆66Apr 4, 2025Updated last year