[ECCV 2024] Official PyTorch implementation of RoPE-ViT "Rotary Position Embedding for Vision Transformer"
☆470Oct 29, 2025Updated 9 months ago
Alternatives and similar repositories for rope-vit
Users that are interested in rope-vit are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- N-dimensional Rotary Position Embeddings for PyTorch☆84Feb 14, 2024Updated 2 years ago
- [ECCV 2024] Official PyTorch implementation of LUT "Learning with Unmasked Tokens Drives Stronger Vision Learners"☆14Dec 1, 2024Updated last year
- [ICLR'25 Oral] Representation Alignment for Generation: Training Diffusion Transformers Is Easier Than You Think☆1,701Mar 16, 2025Updated last year
- Implementation of Rotary Embeddings, from the Roformer paper, in Pytorch☆822Jun 20, 2026Updated 2 months ago
- This repo contains the code for 1D tokenizer and generator☆1,172Mar 20, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICML-2025] We introduce Lie group Relative position Encodings (LieRE) that goes beyond RoPE in supporting n-dimensional inputs.☆14Aug 8, 2025Updated last year
- [ECCV2024][ICCV2023] Official PyTorch implementation of SeiT++ and SeiT☆56Aug 12, 2024Updated 2 years ago
- [ECCV 2024] Official PyTorch implementation of "HYPE: Hyperbolic Entailment Filtering for Underspecified Images and Texts"☆20Nov 22, 2024Updated last year
- ☆58Aug 16, 2025Updated last year
- Masked Diffusion Transformer is the SOTA for image synthesis. (ICCV 2023)☆596Apr 23, 2024Updated 2 years ago
- [ICML 2024 Spotlight] FiT: Flexible Vision Transformer for Diffusion Model☆436Nov 10, 2024Updated last year
- Autoregressive Model Beats Diffusion: 🦙 Llama for Scalable Image Generation☆1,966Aug 15, 2024Updated 2 years ago
- PyTorch implementation of MAR+DiffLoss https://arxiv.org/abs/2406.11838☆1,951Feb 20, 2026Updated 6 months ago
- VisionLLaMA: A Unified LLaMA Backbone for Vision Tasks☆391Jul 9, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official PyTorch Implementation of "Scalable Diffusion Models with Transformers"☆8,689May 31, 2024Updated 2 years ago
- Model Stock: All we need is just a few fine-tuned models☆129Aug 9, 2025Updated last year
- Official codebase used to develop Vision Transformer, SigLIP, MLP-Mixer, LiT and more.☆3,530May 19, 2025Updated last year
- This repository provides the code and model checkpoints for AIMv1 and AIMv2 research projects.☆1,424Aug 4, 2025Updated last year
- [ECCV2024] Official implementation of paper, "DenseNets Reloaded: Paradigm Shift Beyond ResNets and ViTs".☆154Aug 8, 2024Updated 2 years ago
- Official Pytorch implementation of MuCo: Multi-turn Contrastive Learning for Multimodal Embedding Model (CVPR 2026)☆15Apr 16, 2026Updated 4 months ago
- ImageNet-12k subset of ImageNet-21k (fall11)