Research code for pixel-based encoders of language (PIXEL)
β348Jul 15, 2025Updated last year
Alternatives and similar repositories for pixel
Users that are interested in pixel are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This repository contains an extension of fairseq for pixel / visual representations of text for machine translation.β37Feb 2, 2024Updated 2 years ago
- π VITRina: VIsual Token Representationsβ11Jun 15, 2023Updated 3 years ago
- β85Dec 4, 2022Updated 3 years ago
- Textual Visual Semantic Dataset for Text Spotting. CVPRW 2020β12Jul 2, 2022Updated 4 years ago
- Code release for SLIP Self-supervision meets Language-Image Pre-trainingβ792Feb 9, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Official pytorch implementation of I2I translation with low resolution conditioningβ23Sep 2, 2021Updated 4 years ago
- Large-scale Self-supervised Pre-training Across Tasks, Languages, and Modalitiesβ81Jan 7, 2026Updated 6 months ago
- Omnivore: A Single Model for Many Visual Modalitiesβ573Nov 12, 2022Updated 3 years ago
- Official PyTorch implementation of the paper "In-Context Learning Unlocked for Diffusion Models"β414Mar 25, 2024Updated 2 years ago
- Language Models Can See: Plugging Visual Controls in Text Generationβ261Jun 1, 2022Updated 4 years ago
- PyTorch code for "Perceiver-VL: Efficient Vision-and-Language Modeling with Iterative Latent Attention" (WACV 2023)β34Feb 5, 2023Updated 3 years ago
- Official implementation and data release of the paper "Visual Prompting via Image Inpainting".β319Aug 7, 2023Updated 2 years ago
- A Multilingual Keyboard Layout-Based Typo Generatorβ17Nov 23, 2025Updated 7 months ago
- A Structured Span Selector (NAACL 2022). A structured span selector with a WCFG for span selection tasks (coreference resolution, semantiβ¦β21Jul 11, 2022Updated 4 years ago
- Proton VPN Special Offer - Get 70% off β’ AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Pix2Seq codebase: multi-tasks with generative modeling (autoregressive and diffusion)β945Nov 7, 2023Updated 2 years ago
- β88Jan 10, 2024Updated 2 years ago
- Python source code for EMNLP 2021 Findings paper: "Subword Mapping and Anchoring Across Languages".β13Sep 17, 2021Updated 4 years ago
- β43Aug 9, 2022Updated 3 years ago
- Masked Siamese Networks for Label-Efficient Learning (https://arxiv.org/abs/2204.07141)β463May 9, 2022Updated 4 years ago
- Patching open-vocabulary models by interpolating weightsβ91Sep 28, 2023Updated 2 years ago
- Code for paper βLanguage Versatilists vs. Specialists: An Empirical Revisiting on Multilingual Transfer Abilityββ15Jun 13, 2023Updated 3 years ago
- Implementation of N-Grammer, augmenting Transformers with latent n-grams, in Pytorchβ81Dec 4, 2022Updated 3 years ago
- β46Jul 5, 2022Updated 4 years ago
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [NeurIPS2023] Official implementation and model release of the paper "What Makes Good Examples for Visual In-Context Learning?"β182Mar 4, 2024Updated 2 years ago
- A Unified Library for Parameter-Efficient and Modular Transfer Learningβ2,822Apr 26, 2026Updated 2 months ago
- β15Jan 14, 2026Updated 6 months ago
- PyTorch code for "Unifying Vision-and-Language Tasks via Text Generation" (ICML 2021)β372Jul 29, 2023Updated 2 years ago
- Official codebase used to develop Vision Transformer, SigLIP, MLP-Mixer, LiT and more.β3,497May 19, 2025Updated last year
- β12Mar 12, 2023Updated 3 years ago
- Paper List for In-context Learning π·β19Jan 3, 2023Updated 3 years ago
- Exploring Visual Prompts for Adapting Large-Scale Modelsβ292Jun 6, 2022Updated 4 years ago
- UDapter is a multilingual dependency parser that uses "contextual" adapters together with language-typology features for language-specifiβ¦β31Dec 5, 2022Updated 3 years ago
- Managed Database hosting by DigitalOcean β’ AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Supervision Exists Everywhere: A Data Efficient Contrastive Language-Image Pre-training Paradigmβ678Sep 19, 2022Updated 3 years ago
- Official repository of OFA (ICML 2022). Paper: OFA: Unifying Architectures, Tasks, and Modalities Through a Simple Sequence-to-Sequence L β¦β2,557Apr 24, 2024Updated 2 years ago
- Official Pytorch implementation for Distilling Image Classifiers in Object detection (NeurIPS2021)β32Feb 10, 2022Updated 4 years ago
- PyTorch code for "VL-Adapter: Parameter-Efficient Transfer Learning for Vision-and-Language Tasks" (CVPR2022)β212Dec 18, 2022Updated 3 years ago
- Code for pre-training CharacterBERT models (as well as BERT models).β34Sep 6, 2021Updated 4 years ago
- Improving Language Understanding from Screenshots. Paper: https://arxiv.org/abs/2402.14073β32Jul 9, 2024Updated 2 years ago
- Datasets for compositional learningβ11Nov 28, 2018Updated 7 years ago