Research code for pixel-based encoders of language (PIXEL)
β348Jul 15, 2025Updated last year
Alternatives and similar repositories for pixel
Users that are interested in pixel are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This repository contains an extension of fairseq for pixel / visual representations of text for machine translation.β37Feb 2, 2024Updated 2 years ago
- π VITRina: VIsual Token Representationsβ11Jun 15, 2023Updated 3 years ago
- β85Dec 4, 2022Updated 3 years ago
- Textual Visual Semantic Dataset for Text Spotting. CVPRW 2020β12Jul 2, 2022Updated 4 years ago
- Code release for SLIP Self-supervision meets Language-Image Pre-trainingβ791Feb 9, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [EMNLP'23] Official Code for "FOCUS: Effective Embedding Initialization for Monolingual Specialization of Multilingual Models"β37Jun 7, 2025Updated last year
- Official pytorch implementation of I2I translation with low resolution conditioningβ23Sep 2, 2021Updated 4 years ago
- Large-scale Self-supervised Pre-training Across Tasks, Languages, and Modalitiesβ81Jan 7, 2026Updated 7 months ago
- Omnivore: A Single Model for Many Visual Modalitiesβ573Nov 12, 2022Updated 3 years ago
- Official PyTorch implementation of the paper "In-Context Learning Unlocked for Diffusion Models"β414Mar 25, 2024Updated 2 years ago
- Language Models Can See: Plugging Visual Controls in Text Generationβ260Jun 1, 2022Updated 4 years ago
- PyTorch code for "Perceiver-VL: Efficient Vision-and-Language Modeling with Iterative Latent Attention" (WACV 2023)β34Feb 5, 2023Updated 3 years ago
- Official implementation and data release of the paper "Visual Prompting via Image Inpainting".β320Aug 7, 2023Updated 3 years ago
- A Multilingual Keyboard Layout-Based Typo Generatorβ17Nov 23, 2025Updated 8 months ago
- Managed Database hosting by DigitalOcean β’ AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- A Structured Span Selector (NAACL 2022). A structured span selector with a WCFG for span selection tasks (coreference resolution, semantiβ¦β21Jul 11, 2022Updated 4 years ago
- Pix2Seq codebase: multi-tasks with generative modeling (autoregressive and diffusion)β945Nov 7, 2023Updated 2 years ago
- β88Jan 10, 2024Updated 2 years ago
- Code for WECHSEL: Effective initialization of subword embeddings for cross-lingual transfer of monolingual language models.β93Sep 12, 2024Updated last year
- Python source code for EMNLP 2021 Findings paper: "Subword Mapping and Anchoring Across Languages".β13Sep 17, 2021Updated 4 years ago
- β43Aug 9, 2022Updated 4 years ago
- Masked Siamese Networks for Label-Efficient Learning (https://arxiv.org/abs/2204.07141)β463May 9, 2022Updated 4 years ago
- Patching open-vocabulary models by interpolating weightsβ91Sep 28, 2023Updated 2 years ago
- Code for paper βLanguage Versatilists vs. Specialists: An Empirical Revisiting on Multilingual Transfer Abilityββ15Jun 13, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Implementation of N-Grammer, augmenting Transformers with latent n-grams, in Pytorchβ82Dec 4, 2022Updated 3 years ago
- β45Jul 5, 2022Updated 4 years ago
- [NeurIPS2023] Official implementation and model release of the paper "What Makes Good Examples for Visual In-Context Learning?"β182Mar 4, 2024Updated 2 years ago
- A Unified Library for Parameter-Efficient and Modular Transfer Learningβ2,826Apr 26, 2026Updated 3 months ago
- β15Jan 14, 2026Updated 6 months ago
- PyTorch code for "Unifying Vision-and-Language Tasks via Text Generation" (ICML 2021)β372Jul 29, 2023Updated 3 years ago
- Official codebase used to develop Vision Transformer, SigLIP, MLP-Mixer, LiT and more.β3,513May 19, 2025Updated last year
- β12Mar 12, 2023Updated 3 years ago
- Paper List for In-context Learning π·β19Jan 3, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Exploring Visual Prompts for Adapting Large-Scale Modelsβ292Jun 6, 2022Updated 4 years ago
- UDapter is a multilingual dependency parser that uses "contextual" adapters together with language-typology features for language-specifiβ¦β31Dec 5, 2022Updated 3 years ago
- Supervision Exists Everywhere: A Data Efficient Contrastive Language-Image Pre-training Paradigmβ677Sep 19, 2022Updated 3 years ago
- Official repository of OFA (ICML 2022). Paper: OFA: Unifying Architectures, Tasks, and Modalities Through a Simple Sequence-to-Sequence Lβ¦β2,557Apr 24, 2024Updated 2 years ago
- PyTorch code for "VL-Adapter: Parameter-Efficient Transfer Learning for Vision-and-Language Tasks" (CVPR2022)β212Dec 18, 2022Updated 3 years ago
- Code for pre-training CharacterBERT models (as well as BERT models).β34Sep 6, 2021Updated 4 years ago
- Improving Language Understanding from Screenshots. Paper: https://arxiv.org/abs/2402.14073β32Jul 9, 2024Updated 2 years ago