code for reproducing some of the diagrams in the paper "Multimodal Neurons in Artificial Neural Networks"
☆310Mar 21, 2021Updated 5 years ago
Alternatives and similar repositories for CLIP-featurevis
Users that are interested in CLIP-featurevis are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Code associated with our paper "Learning Group Structure and Disentangled Representations of Dynamical Environments"☆15Dec 8, 2022Updated 3 years ago
- A GPT, made only of MLPs, in Jax☆59Jun 23, 2021Updated 5 years ago
- ☆2,096Apr 29, 2022Updated 4 years ago
- Chef cookbooks for managing a Ceph cluster☆11Apr 2, 2023Updated 3 years ago
- Implementation / replication of DALL-E, OpenAI's Text to Image Transformer, in Pytorch☆5,626Feb 17, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [CVPR 2021] VirTex: Learning Visual Representations from Textual Annotations☆563Aug 22, 2025Updated last year
- SentAugment is a data augmentation technique for NLP that retrieves similar sentences from a large bank of sentences. It can be used in c…☆358Feb 22, 2022Updated 4 years ago
- [CVPR 2021 Best Student Paper Honorable Mention, Oral] Official PyTorch code for ClipBERT, an efficient framework for end-to-end learning…☆728Aug 8, 2023Updated 3 years ago
- Retryable HTTP client in Go☆13Apr 2, 2023Updated 3 years ago
- Code for EMNLP2021 paper "Allocating Large Vocabulary Capacity for Cross-lingual Language Model Pre-training"☆20Nov 12, 2021Updated 4 years ago
- Highly specialized crate to parse and use `google/sentencepiece` 's precompiled_charsmap in `tokenizers`☆23Jun 9, 2026Updated 2 months ago
- VISSL is FAIR's library of extensible, modular and scalable components for SOTA Self-Supervised Learning with images.☆3,293Mar 3, 2024Updated 2 years ago
- Code for the Shortformer model, from the ACL 2021 paper by Ofir Press, Noah A. Smith and Mike Lewis.☆147Jul 26, 2021Updated 5 years ago
- Fluentd output plugin that sends events to Amazon Kinesis Streams and Amazon Kinesis Firehose.☆12Apr 2, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Flexible Feature visualization on PyTorch, for research and art☆247Mar 17, 2025Updated last year
- Repository for "Generating images from caption and vice versa via CLIP-Guided Generative Latent Space Search"☆179Sep 30, 2021Updated 4 years ago
- Oscar and VinVL☆1,053Aug 28, 2023Updated 3 years ago
- Implementation of Token Shift GPT - An autoregressive model that solely relies on shifting the sequence space for mixing☆49Jan 27, 2022Updated 4 years ago
- VQGAN+CLIP with some additional tuning. For notebooks and the command line.☆50Aug 20, 2021Updated 5 years ago
- Lifelong Variational Autoencoder☆15Dec 6, 2017Updated 8 years ago
- ☆21Apr 16, 2022Updated 4 years ago
- Code for ACL 2023 Oral Paper: ManagerTower: Aggregating the Insights of Uni-Modal Experts for Vision-Language Representation Learning☆12Aug 23, 2025Updated last year
- WIT (Wikipedia-based Image Text) Dataset is a large multimodal multilingual dataset comprising 37M+ image-text sets with 11M+ unique imag…☆1,113Sep 27, 2024Updated last year
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Official codebase for Pretrained Transformers as Universal Computation Engines.☆245Jan 14, 2022Updated 4 years ago
- Repository for out-of-tree scheduler plugins based on scheduler framework.☆13Apr 2, 2023Updated 3 years ago
- CLIP (Contrastive Language-Image Pretraining), Predict the most relevant text snippet given an image☆34,272Mar 25, 2026Updated 5 months ago
- Code release for SLIP Self-supervision meets Language-Image Pre-training☆788Feb 9, 2023Updated 3 years ago
- PyTorch code for EMNLP 2020 Paper "Vokenization: Improving Language Understanding with Visual Supervision"☆190Mar 8, 2021Updated 5 years ago
- Open-AI's DALL-E for large scale training in mesh-tensorflow.☆432Feb 12, 2022Updated 4 years ago
- ☆41Apr 2, 2023Updated 3 years ago
- [Findings of NAACL2022] A Dog Is Passing Over The Jet? A Text-Generation Dataset for Korean Commonsense Reasoning and Evaluation☆11May 27, 2022Updated 4 years ago
- ↔️ Utilizing RBERT model structure for KLUE Relation Extraction task☆15Nov 15, 2022Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- I have created a dataset of Image-Text-Pairs by using the cosine similarity of the CLIP embeddings of the image & it's caption derrived f…☆18Apr 22, 2021Updated 5 years ago
- [ICLR 2022] code for "How Much Can CLIP Benefit Vision-and-Language Tasks?" https://arxiv.org/abs/2107.06383☆420Oct 28, 2022Updated 3 years ago
- This repository contains the code and datasets for our ICCV-W paper 'Enhancing CLIP with GPT-4: Harnessing Visual Descriptions as Prompts…☆30Feb 21, 2024Updated 2 years ago
- Examples of using sparse attention, as in "Generating Long Sequences with Sparse Transformers"☆1,613Aug 12, 2020Updated 6 years ago
- Implementations of GANs in Tensorflow 2.x☆16Feb 12, 2022Updated 4 years ago
- Generate a denotation graph from a set of image captions☆16Sep 4, 2018Updated 8 years ago
- Conceptual 12M is a dataset containing (image-URL, caption) pairs collected for vision-and-language pre-training.☆426Jul 14, 2025Updated last year