Bling's Object detection tool
☆55Jan 9, 2023Updated 3 years ago
Alternatives and similar repositories for BriVL-BUA-applications
Users that are interested in BriVL-BUA-applications are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Bridging Vision and Language Model☆287Mar 27, 2023Updated 3 years ago
- The Document of WenLan API, which was used to obtain image and text feature.☆41Jan 10, 2023Updated 3 years ago
- Official implementation of the ICASSP-2022 paper "Text2Poster: Laying Out Stylized Texts on Retrieved Images"☆214Dec 18, 2023Updated 2 years ago
- CVPR 2021 Official Pytorch Code for UC2: Universal Cross-lingual Cross-modal Vision-and-Language Pre-training☆34Nov 9, 2021Updated 4 years ago
- ☆18Mar 20, 2022Updated 4 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Southeast University Knowledge Graph-OpenRichpedia☆41Aug 28, 2021Updated 5 years ago
- [ACL 2024] Benchmarking Knowledge Boundary for Large Language Models: A Different Perspective on Model Evaluation☆10May 26, 2024Updated 2 years ago
- TaiSu(太素)--a large-scale Chinese multimodal dataset(亿级大规模中文视觉语言预训练数据集)☆192Nov 17, 2023Updated 2 years ago
- 基于方差权重因子选词的SIF句向量模型-实验源码☆11Mar 8, 2020Updated 6 years ago
- Implementation of Relation Extraction with Multi-instance Multi-label Convolutional Neural Networks in tensorflow☆15Apr 2, 2017Updated 9 years ago
- I have created a dataset of Image-Text-Pairs by using the cosine similarity of the CLIP embeddings of the image & it's caption derrived f…☆18Apr 22, 2021Updated 5 years ago
- ☆15Jul 24, 2017Updated 9 years ago
- ☆17Oct 15, 2023Updated 2 years ago
- A subset of YFCC100M. Tools, checking scripts and links of web drive to download datasets(uncompressed).☆19Aug 5, 2026Updated last month
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- This repo contains codes and instructions for baselines in the VLUE benchmark.☆41Jul 16, 2022Updated 4 years ago
- PyTorch source code for "Stacked Cross Attention for Image-Text Matching" (ECCV 2018)☆580May 18, 2023Updated 3 years ago
- Small Flask-based apps to browse the Flickr30k dataset.☆20Mar 3, 2017Updated 9 years ago
- cpp write language detect model☆11Sep 22, 2021Updated 4 years ago
- Code for AAAI 2022 paper Unsupervised Sentence Representation via Contrastive Learning with Mixing Negatives☆23Jun 14, 2022Updated 4 years ago
- Code recipe for "Multimodal One-Shot Learning of Speech and Images"☆11Nov 21, 2018Updated 7 years ago
- Dataset and codes for the paper "Product-oriented Machine Translation with Cross-modal Cross-lingual Pre-training".☆25Mar 6, 2022Updated 4 years ago
- Evaluating Visual Fidelity of Image Descriptions☆11Aug 15, 2019Updated 7 years ago
- ☆11Apr 19, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 模式识别课设代码:图文生成(CLIP+DALLE+BriVL)☆20Jun 15, 2023Updated 3 years ago
- ViT models pretrained with up to ~5k hours of human-like video data☆14Aug 10, 2023Updated 3 years ago
- X-VLM: Multi-Grained Vision Language Pre-Training (ICML 2022)☆506Nov 25, 2022Updated 3 years ago
- ☆15Apr 30, 2022Updated 4 years ago
- Convert 3D Human Pose to VMD file☆14Apr 21, 2019Updated 7 years ago
- An open-source library for contamination detection in NLP datasets and Large Language Models (LLMs).☆61Aug 13, 2024Updated 2 years ago
- ☆12May 3, 2024Updated 2 years ago
- Enriching MS-COCO with Chinese sentences and tags for cross-lingual multimedia tasks☆215Feb 12, 2025Updated last year
- 專題論文☆10Jul 27, 2013Updated 13 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆83Jul 3, 2023Updated 3 years ago
- [ACM MM 2022] (Oral): Multi-Modal Experience Inspired AI Creation☆21Nov 27, 2024Updated last year
- ☆11Mar 13, 2023Updated 3 years ago
- Video-Language Alignment via Spatio–Temporal Graph Transformer; ArXiv: https://arxiv.org/abs/2407.11677☆15Jul 24, 2024Updated 2 years ago
- 使用springboot+minio+elasticsearch+webuploader实现图床,支持给图片打标签,使用elasticsearch搜索,支持图片压缩,支持分片上传,秒传,断点续传☆19Oct 3, 2024Updated last year
- OpenAI CLIP text encoders for multiple languages!☆832May 15, 2023Updated 3 years ago
- 500,000 multimodal short video data and baseline models. 50万条多模态短视频数据集和基线模型(TensorFlow2.0)。☆136Jul 23, 2019Updated 7 years ago