A project that can generate ancient poems based on pictures, including CLIP, T5, GPT2 models
☆21Feb 16, 2025Updated last year
Alternatives and similar repositories for Image2Poem
Users that are interested in Image2Poem are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official implementation for P2SAM (ACM MM 2024)☆14Dec 7, 2024Updated last year
- 【ICME2025 Oral】Offical Pytorch Code for "Fraesormer: Learning Adaptive Sparse Transformer for Efficient Food Recognition"☆13Mar 21, 2025Updated last year
- [EMNLP 2022] Language Model Pre-Training with Sparse Latent Typing☆14Feb 10, 2023Updated 3 years ago
- ☆15Jun 19, 2024Updated 2 years ago
- An implementation of online data mixing for the Pile dataset, based on the GPT-NeoX library.☆14Jan 9, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Using machine learning techniques for prediction and modelling non linear dynamic systems.☆11Jun 29, 2018Updated 8 years ago
- pre-training llama3 using chinese☆13May 1, 2024Updated 2 years ago
- Multiple Attractors simulation with customization☆14Feb 22, 2026Updated 6 months ago
- 毕业设计(医疗问答系统)☆28Mar 28, 2022Updated 4 years ago
- [ACM'MM 2025] UAV Street-Satellite matching workshop Challenging paper, SkyLink: Unifying Street-Satellite Geo-Localization via UAV-Media…☆28Dec 9, 2025Updated 8 months ago
- ☆16Sep 4, 2025Updated last year
- A matlab package for analyzing chaotic properties of time series data☆11Jun 29, 2018Updated 8 years ago
- LongAttn :Selecting Long-context Training Data via Token-level Attention☆15Jul 16, 2025Updated last year
- A ComfyUI extension for StyleShot.☆16Apr 23, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆16Aug 28, 2024Updated 2 years ago
- Video-Language Alignment via Spatio–Temporal Graph Transformer; ArXiv: https://arxiv.org/abs/2407.11677☆15Jul 24, 2024Updated 2 years ago
- ☆18May 17, 2022Updated 4 years ago
- 数据库课程设计,Python+SQLServer实现疫情医疗信息管理系统☆28Nov 28, 2024Updated last year
- ☆18Sep 29, 2022Updated 3 years ago
- WebRED is a large and diverse manually annotated dataset for extracting relationships from a variety of text found on the World Wide Web.☆22Mar 11, 2021Updated 5 years ago
- 中文文本合成 for OCR☆12Mar 14, 2023Updated 3 years ago
- 前沿论文持续更新--视频时刻定位 or 时域语言定位 or 视频片段检索。☆14Nov 8, 2023Updated 2 years ago
- [CVPR 2025 GMCV] Test-Time Frequency Scaling: Instant Frequency Control for Any Diffusion Model☆56May 31, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [ACL 2026] WildGraphBench: Benchmarking GraphRAG with Wild-Source Corpora☆17May 11, 2026Updated 3 months ago
- Code for CVPR 2023 paper "SViTT: Temporal Learning of Sparse Video-Text Transformers"☆21Jun 16, 2023Updated 3 years ago
- vue集成海康官方控件webVideoCtrl.js,运行在IE11环境下。☆16Mar 26, 2021Updated 5 years ago
- A roadmap of artificial intelligence☆17Sep 17, 2022Updated 3 years ago
- 通过阿里云盘,colab,国内下载huggingface大模型轻轻松松☆42Apr 15, 2026Updated 4 months ago
- GFPGAN face reconstruction with ncnn on a bare Raspberry Pi☆14Jan 4, 2023Updated 3 years ago
- [NAACL 2024] LaDiC: Are Diffusion Models Really Inferior to Autoregressive Counterparts for Image-to-text Generation?☆42Jun 9, 2024Updated 2 years ago
- Unofficial implementation of 'Large-Scale Text-to-Image Model with Inpainting is a Zero-Shot Subject-Driven Image Generator'☆10Dec 10, 2024Updated last year
- In-sensor reservoir computing for language learning via two-dimensional memristors☆26Feb 16, 2023Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Text Detection by RetinaNet with PyTorch (Code will be released soon)☆10Dec 1, 2018Updated 7 years ago
- Code for ACL 2020 paper "Rigid Formats Controlled Text Generation":https://www.aclweb.org/anthology/2020.acl-main.68/☆236May 25, 2021Updated 5 years ago
- Official PyTorch implementation of ResFormer: Scaling ViTs with Multi-Resolution Training, CVPR2023☆30Jun 22, 2023Updated 3 years ago
- ☆22May 30, 2023Updated 3 years ago
- Create After Effects scripts in Python.☆12Jan 29, 2021Updated 5 years ago
- 如何在github上传本地项目代码(新手使用)☆23Sep 30, 2019Updated 6 years ago
- Official implementation for Dynamically Instance-Guided Adaptation: A Backward-free Approach for Test-Time Domain Adaptive Semantic Segme…☆13Mar 19, 2024Updated 2 years ago