Text detection and recognition in natural videos
☆28Feb 7, 2018Updated 8 years ago
Alternatives and similar repositories for VideoText
Users that are interested in VideoText are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A Caffe Fully Convolutional Network that detects any kind of text and generates pixel-level heatmaps.☆23Nov 24, 2017Updated 8 years ago
- A modified SSD model for text detection☆91Jan 23, 2017Updated 9 years ago
- ☆12Apr 25, 2017Updated 9 years ago
- Character-level conversion between Hebrew text and Latin transliteration using deep learning - a demonstration of seq2seq training.☆16Jun 27, 2023Updated 3 years ago
- Large-scale city camera video dataset☆11Jul 20, 2020Updated 6 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Trial for using YOLO v3 with aiortc (WebRTC implementation with Python)☆17Jul 15, 2018Updated 8 years ago
- Open tools and data for cloudless automatic speech recognition☆13Oct 1, 2019Updated 6 years ago
- python image data augmentation☆12Jul 24, 2017Updated 9 years ago
- VGG16 architecture with BatchNorm☆14Apr 4, 2017Updated 9 years ago
- ☆14Jun 16, 2023Updated 3 years ago
- Generating video descriptions using deep learning in Keras☆25Sep 8, 2020Updated 5 years ago
- Tool for creating scenarios with py-faster-rcnn☆12Nov 2, 2017Updated 8 years ago
- This is the official repo for "S3D: Single Shot multi-Span Detector via Fully 3D Convolutional Network"☆15Jan 23, 2019Updated 7 years ago
- Master in Computer Vision - M5 Visual recognition☆13May 5, 2017Updated 9 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- this is the code release for ''Weakly Supervised PatchNets: Describing and Aggregating Local Patches for Scene Recognition''☆43Feb 27, 2018Updated 8 years ago
- Deep colorizer based on unet architecture (Encoder decoder with skip connections)☆14Jan 7, 2017Updated 9 years ago
- ☆22Jun 9, 2015Updated 11 years ago
- This is my attempt at the ActivityNet Challenge 2017. Thanks to the organizers for providing the boilerplate code and annotated datasets.…☆10Jul 19, 2017Updated 9 years ago
- C++ version of pyannote audio overlapped speech detection pipeline☆13Feb 14, 2024Updated 2 years ago
- マウスクリックで指定した座標を矩形に射影変換するプログラム。☆10Jul 9, 2020Updated 6 years ago
- Extraction and removal of a video background using k-means.☆19Feb 15, 2020Updated 6 years ago
- ☆17Aug 25, 2017Updated 9 years ago
- Export Mysql/Oracle Data to Excel CSV☆14Dec 13, 2013Updated 12 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- we explores the fascinating domain of text-to-image generation using the powerful capabilities of the Flux API. The objective is to trans…☆12Aug 14, 2024Updated 2 years ago
- Visual Hash for matching copies of visually similar images.☆16Mar 17, 2025Updated last year
- ☆10Sep 26, 2019Updated 6 years ago
- ☆21Mar 4, 2024Updated 2 years ago
- Visual Attention based OCR☆1,117Nov 8, 2018Updated 7 years ago
- Docker Image for Darknet☆10Dec 17, 2016Updated 9 years ago
- videojs plugin, display marker point on progress bar.☆11Oct 25, 2023Updated 2 years ago
- Data Dialogue enables natural language querying of databases by integrating LLMs with SQL databases.☆15May 3, 2025Updated last year
- Jump through Go stacktraces as easily as grep-mode☆13Apr 30, 2015Updated 11 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- 本项目数据库设计、技术选型、前后端代码编写等,全部由本人完成。本人特别爱刷抖音,有一天突发奇想我能不能自己做个抖音,恰好我学习的软件开发技能还没什么用武之地,于是这个项目便诞生了。☆10Apr 7, 2026Updated 4 months ago
- Speaker-aware CTC (SACTC) for multi-talker overlapped speech recognition.☆22May 26, 2025Updated last year
- This tool can convert picture format(NV12/YUYV/UYVY...) to (png/jpg/bmp)☆10Jul 14, 2018Updated 8 years ago
- We ❤️ smolagents - Building Agents Approaches ... to solve use-cases☆13Jan 4, 2025Updated last year
- pytorch实现的Pyramidbox 人脸检测模型, 对原来代码的部分模块进行了修改,更简洁高效☆22Dec 8, 2020Updated 5 years ago
- open-deepsearch☆12Mar 3, 2025Updated last year
- 以500px为图源的Android壁纸应用☆11Mar 12, 2023Updated 3 years ago