Analyze videos using LLMs, Computer Vision and Automatic Speech Recognition
☆1,536Apr 19, 2026Updated 3 months ago
Alternatives and similar repositories for video-analyzer
Users that are interested in video-analyzer are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 利用 AI 大模型,一键解说并剪辑视频☆10,700Jul 23, 2026Updated 3 weeks ago
- FunASR-powered video transcription, subtitle generation, and LLM-assisted clipping tool with a local Gradio UI.☆6,138Aug 3, 2026Updated last week
- Taming Stable Diffusion for Lip Sync!☆5,990Jun 20, 2025Updated last year
- Netflix-level subtitle cutting, translation, alignment, and even dubbing - one-click fully automated AI video subtitle team | Netflix级字幕切…☆18,143Jul 2, 2026Updated last month
- Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenA…☆19,820Updated this week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.☆22,739May 25, 2026Updated 2 months ago
- Official implementation of "Sonic: Shifting Focus to Global Audio Perception in Portrait Animation"☆3,269Jan 8, 2026Updated 7 months ago
- Clapper.app, a video synthesizer and sequencer designed for the age of AI cinema☆2,330Aug 1, 2025Updated last year
- Open-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio…☆9,065Updated this week
- Valley is a cutting-edge multimodal large model designed to handle a variety of tasks involving text, images, video, and audio data.☆295May 8, 2026Updated 3 months ago
- Real time interactive streaming digital human☆8,749Updated this week
- 一键将视频转换为优质小红书笔记,自动优化内容和配图☆1,800Oct 30, 2025Updated 9 months ago
- 利用 AI 大模型和自动化工作流,根据主题或关键词一键生成高清短视频。Generate HD short videos from a topic or keyword with an automated AI workflow.☆103,011Updated this week
- SOTA Open Source TTS☆32,162Aug 3, 2026Updated last week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 快速提取音视频内容,整理成一份结构化的markdown笔记☆2,257Updated this week
- AI video agents framework for next-gen video interactions and workflows.☆1,511Jan 23, 2026Updated 6 months ago
- 🌐 Make websites accessible for AI agents. Automate tasks online with ease.☆109,098Updated this week
- AutoClip: AI-powered video clipping and highlight generation · 一款智能高光提取与剪辑的二创工具☆980Sep 15, 2025Updated 10 months ago
- 🚀 Truly open-source AI avatar(digital human) toolkit for offline video generation and digital human cloning.☆14,440Apr 21, 2026Updated 3 months ago
- 🚀🤖 Crawl4AI: Open-source LLM Friendly Web Crawler & Scraper. Don't be shy, join here: https://discord.gg/jP8KfhDhyN☆78,029Updated this week
- Build Agentic workflows, RAG pipelines, with rich AI model and tool support on one collaborative workspace. Deploy on cloud, VPC, or self…☆152,344Updated this week
- Faster Whisper transcription with CTranslate2☆24,874Nov 19, 2025Updated 8 months ago
- [ICLR2025] DisPose: Disentangling Pose Guidance for Controllable Human Image Animation☆377Nov 20, 2025Updated 8 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- 🔥🔥First-ever hour scale video understanding models☆626Jul 14, 2025Updated last year
- No fortress, purely open ground. OpenManus is Coming.☆57,956Feb 11, 2026Updated 6 months ago
- AIGCPanel 是一个简单易用的一站式AI数字人系统,支持视频合成、声音合成、声音克隆,简化本地模型管理、一键导入和使用AI模型。☆5,443Jul 16, 2026Updated 3 weeks ago
- Translate the video from one language to another and embed dubbing & subtitles.☆18,673Updated this week
- Open-source alternative to Opus Clip, Vidyo.ai, Klap & SubMagic. Turn long-form YouTube videos into viral 9:16 shorts using LLM highlight…☆4,584Jul 29, 2026Updated 2 weeks ago
- 小红书笔记 | 评论爬虫、抖音视频 | 评论爬虫、快手视频 | 评论爬虫、B 站视频 | 评论爬虫、微博帖子 | 评论爬虫、百度贴吧帖子 | 百度贴吧评论回复爬虫 | 知乎问答文章|评论爬虫☆62,308Updated this week
- A simple screen parsing tool towards pure vision based GUI agent☆25,252Jul 20, 2026Updated 3 weeks ago
- 🔍 An LLM-based Multi-agent Framework of Web Search Engine (like Perplexity.ai Pro and SearchGPT)☆6,912Jul 4, 2025Updated last year
- A generative speech model for daily dialogue.☆39,775Apr 10, 2026Updated 4 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface.☆127,356Updated this week
- Frontier Multimodal Foundation Models for Image and Video Understanding☆1,174Aug 14, 2025Updated 11 months ago
- [KDD'2026] "VideoRAG: Chat with Your Videos"☆3,280Mar 18, 2026Updated 4 months ago
- MuseTalk: Real-Time High Quality Lip Synchorization with Latent Space Inpainting☆6,377Sep 26, 2025Updated 10 months ago
- A Pocket-Sized MLLM for Ultra-Efficient Image and Video Understanding on Your Phone☆26,161Updated this week
- 实时交互数字人,可自定义形象与音色,支持音色克隆,对话延迟低至3s。Real-time voice interactive digital human, customizable appearance and voice, supporting voice cloning,…☆1,300Dec 18, 2025Updated 7 months ago
- An AI-Powered Speech Processing Toolkit and Open Source SOTA Pretrained Models, Supporting Speech Enhancement, Separation, and Target Spe…☆4,404Aug 14, 2025Updated 11 months ago