This project provides a production-ready, real-time inference server for LatentSync, enabling high-quality, low-latency 2D digital human live streaming. It features a robust multi-process architecture, seamless idle/formal stream switching with smooth crossfade transitions, and is optimized for real-world deployment scenarios.
☆30Aug 16, 2025Updated last year
Alternatives and similar repositories for LatentsyncRealtime
Users that are interested in LatentsyncRealtime are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- 使用OpenCV部署图像描述Image_Captioning,包含C++和Python两个版本的程序☆12Dec 22, 2023Updated 2 years ago
- ☆19Mar 11, 2026Updated 5 months ago
- 使用ONNXRuntime部署Detic检测2万1千种类别的物体,包含C++和Python两个版本的程序☆17Aug 29, 2023Updated 3 years ago
- The official implement of Freeze-Omni.☆16Jul 10, 2025Updated last year
- 使用ONNXRuntime部署DeDoDe:"局部特征匹配:检测,不要描述——描述,不要检测"。依然是C++和Python两个版本的程序☆23Dec 22, 2023Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆15Apr 28, 2023Updated 3 years ago
- 使用onnxruntime部署实时视频帧插值,包含C++和Python两个版本的程序☆29Feb 14, 2024Updated 2 years ago
- phonetic similarity algorithms☆13Jun 19, 2018Updated 8 years ago
- fd-sds☆21Apr 8, 2026Updated 5 months ago
- 这是一款视频分析处理工具,目前嵌入了Visual Tracking功能,手动勾选视频中第一帧的某个物体,程序自动跟踪该物体在整个视频序列中的位置☆20Mar 30, 2017Updated 9 years ago
- This is application for dysarthria to improve their pronunciation by using deep learning☆10Dec 29, 2020Updated 5 years ago
- ☆10Dec 14, 2020Updated 5 years ago
- Perform the forced decoding with target transcription☆11Sep 12, 2018Updated 7 years ago
- Speech recognition module for Python, supporting several engines and APIs, online and offline.☆13Mar 9, 2022Updated 4 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Analytical methods for efficient inference of integrate-and-fire circuit models from single-trial spike trains☆11Oct 30, 2019Updated 6 years ago
- This demo showcases different approaches to handling the delay during RAG (Retrieval-Augmented Generation) lookups in a voice-enabled AI …☆20Jan 23, 2025Updated last year
- 洛曦 数字人视频播放器,带HTTP API,使用gradio api对接Easy-Wav2Lip、Sadtalker、GeneFacePlusPlus、MuseTalk,也可以用于播放本地视频☆173Oct 20, 2024Updated last year
- TEN VAD low-latency voice activity detection for real-time streaming, integrated with livekit-agents☆26Nov 13, 2025Updated 9 months ago
- This repository features projects that track physical activities using computer vision, integrating OpenCV with MediaPipe's Pose Estimati…☆15Oct 21, 2024Updated last year
- A simple script to prepare dataset for training with TTS Tortoise model via https://git.ecker.tech/mrq/ai-voice-cloning☆12Jan 12, 2024Updated 2 years ago
- This project provides a Flask-based API for generating high-quality text-to-speech (TTS) audio using F5-TTS, a flexible and powerful TTS …☆16Aug 14, 2026Updated 3 weeks ago
- ☆13Oct 27, 2021Updated 4 years ago
- ☆13Apr 9, 2021Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Using Kaldi (Automatic Speech Recognition) and Gentle (Forced Word Aligner), this script finds both rhymes and alliteration in speeches w…☆13May 4, 2018Updated 8 years ago
- PersonaTalk Hack☆15Jan 10, 2025Updated last year
- Speech Processing & Linguistic Analysis Tool☆11Jun 30, 2019Updated 7 years ago
- A powerful AI chat application that enables human-like conversation with fully animated AI characters.☆13Dec 7, 2024Updated last year
- 一款智能合同审查工具——an intelligent contract review tool that extracts text and structure from PDFs/images to identify risky clauses.☆28Mar 9, 2026Updated 5 months ago
- 基于webrtc修改kurento源码实现one2many实时直播☆11Aug 9, 2026Updated 3 weeks ago
- SPIE Medical Imaging 2019 Notes By Hao☆16Feb 26, 2019Updated 7 years ago
- An open source chat bot architecture for voice/vision (and multimodal) assistants, local(CPU/GPU bound) and remote(I/O bound) to run.☆89Dec 28, 2025Updated 8 months ago
- [ICASSP'25] DEGSTalk: Decomposed Per-Embedding Gaussian Fields for Hair-Preserving Talking Face Synthesis☆55Oct 25, 2025Updated 10 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- This project is based on an improved Wav2Lip model, achieving synchronization between audio and video lip movements to enhance video prod…☆17Jul 10, 2024Updated 2 years ago
- VoxLingua107 recipe for SpeechBrain☆13Jul 3, 2021Updated 5 years ago
- Dockerfiles for building an image including Stable Diffusion with Automatic1111 UI and kohya_ss running with the required ROCm software f…☆15Apr 19, 2024Updated 2 years ago
- This project used Yolov8/AnimeGAN and Flask to accomplish the task of background segmentation , background remove and background replacem…☆12Apr 12, 2024Updated 2 years ago
- The source code of IEEE TPAMI 2025 "Hyper-YOLO: When Visual Object Detection Meets Hypergraph Computation".☆120Dec 16, 2024Updated last year
- Paper summary of 2D video generation. Updated 2021.06☆16Apr 30, 2021Updated 5 years ago
- Fraud detection using lgb, catboost, rf, etc☆12May 4, 2020Updated 6 years ago