Digital Avatar Conversational System - Linly-Talker. 😄✨ Linly-Talker is an intelligent AI system that combines large language models (LLMs) with visual models to create a novel human-AI interaction method. 🤝🤖 It integrates various technologies like Whisper, Linly, Microsoft Speech Services, and SadTalker talking head generation system. 🌟🔬
☆3,441Feb 10, 2026Updated 6 months ago
Alternatives and similar repositories for Linly-Talker
Users that are interested in Linly-Talker are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Real time interactive streaming digital human☆9,285Updated this week
- [ICCV'23] Efficient Region-Aware Neural Radiance Fields for High-Fidelity Talking Portrait Synthesis☆1,260Mar 14, 2025Updated last year
- MuseTalk: Real-Time High Quality Lip Synchorization with Latent Space Inpainting☆6,474Sep 26, 2025Updated 11 months ago
- 每个人都能用的数字人☆2,133Aug 16, 2026Updated 2 weeks ago
- Real time streaming talking head☆478May 17, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [CVPR 2024] This is the official source for our paper "SyncTalk: The Devil is in the Synchronization for Talking Head Synthesis"☆1,626Sep 18, 2025Updated 11 months ago
- [AAAI 2025] EchoMimic: Lifelike Audio-Driven Portrait Animations through Editable Landmark Conditioning☆4,294Apr 7, 2026Updated 4 months ago
- MuseV: Infinite-length and High Fidelity Virtual Human Video Generation with Visual Conditioned Parallel Denoising☆2,848Jun 28, 2024Updated 2 years ago
- 一个超轻量级、可以在移动端实时运行的数字人模型☆2,631Jul 22, 2026Updated last month
- GeneFace++: Generalized and Stable Real-Time 3D Talking Face Generation; Official Code☆1,809Oct 18, 2024Updated last year
- 实时交互数字人,可自定义形象与音色,支持音色克隆,对话延迟低至3s。Real-time voice interactive digital human, customizable appearance and voice, supporting voice cloning,…☆1,303Dec 18, 2025Updated 8 months ago
- fay是一个帮助数字人(2.5d、3d、移动、pc、网页)或大语言模型(openai兼容、deepseek)连通业务系统的agent框架。☆13,466Aug 7, 2026Updated 3 weeks ago
- [SIGGRAPH Asia 2022] VideoReTalking: Audio-based Lip Synchronization for Talking Head Video Editing In the Wild☆7,282Aug 5, 2024Updated 2 years ago
- Taming Stable Diffusion for Lip Sync!☆6,040Jun 20, 2025Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [CVPR 2025] EchoMimicV2: Towards Striking, Simplified, and Semi-Body Human Animation☆4,650Feb 23, 2026Updated 6 months ago
- [CVPR 2023] SadTalker:Learning Realistic 3D Motion Coefficients for Stylized Audio-Driven Single Image Talking Face Animation☆14,052Jun 26, 2024Updated 2 years ago
- AI Vtuber是一个由 【ChatterBot/ChatGPT/claude/langchain/chatglm/text-gen-webui/闻达/千问/kimi/ollama】 驱动的虚拟主播【Live2D/UE/xuniren】,可以在 【Bilibili/抖音/…☆4,439Jul 29, 2025Updated last year
- Real3D-Portrait: One-shot Realistic 3D Talking Portrait Synthesis; ICLR 2024 Spotlight; Official code☆1,091Oct 18, 2024Updated last year
- 本项目基于SadTalkers实现视频唇形合成的Wav2lip。通过以视频文件方式进行语音驱动生成唇形,设置面部区域可配置的增强方式进行合成唇形(人脸)区域画面增强,提高生成唇形的清晰度。使用DAIN 插帧的DL算法对生成视频进行补帧,补充帧间合成唇形的动作过渡,使合成的唇…☆2,009Jun 4, 2023Updated 3 years ago
- ☆3,725Jul 31, 2026Updated last month
- AniPortrait: Audio-Driven Synthesis of Photorealistic Portrait Animation☆5,020Jul 2, 2024Updated 2 years ago
- JoyHallo: Digital human model for Mandarin☆521Sep 21, 2025Updated 11 months ago
- Streamer-Sales 销冠 —— 卖货主播 LLM 大模型🛒🎁,一个能够根据给定的商品特点从激发用户购买意愿角度出发进行商品解说的卖货主播大模型。🚀⭐内含详细的数据生成流程❗ 📦另外还集成了 LMDeploy 加速推理🚀、RAG检索增强生成 📚、TTS文…☆3,763Mar 8, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Colab for making Wav2Lip high quality and easy to use☆855May 17, 2024Updated 2 years ago
- Wav2Lip version 288 and pipeline to train☆649Aug 13, 2025Updated last year
- V-Express aims to generate a talking head video under the control of a reference image, an audio, and a sequence of V-Kps images.☆2,360Jan 24, 2025Updated last year
- This repository contains the codes of "A Lip Sync Expert Is All You Need for Speech to Lip Generation In the Wild", published at ACM Mult…☆13,184Jun 22, 2025Updated last year
- [ACM MM 2024] This is the official code for "AniTalker: Animate Vivid and Diverse Talking Faces through Identity-Decoupled Facial Motion …☆1,598Aug 15, 2024Updated 2 years ago
- CVPR2023 talking face implementation for Identity-Preserving Talking Face Generation With Landmark and Appearance Priors☆734Jan 6, 2024Updated 2 years ago
- The fastest digital human algorithm, now on your desktop.☆586Sep 29, 2025Updated 11 months ago
- Official implementations for paper: DreamTalk: When Expressive Talking Head Generation Meets Diffusion Probabilistic Models☆1,787Jan 15, 2024Updated 2 years ago
- GeneFace: Generalized and High-Fidelity 3D Talking Face Synthesis; ICLR 2023; Official code☆2,657Oct 18, 2024Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- MimicTalk: Mimicking a personalized and expressive 3D talking face in minutes; NeurIPS 2024; Official code☆830Oct 16, 2024Updated last year
- 🚀 The best real-time interactive AI avatar(digital human) with on-premise deployment and <1.5 s latency.☆8,203Aug 5, 2026Updated 3 weeks ago
- ☆245Dec 26, 2023Updated 2 years ago
- Multi-lingual large voice generation model, providing inference, training and deployment full-stack ability.☆23,325May 25, 2026Updated 3 months ago
- [ICLR 2025 Oral] TANGO: Co-Speech Gesture Video Reenactment with Hierarchical Audio-Motion Embedding and Diffusion Interpolation☆1,163Aug 24, 2025Updated last year
- Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenA…☆20,081Updated this week
- ☆472Jun 30, 2025Updated last year