Google Gemini live voice to text realtime stream in the browser
☆16Dec 13, 2025Updated 7 months ago
Alternatives and similar repositories for gemini-live
Users that are interested in gemini-live are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆16Mar 19, 2021Updated 5 years ago
- Speech recognition in JavaScript☆18Oct 27, 2018Updated 7 years ago
- PocketSphinx phonetic feature extraction for intelligibility prediction and remediation☆29May 28, 2020Updated 6 years ago
- Speech recognizer in JavaScript, base library☆19Jul 24, 2014Updated 12 years ago
- Discussion and project code for the Deep Learning class☆12Mar 9, 2023Updated 3 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- self-hosted workspace for Google's Jules AI agent☆17Jul 4, 2026Updated last month
- This repo will not not be updated anymore, please see https://github.com/wikiwho/WikiWho from now on.☆16Mar 31, 2017Updated 9 years ago
- simple streamlit chat app to talk with MistralAI's API.☆15Dec 15, 2023Updated 2 years ago
- SuggestBot is an article recommender for Wikipedia☆21Dec 29, 2024Updated last year
- Creating Container Images in AWS Fargate with Kaniko☆17Jan 15, 2024Updated 2 years ago
- ⚙️ Powerful JS library to manage audio recording : intelligent cutting, saturation control, various export options...☆44Jan 16, 2026Updated 6 months ago
- scikit-learn: machine learning in Python☆12Oct 17, 2016Updated 9 years ago
- A plain HTML/JS demo that uses vmsg (a modern WebAssembly version of LAME) to record and encode mp3 audio in the browser☆23Apr 25, 2026Updated 3 months ago
- Code for the Paper: [ECCV2022] Sound Localization by Self-Supervised Time-Delay Estimation☆29Mar 15, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Extract frequency, power, width and dissonance of formants from wav files☆28Jun 3, 2022Updated 4 years ago
- Run SQL queries and send the results to Geckoboard Datasets☆26Jul 7, 2023Updated 3 years ago
- Probabilistic modelling with tensor networks☆29Nov 29, 2019Updated 6 years ago
- A Python package to facilitate research on building and evaluating automated scoring models.☆71Dec 27, 2024Updated last year
- A simple but complete web application skeleton that runs 'out of the box'.☆38Jul 4, 2017Updated 9 years ago
- Material for SF ACM Chapter talk on 10/28☆27Oct 21, 2020Updated 5 years ago
- 夏目悠李/男声歌声データベースの最新ラベルデータ☆12Sep 2, 2020Updated 5 years ago
- Speech Recognition implementation using Artificial Neural Networks☆10Sep 7, 2015Updated 10 years ago
- Visualization for hidden Markov model computations☆14Dec 19, 2014Updated 11 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Expected edit distance implementation using OpenFst tools☆11May 13, 2015Updated 11 years ago
- CQT-DTW Score Follower☆17Sep 24, 2022Updated 3 years ago
- Unsupervised speech activity detection system.☆11Jul 2, 2018Updated 8 years ago
- Lie Detection by voice and heart rate☆10Dec 20, 2017Updated 8 years ago
- Automatically exported from code.google.com/p/transducersaurus☆11Apr 1, 2015Updated 11 years ago
- This is an extension of kaldi speech recognition software which allows to perform decoding of speech with hybrid word and phoneme graphs.…☆11Feb 4, 2020Updated 6 years ago
- Multiobjective Optimization Training of PLDA for Speaker Verification☆10Jun 14, 2018Updated 8 years ago
- Open source image editor for windows 10. Can be controlled by voice commands and Cortana.☆17Nov 14, 2017Updated 8 years ago
- EditEvo is a browser-based video editor designed to provide users with editing capabilities directly within their web browser. The appli…☆11May 7, 2024Updated 2 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Text-Dependent Speaker Recognition System with Machine Learning Techniques☆10Dec 31, 2017Updated 8 years ago
- JavaScript libraries to interact with the Ispikit pronunciation assessment server☆11Nov 16, 2016Updated 9 years ago
- A simple pyaudio microphone interface☆11Jul 27, 2018Updated 8 years ago
- Hadoop-based tool for extraction of large scale synchronous grammars for paraphrasing and machine translation☆15Dec 2, 2016Updated 9 years ago
- This is application for dysarthria to improve their pronunciation by using deep learning☆10Dec 29, 2020Updated 5 years ago
- Perform the forced decoding with target transcription☆11Sep 12, 2018Updated 7 years ago
- This repository demonstrates how it is possible to record and playback audio in most modern browsers, and most notably - Safari 11 on bot…☆44Feb 21, 2018Updated 8 years ago