Open-source voice data collection platform for building inclusive voice datasets. Collaborative transcription with quality consensus. FastAPI + React + PostgreSQL.
-
Updated
Apr 9, 2026 - Python
Open-source voice data collection platform for building inclusive voice datasets. Collaborative transcription with quality consensus. FastAPI + React + PostgreSQL.
Baseline Emotion Recognizer recognizes emotion from audio data. It uses openSMILE tool kit to extract the features. For classification it uses SVM machine learning technique. For the classification it uses scikit-learn library.
Real-time speech transcription client integrating with the Speechmatics API. Streams audio over WebSocket, delivers live word-by-word transcripts via Server-Sent Events, and surfaces confidence scoring, speaker diarisation, and session analytics in a vanilla JavaScript dashboard. Built with Python and FastAPI.
To associate your repository with the speech-technology topic, visit your repo's landing page and select "manage topics."