Skip to content

Navigation Menu

Sign in
Appearance settings

Search code, repositories, users, issues, pull requests...

Provide feedback

We read every piece of feedback, and take your input very seriously.

Saved searches

Use saved searches to filter your results more quickly

Appearance settings
#

voice-processing

Here are 34 public repositories matching this topic...

AI-powered disaster response platform with offline-first architecture using Gemma 3n. Provides computer vision hazard detection, voice analysis with emergency keywords, PDF report generation, and multi-user coordination - all working without internet access.

  • Updated Oct 23, 2025
  • HTML

A comprehensive AI companion leveraging advanced semantic analysis, sentiment detection, and voice processing to provide personalized and context-aware interactions using Autogen, semantic-router, and VoiceProcessingToolkit.

  • Updated Jun 2, 2024
  • Python

The VoiceProcessingToolkit is an all-encompassing suite designed for sophisticated voice detection, wake word recognition, text-to-speech synthesis, and advanced audio processing. It offers intuitive interfaces to streamline the integration of voice processing capabilities into your applications

  • Updated Jun 5, 2025
  • Python
AI-CallConnect

A cutting-edge AI-powered phone agent designed for seamless voice interactions, dynamic data handling, and scalable communication. Perfect for modern sales and customer engagement solutions.

  • Updated Jan 15, 2025
  • Jupyter Notebook

🎙️ One-click deploy AI voice platform to Cloudflare Workers! 🚀 TTS & STT with 650+ voices across 154 languages. Microsoft Edge TTS + SiliconFlow API. Free, serverless, edge-powered. Optional statistics (KV/D1), multilingual UI (9 languages). No registration required.

  • Updated May 20, 2026
  • JavaScript

🎧 Transcribe any audio to text in seconds using OpenAI Whisper — right in Google Colab. No setup needed! Upload your MP3, WAV, M4A, or FLAC file and get accurate, multilingual transcriptions powered by Whisper’s medium model — all free in the cloud. ☁️

  • Updated Nov 8, 2025
  • Jupyter Notebook

This repository is made in lieu of submission towards the solution of problem statement 2 of the OPEN AI NLP hackathon. The objective here is to classify the voice recordings of a call center proceeding by treating them as consumer complaints into the said categories of the automotive industry.

  • Updated Sep 8, 2019
  • Jupyter Notebook

Improve this page

Add a description, image, and links to the voice-processing topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the voice-processing topic, visit your repo's landing page and select "manage topics."

Learn more

Morty Proxy This is a proxified and sanitized view of the page, visit original site.