Category · Speech & Audio

AI projects for Speech & Audio

Speech recognition, text-to-speech and voice cloning you can run yourself.

33 Projects
12 Categories
20 Technologies
32 Open source
2 projects found
Updating…
Sort projects by
Editor’s picks first Hand-picked projects lead, then the rest of the catalogue in curated order.
Featured

Whisper Verified

Speech & Audio

Speech recognition that holds up on accents, noise and 90+ languages.

Python PyTorch Hugging Face MIT
Intermediate Notebook Open Source
View project

Coqui TTS Verified

Speech & Audio

A text-to-speech toolkit with voice cloning from seconds of audio.

Python PyTorch Hugging Face MPL-2.0
Intermediate Notebook Open Source
View project
Explore by category

Browse Project Categories

Twelve build areas, from chatbots and agents to vector search and fine-tuning.

About Speech & Audio

Every project in Speech & Audio is open source and buildable — each one links straight to its repository so you can read the code before you commit an afternoon to it.

  • Filter by difficulty to find something that matches where you are
  • Every project lists its real tech stack, not a vague category
  • Licences are stated up front so you know what you can ship