AI model purpose-built for tasks that need a consistent, reliable answer every time — reading documents, classifying content, transcribing speech.
Results for “speech”
20 tools found
Real-time voice AI models built for near-instant, natural-sounding speech — the audio layer behind voice agents that need to feel like a real conversation.
Speech recognition and voice-generation API that lets developers add accurate, fast transcription and natural-sounding speech to any app.
No-code-friendly platform for building AI phone agents that handle customer service and sales calls, with drag-and-drop configuration.
Text-to-speech API built by linguists for hundreds of natural-sounding voices across 50+ languages, tuned for real conversation, not audiobooks.
Developer platform that wires together speech recognition, an AI model and speech generation into a working phone-answering voice agent.
An AI phone agent that answers a Shopify store's inbound support calls 24/7 — handling 'where's my order', returns and product questions so a human team doesn't have to.
An app that turns your spoken words into clean, properly punctuated text in any app you're typing into — like having a fast, tidy typist sitting next to you all day.
An AI platform that conducts adaptive voice and video job interviews at scale and scores candidates automatically.
A free, open-source Mac app that turns your voice into text and can run simple AI agent tasks, all without touching the internet.
A bot-free AI meeting assistant for Mac that listens to your calls and writes the notes, without ever showing up as a visible participant.
Dictée vocale en ligne — convertir la voix en texte
IBM Text to Speech
Local Whisper-based speech-to-text with a global hotkey. [![Open-Source Software][OSS Icon]](https://github.com/TypeWhisper/typewhisper-mac) ![Freeware][Freeware Icon]
Real-time speech-to-text app. [![Open-Source Software][OSS Icon]](https://github.com/Beingpax/VoiceInk) ![Freeware][Freeware Icon]
Multi-provider speech-to-text with AI transformations and keyboard shortcuts. [![Open-Source Software][OSS Icon]](https://github.com/EpicenterHQ/epicenter/tree/main/apps/whispering) ![Freeware][Freeware Icon]
Open Source Toolkit For Speech Recognition purely based on Java speech recognition library.
A free codec for free speech. Obsoleted by Opus. [BSD]
Open-source on-device AI runtime that runs LLMs, vision models and speech models directly on phones, laptops and wearables — no cloud round-trip required.
AI content moderation API that detects harmful text, images, video and audio (NSFW, hate speech, violence, CSAM, AI-generated content) at scale.