← All Categories

Text-to-Speech

45 tools

Voice Inbox

Voice Inbox is a tool designed for quickly capturing thoughts on the go. It transcribes spoken words with human-level accuracy and saves them to a journal, allowing users to focus on expressing themselves and managing tasks. It integrates with Obsidian for seamless note-taking.

FreemiumFree tier

ClearCypherAI

ClearCypher LLC is a company that builds Generative AI products, including Audio to Audio (T2T) speech engine, Text to Audio (T2A) speech engine, and Audio to Text (A2T) transcription engine. They offer machine learning solutions specializing in automatic speech recognition, machine translation, optical character recognition, and speaker identification. Their platform provides language technology solutions for processing audio, video, image, and text content, delivering enterprise-grade language translation and voice biometrics.

FreemiumFree tier▴ 2

Audiosonic

Transform Text into Realistic Audio with Audiosonic by Writesonic

Audiosonic is an AI-powered text-to-speech tool developed by Writesonic that transforms written text into realistic, human-like audio 1. Its core purpose is to provide high-quality, engaging audio content quickly and easily, eliminating the need for expensive voice actors and recording studios 12. Key features include: Realistic, human-like audio generation using advanced deep learning algorithms 12 Support for over 30 languages and dialects 18 Customizable voice settings (gender, accent, tone, speed, pitch) 112 Instant AI voice generation 12 Commercial use clearance for generated audio 12 Potential applications include marketing and advertising, education, podcast production, accessibility solutions, software demos, and content repurposing. Audiosonic's unique selling points are its high-quality natural-sounding audio, extensive multilingual support, ease of use, instant audio generation, and seamless integration with Writesonic 112. Technically, Audiosonic is a cloud-based SaaS application requiring an internet connection 15. It integrates fully within the Writesonic platform, streamlining the content creation process 112. While specific awards or recognition are not documented, Audiosonic was released in September 2023 and continues to be improved 612. Its advanced capabilities and integration with Writesonic position it as a powerful tool for businesses and content creators seeking to efficiently produce high-quality audio content.

FreemiumFree tier

OpenWispr

Open source voice-to-text assistant, 3x faster than typing.

OpenWispr is an open-source, AI-powered voice dictation tool that converts your voice into formatted text instantly. It runs 100% locally, ensuring full privacy, and is designed to be 3-5x faster than typing. It's especially useful for prompting LLMs, writing emails, sending texts, and works seamlessly across various applications, allowing users to pick their preferred model and even edit the system prompt for full control.

FreemiumFree tier

SpeechLab

SpeechLab's Natural-Sounding AI Voice Solutions

SpeechLab provides a cutting-edge AI platform that overcomes language barriers using sophisticated speech-to-speech translation and dubbing tools. Supported by Andrew Ng’s AI Fund and leading investors, it delivers top-tier features like superior transcription, context-aware translation, and dubbed audio that sounds almost identical to human voices. Users can translate, transcribe, and dub material across various languages and dialects, achieving a flexible, detailed conveyance of ideas and feelings with lifelike accuracy. The service emphasizes ethical standards, mandating that users possess rights to any voices utilized, and follows rigorous protocols to prevent unauthorized voice cloning. Perfect for media, business, and education fields, SpeechLab fits effortlessly into current processes, offering a scalable, team-oriented platform customized for content producers, companies, and schools. Pricing options range from a free initial trial to full-service white-glove support, rendering premium dubbing and translation available to everyone.

FreemiumFree tier▴ 4

Podbrews

AI-Powered Document-to-Podcast Conversion with Podbrews

Podbrews is an innovative platform designed to convert your written documents into engaging podcast-style audio files using advanced AI technology. Users can seamlessly transform any PDF into a compelling audio experience, perfect for on-the-go listening or accessibility needs. Podbrews stands out by offering a plethora of audio styles to choose from, including sci-fi, fantasy, and public radio, ensuring that the final product matches the user's desired aesthetic. This tool is perfect for professionals, educators, and content creators looking to reach a wider audience through audio formats.

FreemiumFree tier▴ 1

SpeechFlow - Advanced Speech-to-Text API

SpeechFlow is a multilingual Speech-to-Text API that offers state-of-the-art accuracy in 14 languages. It converts sound to text, speech to text, and audio to text with high accuracy. SpeechFlow supports both cloud and on-prem deployment.

FreemiumFree tier▴ 7

luvvoice

Luvvoice: Free AI Text‑to‑Speech with 200+ Voices, 70+ Languages, and Voice Cloning

Luvvoice is a free online AI text‑to‑speech (TTS) platform that converts text and documents into natural‑sounding audio using real AI voices. With 200+ AI voices across 70+ languages and dialects, it supports advanced voice cloning, easy text‑to‑audio, and document‑to‑voice (including PDF). Luvvoice offers generous usage with no ads or CAPTCHA, extended character limits (up to 20,000 per conversion and 20,000,000 per month for standard voices), and flexible Free, Basic, and Pro plans—positioning it as a leading ElevenLabs alternative for 2025.

FreemiumFree tier▴ 1
W

WhisperUI - Text to Speech

WhisperUI is a text to speech and speech to text service powered by OpenAI Whisper API. With WhisperUI you can use your OpenAI api keys to get affordable text to speech and speech to text services. It allows users to convert audio files to text and SRT files using OpenAI Whisper Speech to Text.

FreemiumFree tier

Dictato

Dictato is a private, fast voice-to-text dictation application specifically built for macOS. It allows users to transcribe speech directly into any application—such as Gmail, Slack, or VS Code—using a global hotkey. The app operates 100% on-device, meaning no audio data is ever sent to the cloud, ensuring total privacy. It features three different transcription engines (Whisper, Parakeet, and Apple) to balance speed and language support, and it bypasses the standard 60-second limitation found in Apple's built-in dictation. It is designed for professionals who need to capture ideas at the speed of thought without compromising security.

FreemiumFree tier

Sayline

Sayline is a native macOS application designed for private, local voice dictation in any text field. It allows users to replace manual typing with voice commands using global hotkeys across various applications like Gmail, Slack, VS Code, or Notes. Utilizing on-device processing technologies (NVIDIA Parakeet and MLX), Sayline ensures uncompromised security and privacy by keeping all audio and data local to the user's Mac, never sending it to the cloud. Sayline is engineered to boost productivity, claiming to be 4x faster than manual typing.

FreemiumFree tier

babbly.co

Babbly is an early speech therapy tool that transforms playtime into progress. It uses AI-powered infant speech and brain development monitoring to identify the risk of developmental delays as early as 9 months. Babbly helps parents understand their child’s development by analyzing and monitoring their language progression and recommending activities to accelerate their development. It provides objective data to inform parental intuition and helps parents find out if their child is at risk of speech and language delays, which can be a sign of developmental conditions such as autism.

FreemiumFree tier

ListenRobo

ListenRobo is an AI-powered transcription platform that accurately transcribes, summarizes, and translates media files (audio & video) into text or subtitles for content creators. It supports 92 languages and offers features like fast and accurate transcription, privacy and security, and translation options. Users can transcribe audio and video to text or subtitles, generate English subtitles online, and download subtitles in various formats.

FreemiumFree tier

reccloud.cn

新一代AI音视频处理平台 | Next-gen AI audio/video processing platform

RecCloud is a leading AI audio and video processing platform that offers a range of tools for content creation and editing. It includes features like AI speech-to-text, AI subtitles, AI text-to-speech, and AI video translation. The platform is designed to be user-friendly and accessible online.

FreemiumFree tier▴ 2

Smart Dictate

Smart Dictate is a context-aware dictation and AI chat tool designed to enhance dictation and information extraction experiences across any website. The app analyses the content of the website and uses it as context for the next user operations. It's an AI-powered dictation tool that understands context, technical terms, and industry jargon, saving time with accurate voice-to-text across all websites.

FreemiumFree tier

Voicetypr

VoiceTypr is an offline AI voice-to-text application designed for founders and builders. It runs locally on your computer, ensuring privacy by default, and operates on a pay-once, use-forever model without subscriptions. It allows users to dictate text into various applications like ChatGPT, Claude, Cursor, VS Code, email, and more, supporting over 99 languages and offering features like smart formatting, high accuracy, and audio/video file transcription.

FreemiumFree tier

VoiceNovel

VoiceNovel is an advanced AI voice synthesis platform that transforms novels into high-quality voice novels and audiobooks. It leverages AI technology to convert text into natural-sounding speech, supporting multiple voice styles to give each character a unique voice and create an immersive listening experience. The platform offers features for novel upload and analysis, a personal library for converted audiobooks, and an audio player with download options for premium users.

FreemiumFree tier

VoiceRec: AI Vocal Recorder

AI-powered vocal recorder for capturing, transcribing, and sharing audio recordings.

FreemiumFree tier▴ 1

LMNT

Next-Level AI Text-to-Speech Solutions

Next Level AI Text to Speech. Ultrafast. Lifelike. Reliable. Experience low latency streaming designed for conversational apps, agents, and games, built from the ground up. Create remarkably authentic, expressive voices with studio-quality voice clones from just a 5-minute recording, or instant voice clones from 15 seconds. Or choose a voice from our library. Engineered by an ex-Google team. Handle unbelievable scale without a sweat and enjoy consistent low latency and high availability.

FreemiumFree tier▴ 1

DoThread

Simple, Flexible Life Journaling

DoThread is a simple, privacy-first app designed to help users organize notes, tasks, and ideas into time-based "Threads". It offers a fast, secure, and distraction-free environment to track progress and stay consistent. The app allows users to capture thoughts, whether spoken, written, or quickly jotted down, and organizes them into threads for easy navigation. DoThread emphasizes simplicity and privacy, storing all data securely in the user's personal iCloud account without keeping any information on its servers.

FreemiumFree tier

Cheetu AI

Your Lightweight Interpreter and AI Notetaker

Cheetu AI provides real-time transcription, live translation, and instant AI summaries for every meeting, lecture, or interview.

FreemiumFree tier▴ 1
PreviousPage 2 of 2