Found 60 results for “voice”
MyVocal.ai is a tool that provides voice synchronization and voice cloning features. Users can synchronize their own voice with popular music and complete voice cloning in a relatively short time.
Lovo is an AI voice generation and text-to-speech tool that supports converting text into natural speech, suitable for audio content production, voiceover, and various creative scenarios, helping reduce manual recording costs and time investment.
PolyAI is a company that provides enterprise-grade voice assistant solutions, focusing on handling customer calls through natural conversational AI to help businesses improve phone service efficiency and automation.
Neon AI is an open-source voice application development platform that combines artificial intelligence and natural language understanding capabilities, supporting the creation of voice-interaction applications similar to smart voice assistants.
MyShell is an AI chat and character conversation tool that supports logging in and using it in multiple ways, making it suitable for communicating with different AI characters and experiencing conversational interaction.
Procys is a data extraction tool for invoice and bill processing that uses machine learning to automatically identify and extract key information, reducing manual entry and organization work.
Koe Recast is an AI voice conversion tool that can transform a user's voice into different styles, suitable for scenarios such as voice chat, virtual social interaction, and real-time interaction.
Open Voice OS is an open-source project for voice artificial intelligence, operating as a community-driven Linux distribution designed to enable voice AI to be deployed and customized across devices such as smart speakers, phones, and TVs.
Translate.video is an AI translation tool for video content, supporting video translation, subtitle translation, dubbing, AI voice conversion, recording, and text generation to help distribute video content in multiple languages.
AI Voice Detector is an audio authenticity detection tool used to identify whether speech is generated by AI. Users can upload audio files for verification, making it suitable for scenarios involving evidence review, media judgment, and authenticity analysis in customer communications.
Article.Audio is an online service that converts article content into spoken audio, supporting the transformation of text articles into listenable audio for convenient access to information when reading is inconvenient.
HeyGen is an online AI video generation tool that supports creating talking avatar videos and provides customizable avatars and voiceover features, suitable for content production scenarios such as training, teaching, explanations, and marketing.
Artflow.ai is an AI tool for animated story creation, supporting the generation of characters, scenes, and voices, and can also write dialogue, helping users build original visual narrative content.
Adobe Speech Enhancer is an AI audio enhancement tool for improving the quality of voice recordings. It can reduce background noise and highlight voices, making ordinary spoken recordings sound clearer and closer to a studio effect.
GitHub Next is a collection of AI experimental projects for development scenarios, among which Copilot Voice supports coding and operations through voice commands. Users can use "Copilot" as the wake word and replace keyboard input with voice in parts of the programming workflow.
Kore.ai is an enterprise service product focused on conversational AI, helping businesses build intelligent assistants and automated workflows for customers, employees, and agents across voice and digital channels.
Novels AI is a tool for generating personalized audio adventure stories, allowing users to customize characters and plot choices and experience AI-driven immersive story content in audiobook form.
AItoGrow is an AI tool directory and information website for startup teams, bringing together categories such as marketing, text generation, image generation, voice generation, productivity improvement, and personal assistants, while also providing AI market statistics and industry trend information.
AIDev.Codes is a tool for generating interactive web pages by conversing with AI, supporting text generation, image generation, an optional voice interface, as well as free hosting and custom subdomains.
VoiceLine is a website that provides hybrid voice and text messaging services. It allows users to send voice notes of up to 60 seconds, providing better understanding, tone, and connection. VoiceLine also offers the latest AI insights to help oversee sales. The website was developed in collaboration with more than 300 of the world's most innovative companies. Using VoiceLine can make input faster and more efficient, freeing up more time to focus on work while minimizing barriers to collaboration and connection.
Murf AI is an AI voice generation tool that converts text into natural, lifelike human speech, suitable for creating podcasts, video voiceovers, presentation narration, and other audio content.
NarrationBox is an AI voice generation tool that offers more than 700 AI narrator voices for creating audio content such as podcasts, audiobooks, and dubbing.
AiCogni is an AI voice and writing assistant powered by ChatGPT, supporting multilingual understanding and responses. It can be used for Q&A, creative generation, and writing assistance, and provides voice control and Wear OS device support.
Krisp is an AI-based noise cancellation app mainly used to improve the quality of online meetings and voice communication. It supports Mac and Windows, and offers voice productivity-related features and a free version.
Revocalize AI is an AI voice synthesis tool that supports voice cloning, voice protection, and voice creation, offers multilingual voice options, and is suitable for audio content production and personalized voice applications.
Groot Music is a music bot that runs on Discord, supports multilingual use, and provides advanced features including AI tools, making it suitable for voice interaction and music playback needs in communities.
Voiceful provides game character voice generation and speech synthesis demos, and supports integration into Unity via SDK, making it suitable for development and testing scenarios that require character voice capabilities.
FineShare FineVoice is an AI real-time voice changer that supports instant voice adjustment and personalized processing during meetings, live streams, chats, and gaming.
Parseur is an intelligent processing tool for document data extraction that can automatically capture text and structured information from emails, invoices, forms, and other content to reduce manual data entry work.
Spatial.ai is a tool service provider that uses social data for consumer insights and behavioral analysis, helping teams incorporate users' voices from social media into market research and decision-making.
Voiceflow is a collaborative tool for building conversational assistants, supporting the design, prototyping, and publishing of chatbots, voice assistants, IVR, and Web Chat experiences through drag-and-drop.
Replica Studios is an AI voice generation tool for creative projects, offering virtual voice actors with emotional expression that can be used to create more natural voice performance content.
Powtoon is a website for creating videos and animations online. They provide professionally designed templates as well as useful tips, training courses, and guides to shorten the learning curve. Users can create stunning videos and presentations on the website, and before exporting the final video they can benefit from royalty-free video footage, images, animations, characters, voiceovers, or music. Powtoon can also be used to import PowerPoint presentations and convert them into videos.
This is a how-to guide on using Siri voice commands to call ChatGPT, covering how to ask questions, get answers, and configure saving the results to Notes.
Voicemod is a website that provides real-time voice changing and custom sound effects, serving desktop applications and games such as Discord, ZOOM, Google Meet, and Minecraft.
Hirex.ai is a no-code voice interview robot platform that can automate large-scale recruitment assessments. It supports voice interviews and integrates coding tests, multiple-choice questions, video interviews, hackathons, and WhatsApp chatbots.
WhatsGPT is a GPT chat tool that can be used in WhatsApp and Telegram, supporting text and image processing, allowing users to have AI conversations directly within familiar messaging apps.
Suki is an AI voice assistant for doctors, used to reduce documentation and administrative burdens so clinical staff can devote more time to patient care.
Voicera is an article-to-speech tool that can automatically detect content and generate a playable audio version. It supports multiple languages and voice options, making it convenient for users to access information by listening.
Poised is an AI communication coaching tool that provides real-time feedback, helping users improve spoken expression and overall communication skills during voice interactions.
Krater.AI is an AI-driven content creation tool that provides a suite of tools for marketers and content creators. It offers solutions such as ad copy generation, creating stunning images, and converting audio content into written content or realistic voiceovers. The website claims to have advanced technology comparable to Jasper, Midjourney, and Writer.com. Krater.AI is designed to be user-friendly and intuitive, offering features such as image generation, copywriting, chat, speech-to-text, code, and more. The website also has a Twitter account where they share updates about their product.
Tavus is an AI personalized video generation tool for product, marketing, and sales teams. It can mass-produce customized videos for different audiences based on templates, and use voice variables to deliver communication content that better matches the recipient.
Abei Intelligence is a one-stop AI picture book creation platform designed for children's education. Through three simple steps—story creation, image generation, and intelligent voice-over—users can quickly create personalized picture books. Abei Intelligence encourages parent-child interaction, cultivates children's creativity, emotional expression, and language skills, while incorporating science, moral education, and physical activities to spark children's interest in technology and help them thrive in the intelligent era.
AiPy is a free and open-source AI agent factory, a local version of Manus, built on large language models (LLMs) and Python capabilities. It supports local deployment to ensure data privacy and security. Through the "Python-Use" paradigm, AiPy gives AI "hands," enabling it to analyze local data, operate local applications, and execute complex tasks such as controlling phones, generating multi-voice speech, analyzing medical test reports, extracting speech from videos, and sending scheduled emails.
AnyGen is an AI office agent launched by ByteDance that improves office efficiency through voice input and AI technology. Users can press and hold the recording button to quickly convert speech into text, with support for adding photos, screenshots, and links, avoiding the tedious organization required after traditional note-taking.
Deepgram is a platform that provides advanced AI speech recognition and natural language processing technology. Its core products are powerful Speech-to-Text (STT) and Text-to-Speech (TTS) APIs, enabling developers to quickly integrate voice transcription and understanding capabilities into their own applications and services.
Doudou AI is an AI gaming companion launched by Xinying Suixing. Users can interact intimately with virtual characters such as catgirls, such as rubbing cheeks and touching tails, and increase intimacy through voice chat. Through dual modes of desktop pet and floating ball, Doudou provides non-intrusive companionship and offers help only when users need it.
ElevenLabs is an AI text-to-speech platform that provides realistic voice synthesis solutions for developers, creators, and enterprises. Its core products include text-to-speech (supporting 29+ languages including Chinese and 10,000+ voices), AI dubbing, voice cloning, music generation, and more.
JoyPix is an AI creation tool focused on digital humans and speech synthesis. Users can create personalized virtual avatars by uploading photos, with support for voice conversations with virtual avatars.
LangLang Voiceover is an intelligent text-to-speech tool that provides voice synthesis services. It supports more than 30 languages, including Chinese, English, German, and French, as well as more than 10 emotional styles such as happy, sad, and excited. The platform is feature-rich and easy to use, supporting SSML tags to enable advanced functions such as polyphonic character handling and multi-speaker dubbing.
AI real-time voice changing tool
Miaochuang (formerly "Yizhen Miaochuang") is an intelligent AI content generation platform based on the Miaochuang AIGC engine, providing creators and organizations with AI generation services including text continuation, text-to-speech, text-to-image, and image-and-text-to-video. Yizhen Miaochuang intelligently analyzes copy, assets, AI voice, subtitles, and more to quickly produce finished videos, enabling zero-threshold video creation.
Xmov Nebula is an embodied intelligent 3D digital human open platform launched by Xmov Technology, dedicated to upgrading AI from “having a brain” to “having a body” to enable natural expression and interaction. Based on text input, Xmov Nebula can generate a 3D digital human’s voice, expressions, and movements in real time, supporting multimodal generation, low-cost operation, low-latency interaction, and multi-terminal adaptation.
Moyin Workshop is a professional AI voiceover tool with more than 800 voices and over 1,000 styles, meeting a wide range of needs from video dubbing to audiobooks. Moyin Workshop offers rich features, including speech rate adjustment, polyphonic character selection, and pause control, ensuring realistic and natural text-to-speech results. Users can easily download lossless audio files and enjoy a convenient voiceover experience.
Octane AI is a product quiz and zero-party data tool for Shopify brands, helping merchants collect user preferences through Q&A and provide more personalized shopping experiences and marketing support.
Tingnao AI is an AI-powered intelligent voice assistant focused on speech-to-text and real-time recording summaries, offering audio/video transcription, real-time recording-to-text, AI summaries, chapter overview, and other features. Users can freely drag text to view audio/video progress and enjoy a convenient intelligent recording experience.
Uberduck is an open-source community for AI voice generation and synthesis. The platform offers more than 5,000 voices to help users create AI dubbing and speech, and you can even use your own custom voice clone for synthesis.
AI text-to-speech generation tool
Wondershare Virbo is an AI digital human talking-video marketing tool launched by Wondershare Technology, focused on providing video creators and cross-border e-commerce practitioners with a full-chain AIGC creation experience. The software uses advanced AI technology to allow users to quickly generate HD videos containing digital human characters, dynamic scenes, and precise backgrounds through simple text input or voice files.
Wondercraft is a versatile AI audio content creation platform that uses generative AI voice technology to allow users to quickly convert text content into podcasts, audiobooks, ads, and other audio formats.
