Search results for "TTS"

Found 14 results (11 tools · 0 articles · 3 skills). Sorted by relevance by SeoAIu.

AI Tools (11)

Vocu AI — A domestic large-scale speech model ranked #1 globally on the HuggingFace TTS Arena, capable of infusing AI-generated speech with a human touch and genuine emotion.

153
Vocu AI — A domestic large-scale speech model ranked #1 globally on the HuggingFace TTS Arena, capable of infusing AI-generated speech with a human touch and genuine emotion.
AI Tools

Vocu AI is an AI speech synthesis platform developed in-house by Guangzhou Shuogu Technology; its V3 series models ranked first globally in the Hugging Face TTS Arena blind tests. It offers both instant cloning (requiring a 3-second sample with over 95% similarity) and professional-grade cloning. Supporting more than 30 languages ​​and dialects, the platform delivers precise emotional expression and cinematic-quality performance. API access is available to facilitate efficient integration for developers.

2026-07-11

PubMedQA—an "AI benchmark" specifically designed for biomedical question answering; only by comprehending research papers can an AI truly pass the "Medical Turing Test."

136
PubMedQA—an "AI benchmark" specifically designed for biomedical question answering; only by comprehending research papers can an AI truly pass the "Medical Turing Test."
AI Tools

PubMedQA is the first question-answering dataset requiring reasoning over biomedical research texts; it was released in 2019 by institutions including the University of Pittsburgh. The task involves answering "Yes," "No," or "Maybe" questions based on PubMed abstracts. Comprising 1,000 expert-annotated samples and 211,000 artificially generated ones, the dataset aims to evaluate the ability of AI models to comprehend and reason about complex medical literature. It is widely used to benchmark the performance of large language models in the medical domain.

2026-07-10

Deepgram—a leading enterprise-grade voice AI platform offering real-time speech recognition, synthesis, and fully managed voice agent APIs.

112
Deepgram—a leading enterprise-grade voice AI platform offering real-time speech recognition, synthesis, and fully managed voice agent APIs.
AI Tools

Deepgram is a leading voice AI platform that provides developers with high-accuracy, cost-effective real-time speech-to-text (STT), text-to-speech (TTS), and a unified voice agent API. Its Nova series models outperform competitors in both accuracy and speed; Aura-2 TTS offers latency under 200 milliseconds, and the voice agent API is priced at just $4.50 per hour. Supporting both cloud and self-hosted deployments, the platform is trusted by over 200,000 developers.

2026-07-11

TTSMaker—A free online text-to-speech tool; generate over 200 voice styles with a single click.

107
TTSMaker—A free online text-to-speech tool; generate over 200 voice styles with a single click.
AI Tools

TTSMaker is a free online AI text-to-speech tool that supports multiple languages ​​and over 200 voice styles. No registration is required; simply enter text to generate natural, realistic speech, adjust speed and volume, and download MP3 files. Ideal for video voiceovers, audiobook production, and educational content creation, it allows text to easily "come to life" with a voice.

2026-07-11

Murf AI—From voiceover studio to real-time voice agent: Generate professional-grade voices with AI.

105
Murf AI—From voiceover studio to real-time voice agent: Generate professional-grade voices with AI.
AI Tools

Murf AI is an AI voice generation and conversational platform offering over 200 ultra-realistic voices, support for more than 20 languages, and voice cloning capabilities. Its Falcon TTS API enables real-time voice agent deployment with an industry-leading 55ms latency, while the Studio editor allows for audio-video synchronization, making it suitable for applications such as video voiceovers, course creation, and intelligent customer service.

2026-07-11

Moyin (Moyin Workshop)—an AI voiceover powerhouse featuring over 800 voices and 1,000 styles, trusted by creators of short videos and audiobooks.

100
Moyin (Moyin Workshop)—an AI voiceover powerhouse featuring over 800 voices and 1,000 styles, trusted by creators of short videos and audiobooks.
AI Tools

Moyin (Moyin Workshop) is an AI voice synthesis platform under Mobvoi. Powered by the proprietary "Sequence Monkey" (Xulie Houzi) large model and a fifth-generation TTS engine, it offers over 800 voice profiles, more than 1,000 styles, and nearly 20 fine-tuning features. The platform supports multiple languages ​​and dialects, voice cloning, cloud-based video editing, and multi-user collaboration; it is widely used in applications such as short videos, audiobooks, and film/TV commentary, having served over 6 million users to date.

2026-07-11

Beatoven.ai—an AI music generator built for creators; create emotion-driven soundtracks and ensure copyright never stands in the way of your creativity.

99
Beatoven.ai—an AI music generator built for creators; create emotion-driven soundtracks and ensure copyright never stands in the way of your creativity.
AI Tools

Beatoven.ai is an AI-powered, royalty-free music generation platform specializing in creating emotive soundtracks for videos, podcasts, games, and more. It generates unique background music from text, images, or videos, offering customization across 16 moods and various styles. Its Maestro model is trained on fully licensed data and provides royalty sharing for musicians, having already helped creators worldwide generate millions of tracks.

2026-07-11

LOVO AI — A lifelike TTS platform with an integrated AI video editor, featuring over 500 voices and more than 100 languages.

99
LOVO AI — A lifelike TTS platform with an integrated AI video editor, featuring over 500 voices and more than 100 languages.
AI Tools

LOVO AI is a high-fidelity text-to-speech and video creation platform. Its core product, Genny, integrates voice generation with online video editing, offering over 500 voices and 100 languages, alongside features such as AI scriptwriting and automatic subtitling. Supporting voice cloning and team collaboration, the platform is ideal for content creators, marketers, and educators.

2026-07-11

TTSMaker — A free, commercially usable AI text-to-speech tool offering a choice of over 300 voices across more than 50 languages.

96
TTSMaker — A free, commercially usable AI text-to-speech tool offering a choice of over 300 voices across more than 50 languages.
AI Tools

TTSMaker is a free, no-registration AI text-to-speech tool that supports over 50 languages ​​and more than 300 voice styles. You can convert text into natural-sounding speech in just three steps without signing up, and the generated audio can be used for commercial purposes—such as video voiceovers and audiobook production—free of charge. It offers a free monthly allowance of 30,000 characters, with unlimited usage available for select voices.

2026-07-11

Voicemaker — A cost-effective AI voiceover platform featuring over 130 languages ​​and 1,500+ voices, with unlimited use of free standard TTS.

92
Voicemaker — A cost-effective AI voiceover platform featuring over 130 languages ​​and 1,500+ voices, with unlimited use of free standard TTS.
AI Tools

Voicemaker is an AI text-to-speech tool offering a selection of over 130 languages ​​and 1,500+ voices, with support for emotion control, voice cloning, voice-to-voice conversion, and API integration. Its ProPlus models support SSML and voice effects, while the free version offers unlimited standard voiceovers. It is suitable for video narration, course creation, and IVR systems.

2026-07-11

Mobvoi Open Platform - One-stop AI Voice Interaction and NLP Services for Speech Recognition, Semantic Understanding and TTS

87
Mobvoi Open Platform - One-stop AI Voice Interaction and NLP Services for Speech Recognition, Semantic Understanding and TTS
AI Tools

Mobvoi Open Platform is a one-stop AI capability platform launched by Mobvoi, offering a range of artificial intelligence services including speech recognition, semantic understanding, text-to-speech (TTS), and wake-up word customization. Designed for developers, enterprises, and partners, the platform supports rapid integration of intelligent voice interaction features, widely used in smart hardware, in-car systems, smart home, and customer service robots. With rich APIs and SDKs, it helps users lower the barrier to AI development and achieve efficient, stable voice interaction experiences.

2026-07-15

Career Skills (3)

Eldercare Exercise: Voice-Guided Bedridden Rehabilitation & Home Assistant Integration

Eldercare Exercise: Voice-Guided Bedridden Rehabilitation & Home Assistant Integration

A lightweight, TTS-guided exercise program for bedridden or seated elderly. Features 3 levels (super light, light, moderate), integrates with Home Assistant for cron scheduling, motion detection, SOS conflict prevention, multi-elder support, and daily reporting. Safe for 90+ with physician approval required.

护理助理技能
0 2026-08-24
Eldercare Profiles — Multi-elderly Care Profile Manager with Chat-based Management

Eldercare Profiles — Multi-elderly Care Profile Manager with Chat-based Management

A Home Assistant skill designed for family eldercare, allowing creation and management of multiple elderly profiles (grandma, grandpa, aging parents) with individual sensor mappings, contacts, TTS settings, and room-based configuration. Supports chat commands for adding/modifying/deactivating profiles, and auto-migration from existing hardcoded setups. Ideal for multi-elderly households needing centralized monitoring.

家庭健康助手技能
0 2026-07-30
Eldercare Exercise — Voice-Guided Bedridden Workout with Auto Timer and Level Adjustment

Eldercare Exercise — Voice-Guided Bedridden Workout with Auto Timer and Level Adjustment

A lightweight exercise skill designed for elderly individuals who are bedridden or can sit. It uses TTS to guide each movement step by step, with automatic counting, rest breaks, and three difficulty levels (bedridden, sitting, standing with support). Supports daily cron at 9 AM or chat triggers, integrates with Home Assistant media player for voice output. Disabled by default; requires family activation and medical clearance.

家庭健康助手技能
0 2026-08-02