FakeYou - Deep Fake Text to Speech

FakeYou - Deep Fake Text to Speech

Learning Machines, Inc. (Storyteller) · 音声・言語

FakeYou - Deep Fake Text to Speech is a browser-based AI voice generator that turns typed text into audio using more than 3,500 community-built celebrity, character, and custom voices. It also handles voice conversion, custom voice design, and lip-sync video, so you can re-voice an existing clip or pair generated speech with a moving face. The core tools are free to use, with Plus, Pro, and Elite subscriptions adding faster queues, longer clips, and private model uploads.

FakeYou - Deep Fake Text to Speech の画面プレビュー

FakeYou - Deep Fake Text to Speech について

What Is FakeYou

FakeYou is a deepfake text to speech platform built around a community library of voice models. Type a line, pick a voice like a cartoon character or a public figure, and the site renders an audio clip you can download or share. The same account gives you voice conversion, which takes a recording of your own voice and outputs it in someone else's timbre.

The appeal is breadth, not polish. FakeYou launched in 2020 as Vocodes and rebranded under its current name while the model catalog grew through user contributions rather than an in-house studio team. That means the collection is huge. It's also inconsistent.

So what do you actually get? A huge catalog where some models land close to the original and others sound compressed and flat. You learn to spot the good ones fast.

Two rules shape how you can use the output. Every generated audio file carries a watermark, and the terms prohibit commercial use unless you pick a voice that's specifically marked for it. Published clips must also be labeled as deepfakes. The platform is owned by Learning Machines, Inc., which operates under the Storyteller brand.

Getting Started

  1. Create a free account at fakeyou.com, or sign in with an existing one to save your generation history.
  2. Open the Text to Speech page and search the voice library, filtering by language or community rating until you find a model worth testing.
  3. Type or paste your script into the text box, keeping an eye on the clip length limit that comes with your plan tier.
  4. Hit generate and wait for the queue. Free users sit behind paying users during busy periods, so render time climbs at peak hours.
  5. Preview the result, regenerate if the tone lands wrong, then download the MP3 or WAV and label it as a deepfake before you publish.

AIツール情報

FakeYou - Deep Fake Text to Speechの料金、対応プラットフォーム、性能を簡単に確認できます。

無料プランはい
有料プラン$12 - $40/mo
プラットフォームWeb
開発元Learning Machines, Inc. (Storyteller)
カテゴリ音声・言語
リリース日May 2020
最終更新Sep 2026
サイト訪問数436.6K
サイト世界ランキング93.6K
API提供状況はい

こんな人におすすめ

このツールが最も力を発揮するユーザー、タスク、シーン。

ユーザー

  • Content creators and meme makers
  • Fan dubbers and hobby voice actors
  • Developers building bots

タスク

  • Generating a quick character line for a video intro
  • Turning a recorded voice memo into another persona
  • Cloning a signature voice for a personal project

シーン

  • Producing fan content on a hobby budget
  • Experimenting with voice cloning before committing to paid software
  • Making lip-sync videos for social posts

主な機能

Community Voice Library

More than 3,500 voice models sit behind the search box, spanning celebrities, cartoon characters, anime figures, historical names, and original personas. Featured picks and community ratings help you sort through the pile instead of guessing. Quality is a lottery. One model sounds near-perfect, the next sounds like a phone call from 2004.

Text to Speech Generation

This is the core tool. You type a script, choose a voice, and the platform renders speech you can replay or download. Short lines are where it shines. Longer passages take more time and can lose the emotional nuance the shorter clips carry. Output comes out as MP3 or WAV.

Voice to Voice Conversion

Voice conversion takes an existing recording and outputs it in a target voice from the library, keeping much of the original pacing and delivery. You can adjust pitch shift, pitch estimation, and automatic F0 conversion before you run it, and pitch can move by as much as 36 semitones, which opens up voices well outside your own range. Results vary depending on how close your source recording is to the target model's training data.

Voice Designer

The Voice Designer builds a custom AI voice from audio samples you upload, guided step by step through the process. It suits creators who want a signature sound rather than borrowing someone else's. The tool is still marked beta, so expect rough edges. Pro and Elite subscribers get private model storage, and Elite adds sharing options on top. AI voice cloning through this route keeps the resulting model in your own account until you decide to publish it.

F5-TTS and Seed-VC Engines

FakeYou runs a newer generation of engines alongside its classic TTS system. F5-TTS handles zero-shot cloning with more natural prosody and emotion, and it swaps between English and Chinese mid-sentence. Seed-VC powers real-time voice conversion. It keeps the speaker's emotion and timing intact while changing the vocal identity. Both sit on dedicated pages separate from the main text to speech flow.

Lip-Sync Video

Generated audio can be pushed into lip-sync video, making a character's mouth move to match the line you produced. It's a practical shortcut for memes, dubbing, and short-form clips where paying for studio animation isn't an option, and you can pair it with a downloaded voice clip for a finished piece without leaving the browser.

Public API and Documentation

An API sits behind the consumer site with documentation covering endpoints, response codes, and authentication. Optional API tokens bypass the default IP rate limit and open up privately uploaded voice models. Developers have used it to power community bots and integrations. The docs also cover additional endpoints on request.

メリットとデメリット

メリット

  • A free tier covers the core text to speech and voice conversion tools, so you can test the workflow without a subscription.
  • The community voice library spans more than 3,500 models and beats most commercial TTS catalogs on sheer variety.
  • API access with optional tokens lets developers build bots and integrations on top of the same voice models.
  • Voice conversion keeps your original delivery and pacing while changing the vocal identity, which saves re-recording.
  • Downloadable MP3 and WAV output means generated clips drop straight into editing software.

デメリット

  • Audio quality swings widely because models come from different community trainers. A voice that sounds great in a preview can fall apart on full sentences.
  • Free users queue behind paid subscribers during peak hours. That turns a quick clip into a five-minute wait when the site is busy.
  • Every generated file is watermarked, and the terms block commercial use unless the voice is specifically flagged for it. That rules FakeYou out for client work.
  • Longer clips sit behind higher tiers. Unlimited voice conversion is reserved for the top Elite plan.
  • Community-uploaded celebrity and public-figure models raise recurring likeness and consent questions that you have to weigh before publishing.

よくある質問

Yes. The core text to speech and voice conversion tools work at standard speed without paying. Free output does queue behind paid users when traffic spikes. Plus, Pro, and Elite subscriptions add faster priority, longer clips, and private model uploads.

関連コンテンツ

FakeYou - Deep Fake Text to Speechに関連するツール、スキル、記事を探す。

FakeYou - Deep Fake Text to Speechの代替ツール

Prosp

Prosp

Prosp · ライティング · 音声・言語 · マーケティング

Prospは、代理店や営業チーム向けに作られたAI LinkedInアプローチツールだ。見込み客ごとに自分の声でメッセージとボイスメッセージを書き、相手が本当に返信するように仕向ける。アカウントをつなぎ、リードを見つけ、AIにパーソナライズされたメッセージを大量に下書きさせて送信させる。すべて1つの受信トレイから行える。大量のアウトバウンドを回す人向けにできている。それがこの製品の売り込みのすべてだ。それでも各接点は人間らしく感じられなければならない。

有料 / $30.99 - $79.99 per account/mo詳細を見る
Wordly AI Translation

Wordly AI Translation

Wordly · 音声・言語 · 生産性向上

Wordly AI Translationは、会議・カンファレンス・イベント向けに作られたリアルタイムAI翻訳・字幕プラットフォームだ。60以上の言語でライブ翻訳、字幕、文字起こし、要約を提供し、参加者は専用ヘッドセットを使わず、QRコードを読み取るかリンクを開くだけで参加できる。Zoom、Microsoft Teams、Google Meet、Webexと連携し、セッションごとに人間の通訳を雇わずに多言語対応したい組織に向く。それだけのことだ。

有料 / $0 - $150/mo詳細を見る
Musicful

Musicful

Musicful AI · 音声・言語 · 動画

Musicfulは、テキストを数分で音楽に変えるAI音楽ジェネレーター兼AIミュージックビデオ作成ツール。テキストの指示、歌詞、あるいは口ずさんだメロディーを渡すと、ボーカルと楽器が入った完成したトラックを返す。AI楽曲ジェネレーターとしても機能し、自分が作った曲からミュージックビデオを生成し、AIカバーツールと開発者向けAPIも備える。ウェブブラウザとAndroidアプリで動くので、パソコンで曲を始めてスマホで続きを進められる。

無料 / $0 - $20/mo詳細を見る