Ghostcut

Ghostcut

JollyToday · Voice & Language · Video

Ghostcut is an AI-powered localization platform from JollyToday that rebuilds a video for a new language in a single workspace. It generates subtitles, erases burned-in text, translates speech, and dubs it in a cloned voice. The product targets short-drama studios, ecommerce sellers, and creators who push out hundreds of clips a month and need the process to run in batches.

About Ghostcut

What Is Ghostcut

Ghostcut is a web-based video localization tool built by JollyToday, a Guangzhou company whose Chinese product is known as Guishou Jianji. It bundles four jobs that usually live in separate apps: subtitle extraction, subtitle removal, translation, and voice dubbing. The pitch is simple. Instead of paying a subtitling house for every episode, you upload your files and let the platform do the repetitive work overnight.

The main problem it solves is the gap between a video that's finished in one language and a version that works in another. Burned-in Chinese subtitles ruin an English upload. A voice track in Mandarin won't land with a Spanish audience. Ghostcut attacks both ends: it wipes the original text off the frame and rebuilds the audio in the target language.

There are limits worth knowing up front. The heavy lifting still needs a human pass. Machine translation gets the timing and the raw lines in place, but names, jokes, and cultural references often need a check before publishing. The final look of a subtitle removal also depends on the source footage, so busy backgrounds with moving text can leave more visible traces than clean static ones.

Getting Started

  1. Create an account on jollytoday.com and confirm your email to open the workspace.
  2. Upload a video file, or point the batch manager at a folder of episodes if you're working at scale.
  3. Let the OCR and ASR engines detect the on-screen text and spoken lines, then pick the ones you want handled.
  4. Choose your target language and dubbing voice, and run translation and voice cloning on the selected track.
  5. Review the result, adjust subtitle timing or speaker labels where needed, then render and export.

Product Information

A quick look at Ghostcut's pricing, supported platforms, and performance.

Free PlanYes
Paid Plans$0 - $100/mo
PlatformWeb
DeveloperJollyToday
CategoryVoice & Language · Video
Release DateMay 2021
Latest UpdatedSep 2025
Website Visits76K
Website Global Rank401.6K
API AvailabilityYes

Best for

The users, tasks, and scenarios where this tool fits best.

Users

  • Short-drama studios shipping episodic content to YouTube and overseas apps
  • Ecommerce sellers running video ads on TikTok and similar platforms
  • Solo creators and UGC channels

Tasks

  • Removing hard-coded subtitles and watermarks
  • Translating and dubbing dialogue
  • Localizing a full season in one pass

Scenarios

  • Taking a finished Chinese drama to an English audience
  • Repurposing a single product video across five markets
  • Rebuilding old library content that only exists with burned-in captions

Key features

Traceless Subtitle and Watermark Removal

Ghostcut uses OCR to find text on screen and then paints over it, rebuilding the area underneath so the removal blends into the footage. It handles hard subtitles, captions, lower thirds, watermarks, and logos, and supports Chinese, English, Japanese, Korean, and Arabic text. You can also mark specific regions to remove or protect before the engine runs, which helps when a watermark sits next to something you need to keep.

Dual ASR and OCR Subtitle Generation

The platform reads both audio and picture. Automatic speech recognition pulls lines from the spoken track while optical character recognition catches text already on screen, and the two are matched into a timed subtitle file. It supports 100+ languages according to JollyToday, with noise and interference reduction built in. Once the lines are extracted, you can edit them online and tag which speaker said what.

AI Translation with a Multi-Agent Review Step

Translation runs through large models such as DeepSeek, and JollyToday describes the flow as multi-agent, meaning one pass drafts and another checks. That setup targets the place where raw machine translation usually falls down: character names, tone, and dialogue that only makes sense in the original language. It's still worth a human read before you publish, because the review step smooths accuracy rather than guaranteeing it.

Voice Cloning and AI Dubbing

Ghostcut generates dubbing in dozens of languages, covering US, European, and Asian voices, and supports cloning a speaker's voice from a sample. The goal is a dub that keeps the original performance instead of a flat read. Quality varies by language, and less common languages are where you'll hear the seams most clearly.

Batch Project Management

Projects and assets are managed in one place, and the platform is built to upload and translate hundreds of videos at once. For teams working through a season or a campaign, that changes the job from a per-file task into a queue you set up once and monitor.

Editing and Rendering in One Place

After translation, you composite subtitles, audio, and music back onto the video with automatic alignment, then render the final file. Editing project files can be exported too, so you're not locked into the platform if you want to finish in another editor.

Subtitle Removal API

Beyond the web app, Ghostcut offers an API for subtitle removal, aimed at developers and teams who want to wire the capability into their own pipeline. That's the path for anyone pushing volume high enough that manual uploads stop making sense.

Pros and cons

Pros

  • One workspace for subtitle generation, removal, translation, and dubbing, so you skip juggling four separate subscriptions.
  • Batch processing is a first-class feature, which matters for anyone working through many episodes rather than a single clip.
  • Free tier and low per-minute pricing make it easy to test before committing.
  • Removal supports Chinese, English, Japanese, Korean, and Arabic text, a wider range than many competing tools.
  • An API covers the subtitle-removal step for teams that want to automate the pipeline.

Cons

  • Translation and dubbing still need a human review pass, so it won't fully replace an editor for name-heavy or culturally specific dialogue.
  • Removal quality depends on the source footage, and busy or moving backgrounds can leave visible traces.
  • Dubbing quality is stronger in major languages than in less common ones, which can be a real issue if your audience is small.
  • Pricing is quoted per minute in different currencies by region, so you need to check the plan that applies to you before budgeting.

Frequently asked questions

It's a video localization platform. You upload footage, and it generates subtitles, removes burned-in text and watermarks, translates the dialogue, and dubs it in a cloned voice. The point is to turn one finished video into versions that work in other languages, without a full re-edit.

Related content

Explore related tools, skills, and articles for Ghostcut.

Ghostcut Alternatives

Prosp

Prosp

Prosp · Writing · Voice & Language · Marketing

Prosp is an AI LinkedIn outreach tool built for agencies and sales teams, and it writes the message and the voice note in your own voice for each prospect so they actually reply. You connect your accounts, find leads, and let the AI draft and send personalized messages at scale, all from one inbox. It's built for people running outbound at volume. That's the whole pitch. Every touchpoint still has to feel human.

Paid / $30.99 - $79.99 per account/moView details
Wordly AI Translation

Wordly AI Translation

Wordly · Voice & Language · Productivity

Wordly AI Translation is a real-time AI translation and captioning platform built for meetings, conferences, and events. It delivers live translation, captions, transcripts, and summaries in more than 60 languages, and attendees join by scanning a QR code or opening a link instead of using dedicated headsets. The platform works with Zoom, Microsoft Teams, Google Meet, and Webex, and it's designed for organizations that want multilingual access without hiring human interpreters for every session. Simple as that.

Paid / $0 - $150/moView details
Musicful

Musicful

Musicful AI · Voice & Language · Video

Musicful is an AI music generator and AI music video maker that turns text to music in minutes. Give it a text prompt, a set of lyrics, or a hummed melody and it returns a finished track with vocals and instruments. It also doubles as an AI song generator, produces music videos from the songs you create, and offers an AI cover tool plus a developer API. The platform runs in a web browser and through an Android app, so you can start a song on desktop and pick it up on your phone.

Free / $0 - $20/moView details