Add your video
Paste a link or upload a file.
BeamTake transcribes English speech, translates it into Japanese subtitles and burns them into the video. The audio stays the same, so the speaker keeps their own voice; only the text on screen changes.
Paste a link or upload a file.
Speech is transcribed with word-level timing.
Subtitles are translated and their timings rebalanced for the target language.
Japanese subtitles are rendered into the video, and .srt files in both languages download too.
Speech is fast, with lots of contractions and connected forms (gonna, wanna). Names and brand names get spelled by ear; downloading the transcript and fixing them is the fastest route.
Spoken English uses connected forms: “going to” becomes “gonna”. The transcript writes what was said, not written English; you can tidy it up in the .srt file.
English captions open content up to the global feed. Most viewers whose first language isn't English still follow English captions easily, so it's usually the first extra language to add.
Translation follows the flow of speech, not word by word: a English sentence can come out longer or shorter in Japanese, so subtitle timings are rebalanced. Small kana and punctuation shouldn't start a new line; character-based breaking is essential.
Line breaks follow the target language: 14–18 characters per line. Both languages read in the same direction, so alignment needs no extra work.
Japanese viewers are used to on-screen text: even Japanese TV uses it heavily, so captioned clips feel natural.
| Source language | English (English) |
|---|---|
| Target language | Japanese (日本語) |
| Source direction | left to right |
| Target direction | left to right |
| Target line length | 14–18 characters per line. |
| Audio | unchanged, the speaker's own voice |
| Output | burned-in subtitles + .srt in both languages |
| Aspect ratios | 9 ratios, 1080 on the short side (2160 on Studio) |
We don't generate synthetic voices. The speaker's own voice and intonation stay, and viewers read the text. For interviews, lectures, podcasts and other speech-heavy content, that's more honest than imitating a voice, and causes fewer problems in practice.
No. The audio stays as it is and the translation goes out as subtitles, so the speaker's voice and tone are preserved.
Spoken English uses connected forms: “going to” becomes “gonna”. The transcript writes what was said, not written English; you can tidy it up in the .srt file.
Yes. Small kana and punctuation shouldn't start a new line; character-based breaking is essential.
Yes. Both .srt files download with the video; fix the translated one in any subtitle editor and upload it wherever you publish.
Burn word-by-word captions into the whole video, with .srt and .txt files.
OpenTranscriptTurn video or audio into text: clean paragraphs, .srt, .vtt and an optional summary with chapters.
OpenCrop and resizeTurn a horizontal video vertical or square, following the speaker. Framing only, no captions.
OpenTrim videoSet a start and end and keep that part. Change the aspect ratio too if you like.
OpenSplit videoSplit a long video into equal parts, each labeled 'Part 1/5'.
OpenAdd text and imagesAdd text and/or a logo on top of your video. Choose position, size and style.
OpenProcess 30 minutes of video free every month and see the results. Upgrade if you like it.
No credit card · Cancel anytime