First, confirm the three formats of subtitles during the project initiation phase.

Subtitles for cross-border product videos are not as simple as adding a line of English in post-production. Subtitles may be burned into the video, delivered as separate subtitle files, or not appear on screen at all. The focus of a textless version is to keep the footage clean from shooting to color grading. Brands must decide which subtitle format to use for each platform during the project initiation phase, as this decision directly affects shot composition during filming and the post-production workflow.

Product video shooting footage from the case study, observing the relationship between the shot, subject, and lighting.
Frame capture from the case study, sourced from the research material "RSP's trees, bees and birds in Mysterious Island". This footage is used solely to observe shots and production methods and does not represent an ONCE client project. Source page Case study page。

Brands are advised to clearly answer three questions at the project kickoff meeting. First, does the video need to retain foreign text on product labels or packaging within the frame? Second, are the English subtitles for native speakers or non-native buyers, which determines the vocabulary difficulty and display duration? Third, should the textless version retain graphic logos or watermarks in the frame, which involves copyright and brand exposure scope? The production team must record these three answers in the project memo as a baseline for all subsequent stages.

If the brand is temporarily unable to determine platform requirements, the production team should proactively provide a platform subtitle specification comparison chart, reminding the client to refer to the latest official requirements prior to publishing. Do not assume that all platforms support the same subtitle file format, nor assume that a textless version can be created by removing text from a texted version with one click in post-production. Burned-in subtitles and independent subtitle tracks follow completely different paths in post-production; a single wrong choice requires redoing the entire export process.

Reserve safe areas for subtitles and plan the visual center of gravity during the storyboard phase.

The visual composition of product videos must reserve a safe area for subtitles. English subtitles typically appear in the lower third of the frame; if key product information or selling point text is placed in this area during shooting, adding subtitles in post-production will obscure the main subject. Storyboard artists must mark the subtitle safe area for each shot and ensure that the main product, functional demonstration actions, and key explanatory text avoid this area.

For textless versions, the storyboard phase requires an extra check for any temporary text within the frame. For example, small text on product labels, slogans on backdrops, or printed content on props may seem insignificant on set, but when a clean frame is required for the textless version in post-production, these elements become irremovable flaws. The production team should color-code all in-frame text elements in the storyboard and confirm item by item which can be avoided and which must be retained.

If the product itself needs to display packaging text or label information, it is recommended to shoot a separate set of close-up shots for the texted version, while simultaneously shooting a set of textless product B-roll for the textless version. This allows editors to freely switch footage according to version requirements during post-production, without needing to both retain text and demand a clean frame in the same shot. This practice adds a small amount of shooting time but avoids the embarrassment of unmanageable post-production issues.

Record subtitle-related assets on set and control frame dynamics.

During the shooting execution phase, one person on set must be specifically designated to record subtitle assets. This person should shoot all clean frames that might be used as subtitle backgrounds, including static product states, scene B-roll, and shallow depth-of-field shots, as these assets can serve as subtitle backdrops when producing English subtitles in post-production, preventing subtitles from being placed directly over cluttered frames. When recording, pay attention to the shot size and camera movement of each shot to facilitate post-production matching.

Frame dynamics greatly affect subtitle readability. If the camera movement is too fast or the subject moves frequently, subtitles placed in a fixed position will appear jarring. When shooting product close-ups and functional demonstrations, cinematographers should control the camera movement speed to leave a stable visual area for subtitles. For shots requiring rapid transitions, it is recommended to shoot a few extra seconds of the starting and ending frames, allowing subtitles to be added when the shot is static in post-production and avoiding subtitles bouncing along with the moving frame.

The audio recording phase must also serve the subtitles. If the video contains product usage sounds or ambient noise, these sounds will occupy the audience's attention, and the subtitle display duration must be adjusted accordingly. The on-set sound mixer should record the content and duration of each sound segment to facilitate post-production decisions on when subtitles should appear and disappear. If the video is completely silent or only has background music, the subtitle rhythm can be freely controlled by the editor, and in this case, the team must confirm with the brand in advance whether a music-only version is acceptable.

In post-production, create the textless version first, then the texted version, and manage them in layers.

The correct sequence for post-production editing is to first complete the textless version, and then add subtitles based on the textless version to generate the subtitled version. The benefit of this approach is that the textless version can serve as a master for all platforms, while the subtitled version is merely a derivative output of the master. If the subtitled version is created first and the subtitles are deleted later, it is easy to accidentally damage the visual content or affect the timeline structure during the deletion process.

A clear layer structure must be established in the editing software. The video layer, color grading layer, subtitle layer, and graphics layer should be managed separately, with each layer having a clear naming convention. Within the subtitle layer, English subtitles and potential subtitles in other languages must be distinguished to facilitate future expansion. Color grading should be completed on the textless version, as color grading affects the brightness and contrast of the image, which in turn affects the clarity of the subtitles. After color grading is complete, the subtitled version directly applies the same color grading results without requiring separate adjustments.

The translation and proofreading of English subtitles must be independent of the editing process. The brand should provide accurate English expressions for the product's selling points, rather than having the editor translate them independently. The translated content must undergo at least two rounds of proofreading: the first round checks grammar and wording, and the second round checks the time correspondence between the subtitles and the visuals. The display duration of the subtitles should follow standard reading speeds, but specific values should be based on the actual content, without applying a fixed formula. Every subtitle must be played back on the textless version for verification to ensure it does not obscure key visuals.

The export phase generates master, subtitled, and textless files respectively.

The export phase must generate three files for different purposes. The master is the highest quality textless version, containing complete video and audio information, used for archiving and future remastering. The subtitled version is the delivery file with English subtitles, used for direct upload to platforms that require subtitles. The textless version is a clean video file, used for platforms to automatically generate subtitles or for subsequently adding subtitles in other languages. The resolution and encoding format of the three files can be the same, but the presence or absence of the subtitle track must be strictly distinguished.

For the subtitled version, it must also be confirmed whether the subtitles are burned into the video or exist as an independent subtitle track. Burned-in subtitles are suitable for short video platforms, as the platforms may not support uploading external subtitle files. Independent subtitle tracks are suitable for video websites and streaming platforms, where users can toggle subtitles on and off themselves. If the brand targets multiple platforms simultaneously, it is recommended to export both formats and clearly label them in the file names. Independent subtitle tracks must use standard formats, and the timeline must be checked to ensure it is perfectly synchronized with the video.

When exporting the textless version, pay attention to whether the audio tracks are complete. Some production teams accidentally delete sound effect or background music tracks when removing subtitles, resulting in missing audio in the textless version. Before exporting, check the integrity of the video tracks, audio tracks, and subtitle tracks item by item, and record the encoding parameters and file size of each file. Upon delivery, attach a technical specification document detailing the purpose, applicable platforms, and re-encoding recommendations for each file.

The acceptance checklist verifies the visuals, subtitles, audio, and file specifications item by item.

When the brand conducts acceptance, they should not just look at the overall effect, but verify it by item. The first check is visual integrity; no residual subtitles or text traces can appear in the textless version, and subtitles in the subtitled version must not obscure the main product and key information. The second check is subtitle accuracy, comparing the English translation with the actual product functions line by line, paying special attention to the accuracy of professional terminology and units of measurement.

The third check is audio consistency. The audio content of the subtitled version and the textless version must be completely identical, including background music, sound effects, and voiceovers. If the audio in the subtitled version is adjusted to match the rhythm of the subtitles, the same adjustment must be made to the textless version; otherwise, switching between the two versions will create a disjointed experience. The fourth check is file naming and format. All delivery files must be named according to agreed-upon rules, including version numbers, dates, and purpose descriptions, to avoid confusion.

During the acceptance process, each file must be played back in full, rather than just viewing screenshots or preview windows. During playback, pay attention to whether the timing of subtitle appearance and disappearance corresponds to the spoken content, and check if the textless version exhibits jitter or tearing during fast-moving scenes. If issues are found, record the specific timecode and problem description, and provide feedback to the production team for revision. After revisions, the entire file must be re-verified; do not just check the modified segments, as other parts may develop new issues due to re-rendering.

Explanation of Inapplicable Conditions and Alternative Solutions

Producing both English subtitled and textless versions simultaneously is not suitable for all cross-border product videos. If the video content primarily consists of character dialogue, subtitles are an essential tool for understanding the content, and the textless version has almost no use cases; in this case, only the subtitled version can be produced. If the video is a pure product showcase without voiceover, the textless version may be more suitable as the master version, making the subtitled version seem redundant.

For videos requiring frequent subtitle language updates, such as those targeting multiple non-English markets, it is recommended to produce only a textless master version and then add subtitles individually according to different platform requirements. This avoids re-rendering the entire video every time a language is updated. For AIGC-generated video content, the production process for subtitled and textless versions differs from traditional shooting, requiring separate confirmation of whether the generation tool supports layered output.

If the video contains a large amount of dynamic text or effect subtitles, such as pop-up product selling points or price tag animations, these elements belong to the visual content rather than subtitles and cannot be simply deleted in the textless version. In this case, the brand must re-evaluate the definition of the textless version: whether it means absolutely no text or just no dialogue subtitles. The production team must clarify this boundary during the project initiation phase to avoid later misunderstandings.

The next step is recommended to start with the asset list and platform confirmation.

When preparing to launch a cross-border product video project, it is recommended that the brand first compile a complete asset list, including physical products, packaging, usage scenarios, competitor reference videos, and screenshots of platform requirements. Simultaneously, confirm the latest subtitle support formats and file specifications with the target platforms, and do not rely on outdated experience. After the production team receives these materials, they can more accurately evaluate the project scope and develop a storyboard plan.

Before formal shooting, first produce a test shot that includes the subtitle safe area and textless version composition to verify if the framing is reasonable using actual footage. This test has a very low cost but can prevent the discovery of composition issues after large-scale shooting. The test shot is also used to confirm the visual effects of the subtitle font, size, and position, ensuring the final video has good readability on both mobile and desktop devices.

If you are preparing a product video shooting project, you can first organize the brief, reference footage, product or company materials, delivery platforms, and copyright scope, and then view thee-commerce product video service page, bringing communication from abstract preferences to actionable production boundaries.