In a video editing method, an editing interface includes visualized representations of a video track and a dubbing subtitle track. In the method, recording of a voice clip is started when a dubbing start operation is triggered in association with a first time point. In the method, the recording of the voice clip is stopped when a dubbing end operation is triggered in association with a second time point. In the method, the recorded voice clip is loaded into the dubbing subtitle track from the first time point to the second time point of the time progress of the dubbing subtitle track. In the method, a voice recognition text recognized from the voice clip is displayed in the editing interface. In the method, a fused video is generated based on a video editing completion operation, the fused video including the video, the voice clip, and the voice recognition text.
Full Text
What is claimed is: