Workflow overview
Images, video, and audio
Generate covers, images, and videos in a chat, edit talking-head videos, handle voiceovers and subtitles, and find the results.
Checked 2026-10-07. Applies to Beav desktop 2.8.13.
Beav has no separate media creation page. Images, video, and audio are all done in a New chat and its right workspace: describe the purpose, platform, aspect ratio, and reference material, and the AI chooses suitable tools. Media tasks usually take longer than text and may use credits or external services.
Images and covers
- Describe the image you want directly, or attach reference images (paste them, drag them in, or drag them from the asset library on the right).
- For Xiaohongshu (RedNote) covers and the cover of a complete image-and-text post, the AI considers 12 types of composition (single image, image with text, comparison, collage, and more) and chooses based on the content. This does not mean it generates 12 images each time.
- For a fixed visual style, use the creation templates below a new chat. You can also turn an image you like into a personal template.
- When making a complete Xiaohongshu image-and-text post, if you have not authorized generating all images at once, the AI first makes two for you to confirm by default, then continues with the remaining pages.
- For an existing image, you can ask the AI about specific content, text placement, or whether labels are covered. It inspects the relevant area without changing the original image.
Video
For a talking-head video, provide the video in the chat and state your editing goal. The AI creates an editable project, opens the editor on the right, and handles cleanup, subtitles, supplementary footage, and animation. An open editor only means the project is ready; play through the whole video before exporting. For full usage, see Video editing and smart talking-head cuts.

For an explainer video, title cards, subtitles, narration, and chapter visuals can all go on the same timeline. The project below already includes a title card, automatic subtitles, and a narration track. Before export, check the pacing, subtitle legibility, and audio-video sync section by section.

When you ask the AI to turn images, a voiceover, and subtitles into a video, it also uses the editor to create an editable project. Image push-ins, voiceover, and subtitles stay on separate tracks, so you can change text and timing later. To generate a video, use a video template or describe it directly; see Creation templates.
Reference video and speech timing
Provide an authorized local video or a video asset from the current workspace, and you can have the AI look at a specific moment or short interval to answer a concrete question, without creating an editing project just to view the reference. Sampled frames are saved in the current chat and are not automatically imported into the asset library. Still frames cannot replace checking the sound and continuous motion.
When an animation already has a voiceover and a word-level transcript, the AI can locate key phrases by sound and place subtitles, reveals, and visual changes at the matching points. After changing the language or the voiceover, recheck the affected timings.
Audio and subtitles
- Voiceover: tell the AI the script, language, and voice, or open the AI Voiceover mini tool, which lets you set language, voice, and pauses per segment and can also clone voices.
- Transcription: use Transcribe Media to choose a transcription model and export TXT, SRT, VTT, or JSON. To extract transcripts from Douyin, WeChat Channels, or TikTok links, use Extract Video Script.
- Waveform: you can ask the AI to look at the waveform of an audio file or a video's audio track to see pauses and volume changes, without creating an editing project first.
Cost and duplicate submissions
Generation uses credits or external services you have configured, and available parameters vary by model. A successful submission does not mean generation is finished. Do not resubmit the same task just because there is no result yet; check the original task's status first. When a result cannot be confirmed, Beav keeps the failure information and any outputs already produced, and does not generate again automatically. An external Agent that hits awaiting_approval must wait for you.
Find the results
If you switch workspaces, generation tasks keep running in the original workspace and their results are saved there. Open the original chat and go to Manuscripts → This chat on the right to see that chat's attachments, generated results, and asset packages. For images or videos you want to reuse long term, save them to the asset library, or explicitly ask the AI to save them.