A face introduces your AI influencer. A voice can explain an idea, welcome an audience or give a product demonstration a more personal rhythm. CamTalk is a coming-soon video feature being designed to bring those choices together in FanPro Studio.

This preview walks through the planned workflow in the supplied designs: choose a character image, prepare the speech and audio, then describe how the character should perform. It also explains the difference between using an influencer’s assigned voice and creating a freeform clip with another voice or uploaded audio.

What CamTalk is being designed to do

The designs place CamTalk inside Videos. Instead of putting everything into one prompt, they separate the character image, spoken words, voice and background sound from the scene direction. That gives each part of a talking video a clear role.

The planned creative inputs
InputWhat it contributes
Character imageThe visible person and the starting appearance.
Speech or uploaded audioThe words and delivery the viewer hears.
VoiceThe sound of the speaker, using the selected influencer’s voice or a freeform choice.
Background soundAudio that supports the setting.
PromptScene direction: expression, gestures and the camera.
Optional presetA starting point that can fill in a prompt and associated settings.

The useful starting point is a short idea with one purpose. An introduction, a product explanation or a single answer to a common question gives you a focused first clip to review.

Start with a clear character image

The planned Character Image field asks for the full face to be visible. Choose a reference where the expression, face and framing are easy to read. If your intended scene includes hand gestures, leave enough room in the composition for those actions.

CamTalk design preview: an influencer is selected and the Images Library shows that influencer’s generated images. The character image and optional preset have separate inputs.

The image picker includes My Generations and Uploads. The selected-influencer design filters the library to that influencer, making it easier to find a suitable reference without mixing identities. The designs also provide a freeform path without a selected influencer.

  1. Decide which character the clip is for.
  2. Choose a clear image with the full face visible.
  3. Check whether the framing supports your intended gesture and camera direction.
  4. Keep the outfit and setting consistent with the story you want the character to tell.

Understand the two planned voice paths

Influencer voice and freeform audio
PathPlanned behavior
An influencer is selectedTheir assigned voice is used automatically. The design asks you to deselect the influencer before choosing another voice or uploading speech audio.
No influencer is selectedThe freeform flow lets you choose a library voice for typed speech or supply your own audio.
Uploaded speech is activeRemove the uploaded audio before returning to a library voice and a typed script.

These choices answer different creative needs. An assigned voice can help a recurring influencer sound consistent from one clip to the next. The freeform path gives you room to choose a different voice for a separate concept or work from an audio recording you already have.

Voice Selection design preview: search, use-case filters, voice previews, language and speed controls. Names, catalog contents and settings shown here are design examples.

The planned voice library includes preview playback and filters for qualities such as style, tone, age and language. Listen for whether the voice fits the character and the message. An energetic opening and a calm explanation may need different delivery even when they share the same subject.

The Upload Audio designs list MP3, WAV and FLAC. Final file limits and supported options will be confirmed with the release. The preview does not establish a voice-cloning feature.

Write the speech and scene direction separately

The speech field is for the words the character should say. The Prompt field is for what the viewer should see: expression, movement and camera direction. Keeping them separate makes the brief easier to read and revise.

These are illustrative writing examples, not tested CamTalk prompts. Notice that the speech explains the subject while the direction describes a performance. A stage instruction such as “smile and turn toward the camera” belongs in the direction, rather than being mixed into words that are meant to be spoken.

  1. Write one clear opening sentence.
  2. Give the clip one main point before adding supporting detail.
  3. Read the script aloud and leave space for pauses.
  4. Describe a small number of gestures that support the words.
  5. Check that the reference image can support the planned action.

Add background sound or start from a preset

Background Sound is a separate part of Audio Setup in the designs. Think about what the setting needs: quiet room tone, an outdoor atmosphere or another appropriate background. Supporting audio should leave the speech easy to understand.

The optional CamTalk to Copy input opens CamTalk Presets. The planned library separates Presets from My Templates. Selecting a preset can fill the prompt and associated settings, which you can then review and edit for your own character and script.

CamTalk design with a preset applied: the prompt and audio choices sit together in the setup. The gallery beside it is part of the design and is not presented as a CamTalk result.

Treat a preset as a starting point. A useful template still needs to fit the image, words and intended audience. Check any filled direction for references to props, gestures or a setting that do not belong in your version.

Ideas to prepare before CamTalk launches

A focused first-clip plan
Use caseWhat to prepareWhat the clip should communicate
Influencer introductionA clear portrait and a short welcome.Who the character is and what the audience can expect.
Product explainerA suitable character image and accurate product details.One feature or detail in straightforward language.
Social video openingA concise hook and one follow-up point.Why the next few seconds are worth watching.
Guide introductionThe goal of the guide and its starting context.What the viewer will learn before the instructions begin.

Prepare the material you control now: the character reference, script, intended audience and any audio you have permission to use. A small, organized brief will be easier to adapt when the final controls are available.

The designs include duration, resolution, generation and result states. Their displayed numbers are mockup values, so this preview does not announce final prices, credit charges, clip-length limits or output resolutions.

What to check when the feature becomes available

Once you can generate a CamTalk clip, review the whole performance. A talking video needs the image, sound and timing to work together. A convincing first frame alone will not show whether the complete delivery fits the brief.

  • Check that the character remains recognizable throughout the clip.
  • Compare mouth movement and spoken audio for timing.
  • Listen for clear wording, appropriate pronunciation and natural pauses.
  • Watch whether expressions and gestures support the message.
  • Check that background sound leaves the voice understandable.
  • Review the first and last frames and the actual output format.

Change one part of the brief at a time when refining a result. If the delivery feels rushed, simplify the script before adding more actions. If the gesture distracts from the message, reduce the movement while keeping the words steady.

Explore the video workflows available today

CamTalk is coming soon. While its workflow is being developed, you can explore the existing Studio Marketing examples below to practice organizing character references, dialogue and scene direction.

Create a bag-review video in Studio Marketing

Plan a short reveal with one character and several product views.

Build a 30-second Porsche review

Connect dialogue, performance and references across five scenes.


FanPro Editorial

FanPro team