Claude MCP tutorial · generate

How to create a text-driven avatar video with Claude MCP and Medux

Use Claude and Medux through MCP to generate an avatar video from text with reusable avatar and voice assets. This guide keeps file roles and constraints in context, reviews the proposed call, and validates the returned result.

POST/api/app/v1/generate/video/text_driven

Generate an avatar video from text using a reusable avatar and voice.

Open the Create Text Driven Video API reference →

Claude-specific workflow

Run Create a Text-Driven Avatar Video as a Claude MCP workflow

Keep the goal, source assets, and constraints together in the conversation. Ask Claude to restate the script, avatar ID, voice ID, and supported video settings before proposing medux_video_create_text_driven. That context checkpoint makes the file roles and desired outcome easy to correct before any Medux call is approved.

Claude prompt pattern

Using the assets and requirements in this conversation, help me create a text-driven avatar video with Medux MCP.
Restate which input fulfills each role: the script, avatar ID, voice ID, and supported video settings.
Propose the medux_video_create_text_driven call and summarize its arguments before asking for approval.
When it finishes, return the generated talking-avatar video with a checklist confirming that the avatar, voice, and spoken script match the request and the full video plays correctly.

Context checkpoint

Have Claude summarize the intended transformation, identify every source asset by role, and list the important constraints. Review that summary together with the proposed tool arguments; correct the conversation first if a file, order, time range, or setting is ambiguous.

Result verification

Keep the Medux task ID in the conversation while Claude checks progress. When processing ends, ask for the generated talking-avatar video plus a concise validation checklist confirming that the avatar, voice, and spoken script match the request and the full video plays correctly.

What this tutorial does

Medux operation

Generate an avatar video from text using a reusable avatar and voice.

Agent workflow

Claude selects medux_video_create_text_driven, presents the arguments for review, calls Medux, and reports the response.

Before you start

Authentication

  • Authentication is required via Authorization header.

Preconditions

  • Prepare avatar and voice assets before calling this endpoint.
  • Requires avatar asset id, voice asset id, and non-empty text.
  • Avatar and voice IDs can come from your Media Assets or official public assets.
  • Also accepts optional duration seconds, links, clarity, aspect ratio, and model.

Prompt Claude

Start with a clear instruction that names the Medux operation and the intended inputs.

Create a 7 second 1080P 16:9 talking video from this script using avatar asset id {avatar asset id}, voice asset id {voice asset id}, and model id talking-video-text-standard.

Request fields to review

Confirm these values before approving the MCP tool call. The linked API page remains the source of truth for current limits and response semantics.

FieldTypeRequirementPurpose
aspect_ratiostringOptionalRequested output aspect ratio, such as 16:9 or 9:16.
avatar_asset_idstringRequiredReusable avatar asset ID owned by the current user or listed as an official public asset.
claritystringOptionalRequested output resolution label, such as 1080P or 2K.
duration_secondsint32OptionalRequested output duration, in seconds.
linksarray<string>OptionalOptional reference links submitted from the Playground form.
modelstringOptionalGeneration model label or version selected in the Playground form.
model_idstringOptionalStable public Medux model identifier.
textstringOptionalProvide the value described in the API reference.
titlestringOptionalProvide the value described in the API reference.
voice_asset_idstringRequiredReusable voice asset used for text-driven lip sync.

Step-by-step workflow

01

Connect Medux MCP to Claude

Add the Medux remote MCP endpoint to Claude, provide an authorized credential, and confirm that the medux_video_create_text_driven tool is available.

02

Prepare the required inputs

Prepare the source files, asset IDs, text, or settings required by the operation. Upload local media first when the tool requests file IDs.

03

Ask Claude to run Create Text Driven Video

Use one direct instruction with the intended values. Claude should select the Medux tool and show the proposed arguments before execution.

04

Review and approve the tool call

Check that the tool is medux_video_create_text_driven and that each file ID, task ID, option, and output setting matches your request before approving it.

05

Check the Medux response

If Medux returns an asynchronous task, let the agent monitor it until completion; otherwise review the immediate response and save any result identifiers or output URLs.

Workflow recap

User goal
  ↓
Claude selects medux_video_create_text_driven
  ↓
Review inputs and approve the tool call
  ↓
Medux runs Create Text Driven Video
  ↓
Claude reports the response and output

Frequently asked questions

Can Claude use the Medux Create Text Driven Video API through MCP?

Claude can call the medux_video_create_text_driven Medux MCP tool to generate an avatar video from text using a reusable avatar and voice. This guide covers the required inputs, approval flow, result handling, and the matching API reference.

Which MCP tool does this tutorial use?

This workflow uses medux_video_create_text_driven. Review its arguments before approval and consult the API reference for the current contract.

Quick answer

How do I run Create a Text-Driven Avatar Video with Claude MCP?

Claude can use medux_video_create_text_driven through Medux MCP to generate an avatar video from text with reusable avatar and voice assets. The workflow keeps input roles and constraints in context, reviews the proposed call, and checks the returned result.

What does this Medux tutorial cover?

Claude can use medux_video_create_text_driven through Medux MCP to generate an avatar video from text with reusable avatar and voice assets. The workflow keeps input roles and constraints in context, reviews the proposed call, and checks the returned result.

Related topics: Create a Text-Driven Avatar Video · Create a Text-Driven Avatar Video with Claude MCP · Medux MCP tutorial