Connect Medux MCP to Codex
Add the Medux remote MCP endpoint to Codex, provide an authorized credential, and confirm that the medux_speech_create_audio tool is available.
Learn to generate speech from text with the Medux API through Codex or Claude MCP. Check inputs, tool calls, task status, and output.
POST/api/app/v1/generate/audio
Generate speech audio from text using a reusable voice asset.
Open the Create Audio API reference →The Medux operation and request fields are shared. Switch tabs for the client-specific prompt, approval pattern, screenshots, and walkthrough.
Codex-specific workflow
Use the active workspace as the source of truth. Ask Codex to identify the text, voice asset ID, language, and supported speech settings, select medux_speech_create_audio, and keep the Medux task ID and returned output with the rest of the project. Naming the input roles and the expected result prevents an agent from guessing which file should be used where.
In this workspace, use medux_speech_create_audio through Medux MCP to generate speech audio from text.
First inspect and identify the text, voice asset ID, language, and supported speech settings.
Show me the exact tool arguments before execution.
After the task completes, return the generated speech audio and verify that the audio speaks the complete text with the intended reusable voice.
Before approving the call, compare the selected tool, file or asset IDs, ordering, timing, and output settings with the request. If a local filename was mapped to an uploaded Medux file ID, keep that mapping visible so it can be audited later.
Let Codex monitor an asynchronous task until it reaches a terminal state, then save or report the generated speech audio. The final check is explicit: confirm that the audio speaks the complete text with the intended reusable voice.
Generate speech audio from text using a reusable voice asset.
Codex selects medux_speech_create_audio, presents the arguments for review, calls Medux, and reports the response.
Start with a clear instruction that names the Medux operation and the intended inputs.
Generate speech audio from this text using voice asset id {voice asset id} and model id speech-standard.Confirm these values before approving the MCP tool call. The linked API page remains the source of truth for current limits and response semantics.
| Field | Type | Requirement | Purpose |
|---|---|---|---|
links | array<string> | Optional | Optional reference links submitted from the Playground form. |
model | string | Optional | Generation model label or version selected in the Playground form. |
model_id | string | Optional | Stable public Medux model identifier. |
text | string | Optional | Provide the value described in the API reference. |
title | string | Optional | Provide the value described in the API reference. |
voice_asset_id | string | Required | Reusable voice asset used for text-to-speech generation. |
Add the Medux remote MCP endpoint to Codex, provide an authorized credential, and confirm that the medux_speech_create_audio tool is available.
Prepare the source files, asset IDs, text, or settings required by the operation. Upload local media first when the tool requests file IDs.
Use one direct instruction with the intended values. Codex should select the Medux tool and show the proposed arguments before execution.
Check that the tool is medux_speech_create_audio and that each file ID, task ID, option, and output setting matches your request before approving it.
If Medux returns an asynchronous task, let the agent monitor it until completion; otherwise review the immediate response and save any result identifiers or output URLs.
User goal
↓
Codex selects medux_speech_create_audio
↓
Review inputs and approve the tool call
↓
Medux runs Create Audio
↓
Codex reports the response and outputCodex can call the medux_speech_create_audio Medux MCP tool to generate speech audio from text using a reusable voice asset. This guide covers the required inputs, approval flow, result handling, and the matching API reference.
This workflow uses medux_speech_create_audio. Review its arguments before approval and consult the API reference for the current contract.
Claude-specific workflow
Keep the goal, source assets, and constraints together in the conversation. Ask Claude to restate the text, voice asset ID, language, and supported speech settings before proposing medux_speech_create_audio. That context checkpoint makes the file roles and desired outcome easy to correct before any Medux call is approved.
Using the assets and requirements in this conversation, help me generate speech audio from text with Medux MCP.
Restate which input fulfills each role: the text, voice asset ID, language, and supported speech settings.
Propose the medux_speech_create_audio call and summarize its arguments before asking for approval.
When it finishes, return the generated speech audio with a checklist confirming that the audio speaks the complete text with the intended reusable voice.
Have Claude summarize the intended transformation, identify every source asset by role, and list the important constraints. Review that summary together with the proposed tool arguments; correct the conversation first if a file, order, time range, or setting is ambiguous.
Keep the Medux task ID in the conversation while Claude checks progress. When processing ends, ask for the generated speech audio plus a concise validation checklist confirming that the audio speaks the complete text with the intended reusable voice.
Generate speech audio from text using a reusable voice asset.
Claude selects medux_speech_create_audio, presents the arguments for review, calls Medux, and reports the response.
Start with a clear instruction that names the Medux operation and the intended inputs.
Generate speech audio from this text using voice asset id {voice asset id} and model id speech-standard.Confirm these values before approving the MCP tool call. The linked API page remains the source of truth for current limits and response semantics.
| Field | Type | Requirement | Purpose |
|---|---|---|---|
links | array<string> | Optional | Optional reference links submitted from the Playground form. |
model | string | Optional | Generation model label or version selected in the Playground form. |
model_id | string | Optional | Stable public Medux model identifier. |
text | string | Optional | Provide the value described in the API reference. |
title | string | Optional | Provide the value described in the API reference. |
voice_asset_id | string | Required | Reusable voice asset used for text-to-speech generation. |
Add the Medux remote MCP endpoint to Claude, provide an authorized credential, and confirm that the medux_speech_create_audio tool is available.
Prepare the source files, asset IDs, text, or settings required by the operation. Upload local media first when the tool requests file IDs.
Use one direct instruction with the intended values. Claude should select the Medux tool and show the proposed arguments before execution.
Check that the tool is medux_speech_create_audio and that each file ID, task ID, option, and output setting matches your request before approving it.
If Medux returns an asynchronous task, let the agent monitor it until completion; otherwise review the immediate response and save any result identifiers or output URLs.
User goal
↓
Claude selects medux_speech_create_audio
↓
Review inputs and approve the tool call
↓
Medux runs Create Audio
↓
Claude reports the response and outputClaude can call the medux_speech_create_audio Medux MCP tool to generate speech audio from text using a reusable voice asset. This guide covers the required inputs, approval flow, result handling, and the matching API reference.
This workflow uses medux_speech_create_audio. Review its arguments before approval and consult the API reference for the current contract.