talking-avatar
Prototype a short spoken introduction with a fictional presenter. Muse makes the portrait, MiniMax Speech reads the script, and LongCat animates speech at 480p. Keep scripts under 30 seconds. Cost … text:Presenter look -> text:Spoken script -> text:Delivery -> image:face -> tts -> lipsync; returns video (longcat-avatar-1.5) saved to disk (file path in result). Runs on NanoGPT — $0.53 deposit per call, paid in Nano (XNO) — settles at actual model cost + 20%, change returned; no account needed. Example: Video sample (https://nanoodle.com/examples/gallery/#talking-avatar). A fictional presenter reads a short workshop introduction. Sampled frames preserve the presenter; check precise lip-sync timing in playback. Generation can take several minutes.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| Delivery | No | Text; default: "Natural calm delivery, subtle blinks and small head movements. Keep the face and mouth visible, eyes toward the camer..."; optional | |
| _payment_id | No | Payment id from this tool's previous payment-required response. Phase 3 only: after /x402/watch closes with status paid, call again with the same arguments plus this id to open the RESULTS stream. Do not pass it while payment is still pending — monitor the watch SSE first. | |
| Spoken_script | No | Text; default: "Welcome to the workshop. Start with the blue tray on your table. Inside you will find paper, a pencil, and the parts ..."; optional | |
| Presenter_look | No | Text; default: "A fictional friendly male museum guide in his thirties, wearing a plain navy shirt, shoulders-up front-facing portrai..."; optional |