talking-avatar
TTS speaks your line, and the lipsync model animates the portrait to match. upload -> text:speak -> text:movement -> tts -> lipsync; returns video (longcat-avatar-1.5) saved to disk (file path in result). Runs on NanoGPT — $0.61 deposit per call, paid in Nano (XNO) — settles at actual model cost + 20%, change returned; no account needed.
Input Schema
| Name | Required | Description | Default |
|---|---|---|---|
| Image | Yes | * required; image — file path or https URL | |
| speak | No | Text; default: "Hi there! I used to be a still photo — then somebody wired three nodes together, and now I won't stop talking."; optional | |
| movement | No | Text; default: "exaggerated head movement"; optional | |
| _payment_id | No | Payment id from this tool's previous payment-required response. Phase 3 only: after /x402/watch closes with status paid, call again with the same arguments plus this id to open the RESULTS stream. Do not pass it while payment is still pending — monitor the watch SSE first. |