avots-mcp
avots-mcp is a multi-provider AI platform MCP server giving you access to image generation, video creation, audio production, avatars, chat, and more through a single connection and balance.
πΌοΈ Image Generation & Editing: Generate images from text prompts using models like Nano Banana (Gemini), GPT-5 Image, FLUX, Recraft (native SVG/vector), and Ideogram (best-in-class typography). Supports multiple aspect ratios and styles.
π¬ Video Generation & Editing: Generate videos from text or images using Veo 3.1, Seedance 2.0, Kling v3.0 Pro, Sora 2 Pro, and Grok Imagine. Edit existing clips via text prompts, perform face swaps, and re-voice/lip-sync existing videos with new speech.
π§π€ Avatars & Talking Heads: Create and save reusable face identities, generate talking-avatar videos from a portrait + text, and produce vertical AI-vlog clips for Shorts/TikTok.
π΅ Music & Audio: Generate music across genres (ElevenLabs Music, ACE-Step, Stable Audio, MusicGen) and text-to-speech narration with preset or cloned voices, including multi-language support.
π¨ Studio Templates: Assemble photo montage slideshow reels (4β25 photos), generate vintage travel posters from a face photo, and put a face into viral trend templates.
π¬ Chat with 300+ Models: Send prompts to Claude, GPT-5, Gemini, DeepSeek, Sonar, and more β all billed through one balance.
π§ Utilities: Check token balance and pricing, list available models/avatars/trends, poll async job status, and schedule calendar events from natural language (Apple/Google Calendar).
Allows creating calendar events on a linked Apple Calendar from natural language descriptions.
Provides text-to-speech and music generation using ElevenLabs voices and music models.
Allows creating calendar events on a linked Google Calendar from natural language descriptions.
Provides a bot for billing, models, and web app inquiries, but no direct MCP tools.
avots-mcp
Official MCP (Model Context Protocol) server for avots.ai - a multi-provider AI platform.
One connection gives you:
πΌ Image generation & editing - Nano Banana (Gemini 3 Pro / 3.1 Flash), GPT-5 Image, FLUX, Recraft (incl. native vector SVG), Ideogram
π¬ Video generation & editing - Veo 3.1, Seedance 2.0, Kling v3.0 Pro, Sora 2 Pro, Grok Imagine, Gemini Omni Flash (async, 1-8 min); plus scene edit, face swap and lip-sync re-voicing of existing clips
π§βπ€ Talking heads & avatars - saved reusable face identities, talking-avatar videos, vertical AI-vlogs for Shorts / TikTok
π΅ Music & audio - ElevenLabs Music, ACE-Step, Stable Audio, TTS narration (incl. cloned voices)
π¨ Studio templates - photo-montage reels, vintage travel posters, viral trend recreation with your face
π¬ 300+ chat models - Claude (Sonnet / Opus), GPT-5, Gemini 3, DeepSeek, Sonar, and more - billed through one balance
The server lives at https://mcp.avots.ai/ and speaks the MCP 2025-06-18 spec over Streamable HTTP. Tools are billed per call against your avots balance (same balance you'd see on the web app or the Telegram bot).
Quick start
Sign up at avots.ai and mint an MCP key at Settings β Integrations (it looks like
av_mcp_<48hex>).Pick your client from the table below and follow the linked guide.
Try it - ask your client "generate an image of a fox in a snowy forest" and watch tokens get spent.
Client | Guide | Auth model |
Claude.ai web | OAuth - paste URL, click Connect, sign in (no token copy-paste) | |
Claude Desktop | Bearer token via | |
Claude Code (CLI) | Bearer token via | |
Cursor | Bearer token via | |
Cline | Bearer token via | |
Any other MCP client | docs/tools.md - endpoint + tool list | Bearer header |
Ready-to-paste mcp.json snippets live under examples/.
Related MCP server: Muapi
What's in the server
Nineteen tools, all documented in docs/tools.md:
Tool | Cost | What it does |
| free | Current tokens, subscription tier and the full pricing catalog. |
| free | All active models with per-call cost (filter by |
| free | The user's saved reusable face identities. |
| free | The avots Studio catalog of viral templates (face-swap videos, ads, animations). |
| free | Poll any async job by |
| ~10-1000 β‘ | Send a prompt to any chat model. Useful for delegating to GPT, DeepSeek, Sonar, etc. |
| ~200-500 β‘ | Sync image gen AND photo editing ( |
| ~200-5000 β‘ | Async video gen with two-step confirmation; supports i2v, multi-photo character refs and motion-reference clips. |
| ~500-2000 β‘ | Swap the face in an existing video with a photo (or saved avatar). Two-step. |
| ~450 β‘/sec | Scene edit of an existing clip by text prompt; person, motion and original audio preserved. Two-step. |
| ~300-600 β‘ | Re-voice an existing talking video: new text (TTS) or a ready audio track. Two-step. |
| free / ~200-500 β‘ | Save a reusable face identity from a photo (free) or generate one from a description. |
| ~600-2500 β‘ | A portrait speaks your exact text: TTS + lip-sync, quality/fast tiers. Two-step. |
| preview shows price | Vertical AI-influencer clip for Shorts / TikTok: the server writes the line from your topic. Two-step. |
| ~50-800 β‘ | Music (ElevenLabs Music, ACE-Step, Stable Audio) or TTS narration incl. cloned voices. Async. |
| ~200 β‘ | Slideshow reel from 4-25 photos: Ken Burns + crossfades + music. |
| ~200-500 β‘ | Face photo β vintage travel poster of any country. Synchronous. |
| per trend | Put the user's face into a viral template from |
| ~5 β‘ | Natural language β event in the linked Apple/Google calendar (or an .ics download). |
About the two-step video flow. Video is the most expensive tool. To avoid surprise spend,
generate_videoreturns a preview card the first time it's called (no submit, no reserve). The client (e.g. Claude) shows the cost + alternative models with prices and asks the user. The user confirms, the client re-calls with the chosenmodel+confirmed: true, and only then does the job get submitted. On submit error the server returns the same alternatives card - it never silently swaps to a pricier model.
What you can build
A few things this is actually useful for. Each one chains two or more tools through one connection and one balance.
Social ad creative in one prompt β hero image, animated variant, music bed.
Product-photo angle pack β one product shot in, four angles + a rotation clip out.
Storyboard to animatic β four script-driven frames animated into a 12-second rough cut.
Vertical Reels / Shorts factory β 9:16 clip + matching 15-second music bed, repeatable per video.
Podcast cover art + show notes β four cover variations and a written episode description for each.
Second-opinion delegation β forward a tricky problem to a model from a different lineage and compare.
Localized brand assets β translated copy and locale-tuned visuals across markets.
Each of these flows is written out as a runnable script β exact prompt, tool sequence, model picks, cost β in docs/recipes.md.
Cost ranges from ~200β‘ for a single image to ~5000β‘ for a 10-second 1080p Kling Pro clip β run list_models (free) at any time for live per-call prices.
Billing
All tool calls bill against your avots balance, just like the web app and Telegram bot. No separate metering. Daily USD cap (set in Settings) and per-key rate limits apply.
See pricing.avots.ai for token packs and subscriptions.
Troubleshooting
Common cross-client issues (401s, daily caps, two-step video flow, image rendering, npx PATH gotchas) are collected in docs/troubleshooting.md. For client-specific setup, see the per-client guide linked in the table above.
Issues & feedback
Open an issue here, or write to hello@avots.ai. For platform questions (billing, models, web app) the Telegram bot @AvotsAIbot is the fastest channel.
License
MIT - feel free to fork the docs, the examples, and anything else here.
Maintenance
Resources
Unclaimed servers have limited discoverability.
Looking for Admin?
If you are the server author, to access and configure the admin panel.
Related MCP Servers
- Alicense-qualityAmaintenanceMCP server for generating, editing, and processing images via multiple providers including Kilo, OpenRouter, OpenAI, and Gemini, with local tools for background removal, resizing, and cropping.Last updated542MIT
- Alicense-qualityDmaintenanceAccess 400+ generative AI models directly from your AI assistant β generate images (FLUX, Midjourney, GPT-4o), create videos (Veo3, Kling), make music (Suno), and enhance photos, all via a single MCP server.Last updated8MIT
- AlicenseAqualityBmaintenanceOne MCP server for music, image, video, and audio generation across Suno, Grok Imagine, Seedance, Kling, Hailuo, Wan, VEO, Ideogram, and GPT Image 2. Generate, edit, upscale, reframe, and master through one API key and one credit pool.Last updated161735MIT
- AlicenseAqualityCmaintenanceMulti-provider media generation MCP server that generates images, videos, audio, and transcriptions from text prompts using OpenAI, xAI, Gemini, ElevenLabs, and BFL through a single unified interface.Last updated6911MIT
Related MCP Connectors
One API key for 6 AI models. Pay-per-use. MCP protocol support with web search.
Run 100+ AI models β image, video, audio, 3D β through one API with pay-per-use billing.
MCP server for Hailuo (MiniMax) AI video generation
Latest Blog Posts
- Who's Calling? MCP Hosts Are an Identity Blind Spot (And the Spec Knows It)By Om-Shree-0709 on .mcpAgent IdentityOAuth 2.1
- Your AI Chatbot Just Exposed Your CEO's Salary to an InternBy Om-Shree-0709 on .Agent IdentityMCP SecurityOAuth Delegation
- Why MCP Servers Need Execution Sandboxing (And Why Your Current Stack Isn't Enough)By Om-Shree-0709 on .Agentic AiPrompt InjectionWebAssembly
MCP directory API
We provide all the information about MCP servers via our MCP API.
curl -X GET 'https://glama.ai/api/mcp/v1/servers/avotsai/avots-mcp'
If you have feedback or need assistance with the MCP directory API, please join our Discord server