Enables AI image and video generation using Google Nano Banana and Veo 3.1 via a LiteLLM gateway, providing tools for synchronous image generation and asynchronous video generation with polling, returning public URLs.
Enables image and video generation across GPT-Image, Gemini, Grok, and Jimeng with file-based outputs, multi-reference support, and model capability lookup.
Enables generating images, video, and audio through a single capability-oriented interface, with server-side routing, safety screening, job lifecycle management, concurrency limits, and retention.
Bridges Hermes Agent to the Hermes Intelligence Platform API, enabling tools to read/write learning loop data (context, feedback, signals, memory, etc.) via stdio.
MCP server for generating and editing images using OpenAI, and creating videos using OpenAI Sora and Google Veo. Enables fetching media from URLs or disk with smart output placement.
Provides tools for agent-driven video creation, including image generation via Google's GenAI, video generation, and local video stitching with FFmpeg.
Enables generating images from text prompts using various AI models (FLUX, Recraft, GPT Image, etc.) and automatically delivering them through Cloudinary's platform.
Enables text-only coding models to read images, PDFs, presentations, spreadsheets, and other non-text files through a single analyze_media tool, combining local document extraction, OCR, and optional vision models with clear evidence labeling.
MCP server that enables AI assistants to generate images and videos on demand using free-tier providers like Agnes AI, Cloudflare Workers AI, Hugging Face, and Google Gemini, with tools for creating media, polling async jobs, and listing providers.
Enables agents to query historical call-center call reason notes and tags on demand, detect statistically significant emerging problem spikes against baselines, and evaluate automated categorizers against ground truth.