Ideogram MCP Server

📦 Обзор проекта
Инструмент TypeScript, позволяющий использовать Ideogram API (v3.0) через сервер MCP.
Многофункциональность, включая генерацию изображений, ссылку на стиль, волшебную подсказку, соотношение сторон, выбор модели и т. д.
Можно использовать немедленно с Claude Desktop и другими клиентами MCP.
Related MCP server: OpenAI MCP
⚡️ Быстрый старт
Если вы хотите подключиться к Claude Desktop или другим клиентам MCP с молниеносной скоростью,
Просто скопируйте и вставьте приведенный ниже фрагмент JSON в свой файл конфигурации! ✨
{
"mcpServers": {
"ideogram": {
"command": "npx",
"args": [
"@sunwood-ai-labs/ideagram-mcp-server"
],
"env": {
"IDEOGRAM_API_KEY": "your_api_key_here"
}
}
}
}🛠️ Характеристики инструмента MCP
сгенерировать_изображение
Список параметров (последняя версия)
Параметры | Тип | объяснение | Обязательно/Необязательно | замечания |
быстрый | нить | Запрос на создание изображения (рекомендуется английский) | Необходимый | |
соотношение сторон | нить | Соотношение сторон (например, «1x1», «16x9», «4x3» и т. д.) | любой | 15 видов |
разрешение | нить | Разрешение (см. официальную документацию, всего 69 типов) | любой | |
семя | целое число | Случайное числовое значение (для обеспечения воспроизводимости) | любой | 0 до 2147483647 |
magic_prompt | нить | Волшебная подсказка ("АВТО" | "НА" | "ВЫКЛЮЧЕННЫЙ" |
скорость_рендеринга | нить | Скорость рендеринга для v3 ("TURBO" | "ПО УМОЛЧАНИЮ" | "КАЧЕСТВО" |
коды_стилей | нить[] | Последовательность кода в стиле 8 символов | любой | |
тип_стиля | нить | Тип стиля ("АВТО" | "ОБЩИЙ" | "РЕАЛИСТИЧЕСКИЙ" |
отрицательный_запрос | нить | Исключения (рекомендуется английский) | любой | |
num_images | число | Количество сгенерированных изображений (от 1 до 8) | любой | |
style_reference | объект | Справочник стилей (Новое в Ideogram 3.0) | любой | Подробности ниже |
└ URL-адреса | нить[] | Массив URL-адресов справочных изображений (до 3) | любой | |
└ код_стиля | нить | Код стиля | любой | |
└ случайный_стиль | булев | Использовать случайный стиль | любой | |
выходной_каталог | нить | Каталог хранения изображений (по умолчанию: «docs») | любой | |
базовое_имя_файла | нить | Основа для имени сохраненного файла (по умолчанию: «ideogram-image») | любой | Присвоение метки времени и идентификатора |
размытие_маски | булев | Размыть края изображения (установить значение true для наложения масок) | любой | По умолчанию: ложно |
📝 Пример использования
const result = await use_mcp_tool({
server_name: "ideagram-mcp-server",
tool_name: "generate_image",
arguments: {
prompt: "A beautiful sunset over mountains",
aspect_ratio: "16x9",
rendering_speed: "QUALITY",
num_images: 2,
style_reference: {
urls: [
"https://example.com/ref1.jpg",
"https://example.com/ref2.jpg"
],
random_style: false
},
blur_mask: true
}
});🧑💻 Разработка, сборка и тестирование
npm run build... сборка TypeScriptnpm run watch... режим разработки (автоматическая сборка)npm run lint... Анализ кодаnpm test... запустить тесты
🗂️ Структура каталога
ideagram-mcp-server/
├── assets/
├── docs/
│ └── ideogram-image_2025-05-18T06-31-45-777Z.png
├── src/
│ ├── tools/
│ ├── types/
│ ├── utils/
│ ├── ideogram-client.ts
│ ├── index.ts
│ ├── server.ts
│ └── test.ts
├── .env.example
├── package.json
├── tsconfig.json
├── README.md
└── ...(省略)📝 Вклады
Форк этого репозитория
Создайте новую ветку (
git checkout -b feature/awesome)Внесение изменений (сообщения о внесении изменений должны быть на японском языке, рекомендуется использовать эмодзи!)
Создание push- и pull-запросов
🚀 Развертывание и выпуск
Автоматическая публикация npm с помощью GitHub Actions
Обновление версии → Автоматическое развертывание путем отправки тегов
npm version patch|minor|major
git push --follow-tagsПодробности смотрите в docs/npm-deploy.md !
📄 Лицензия
Массачусетский технологический институт
Available Tools
1 toolgenerate_imageC
Generate an image using Ideogram AI
| Name | Required | Description | Default |
|---|---|---|---|
| prompt | Yes | The prompt to use for generating the image (must be in English) | |
| aspect_ratio | No | The aspect ratio for the generated image (see official docs for all 15 values) | |
| resolution | No | The resolution for the generated image (see official docs for all 69 values) | |
| seed | No | Random seed. Set for reproducible generation. | |
| magic_prompt | No | Whether to use magic prompt | |
| rendering_speed | No | Rendering speed for v3 (TURBO/DEFAULT/QUALITY) | |
| style_codes | No | Array of 8-char style codes | |
| style_type | No | The style type for generation | |
| style_reference_images | No | A set of images to use as style references (max 10MB, JPEG/PNG/WebP) | |
| negative_prompt | No | Description of what to exclude from the image (must be in English) | |
| num_images | No | Number of images to generate (1-8) | |
| style_reference | No | Style reference options for Ideogram 3.0 | |
| output_dir | No | Directory to save generated images (default: 'docs'). | |
| base_filename | No | Base filename for saved images (default: 'ideogram-image'). Timestamp and image ID will be appended automatically. | |
| blur_mask | No | Apply a blurred mask to the image edges (using a fixed mask image). If true, the output image will have blurred/feathered edges. (default: false) |
TDQS
Does the description disclose side effects, auth requirements, rate limits, or destructive behavior?
No annotations are provided, so the description carries full burden for behavioral disclosure. The description only states the basic action without mentioning rate limits, authentication needs, output format, error conditions, or cost implications. For a complex image generation tool with 15 parameters, this leaves significant behavioral gaps.
Agents need to know what a tool does to the world before calling it. Descriptions should go beyond structured annotations to explain consequences.
Is the description appropriately sized, front-loaded, and free of redundancy?
The description is a single, efficient sentence that states the core purpose without unnecessary elaboration. It's appropriately sized for a tool name that clearly indicates its function, and there's no wasted verbiage or structural issues.
Shorter descriptions cost fewer tokens and are easier for agents to parse. Every sentence should earn its place.
Given the tool's complexity, does the description cover enough for an agent to succeed on first attempt?
Given the tool's complexity (15 parameters, no output schema, no annotations), the description is inadequate. It doesn't explain what the tool returns, error handling, performance characteristics, or typical use patterns. For an image generation tool with many configuration options, more context is needed to help the agent use it effectively.
Complex tools with many parameters or behaviors need more documentation. Simple tools need less. This dimension scales expectations accordingly.
Does the description clarify parameter syntax, constraints, interactions, or defaults beyond what the schema provides?
Schema description coverage is 100%, so the schema already documents all parameters thoroughly. The description adds no parameter information beyond what's in the schema. According to scoring rules, when schema coverage is high (>80%), the baseline is 3 even with no param info in the description.
Input schemas describe structure but not intent. Descriptions should explain non-obvious parameter relationships and valid value ranges.
Does the description clearly state what the tool does and how it differs from similar tools?
The description 'Generate an image using Ideogram AI' states the basic action (generate) and resource (image) but lacks specificity. It doesn't mention what kind of images, quality levels, or typical use cases. Without sibling tools, differentiation isn't needed, but the purpose remains vague beyond the basic verb-noun pairing.
Agents choose between tools based on descriptions. A clear purpose with a specific verb and resource helps agents select the right tool.
Does the description explain when to use this tool, when not to, or what alternatives exist?
No guidance is provided on when to use this tool versus alternatives. The description doesn't mention prerequisites, ideal scenarios, or limitations. Without sibling tools, there's no need for differentiation, but the absence of any usage context leaves the agent with no guidance on appropriate application.
Agents often have multiple tools that could apply. Explicit usage guidance like "use X instead of Y when Z" prevents misuse.
Tool Schema Changelog
Recent tool additions, removals, and schema changes observed during successful MCP inspections.
1 tool update
v1.0.0- First observed
generate_image
TDQS
Scored across 1 tool
With only one tool, there is no possibility of confusion or overlap between tools. The single tool 'generate_image' has a clearly distinct and unambiguous purpose.
A single tool inherently has perfect naming consistency, as there are no other tools to compare it against. The name 'generate_image' follows a clear verb_noun pattern.
One tool is too few for a server named 'Ideogram MCP Server', which suggests a broader scope for image generation or AI tasks. A single tool feels thin and limited for such a domain.
The tool surface is severely incomplete for an image generation server. It only offers generation with no options for editing, listing, deleting, or managing images, creating significant gaps in functionality.
Maintenance
Related MCP Connectors
MCP server for Qwen Image 3 AI image generation
Multi-model AI image and video generator. 14 models behind one OAuth-secured MCP endpoint.
Generate AI images and videos from any compatible MCP client.
MCP server for Flux AI image generation
Related MCP Servers
- AlicenseNot gradedqualityFmaintenanceA server that provides AI-powered image generation, modification, and processing capabilities through the Model Context Protocol, leveraging Google Gemini models and other image services.18MIT
- AlicenseNot gradedqualityDmaintenanceA Model Context Protocol server enabling AI assistants to generate images through OpenAI's DALL-E API with full support for all available options and fine-grained control.15 npm1MIT
- AlicenseDqualityCmaintenanceA Model Context Protocol server that provides image generation capabilities using Google's Gemini 2 API, allowing users to generate multiple images with customizable parameters like prompts, aspect ratios, and person generation settings.1110 npm5MIT
- AlicenseDqualityFmaintenanceA Model Context Protocol server that enables generating and editing images using OpenAI's gpt-image-1 model, allowing AI assistants to create and modify images from text prompts.270 npm18MIT