Loometo
Allows auto-copying generated assets to a Cloudflare R2 bucket for permanent storage.
Allows text-to-speech and voice cloning via ElevenLabs API.
Allows pulling frames from Figma onto the canvas for image generation pipelines.
Allows using Google AI models (Imagen 4, Veo 3, Gemini) directly for image and video generation.
Allows using OpenAI models (GPT-Image / DALL·E, LLM nodes) for image generation and chat.
Click on "Install Server".
Wait a few minutes for the server to deploy. Once ready, it will show a "Started" state.
In the chat, type
@followed by the MCP server name and your instructions, e.g., "@Loometobuild me a UGC pipeline"
That's it! The server will respond to your query, and you can continue using it as needed.
Here is a step-by-step guide with screenshots.
An open-source, node-based AI creative studio. Bring your own keys.
Wire image, video, 3D, and audio models into visual pipelines on a canvas — then let your AI coding agent build those pipelines for you.
Made by Dinimiciuil Labs.
A real pipeline: reference boards + Director Formula feeding image → video, with finished outputs at every stage.
Loometo is a visual canvas for AI generation. Drag nodes, connect them, hit Run. It's model-agnostic — one MuAPI key routes to hundreds of models (Kling, Veo, Sora, Wan, Flux, Nano Banana, Tripo, and more), and you pay providers directly at cost instead of a per-credit platform markup.
Think ComfyUI, but friendlier — with audio, 3D, ready-made templates, prompting help built in, and one thing no other node canvas has: your coding agent can drive it.
Table of contents
Related MCP server: excalidraw-mcp
What makes it different
Batteries included. Not just image + video. Text-to-speech and voice cloning, image → 3D (rotate in-node + export GLB/STL), 22 pre-wired pipeline templates, a director-formula prompt builder, and model-recommendation hints — out of the box.
Your agent builds the canvas. Point Claude Code, Codex, or Cursor at the built-in MCP server and say "build me a UGC pipeline." Nodes appear and wire themselves live in your browser. (See
mcp/.)Model-agnostic, no lock-in. Swap any model the day it launches. Loometo is the layer over the whole market, not a bet on one lab.
Your keys, your machine. No accounts, no database, no telemetry. Keys stay in your browser or your
.env.local. It runs entirely locally.
Quick start
git clone https://github.com/dinimiciuillabs-cloud/loometo.git
cd loometo
npm install
npm run dev # http://localhost:3005Then open the app, hit Setup Keys (top right), paste your MuAPI key, and Save. That's it — no config files needed. That single key unlocks ~80% of Loometo.
Prefer config files? cp .env.example .env.local, add MUAPI_API_KEY,
restart the dev server. .env.local always wins over pasted keys.
API keys — where each one comes from
Every key can be pasted in-app under Settings → API Keys (stored in your
browser, sent only to your own localhost server). Three can also go in
.env.local:
Provider | Powers | Get your key | In-app paste |
|
MuAPI (required) | Video, 3D, upscale, image edit, face swap, lip sync — the ~250-model catalog | ✅ | ✅ | |
Google AI | Imagen 4, Veo 3, Gemini — direct, no reseller markup | ✅ | ✅ | |
ElevenLabs | TTS + voice cloning (free tier available) | ✅ | ✅ | |
OpenAI | GPT-Image / DALL·E, LLM nodes | ✅ | — | |
Meshy | Direct image → 3D (cheapest 3D path) | ✅ | — | |
Figma | Pull frames onto the canvas | ✅ | — |
Optional: set Cloudflare
R2_*keys in.env.localand every generated asset is auto-copied to your own bucket — provider URLs expire in 24–72h, this makes them permanent.
The tour, in five screens
You're never staring at an empty grid
First launch: a three-step explainer, all 22 templates one click away, and a search over all 57 nodes.
Model intelligence, inside the node
Searchable picker with ★ PICK recommendations and $ / $$ / $$$ cost tiers on every model — before you spend a cent. Parameters below it are schema-driven, so you can never set a value the model rejects.
Every node explains itself
Tap the (i) on any of the 57 nodes: plain English, what goes in, what comes
out. Typed handles mean wires only connect where the types match.
Prompt like a film director
The Director Formula node encodes real cinematography discipline — shot codes (ECU → EWS), annotated lens bands, lighting split into source / direction / quality like a DoP would, plus textures, composition, style reference and emotional vision. Eleven fields in, one director-grade prompt out, wired into any generator. Save the whole configuration as a preset.
Drive it from Claude Code / Codex / Cursor
Loometo ships a zero-dependency MCP server (mcp/loometo-mcp.mjs,
Node 18+) wrapping its HTTP canvas bridge. Your agent reads the graph, adds
nodes, wires them, and fires runs — live in your browser tab.
you › build me a UGC pipeline for a street-food ad
→ snapshot() reads empty canvas
→ add_node(character-board) ✓ char-board-1
→ add_node(image-to-video, kling) ✓ i2v-1
→ connect(char-board-1 → i2v-1) ✓
→ set_param(i2v-1, aspect 9:16) ✓
→ run_all() ✓ generating…Setup (3 steps):
npm run devand leave the browser tab open — set your keys first (Setup Keys → paste → Save) or every run the agent fires will fail.Register the server in your agent — use the absolute path:
// Claude Code (.mcp.json or ~/.claude.json) — or: // claude mcp add loometo -- node /abs/path/mcp/loometo-mcp.mjs "mcpServers": { "loometo": { "command": "node", "args": ["/abs/path/to/loometo/mcp/loometo-mcp.mjs"], "env": { "LOOMETO_URL": "http://localhost:3005" } } }Codex uses the same shape in
~/.codex/config.toml; Cursor / Cline in their MCP settings. Full walkthrough:mcp/README.md.Reload your agent and just ask. Nodes are addressable by the name on the card — you can literally say "run char-board-1".
The 11 tools: snapshot · add_node · connect · set_param ·
run_node · run_all · delete_node · clear · select · pulse ·
list_models
22 one-click pipeline templates
Every template is a pre-wired canvas — models pre-picked, prompts pre-filled with director-formula briefs you edit in place.
Template | What it wires | Template | What it wires |
UGC Clip | Person + place → video | Voiceover Reel | Script + face → talking clip |
Product Shots | One photo → angle board + cutout | Animated Logo | Logo → motion sting |
Link → Carousel | Paste a link, post a carousel | Sticker Pack | Photo → 4 die-cut stickers |
Launch Kit | One hero → every channel | Background Swap | Cut out → drop in a scene |
3D Product Spin | Photo → spinnable 3D | Relight Product | Phone photo → studio light |
Style Remix | Photo + idea → upscaled art | Virtual Try-On | Person + outfit idea → look |
Face Swap | Two photos → swapped | Headshot Studio | Selfie → pro headshots |
Video Restyle | Clip + idea → new look | Photo Restore | Old / blurry → sharp |
Talking Avatar | Photo + script → talking video | Object Remover | Delete anything from a photo |
Full UGC Campaign | Character + place + product → 3 clips | Ad Variant Factory | One hero → 6 ad angles |
Carousel from Idea | Topic → carousel | Blog Hero | Article link → hero image |
All 57 nodes, in plain English
The same explanations behind each node's (i) button. Click a group to expand.
Node | What it does |
Prompt | Type what you want to create here. Like telling the AI 'make me a sunset on a beach' — it's the starting point of everything. |
Prompt Enhancer | Your prompt goes in, a better prompt comes out. The AI rewrites it with more detail so your image or video looks amazing. |
Audio Transcriber | Listens to your uploaded audio and writes out the spoken script. Feed the transcript into a Prompt Combiner so the video model knows what is being said — not just lip shapes. |
Prompt Combiner | Connect two prompts and this merges them into one. Useful when you have different ideas you want to blend together. |
Director Formula | Fill the 7 fields like a creative director writing a brief — subject + action, shot type, lens, lighting (3 facets), composition, style ref, vision. The node assembles a director-grade prompt that wires into any image generator. Vocabulary taken verbatim from cinematography sources. |
Run Any LLM | Talk to any AI brain (ChatGPT, Gemini, Claude) directly inside your workflow. Ask questions, get descriptions, or have it write prompts. |
Image Describer | Give it a photo and it tells you exactly what's in it as a detailed prompt. Perfect for recreating a style. |
Image → JSON | Vision LLM produces a strict JSON description of one or several images of the same subject. Wire optional Extra Instructions (positive direction: 'include hex codes', 'list every logo separately') and Negative (avoid: 'don't speculate about brands not visible'). Output JSON wires into any text input for downstream identity lock. |
Node | What it does |
Text → Speech | Turns a script into a spoken audio clip. Wire a Prompt with your script in, get audio out. For ElevenLabs free tier, wire a Voice Cloner in (library voices are paid-only). Pipe that into Lip Sync to make a video character say it. |
Voice Cloner | Wire in a 30+ second audio sample (your voice, a podcast clip, anything clean). Give it a name. Click Clone. ElevenLabs returns a voice_id you can wire into Text → Speech for unlimited future TTS in that voice. Free tier compatible. |
Lip Sync | Upload a video of a person and an audio clip — it animates their lips to match the audio. |
Speech → Video | Take a portrait photo, add a voice recording, and get a video of that person speaking. Perfect for AI avatars. |
Upload Audio | Pick an audio file (mp3, wav). Feed it into Lip Sync or Speech → Video. |
Node | What it does |
Video Describer | Same as Image Describer but for videos. Watches your video and writes out what's happening so you can recreate or remix it. |
Text → Video | Type what you want to see and it creates a video. The dropdown shows only what each model truly supports. |
Image → Video | Animate a single photo (First Frame), wire First + Last for transition (Veo 3.1), or wire multiple references for reference-conditioned models. Ref handle count is set by the chosen model: Sora 2 / Veo 3 base = 1, Pixverse = 2, Veo 3.1 Reference = 3, Wan 2.1 Reference = 5, Vidu Q1/Q2 + Kling O1 Reference = 7, Seedance Omni = 9. Address each in the prompt as @image1..@imageN. |
Video → Video | Transform a video's look completely. 'Make this look like anime' or 'make it look like 1970s film.' Same motion, new style. |
Video Upscale | Makes blurry or low-quality videos look sharp and crisp. |
Video Matte | Removes the background from every frame of a video automatically. Tracks your subject — like a green screen without the green screen. |
Upload Video | Pick a video file (mp4, mov, webm). Feed it into Video → Video, Lip Sync, or other video tools. |
Node | What it does |
Image → 3D | Generate a 3D mesh from one image, multiple angle photos, or a text prompt. Tripo3D H31 ($0.20-$0.30) is the cheapest; Meshy 6 ($0.50) is most detailed. Drag the mesh to rotate. After the GLB loads, a 2D snapshot is auto-captured so downstream image nodes have a usable image to consume; download serves the .glb. |
Node | What it does |
Extract Frame | Grab any single frame from a video and save it as a still image. Pick the exact moment you want — frame by frame. |
Painter | Brush directly on the canvas. Paint MASKS for Inpaint / Object Remove, or sketch overlays that flow downstream as images. |
Import from Figma | Paste a Figma frame URL. Right-click any frame in Figma → Copy link to selection. Hit Run to pull it in as an image. |
Upload Image | Pick an image from your computer. It becomes the starting point for any image workflow. |
Import from URL | Paste any website link. Free mode extracts the readable text directly ($0). Google mode uses Gemini to clean the text AND recover the page's images and logo. Tap the recovered assets to pick which ones flow into your carousel as references — a tick marks the chosen ones. |
Compare | Put two images side by side with a draggable slider. Perfect for before/after. |
Iterator | Run the same process on multiple images at once. Instead of generating 10 images one by one, the Iterator does them in a batch. |
Output | The finish line. Connect image, video, and/or audio here to preview and download each. Pair Veo video + TTS audio here, then combine in your editor (CapCut / DaVinci). |
Node | What it does |
Text → Image | Type a description and it draws a picture. Models that accept a STYLE REFERENCE image (Nano Banana, Flux Kontext, GPT-Image) expose a second Style Ref handle. Pure T2I models hide it. |
Image → Image | Give it a photo + a description and it transforms the photo. Like 'make this look like an oil painting.' |
Upscale | Makes small or blurry images bigger and sharper. It actually invents new detail — turn a tiny image into a crisp 4K masterpiece. |
Remove Background | Cuts out the background of any photo automatically. Subject stays, background disappears. Perfect for cutout images. |
Relight | Changes where the light is coming from in a photo. Make daytime look like sunset, or add dramatic studio lighting — without a studio. |
Inpaint | Paint over something you don't want and the AI fills it in naturally. Remove people, erase logos — it blends in seamlessly. |
Outpaint | Makes your image wider or taller by generating what would be outside the frame. Like zooming out — the AI imagines what's beyond the edges. |
Face Swap | Put your face (or anyone's face) onto a different body or scene. Handles skin tone, lighting and angle automatically. |
Object Remove | Point at anything and poof — it's gone. The AI fills in what the background looks like. Remove tourists, wires, anything. |
Compositor | Visual layer editor. Wire a SCENE + up to 4 SUBJECTS, then drag / resize / reorder them in the polaroid. Two output modes: FLAT (exact pixel composite, no AI) or AI BLEND (sends the arrangement to Nano Banana Pro Edit which re-renders with matched lighting and perspective). |
Reference Sheet | Feed up to 6 photos of the SAME thing (a person, a place, a product) and one prompt. The model uses them all as identity / location anchors. Better than asking one model to 'imagine four angles' from a single photo. |
Character Board | Wire 1-5 photos of a person AND optionally a JSON spec from Image→JSON. Generates a studio model board (8 head expressions, 4 full-body angles, hand/eye/shoulder close-ups, fabric and skin-tone swatches) with anti-fake realism + director-formula lighting baked in. Feed the output into Image→JSON downstream to lock identity for compositors. |
Location Board | Wire 1-5 photos of a place AND optionally a JSON spec. Generates a location reference board — exterior angles, time-of-day variants (golden/midday/blue/night), signage close-ups, brand swatches. Director-formula lens + 4-facet lighting baked per panel. Feed the output into Image→JSON downstream. |
Product Board | Wire 1-5 photos of a product AND optionally a JSON spec. Generates a spec-board — hero angles, in-hand shots, brand-mark and material close-ups, colourway swatches. Director-formula lens + 4-facet softbox lighting baked per panel. Feed the output into Image→JSON downstream. |
B-roll Board | Wire 1-5 photos of your scene / subject / product. Generates 6-9 cutaway-style frames sharing one colour grade and lighting era — hands, atmosphere details, environment beats, lifestyle moments. Use for video edit asset library or campaign supporting shots. |
Storyboard | Wire 1-5 photos of your character / scene / setting + an extra-direction line describing the story beat-by-beat. Generates 6-9 numbered panels, each labelled with shot type (ECU / CU / MS / FS / LS) and a one-line action. Photoreal hero frames at pre-vis quality. |
Mascot Board | Wire 1-5 photos / concept art of a brand mascot. Generates a mascot bible — 4 orthographic views (front / 3-4 / side / back), 6 expressions, 4-5 poses, brand palette swatches, face + hand detail callouts, scale reference next to a human. Locks the mascot for consistent reuse across campaigns. |
Creature Board | Wire 1-5 photos / concept art of a creature, monster, or designed animal. Generates a creature concept-art bible — 5 orthographic views (front / 3-4 / side / back / top-down), face / limb / signature-feature / texture / joint detail callouts, material swatches (skin / fur / scales / feathers), threat-display + resting + hunting pose silhouettes, scale reference next to a human. |
Outfit Change | Virtually try on different outfits. Describe the clothing and the AI puts it on the person. Great for fashion and e-commerce. |
Crop | Cut your image or video to a specific size. Choose 16:9 for YouTube, 9:16 for TikTok, or 1:1 for Instagram. Instant, no AI needed. |
Blur | Makes things fuzzy. Blur backgrounds, censor things, or create a dreamy soft-focus effect. Slide to control how blurry. |
Levels | Make your image brighter, darker, or more contrasty. Like the basic sliders in your phone's photo editor but in your workflow. |
Invert | Flips all colors to their opposite — like a photo negative. Black becomes white. Very useful for flipping masks. |
Mask Extractor | Click on objects in your image and it automatically traces around them. Magic scissors that cut out exactly what you want. |
Mask by Text | Describe what to select — 'the person's hair' or 'the car' — and it draws the selection automatically. No clicking needed. |
Matte Grow/Shrink | Makes your selection slightly bigger or smaller. Use this to clean up rough edges. |
Merge Alpha | Takes your image and your mask and combines them — the masked area becomes transparent. How you get cutout images with no background. |
Carousel Board | Turns any text (wire in an Import from URL, or paste directly) into a ready-to-post social carousel: hook slide, point slides, CTA slide. Wire a Logo in to brand every slide and a Reference Image for the hook background. Finish 'Flat' renders locally for free; 'AI Enhance' passes each slide through Nano Banana Pro Edit (~$0.12/slide). One output handle per slide. |
The model catalog — all 276
The catalog lives in public/muapi-schema.json —
selectors read allowed values from it, so the UI never offers a parameter a
model rejects. New model out today? Add one schema entry and it appears.
Routing: models run either direct (Google, OpenAI, Meshy, ElevenLabs —
your key, raw cost) or via the MuAPI reseller (one key, ~250 models).
lib/api/router.ts picks the route per model.
Cost tiers — $ cheap · $$ medium · $$$ premium — show on every model in
the picker.
Model slug | Variant | Extra parameters |
| AI Video Effects | name, aspect_ratio, resolution, quality |
| Motion Controls | name, aspect_ratio, resolution, quality |
| VFX | name, aspect_ratio, resolution, quality |
| Image to Video | images_list, aspect_ratio |
| Image to Video [Fast] | images_list, aspect_ratio |
| Image to Video | aspect_ratio, resolution, duration |
| Image to Video | aspect_ratio, resolution, quality, duration |
| Image to Video | aspect_ratio, resolution, num_videos, variety |
| Image to Video | aspect_ratio |
| Seedance 2.0 Omni Reference | images_list, aspect_ratio, quality, duration |
| Seedance 2 VIP Omni Reference Fast | images_list, aspect_ratio, duration |
| Lite Image to Video | last_image, resolution, duration, camera_fixed |
| Pro Image to Video | resolution, duration, camera_fixed |
| Master Image to Video | aspect_ratio, duration |
| Standard Image to Video | aspect_ratio, duration |
| Pro Image to Video | last_image, aspect_ratio, duration |
| Image to Video | last_image, aspect_ratio, resolution, quality |
| Act 2 Image to Video | reference_video_url, aspect_ratio |
| Image to Video | images_list, aspect_ratio, resolution, duration |
| Image to Video | images_list, aspect_ratio, resolution, duration |
| Reference I2V | images_list, aspect_ratio |
| Standard I2V | end_image_url, duration, resolution |
| Pro I2V | end_image_url, duration, resolution |
| Video Effects | name |
| Image to Video | images_list, aspect_ratio, resolution, duration |
| Lite Reference to Video | images_list, resolution, duration |
| Reference to Video | images_list, resolution, aspect_ratio, duration |
| Pro Image to Video | duration |
| Image to Video | resolution, duration |
| Image to Video (Fast) | resolution, duration |
| Sora 2 Image to Video | images_list, aspect_ratio, duration, remove_watermark |
| Image to Video | prompt only |
| Sora 2 Pro Image to Video | images_list, aspect_ratio, duration, resolution |
| Motion 2.0 I2V | aspect_ratio |
| Image to Video | last_image, motion, strength, options |
| Image to Video | last_image, aspect_ratio, duration, resolution |
| Image to Video [Fast] | last_image, aspect_ratio, duration, resolution |
| Reference to Video | images_list, resolution, duration, generate_audio |
| Pro Image to Video Fast | resolution, duration, camera_fixed |
| Pro Image to Video | duration, generate_audio |
| Fast Image to Video | duration, generate_audio |
| Reference I2V | images_list, resolution, aspect_ratio, duration |
| Turbo I2V | last_image, resolution, duration, bgm |
| Pro I2V | last_image, resolution, duration, bgm |
| Pro I2V | resolution |
| Standard I2V | duration |
| Fast I2V | duration, go_fast |
| Standard Image to Video | duration |
| Image to Video | images_list, mode, duration |
| Image to Video [Pro] | last_image, aspect_ratio, duration |
| Reference to Video [Pro] | images_list, aspect_ratio, duration, keep_original_sound |
| Image to Video | duration, sound |
| Image to Video | images_list, style, thinking, aspect_ratio |
| Spicy Image to Video | resolution, duration |
| Image to Video | resolution, duration, shot_type |
| Image to Video [Standard] | last_image, duration |
| Reference to Video [Standard] | images_list, aspect_ratio, duration |
| Image to Video | last_image, aspect_ratio, resolution, duration |
| Image to Video [Fast] | last_image, aspect_ratio, resolution, duration |
| Std Image to Video | resolution, duration |
| Image to Video [Pro] | last_image, duration, generate_audio |
| Image to Video [standard] | last_image, duration, generate_audio |
Model slug | Variant | Extra parameters |
| Text to Video | aspect_ratio |
| Text to Video [Fast] | aspect_ratio |
| Text to Video | aspect_ratio, resolution, duration |
| Text to Video | aspect_ratio, resolution, quality, duration |
| Text to Video | aspect_ratio |
| Fast Text to Video | aspect_ratio |
| Lite Text to Video | aspect_ratio, resolution, duration, camera_fixed |
| Pro Text to Video | aspect_ratio, resolution, duration, camera_fixed |
| Master Text to Video | aspect_ratio, duration |
| Text to Video | aspect_ratio, resolution, quality, duration |
| Text to Video | aspect_ratio, resolution, duration |
| Text to Video | aspect_ratio, resolution, duration |
| Fast Text to Video | aspect_ratio, resolution |
| Standard T2V | duration, resolution |
| Pro T2V | duration, resolution |
| Text to Video | aspect_ratio, resolution, duration |
| Pro Text to Video | aspect_ratio, duration |
| Text to Video | aspect_ratio, resolution, duration |
| Text to Video (Fast) | aspect_ratio, resolution, duration |
| Sora Text to Video | aspect_ratio, resolution |
| Sora 2 Text to Video | aspect_ratio, duration, remove_watermark |
| Text to Video | aspect_ratio |
| Sora 2 Pro Text to Video | aspect_ratio, duration, resolution, remove_watermark |
| Text to Video | aspect_ratio, duration, resolution |
| Text to Video [Fast] | aspect_ratio, duration, resolution |
| Sora 2 Pro Storyboard | shots, duration, images_list, aspect_ratio |
| Extend Video | request_id |
| Pro Text to Video Fast | resolution, duration, aspect_ratio, camera_fixed |
| Pro Text to Video | duration, generate_audio |
| Fast Text to Video | duration, generate_audio |
| Pro T2V | resolution |
| Standard T2V | duration |
| Text to Video | aspect_ratio, mode, duration |
| Text to Video [Pro] | aspect_ratio, duration |
| Text to Video | aspect_ratio, duration, sound |
| Text to Video | style, thinking, aspect_ratio, resolution |
| Text to Video | aspect_ratio, resolution, duration, shot_type |
| Text to Video | aspect_ratio, resolution, duration, generate_audio |
| Text to Video [Fast] | aspect_ratio, resolution, duration, generate_audio |
| Std Text to Video | aspect_ratio, resolution, duration |
| 4k Video | request_id |
| Text to Video [Pro] | aspect_ratio, duration, generate_audio |
| Text to Video [standard] | aspect_ratio, duration, generate_audio |
| Seedance 2.0 | aspect_ratio, duration, quality |
Model slug | Variant | Extra parameters |
| Image Upscaler | prompt only |
| Image Faceswap | swap_url, target_index |
| Dress Change | model_image_url, garment_image_url |
| Background Remover | prompt only |
| Product Shot | scene_description |
| Skin Enhancer | prompt only |
| Color Photo | prompt only |
| Kontext Dev I2I | images_list, aspect_ratio, num_images |
| Product Photography | person_image_url, product_image_url |
| Ghibli Style | prompt only |
| Image Extension | prompt only |
| Object Eraser | mask_image_url |
| Kontext Pro I2I | images_list, aspect_ratio |
| Kontext Max I2I | images_list, aspect_ratio |
| Image to Image | images_list, aspect_ratio, num_images |
| Edit Image | mask_image_url, aspect_ratio, num_images |
| Image to Image | speed, aspect_ratio, variety, stylization |
| GPT Image 2 (Image to Image) | images_list, aspect_ratio, resolution, quality |
| Edit Image v3 | prompt only |
| Style Reference | speed, aspect_ratio, variety, stylization |
| Omni Reference | speed, aspect_ratio, weight, variety |
| Subject Reference | aspect_ratio, num_images |
| Character | render_speed, style, aspect_ratio, num_images |
| Pulid Image to Image | aspect_ratio |
| Edit Image | aspect_ratio |
| Image Effects | name |
| Edit Image | images_list, aspect_ratio |
| v3 Reframe | aspect_ratio, render_speed, style, num_images |
| Edit Image v4 | images_list, aspect_ratio, resolution, num_images |
| Image Effects | name, aspect_ratio |
| Image Effects | name |
| Redux Image to Image | aspect_ratio, num_images |
| Edit Image Plus | images_list, width, height |
| Edit Image | images_list, width, height |
| Image to Image | style, aspect_ratio, strength, quality |
| Edit Image | prompt only |
| Image Upscale | upscale_factor |
| Image Upscale | resolution |
| Edit Image Plus Lora | images_list, rotate_right_left, move_forward, vertical_angle |
| Pro Edit Image | images_list, aspect_ratio, resolution |
| Image to Image | make_input |
| Edit Image [Pro] | images_list, aspect_ratio, resolution |
| Edit Image [Dev] | images_list, width, height |
| Edit Image [Flex] | images_list, aspect_ratio, resolution |
| Edit Image [Pro] | images_list, aspect_ratio, resolution |
| Reference to Image | images_list, aspect_ratio, resolution |
| Edit Image | images_list, aspect_ratio, quality |
| Edit Image 2511 | images_list, width, height |
| Edit Image | images_list |
| Text to Image 2512 | width, height |
| Image to Image | images_list, aspect_ratio, quality |
| Image to Image | prompt only |
| Image to Image | model_url, api_key, params |
| Edit Image [Klein 4B] | images_list, aspect_ratio |
| Edit Image [Klein 9B] | images_list, aspect_ratio |
| Add Image Watermark | watermark_image_url, position, opacity, scale |
Model slug | Variant | Extra parameters |
| Video Faceswap | target_gender, target_index |
| v2 Video to Video | duration |
| Act 2 Video to Video | reference_video_url, aspect_ratio |
| Aleph Video to Video | aspect_ratio |
| Modify V2V | prompt only |
| Flash Reframe V2V | aspect_ratio, duration |
| AI Dance Effects | resolution |
| Audio to Video | resolution |
| Video Upscaler | resolution, copy_audio |
| Edit Video | resolution |
| Video Translate | language |
| Anime Video | mode, resolution |
| Video Upscale | upscale_factor |
| Video Upscaler Pro | resolution |
| Watermark Remover | prompt only |
| Remix Video | aspect_ratio |
| Video to Video | make_input |
| Edit Video [Pro] | images_list, aspect_ratio, keep_original_sound |
| Edit Video Fast [Pro] | images_list, aspect_ratio, keep_original_sound |
| Spicy Video Extend | resolution, duration |
| Edit Video [Standard] | images_list, keep_original_sound |
| Pro Motion Control | prompt only |
| Video Extend | resolution, duration, generate_audio, camera_fixed |
| Video Extend [Fast] | resolution, duration, generate_audio, camera_fixed |
| Std Motion Control | prompt only |
| Add Video Watermark | watermark_image_url, position, opacity, scale |
| AI Clipping | num_highlights, aspect_ratio, return_coordinates_only |
Model slug | Variant | Extra parameters |
| v2 Text to Audio | duration |
| Create Music | style, model, instrumental, negative_tags |
| Remix Music | style, model, instrumental, negative_tags |
| Extend Music | style, model, continue_at, instrumental |
| Voice Clone | custom_voice_id, model, need_noise_reduction, need_volume_normalization |
| Speech HD | voice_id, speed, volume, pitch |
| Speech Turbo | voice_id, speed, volume, pitch |
| Text to Audio | make_input |
Model slug | Variant | Extra parameters |
| Dev | width, height, num_images |
| Kontext Dev T2I | aspect_ratio, num_images |
| Fast | width, height, num_images |
| Dev | width, height, num_images |
| Full | width, height, num_images |
| Anime Generator | width, height |
| Text to Image | width, height |
| Kontext Pro T2I | aspect_ratio |
| Kontext Max T2I | aspect_ratio |
| Text to Image | aspect_ratio, num_images |
| Text to Image | speed, aspect_ratio, variety, stylization |
| Schnell | width, height, num_images |
| GPT Image 2 (Text to Image) | aspect_ratio, resolution, quality |
| Text to Image v3 | aspect_ratio |
| Text to Image | aspect_ratio, num_images |
| v3 Text to Image | render_speed, style, aspect_ratio, num_images |
| Text to Image | aspect_ratio |
| Imagen 4 | aspect_ratio, num_images |
| Imagen 4 Fast | aspect_ratio, num_images |
| Imagen 4 Ultra | aspect_ratio |
| Text to Image | width, height |
| Text to Image v4 | aspect_ratio, resolution, num_images |
| Text to Image v2.1 | width, height |
| Text to Image | width, height |
| Krea Dev | aspect_ratio, num_images |
| Text to Image | width, height |
| Text to Image | width, height |
| Text to Image | width, height |
| Text to Image v3.0 | width, height |
| Phoenix 1.0 T2I | aspect_ratio |
| Lucid Origin T2I | aspect_ratio |
| Text to Image | aspect_ratio |
| Text to Image | aspect_ratio |
| Text to Image Pro | aspect_ratio, resolution |
| Text to Image [Pro] | aspect_ratio, resolution, num_images |
| Text to Image Turbo | width, height |
| Text to Image [Dev] | width, height |
| Text to Image [Flex] | aspect_ratio, resolution |
| Text to Image [Pro] | aspect_ratio, resolution |
| Text to Image | aspect_ratio, resolution |
| Text to Image | aspect_ratio, quality |
| Text to Image | aspect_ratio, quality |
| Text to Image | width, height |
| Text to Image [Klein 4B] | aspect_ratio |
| Text to Image [Klein 9B] | aspect_ratio |
| Text to Image Base | aspect_ratio, strength |
| Seedream 5.0 | prompt only |
Model slug | Variant | Extra parameters |
| Dev LoRA | model_id, width, height, num_images |
| Image to Video (LoRA) | lora_list, aspect_ratio, resolution, quality |
| Text to Video (LoRA) | lora_list, aspect_ratio, resolution, quality |
| LoRA | lora_list, width, height |
Model slug | Variant | Extra parameters |
| sync | prompt only |
| latent | prompt only |
| Creatify | prompt only |
| Veed | prompt only |
| Audio to Video | resolution |
| Image to Video | resolution |
| Standard A2V | prompt only |
| Pro A2V | prompt only |
| Standard A2V | prompt only |
| Pro A2V | prompt only |
| Audio to Video | resolution |
Model slug | Variant | Extra parameters |
| GPT5 Nano Text | prompt only |
| Storyboard | duration |
| GPT5 Mini Text | prompt only |
| Text to Text | make_input |
| Text to Text | system_prompt, model, reasoning, priority |
| Image to Text | images_list, system_prompt, model, reasoning |
| Architect | prompt only |
| Agent Chat | message, conversation_id |
Model slug | Variant | Extra parameters |
| Image to 3D | should_texture, topology, target_polycount, should_remesh |
| Image to 3D | images_list, should_texture, topology, target_polycount |
| Image to 3D | mode, topology, target_polycount, should_remesh |
| Image to 3D | texture, texture_quality, geometry_quality, pbr |
| Image to 3D | images_list, texture, texture_quality, geometry_quality |
| Image to 3D | texture, texture_quality, geometry_quality, pbr |
| Image to 3D | texture, face_limit |
| Image to 3D | texture, face_limit |
| Image to 3D (Meshy direct) | prompt only |
Architecture & security
app/api/proxy/* server-side key-holding proxies (muapi, google, elevenlabs, r2, fetch)
app/api/canvas/* the agent bridge — HTTP ops + SSE live-sync
components/nodes/ all 57 node implementations
lib/api/ dispatch, per-model routing, model catalog
lib/flow/ templates, propagation, GLB→STL converter
mcp/ the zero-dependency MCP serverLocalhost-only API routes — every proxy and bridge route rejects non-local requests with a 403.
CSRF-hardened — mutating endpoints require
application/json, forcing a failing preflight on any cross-origin attempt.Keys stay yours —
.env.localkeys never reach the browser; browser-pasted keys travel only to your own localhost server.Zero telemetry — no analytics, no accounts, no database, no phone-home.
The dev server has no auth by design — it's built for your own machine. Don't expose it to the public internet without adding your own auth layer.
License & branding
Code: Apache-2.0. Use it, change it, ship it commercially, whatever you like.
Brand: "Loometo" and the Loometo logo are trademarks of Dinimiciuil Labs and are not covered by the code license — see TRADEMARK.md. Fork the code freely, but rebrand your fork: don't ship it as "Loometo" or use the logo. Credit travels with the code via NOTICE.
Made by Dinimiciuil Labs.
This server cannot be installed
Maintenance
Resources
Unclaimed servers have limited discoverability.
Looking for Admin?
If you are the server author, to access and configure the admin panel.
Latest Blog Posts
- Who's Calling? MCP Hosts Are an Identity Blind Spot (And the Spec Knows It)By Om-Shree-0709 on .mcpAgent IdentityOAuth 2.1
- Your AI Chatbot Just Exposed Your CEO's Salary to an InternBy Om-Shree-0709 on .Agent IdentityMCP SecurityOAuth Delegation
- Why MCP Servers Need Execution Sandboxing (And Why Your Current Stack Isn't Enough)By Om-Shree-0709 on .Agentic AiPrompt InjectionWebAssembly
MCP directory API
We provide all the information about MCP servers via our MCP API.
curl -X GET 'https://glama.ai/api/mcp/v1/servers/dinimiciuillabs-cloud/loometo'
If you have feedback or need assistance with the MCP directory API, please join our Discord server