Skip to main content

45,132 servers. Last updated 2026-06-22 03:42

"author:arizawan" matching MCP servers:

vidlizer
Image & Video Processing Multimedia Processing Speech Processing
arizawan
A
license
-
quality
B
maintenance
vidlizer pulls frames out of any video, image, or PDF using ffmpeg, sends them to a vision LLM, and returns a flow array — one entry per scene. Each entry tells you what happened, who was on screen, what text was visible, and what changed. If the video has audio, it transcribes it with Apple MLX Whisper and merges the speech into each step.
Last updated 2026-04-29
MIT