Skip to main content
Glama
22,816 servers. Last updated

"author:arizawan" matching MCP servers:

  • A
    license
    -
    quality
    C
    maintenance
    vidlizer pulls frames out of any video, image, or PDF using ffmpeg, sends them to a vision LLM, and returns a flow array — one entry per scene. Each entry tells you what happened, who was on screen, what text was visible, and what changed. If the video has audio, it transcribes it with Apple MLX Whisper and merges the speech into each step.
    Last updated
    1
    MIT