Skip to main content
Glama
93,551 servers. Updated

Matching MCP tools:

Matching MCP Connectors:

"Using Knowledge Base Files for Drafting, Writing, and Editing Documents" matching MCP servers:

GET /v1/servers – MCP directory API reference
  • A
    license
    A
    quality
    D
    maintenance
    Let your AI agent call your phone and talk to you — MCP servers for live, interruptible voice calls + tiered alerts, using free self-hosted pieces (pjsua2 + whisper.cpp + Linphone). No paid telephony, no extra API key.
    3
    22
    Apache 2.0
  • A
    license
    A
    quality
    C
    maintenance
    MCP server for local speech-to-text using Whisper Large V3 (MLX), enabling audio transcription with text/timestamps/SRT output and LLM-based correction, all running offline on Apple Silicon.
    2
    MIT
  • A
    license
    A
    quality
    B
    maintenance
    Enables generating images, video, audio, and speech from MCP clients using your own Vidofy account, with access to hundreds of models for text-to-video, image-to-video, image editing, lipsync, text-to-speech, and voice cloning.
    9
    109 npm
    1
    MIT
  • A
    license
    A
    quality
    D
    maintenance
    Enables local text-to-speech synthesis for Claude and Cursor using Supertonic 3, with support for multiple voices, expressions, and languages. No API key or cloud required.
    3
    MIT
  • F
    license
    A
    quality
    D
    maintenance
    Voice input for Claude Code — speak Vietnamese or English; hold a hotkey, speak, release, and the transcribed text is cleaned and pasted into the input box for editing before sending.
    2
    -
  • A
    license
    A
    quality
    B
    maintenance
    Translate PDF, Word (DOCX), Excel (XLSX) and PowerPoint (PPTX) files, images, subtitles and text. Turn audio and video into transcripts or translated subtitles. Preserve layout where supported. Requires an Equalang API key and credits.
    8
    512 npm
    1
    Apache 2.0
  • A
    license
    A
    quality
    A
    maintenance
    Enables AI agents to speak using MacOS native text-to-speech, with support for blocking and non-blocking speech and a sequential queue.
    2
    2
    15
    MIT
  • A
    license
    A
    quality
    D
    maintenance
    An MCP server that enables transcribing local audio files and Telegram voice messages using OpenAI's Whisper via local inference or cloud API. It supports multiple audio formats, automatic language detection, and optional word-level timestamps for AI-powered audio analysis.
    5
    1
    MIT
  • A
    license
    A
    quality
    D
    maintenance
    Extract structured knowledge from voice recordings. Transcribes audio using Mistral's Voxtral model and lets your LLM agent handle post-processing within your existing setup.
    16
    MIT
  • A
    license
    A
    quality
    C
    maintenance
    Enables MCP-capable assistants to transcribe local audio files or URLs and perform speaker diarization for Spanish and Portuguese audio, with options for speaker count hints, domain prompts, and transcript retrieval in multiple formats.
    4
    7 npm
    MIT
  • A
    license
    A
    quality
    D
    maintenance
    Local, private audio transcription MCP server enabling AI agents to transcribe audio files entirely on-device without uploading data.
    3
    MIT