mistral-ai-mcp
Click on "Install Server".
Wait a few minutes for the server to deploy. Once ready, it will show a "Started" state.
In the chat, type
@followed by the MCP server name and your instructions, e.g., "@mistral-ai-mcpocr this pdf: https://example.com/doc.pdf"
That's it! The server will respond to your query, and you can continue using it as needed.
Here is a step-by-step guide with screenshots.
mistral-ai-mcp
Mistral AI MCP server + CLI. OCR documents, TTS text-to-speech, STT speech-to-text via Voxtral.
Features
OCR - Extract text from PDFs, DOCX, PPTX, XLSX, images
TTS - Generate speech with preset voices or voice cloning
STT - Transcribe audio (batch or realtime)
Related MCP server: speaches-mcp
Install
npm install -g mistral-ai-mcpRequires Node.js 18+.
CLI Usage
mistral-ai ocr <file-or-url> # Extract text from documents/images
mistral-ai tts <text> # Generate speech from text
mistral-ai stt <audio> # Transcribe audio to text
mistral-ai config ... # Manage configurationOCR Examples
# Local PDF
mistral-ai ocr ./document.pdf > output.md
# From URL
mistral-ai ocr https://arxiv.org/pdf/2301.00001.pdf > paper.md
# Tables
mistral-ai ocr ./document.pdf --table-format htmlTTS Examples
# Preset voice
mistral-ai tts "Hello, world!" --voice-id alice
# Voice cloning via reference audio
mistral-ai tts "Hello from me!" --ref-audio ./my-voice.wav
# Output format
mistral-ai tts "Hello!" --voice-id bob --format wav > output.wavSTT Examples
# Basic transcription
mistral-ai stt ./audio.mp3
# Realtime mode (low latency)
mistral-ai stt ./audio.mp3 --realtime
# Speaker diarization
mistral-ai stt ./meeting.mp3 --diarize
# Specific language
mistral-ai stt ./audio.mp3 --language enSupported Formats
Tool | Input Formats | Output Formats |
OCR | PDF, DOCX, PPTX, XLSX, PNG, JPEG, AVIF | Markdown + YAML |
TTS | Text | MP3, WAV, PCM, FLAC, Opus |
STT | MP3, WAV, FLAC, OGG, WebM | Text + JSON |
Configuration
Set API Key
mistral-ai config api_key <your-key>Or via environment:
export MISTRAL_API_KEY=your-keyConfig Location
Linux/macOS:
~/.mistral-ai/config.jsonWindows:
%USERPROFILE%\.mistral-ai\config.json
Override with MISTRAL_AI_CONFIG_DIR env var.
Show Config
mistral-ai config showConfig Options
mistral-ai config api_key <key> # Set API key
mistral-ai config base_url <url> # API endpoint (default: https://api.mistral.ai/v1)
mistral-ai config model <model> # Default OCR modelMCP Server
Start server with stdio transport:
mistral-ai-mcpMCP Tools
Tool | Description |
| Extract text from PDF sources |
| Generate speech from text |
| Transcribe audio to text |
Example (Node.js client)
import { StdioClientTransport } from '@modelcontextprotocol/sdk/client/stdio.js';
import { Client } from '@modelcontextprotocol/sdk/client/index.js';
const transport = new StdioClientTransport({
command: 'node',
args: ['node_modules/.bin/mistral-ai-mcp'],
env: { MISTRAL_API_KEY: 'your-key' },
});
const client = new Client({ name: 'test', version: '1.0.0' }, { capabilities: {} });
await client.connect(transport);
// OCR
const ocrResult = await client.callTool({
name: 'ocr_pdf_url',
arguments: { pdf_url: 'https://example.com/doc.pdf' },
});
// TTS
const ttsResult = await client.callTool({
name: 'tts_speech',
arguments: { text: 'Hello!', voice_id: 'alice' },
});
// STT
const sttResult = await client.callTool({
name: 'stt_transcribe',
arguments: { audio_source: './audio.mp3' },
});Development
Setup
npm installScripts
npm run dev # Run with tsx (watch mode)
npm run build # Compile TypeScript
npm run clean # Remove dist/
npm run lint # ESLint check
npm run format # Prettier check
npm run format:write # Prettier fix
npm run typecheck # Type check only
npm run test # Run testsGit Hooks
Husky hooks enforce code quality:
pre-commit: Runs
npm run lint && npm run formatpre-push: Runs
npm run typecheck && npm run test
CI/CD
GitHub Actions workflows:
CI (
.github/workflows/ci.yml): Lint, build, typecheck on push/PRTests (
.github/workflows/test.yml): Run tests on push/PR
License
ISC
This server cannot be installed
Maintenance
Related MCP Servers
- Flicense-qualityDmaintenanceAn MCP server that enables Claude to perform OCR on local files using Mistral AI's document processing capabilities. It converts documents and images into markdown format for seamless analysis and interaction.Last updated
- FlicenseCqualityCmaintenanceAn MCP server that exposes speech-to-text and text-to-speech capabilities using a local speaches instance, allowing AI assistants to transcribe audio and generate speech.Last updated2
- Alicense-qualityAmaintenanceMCP server for offline speech-to-text and speaker diarization, enabling AI agents to transcribe audio locally without cloud APIs.Last updated3MIT
- Alicense-qualityDmaintenanceMinimal MCP server for local speech recognition using faster-whisper. Runs on CPU, no cloud required.Last updatedMIT
Related MCP Connectors
OCR, transcription, file extraction, and image generation for AI agents via MCP.
MCP server for Wan AI video generation
MCP server for AI dialogue using various LLM models via AceDataCloud
Latest Blog Posts
- Who's Calling? MCP Hosts Are an Identity Blind Spot (And the Spec Knows It)By Om-Shree-0709 on .mcpAgent IdentityOAuth 2.1
- Your AI Chatbot Just Exposed Your CEO's Salary to an InternBy Om-Shree-0709 on .Agent IdentityMCP SecurityOAuth Delegation
- Why MCP Servers Need Execution Sandboxing (And Why Your Current Stack Isn't Enough)By Om-Shree-0709 on .Agentic AiPrompt InjectionWebAssembly
MCP directory API
We provide all the information about MCP servers via our MCP API.
curl -X GET 'https://glama.ai/api/mcp/v1/servers/tan-yong-sheng/mistral-ai-mcp'
If you have feedback or need assistance with the MCP directory API, please join our Discord server