"Information about AliDocs (Alibaba's document collaboration platform)" matching MCP connectors:
GET /v1/connectors — MCP directory API referenceMatching Connector Tools:
Unstructured document processing for LLM pipelines. Upload as PDF/DOCX/TXT any supported files, extract structured data (PII-redacted), build LLM-ready datasets, and search/export results — all via MCP tools (document.process, job.status, job.result, dataset.build, dataset.search, dataset.export).
AI-first file sharing and collaboration. 251 tools give agents a full workspace: file storage, branded shares, comments, workflows, and built-in RAG. 50GB free, no credit card.
SENATRAN: Recall, official-source lookup. Platform-hosted, pay per query with prepaid credit.
Your private knowledge base: upload documents (.md, .txt, .docx, PDF, images), the platform indexes
Document-to-Markdown MCP server — convert PDF, Office and HTML into LLM-ready Markdown.
Manage your Mistral platform — models, files, batch jobs, agents and RAG document libraries.
CareerProof MCP gives AI agents direct access to a professional-grade career and workforce intelligence platform. Two namespaces: atlas_* for HR/TA teams (candidate evaluation, batch shortlisting, competency scoring, interview generation, JD analysis, custom eval frameworks, research reports) and ceevee_* for professionals (CV optimization, career positioning, salary intelligence, market reports). Backed by RAG knowledge from 50+ premium research sources (McKinsey, BCG, HBR, Gartner, WEF)
Recommendations, search, catalogue, analytics, and platform admin tools for NeuronSearchLab
Rafter holds a team's durable knowledge — skills, agents and memory files, each versioned — and serves it to AI tools over MCP. Agents search across the team's artifacts before answering questions about how the team works or what was decided, fetch full artifact text along with its citation edges (cites, cited_by, links) to explore related material, and write new learnings back as memory. Also covers workspace, team and membership management.
Search 18M+ legal documents across 110+ countries. Case law, legislation, and doctrine with semantic + keyword hybrid search. Supports tool discovery, multi-jurisdictional queries, citation resolution, and full document retrieval. Requested missing datasets can be fully indexed within 48h.
VaultCrux Platform — 60 tools: retrieval, proof, intel, economy, watch, org
Answers questions about a business using its indexed website content, with cited sources.
LLMtoMD is the memory layer for AI coding agents. It converts any document — PDF, DOCX, slides, spreadsheets, images, audio, even whole websites — into clean, structured Markdown, then exposes it over MCP so your agent can search your FRDs, specs, and API docs on demand instead of re-reading (or forgetting) them.
Agentic search over your Dewey document collections from any MCP-compatible client.
ContextBook is an open-source MCP server that gives AI tools a persistent, searchable context library. Store information as Books and Pages, retrieve exactly what's needed via natural-language semantic search - injected on demand, not pre-loaded. Works with Cursor, Claude, Windsurf, and any MCP-compatible client. Self-hostable, MIT licensed.
Real-time web search, reasoning, and research through Perplexity's API
The Needle MCP server enables semantic search on documents stored in files like PDFs, DOCX, and XLSX by connecting AI applications to external data sources. It provides capabilities to create and manage document collections, perform natural language searches on stored content, and retrieve relevant information without requiring exact keyword matches.
Make your knowledge agent-ready. Connect docs from Confluence, Notion, GitHub, Dropbox, or Google Drive — any AI agent searches them via one MCP endpoint. 3 retrieval modes: vector search, broad search, and full document access. The agent decides how deep to dig.
Provides metadata information to AI agents through the search API.