mcp-textbrowser
Click on "Install Server".
Wait a few minutes for the server to deploy. Once ready, it will show a "Started" state.
In the chat, type
@followed by the MCP server name and your instructions, e.g., "@mcp-textbrowseropen bytesbrains.io and summarize the page content"
That's it! The server will respond to your query, and you can continue using it as needed.
Here is a step-by-step guide with screenshots.
@bytesbrains/mcp-textbrowser
MCP server — text-first headless browser for Claude Code and any MCP host. DOM + OCR text maps. Zero image tokens by default. 5-15x cheaper than screenshot-based browser MCPs.
browser_navigate(url) → DOM elements + OCR text (~200 tokens)
browser_navigate(url, visual=true) → text + PNG (use for layout/color only)Why
Every screenshot-based browser MCP sends a PNG to the AI on every action. At 1280×800 that's ~1,300 image tokens per page — 5-15x more expensive than reading the same content as text.
mcp-textbrowser captures a screenshot for OCR, extracts the text, then discards the image. Only structured DOM elements and OCR text reach the model. Switch to visual=true only when you genuinely need pixels (layout checks, color, design review).
Mode | Tokens per action | When to use |
text-only (default) | ~150–400 | Everything: navigation, forms, data extraction, workflows |
| ~1,500–3,000 | Layout, colors, CSS, design review |
Related MCP server: MCP Master Puppeteer
Install
Step 1 — install Chromium (one-time, ~130MB):
npx playwright install chromiumStep 2 — add the MCP server:
Claude Code (CLI) — one command
claude mcp add textbrowser -- npx -y @bytesbrains/mcp-textbrowserThat's it. Restart Claude Code and the tools are ready.
Claude Desktop
Add to ~/Library/Application Support/Claude/claude_desktop_config.json (Mac) or %APPDATA%\Claude\claude_desktop_config.json (Windows):
{
"mcpServers": {
"textbrowser": {
"command": "npx",
"args": ["-y", "@bytesbrains/mcp-textbrowser"]
}
}
}Restart Claude Desktop.
Manual (any MCP host)
Use command: npx, args: ["-y", "@bytesbrains/mcp-textbrowser"] in your host's MCP server config.
Tools
Tool | What it does |
| Open a URL, return page context |
| Click by CSS selector / XPath / visible text |
| Fill an input field |
| Scroll page or element into view |
| Capture current page context |
| Read current page without navigating |
| Run JS in the page (safe DOM ops only) |
All tools default to text-only — pass visual: true to any tool to also receive the PNG.
Example output
Page: https://example.com/
Title: Example Domain
Viewport: 1280x800
Elements (3 interactive of 14 total):
[1] <a> href="https://iana.org/domains/example" bbox=(133,175,254,195)
text: "More information..."
OCR (full page screenshot):
Example Domain
This domain is for use in illustrative examples in documents.
You may use this domain in literature without prior coordination or asking for permission.Requirements
Node.js 18+
Chromium:
npx playwright install chromium(one-time, ~130MB)
License
MIT © BytesBrains
Built by Agent, for Agents 🤖
Built and maintained by BytesBrains — AI automation & agents, engineered to production standards. The model proposes, code guarantees.
Maintenance
Resources
Unclaimed servers have limited discoverability.
Looking for Admin?
If you are the server author, to access and configure the admin panel.
Related MCP Servers
- AlicenseBqualityDmaintenanceA headless browser MCP server that allows AI agents to fetch web content and perform Google searches without API keys, supporting various output formats like Markdown, JSON, HTML, and text.Last updated224MIT
- Flicense-qualityDmaintenanceAn advanced MCP server for browser automation using Puppeteer, specifically optimized for token efficiency through minimal data returns and progressive enhancement. It enables agents to navigate pages, capture LLM-optimized screenshots, extract structured content, and perform batch interactions.Last updated3

Plasmateofficial
Alicense-qualityCmaintenanceAgent-native headless browser for AI agents. Converts web pages to a Semantic Object Model (SOM) instead of raw HTML — 17x average token reduction across real-world sites (up to 117x on complex pages). Native MCP server with fetch_page, extract_text, extract_links, and full browser automation. No API key required.Last updated77Apache 2.0- AlicenseAqualityBmaintenanceA token-efficient MCP server that gives AI agents structured access to the web, returning compact page summaries and targeted queries instead of full accessibility dumps.Last updated23334169MIT
Related MCP Connectors
Headless-browser-as-JSON with memorymarket cache economics. Real Chromium, crypto settlement.
A paid remote MCP for AI agent browser approval MCP, built to return verdicts, receipts, usage logs,
A paid remote MCP for AI agent browser DevTools MCP, built to return verdicts, receipts, usage logs,
Latest Blog Posts
- Who's Calling? MCP Hosts Are an Identity Blind Spot (And the Spec Knows It)By Om-Shree-0709 on .mcpAgent IdentityOAuth 2.1
- Your AI Chatbot Just Exposed Your CEO's Salary to an InternBy Om-Shree-0709 on .Agent IdentityMCP SecurityOAuth Delegation
- Why MCP Servers Need Execution Sandboxing (And Why Your Current Stack Isn't Enough)By Om-Shree-0709 on .Agentic AiPrompt InjectionWebAssembly
MCP directory API
We provide all the information about MCP servers via our MCP API.
curl -X GET 'https://glama.ai/api/mcp/v1/servers/bytesbrains/mcp-textbrowser'
If you have feedback or need assistance with the MCP directory API, please join our Discord server