Browser MCP Server

Browser MCP Server

Enables AI agents to control browsers with human-like behavior, stealth anti-detection, and 70 tools for navigation, interaction, and monitoring.

Category
Visit Server

README

🌐 Browser MCP Server

The most comprehensive Browser MCP Server for AI agents — 70 tools with human-like behavior, stealth anti-detection, session reuse, multi-browser support, 50+ device presets, and token-optimized outputs.

Built on Playwright + Model Context Protocol.


✨ Key Features

Feature Description
šŸ¤– Human-Like Behavior Bezier curve mouse movements, variable-speed typing, smooth scrolling — evades bot detection
šŸ›”ļø Stealth Anti-Detection WebGL spoofing, canvas fingerprint noise, navigator.webdriver hidden, plugin/language spoofing
šŸ”„ Session Reuse Connect to existing Chrome via CDP or user-data-dir — keeps cookies, logins, tabs
šŸŒ Multi-Browser Chromium, Firefox, WebKit — test across all engines
šŸ“± 50+ Device Presets iPhone, iPad, Pixel, Galaxy, OnePlus, Desktop — full viewport/UA/touch emulation
🌐 Proxy Support HTTP/HTTPS/SOCKS5 proxy with bypass rules
šŸ“Š Network Monitoring Auto-capture all requests/responses, console logs, filter by type
šŸ“ø Visual Annotation mark_page: numbered boxes on interactive elements — massive token savings
šŸ’¾ Storage State Save/load cookies + localStorage for cross-session persistence
šŸŽÆ Smart Actions Natural language commands: "click Login", "type hello into search"
šŸ“‰ Token Optimized Compact outputs, configurable truncation — saves 60-80% tokens vs competitors
šŸŽ¬ Video Recording Record browser sessions for debugging and replay
šŸ“ Geolocation Override latitude/longitude for location-based testing
šŸ”’ HTTPS Errors Ignore certificate errors for internal/staging sites

šŸš€ Quick Start

Install

git clone https://github.com/user/browser-mcp-server.git
cd browser-mcp-server
npm install
npx playwright install
npm run build

Run

# Auto-detect: tries CDP on port 9222, falls back to launching a new browser
node dist/index.js

# Connect to existing Chrome (preserves your logged-in sessions)
chrome --remote-debugging-port=9222
node dist/index.js --cdp http://localhost:9222

# Reuse Chrome profile (keeps all cookies, logins, extensions)
node dist/index.js --user-data-dir "C:\Users\you\AppData\Local\Google\Chrome\User Data"

# Launch Firefox with iPhone emulation and proxy
node dist/index.js --browser firefox --device "iPhone 15" --proxy-server http://proxy:8080

VS Code MCP Configuration

Add to .vscode/mcp.json:

{
  "servers": {
    "browser": {
      "type": "stdio",
      "command": "node",
      "args": ["path/to/browser-mcp-server/dist/index.js"]
    }
  }
}

Claude Desktop Configuration

Add to claude_desktop_config.json:

{
  "mcpServers": {
    "browser": {
      "command": "node",
      "args": ["path/to/browser-mcp-server/dist/index.js"],
      "env": {}
    }
  }
}

šŸ› ļø All 70 Tools

Connection (1)

Tool Description
browser_connect Connect to browser (CDP/profile/temp launch). Multi-browser, device emulation, proxy.

Navigation (5)

Tool Description
navigate Navigate to URL with configurable wait conditions
go_back Go back in browser history
go_forward Go forward in browser history
reload Reload current page
wait_for_navigation Wait for page navigation to complete

Page/Tab Management (4)

Tool Description
list_pages List all open tabs
new_page Open new tab (optionally with URL)
close_page Close a tab by ID
focus_page Bring a tab to front/focus

Interaction — Human-Like (14)

Tool Description
click Click element with Bezier curve mouse movement
type_text Type with human-like keystroke timing
fill_form Fill multiple form fields with natural pauses
select_option Select dropdown option by value or label
check_checkbox Check/uncheck checkbox or radio button
hover Hover with natural mouse movement
press_key Press key or combo (Enter, Ctrl+C, Tab, etc.)
scroll Scroll with smooth human-like motion
scroll_to_element Scroll element into view
wait_for_element Wait for element to appear/disappear
wait_for_text Wait for specific text on page
drag_and_drop Drag element to target with natural motion
upload_file Upload file(s) to input
handle_dialog Accept or dismiss browser dialogs

Reading/Observation — Token-Optimized (14)

Tool Description
get_page_content Get visible text (default 4000 chars)
get_page_html Get HTML (truncated to save tokens)
get_element_text Get element text content
get_element_attribute Get element attribute value
get_element_value Get form element value
element_exists Quick true/false existence check
element_count Count elements matching selector
get_bounding_box Get element position and size
take_screenshot Screenshot page or element (PNG/base64)
get_accessibility_tree Compact ARIA accessibility tree
get_form_elements List form elements (compact format)
get_links List page links (deduplicated)
get_table_data Extract table data as structured JSON
evaluate_javascript Run JavaScript in page context

Cookies & Storage (10)

Tool Description
get_cookies Get cookies (compact format)
set_cookies Set cookies
clear_cookies Clear all cookies
get_local_storage Get localStorage
set_local_storage Set localStorage key-value pairs
get_session_storage Get sessionStorage
set_session_storage Set sessionStorage key-value pairs
clear_storage Clear localStorage, sessionStorage, or both
save_storage_state Save cookies + localStorage to JSON file
load_storage_state Load cookies + localStorage from JSON file

Network Monitoring & Control (8)

Tool Description
wait_for_network_idle Wait for all network requests to finish
wait_for_url Wait for URL to match pattern
wait_for_response Wait for network response matching URL
wait_for_request Wait for network request matching URL
get_network_log Get captured network requests/responses
get_console_logs Get captured console messages
block_urls Block URLs by pattern (ads, trackers, etc.)
set_extra_headers Set custom HTTP headers

Device & Viewport (3)

Tool Description
set_viewport Set viewport size for responsive testing
emulate_device Emulate device (50+ presets)
list_devices List all available device presets

Geolocation & Permissions (3)

Tool Description
set_geolocation Override browser geolocation
grant_permissions Grant browser permissions
clear_permissions Clear all granted permissions

Frame/iFrame Support (3)

Tool Description
list_frames List all frames on the page
execute_in_frame Run JavaScript in a specific frame
click_in_frame Click element inside a frame

PDF (1)

Tool Description
save_as_pdf Save page as PDF (A4, Letter, etc.)

Visual Annotation — mark_page (5)

Tool Description
mark_page Annotate page with numbered boxes on interactive elements
unmark_page Remove annotations
click_element Click by index from mark_page
type_into_element Type into element by index from mark_page
mark_page_and_screenshot Mark + screenshot (best for visual agents)

Smart Tools (2)

Tool Description
get_page_summary Ultra-compact page summary (minimal tokens)
smart_action Natural language browser action

Browser Management (2)

Tool Description
get_browser_info Get browser config details
close_browser Close browser and cleanup

šŸ“± Supported Device Presets (50+)

iPhones: SE, 12 Mini, 12, 12 Pro, 12 Pro Max, 13, 13 Pro, 14, 14 Pro, 14 Pro Max, 15, 15 Pro, 15 Pro Max, 16 Pro Max

iPads: Mini, Air, Pro 11, Pro 12.9, 10th Gen

Android: Pixel 5, 6, 7, 8 Pro, Galaxy S21, S23 Ultra, Z Fold 5, A54, OnePlus 12, Moto G, Xiaomi 14, Huawei P60

Desktop: Chrome 1920x1080, Chrome 1440, Firefox, Safari, Edge, Small Laptop, Ultrawide

Landscape: iPhone Landscape, iPad Landscape, Android Landscape


šŸ”§ CLI Options

CONNECTION:
  --cdp <url>                 Connect via Chrome DevTools Protocol
  --user-data-dir <path>      Launch browser with user profile
  --executable-path <path>    Path to browser executable

BROWSER:
  --browser <engine>          chromium, firefox, or webkit (default: chromium)
  --headless                  Run in headless mode
  --timeout <ms>              Default action timeout (default: 30000)

VIEWPORT & DEVICE:
  --viewport <WxH>            Viewport size, e.g. 1920x1080
  --device <name>             Emulate device preset

NETWORK:
  --proxy-server <url>        Proxy URL
  --proxy-bypass <domains>    Domains to bypass proxy
  --ignore-https-errors       Ignore HTTPS certificate errors
  --block-service-workers     Block service workers

CONTEXT:
  --user-agent <string>       Custom user agent
  --locale <locale>           Browser locale (e.g. en-US)
  --timezone <tz>             Timezone (e.g. America/New_York)
  --color-scheme <scheme>     light, dark, or no-preference
  --geolocation <lat,lng>     Override geolocation
  --permissions <list>        Comma-separated permissions

SESSION:
  --storage-state <path>      Load session from JSON file

VIDEO:
  --record-video              Record browser session
  --video-dir <path>          Video output directory

Environment Variables

Variable CLI Equivalent
BROWSER_CDP_URL --cdp
BROWSER_USER_DATA_DIR --user-data-dir
BROWSER_EXECUTABLE_PATH --executable-path
BROWSER_ENGINE --browser
BROWSER_PROXY --proxy-server

🐳 Docker

docker build -t browser-mcp-server .
docker run -i browser-mcp-server
docker run -i browser-mcp-server --headless --browser chromium

šŸ—ļø Architecture

src/
ā”œā”€ā”€ index.ts           # CLI entry point, MCP server setup
ā”œā”€ā”€ browser-manager.ts # Browser lifecycle, multi-engine, stealth, tracking
ā”œā”€ā”€ tools.ts           # All 70 MCP tools with human-like behavior
ā”œā”€ā”€ resources.ts       # 6 MCP resources (page, cookies, console, network, etc.)
ā”œā”€ā”€ human.ts           # Human behavior engine (Bezier curves, natural typing)
ā”œā”€ā”€ devices.ts         # 50+ device emulation presets
ā”œā”€ā”€ utils.ts           # Shared utilities (selectors, accessibility, text extraction)
└── mark_page.js       # Page annotation script (numbered interactive elements)

šŸ”’ Anti-Detection Features

  • navigator.webdriver set to false
  • Chrome runtime spoofed
  • Permissions query overridden
  • Plugins and languages spoofed
  • WebGL vendor/renderer overridden (Intel Iris)
  • Canvas fingerprint noise injection
  • connection.rtt consistency
  • Automation command-line flags disabled
  • Human-like timing on all interactions

šŸ“Š Competitive Advantage

Feature This Server Playwright MCP (Microsoft) mcp-playwright Browserbase MCP
Human-like behavior āœ… Bezier curves āŒ āŒ āŒ
Stealth anti-detection āœ… Full suite āŒ āŒ āŒ
Session reuse (CDP) āœ… āœ… āŒ āŒ (cloud)
Multi-browser āœ… 3 engines āœ… āœ… āŒ
Device emulation āœ… 50+ presets āŒ āœ… 143 āŒ
Token optimization āœ… Compact outputs āŒ āŒ āŒ
Network monitoring āœ… Auto-capture āŒ āŒ āŒ
mark_page annotation āœ… āŒ āŒ āŒ
Smart natural actions āœ… āŒ āŒ āŒ
Console log capture āœ… Auto āœ… āŒ āŒ
Video recording āœ… āŒ āœ… āœ…
iframe support āœ… āœ… āŒ āŒ
Total tools 70 ~20 ~30 ~15

šŸ“„ License

MIT

Recommended Servers

playwright-mcp

playwright-mcp

A Model Context Protocol server that enables LLMs to interact with web pages through structured accessibility snapshots without requiring vision models or screenshots.

Official
Featured
TypeScript
Magic Component Platform (MCP)

Magic Component Platform (MCP)

An AI-powered tool that generates modern UI components from natural language descriptions, integrating with popular IDEs to streamline UI development workflow.

Official
Featured
Local
TypeScript
Audiense Insights MCP Server

Audiense Insights MCP Server

Enables interaction with Audiense Insights accounts via the Model Context Protocol, facilitating the extraction and analysis of marketing insights and audience data including demographics, behavior, and influencer engagement.

Official
Featured
Local
TypeScript
VeyraX MCP

VeyraX MCP

Single MCP tool to connect all your favorite tools: Gmail, Calendar and 40 more.

Official
Featured
Local
graphlit-mcp-server

graphlit-mcp-server

The Model Context Protocol (MCP) Server enables integration between MCP clients and the Graphlit service. Ingest anything from Slack to Gmail to podcast feeds, in addition to web crawling, into a Graphlit project - and then retrieve relevant contents from the MCP client.

Official
Featured
TypeScript
Kagi MCP Server

Kagi MCP Server

An MCP server that integrates Kagi search capabilities with Claude AI, enabling Claude to perform real-time web searches when answering questions that require up-to-date information.

Official
Featured
Python
E2B

E2B

Using MCP to run code via e2b.

Official
Featured
Neon Database

Neon Database

MCP server for interacting with Neon Management API and databases

Official
Featured
Exa Search

Exa Search

A Model Context Protocol (MCP) server lets AI assistants like Claude use the Exa AI Search API for web searches. This setup allows AI models to get real-time web information in a safe and controlled way.

Official
Featured
Qdrant Server

Qdrant Server

This repository is an example of how to create a MCP server for Qdrant, a vector search engine.

Official
Featured