browser_expand
Re-serialize one collapsed region, dedup template, or handle at higher detail. Use after browser_observe when a region was folded for the token budget. Requires render_mode 'agent' or 'both'.
When to use browser_expand
Use browser_expand when you need to render a page as a compact, handle-addressable view built for AI agents. It is part of Owl Browser's Agent Rendering toolset and runs inside a self-hosted, source-level stealth engine, so every call inherits the same undetectable browser fingerprint as the rest of your automation — no separate anti-detect setup required.
Usage Example
Parameters
Required
context_idstringrequiredThe unique identifier of the browser context (e.g., 'ctx_000001')
selectorstringrequiredThe OwlMark handle token of the collapsed region to expand, from a prior browser_observe (e.g. 'R1', 'b7'), or a dedup template id like 'T1'. Same token, same parameter, as browser_click/browser_type take.
Optional
detailenumminnormalfullDetail level for the expanded region
Response
Returns a JSON object with the operation result.
{
"success": true,
"result": <value>
}Frequently Asked Questions
What does browser_expand do?
Re-serialize one collapsed region, dedup template, or handle at higher detail. Use after browser_observe when a region was folded for the token budget. Requires render_mode 'agent' or 'both'. It belongs to Owl Browser's Agent Rendering category and is available through the REST API, the Python SDK (browser.expand()), the Node.js SDK, and the MCP server.
What parameters does browser_expand accept?
browser_expand accepts 2 required parameters (context_id, selector) and 1 optional parameter. All parameters are sent as JSON in a POST request to /api/execute/browser_expand.
Is browser_expand detectable by anti-bot systems like Cloudflare or DataDome?
No. browser_expand executes inside Owl Browser's Chromium engine, which applies fingerprint spoofing at the C++ source level rather than through JavaScript patches. Every tool call shares the same consistent, human-like fingerprint, so anti-bot systems such as Cloudflare, DataDome, and Akamai see an ordinary browser.
Related Tools
browser_observeAgent-native page observation. Returns the compacted OwlMark render (text-only structural view of the page), a handle table of interactive elements with stable tokens, page metadata, and a token estimate. Pass a handle token (e.g. 'b3') or 'pm:N' to browser_click/browser_type. Requires the context to be created with render_mode 'agent' or 'both'. ~20-100x fewer tokens than a screenshot for AI agent page understanding.
browser_read_nodeRead the raw accessible name, value, description, and attributes for one handle, uncompacted. Use after browser_observe to inspect a single element in full. Requires render_mode 'agent' or 'both'.
browser_create_contextCreate a new isolated browser context with its own cookies, storage, and optional proxy configuration. Each context acts as an independent browser session. Use this to create multiple isolated browsing sessions, configure proxy/Tor connections, load browser profiles with saved fingerprints, and enable/disable LLM features. Returns a context_id to use with other browser tools.
browser_navigateNavigate the browser to a specified URL. This is a non-blocking operation that starts navigation and returns immediately. Use browser_wait_for_network_idle or browser_wait_for_selector to wait for the page to fully load. Supports HTTP, HTTPS, file, and data URLs. When wait_until is set (load, networkidle, fullscroll, domcontentloaded) and the page declares WebMCP tools, the response includes a webmcp_tools array containing the full tool definitions (name, description, inputSchema). Use browser_webmcp_call_tool to execute any of these tools directly.
browser_clickClick on an element using CSS selector, XY coordinates, or natural language description. Supports semantic element finding using AI - describe what you want to click (e.g., 'login button', 'search icon') and the system will locate the right element. Simulates a real mouse click with proper event dispatch. Optionally hold the mouse button for press-and-hold interactions using hold_ms.
browser_typeType text into an input field with human-like keystroke simulation. Target the field using CSS selector, coordinates, or natural language (e.g., 'email field'). When selector is omitted, types into the currently focused element. Note: Does NOT clear existing content - use browser_clear_input first if you need to replace text rather than append.