WebBrain icon

WebBrain

WebBrain is a free, open-source AI browser extension for Chrome, Edge, and Firefox that helps users read, extract, and automate web pages. It supports local and cloud LLMs, multilingual UI, and a read-only Ask Mode alongside an action-oriented Act Mode.

WebBrain

What WebBrain is

WebBrain is a free, open-source AI browser extension for Chrome, Edge, and Firefox that lets you read, extract, and automate web pages from a sidebar inside the browser. It is presented as a self-hostable alternative to proprietary browser AI plugins, including Claude’s browser plugin.

The product is built around two modes: Ask Mode for read-only questions, summaries, and extraction from the current page, and Act Mode for browser actions such as clicking, typing, scrolling, navigating, and running scripts. The site emphasizes control and safety by default, including prompt-injection defenses, stricter handling for sensitive fields, and confirmation before consequential actions.

WebBrain supports multiple LLM providers, including local and cloud options, so users can choose a model that fits their privacy, latency, or cost needs. The site also highlights multilingual UI support and documentation for using WebBrain with MCP, extending it into a tool that can fit both individual browsing tasks and more structured automation workflows.

Features

Page understanding

Read articles, docs, dashboards, and forms from the current page and get answers based on what is already visible in the browser.

Ask and act workflows

Use natural-language instructions to click, type, scroll, navigate, and interact with pages. The extension asks before consequential actions, and Ask Mode stays read-only by default.

Data extraction

Extract structured data from tables, lists, links, forms, and PDFs, with examples on the site showing product catalogs and search results being captured into usable output.

Multi-provider LLM support

Connect to local or cloud models through llama.cpp, OpenAI, Claude, OpenRouter, Ollama, LM Studio, vLLM, Grok, Gemini, DeepSeek, or Mistral, and switch providers from extension settings.

Multilingual UI

Use English, Español, Français, Türkçe, or 中文, with automatic language detection on first use and manual switching from the sidebar.

Privacy and context controls

Run with a local model for reduced data exposure. The site also notes token-conscious screenshots, automatic context trimming, and optional profile autofill stored locally in the browser.

Use cases

  • Read and summarize pages

    Ask questions about a news article, documentation page, or dashboard and get answers based on the page content without leaving the browser.

  • Extract data from web pages

    Pull product names, prices, links, or form fields from tables and lists, including pages and PDFs that need structured extraction.

  • Automate browser workflows

    Fill out multi-step forms and other repetitive browser tasks by instructing the agent to click, type, scroll, and navigate for you.

  • Choose the model setup you need

    Use a local model when you want the browser agent to stay offline, or connect a hosted provider when convenience or stronger model quality matters more.

  • Work in multilingual or developer setups

    Use the extension in supported languages and pair it with MCP-based workflows when you want browser actions to fit into a larger toolchain.

Pros and Cons

Pros

  • Free, open-source, and self-hostable, with no telemetry, no tracking, and no accounts required.
  • Supports both local and cloud LLMs, letting users choose between offline privacy and hosted convenience.
  • Covers multiple browser tasks in one extension: reading, summarizing, extracting structured data, and acting on pages.
  • Works with Chrome, Edge, Firefox, and other Chromium-compatible browsers.
  • Provides multilingual UI support and auto-detects the browser language on first use.

Cons

  • The site warns that local models should have enough context window for best results, and notes that 16k is recommended while 8k can work in Compact mode.
  • CAPTCHA solving is optional and requires a BYO CapSolver API key; it is not shipped or enabled by default.
  • Some actions are intentionally gated: WebBrain starts in read-only Ask Mode and asks before consequential actions, so users still need to review sensitive steps.

FAQ

Which browsers does WebBrain support?

WebBrain is available as a browser extension for Chrome, Edge, and Firefox. The site says it also works with Brave, Opera, Vivaldi, and other Chromium-compatible browsers.

What can I do in Ask Mode?

It starts in Ask Mode, which is read-only. In that mode you can ask questions about the current page, extract information, and summarize content without changing the page.

What is Act Mode used for?

Act Mode gives the agent full browser actions: clicking, typing, scrolling, navigating, and running scripts to complete multi-step workflows.

Which AI providers can WebBrain use?

The site says WebBrain works with local llama.cpp and OpenAI, Claude, and OpenRouter, and it also lists Ollama, LM Studio, vLLM, Grok, Gemini, DeepSeek, and Mistral as supported provider options.

Is WebBrain free and self-hostable?

Yes. The site describes WebBrain as free, open-source, and self-hostable, with no telemetry, no tracking, and no accounts required.

Quick Facts

Category
AI browser agent / browser extension
Platforms
Chrome, Edge, Firefox; also works with Brave, Opera, Vivaldi, and other Chromium-compatible browsers
Pricing
Free forever, open-source
Integrations
Local and cloud LLMs; MCP documentation available
Source domain
webbrain.one
Primary use
Read, extract, and automate web pages