The open-source library that lets AI agents actually click buttons on websites.
Browser Use is the open-source library that turned 'AI agents can browse the web' from a demo into a production reality. Released in late 2024, it became the de-facto primitive underneath dozens of AI agent products (Manus, Lindy, various research bots). It exposes any LLM to a real Chromium browser via a clean Python API, with DOM-as-text for navigation and Playwright for execution.
Who it's for: Engineers building AI agents that need to interact with real websites — scraping, form-filling, login flows, multi-step research. The go-to primitive under agents like Manus, Lindy, and dozens of vertical SaaS bots.
Converts the messy DOM into a clean, LLM-readable representation. The agent sees a structured page with element IDs, not raw HTML. This is the magic that makes web agents actually work.
Drives a real Chromium via Playwright. The agent can click, type, scroll, switch tabs, handle popups, and recover from errors. Real browser, not a fake one.
from browser_use import Agent; agent = Agent(task='...'); await agent.run(). That's it. Works with any LangChain or direct LLM call.
Hosted version with stealth proxies, captcha handling, and persistent browser sessions for production. Avoids the bot-detection nightmare of self-hosting headless Chromium.
If you're building an AI agent that needs to do anything on the web, Browser Use is the default primitive. Don't try to roll your own Playwright wrapper — the prompt-to-DOM conversion alone saves you weeks of debugging. For production, budget for the Cloud tier; the self-hosted version is fine for prototyping but gets detected quickly.
Web scraping API. Better for static content extraction, worse for interactive flows.
Neural search engine. Pre-built search results, not a browser automation layer.
Browser with built-in AI sidebar. Good for the human side of agent research.
Search API built for agents. Returns clean results without a browser session.