browser-use/web-ui

🖥️ Run AI Agent in your browser.

What it solves

This project provides a user-friendly graphical interface for the browser-use library, allowing users to interact with AI agents that can navigate and use websites. It removes the need for writing code to trigger agentic browser automation tasks.

How it works

Built on Gradio, the WebUI serves as a control panel for the browser agent. It integrates with various Large Language Models (LLMs) such as Google, OpenAI, Azure OpenAI, Anthropic, DeepSeek, and Ollama. Users can configure the agent to use a custom local browser instance to maintain authentication sessions and avoid re-logging into websites.

Who it’s for

Users who want to operate AI browser agents without needing to write Python scripts, as well as developers who want a visual way to monitor and manage their browser-automation agents.

Highlights

  • Broad LLM Support: Compatible with a wide range of providers including DeepSeek, Anthropic, and Ollama.
  • Custom Browser Integration: Ability to use a local browser executable and user data directory to bypass authentication hurdles.
  • Persistent Sessions: Option to keep the browser window open between different AI tasks to maintain state.
  • Detailed Monitoring: Supports high-definition screen recording and VNC viewer access when running via Docker.

Related

  • Project
  • Project
  • Project
  • Project
  • Project