Building Browser AI: Leo’s Development Progress and Plans | Brave

Authors

Matt McAlister

By Matt McAlister, Senior Director, Premium Products and Services at Brave

In late 2023 we published our first roadmap for Leo, Brave’s built-in AI browser assistant. In that post, we outlined our vision for Leo as a companion that helps you make sense of the Web, while still preserving your control and privacy.

Since then, Leo has been evolving from a helpful browsing companion toward a smart, personalized collaborator—a shift that reflects broader changes in how we think about AI and the Web. While we’ve delivered on our core promises around model quality, search capabilities, and multi-platform support, several of our most exciting innovations weren’t even on our original roadmap. Features like Multi-Tab Context, Tab Focus Mode, and Agentic AI emerged from user feedback and our ongoing commitment to making AI fundamental to how browsers work.

What hasn’t changed is our commitment to privacy-by-design and user choice. Given the product’s evolution since late 2023, we thought it was time to document Leo’s journey and share where we’re heading next.

What we’ve accomplished

Over the past 18 months, we’ve made progress with both foundational capabilities and innovative features:

Model improvements and performance

When we first launched Leo, we started with Llama 2 models. Since then, we’ve expanded our model offerings and given you more choice and options:

Learn more about the differences between Leo’s AI models .

Context and content

Going native

Learn how to use Brave Leo on desktop and mobile.

Learn how Leo gets current information with the Brave Search API.

User experience improvements

Privacy and security enhancements

We’ve added to our original privacy protections with better documentation and consistency in our messaging when new features are added:

Innovative new features: beyond the original roadmap

Building on this foundation, we’re currently working on new features that transform what a web browser can do:

Agentic AI…with guardrails

As the browser evolves with agentic AI features, it will become a more active partner in using the Web. We’re still refining the safety features before releasing our implementation, but it’s plain to see how helpful agentic AI will be, with the browser acting on your behalf to:

While some browsers give their AI access to your data and logged-in webpage sessions (such as your email, calendar, bank accounts, or social media accounts), Brave is working to ensure you retain full control of your AI’s activities. We’re implementing several safeguards and controls such as restrictions on access to extensions, separate storage partitions, and clear warnings and visual indicators while it’s activated. Our implementation will ensure you consent before letting AI act on your behalf.

Context management

Browser-based AI can work within your existing workflow when you choose to activate it, without requiring you to copy/paste and explain your intentions as you do with cloud-based or standalone AI chat apps. Brave’s multi-tab context feature lets you ask questions that span resources already open in your browser–PDFs, articles, spreadsheets, and more–whether you’re working on school assignments, company reports, or research projects.

Soon you’ll be able to add tabs, bookmarks and history as attachments by mentioning them with the “@” symbol, simplifying context management even further.

Improved context handling

We have also implemented many enhancements and continue improving the way content is processed to provide the highest quality insights and output:

Vision support

Leo now understands and analyzes images, bringing multi-modal capabilities to your browsing experience:

Tab organizer

Developers and business users often find themselves with numerous tabs open in multiple windows across client projects, research, and personal browsing. Rather than hunting manually through this digital mess, they now ask Leo to organize tabs by workflow stage, documentation type, language, or many other custom structures:

Learn how to use tab focus mode.

Looking ahead

As browsers evolve with more AI capabilities, they become the natural platform for AI to orchestrate your digital activities rather than simply assist with individual tasks. So, in addition to these ongoing improvements and upcoming releases, our near term plans include capabilities like:

Tasks and scheduling

Building on the agentic foundations referenced above, we’ll also add a task scheduling system that will let you:

Imagine setting up Leo to check for concert tickets within your price range, monitor GitHub issues that mention you, or collect news on topics you care about—all automatically and on a schedule you define.

On-device AI

We’re also exploring ways to expand on Leo’s initial BYOM capabilities for more on-device AI. This includes fully integrated local model support for running models, and enabling offline AI capabilities for tasks with even higher security needs on both desktop and mobile devices.

Better outputs

Using the latest models and tools such as the Brave Search API, users will be able to generate higher quality outputs. Product data, sports data, cryptocurrency prices, and richer outputs including images, tables, and other content formats will help you learn more about your personal interests. Audio mode will respond to you with a voice of your choice. You’ll also be able to create more robust and professional documents for work, school, or home.

Longer term

With these strong foundations in place, we’re changing the way we operate and deliver AI features in the browser.

We listen closely to what users are asking for. They’re telling us about what models they use, where to integrate them in the browser, and what would help them in their work. They’re telling us how to enable more control over AI.

We’re also our own power users, using AI tools internally to learn faster and achieve more. This experience drives our conviction that AI should work the way you do.

That’s exactly what we’re building: AI that adapts to your workflow, not the other way around, optimized with tools and services that help you learn faster, do more, and make better decisions. We’ll deliver this through a comprehensive AI operating environment, guided by your feedback, our focus on useful tools over flashy features, and our unwavering commitment to user choice and privacy.


Roadmap progress and future direction

As we reflect on our journey so far, the table below shows what we’ve delivered, what we haven’t completed yet, and how our understanding has evolved as we’ve released features and capabilities.

Strategic Direction Successfully achieved Incomplete Beyond original roadmap Future
Privacy-First Architecture - Opt-in consent
- Settings for disabling features
- Reverse proxy to anonymize IP addresses
- Zero data retention
- Privacy-preserving subscriptions
- Agentic AI guardrails (in development)
- Task progress indicators
- Activity previewer
- User intervention controls
- User consent mechanisms
- Access limitations
- Dedicated agent execution environment
- Enterprise group policy
- Chat history delete, disable
- Temporary chats
- More agentic AI guardrails
- Task management UI
- Permission management
- Activity/audit trails
- Rollback/undo
- More enterprise controls
- Agent-to-agent guardrails
- Auth for 3rd party APIs
- More controls for contexts, memories, models, local storage
- AI privacy policy summaries
- AI security warnings, alerts for suspicious ads, sites, network activity, etc
AI Model Infrastructure - Better models
- Vision model support
- Rate limiter
- Multi language support (limited number of languages)
- Integrated, pre-configured client-side models
- Many languages
- Image generation
- Bring your own model (BYOM) capability
- Rapid model deployment
- Performance improvements
- Reasoning model support
- Support for multiple images, screenshots
- Automatic model selection
- Automatic prompt routing
- MCP interfaces
- Integrated, pre-configured client-side models
- WebLLM exploration (in development)
- Improved model routing
- Multi-model, multi-agent orchestration
- More languages
- Text-to-speech
- Translation for live streams
- Image generation
Context & Understanding - Large context sizes
- Improved context extraction
- Context awareness for different content
- DOM tree parsing, accessibility tree
- Video transcript support
- PDF summarization
- Image understanding
- Brave Search augmented answers
- Video conference (Brave Talk) transcription understanding
- Entity highlighting
- Content suggestions/auto-completion
- User-defined boundaries, mood, and context awareness
- Customized tone and summary styles
- Drag and drop context
- Smart truncation
- Structured language handling
- Support for Google Docs
- Help Docs used to answer user questions
- Prompt classification (in development)
- Multiple attachments
- Screenshot upload
- Bookmarks (in development)
- Browser history (in development)
- Memory (in development)
- Larger contexts
- Context optimizations
- Specialized understanding (e.g. weather, cryptocurrency, products)
- Agent-to-agent context pipelines
- 3rd party APIs (Google, Notion, GitHub, etc.)
- Personalized, adaptive memories
- Automatic context updates
- Pre-emptive context loading
- Data, visualization, complex interface understanding
- Drag and drop context
- Live video stream comprehension
- Real-time grammar and style suggestions
- Intelligent query refinement and search suggestions
- Customized, adaptive AI personality and response styles
- Custom prompt templates and shortcuts
- User-editable system prompts
- Smart next action suggestions
- Personalized task suggestions
Productivity and utility - Right-click tools for text generation
- Conversation persistence
- Filling in editable fields
- Browser developer tools integration
- News integration
- Financial tools (e.g., cryptocurrency context)
- Plugin/extension framework
- Agentic actions (in development)
- Navigation
- Page interactions
- Form automation
- Tab controls
- Multi-site workflows
- Background tasks
- Multi-context conversations – tabs, bookmarks, history, uploads, etc.
- Right-click tools menu improvements
- Tab focus mode
- Conversation starters
- Editable prompts and responses
- Response regeneration
- Import SERP to Leo
- Chat with Brave Talk video transcription
- Writing assistant (in development)
- Smart Tasks
- Scheduling
- Progress tracking
- Task history
- Parallel tasks
- Auto-completion features
- Search tools
- MCP tools
- Browser developer tools
- Plugin/extension framework (incl MCP tools)
- Deep research (business analysis, academic papers, planning documents)
- Smart shopping
- Comparison tools
- Price tracking agents
- Suggestions
- Cart management
- Booking automation
- Personalized news insights
- Improved content generation (documents, code, images)
- Content generation versioning
- Smart bookmarks, browser history tools
User Experience - Full-page UI via brave://leo-ai/
- Sources and citations
- Mobile parity on Android and iOS
- Conversation UI
- Response formatting UI
- Code, markdown output (w copy/paste)
- Enhanced Feedback
- Enhanced mobile UI (tab-based)
- Full UI language translation
- Voice-controlled prompt entry
- Rich-media responses
- Voice responses with emotion, accents and laughing
- 3rd party content spaces
- Agent-to-agent management interfaces
- Cross-device sync
Business model - Premium subscriptions - Ad-supported option
- Brave Rewards integration
- Crypto payments for subscriptions (in development) - Ad-supported option
- Brave Rewards integration
- Affiliate partnerships
- Agent marketplaces
- Payments

For users who prefer not to use AI features, you can choose not to opt in, and if you wish to turn it off after trying it you can always disable Brave Leo completely.