Multi-model workflows
Choose a model per chat, consult another model, or delegate a bounded task. Keep each model’s output and context visible.
Built for small local workloads and large, long-term projects. Learn agentic coding for free in Chrome, or use the Agent for multi-model orchestration.
Use local models within Locale to help the environment. Smaller models can use less energy and run without cloud compute.
I’ll create the page in your workspace and open a preview.
The homepage is ready to review.
A harness connects models to context, tools, and a workspace. For large-scale projects, use the Agent to coordinate model groups, delegate tasks, and retain project context.
Choose a model per chat, consult another model, or delegate a bounded task. Keep each model’s output and context visible.
Save a primary model, advisors, and shared instructions. Apply the group to project chats for a consistent workflow.
Organize chats, files, and tasks. Build in a browser Sandbox with live previews, or work in a Device folder with the Agent.
Plan an approach, implement with tools, or discuss a problem in read-only Ducky mode. Fork chats to explore alternatives.
Review command approvals and scope file access to selected folders. Inspect tool activity, changed files, and milestones.
Earn money from optional sponsored placements through Kickbacks.ai. Use local models to reduce paid inference costs. Earnings vary.
Model consultation, Agent Groups, connected providers, Kickbacks.ai, and Device tools require Locale Agent. Browser features remain available in the free trial.
Import text and binary files, edit code, preview HTML and app output, inspect the browser console, and export source as a ZIP. Browser snapshot history supports commits and branches. Device workspaces use actual disk folders and real Git through the Agent.
Search chats and projects, pin and archive conversations, fork a discussion, inspect reply versions, queue messages, and track tasks and milestones. Saved notes and context compaction help retain relevant information. Deep completion and bounded loops support longer work.
Choose and favorite models, compare observed throughput, track context usage, and set supported reasoning controls. Image attachments and image descriptions depend on model capabilities. With the Agent, consult other models or delegate bounded tasks with separate transcripts and inherited permissions.
The Agent adds file and command tools, provider credentials in the operating system credential manager, model management, GPU/VRAM readings, and encrypted local-network pairing with reconnect support. Experimental opt-in workers handle independent text requests on trusted local networks; they do not combine GPU memory into a single model.
Keep working in Locale from the comfort of your browser. The Windows Agent connects local Ollama models, project folders, approved commands, model groups, and Kickbacks.ai.
Pair your phone or another browser on your trusted local network to work with the same machine. The Agent runs the tools and model connections while your browser stays the interface.
Windows x64 · Portable · Keep the Agent open and your machine online. Current phone pairing uses your trusted local network; internet device access is planned.
Run LocaleAgent.exe. The Agent stays available in the system tray.
Open Locale in your browser. Pair another browser or your phone with the Agent’s code on your trusted network.
Connect local or provider models, select a project folder, and use tools from the browser. Review command approvals and keep the Agent open.
A free starting point for beginners learning agentic coding, with orchestration tools for power users managing larger projects.
Learn agentic coding with Chrome’s on-device Gemini Nano foundation model. Create files and preview small projects for free, with no Agent, account, or API key.
Use model groups and task delegation for larger projects. Locale Agent adds real project folders, approved commands, local Ollama models, and optional provider connections.
Chrome and local Ollama models run on your hardware. Smaller models can reduce resource use and reliance on remote servers. Environmental impact depends on hardware, workload, and energy source; carbon savings have not been measured.
Start with a local model. Add a specialist for review or a group for repeatable tasks. You control which providers receive your requests.
Gemini Nano, directly in supported desktop Chrome.
Run installed local models through the Agent on your own hardware.
Connect supported Zen or Go models with your own credentials.
Choose an xAI API key or supported Grok Build subscription sign-in.
Use installed local models or connect a provider through the Agent. These examples are separate from Cloud’s upcoming model catalog.
On-device in supported desktop Chrome.
Provider details ↗Local models from Locale’s model library. Hardware requirements vary.
Provider details ↗Examples from connected provider catalogs. Access depends on your account.
Provider details ↗Through the Agent with an API key or supported Grok Build sign-in.
Provider details ↗Provider availability, account requirements, usage limits, and fees apply separately. The free Agent supports your own provider connections. The approved plans add account-based inference credits to Sync and Cloud. Keep uses Chrome’s foundation model or your Agent connections. Cloud account inference is still in development.
Save model roles and instructions once, then reuse them across project chats. Consultation requests go only to the models you select.
Locale exposes tools according to the model, workspace, and work mode. Plan and Ducky restrict changes. Build carries out work with the required permissions.
List a workspace, find matching text, read selected lines, and write changes. Browser files stay in Sandbox; Device files stay within the folders you select.
list_dirread_fileread_linesload_filegrep_filewrite_fileRun JavaScript, open a page preview, inspect its controls, and test interactions. Supported image models can describe a preview so the model can review the result.
scratchrun_javascriptpreview_pageinspect_previewwebsite_previewReview changes, save snapshots, create branches, and export your files. Browser snapshot history stays separate from real Git repositories on your machine.
sandbox_gitbrowser_exportMaintain a task list, record milestones and project notes, check context usage, and ask for decisions when work needs your input.
task_listtask_updatemilestoneproject_notecontext_usageask_userRequest a second opinion, delegate a bounded task, and collect its result. Reusable Agent Groups define the primary model, advisors, and shared instructions.
consult_modeldelegate_taskagent_tasksRun shell, PowerShell, or Bun commands in your project and initialize real Git repositories. Commands follow Locale’s approval rules and workspace permissions.
run_commandrun_powershellrun_buninit_gitCheck installed models, download an Ollama model, and send work to a local model. Model setup and inference run on the computer hosting your Agent.
ollama_statusollama_pullollama_chatDescribe image attachments with a supported model. The Agent adds web search when current information is needed. Web access requires an internet connection.
describe_imageweb_searchThe free Chrome version uses the foundation model and browser tools. Agent connections add local and connected models, device tools, and Kickbacks.ai. Sync and Cloud include account inference credits at launch. Cloud adds a worker. Keep uses Chrome Nano or your own model connection. Cloud services are in development.
Local use is free. Choose the cloud services you need.
Cloud services in development. Mock checkout only, no charges.Store your work and publish a personal site.
$5 every month
Site files and databases use this storage.
Use Chrome Nano or your own model connection.
Account inference, websites, and device sync.
Monthly subscription
Site files and databases use this storage.
Use cloud models through your Locale account. Device tools require an open Agent.
Models and a worker for independent cloud tasks.
Monthly subscription
Site files and databases use this storage.
One worker: 2 shared vCPUs, 4 GB RAM, at least 12 GB working space. Pauses when idle.
Inference through your account. Sync and Cloud include model access. 100 inference credits cover $1 of published model usage. Tokens used for input, output, and reasoning draw from the same balance. Model availability varies by provider. Device files and local models require an open Agent.
Sites with room to grow. Publish at a shareable HTTPS URL. *Database, site sign-in, and site credit allowances are planned launch targets. They cover managed backend functions, not arbitrary server processes. Access is static-only.
100 site credits represent $1 of metered backend usage, separate from inference. Planned optional top-up: $5 for 400 site credits. No automatic charges. Site request limits reset each Monday at 00:00 UTC. Credits, active-user counts, and worker hours reset monthly without rollover. Purchased site credits are planned to last 12 months.
Allowances are per account across its sites. Each site supports up to 100 MB, files up to 20 MB, counted within storage. Site files, database contents and indexes count toward your cloud storage allocation. Database scans, writes, and function execution consume site credits. New work pauses at the limit. Site accounts, metering, publishing, and cloud services are not active yet.
The browser harness and Windows Agent remain free.
The static browser version uses only Chrome’s foundation model. It supports chat and browser workspaces, including files, previews, tasks, and export. It cannot connect to Ollama, external model providers, device folders, Kickbacks.ai, or remote sync.
No. It needs desktop Chrome with an available Prompt API and supported hardware. Chrome may download its model the first time you use it. Unsupported browsers show setup information instead of sending your prompt to a cloud fallback.
Browser-only chats and workspace files are saved in your browser. Local Ollama inference runs on your machine through the Agent. If you select an online provider, that provider receives the context sent to it. Cloud storage will be a separate subscription service.
Downloaded Chrome and local Ollama models can perform inference offline. The website and required assets must already be available. Downloads, connected providers, web tools, Kickbacks.ai, and future cloud services need an internet connection.
Cloud is $25/month for 256 GB, sync, ten static sites, 800 monthly inference credits ($8 of model usage), and 60 small Box VM hours. Cloud services remain in development. Agent Groups and your own provider connections remain available with the free Agent.
Sync is $10/month with 50 GB, device sync, three static sites, and 300 monthly inference credits ($3) through your Locale account. Cloud adds a larger model budget and a Box worker. Device tools still need the Agent. These offers are not active yet.
Keep is $5 every month for 10 GB and a Personal site, without device sync or inference credits. Model tasks require Chrome’s foundation model or an Agent connection.
Agent tools and local model connections stop being available. Keep the Agent open and the host online for its models, device tools, and paired access. The browser version runs independently. Planned Cloud work can continue without a local Agent.
You can create an account now. Approved offers are available to review in mock checkout. No payment is collected or cloud service activated. The browser version and Agent remain free.