PAVONA

Your AI workspace, on your terms.

Chat, write code, and work with files in one place. Choose the model and decide what its tools may do.

Windows beta · Linux CLI setup

PAVONA's local workspace with conversation starters, model choice, and a permission control beside the message box.
Inside PAVONANight appearance · local candidate 446af1601eca · conversation, models, and permissions.Explore the workspace
Appearance Explore all 12 themes
01Find a guide or check your computerSearch the handbook, or see what fits your memory and system

Getting started

Find the guide you need.

These tools run in your browser. Your entries are not sent anywhere.

02What it doesChat, agents, models and spending limits, in one window

01 / Your workspace

A look at
the workspace.

The workspace keeps your conversations, models and agents in one place.

Inside the workspace
PAVONA Interface preview
Where chats are keptLocal model
Where does this conversation go when I close the window?
PAVONA · example reply

It stays on this machine.

Your chat history is stored locally, in a folder you can open. Cloud models and connected tools send the content they need to the services you choose.

01History you can readPlain JSON on your own disk
02Connections you chooseLocal inference or your cloud provider
Illustrative interface · example conversation, not a live session.

Features

Chat, agents, spending limits and model discovery.

01 · Conversation

Chat with
the model you choose.

Pin the model you want. Your history stays on your disk, and tools can do only what you allow.

Local modelsCloud providersPermission modes
How chat works

The Instructions box ships empty. Your instructions are sent as written. Tools add the descriptions a model needs to use them. You can inspect the request in the app under “What the model actually receives”.

You choose a permission mode for each conversation, from Plan (nothing changes) to Bypass (everything runs). Read the handbook.

02 · Assistants

Give each agent
its own model, tools and instructions.

ModelToolsInstructions

Each agent works from instructions in a file you own. Choose its model and tools to fit the work, then set its permissions.

Inside an agent

Instructions are plain-text files you can edit. Each agent has a model, a colour and a configured tool set. Review those settings before allowing an agent to act on your files. Read about the controls.

03 · Spending

Set a spending limit
for cloud models.

Paid calls stay off until you set a monthly budget. Local inference has no provider fee.

Usage and limits

Usage is recorded from the provider's own figures. PAVONA warns you at four fifths of the limit and stops paid calls when you reach it. Provider billing can lag, so check your provider account too.

04 · Discovery

Find new models
inside the app.

Discover models from Ollama and OpenRouter without waiting for a new app release.

Model discovery and checks

Local model names come from Ollama, and cloud names come from the live OpenRouter catalogue. Availability depends on the provider and your account. The included prompt-injection probes are a small mechanical check, not a certification.

03For companiesPilot PAVONA on your own workflow, with the terms agreed first

PAVONA for companies

Try PAVONA
with your team.

Start with a workflow your team already runs. Agree who owns it, what data it can use and how you will judge the result.

Pilot scope, pricing and support are agreed before work begins.

01 / DevAd

Build, check and review.

DevAd builds in a separate copy of the code. You decide what success means before a change is promoted.

Engineering workflow

Evaluate front ends, back ends and client-specific software against your acceptance criteria.

  • Readable plans and a change history
  • Checks and review before promotion
  • Reports that name unfinished work
How DevAd works

02 / VERIDITAS

Security checks for what you own.

Use it for defensive checks and authorised testing. Written scope and asset-owner permission come first.

Security workflow

Agree the assets, permitted actions and evidence to retain before testing begins.

  • Permission and data-flow review
  • Release integrity and local checks
  • Findings linked to a proposed fix
Review permissions

03 / Support

Support your team.

Set up local assistants for your team. You review each problem report before anything is sent.

Support workflow

The app composes reports with version and system details for you to inspect.

  • Instructions your team can maintain
  • A report you inspect before sending
  • Release notes that explain changes
See the support path
What to establish before an enterprise rollout

PAVONA is a Windows beta. Its enterprise policy catalogue, provider restrictions and audit records cover specific execution paths; they are not a complete managed enterprise service. Central administration, SSO, tenant isolation, contractual support and deployment requirements need a separate review for your use case.

Security testing requires the asset owner's authorisation. No certification, compliance status or third-party partnership is claimed here. A pilot should produce evidence and a list of remaining gaps, not an assumed approval.

Privacy

What stays on your computer,
and what leaves it.

Local models run on your hardware. Cloud models and connected tools send data to the services you configure.

Read the network details

Web research, model downloads, cloud providers and other connected tools use the network. Keeping files on your own disk does not protect them from other people or software on the same machine.

The update check reads the release list every six hours while the app is open. Updates install only with your confirmation. What leaves your computer.

04Set it upWindows today: three steps from download to first conversation

From download to first conversation

Set up PAVONA
on Windows.

Installation guide
System
Windows 10 or 11
Free disk space
About 8 GB
Memory
8 GB to start
Graphics card
Optional
  1. 1

    Download

    One file, 48 MB. Get the current installer and its checksum from the release page.

  2. 2

    Run setup

    Double-click it. Setup downloads the runtime and a first model: about 2 to 7 GB.

  3. 3

    Open the app

    Choose your model and start a conversation.

Hardware, model fit and the Windows warning

PAVONA is a Windows beta today. Linux Mint is in validation, and native Apple validation is still pending. About 8 GB of disk space covers the installer, the runtime and a first model of 1 to 5 GB. More models need more room.

A model needs about its own file size in free memory, plus a gigabyte for the conversation. With 8 GB, start with a small model. A graphics card helps but is optional.

Setup measures model fit and reports it as runs well, runs acceptably, runs with limitations or cannot run. Unknown results are labelled.

The installer is unsigned, so Windows may warn once. Check the checksum, then decide. The install guide explains the steps.

05Benchmarks and updatesResults against other agents on the same tasks, and updates that install only when you approve them

New releases

How updates
reach you.

While the app is open, it checks the release list every six hours. Nothing installs until you approve it.

Read the patch notes
Updates and problem reports

Patch notes appear in the News panel. “Report a problem” writes a report that you read before it is sent anywhere. Get support.

Benchmarks

Compare the results
and their limits.

Every agent gets the same tasks and the same checkers. The results were recorded on one laptop and are not a promise for yours.

See the results
Open the recorded comparison

Measured, not claimed

Against the others, on the same tasks

Recorded task outcomes from 2026-09-24. These exploratory results do not establish product superiority: matching model and resource budgets have not been verified. It trails on time per verified task, and the team's standing card is to close that gap and measure again. The whole page, with every task and every weakness.

AgentTasks passedSeconds per verified taskFiles left behind per task
PAVONA94%43.090.057
Codex94%107.1470.371
Claude Code91%38.9140.114
OpenClaw91%31.8620.057
Aider20%63.9390.971
Crush83%29.0133.143
Goose89%43.9090.371

"Not run: billing" (or auth, model, engine) means the agent's own program refused every task for that reason and nothing was measured. "No verified work" means the agent passed no task, so its speed is not a figure. "Not measured" means the harness cannot observe that figure from outside an external agent.

Benchmark comparison

Latest published measurements

Select a category, then focus or tap a bar to see its evidence. Highlighted segments show a measured advantage.

Loading the published comparison…

Read the recorded benchmark tables.

Published results refresh here every minute. A new score appears only after a benchmark has finished and been published. The animation does not mean there is a new measurement.

Check it yourself

Read the source and run the tests.

The download includes readable source and the test suite.

Test runner · command, not a result
python run-tests.py

Some tests use models, the network or local files. This page does not claim a passing run.

Open the handbook

Free while in beta

Download the beta.

Your beta copy stays yours for good.
The source is readable and comes under the PAVONA licence. It is not open source.

Now playing in the corner: Feast Your Eyes by Mastodon