Zuvra

a quiet, private AI companion

Your own AI, living in your computer — not in a data center. No accounts. No cloud. No one listening.

Get the engine
1.08B on-device model 131,072 context tokens 100% local & private Apache-2.0 model Free to use

First sight

Your first screen. Calm, empty, waiting.

The Zuvra screen

Why Zuvra

Six quiet luxuries — crafted, private, and already yours.

Privacy, built in

Chats live only on this device. One tap deletes — wiped from memory, screen, and disk. Truly gone.

Watch it think

Reasoning unfolds live in its own panel. Follow it, read the answer, or stop it mid-thought.

Answers, as they happen

Words stream live; your tab announces each reply.

Every token, accounted for

Each reply shows its token cost. Hover for the breakdown.

Your past, at hand

Search, rename, organize — in a drawer that never interrupts.

Code highlighting

Code highlighting for 23 languages and formats. Python · JavaScript · TypeScript · Java · C · C++ · C# · Go · Rust · PHP · Ruby · Swift · Kotlin · SQL · HTML · CSS · Shell · JSON · YAML · TOML · INI · Dockerfile · Diff. Copy any block with one tap — HTML and JavaScript open as runnable previews.

Run it on your machine in 3 steps

The chat talks only to your computer. Its engine is the free MiniCPM Desk Pet — install it, keep it running, done.

1Get the engine

Download the official desktop app (v0.10.0). One click starts the download from GitHub Releases:

Detecting your device…

All versions: official releases page.

2Install & run it

Install, launch, leave it running. Settings and models survive upgrades.

3Open your chat

Download Zuvra above, open Zuvra.html — the same interface, no internet needed. Nothing is ever uploaded.

Requirement, plainly stated: the chat file needs MiniCPM Desk Pet running (the desktop app and local engine). Without it you will see “Offline — open MiniCPM Desk Pet and retry.” The chat itself never contacts any external server.
Take everything with you: one-click copy everywhere, runnable previews, searchable history. On your device, free.
Answers begin at onceThinking first, live — never a blank screen.
Light on your device1.08B Q8 — sips CPU and RAM on ordinary computers.
Zero network waitingOn-device: no queues, no accounts. Speed depends on your hardware.
Small but serious1B-class SOTA (per OpenBMB), hybrid reasoning — more with less.
Built with MiniCPM

Powered inside by MiniCPM 5

Zuvra's inner engine is MiniCPM5-1B — a 1.08B on-device model by OpenBMB, running locally as a GGUF quant. 131,072 tokens of memory in a 1B body. Text only — it understands written words, not images, audio, or video.

1.08Bparameters 131Kcontext length Q8_0local GGUF quant Thinkhybrid reasoning Textonly · no images
License & attribution. MiniCPM5-1B © OpenBMB — this repository and the MiniCPM model weights are released under the Apache License, Version 2.0. You may obtain a copy of the License at http://www.apache.org/licenses/LICENSE-2.0. Model: openbmb/MiniCPM5-1B · Source: OpenBMB/MiniCPM · Paper: MiniCPM Tech Report · Desk Pet app: MiniCPM-Desk-Pet (AGPL-3.0). Built by ModelBest (Beijing) with Tsinghua NLP and the OpenBMB open-source community. The Zuvra interface is a separate original work; the model weights run locally on-device and this page distributes no model files.