Skip to main content

Settings

Phlox is configured almost entirely through the in-app Settings page. Configuration lives in the encrypted database (not in files), so changes persist with your data. Most settings autosave.

Settings is divided into two top-level panels: User Settings (per user) and Admin Settings, which is only visible to admins — global model endpoints, system prompts, and policies are admin-only writes (see Users & Authentication). On the single-user desktop app you are always the admin, so everything shows.

Settings pageSettings page

User Settings​

TabWhat you configure
GeneralYour name and specialty (provided as context to the LLM and used for specialty-aware suggestions); your default note template and default letter template; your preferred language (per user — see Language Support).
Note TemplatesCreate, edit, and delete note templates.
Letter TemplatesManage correspondence templates, including the protected "Dictation" template.
Quick ChatConfigure up to three quick-chat buttons (label + prompt) shown in the chat surfaces.

Language​

The General tab also sets your preferred language — a per-user setting that drives transcription, note/letter/chat output, and date formatting. On desktop, selecting a non-English language offers to download a multilingual transcription model (available once the language appears in the selector). See Language Support for details and limitations.

Admin Settings​

The Admin Settings panel (formerly "Model Settings") contains the model configuration and the Users tab; it is only visible to admins. It behaves differently depending on whether you run a desktop (Tauri) build or a Docker/web build.

Users​

Admin-only. Create user accounts (admin or clinician), reset passwords, and disable accounts. See Users & Authentication.

Local mode (desktop only)​

On desktop you can run inference locally. The panel toggles between Local and Remote inference.

  • Models tab — download and manage the bundled models: the primary AI model (with smart recommendations based on your RAM/cores, plus an estimated processing time), the transcription model (Omi Med STT v1), and the optional embeddings model (Qwen3-Embedding) used for RAG. A system-specs banner shows your available RAM and CPU cores, and a "Reset All Models" action is available.

Local Model ManagerLocal Model Manager

Remote mode​

When using external endpoints (required for Docker; optional on desktop):

  • Whisper tab — transcription API base URL, model name, and API key. A live status indicator shows endpoint reachability.
  • LLM tab — OpenAI-compatible / Ollama base URL, API key, primary model, secondary model, the document/image processing mode, a vision capability probe, and Letter Generation Temperature (0–2, default 0.6 — the sampling temperature used when generating letters; higher values produce more varied writing, with a reset-to-default).
  • RAG tab — choose the embedding model. Changing it triggers a re-embedding pass across all document collections, with a progress indicator.
  • Tools tab — see Tools below.

Policy​

Practice-level system policy, applied to every user:

  • Store Original PDFs — keep uploaded PDF binaries in the database (increases storage usage).
  • Require patient consent for ambient scribing — prompt for the patient's consent before any consultation recording (Ambient or Live agent). Consent is remembered per patient. See Patients → Ambient Scribe Consent.
  • Streaming capture — transcribe Ambient and Dictate recordings utterance-by-utterance while recording, with speaker labels from Phlox itself, so the note is ready faster at stop. Docker-only setting (always on in the desktop app); see Streaming capture.
  • Prompt cache warming — pre-warm the note-generation prompt with the transcript as it grows. On by default with the bundled local model; enabling it overrides that for any endpoint. Useful for self-hosted inference (vLLM, Ollama, llama.cpp servers), but be aware it multiplies billed input tokens on metered cloud APIs.

Document and Image Processing Mode​

Controls how uploaded PDFs and images are processed for notes, demographics, and document chat:

  • Auto (default) — probes the model for vision support and uses vision if available, falling back to text extraction then OCR.
  • Vision only — always sends page images to the vision model. Requires a vision-capable model.
  • OCR only — text extraction (pypdf) with Tesseract OCR fallback. Works with any model but may miss content in scanned documents or images.

In desktop (local) builds, vision is enabled automatically when a multimodal projector (mmproj) file is present alongside the model.

Vision Capability Probe​

The Test Vision Support button sends a tiny test image to the configured model and reports whether it is vision-capable. The result is cached per provider/URL/model and shown as a badge (Vision capable / Not vision-capable / Unknown, with source and timestamp).

Tools​

The Tools tab controls the agentic tool-calling system.

Built-in Tools​

Toggle each built-in tool on or off. Tools that call external services are flagged and ship off by default:

  • Transcript Search, Literature Search, Previous Encounters, Patient Search, Search by Condition, Patient Note Search, Patient Jobs, Outstanding Jobs, Job Completion, Note Creation, Todo List, PDF form tools
  • ⚠️ PubMed Search and Wikipedia Search — disabled by default (external; may transmit PHI)

Tool Servers (MCP)​

Connect external tool servers via the Model Context Protocol (SSE transport):

  1. Click Add Server and provide a name and HTTP URL.
  2. Optionally enable Allow sensitive data (PHI) per server. When off (the default), Phlox strips identifying information from arguments before calling the server. See Security → Tool & MCP data safety.
  3. Test connection to verify reachability and discover the server's tools.
  4. Toggle the server on/off as needed.

Discovered MCP tools appear alongside the built-in tools in chat.

System Prompts​

Power users can customise the system prompts Phlox sends to the LLM. The editors live as tabs inside the Admin Settings panel — Chat (chat interactions) and Summary (patient summary generation) — each an editable system-prompt text area with a Reset to Default button.

A warning in these tabs reminds you that the defaults are carefully tuned. Change them only if you understand the impact on output quality.