- Transcription turns meeting audio into text.
- Intelligence turns text into summaries, titles, and chat responses.
How a meeting moves through AI
Recording, editing, and local storage do not require an AI model. A hosted Transcription provider receives audio. A hosted Intelligence provider receives the text and other context needed for that request.
Anarlog Cloud
Anarlog Cloud uses managed routes. Settings displays the managed selection as Pro (Cloud). Transcription uses that Cloud route. Intelligence offers Auto, displayed as Pro (Cloud), or a curated model choice. Auto lets Anarlog choose a suitable upstream model and retry a compatible provider when needed.The routes below describe the current implementation, not a permanently pinned model contract. Anarlog may update a managed route to improve quality, language coverage, latency, or availability. Select your own provider and model when you need a fixed model ID.
Cloud transcription
The managed Transcription route chooses a provider from the meeting’s languages and whether Anarlog needs live or after-recording transcription.- Soniox is the primary provider for live and after-recording transcription when it supports the selected languages. It uses Soniox 5:
stt-rt-v5live andstt-async-v5after recording. - Deepgram is an eligible fallback and resolves the route to Nova 3 or Nova 2, depending on language support.
- AssemblyAI currently resolves it to Universal 3.5 Pro:
universal-3-5-pro-realtimefor live transcription anduniversal-3-5-proafter recording. - The eligible fallback pool also includes configured routes for Gladia (currently Solaria 1), ElevenLabs, OpenAI, Mistral, and Alibaba Cloud Model Studio.
Cloud Intelligence
For text tasks, Anarlog → Auto currently routes summaries, chat, and generated titles through the~anthropic/claude-sonnet-latest alias on OpenRouter.
The alias tracks the current Claude Sonnet release instead of pinning a dated model. OpenRouter receives the prompt and chooses an available endpoint for that model.
Under Settings → Intelligence, Pro users can also choose Claude Sonnet, Claude Opus, GPT Sol, Gemini Pro, or Gemini Flash. Anarlog tries the selected model first, then its managed fallback route if needed. These choices track model families; use your own provider when you need a fixed model ID. Requests containing audio continue to use the audio route below.
Hosted Intelligence requests that contain audio use a separate ordered model pool:
google/gemini-3.1-pro-previewgoogle/gemini-3.6-flashmistralai/voxtral-small-24b-2507
On-device models
Anarlog currently offers three built-in on-device Transcription providers. Settings shows each provider only when it is supported by your Mac and release.
Parakeet Batch does not always provide word-level audio timestamps. When it returns text without a usable speech alignment, Anarlog keeps the individual words editable and groups the display by audio channel and processing chunk. The order between channels within a chunk is uncertain, and those words cannot be used for precise audio seeking. Speech-aligned words retain their timing.
You can also keep audio on your network by pointing Custom, or Deepgram’s Advanced base URL, at a Deepgram-compatible server on this computer or your local network.
Compatibility code still recognizes some older local model IDs, but they are not offered as new choices. The table above is the current user-facing model list.
Other model-backed features
Not every AI-adjacent feature is controlled by AI Settings:When chat uses web search, Anarlog sends the search query to its hosted research endpoint, which currently uses Exa, even when your Transcription and Intelligence models are local. The returned text can then become context for your selected Intelligence model.
Your own provider
When you configure a provider with your own API key:- Anarlog sends requests directly from the desktop app to the configured base URL.
- The provider and exact model shown in Settings handle the request.
- Anarlog stores the API key in your operating system’s secure credential store, not in the app database. Keys stay on that computer and do not sync to your other devices.
- The provider’s pricing, retention, and training policies apply.
muse-spark-1.3 and the other muse-spark models), configured with a Meta Model API key.
Recent bring-your-own Transcription models in Settings include:
- Google Gemini 3.5 Transcribe (
gemini-3.5-transcribe-liveduring the meeting,gemini-3.5-transcribeafter recording) - AssemblyAI Universal 3.5 Pro (
universal-3-5-pro-realtimeduring the meeting,universal-3-5-proafter recording) - Smallest AI Pulse (
pulse) and Pulse Pro (pulse-pro) for live and after-recording transcription - Meta Muse Voice Transcribe (
muse-voice-transcribe-1.0) for live and after-recording transcription with speaker labels - Nari Labs for live transcription with a Nari API key. Saved beta model IDs that ended in
:freemove to the current models. - Wispr Flow for transcription and dictation with your own API key. Its after-recording adapter accepts at most six minutes per request; choose another provider for longer files.
- Inworld, Gradium, Modulate, and Alebex for live transcription with your own provider credentials.
- NVIDIA Speech NIM for live transcription through your own deployed NIM server. NVIDIA’s hosted gRPC endpoint is not supported.
- Amazon Bedrock for after-recording transcription through an OpenAI-compatible gateway exposing
/audio/transcriptions. Native Bedrock endpoints are not supported.
Choose a setup
To compare on-device, Cloud, and bring-your-own-key cost, see How can I compare the costs of all the transcription options?. For every on-device path, see How can I use on-device models?. For a fully local checklist, see Use Anarlog offline. For the data sent by each feature, see Data, privacy, and retention.