Live transcription
press REC and speakβ¦
Self-talk reasoning
step-by-step Β· same sequenceLIVE process flow β every step lights while it runs on ANY tab or device
(server-driven), visualizer trace-mode style.
π€ MIC
device microphone Β· mob/pc
VAD
voice gate Β· chunker
STT
β
SPEAKER ID
voice match Β· runs with STT
THINK
β
TTS
β
π OUT
audio playback
HEARD (STT)
β
REPLY (THINKING)
β
VOICE (TTS)
β
TRACE β server pipeline timeline (Ξms between stages)
β
total latencyβ
Activity feed
stage transitions appear here liveβ¦
Active pipeline (tap a slot β highlight β pick ANY model to substitute)
TTS β²β
THINKINGβ
STT
βΌβ
Open chat context usage
β
Models (type = instant local filter Β· GO = search Ollama + HuggingFace)
Thinking models (LLM)
STT models (speech β text)
TTS models (text β speech)
Generated voice storage
WHO IS TALKING
Each profile = a person. Enroll 3+ voice samples per person.
Gabriel recognizes them automatically and adapts its response voice.
GPU
loadingβ¦
CPU / RAM / DISK
loadingβ¦
Server log
System prompt
Voice
Speaker recognition
Lower = merges more aggressively (one person
stays one profile). Higher = stricter (may split a person into many). Defaults: match 55, cluster 50.
Version
App β
Server β
Password
Notifications
β
Server
Permanent memory accumulated from chats & live
(RAG). Editable. Facts inject into future conversations.
load to viewβ¦