Β·
STT
β€”
THINK
β€”
TTS
β€”
GPUβ€”
idle
LANG
VOICE
SPEED 1x
TONE
ready

Live transcription

press REC and speak…
idle

Self-talk reasoning

step-by-step Β· same sequence
idle
LIVE process flow β€” every step lights while it runs on ANY tab or device (server-driven), visualizer trace-mode style.
🎀 MIC
device microphone Β· mob/pc
VAD
voice gate Β· chunker
STT
β€”
SPEAKER ID
voice match Β· runs with STT
THINK
β€”
TTS
β€”
πŸ”Š OUT
audio playback
HEARD (STT)
β€”
REPLY (THINKING)
β€”
VOICE (TTS)
β€”
TRACE β€” server pipeline timeline (Ξ”ms between stages)
β€”
total latencyβ€”

Activity feed

stage transitions appear here live…

Active pipeline (tap a slot β†’ highlight β†’ pick ANY model to substitute)

TTS β–²β€”
THINKINGβ€”
STT β–Όβ€”
Open chat context usage β€”

Models (type = instant local filter Β· GO = search Ollama + HuggingFace)

Thinking models (LLM)

STT models (speech β†’ text)

TTS models (text β†’ speech)

Generated voice storage

WHO IS TALKING

Each profile = a person. Enroll 3+ voice samples per person. Gabriel recognizes them automatically and adapts its response voice.

GPU

loading…

CPU / RAM / DISK

loading…

Server log


      

System prompt

Voice

Speaker recognition

Lower = merges more aggressively (one person stays one profile). Higher = stricter (may split a person into many). Defaults: match 55, cluster 50.

Version

App β€”
Server β€”

Password

Notifications

β€”

Server

Permanent memory accumulated from chats & live (RAG). Editable. Facts inject into future conversations.
load to view…
0
🧠 COLLECTED MEMORY (this session)