← Command Center

Voice Session

J.A.R.V.I.S. Voice Layer

Command Core

Reference arc-reactor

STT — speech capturevoice-to-voice sessionsTTS — agent voice synthesisprovider synthesis approval-gated

Primary interface

Talk to J.A.R.V.I.S. here

Voice and typed prompts use the protected J.A.R.V.I.S. backend, approval queue, scoped “Jarvis approved” flow, and spoken/browser speech response path so this screen can replace the external chat for day-to-day command.

Real-time voice

J.A.R.V.I.S. Conversation Mode

○ idle

Command Core

Reference arc-reactor

Press to activate

Uses the same protected backend, approval tickets, and response path as voice.

Device capture scaffold

Microphone Permission Console

Requests browser microphone access only after Stufio taps the button. This local scaffold checks device support, captures an active stream, previews audio level, attempts browser-native STT, and keeps paid STT/TTS providers behind server-side token gates.

mic unknown

Secure context

missing

MediaDevices microphone

unknown

MediaRecorder scaffold

unknown

Browser STT

unknown

Native TTS playback

unknown

Active input

No microphone stream opened yet

Local level meter: 0%

STT transcript preview

idle

Transcript preview will appear here when browser STT is available.

Safe local scaffold: nothing is uploaded, synthesized, stored, or sent to a provider.

Safety rule: browser microphone access is user-triggered and revocable. Provider STT/TTS integration must run through server routes with environment variables; keys are never hardcoded in client code or printed to logs.

Provider Adapter Readiness

Server-side STT/TTS gates

This panel shows which provider adapters are selected, which environment variable names they expect, and which safety gates are closed. Token values stay server-side and are never rendered in the browser.

token values hidden

STT adapter

openai

missing env
Token env name
OPENAI_API_KEY
Token configured
no
Adapter gate
enabled
Live-call approval
present
Can call provider
no

Server-side placeholder only. The adapter reports environment variable names and gate state, never token values, and it does not make live STT/TTS provider calls.

TTS adapter

fishaudio

placeholder blocked
Token env name
FISHAUDIO_API_KEY
Token configured
yes — value hidden
Adapter gate
enabled
Live-call approval
present
Can call provider
no

Server-side placeholder only. The adapter reports environment variable names and gate state, never token values, and it does not make live STT/TTS provider calls.

Token exposure

blocked

Env reporting

names only

Live provider calls

not implemented

Fish Audio · Live TTS

J.A.R.V.I.S. Voice Test

ready

Voice Configuration

Per-agent voice slots

Each agent in the command system has a dedicated voice identity slot. Configure ElevenLabs voice IDs here; no synthesis calls are made until an API key is present and a test is explicitly approved.

no paid calls without approvalapproval-gated

ElevenLabs provider

API key not configured

key absent

ElevenLabs API key is not yet configured. Add ELEVENLABS_API_KEY to the server environment to enable per-agent voice synthesis. No paid calls will be made until a key is present and a test is explicitly run.

Safety — paid API calls

ElevenLabs charges per character synthesized. This panel configures voice identity only. No synthesis request is made by loading this page, saving a voice ID, or viewing test status. A test call will be displayed as a destructive action requiring explicit confirmation before any API call is sent.

J.A.R.V.I.S.

voice slot
Voice ID
jarvis-signature-voice
Style / tone
Calm, authoritative commander — measured cadence, slight synthetic warmth
Purpose
Central operating partner — primary voice for command layer, strategy, reporting
Test status
pending provider
Fallback provider
native-tts

D.E.V.

voice slot
Voice ID
dev-technical-voice
Style / tone
Precise, analytical — clipped technical delivery, confident and direct
Purpose
Engineering and architecture voice — CTO lane
Test status
pending provider
Fallback provider
native-tts

A.C.E.

voice slot
Voice ID
ace-production-voice
Style / tone
Energized production director — crisp, fast, action-oriented
Purpose
Events, logistics, run-of-show execution voice
Test status
pending provider
Fallback provider
native-tts

N.E.X.U.S.

voice slot
Voice ID
nexus-innovation-voice
Style / tone
Creative and expansive — visionary tone, open and exploratory
Purpose
Creative strategy and innovation mapping voice
Test status
pending provider
Fallback provider
native-tts

Provider-ready

ElevenLabs-ready agent voice map

approval-gated

ElevenLabs is the candidate for high-quality custom TTS. The app maps each agent to a voice ID slot. No voice provider API key is stored or requested here; credentials stay approval-gated until the provider and setup path are authorized.

J.A.R.V.I.S.

voice slot

Calm, authoritative commander

Central operating partner — primary voice for command layer, strategy, reporting

D.E.V.

voice slot

Precise, analytical

Engineering and architecture voice — CTO lane

A.C.E.

voice slot

Energized production director

Events, logistics, run-of-show execution voice

N.E.X.U.S.

voice slot

Creative and expansive

Creative strategy and innovation mapping voice

Unified comms

Command Feed — not a Slack clone

The app pulls the important feed and information into a clean command bowl, connected to tasks, approvals, rooms, revenue, and voice sessions — not a recreation of Slack.

Initial source layer

Slack gatewayHermes chat sessionsbridge notescommand historyapproval queueagent rooms

Live Slack API ingestion is a separate approved integration decision. Current safe path: bridge notes and local command history first.

Voice generator requirement

Final responses must not use a generic AI voice. Each agent needs a configured voice identity, STT/TTS handoff, fallback policy, and visible speaking state so the animated core reacts when the correct agent talks.