Nexus // TTS Protocol

Neural Synthesis/Client-Side/16 kHz Mono

Models not loaded
Model Download
Waiting...

Neural models are required before synthesis can begin. This will download approximately 300 MB of model weights from the HF Hub and cache them locally.

PROCESS: Client-side synthesis using Xenova/speecht5_tts (16 kHz) via @huggingface/transformers. Speaker embeddings are pre-computed 512-dim x-vectors from the CMU ARCTIC dataset.

VOICE CLONING: Not supported client-side. Requires a dedicated x-vector/ECAPA speaker encoder whose output distribution matches SpeechT5's training data. General-purpose wav2vec2 feature extractors produce incompatible embeddings.

access https://inky9.lovable.app/ for more options

All operations remain within the local browser nexus. No data is transmitted to external servers.