Inkling is receiving your audio directly, and replying with GGML's Qwen3TTS.
Tap to start
This demo sends your audio directly to Inkling on Hugging Face Inference Endpoints, then voices Inkling's reply with Qwen3-TTS through GGML.
The pipeline
http://localhost:8080
/v1/realtime
Let the assistant act during the conversation. Changes apply live.