Lead story
Models & availability
Latest
Lead story
Models & availability
Latest
This blog post provides a step-by-step guide to deploying the Speech-to-Speech (S2S) pipeline on Hugging Face Inference Endpoints using a custom Docker image. S2S combines Voice Activity Detection, Speech-to-Text, a Language Model, and Text-to-Speech to create a seamless speech interaction experience, with support for multiple languages.
From the source
In this blog post, we’ll guide you step by step to deploy Speech-to-Speech to a Hugging Face Inference Endpoint.
huggingface.co