Lead story
Models & availability
Latest
Lead story
Models & availability
Latest
Hugging Face introduces a remote VAE decoding feature for Inference Endpoints, allowing users to offload the VAE decoder to a separate endpoint to reduce memory usage and improve latency, with support for models like SD v1.5, Flux, and HunyuanVideo.
From the source
we want to pilot an idea with the community — delegating the decoding process to a remote endpoint.
huggingface.co