From the source
Lead story
Ask your AI
Top stories
Models & availability
Latest
Lead story
Top stories
Models & availability
Latest
From the source
AWS announced the Ray Serve Deep Learning Container (DLC), a pre-built Docker image for model inference that bundles PyTorch, Ray Serve, and the GPU stack.
The post demonstrates deploying a Qwen3-VL-2B vision-language model on Amazon EKS using the new container.
From the source
With the launch of the Ray Serve DLC , that same approach now extends to inference. You get a container purpose-built for serving models behind an HTTP endpoint, maintained and tested by AWS, with the full inference stack already assembled.
aws.amazon.com