Generate images and video with vLLM-Omni on SageMaker AI – Part 2

What happened
Deploy two generative media models from one AWS vLLM-Omni Deep Learning Container on Amazon SageMaker AI. Generate an image with FLUX.2-klein through real-time inference, then animate it into video with Wan2.1-VACE through asynchronous inference, and retrieve the MP4 from Amazon S3. Amazon Web Services is a cloud computing company based in Seattle.
You deploy two endpoints from the same AWS vLLM-Omni Deep Learning Container (DLC): a real-time endpoint for FLUX.2-klein-4B image generation and an asynchronous endpoint for Wan2.1-VACE-1.3B video generation. The workflow sends a text prompt to generate a still image, then passes the image and a motion prompt to the video endpoint.
It retrieves the MP4 from Amazon Simple Storage Service (Amazon S3) and provides an optional Streamlit interface. AWS Deep Learning Containers package frameworks and dependencies for training and inference on AWS.
Sources & evidence
- AWS Machine Learning Blog Primary / official
Generate images and video with vLLM-Omni on SageMaker AI – Part 2 ↗
https://aws.amazon.com/blogs/machine-learning/generate-images-and-video-with-vllm-omni-on-sagemaker-ai-part-2/