WORLDTECH NEWS Global technology intelligence.Contact
← Back to WORLDTECH

Generate images and video with vLLM-Omni on SageMaker AI – Part 2

Close-up of a GeForce RTX graphics card on a desk, showcasing its design and technology.
Illustrative photo.Photo by Trần Chính on PexelsAmazon Web Services logo shown for identification only; no affiliation with or endorsement of WORLDTECH is implied.

What happened

Deploy two generative media models from one AWS vLLM-Omni Deep Learning Container on Amazon SageMaker AI. Generate an image with FLUX.2-klein through real-time inference, then animate it into video with Wan2.1-VACE through asynchronous inference, and retrieve the MP4 from Amazon S3. Amazon Web Services is a cloud computing company based in Seattle.

You deploy two endpoints from the same AWS vLLM-Omni Deep Learning Container (DLC): a real-time endpoint for FLUX.2-klein-4B image generation and an asynchronous endpoint for Wan2.1-VACE-1.3B video generation. The workflow sends a text prompt to generate a still image, then passes the image and a motion prompt to the video endpoint.

It retrieves the MP4 from Amazon Simple Storage Service (Amazon S3) and provides an optional Streamlit interface. AWS Deep Learning Containers package frameworks and dependencies for training and inference on AWS.

Sources & evidence