Shipping costs will be calculated based on this address throughout the site.
Select your country
Americas
Argentina
Brazil
Canada
Chile
Colombia
Costa Rica
Dominican Republic
Ecuador
El Salvador
Mexico
Peru
U.S.A.
Uruguay
Europe
Austria
Belgium
Croatia
Czech Republic
Denmark
Finland
France
Germany
Greece
Hungary
Ireland
Italy
Latvia
Malta
Netherlands
Norway
Poland
Portugal
Serbia
Slovakia
Slovenia
Spain
Sweden
Switzerland
United Kingdom
Rest of the world


Video Generation with AI. Working with Diffusion Transformers and Multimodal Learning
Joseph Enochs (Author) · O'Reilly Media · Paperback
Video generation is rapidly becoming a key area in generative AI—combining spatial, temporal, and multimodal reasoning to produce moving images that are both coherent and creative. For many practitioners, however, understanding how these models function and implementing them remains a significant challenge. Video Generation with AI offers a straightforward guide for exploring this new terrain.
Author Joseph Enochs leverages his experience leading enterprise AI projects to clarify how diffusion transformers, multimodal large language models, and spatiotemporal architectures combine to create high-quality video. Blending technical detail with practical examples, he demonstrates how to transition from isolated experiments to production-ready systems that transform creative fields, media, and human-machine collaboration.
Do you have a question about the book? Login to be able to add your own question.

