Fabian Pottbäcker, Thomas Endres & Martin Foertsch
AI'll Be Back: Generative AI in Image, Video, and Audio Production
#1about 2 minutes
The hype and promise of generative AI
Generative AI is at the peak of the Gartner Hype Cycle, with applications spanning text, image, audio, and video generation.
#2about 1 minute
How large language models generate text
Large language models (LLMs) function as next-word predictors, generating text token by token in a process that creates a typewriter-like effect.
#3about 3 minutes
Understanding tokenization and semantic embeddings
Text is broken down into numerical tokens and then mapped into a multi-dimensional vector space where semantically similar words are located close together.
#4about 3 minutes
The role of transformers and the attention mechanism
The transformer architecture uses an attention mechanism to weigh the importance of different words in the input sequence to understand context and resolve ambiguity.
#5about 2 minutes
Connecting text and images with the CLIP model
The CLIP model establishes a shared embedding space for text and images, enabling the system to measure the semantic similarity between a text description and a picture.
#6about 7 minutes
How diffusion models create images from noise
Diffusion models generate images through an iterative process of predicting and subtracting noise from a random starting point, guided by a text prompt's embedding.
#7about 5 minutes
Applying diffusion transformers to video generation
Video generation uses a diffusion transformer to maintain coherence across frames by processing video in patches and applying the denoising process to the entire sequence.
#8about 1 minute
Advanced techniques for video manipulation and editing
Beyond simple generation, models can perform image-to-video conversion, extend existing clips, interpolate between two different videos, or edit specific regions.
#9about 2 minutes
Current limitations and physical inconsistencies in AI video
Generative video models still struggle with understanding cause and effect, leading to physically impossible events and objects appearing or behaving illogically.
#10about 3 minutes
Ethical challenges of generative AI training data
Major ethical concerns include the use of copyrighted or publicly available data without consent for training models, leading to legal challenges and questions about ownership.
Related jobs
Jobs that call for the skills explored in this talk.
Wilken GmbH
Ulm, Germany
Senior
Kubernetes
AI Frameworks
+3
ROSEN Technology and Research Center GmbH
Osnabrück, Germany
Senior
TypeScript
React
+3
Matching moments
14:06 MIN
Exploring the role and ethics of AI in gaming
Devs vs. Marketers, COBOL and Copilot, Make Live Coding Easy and more - The Best of LIVE 2025 - Part 3
04:57 MIN
Increasing the value of talk recordings post-event
Cat Herding with Lions and Tigers - Christian Heilmann
06:44 MIN
Using Chrome's built-in AI for on-device features
Devs vs. Marketers, COBOL and Copilot, Make Live Coding Easy and more - The Best of LIVE 2025 - Part 3
09:10 MIN
How AI is changing the freelance developer experience
WeAreDevelopers LIVE – AI, Freelancing, Keeping Up with Tech and More
04:28 MIN
Building an open source community around AI models
AI in the Open and in Browsers - Tarek Ziadé
04:05 MIN
How AI code generators have become more reliable
AI in the Open and in Browsers - Tarek Ziadé
04:17 MIN
Playing a game of real or fake tech headlines
WeAreDevelopers LIVE – You Don’t Need JavaScript, Modern CSS and More
03:31 MIN
Using AI to make work more human, not replace humans
Turning People Strategy into a Transformation Engine
Featured Partners
Related Videos
Your imaginations is (no longer) the limit: how Generative AI empowers people to be creative
David Estevez
Multimodal Generative AI Demystified
Ekaterina Sirazitdinova
AI: Superhero or Supervillain? How and Why with Scott Hanselman
Scott Hanselman
In the Dawn of the AI: Understanding and implementing AI-generated images
Timo Zander
GenAI Unpacked: Beyond Basic
Damir
The shadows of reasoning – new design paradigms for a gen AI world
Jonas Andrulis
Should we build Generative AI into our existing software?
Simon Müller
The shadows that follow the AI generative models
Cheuk Ho
Related Articles
View all articles



From learning to earning
Jobs that call for the skills explored in this talk.

Forschungszentrum Jülich GmbH
Jülich, Germany
Intermediate
Senior
Linux
Docker
AI Frameworks
Machine Learning

OpenAI
München, Germany
Senior
API
Python
JavaScript
Machine Learning


RE-INvent Retail GmbH
Azure
Python
Microservices
Google Cloud Platform



BMW AG
München, Germany
Senior
Python
PyTorch
TensorFlow
Computer Vision
Natural Language Processing

TMC
Utrecht, Netherlands
Senior
API
Azure
Python
Docker
FastAPI
+1
