Diffusion Gemma: The First Diffusion Model that "Thinks"
Unlock all features
FREE: Get instant access to 10 AI summaries, chats, or transcripts per day.
Unlock all features
FREE: Get instant access to 10 AI summaries, chats, or transcripts per day.
Unlock all features
FREE: Get instant access to 10 AI summaries, chats, or transcripts per day.
Unlock all features
FREE: Get instant access to 10 AI summaries, chats, or transcripts per day.
Unlock all features
FREE: Get instant access to 10 AI summaries, chats, or transcripts per day.
Related videos
Engineering The Perfect First Date
Mark Rober
209.5k views
Sonnet 4.5 Is Here—And It’s a Beast at Coding
Prompt Engineering
52.0k views
GPT-OSS Jailbreak with this Simple Trick
Prompt Engineering
54.4k views
Context Engineering is All You NEED!
Prompt Engineering
38.7k views
The Only Embedding Model You Need for RAG
Prompt Engineering
35.2k views
Gemini CLI — Google’s Free Open-Source Coding Agent
Prompt Engineering
56.6k views
AI prompt engineering in 2025: What works and what doesn’t | Sander Schulhoff
Lenny's Podcast
68.3k views
The Secret to Perfect Prompts (Without Prompt Engineering)
Futurepedia
55.4k views
Do Anything with Local Agents with AnythingLLM
Prompt Engineering
60.4k views
LightRAG: A More Efficient Solution than GraphRAG for RAG Systems?
Prompt Engineering
84.5k views
Top Comments (10)
This model should be fantastic for simple, high volume tasks like data processing
Thanks for your video, I am interested in fine tuning the diffusion models🙏
This is our most sassy Gemma-4 so far!! People have been trying to abliterate her to remove refusals, but to no avail. 😆 The diffusion transformer safety layers are not directionally projectable (✿◠‿◠)
Perfect, we have MATRIX now.
Definitely interested in a step by step into to LoRA and Unsloth fine-tuning in general.
Always informative! TYVM!
Great video. Thanks.
Nice text animation, really enjoyed the blue packer style pixel revolving materializing text animation. New subscriber.
I've been working on this all week since this model was released, learning how to make fine-tuning adjustments and researching how to apply RL. I'm sure this can be implemented to improve the code.
6:15 you mentioned FP8 on GPUs like L40S, A6000, A100 but isnt FP8 supported from Hopper (cuda cc 9)? Those are Ampere GPUs (cuda cc 8) so they wouldnt work right?
Unlock the Data Inside
Turn Videos into Knowledge
- Get FREE 10/day: transcripts, summaries, chats
- Chat with videos, export text & PDF
- $1 free API credit for RAG, chatbots & research
Free forever plan • All features unlocked
Top Comments (10)
This model should be fantastic for simple, high volume tasks like data processing
Thanks for your video, I am interested in fine tuning the diffusion models🙏
This is our most sassy Gemma-4 so far!! People have been trying to abliterate her to remove refusals, but to no avail. 😆 The diffusion transformer safety layers are not directionally projectable (✿◠‿◠)
Perfect, we have MATRIX now.
Definitely interested in a step by step into to LoRA and Unsloth fine-tuning in general.
Always informative! TYVM!
Great video. Thanks.
Nice text animation, really enjoyed the blue packer style pixel revolving materializing text animation. New subscriber.
I've been working on this all week since this model was released, learning how to make fine-tuning adjustments and researching how to apply RL. I'm sure this can be implemented to improve the code.
6:15 you mentioned FP8 on GPUs like L40S, A6000, A100 but isnt FP8 supported from Hopper (cuda cc 9)? Those are Ampere GPUs (cuda cc 8) so they wouldnt work right?