GLM 5.2: What Makes it So Special?
Unlock all features
FREE: Get instant access to 10 AI summaries, chats, or transcripts per day.
Unlock all features
FREE: Get instant access to 10 AI summaries, chats, or transcripts per day.
Unlock all features
FREE: Get instant access to 10 AI summaries, chats, or transcripts per day.
Unlock all features
FREE: Get instant access to 10 AI summaries, chats, or transcripts per day.
Unlock all features
FREE: Get instant access to 10 AI summaries, chats, or transcripts per day.
Related videos
Sonnet 4.5 Is Here—And It’s a Beast at Coding
Prompt Engineering
52.0k views
GPT-OSS Jailbreak with this Simple Trick
Prompt Engineering
54.4k views
What is AI Engineering
Telusko
67.6k views
Context Engineering is All You NEED!
Prompt Engineering
38.7k views
The Only Embedding Model You Need for RAG
Prompt Engineering
35.2k views
Gemini CLI — Google’s Free Open-Source Coding Agent
Prompt Engineering
56.6k views
AI prompt engineering in 2025: What works and what doesn’t | Sander Schulhoff
Lenny's Podcast
68.3k views
The Secret to Perfect Prompts (Without Prompt Engineering)
Futurepedia
55.4k views
Meet KAG: Supercharging RAG Systems with Advanced Reasoning
Prompt Engineering
63.5k views
Do Anything with Local Agents with AnythingLLM
Prompt Engineering
60.4k views
Top Comments (10)
0:46 An example of necessity winning out. When the chip trade war started and sanctions came I assume they pivoted to efficiency knowing they may become constrained in the near term for getting additional high-compute capacity. The surprising thing is that they're being so open with it (a good thing overall). Compared to the capitalistic approach of the US where we just throw pallets of cash at it to brute force frontier leaps. Their (probably) subsidized budget; granted, a lot of the US companies are subsidized in various other ways like tax breaks or DoD contracts; and focus on efficiency allows them to undercut the rest of the market, especially since they _can_ build out capacity for inference compute.
I will definitely give it a try. Looks promising
Very clear & informative video, thank you! Always great when you explain those important concepts!
I am all for open models and I really applaud the chinese, but the elephant in the room is: how are they going to finance these in the long run...!?
I love glm models. They are prob the best we have in open space. But they are not perfect. I was laughing that today when 5.2 has failed during quite a trivial troubleshooting task, making a very wild guess where the actual issue was on the surface. 5.1 was making the same wild guesses. Still love them for building, but in the debugging they are quite shitty 😂
I can run this in Q4 variant on a dual Xeon with 384GB ram and 2 V100, really up to the brim. Around 50 t/s PP and 2 t/s TG, slow but can be still used as a background agent that does planning and review. Minimax m3 q6 is at 40 t/s PP and 5 t/s TG…. Testing and trying to see which I would replace Qwen 3.5 397B with.
great video as usual can you share please how you created these animations, visuals and handwritten text?
Try with pi harness which seems works well
wow, you dont have to retrain your foundational model to use multi layer attentions?
yeah, sounds impressive. it's great we have access to an actually capable model
Unlock the Data Inside
Turn Videos into Knowledge
- Get FREE 10/day: transcripts, summaries, chats
- Chat with videos, export text & PDF
- $1 free API credit for RAG, chatbots & research
Free forever plan • All features unlocked
Top Comments (10)
0:46 An example of necessity winning out. When the chip trade war started and sanctions came I assume they pivoted to efficiency knowing they may become constrained in the near term for getting additional high-compute capacity. The surprising thing is that they're being so open with it (a good thing overall). Compared to the capitalistic approach of the US where we just throw pallets of cash at it to brute force frontier leaps. Their (probably) subsidized budget; granted, a lot of the US companies are subsidized in various other ways like tax breaks or DoD contracts; and focus on efficiency allows them to undercut the rest of the market, especially since they _can_ build out capacity for inference compute.
I will definitely give it a try. Looks promising
Very clear & informative video, thank you! Always great when you explain those important concepts!
I am all for open models and I really applaud the chinese, but the elephant in the room is: how are they going to finance these in the long run...!?
I love glm models. They are prob the best we have in open space. But they are not perfect. I was laughing that today when 5.2 has failed during quite a trivial troubleshooting task, making a very wild guess where the actual issue was on the surface. 5.1 was making the same wild guesses. Still love them for building, but in the debugging they are quite shitty 😂
I can run this in Q4 variant on a dual Xeon with 384GB ram and 2 V100, really up to the brim. Around 50 t/s PP and 2 t/s TG, slow but can be still used as a background agent that does planning and review. Minimax m3 q6 is at 40 t/s PP and 5 t/s TG…. Testing and trying to see which I would replace Qwen 3.5 397B with.
great video as usual can you share please how you created these animations, visuals and handwritten text?
Try with pi harness which seems works well
wow, you dont have to retrain your foundational model to use multi layer attentions?
yeah, sounds impressive. it's great we have access to an actually capable model