Navigate Select ESC Close

LLM that loops instead of Doing Chain-of-Thought

2026-07-01 Science & Technology
29.0k
1.5k
153
bycloud
bycloud
229.0k subscribers

Unlock all features

FREE: Get instant access to 10 AI summaries, chats, or transcripts per day.

Description

Need to fine-tune a model without the hassle? Try out Crusoe's serverless fine-tuning today! https://www.crusoe.ai/contact-sales/serverless-preview?utm_source=bycloud&utm_medium=influencer&utm_campaign=serverlessfinetuning my latest project: Intuitive AI Academy We just wrote a new piece on Optimization!! https://intuitiveai.academy/ limited time code "LOCKIN" for 35% off yearly plan Chain-of-Thought is ugly, so what if we remove it? In this video, I am diving into the latest architecture experiment that is Looped Transformer. So LLMs think with words, but here Looped Transformer simulates thinking by looping the same few layers. My Newsletter https://mail.bycloud.ai/ My Patreon https://www.patreon.com/c/bycloud Loop, Think, & Generalize: Implicit Reasoning in Recurrent-Depth Transformers [Paper] https://arxiv.org/abs/2604.07822 A Mechanistic Analysis of Looped RLMs [Paper] https://arxiv.org/abs/2604.11791 Parcae: Scaling Laws For Stable Looped Language Models [Paper] https://arxiv.org/abs/2604.12946 Mixture-of-Recursions: Learning Dynamic Recursive Depths for Adaptive Token-Level Computation [Paper] https://arxiv.org/abs/2507.10524 Try out my new fav place to learn how to code https://scrimba.com/?via=bycloudAI This video is supported by the kind Patrons & YouTube Members: 🙏Spam Maj, Alex, Chris LeDoux, DX Research Group, Poof N' Inu, Deagan, Robert Zawiasa, Ryszard Warzocha, Tobe2d, Louis Muk, Akkusativ, Kevin Tai, Mark Buckler, NO U, Tony Jimenez, Ângelo Fonseca, jiye, Anushka, Asad Dhamani, Binnie Yiu, Calvin Yan, Clayton Ford, Diego Silva, Etrotta, Gonzalo Fidalgo, Handenon, Hector, Jake Disco very, Michael Brenner, Nilly K, OlegWock, Daddy Wen, Shuhong Chen, Sid_Cipher, Stefan Lorenz, Sup, tantan assawade, Thipok Tham, Thomas Di Martino, Thomas Lin, Richárd Nagyfi, Paperboy, mika, Leo, Berhane-Meskel, Kadhai Pesalam, mayssam, Bill Mangrum, nyaa, Toru Mon, Lame Plane, Matej Macak, Len Mo, saylikhapekar, ZyanSheep, THEVIERAOS, Ricardo Raphael Corona-Moreno, superchordate [Discord] https://discord.gg/NhJZGtH [Twitter] https://twitter.com/bycloudai [Patreon] https://www.patreon.com/bycloud [Business Inquiries] [email protected] [Other Inquiries] [email protected] [Profile & Banner Art] https://twitter.com/pygm7 [Video Editor] @aduckchicken2 Manim Animations created with Manimate https://www.manimate.ai/ [Ko-fi] https://ko-fi.com/bycloudai

Top Comments (10)

@Sam-te6np 2026-07-01

close enough, welcome back RNNs

394 9 replies
@FredericoOliveira-d8o 2026-07-01

Chain-of-Thought is just the LLM 'yapping' to find the answer. works but sucks

249 14 replies
@carlossierra1988 2026-07-01

You need to cover LFM2.5 8b's way of stopping doom loops.

79 15 replies
@rady7273 2026-07-01

dude imagine every step of thinking you do, you would have to write down what you just thought about, then get amnesia and then have to read what you are supposed to do completely again and what you have already come up with and trust that what you wrote down conveys all the information and all of the subtleties to solve the task correctly. And then you do that 100 times. Like playing telephone with yourself. Actually honestly that's exactly how my ADHD ass does things

52 5 replies
@Ahamshep 2026-07-01

I'm surprised that the idea of telling the model to "think step-by-step" back in 2022, is still one of the largest advancements since transformers themselves. If I recall correctly CNNs were developed by observing that neural networks slowly extracted features. So understanding and improving the mechanisms is the right direction. But other than scaling and training data curation, have they just hit a wall?

49 4 replies
@KyberNuII 2026-07-01

We've already dragged transformer to death. Maybe we should revive RNNs or architectures similar to that.

25 1 replies
@bycloudAI 2026-07-01

Need to fine-tune a model without the hassle? Try out Crusoe's serverless fine-tuning today! https://www.crusoe.ai/contact-sales/serverless-preview?utm_source=bycloud&utm_medium=influencer&utm_campaign=serverlessfinetuning

8 3 replies
@intrinsical 2026-07-02

This is already a year old idea. See "Scaling Latent Reasoning via Looped Language Models", published October 2025.

5
@holopengin 2026-07-02

I was about to comment how choosing the number of loops at the get-go seemed entirely backwards. Glad to see that they thought of that and tried deferring that choice to the end of the inference step, and saw better performance when they did. Neat!

4
@sunain-c-tray 2026-07-02

This video provides the excitement of something mathematically elegant with a slap of reality at the end.

4

Unlock the Data Inside
Turn Videos into Knowledge

  • Get FREE 10/day: transcripts, summaries, chats
  • Chat with videos, export text & PDF
  • $1 free API credit for RAG, chatbots & research

Free forever plan • All features unlocked

App screenshot