Remix.run Logo
rhdunn 41 minutes ago

Welch Labs on YouTube has a great collection of videos on how AI models learn. His recent video [1] covers how image models can learn to encode reasoning in the image processing layers when not given an out of band reasoning set of weights to use instead. I suspect that this applies to LLMs and CoT reasoning vs output token weights.

[1] https://www.youtube.com/watch?v=QgH9sr7G13Q