The Purpose I Write This Blog

   Thinking models are crazily popualr nowadays. The first time I delved in this area was in September, 2023. Later I gradually forgetted this area, until Deepseek came to life. I want to keep to collect information about LLM reasoning (as well as post-training) and share my thoughts here.

💡 This post is mainly focused on general reasoning. For agentic reasoning, please refer to this post.

Reinforcement Learning

Blogs

RL algorithms

Engineering

Analyses

Thinking Models and Methods

models

text-based reasoning

visual reasoning

Speech LLMs

Others

Analyses on reasoning

general

interpretability

theories

safety