Theme of inquiry

How do different training methods affect model reasoning?

A question within its area, explored through 4 lines of inquiry below — each a family of specific questions the research asks.


How does decomposing tasks improve reasoning and prevent failure propagation?

40 specific questions

See all 40 questions in this line of inquiry
Does RL create genuinely new reasoning capabilities or refine existing ones?

114 specific questions

See all 114 questions in this line of inquiry
What training dynamics and scale trigger emergence of reasoning capabilities?

65 specific questions

See all 65 questions in this line of inquiry
How does policy entropy collapse constrain scaling of reasoning-focused RL?

37 specific questions

See all 37 questions in this line of inquiry