Theme of inquiry

What systematic failures and vulnerabilities compromise AI reasoning systems?

A question within its area, explored through 11 lines of inquiry below — each a family of specific questions the research asks.


How does memorization interact with learning and generalization?

19 specific questions

See all 19 questions in this line of inquiry
What determines success in training models on multiple tasks?

34 specific questions

See all 34 questions in this line of inquiry
How does example difficulty affect learning efficiency in language models?

39 specific questions

See all 39 questions in this line of inquiry
How do knowledge injection methods compare across cost and effectiveness?

15 specific questions

See all 15 questions in this line of inquiry
How do self-generated feedback mechanisms enable effective model learning?

42 specific questions

See all 42 questions in this line of inquiry
What makes weaker teacher models effective for stronger student training?

31 specific questions

See all 31 questions in this line of inquiry
Why does finetuning cause catastrophic forgetting of model capabilities?

29 specific questions

See all 29 questions in this line of inquiry
How do training priors constrain what context information can override?

56 specific questions

See all 56 questions in this line of inquiry
What are the consequences of models training on synthetic data?

35 specific questions

See all 35 questions in this line of inquiry
Does alignment training create blind spots in detecting genuine safety threats?

35 specific questions

See all 35 questions in this line of inquiry
Does fine-tuning modify underlying model capabilities or only behavioral outputs?

44 specific questions

See all 44 questions in this line of inquiry