Line of inquiry
Inquiring lines›What makes reasoning better — more…›How do prompts and framing affect…›this line of inquiry
Why can LLMs generate ideas better than they evaluate them?
A broader line of inquiry — a family of 17 specific questions the research asks around this. Follow one into its inquiring-line page, or move sideways to a related line below.
Questions in this line of inquiry 17
Specific inquiring lines the field asks around this — ordered from the most general framing down to the most specific angle.
- Why do LLMs generate novel ideas but struggle to evaluate them?
- Do LLMs generate more novel ideas than they can evaluate?
- How can LLMs evaluate their own creative outputs for utility and novelty?
- Why do LLMs excel at generation but struggle with evaluation?
- What makes novelty assessment harder to automate than idea generation?
- Why do models generate creative ideas but fail to evaluate their legitimacy?
- What structural barriers prevent LLMs from making evaluative judgments about writing?
- Where do LLMs succeed at generation but struggle with evaluation?
- Why do LLMs generate novel ideas but lack evaluative commitment?
- Can critique-only calls in LLMs exploit a measurable gap between generation and evaluation?
- Can LLMs recognize rhetorical devices they cannot actually produce themselves?
- Why do LLMs plateau on creativity tasks while humans reach further?
- Do LLMs match top human creative writers in literary quality?
- What workflow structure pairs LLM generation with human evaluation most effectively?
- Why do review corpora contain biases that affect generated comparisons?
- What role do model-based critics play in validating LLM plans?
- What role should stakeholders play in evaluating LLM fairness?