Area of inquiry
How do we develop coherent and human-aligned AI systems?
One of the field's big questions. It gathers the themes below — all questions the research asks, not subjects it is about. (For subjects, browse Topics.)
Themes within this area 4
Each theme is a narrower question. Follow one down toward its lines of inquiry.
- How do different reward signals and mechanisms drive agent learning?
- How do representation and aggregation choices affect model alignment and reliability?
- How do language models maintain conversational coherence through grounding?
- What psychological and emotional factors determine therapeutic AI effectiveness?