Agent Development Kit

Paper · Source
Foundation ModelsMulti-Agent Architectures

In the Agent Development Kit (ADK), an Agent is a self-contained execution unit designed to act autonomously to achieve specific goals. Agents can perform tasks, interact with users, utilize external tools, and coordinate with other agents.

Introduction. The foundation for all agents in ADK is the BaseAgent class. It serves as the fundamental blueprint. To create functional agents, you typically extend BaseAgent in one of three main ways, catering to different needs – from intelligent reasoning to structured process control. ADK provides distinct agent categories to build sophisticated applications:

Discussion / Conclusion. While each agent type serves a distinct purpose, the true power often comes from combining them. Complex applications frequently employ multi-agent architectures where: Understanding these core types is the first step toward building sophisticated, capable AI applications with ADK.

Agents Multi Architecture

Lines of inquiry this paper opens 24

Research framings built by reading the notes related to this paper — the questions it feeds into.

What are the consequences of models training on synthetic data? Can AI-generated outputs constitute genuine knowledge or valid claims? Do language models learn genuine linguistic structure or just surface patterns? How does AI-generated content transformation affect public discourse quality? What makes AI persuasion effective and how can we counter it? How can AI alignment serve diverse human preferences at scale? How can language models sustain linguistic synchrony and intersubjectivity during dialogue? When does optimizing for quality undermine the value of diversity? Why does reinforcement learning suppress output diversity compared to supervised fine-tuning? Does alignment training create blind spots in detecting genuine safety threats? How do language models inherit human biases from training data? How do multi-agent systems achieve genuine cooperation and reasoning? What determines success in training models on multiple tasks? What factors beyond surface content determine how readers extract meaning differently?