Agent Development Kit

Paper · Source
Foundation ModelsMulti-Agent Architectures

In the Agent Development Kit (ADK), an Agent is a self-contained execution unit designed to act autonomously to achieve specific goals. Agents can perform tasks, interact with users, utilize external tools, and coordinate with other agents.

Introduction. The foundation for all agents in ADK is the BaseAgent class. It serves as the fundamental blueprint. To create functional agents, you typically extend BaseAgent in one of three main ways, catering to different needs – from intelligent reasoning to structured process control. ADK provides distinct agent categories to build sophisticated applications:

Discussion / Conclusion. While each agent type serves a distinct purpose, the true power often comes from combining them. Complex applications frequently employ multi-agent architectures where: Understanding these core types is the first step toward building sophisticated, capable AI applications with ADK.

Agents Multi Architecture

Lines of inquiry this paper opens 24

Research framings built by reading the notes related to this paper — the questions it feeds into.

How do models learn from self-generated outputs without cascading failures? Can AI systems achieve real improvement without external human feedback? Do language models encode knowledge that influences generation, or primarily imitate surface patterns? Are AI-generated articles systematically disadvantaged in search ranking and user engagement? Does preference optimization undermine conversational grounding in language models? Can AI systems participate in genuine communication or only simulate it? Why do LLM research ideation systems generate novelty but lack diversity? How does fine-tuning trade off accuracy against reasoning quality? Can base models hide emergent misalignment through alignment training? What limits language model accuracy in evaluating ideas? When do multi-agent systems improve over single frontier models? Why do multi-agent systems reach premature consensus without genuine deliberation? How does model capacity affect learning performance on diverse downstream tasks? How does diversity prevent model convergence on superficial patterns? Can readers reliably distinguish AI-written text from human writing? How do interpretive frames override surface features in text comprehension? How do training data quality and composition affect downstream model performance? Can smaller specialized models match frontier models on key metrics?