Skip to main content

Transformer Architecture · Chapter 10

Encoder–Decoder & Cross-Attention

The full architecture and three attention types

This chapter is part of the Pro track. Pro beta members receive the complete self-paced architecture sequence.

Related reading · Knowledge Lab

Build the context for this chapter