fixes
This commit is contained in:
@@ -209,6 +209,8 @@ including simpler feedforward networks and more recent attention-based models,
|
||||
to evaluate trade-offs in performance, interpretability, and computational cost.
|
||||
|
||||
The next section introduce the \emph{Transformer} architecture, a more recent alternative that forgoes
|
||||
recurrence in favor of attention mechanisms
|
||||
recurrence in favor of attention mechanisms.
|
||||
|
||||
\subsubsection{Transformer Models}\label{subsubsec:transformer_models}
|
||||
|
||||
|
||||
|
||||
Reference in New Issue
Block a user