fixes
This commit is contained in:
@@ -209,6 +209,8 @@ including simpler feedforward networks and more recent attention-based models,
|
|||||||
to evaluate trade-offs in performance, interpretability, and computational cost.
|
to evaluate trade-offs in performance, interpretability, and computational cost.
|
||||||
|
|
||||||
The next section introduce the \emph{Transformer} architecture, a more recent alternative that forgoes
|
The next section introduce the \emph{Transformer} architecture, a more recent alternative that forgoes
|
||||||
recurrence in favor of attention mechanisms
|
recurrence in favor of attention mechanisms.
|
||||||
|
|
||||||
\subsubsection{Transformer Models}\label{subsubsec:transformer_models}
|
\subsubsection{Transformer Models}\label{subsubsec:transformer_models}
|
||||||
|
|
||||||
|
|
||||||
|
|||||||
Reference in New Issue
Block a user