Model Architecture

Explore Model Architecture through related topics and the articles other pages reference most.

Explore articles

Browse subtopics (29)

Articles that also belong to these categories. Counts cover all of Model Architecture.

Showing 61-65 of 65 articles

Vision Transformer

The Vision Transformer (ViT) is a deep learning architecture that represents an image as a sequence of fixed-size patches and processes that sequence with a Transformer encoder.

Computer Vision

YaRN

YaRN (Yet another RoPE extensioN) is a compute-efficient method for extending the context window of large language models that use Rotary Position Embeddings (RoPE).

AI InferenceDeep Learning

xLSTM

xLSTM (Extended Long Short-Term Memory) is a recurrent neural network architecture introduced in May 2024 by Maximilian Beck, Korbinian Pöppel, Sepp Hochreiter, and collaborators at Johannes Kepler University…

Deep LearningNeural Networks