Browse the Atlas

Start from the kind of thing you already know: a model, a module, a concept, or a glossary term. This page points you to the right part of the reference and gives a few strong starting pages.

Quick routes

Use these entry points when you know how you want to explore, but not the exact page yet.

Models

Model pages explain a concrete model family or checkpoint line, what architecture it uses, and which modules matter most before you go deeper.

Browse all model pages

Model Types

Model-type glossary pages explain model families and structural roles such as encoders, decoders, and generative or multimodal setups.

Browse model-type glossary pages

Modules

Module pages break down the moving parts inside a model, such as attention variants, feed-forward blocks, normalization layers, tokenizers, and positional embeddings.

Browse all module pages

Module Components

Module-component glossary pages explain building blocks such as activations, normalization, embeddings, residual paths, and the tensors and logits that flow through model layers.

Browse module-component glossary pages

Concepts

Concept pages explain broader ideas that span many models or modules, such as transformer structure, quantization, or long-context tradeoffs.

Browse all concept pages

Inference

Inference glossary pages cover decoding, sampling controls, runtime latency, and inference-time optimizations such as KV cache and quantization.

Browse inference glossary pages

Papers

Paper pages distill one publication into the models, modules, training methods, and systems it introduces or strengthens.

Browse all paper pages

Training

Training pages focus on the named regimes and post-training methods that shape model behavior beyond the base architecture.

Browse all training pages

Systems

System pages explain the runtime and serving machinery around models, such as KV-cache handling, distributed overlap, and deployment-oriented data flow.

Browse all system pages

Glossary

Glossary pages define the vocabulary around model architecture, training, inference, and evaluation so the rest of the atlas reads more naturally.

Browse the full glossary