The Field Guide · 524 terms · 8 regions · 61 landmarks
Field Guide
A map of AI, not a syllabus. Pick any destination and follow a route through exactly the ideas you need first.
How to use it: open any term and its route shows every idea to understand first, basics first. Landmarks are the concepts the rest of the map leans on, which makes them good first stops.
Foundations
71 terms · 7 landmarks
Core ML ideas, classic algorithms, and the data they learn from.
All 71 Foundations terms
- Active Learning
- AlphaFold
- Anomaly Detection
- Bayesian Inference
- Bias-Variance Tradeoff
- Causal Inference
- Classification
- Clustering
- Collaborative Filtering
- Concentration of Measure
- Conjugate Prior
- Covariance
- Curse of Dimensionality
- Curse of Dimensionality
- Data Preprocessing
- Dataset
- Decision Tree
- Dimensionality Reduction
- Drug Discovery
- Ensemble Learning
- Entropy
- Evidence Lower Bound
- Expectation-Maximization
- Feature
- Feature Engineering
- Feature Importance
- Feature Selection
- Fraud Detection
- Gaussian Process
- Gibbs Sampling
- Gradient Boosting
- Ground Truth
- Hidden Markov Model
- Imbalanced Dataset
- Inductive Bias
- Information Bottleneck
- Jensen-Shannon Divergence
- KL Divergence
- Knowledge Graph
- Labeled Data
- Latent Variable
- Machine Learning
- Manifold Hypothesis
- Markov Chain Monte Carlo
- Maximum A Posteriori
- Maximum Likelihood Estimation
- Mutual Information
- No Free Lunch Theorem
- Normalization
- Occam's Razor
- One-Hot Encoding
- PAC Learning
- Personalization
- Posterior Distribution
- Principal Component Analysis
- Prior Distribution
- Protein Folding
- Rademacher Complexity
- Random Forest
- Recommender System
- Regression
- Semi-Supervised Learning
- Supervised Learning
- Support Vector Machine
- Synthetic Data
- Time Series Forecasting
- Training Data
- Unsupervised Learning
- Variational Inference
- VC Dimension
- Wasserstein Distance
Neural Networks
69 terms · 9 landmarks
Layers, activations, backprop: the building blocks of deep learning.
All 69 Neural Networks terms
- Activation Function
- Autoencoder
- Backward Pass
- Batch Normalization
- Bias
- Boltzmann Machine
- Bottleneck
- Capsule Network
- Convolution
- Convolutional Neural Network
- Deep Belief Network
- Deep Learning
- Depthwise Separable Convolution
- Dilated Convolution
- Dropout
- Echo State Network
- ELU
- Feature Map
- Feedforward Network
- Filter
- Forward Pass
- Gated Recurrent Unit
- GELU
- Graph Attention Network
- Graph Neural Network
- Group Normalization
- He Initialization
- Hidden Layer
- Hopfield Network
- Kernel
- Latent Space
- Layer
- Leaky ReLU
- Liquid State Machine
- Long Short-Term Memory
- Maxout
- Message Passing
- Mish
- Mixture of Experts
- Multi-Layer Perceptron
- Neural Network
- Neural ODE
- Node Embedding
- Padding
- Parameter
- Perceptron
- Pooling
- PReLU
- Radial Basis Function Network
- Receptive Field
- Recurrent Neural Network
- ReLU
- Representation Learning
- Residual Connection
- Restricted Boltzmann Machine
- Sigmoid
- Skip Connection
- Softmax
- Softmax Temperature
- Spectral Normalization
- Spiking Neural Network
- Stride
- Swish
- Transposed Convolution
- Universal Approximation Theorem
- Variational Autoencoder
- Weight
- Weight Normalization
- Xavier Initialization
Training
77 terms · 11 landmarks
Losses, optimizers, regularization, and how models actually learn.
All 77 Training terms
- AdaGrad
- Adam Optimizer
- AdamW
- Automatic Mixed Precision
- AutoML
- Backpropagation
- Batch Gradient Descent
- Batch Size
- Catastrophic Forgetting
- Continual Learning
- Contrastive Learning
- Contrastive Loss
- Cosine Annealing
- Cross-Entropy Loss
- CTC Loss
- Curriculum Learning
- CutMix
- Cutout
- Cyclical Learning Rate
- Data Augmentation
- Data Parallelism
- Distillation Temperature
- Early Stopping
- Empirical Risk Minimization
- Epoch
- Exploding Gradient
- Fine-Tuning
- Focal Loss
- Gradient Accumulation
- Gradient Checkpointing
- Gradient Clipping
- Gradient Descent
- Hinge Loss
- Huber Loss
- Hyperparameter
- Hyperparameter Tuning
- Knowledge Distillation
- L1 Regularization
- L2 Regularization
- Label Smoothing
- LAMB Optimizer
- Learning Rate
- Learning Rate Decay
- Learning Rate Schedule
- Lookahead Optimizer
- LoRA
- Loss Function
- Mean Squared Error
- Meta-Learning
- Mini-Batch Gradient Descent
- Mixed Precision Training
- Mixup
- Momentum
- Multi-Task Learning
- Nesterov Momentum
- Neural Architecture Search
- Online Learning
- Overfitting
- Pipeline Parallelism
- Pre-training
- Regularization
- RMSprop
- Self-Supervised Learning
- Sharded Data Parallelism
- Step Decay
- Stochastic Gradient Descent
- Student Model
- Teacher Model
- Tensor Parallelism
- Training
- Transfer Learning
- Triplet Loss
- Underfitting
- Vanishing Gradient
- Warmup
- Weight Decay
- ZeRO
Evaluation
45 terms · 7 landmarks
Metrics, benchmarks, and knowing whether a model is any good.
All 45 Evaluation terms
- A/B Testing
- Accuracy
- Attention Visualization
- AUC
- Average Precision
- Baseline Model
- Benchmark
- Benchmark Gaming
- BLEU Score
- Calibration
- Cohen's Kappa
- Confusion Matrix
- Cross-Validation
- Data Leakage
- Dice Coefficient
- Explainable AI
- F1 Score
- False Negative
- False Positive
- False Positive Rate
- Grad-CAM
- Intersection over Union
- Jaccard Index
- LIME
- Log Loss
- Matthews Correlation Coefficient
- Mean Absolute Error
- Out-of-Distribution
- Perplexity
- Precision
- Precision-Recall Curve
- R-squared
- Recall
- Robustness
- ROC Curve
- ROUGE Score
- Saliency Map
- Sensitivity
- SHAP
- Specificity
- Test Set
- Train-Test Split
- True Negative
- True Positive
- Validation Set
Language & LLMs
114 terms · 14 landmarks
Tokens, attention, transformers, prompting, and large language models.
All 114 Language & LLMs terms
- Alignment Tax
- Anaphora Resolution
- Assistant Response
- Attention Head
- Attention Is All You Need
- Attention Mask
- Attention Mechanism
- Attention Score
- Autoregressive Model
- BART
- Beam Search
- BERT
- Bidirectional Attention
- BOS Token
- BPE
- Causal Language Modeling
- Causal Mask
- Chain-of-Thought
- Chatbot
- Cloze Task
- Code Generation
- Constituency Parsing
- Constitutional AI
- Context Window
- Conversation History
- Coreference Resolution
- Cross-Attention
- Decoder-Only Model
- Dependency Parsing
- Dialogue State Tracking
- Embedding
- Emergent Abilities
- Encoder-Decoder
- Encoder-Only Model
- Entity Linking
- EOS Token
- Few-Shot Learning
- Flash Attention
- Foundation Model
- GPT
- Greedy Decoding
- Grounding
- Hallucination
- In-Context Learning
- Information Extraction
- Instruction Following
- Instruction Tuning
- Intent Recognition
- Language Modeling
- Language Understanding
- Large Language Model
- Latent Dirichlet Allocation
- Lemmatization
- Length Penalty
- Machine Translation
- Masked Language Modeling
- Multi-Head Attention
- N-gram
- Named Entity Recognition
- Natural Language Inference
- Natural Language Processing
- Neural Scaling Laws
- Nucleus Sampling
- Padding Token
- Paraphrase Detection
- Part-of-Speech Tagging
- Positional Encoding
- Prompt Engineering
- Prompt Template
- Prompt Tuning
- Query-Key-Value
- Question Answering
- Reading Comprehension
- Relation Extraction
- Repetition Penalty
- Retrieval-Augmented Generation
- Retrieval-Interleaved Generation
- RLHF
- Scaled Dot-Product Attention
- Self-Attention
- Semantic Role Labeling
- Semantic Similarity
- SentencePiece
- Sentiment Analysis
- Sequence-to-Sequence
- Shot
- Slot Filling
- Special Token
- Speech Recognition
- Stemming
- Stop Token
- Stop Words
- Subword Tokenization
- System Prompt
- T5
- Temperature
- Text Classification
- Text Generation
- Text Summarization
- Textual Entailment
- TF-IDF
- Token
- Tokenization
- Top-k Sampling
- Top-p Sampling
- Topic Modeling
- Transformer
- User Prompt
- Vector Database
- Vocabulary Size
- Word Sense Disambiguation
- Word2Vec
- WordPiece
- Zero-Shot Learning
Vision & Multimodal
54 terms · 6 landmarks
Seeing, generating, and mixing images, video, and audio.
All 54 Vision & Multimodal terms
- 3D Reconstruction
- Anchor Box
- Audio Processing
- Autonomous Vehicles
- Bounding Box
- Center Crop
- CLIP
- Color Jittering
- Computer Vision
- Data Augmentation in Vision
- Depth Estimation
- Diffusion Model
- Discriminator
- EfficientNet
- Face Detection
- Facial Recognition
- Faster R-CNN
- Feature Pyramid Network
- Generative Adversarial Network
- Generator
- Image Augmentation
- Image Classification
- Image Generation
- Image Inpainting
- Image Normalization
- Image Preprocessing
- ImageNet
- Inception
- Instance Segmentation
- Mask R-CNN
- Medical Diagnosis
- Multimodal Learning
- Multimodal Model
- Non-Maximum Suppression
- Object Detection
- Optical Character Recognition
- Optical Flow
- Pose Estimation
- R-CNN
- Random Crop
- Region Proposal Network
- ResNet
- RetinaNet
- Semantic Segmentation
- Squeeze-and-Excitation
- Stable Diffusion
- Style Transfer
- Super-Resolution
- Text-to-Speech
- Transfer Learning in Vision
- U-Net
- VGG
- Vision Transformer
- YOLO
Agents & RL
29 terms · 6 landmarks
Reinforcement learning, planning, tool use, and agents that act.
All 29 Agents & RL terms
- Actor-Critic
- Agent
- AI Agent
- AlphaGo
- Behavioral Cloning
- Deep Q-Network
- Domain Randomization
- Environment
- Exploration vs Exploitation
- Hierarchical RL
- Imitation Learning
- Inverse Reinforcement Learning
- Markov Decision Process
- Meta-RL
- Model-Based RL
- Model-Free RL
- Monte Carlo Tree Search
- Multi-Agent RL
- Offline RL
- Policy
- Policy Gradient
- PPO
- Q-Learning
- Reinforcement Learning
- Reward
- Sim-to-Real Transfer
- Tool Use
- Value Function
- World Model
Shipping AI
65 terms · 1 landmarks
Serving, infrastructure, efficiency, safety, and running AI for real.
All 65 Shipping AI terms
- Adversarial Attack
- Adversarial Example
- Adversarial Perturbation
- Adversarial Training
- AI Alignment
- AI Governance
- AI Safety
- Algorithmic Accountability
- Backdoor Attack
- Batch Processing
- Bias in AI
- Black Box
- Blue-Green Deployment
- Canary Deployment
- Certified Robustness
- CI/CD for ML
- Compute
- Content Moderation
- Data Drift
- Data Poisoning
- Data Versioning
- Differential Privacy
- Edge Deployment
- Experiment Tracking
- Explainability
- Fairness
- Feature Store
- Federated Learning
- FLOPS
- GPU
- gRPC
- Homomorphic Encryption
- Inference
- Inference Latency
- Interpretability
- Membership Inference
- MLOps
- Model Caching
- Model Card
- Model Compression
- Model Drift
- Model Endpoint
- Model Extraction
- Model Inversion
- Model Lineage
- Model Monitoring
- Model Performance Degradation
- Model Registry
- Model Reproducibility
- Model Retraining
- Model Serving
- Model Versioning
- Model Watermarking
- Neuromorphic Computing
- Prediction Confidence
- Privacy-Preserving ML
- Pruning
- Quantization
- Request Batching
- REST API
- Secure Multi-Party Computation
- Shadow Deployment
- Throughput
- TPU
- Trusted Execution Environment