Just for you
Your progress
A plain record of where you are — stored only in this browser. Mark concepts as you go; revisit anything flagged. Export a copy if you switch devices.
Concepts understood
0 / 81
Chapters completed
0 / 6
Papers read
0 / 57
Videos watched
0 / 36
Experiments done
0 / 19
Must-know concepts understood: 0 of 68. The site contains only published material so far; totals grow as chapters are added.
Concepts by status
Not started81
- Activation Functions
- AI Winters
- From MLPs to Transformers: The Architecture Story
- The Artificial Neuron
- Attention
- Backpropagation
- Batch Normalization
- Conditional Probability and Bayes' Theorem
- Causal Masking
- Probability of Sequences
- The Chain Rule
- Computational Graphs and Autodiff
- Cross-Entropy Loss
- Data Leakage
- Decision Trees and Random Forests
- Derivatives and Gradients
- Distribution Shift
- Dot Product
- Dropout
- Embeddings
- Entropy
- Evaluation Metrics for Classifiers
- Expected Value and Variance
- Expert Systems
- Hand-Crafted Features vs Learned Features
- Features, Labels and Tasks
- Feed-Forward Sublayer (MLP)
- The Fixed-Vector Bottleneck
- The Forward Pass
- Generalization, Overfitting and Underfitting
- GloVe
- Gradient Descent
- Weight Initialization
- k-Means Clustering
- KL Divergence
- The Knowledge-Acquisition Bottleneck
- Knowledge Representation
- Language Modeling
- Layer Normalization
- Supervised, Unsupervised and Self-Supervised Learning
- Linear Regression
- Logic and Rules
- Logistic Regression
- Loss Functions
- LSTMs and GRUs
- Matrix Multiplication
- Momentum and Adam
- Multi-Head Attention
- Multilayer Perceptron (MLP)
- N-Gram Models
- Naive Bayes
- Neural Language Model
- Neural Machine Translation
- One-Hot Encoding
- Principal Component Analysis (PCA)
- The Perceptron
- Perplexity
- Planning
- Positional Encoding
- Probability and Distributions
- Regularization
- Representation Learning
- Residual Connections
- Recurrent Neural Networks
- From Rules to Learning
- Sampling and Uncertainty
- Search
- Self-Attention
- Sequence-to-Sequence Models
- Softmax
- Stochastic Gradient Descent (SGD)
- Support Vector Machines
- Symbolic AI
- Tensors and Shapes
- Text as Data
- The Transformer Block
- Encoder, Decoder & Encoder–Decoder
- The Turing Test
- Vanishing and Exploding Gradients
- Vectors
- Word2Vec
Learning0
Understood0
Revisit0
Your data
Progress lives in this browser's local storage. Nothing is sent anywhere.