Large language models in depth — tokenization, attention, pretraining, fine-tuning, alignment, and evaluation.