Work with tokenization, next-token prediction, decoding, context windows, instruction following, and evaluation for large language models.