Diffusion models, RLHF, LoRA, and retrieval-augmented generation
Generate images by learning to reverse a gradual noising process
Ground language model outputs in retrieved documents to reduce hallucination
Adapt large models by training tiny low-rank matrices instead of all the weights
Align language models with human preferences using reward models and policy optimization