Shaving Milliseconds Off LoRA Inference for Diffusion Models
Notes from swapping adapters at inference time without eating a full UNet forward pass on every merge.
diffusion-modelslorainferencepytorch
Chapters from the build log — write-ups on AI/ML experiments, systems I've shipped, and whatever rabbit hole I fell into that week.