This hands-on tutorial explores response engineering techniques that improve language-model outputs at inference time, including Best-of-N selection, self-consistency, Reflexion, and Mixture-of-Agents.
Response Engineering for Training-Free Alignment of LLMs
A practical tutorial on Best-of-N, self-consistency, Reflexion, and Mixture-of-Agents for improving model responses without retraining.
Tutorial resources Watch the tutorial segment or explore the slide deck.
Response Engineering · starts at 2:06:50
Open on YouTubeNAIRR 2026 slide deck
Open slides View source
Discussion
Comments & questions
Join the conversation with GitHub.