This blog discusses the challenges and innovations in post-training generative recommender systems, specifically through the lens of a novel algorithm called Advantage-Weighted Supervised Fine-tuning (A-SFT). The authors explore how generative recommenders can improve user experience by integrating user feedback, address issues with traditional reinforcement learning techniques, and demonstrate the effectiveness of A-SFT compared to existing methods in optimizing recommendations.