Amazon researchers introduce Self-Distilled Reasoning to generate reasoning tokens for supervised fine-tuning. The technique tackles the reasoning suppression problem by synthesizing missing thinking traces in training data. Developers can use this approach to improve model performance on reasoning benchmarks when utilizing Amazon Nova.
Opening Kapyn…