The Amazon ML Blog post explores a method for generating 'thinking tokens' for datasets lacking reasoning traces in Supervised Fine-Tuning (SFT) customization. This technique, called Self-Distilled Reasoning (SDR), aims to address the reasoning suppression problem and has been validated across three benchmarks.