Self-Distilled Reasoning for Fine-Tuning Amazon Nova
Amazon has introduced the Self-Distilled Reasoning (SDR) method for supervised fine-tuning (SFT) of Amazon Nova 2 models. SDR generates chain-of-thought (CoT) reasoning paths for datasets lacking them, using the model itself. This addresses the problem of reasoning suppression and reduces catastrophic forgetting, improving target performance while retaining general capabilities.
Amazon Web Services



