Onderzoek RSS

Modellen 🇺🇸

Self-Distilled Reasoning for Fine-Tuning Amazon Nova

Amazon has introduced the Self-Distilled Reasoning (SDR) method for supervised fine-tuning (SFT) of Amazon Nova 2 models. SDR generates chain-of-thought (CoT) reasoning paths for datasets lacking them, using the model itself. This addresses the problem of reasoning suppression and reduces catastrophic forgetting, improving target performance while retaining general capabilities.

Amazon Web ServicesAmazon Web Services
AWS ML blog24.07 · 05:06
Onderzoek 🇺🇸

Simulation for Physical AI: A Review

Gathering data for physical AI in the real world is a slow, expensive, and risky process. Simulation enables generating large volumes of photorealistic physical data through GPU parallelism. This article reviews popular simulation engines: MuJoCo, MuJoCo Warp, NVIDIA Isaac Sim and Isaac Lab, as well as the new physics engine Newton, developed by NVIDIA, Google DeepMind, and Disney Research.

NVIDIANVIDIA Google DeepMindGoogle DeepMind MuJoCoMuJoCo
Hugging Face blog24.07 · 05:05
Vers nieuws