Google Introduces Auto Frame: Neural Network Changes Photo Perspective Retroactively
Google/DeepMind
Google DeepMind
Google Research
Google has launched Auto Frame in Google Photos, a new feature that uses ML models and generative AI to re-angle photos after they are taken. The system reconstructs a 3D scene, adjusts camera pose and focal length, and fills in hidden content via diffusion inpainting. It also automatically detects and corrects wide-angle distortion in portraits.
Google Research and Google DeepMind have introduced Auto Frame, a feature now live in Google Photos that edits photos by changing the camera perspective after capture. Unlike traditional cropping, it interprets a photo as a 3D scene, estimates a 3D point map per pixel along with original focal length, then uses classical 3D rendering to simulate a new camera position and intrinsics. The rendered image inevitably has holes where previously unseen content would appear; these are filled by a specially trained latent diffusion model that learns to reconstruct the target view from a re-rendered source view. For portraits, ML models detect faces and their 3D orientation to automatically suggest optimal framing and correct wide-angle distortion by adjusting virtual camera intrinsics. The fully automatic solution is available as the second rendition option within Auto frame candidates in Google Photos.
- Сокращения
- ML = Machine Learning
- 3D = Three Dimensional
Source: Google Research —
original
