Advanced
Open
Pro
Building an Offline Evaluation Set for Outpainting
Part of the AI Engineer Interview path →
Part of the Generative Vision & Image AI System Design path →
Your team has a solid offline evaluation pipeline for inpainting (removal and generate-with-prompt): synthetically mask a region of a real, held-out image and compare the model's fill against the known original pixels. A PM asks you to build the equivalent evaluation for outpainting before the feature ships.
- Explain why the inpainting evaluation approach can't be applied to outpainting unmodified, and what's missing.
- Propose a concrete construction for an outpainting offline evaluation set that provides genuine ground truth.
- Even with your proposed ground-truth set, what aspect of outpainting quality would still require human evaluation rather than an automated metric, and why?
Share this question