Multi-view diffusion models typically struggle with mirrors, treating reflections as visual noise rather than coherent spatial cues. Ref-GeNVS addresses this limitation without retraining or fine-tuning the underlying backbone. By estimating the mirror plane, it mirrors camera poses to construct virtual viewpoints, treating reflection as an explicit geometric counterpart to the primary scene. A combination of mirror-gated attention and reflection injection then guides the diffusion process, allowing the model to consistently generate novel views and recover occluded structures visible only inside the glass.
No heat snapshots are available in the last 24 hours.