The Proposed GIN Approach
GIN first extracts an instance feature with a backbone and ROI-Align. Its invariant amodal learning branch predicts a transformation, normalizes the feature, and learns an invariant shape prior. A complementary vanilla mask branch preserves residual appearance information. The feature refinement network combines both branches to produce the final amodal mask.