The Proposed A3D Approach
A3D first predicts an amodal bounding box and visible mask. Its category-specific 3D modeling branch reconstructs a complete shape prior, while a viewpoint estimator predicts camera parameters. A differentiable renderer projects the reconstructed model into a coarse 2D amodal mask, which is then combined with image and visible-mask cues for final refinement.