Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

{\Phi}eat: Physically Grounded Material Feature Representation

About

While foundation models have emerged as general-purpose visual backbones, their representations are primarily optimized for semantics and lack explicit modeling of physical factors, such as reflectance, hindering their efficacy in tasks requiring explicit material reasoning. We introduce $\Phi$eat$, a novel material-grounded visual backbone that encourages a representation sensitive to material identity, including reflectance and mesostructure. Instead of relying on generic data augmentations, we pretrain our model by contrasting observations of the same material under controlled variations in lighting and geometry. This encourages invariance to extrinsic factors while preserving sensitivity to intrinsic material properties. We show that the resulting representation provides strong priors for material-centric tasks, including feature-based material selection and classification. Our results demonstrate that physically inspired weak supervision is an effective strategy for learning representations tailored to material perception.

Giuseppe Vecchio, Adrien Kaiser, Claudia Cuttano, Rouffet Romain, Rosalie Martin, Elena Garces, Tamy Boubekeur• 2025

Related benchmarks

TaskDatasetResultRank
Material SelectionDuMaS
L1 Error25.4
10
k-NN classificationSynthetic (test)
Accuracy71.8
8
Material ConsistencyMaterial renderings Illumination variations
Hamming Distance0.221
4
Material ConsistencyMaterial renderings Geometry variations
Hamming Distance30.5
4
Showing 4 of 4 rows

Other info

Follow for update