Model Card: Fashion CLIP Disclaimer: The model card adapts the model card from here. Model Details UPDATE (10/03/23): We have updated the model! We found that laion/CLIP ViT B 32 laion2B s34B b79K checkpoint (thanks Bin!) worked better than original OpenAI CLIP on Fashion. We thus fine tune a newer (and better!) version of FashionCLIP (henceforth FashionCLIP 2.0), while keeping the architecture the same. We postulate that the perofrmance gains afforded by laion/CLIP ViT B 32 laion2B s34B b79K are due to the increased training data (5x OpenAI CLIP data). Our thesis, however, remains the same fine tuning laion/CLIP on our fashion dataset improved zero shot perofrmance across our benchmarks. See the below table comparing weighted macro F1 score across models. Model FMNIST KAGL DEEP OpenAI CLIP 0.66 0.63 0.45 FashionCLIP 0.74 0.67 0.48 Laion CLIP 0.78 0.71 0.58 FashionCLIP 2.0 0.83 0.73 0.62 FashionCLIP is a CLIP based model developed to produce general product representations for fashion concepts. Leveraging the pre trained checkpoint (ViT B/32) released by OpenAI, we train FashionCLIP on a large, high quality novel fashion dataset to study whether domain specific fine tuning of CLIP…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy