A Unified Model with Structured Output for Fashion Images Classification

A picture is worth a thousand words. Albeit a cliché, for the fashion industry, an image of a clothing piece allows one to perceive its category (e.g., dress), sub-category (e.g., day dress) and properties (e.g., white colour with floral patterns). The seasonal nature of the fashion industry creates a highly dynamic and creative domain with evermore data, making it unpractical to manually describe a large set of images (of products). In this paper, we explore the concept of visual recognition for fashion images through an end-to-end architecture embedding the hierarchical nature of the annotations directly into the model. Towards that goal, and inspired by the work of [7], we have modified and adapted the original architecture proposal. Namely, we have removed the message passing layer symmetry to cope with Farfetch category tree, added extra layers for hierarchy level specificity, and moved the message passing layer into an enriched latent space. We compare the proposed unified architecture against state-of-the-art models and demonstrate the performance advantage of our model for structured multi-level categorization on a dataset of about 350k fashion product images.

READ FULL TEXT
research
03/02/2021

Multi-Level Attention Pooling for Graph Neural Networks: Unifying Graph Representations with Multiple Localities

Graph neural networks (GNNs) have been widely used to learn vector repre...
research
09/10/2019

Structured Modeling of Joint Deep Feature and Prediction Refinement for Salient Object Detection

Recent saliency models extensively explore to incorporate multi-scale co...
research
05/09/2021

Dispatcher: A Message-Passing Approach To Language Modelling

This paper proposes a message-passing mechanism to address language mode...
research
10/17/2009

Faster Algorithms for Max-Product Message-Passing

Maximum A Posteriori inference in graphical models is often solved via m...
research
11/02/2016

CRF-CNN: Modeling Structured Information in Human Pose Estimation

Deep convolutional neural networks (CNN) have achieved great success. On...
research
02/22/2022

Equivariant Graph Hierarchy-Based Neural Networks

Equivariant Graph neural Networks (EGNs) are powerful in characterizing ...
research
05/03/2023

Fashionpedia-Ads: Do Your Favorite Advertisements Reveal Your Fashion Taste?

Consumers are exposed to advertisements across many different domains on...

Please sign up or login with your details

Forgot password? Click here to reset