Tag: Fashion Captioning
-
Comparative analysis of fashion captioning and multimodal fashion recommendation
Author: Maria De Los Angeles Gwendolyn Aglae Rippberger Fonseca Supervisor: Julia Neidhardt Abstract This thesis explores two main tasks: (1) fine-tuning image captioning models for fashion datasets and (2) evaluating different feature spaces for personalized fashion recommendations. We fine-tune state-of-the-art vision-language models – BLIP-2 and LLaVA – on two fashion datasets, H&M and FACAD, to…
