food-image-recipe

Tag

Cards List
#food-image-recipe

SIMMER: Cross-Modal Food Image–Recipe Retrieval via MLLM-Based Embedding

arXiv cs.CL · 2026-04-20 Cached

SIMMER proposes a novel MLLM-based embedding approach for cross-modal food image-recipe retrieval, replacing traditional dual-encoder architectures with a unified encoder and achieving state-of-the-art results on the Recipe1M dataset with significant improvements over prior methods.

0 favorites 0 likes
← Back to home

Submit Feedback