Search papers, labs, and topics across Lattice.
2
0
3
Target-aware data selection and VLM-based caption refinement can boost food image retrieval performance by over 19%, revealing the potential of fine-tuning multimodal models with curated data.
Injecting LLM-derived ingredient labels into food image segmentation boosts accuracy to state-of-the-art levels while keeping resource demands low.