Search papers, labs, and topics across Lattice.
2
0
4
Target-aware data selection and VLM-based caption refinement can boost food image retrieval performance by over 19%, revealing the potential of fine-tuning multimodal models with curated data.
Open-KNEAD achieves up to 53% better nutrition estimates than leading closed models, all while keeping user data private and local.