Search papers, labs, and topics across Lattice.
Orange Team, Youku Moku-Lab, HUJING Digital Media & Entertainment Group
3
0
7
Transforming a single image into a fully navigable 3D world in real-time could revolutionize how we interact with visual environments.
By incorporating language guidance into federated learning, SurgFed tackles the long-standing problem of tissue and task heterogeneity in surgical video understanding, leading to improved segmentation and depth estimation across diverse surgical settings.
Current vision-language models are surprisingly bad at surgical safety reasoning, failing to integrate phase information to identify safe operative zones, but a new RLHF-tuned model closes the gap.