Search papers, labs, and topics across Lattice.
3
0
7
7
Decoupling perception from reasoning in visual tasks leads to a remarkable 93.2% accuracy on V-Star, showcasing a new paradigm for fine-grained visual reasoning.
Multimodal models can "see" the image but still fail at reasoning because the visual input distracts the routing mechanism from activating the right experts.
AI agents can now learn durable skills instead of constantly "reinventing the wheel," thanks to SkillNet's infrastructure for creating, evaluating, and connecting AI skills at scale.