Search papers, labs, and topics across Lattice.
Center for Machine Learning, Technical University of Munich Munich
2
0
3
P-JEPA achieves state-of-the-art action classification on long procedural videos while using an order of magnitude fewer parameters than existing models.
Traditional scene graph methods falter in capturing the temporal structure of OR activities, while a new vision-only model achieves superior performance in multi-role action recognition.