Search papers, labs, and topics across Lattice.
3
0
7
Goku redefines the landscape of video editing datasets by enabling complex, multi-task editing capabilities that surpass traditional single-task limitations.
Ditch the VAE bottleneck: Representation Forcing lets you train unified multimodal models to generate high-quality images directly from pixels, rivaling VAE-based approaches without the architectural constraint.
Ditch the textual explanations: symbolic outputs like bounding boxes are the secret sauce for boosting multimodal verifier performance.