Search papers, labs, and topics across Lattice.
7
0
10
Bridging the Context Gap in T2I models, Qwen-Image-Agent achieves state-of-the-art performance by intelligently constructing context from user input and external sources.
A novel reward system boosts Qwen-Image-2.0's performance, achieving a 2.61 point increase in overall quality and significant gains in both text-to-image and image editing tasks.
Qwen-RobotManip achieves a 20% relative improvement over the previous state-of-the-art in robotic manipulation, showcasing unprecedented generalization capabilities from diverse, open-source datasets.
Qwen-RobotNav redefines navigation by allowing real-time reconfiguration of strategies, achieving unprecedented flexibility and performance across diverse tasks.
Language-driven video generation in Qwen-RobotWorld achieves unprecedented accuracy in predicting robotic actions, outperforming existing models across key benchmarks.
Rethinking few-step distillation reveals that the training pipeline's organization is as crucial as the distillation objectives themselves.
Existing text-to-image benchmarks miss the mark on real-world artistic creation, but Qwen-Image-Bench finally provides a creator-centric evaluation that reliably distinguishes state-of-the-art models.