Search papers, labs, and topics across Lattice.
Affiliation:
1
0
2
10
DARD achieves a remarkable 2.71脳 speedup in dLLM inference while enhancing output quality by 4.35 points on CIDEr, redefining the efficiency landscape for large language models.