Search papers, labs, and topics across Lattice.
Affiliation:
1
0
3
15
MoNe slashes compute and memory costs by 80% for long-context inference while enabling Transformers to handle context lengths far beyond their original limits.