Search papers, labs, and topics across Lattice.
1
0
2
Hyper-parallel decoding on a distilled 4B LLM matches frontier model information extraction accuracy while operating at an 8% compute footprint.