Search papers, labs, and topics across Lattice.
3
0
6
0
Despite high diagnostic accuracy, LLMs fail to choose valid recovery actions for over 60% of incidents, exposing a critical flaw in their operational utility.
RATrain achieves a remarkable 1.35x end-to-end speedup for LLM training on bandwidth-limited supercomputers, challenging the notion that high-bandwidth environments are necessary for efficiency.
LLM inference on supercomputers doesn't have to be a bottleneck: THInfer achieves up to 84% higher throughput than A800 GPUs by co-designing hardware-aware kernels and a communication-optimized pipeline.