Search papers, labs, and topics across Lattice.
6
0
9
5
AtumAI slashes the onboarding time for datacenter control-plane policies from months to mere minutes while consistently outpacing expert-engineered solutions.
ACID redefines the caching paradigm in video generation, achieving unprecedented speedups while maintaining visual fidelity by dynamically adjusting thresholds based on drift signals.
Sangam slashes latency for diffusion language models by intelligently managing prefill and decode processes, revealing a new paradigm for efficient LLM serving.
Forget hand-tuning: AutoScout automates ML system configuration, delivering up to 3x speedups over expert settings by jointly optimizing structural and execution parameters.
Stop hand-writing CUDA kernels: CUCo's agent-driven approach co-optimizes computation and communication, slashing LLM training/inference latency by up to 1.57x.
Forget hand-crafted benchmarks: this paper shows how LLMs can continuously generate relevant evaluation datasets for enterprise AI agents from just a few semi-structured documents.