Search papers, labs, and topics across Lattice.
3
0
3
0
CHIARA achieves up to 13.48x speedup in Allreduce operations by intelligently managing hardware hierarchy and communication patterns.
Achieving higher compression rates without sacrificing quality, this method eliminates the need for mesh storage in unstructured volumes.
Datalog on GPUs just got a whole lot faster: SRDatalog achieves up to 47x speedups by finally making worst-case optimal joins practical on GPUs.