Search papers, labs, and topics across Lattice.
2
0
5
2
LLMs struggle to match human decision-making in e-commerce, achieving only 27.3% of the final net assets in a year-long simulation.
Fine-grained reward modeling, achieved by selectively dropping instruction requirements, unlocks substantial improvements in writing-centric generation tasks.