Search papers, labs, and topics across Lattice.
Email:
2
0
5
6
LLMs can outperform random selection in materials optimization, but their effectiveness varies widely across tasks and contexts.
Naive RL fine-tuning for code generation can lead to LLMs regurgitating the same solutions, but penalizing code similarity boosts performance even more than directly optimizing for pass@k.