Search papers, labs, and topics across Lattice.
2
0
4
Auditing Chinese web content reveals pervasive pollution that shifts over time, challenging the integrity of LLM training data.
Prompted LLMs struggle with code error classification, often misclassifying logic errors, while smaller finetuned models lead the way in accuracy.