Search papers, labs, and topics across Lattice.
2
0
5
NebulaExp reveals that a meticulously curated dataset and innovative reinforcement learning strategies can boost LLM performance significantly, achieving up to 4.43 points improvement in instruction-following tasks with minimal data.
GTokenLLMs suffer from a text-dominant bias, but RGLM offers a way to fix this by reconstructing graph information directly from the LLM's graph token outputs.