Search papers, labs, and topics across Lattice.
Affiliation:
2
7
5
20
Transforming zero-reward training instances into valuable learning opportunities, HCGRec cuts down ineffective samples from over 70% to under 20%.
Current LLM evaluation benchmarks often conflate chatbots and true AI agents, leading to misaligned research efforts, but this survey provides a framework for targeted evaluation based on environmental complexity and agent capabilities.