Search papers, labs, and topics across Lattice.
East China Normal University
3
0
5
Performance of large reasoning models drops significantly as logical complexity rises, revealing critical gaps in current evaluation benchmarks.
Ignoring the nuanced interplay between services and hosts in microservice architectures leaves nearly 50% of root causes undiscovered.
R2IF achieves up to 34.62% better performance in function calling accuracy, bridging the gap between reasoning and decision-making in LLMs.