Search papers, labs, and topics across Lattice.
Affiliation:
2
4
5
4
WearableQA, a benchmark comprising 4,084 10-option multiple-choice questions constructed from the wearable time series, blood biomarkers, and demographics of 200 real users, provides a realistic and diagnostic benchmark for evaluating LLM reasoning over real-world wearable data.
Gemma 4's unified architecture and reasoning mode enable it to outperform larger models in human-rated tasks while maintaining high efficiency.