Search papers, labs, and topics across Lattice.
This paper introduces TreeAgent, a multi-agent system that integrates expert decision trees with Vision-Language Models (VLMs) to automate bias labeling in forestry, specifically for tree height classification. By leveraging a Decoupled Declarative Decision (D3) Framework, the system generalizes across various expert-defined decision structures while utilizing multi-agent voting to enhance the reliability of VLM outputs. The results demonstrate that TreeAgent not only surpasses traditional supervised machine learning baselines but also significantly reduces the expert labeling effort required, highlighting its potential for efficient and interpretable annotation processes.
Automated bias labeling in forestry can be achieved with 50% less expert input while outperforming conventional ML methods.
Human-labeled data are widely used as reference annotations in ML, despite known variability across annotators in many expert-driven domains. In addition, expert annotation is slow, inconsistent, and remains a major bottleneck for scaling tasks like tree height bias classification in forestry remote sensing. We propose a multi-agent system (MAS) that orchestrates expert decision trees with Vision-Language Models (VLMs), treating the decision tree as a structural prior while VLMs perform localized semantic perception at individual nodes, with multi-agent voting to mitigate VLM stochasticity. We formalize a Decoupled Declarative Decision (D3) Framework that enables zero-modification generalization across diverse expert-defined decision structures. On a tree bias classification testbed, our framework outperforms supervised ML baselines and reduces the amount of expert labeling effort required. These results suggest that agentic orchestration of VLMs with expert priors can reproduce expert-defined labeling procedures at substantially lower annotation cost while maintaining interpretability.