Search papers, labs, and topics across Lattice.
This paper introduces TrainsBiolab, a comprehensive RGB-D dataset designed to enhance visual perception in autonomous biomedical laboratories by addressing the challenges of cluttered scenes with transparent objects. It comprises 161,315 frames across 98 scenes, featuring detailed annotations for 15 object types, including 6D poses and depth information, which are crucial for tasks like segmentation and pose estimation. The dataset also establishes benchmarks for these tasks, facilitating improved performance in robot manipulation within real-world lab environments.
TrainsBiolab reveals the complexities of manipulating transparent objects in cluttered scenes, providing a goldmine of data that could transform how robots perceive and interact with their environments.
Autonomous biomedical laboratories increasingly rely on visual perception to recognize, localize, and manipulate transparent plasticware, yet high-quality real-world datasets for this setting remain limited. The scarcity of domain-relevant data is particularly restrictive in cluttered multi-object scenes, where mutual occlusion and view-dependent appearance changes remain challenging even for contemporary visual foundation models. Existing transparent-object datasets have advanced segmentation, depth, and pose estimation, but they usually do not evaluate the combined setting of multi-object clutter, occlusion, and calibrated multi-view capture that characterizes real laboratory manipulation scenes. To address this gap, we present TrainsBiolab, a real-world RGB-D dataset of cluttered transparent biomedical objects captured as calibrated multi-view sequences. TrainsBiolab contains 161,315 frames from 98 scenes and 1.03M instance annotations over 15 laboratory object types, including 6D poses, full and visible masks, depth, and per-frame camera calibration. The dataset is organized along three axes that reflect operational difficulty: object category, the total number of objects in a frame, and camera viewpoint. We further define dataset-centric benchmarks for segmentation, depth estimation and completion, and 6D pose estimation, and report a system-level robot manipulation evaluation enabled by the released annotations and calibrations. By focusing on repeated transparent instances, clutter, and multi-view laboratory capture, TrainsBiolab provides a resource for segmentation, depth estimation, 6D pose estimation, and multi-view reasoning in autonomous laboratory manipulation. Project page: https://dualtransparency.github.io/TransBiolab/.