Search papers, labs, and topics across Lattice.
To address the severe scarcity of historical visual records in cross-domain instance retrieval, this work systematically evaluates whether degradation-based synthetic aging of modern images can substitute or augment historical training data using an EfficientNetV2-M backbone. While fully replacing authentic historical images degrades bidirectional mean Recall@1 from 86.56% to 81.27%, synthetically completing missing pairs under scarce conditions boosts R@1 by up to 3.69 percentage points at 25% coverage. These findings reveal that while synthetic aging fails to fully capture genuine temporal domain variability, it effectively bridges cross-domain retrieval gaps specifically by expanding identity coverage rather than increasing raw sample volume.
Synthetic domain shifts cannot replace authentic historical variance, but using synthetic degradation to complete missing paired identities yields up to a 3.69 percentage point R@1 boost under severe data scarcity.
Cultural heritage collections often contain contemporary and historical visual records of the same physical object. Linking these records is difficult because corresponding images may differ in viewpoint, acquisition conditions, color reproduction, framing, resolution, and degradation, while genuine historical images are frequently scarce. This study investigates whether synthetically aged contemporary images can replace or complement missing historical training data in bidirectional instance-level retrieval. Synthetic old-domain images are generated using degradation-oriented transformations. An EfficientNetV2-M model is evaluated on identity-disjoint training, validation, and test sets across three dataset partitions and three training seeds. Mixed real-synthetic training is compared with real-only baselines using proportionally scaled and fixed 300-batch-per-epoch schedules. Complete replacement of genuine historical images reduced bidirectional mean R@1 from 86.56% to 81.27%, showing that synthetic aging does not reproduce the full genuine old-domain variability. Increasing the number of independently generated synthetic variants provided no consistent improvement. Under controlled scarcity, however, synthetic completion improved mean R@1 by 3.69 percentage points at 25% genuine historical coverage and by 2.92 points at 50%, relative to the proportionally scaled real-only baselines. At 75%, the gain decreased to 2.00 points, while performance remained comparable to the complete-real-data reference. Fixed-schedule real-only controls did not reproduce these improvements. The results indicate that genuine and synthetic observations are complementary. Synthetic completion primarily benefits retrieval by extending cross-domain identity coverage rather than by increasing training exposure, with its contribution gradually decreasing as genuine historical coverage increases.