Search papers, labs, and topics across Lattice.
This paper introduces TCA-SIR, a novel approach to Scientific Inspiration Retrieval (SIR) that reformulates the problem by focusing on target-conditioned abstractions rather than mere topical similarity. By learning to generate transferable abstract principles tailored to specific target problems, TCA-SIR significantly enhances the retrieval process, achieving over a 10 percentage point improvement in HitRate@top4% compared to existing methods. The results demonstrate that TCA-SIR not only retrieves more relevant scientific inspirations but also provides clearer interpretability of the mechanisms involved.
Target-conditioned abstractions can boost scientific inspiration retrieval accuracy by over 10%, transforming how we leverage past research for new hypotheses.
Scientific hypothesis generation for AI for Science typically involves Scientific Inspiration Retrieval (SIR) followed by hypothesis composition. Existing SIR methods rank papers by topical similarity and do not explicitly represent how a candidate inspiration transfers to a target problem. This is especially limiting for remote inspirations, whose value often lies in reusable problem-solving principles rather than topical overlap. Motivated by how humans abstract transferable aspects of a source and remap them to a new target, we reformulate SIR as target-conditioned abstraction (TCA). The retrieval object is a transferable abstract principle extracted from a candidate specifically for the target. We present TCA-SIR, which learns to generate target-conditioned abstractions and uses their representations to predict transferability. On ResearchBench, TCA-SIR outperforms prior SIR methods and direct LLM retrieval, improving HitRate@top4% over MOOSE-Chem by more than 10 percentage points. Learned abstractions also recover target-relevant mechanisms more clearly than an untrained TCA prompt, yielding both stronger retrieval and an interpretable rationale for scientific inspiration.