Search papers, labs, and topics across Lattice.
Apodex Discovery introduces a comprehensive framework for evaluating and building discoverative AI, utilizing a heavy-duty solver that integrates a foundation model with problem-scouting, environment-task abstractions, and independent evaluation metrics. By surveying 561 industries to identify 423 real-world problems, the framework enables AI systems to engage in extended, stateful investigations that are both verifiable and focused on genuine discovery. Key results demonstrate that Apodex outperformed existing state-of-the-art methods in AAV capsid design and significantly improved drug repurposing outcomes, showcasing its potential to redefine AI evaluation practices.
Apodex Discovery redefines AI evaluation by enabling verifiable investigations that surpass traditional benchmarks, achieving a 7% improvement in AAV capsid design outcomes.
Apollo did not reach the Moon merely because its engineers could solve difficult equations. It succeeded by turning a distant ambition into a mission architecture of explicit objectives, simulation, verification, and repeated correction. AI now faces a similar transition: frontier models can solve difficult tasks once the problem, tools, and success criteria are specified, yet consequential real-world challenges rarely arrive in an executable or verifiable form. We introduce Apodex Discovery, a framework for building and evaluating discoverative AI through the heavy-duty solver, a system comprising a foundation model, harness, tools, and control policies that pursues extended, stateful, verifiable investigations. It has three core components. First, a problem-scouting process surveyed 561 industries across 16 sectors, assembled 423 high-value real-world problems, and selected 20 for the initial release. Second, a common environment-task-episode abstraction provides data, tools, constraints, feedback, trajectory recording, and verification of intermediate artifacts and final submissions. Third, HDS6 evaluates Tools, Repair, Alternatives, Coherence, Evidence, and Scope independently of final-task success. In AAV capsid design, Apodex surpassed the published state of the art by 7% across viability, tropism, structure prediction, and generative design. In drug repurposing and reformulation, a task-specific biomedical environment improved the mean normalized prediction score of GPT-5.5 and GPT-5.6-sol by 2.5 and 7.6 points over the same closed-book backbone. Controlled ablations show that the fixed TRACES episode interface enables attribution of performance differences to specific solver components. Apodex Discovery moves AI evaluation beyond predefined benchmarks toward verifiable investigations aimed at genuine discovery.