Search papers, labs, and topics across Lattice.
This paper presents the Symbolic Geometric Agent (SGA), a novel module designed to enhance the spatial correctness and visual legibility of educational animations generated by Large Language Models (LLMs). By intercepting LLM-generated code and performing partial execution to extract symbolic scene graphs, SGA effectively identifies and refines geometric occlusions that traditional frameworks overlook. Experimental results demonstrate that SGA achieves a peak Manim Visual Quality Score (MVQS) of 73.11, marking a 16.1% improvement over baseline performance across multiple LLM backbones and animation pipelines.
SGA boosts the visual quality of LLM-generated educational animations by over 16% by tackling geometric occlusions that others ignore.
Recent work leverages Large Language Models (LLMs) to generate executable code for pedagogical animations using libraries such as Manim. However, ensuring spatial correctness and visual legibility remains challenging, as existing frameworks emphasize pedagogical content while overlooking geometric occlusions. We propose the Symbolic Geometric Agent (SGA), a plug-and-play module for code-centric animation pipelines that intercepts LLM-generated code, performs partial execution to extract symbolic scene graphs, and applies targeted refinement when spatial conflicts are detected. We further introduce the Manim Visual Quality Score (MVQS), a deterministic rendering-free proxy for spatial integrity. Experiments on the MMMC-Code benchmark across four LLM backbones and two agentic pipelines show that SGA achieves a peak MVQS of 73.11 (Code2Video + GPT-5.1), corresponding to a 16.1% relative improvement over the raw baseline, and improves MVQS in 7 of 8 backbone x pipeline configurations.