Search papers, labs, and topics across Lattice.
This paper introduces iARCS, an iterative agentic reinforcement learning framework designed to enhance synthetic 3D scene generation by aligning it with natural-language task requirements. By employing a two-stage strategy that combines universal-reward pretraining and task-specific fine-tuning with LLM-generated reward programs, iARCS significantly improves the fidelity of functional constraints such as walkability and reachability. Experimental results demonstrate that scenes generated by iARCS not only meet critical task requirements but also enhance the performance of existing scene generators, highlighting its utility as a synthetic data generation tool.
iARCS transforms 3D scene generation by ensuring that synthetic environments meet essential functional constraints while maintaining diversity and realism.
Synthetic 3D scene generation is increasingly used as a data source for computer vision and embodied AI, but existing generators often optimize perceptual realism without reliably satisfying task-critical functional constraints. This mismatch limits the usefulness of synthetic data for downstream training, where accessibility, traversability, and spatial rule compliance are often essential. We present iARCS, an iterative agentic reinforcement learning framework that adapts a pretrained scene generator to natural-language task requirements. iARCS uses a two-stage strategy: universal-reward pretraining to improve physical plausibility and layout quality, followed by task-specific fine-tuning with LLM-generated reward programs that are iteratively refined from training feedback. Experiments show improved constraint fidelity on walkability, reachability, and clearance-focused tasks, effective task-specific constraint optimization, and competitive scene diversity. We further show that data generated by iARCS improves a base generator, supporting its value as a practical synthetic data generation tool rather than only a controllable scene editing method.