Feb 24, 2026arXiv:2602.21203

Squint: Fast Visual Reinforcement Learning for Sim-to-Real Robotics

Abdulaziz Almuzairee, Henrik I. Christensen

AI Summary

The paper introduces Squint, a visual Soft Actor-Critic method optimized for fast wall-clock training in sim-to-real robotics. Squint leverages parallel simulation, a distributional critic, resolution squinting, layer normalization, a tuned update-to-data ratio, and an optimized implementation to overcome challenges associated with high-dimensional visual inputs. Experiments on the SO-101 Task Set in ManiSkill3 demonstrate that Squint achieves faster training times compared to prior visual off-policy and on-policy methods, with successful sim-to-real transfer to a physical SO-101 robot within 15 minutes using a single RTX 3090.

Key Contribution

Visual robot policies can now be trained in just minutes on a single GPU, thanks to Squint's optimized visual Soft Actor Critic approach.

Abstract

Visual reinforcement learning is appealing for robotics but expensive -- off-policy methods are sample-efficient yet slow; on-policy methods parallelize well but waste samples. Recent work has shown that off-policy methods can train faster than on-policy methods in wall-clock time for state-based control. Extending this to vision remains challenging, where high-dimensional input images complicate training dynamics and introduce substantial storage and encoding overhead. To address these challenges, we introduce Squint, a visual Soft Actor Critic method that achieves faster wall-clock training than prior visual off-policy and on-policy methods. Squint achieves this via parallel simulation, a distributional critic, resolution squinting, layer normalization, a tuned update-to-data ratio, and an optimized implementation. We evaluate on the SO-101 Task Set, a new suite of eight manipulation tasks in ManiSkill3 with heavy domain randomization, and demonstrate sim-to-real transfer to a real SO-101 robot. We train policies for 15 minutes on a single RTX 3090 GPU, with most tasks converging in under 6 minutes.

Computer Vision Robotics & Embodied AI Training Efficiency & Optimization

Citation Metrics

Citations0

Influential citations0

References0

Year2026

VenueN/A

Related Papers

Finding related papers...

Search

Squint: Fast Visual Reinforcement Learning for Sim-to-Real Robotics

Related Papers