Search papers, labs, and topics across Lattice.
This paper introduces LiD-GLM, a novel hybrid model that integrates invertible residual neural networks (i-ResNets) with generalized linear models to achieve a balance between interpretability and flexibility. By constraining the Lipschitz constant of the i-ResNets, the model ensures that deviations from traditional linear assumptions are controlled, allowing for nonlinear parameter estimation while maintaining stochastic monotonicity. The results demonstrate that LiD-GLM can effectively learn complex interactions without sacrificing interpretability, offering a user-defined trade-off between model complexity and clarity.
A controlled deviation from traditional models allows for flexible, interpretable hybrid models that can learn complex interactions without losing clarity.
The combination of traditional statistical models and neural network (NN) components into semi-structured hybrid models is an intriguing approach to construct models that, ideally, combine traditional interpretability with the unprecedented flexibility of NNs. In order to preserve interpretability, it is usually necessary to restrict the NN components to prevent them from dominating the model. However, existing methods that enforce structural constraints on their NN components severely limit their models' flexibility; in contrast, methods that only enforce weak, indirect constraints lose meaningful interpretability. The method we propose therefore leverages invertible residual neural networks (i-ResNets) to equip generalized linear models with both nonlinear parameter estimation and a flexible correction of their distributional assumptions while always retaining stochastic monotonicity of the modeled distribution in the (formerly linear) predictor. The i-ResNets correspond to a controlled deviation from identity and by constraining their Lipschitz constant one can rigorously limit and quantify how far the hybrid model deviates from its traditional counterpart. This enables a user-specifiable compromise between flexibility and interpretability without limiting the structure of nonlinear and interaction effects that can be learned. Furthermore, we develop specific inherent interpretation techniques for our model and enforce model identifiability through an adapted post-hoc orthogonalization.