-
RbX: Region-based explanations of prediction models
Authors:
Ismael Lemhadri,
Harrison H. Li,
Trevor Hastie
Abstract:
We introduce region-based explanations (RbX), a novel, model-agnostic method to generate local explanations of scalar outputs from a black-box prediction model using only query access. RbX is based on a greedy algorithm for building a convex polytope that approximates a region of feature space where model predictions are close to the prediction at some target point. This region is fully specified…
▽ More
We introduce region-based explanations (RbX), a novel, model-agnostic method to generate local explanations of scalar outputs from a black-box prediction model using only query access. RbX is based on a greedy algorithm for building a convex polytope that approximates a region of feature space where model predictions are close to the prediction at some target point. This region is fully specified by the user on the scale of the predictions, rather than on the scale of the features. The geometry of this polytope - specifically the change in each coordinate necessary to escape the polytope - quantifies the local sensitivity of the predictions to each of the features. These "escape distances" can then be standardized to rank the features by local importance. RbX is guaranteed to satisfy a "sparsity axiom," which requires that features which do not enter into the prediction model are assigned zero importance. At the same time, real data examples and synthetic experiments show how RbX can more readily detect all locally relevant features than existing methods.
△ Less
Submitted 16 October, 2022;
originally announced October 2022.
-
A Graph Approach to Simulate Twitter Activities with Hawkes Processes
Authors:
Ao Qu,
Ismael Lemhadri
Abstract:
The rapid growth of social media has been witnessed during recent years as a result of the prevalence of the internet. This trend brings an increasing interest in simulating social media which can provide valuable insights to both academic researchers and businesses. In this paper, we present a step-by-step approach of using Hawkes process, a self-activating stochastic process, to simulate Twitter…
▽ More
The rapid growth of social media has been witnessed during recent years as a result of the prevalence of the internet. This trend brings an increasing interest in simulating social media which can provide valuable insights to both academic researchers and businesses. In this paper, we present a step-by-step approach of using Hawkes process, a self-activating stochastic process, to simulate Twitter activities and demonstrate how this model can be utilized to evaluate the chance of extremely rare web crises. Another goal of this research is to introduce a new strategy that implements Hawkes process on graph structures. Overall, we intend to extend the current Hawkes process to a wider range of scenarios and, in particular, create a more realistic simulation of Twitter activities by incorporating the actual user status and following-follower interactions between users.
△ Less
Submitted 6 August, 2021;
originally announced August 2021.
-
LassoNet: A Neural Network with Feature Sparsity
Authors:
Ismael Lemhadri,
Feng Ruan,
Louis Abraham,
Robert Tibshirani
Abstract:
Much work has been done recently to make neural networks more interpretable, and one obvious approach is to arrange for the network to use only a subset of the available features. In linear models, Lasso (or $\ell_1$-regularized) regression assigns zero weights to the most irrelevant or redundant features, and is widely used in data science. However the Lasso only applies to linear models. Here we…
▽ More
Much work has been done recently to make neural networks more interpretable, and one obvious approach is to arrange for the network to use only a subset of the available features. In linear models, Lasso (or $\ell_1$-regularized) regression assigns zero weights to the most irrelevant or redundant features, and is widely used in data science. However the Lasso only applies to linear models. Here we introduce LassoNet, a neural network framework with global feature selection. Our approach enforces a hierarchy: specifically a feature can participate in a hidden unit only if its linear representative is active. Unlike other approaches to feature selection for neural nets, our method uses a modified objective function with constraints, and so integrates feature selection with the parameter learning directly. As a result, it delivers an entire regularization path of solutions with a range of feature sparsity. On systematic experiments, LassoNet significantly outperforms state-of-the-art methods for feature selection and regression. The LassoNet method uses projected proximal gradient descent, and generalizes directly to deep networks. It can be implemented by adding just a few lines of code to a standard neural network.
△ Less
Submitted 16 June, 2021; v1 submitted 29 July, 2019;
originally announced July 2019.