Matches in SemOpenAlex for { <https://semopenalex.org/work/W2955972601> ?p ?o ?g. }
- W2955972601 abstract "We investigate the complexity of logistic regression models which is defined by counting the number of indistinguishable distributions that the model can represent (Balasubramanian, 1997). We find that the complexity of logistic models with binary inputs does not only depend on the number of parameters but also on the distribution of inputs in a non-trivial way which standard treatments of complexity do not address. In particular, we observe that correlations among inputs induce effective dependencies among parameters thus constraining the model and, consequently, reducing its complexity. We derive simple relations for the upper and lower bounds of the complexity. Furthermore, we show analytically that, defining the model parameters on a finite support rather than the entire axis, decreases the complexity in a manner that critically depends on the size of the domain. Based on our findings, we propose a novel model selection criterion which takes into account the entropy of the input distribution. We test our proposal on the problem of selecting the input variables of a logistic regression model in a Bayesian Model Selection framework. In our numerical tests, we find that, while the reconstruction errors of standard model selection approaches (AIC, BIC, $ell_1$ regularization) strongly depend on the sparsity of the ground truth, the reconstruction error of our method is always close to the minimum in all conditions of sparsity, data size and strength of input correlations. Finally, we observe that, when considering categorical instead of binary inputs, in a simple and mathematically tractable case, the contribution of the alphabet size to the complexity is very small compared to that of parameter space dimension. We further explore the issue by analysing the dataset of the 13 keys to the White House which is a method for forecasting the outcomes of US presidential elections." @default.
- W2955972601 created "2019-07-12" @default.
- W2955972601 creator A5024030376 @default.
- W2955972601 creator A5045481961 @default.
- W2955972601 creator A5051471901 @default.
- W2955972601 date "2019-03-01" @default.
- W2955972601 modified "2023-10-01" @default.
- W2955972601 title "On the complexity of logistic regression models" @default.
- W2955972601 cites W1480376833 @default.
- W2955972601 cites W1531509813 @default.
- W2955972601 cites W1985507521 @default.
- W2955972601 cites W1994029324 @default.
- W2955972601 cites W1999751781 @default.
- W2955972601 cites W2006681603 @default.
- W2955972601 cites W2056099894 @default.
- W2955972601 cites W2058251410 @default.
- W2955972601 cites W2059151864 @default.
- W2955972601 cites W2068782468 @default.
- W2955972601 cites W2075691011 @default.
- W2955972601 cites W2086880647 @default.
- W2955972601 cites W2114766824 @default.
- W2955972601 cites W2119387367 @default.
- W2955972601 cites W2124641450 @default.
- W2955972601 cites W2125389748 @default.
- W2955972601 cites W2135046866 @default.
- W2955972601 cites W2139606141 @default.
- W2955972601 cites W2142331410 @default.
- W2955972601 cites W2142635246 @default.
- W2955972601 cites W2143814058 @default.
- W2955972601 cites W2147238273 @default.
- W2955972601 cites W2168175751 @default.
- W2955972601 cites W2169966170 @default.
- W2955972601 cites W2786580861 @default.
- W2955972601 cites W2801490189 @default.
- W2955972601 cites W2903950532 @default.
- W2955972601 cites W2905074708 @default.
- W2955972601 cites W2926521302 @default.
- W2955972601 cites W3098888484 @default.
- W2955972601 cites W3104720961 @default.
- W2955972601 cites W596170837 @default.
- W2955972601 hasPublicationYear "2019" @default.
- W2955972601 type Work @default.
- W2955972601 sameAs 2955972601 @default.
- W2955972601 citedByCount "0" @default.
- W2955972601 crossrefType "posted-content" @default.
- W2955972601 hasAuthorship W2955972601A5024030376 @default.
- W2955972601 hasAuthorship W2955972601A5045481961 @default.
- W2955972601 hasAuthorship W2955972601A5051471901 @default.
- W2955972601 hasConcept C105795698 @default.
- W2955972601 hasConcept C106301342 @default.
- W2955972601 hasConcept C11413529 @default.
- W2955972601 hasConcept C121332964 @default.
- W2955972601 hasConcept C151956035 @default.
- W2955972601 hasConcept C168136583 @default.
- W2955972601 hasConcept C28826006 @default.
- W2955972601 hasConcept C33923547 @default.
- W2955972601 hasConcept C41008148 @default.
- W2955972601 hasConcept C48372109 @default.
- W2955972601 hasConcept C5274069 @default.
- W2955972601 hasConcept C62520636 @default.
- W2955972601 hasConcept C93959086 @default.
- W2955972601 hasConcept C94375191 @default.
- W2955972601 hasConceptScore W2955972601C105795698 @default.
- W2955972601 hasConceptScore W2955972601C106301342 @default.
- W2955972601 hasConceptScore W2955972601C11413529 @default.
- W2955972601 hasConceptScore W2955972601C121332964 @default.
- W2955972601 hasConceptScore W2955972601C151956035 @default.
- W2955972601 hasConceptScore W2955972601C168136583 @default.
- W2955972601 hasConceptScore W2955972601C28826006 @default.
- W2955972601 hasConceptScore W2955972601C33923547 @default.
- W2955972601 hasConceptScore W2955972601C41008148 @default.
- W2955972601 hasConceptScore W2955972601C48372109 @default.
- W2955972601 hasConceptScore W2955972601C5274069 @default.
- W2955972601 hasConceptScore W2955972601C62520636 @default.
- W2955972601 hasConceptScore W2955972601C93959086 @default.
- W2955972601 hasConceptScore W2955972601C94375191 @default.
- W2955972601 hasLocation W29559726011 @default.
- W2955972601 hasOpenAccess W2955972601 @default.
- W2955972601 hasPrimaryLocation W29559726011 @default.
- W2955972601 hasRelatedWork W1543353791 @default.
- W2955972601 hasRelatedWork W1544709650 @default.
- W2955972601 hasRelatedWork W2038151181 @default.
- W2955972601 hasRelatedWork W2060044229 @default.
- W2955972601 hasRelatedWork W2112792378 @default.
- W2955972601 hasRelatedWork W2142920735 @default.
- W2955972601 hasRelatedWork W2187544847 @default.
- W2955972601 hasRelatedWork W2200392332 @default.
- W2955972601 hasRelatedWork W2781844902 @default.
- W2955972601 hasRelatedWork W2909819554 @default.
- W2955972601 hasRelatedWork W2951376723 @default.
- W2955972601 hasRelatedWork W2963949365 @default.
- W2955972601 hasRelatedWork W2986271303 @default.
- W2955972601 hasRelatedWork W3094514413 @default.
- W2955972601 hasRelatedWork W3098794096 @default.
- W2955972601 hasRelatedWork W3122854700 @default.
- W2955972601 hasRelatedWork W3195203488 @default.
- W2955972601 hasRelatedWork W3202500181 @default.
- W2955972601 hasRelatedWork W2161025652 @default.
- W2955972601 hasRelatedWork W3104799747 @default.
- W2955972601 isParatext "false" @default.