tests logistic regression: Topics by Science.gov

Sample records for tests logistic regression

A Bayesian goodness of fit test and semiparametric generalization of logistic regression with measurement data.

PubMed

Schörgendorfer, Angela; Branscum, Adam J; Hanson, Timothy E

2013-06-01

Logistic regression is a popular tool for risk analysis in medical and population health science. With continuous response data, it is common to create a dichotomous outcome for logistic regression analysis by specifying a threshold for positivity. Fitting a linear regression to the nondichotomized response variable assuming a logistic sampling model for the data has been empirically shown to yield more efficient estimates of odds ratios than ordinary logistic regression of the dichotomized endpoint. We illustrate that risk inference is not robust to departures from the parametric logistic distribution. Moreover, the model assumption of proportional odds is generally not satisfied when the condition of a logistic distribution for the data is violated, leading to biased inference from a parametric logistic analysis. We develop novel Bayesian semiparametric methodology for testing goodness of fit of parametric logistic regression with continuous measurement data. The testing procedures hold for any cutoff threshold and our approach simultaneously provides the ability to perform semiparametric risk estimation. Bayes factors are calculated using the Savage-Dickey ratio for testing the null hypothesis of logistic regression versus a semiparametric generalization. We propose a fully Bayesian and a computationally efficient empirical Bayesian approach to testing, and we present methods for semiparametric estimation of risks, relative risks, and odds ratios when parametric logistic regression fails. Theoretical results establish the consistency of the empirical Bayes test. Results from simulated data show that the proposed approach provides accurate inference irrespective of whether parametric assumptions hold or not. Evaluation of risk factors for obesity shows that different inferences are derived from an analysis of a real data set when deviations from a logistic distribution are permissible in a flexible semiparametric framework. © 2013, The International Biometric Society.
Power and Sample Size Calculations for Logistic Regression Tests for Differential Item Functioning

ERIC Educational Resources Information Center

Li, Zhushan

2014-01-01

Logistic regression is a popular method for detecting uniform and nonuniform differential item functioning (DIF) effects. Theoretical formulas for the power and sample size calculations are derived for likelihood ratio tests and Wald tests based on the asymptotic distribution of the maximum likelihood estimators for the logistic regression model.…
Bias in logistic regression due to imperfect diagnostic test results and practical correction approaches.

PubMed

Valle, Denis; Lima, Joanna M Tucker; Millar, Justin; Amratia, Punam; Haque, Ubydul

2015-11-04

Logistic regression is a statistical model widely used in cross-sectional and cohort studies to identify and quantify the effects of potential disease risk factors. However, the impact of imperfect tests on adjusted odds ratios (and thus on the identification of risk factors) is under-appreciated. The purpose of this article is to draw attention to the problem associated with modelling imperfect diagnostic tests, and propose simple Bayesian models to adequately address this issue. A systematic literature review was conducted to determine the proportion of malaria studies that appropriately accounted for false-negatives/false-positives in a logistic regression setting. Inference from the standard logistic regression was also compared with that from three proposed Bayesian models using simulations and malaria data from the western Brazilian Amazon. A systematic literature review suggests that malaria epidemiologists are largely unaware of the problem of using logistic regression to model imperfect diagnostic test results. Simulation results reveal that statistical inference can be substantially improved when using the proposed Bayesian models versus the standard logistic regression. Finally, analysis of original malaria data with one of the proposed Bayesian models reveals that microscopy sensitivity is strongly influenced by how long people have lived in the study region, and an important risk factor (i.e., participation in forest extractivism) is identified that would have been missed by standard logistic regression. Given the numerous diagnostic methods employed by malaria researchers and the ubiquitous use of logistic regression to model the results of these diagnostic tests, this paper provides critical guidelines to improve data analysis practice in the presence of misclassification error. Easy-to-use code that can be readily adapted to WinBUGS is provided, enabling straightforward implementation of the proposed Bayesian models.
CUSUM-Logistic Regression analysis for the rapid detection of errors in clinical laboratory test results.

PubMed

Sampson, Maureen L; Gounden, Verena; van Deventer, Hendrik E; Remaley, Alan T

2016-02-01

The main drawback of the periodic analysis of quality control (QC) material is that test performance is not monitored in time periods between QC analyses, potentially leading to the reporting of faulty test results. The objective of this study was to develop a patient based QC procedure for the more timely detection of test errors. Results from a Chem-14 panel measured on the Beckman LX20 analyzer were used to develop the model. Each test result was predicted from the other 13 members of the panel by multiple regression, which resulted in correlation coefficients between the predicted and measured result of >0.7 for 8 of the 14 tests. A logistic regression model, which utilized the measured test result, the predicted test result, the day of the week and time of day, was then developed for predicting test errors. The output of the logistic regression was tallied by a daily CUSUM approach and used to predict test errors, with a fixed specificity of 90%. The mean average run length (ARL) before error detection by CUSUM-Logistic Regression (CSLR) was 20 with a mean sensitivity of 97%, which was considerably shorter than the mean ARL of 53 (sensitivity 87.5%) for a simple prediction model that only used the measured result for error detection. A CUSUM-Logistic Regression analysis of patient laboratory data can be an effective approach for the rapid and sensitive detection of clinical laboratory errors. Published by Elsevier Inc.
A Method for Calculating the Probability of Successfully Completing a Rocket Propulsion Ground Test

NASA Technical Reports Server (NTRS)

Messer, Bradley

2007-01-01

Propulsion ground test facilities face the daily challenge of scheduling multiple customers into limited facility space and successfully completing their propulsion test projects. Over the last decade NASA s propulsion test facilities have performed hundreds of tests, collected thousands of seconds of test data, and exceeded the capabilities of numerous test facility and test article components. A logistic regression mathematical modeling technique has been developed to predict the probability of successfully completing a rocket propulsion test. A logistic regression model is a mathematical modeling approach that can be used to describe the relationship of several independent predictor variables X(sub 1), X(sub 2),.., X(sub k) to a binary or dichotomous dependent variable Y, where Y can only be one of two possible outcomes, in this case Success or Failure of accomplishing a full duration test. The use of logistic regression modeling is not new; however, modeling propulsion ground test facilities using logistic regression is both a new and unique application of the statistical technique. Results from this type of model provide project managers with insight and confidence into the effectiveness of rocket propulsion ground testing.
A Note on Three Statistical Tests in the Logistic Regression DIF Procedure

ERIC Educational Resources Information Center

Paek, Insu

2012-01-01

Although logistic regression became one of the well-known methods in detecting differential item functioning (DIF), its three statistical tests, the Wald, likelihood ratio (LR), and score tests, which are readily available under the maximum likelihood, do not seem to be consistently distinguished in DIF literature. This paper provides a clarifying…
Logistic LASSO regression for the diagnosis of breast cancer using clinical demographic data and the BI-RADS lexicon for ultrasonography.

PubMed

Kim, Sun Mi; Kim, Yongdai; Jeong, Kuhwan; Jeong, Heeyeong; Kim, Jiyoung

2018-01-01

The aim of this study was to compare the performance of image analysis for predicting breast cancer using two distinct regression models and to evaluate the usefulness of incorporating clinical and demographic data (CDD) into the image analysis in order to improve the diagnosis of breast cancer. This study included 139 solid masses from 139 patients who underwent a ultrasonography-guided core biopsy and had available CDD between June 2009 and April 2010. Three breast radiologists retrospectively reviewed 139 breast masses and described each lesion using the Breast Imaging Reporting and Data System (BI-RADS) lexicon. We applied and compared two regression methods-stepwise logistic (SL) regression and logistic least absolute shrinkage and selection operator (LASSO) regression-in which the BI-RADS descriptors and CDD were used as covariates. We investigated the performances of these regression methods and the agreement of radiologists in terms of test misclassification error and the area under the curve (AUC) of the tests. Logistic LASSO regression was superior (P<0.05) to SL regression, regardless of whether CDD was included in the covariates, in terms of test misclassification errors (0.234 vs. 0.253, without CDD; 0.196 vs. 0.258, with CDD) and AUC (0.785 vs. 0.759, without CDD; 0.873 vs. 0.735, with CDD). However, it was inferior (P<0.05) to the agreement of three radiologists in terms of test misclassification errors (0.234 vs. 0.168, without CDD; 0.196 vs. 0.088, with CDD) and the AUC without CDD (0.785 vs. 0.844, P<0.001), but was comparable to the AUC with CDD (0.873 vs. 0.880, P=0.141). Logistic LASSO regression based on BI-RADS descriptors and CDD showed better performance than SL in predicting the presence of breast cancer. The use of CDD as a supplement to the BI-RADS descriptors significantly improved the prediction of breast cancer using logistic LASSO regression.
Regression approaches in the test-negative study design for assessment of influenza vaccine effectiveness.

PubMed

Bond, H S; Sullivan, S G; Cowling, B J

2016-06-01

Influenza vaccination is the most practical means available for preventing influenza virus infection and is widely used in many countries. Because vaccine components and circulating strains frequently change, it is important to continually monitor vaccine effectiveness (VE). The test-negative design is frequently used to estimate VE. In this design, patients meeting the same clinical case definition are recruited and tested for influenza; those who test positive are the cases and those who test negative form the comparison group. When determining VE in these studies, the typical approach has been to use logistic regression, adjusting for potential confounders. Because vaccine coverage and influenza incidence change throughout the season, time is included among these confounders. While most studies use unconditional logistic regression, adjusting for time, an alternative approach is to use conditional logistic regression, matching on time. Here, we used simulation data to examine the potential for both regression approaches to permit accurate and robust estimates of VE. In situations where vaccine coverage changed during the influenza season, the conditional model and unconditional models adjusting for categorical week and using a spline function for week provided more accurate estimates. We illustrated the two approaches on data from a test-negative study of influenza VE against hospitalization in children in Hong Kong which resulted in the conditional logistic regression model providing the best fit to the data.
Comparison of IRT Likelihood Ratio Test and Logistic Regression DIF Detection Procedures

ERIC Educational Resources Information Center

Atar, Burcu; Kamata, Akihito

2011-01-01

The Type I error rates and the power of IRT likelihood ratio test and cumulative logit ordinal logistic regression procedures in detecting differential item functioning (DIF) for polytomously scored items were investigated in this Monte Carlo simulation study. For this purpose, 54 simulation conditions (combinations of 3 sample sizes, 2 sample…
Comparison of multinomial logistic regression and logistic regression: which is more efficient in allocating land use?

NASA Astrophysics Data System (ADS)

Lin, Yingzhi; Deng, Xiangzheng; Li, Xing; Ma, Enjun

2014-12-01

Spatially explicit simulation of land use change is the basis for estimating the effects of land use and cover change on energy fluxes, ecology and the environment. At the pixel level, logistic regression is one of the most common approaches used in spatially explicit land use allocation models to determine the relationship between land use and its causal factors in driving land use change, and thereby to evaluate land use suitability. However, these models have a drawback in that they do not determine/allocate land use based on the direct relationship between land use change and its driving factors. Consequently, a multinomial logistic regression method was introduced to address this flaw, and thereby, judge the suitability of a type of land use in any given pixel in a case study area of the Jiangxi Province, China. A comparison of the two regression methods indicated that the proportion of correctly allocated pixels using multinomial logistic regression was 92.98%, which was 8.47% higher than that obtained using logistic regression. Paired t-test results also showed that pixels were more clearly distinguished by multinomial logistic regression than by logistic regression. In conclusion, multinomial logistic regression is a more efficient and accurate method for the spatial allocation of land use changes. The application of this method in future land use change studies may improve the accuracy of predicting the effects of land use and cover change on energy fluxes, ecology, and environment.
Preserving Institutional Privacy in Distributed binary Logistic Regression.

PubMed

Wu, Yuan; Jiang, Xiaoqian; Ohno-Machado, Lucila

2012-01-01

Privacy is becoming a major concern when sharing biomedical data across institutions. Although methods for protecting privacy of individual patients have been proposed, it is not clear how to protect the institutional privacy, which is many times a critical concern of data custodians. Built upon our previous work, Grid Binary LOgistic REgression (GLORE)1, we developed an Institutional Privacy-preserving Distributed binary Logistic Regression model (IPDLR) that considers both individual and institutional privacy for building a logistic regression model in a distributed manner. We tested our method using both simulated and clinical data, showing how it is possible to protect the privacy of individuals and of institutions using a distributed strategy.
A Solution to Separation and Multicollinearity in Multiple Logistic Regression

PubMed Central

Shen, Jianzhao; Gao, Sujuan

2010-01-01

In dementia screening tests, item selection for shortening an existing screening test can be achieved using multiple logistic regression. However, maximum likelihood estimates for such logistic regression models often experience serious bias or even non-existence because of separation and multicollinearity problems resulting from a large number of highly correlated items. Firth (1993, Biometrika, 80(1), 27–38) proposed a penalized likelihood estimator for generalized linear models and it was shown to reduce bias and the non-existence problems. The ridge regression has been used in logistic regression to stabilize the estimates in cases of multicollinearity. However, neither solves the problems for each other. In this paper, we propose a double penalized maximum likelihood estimator combining Firth’s penalized likelihood equation with a ridge parameter. We present a simulation study evaluating the empirical performance of the double penalized likelihood estimator in small to moderate sample sizes. We demonstrate the proposed approach using a current screening data from a community-based dementia study. PMID:20376286
A Solution to Separation and Multicollinearity in Multiple Logistic Regression.

PubMed

Shen, Jianzhao; Gao, Sujuan

2008-10-01

In dementia screening tests, item selection for shortening an existing screening test can be achieved using multiple logistic regression. However, maximum likelihood estimates for such logistic regression models often experience serious bias or even non-existence because of separation and multicollinearity problems resulting from a large number of highly correlated items. Firth (1993, Biometrika, 80(1), 27-38) proposed a penalized likelihood estimator for generalized linear models and it was shown to reduce bias and the non-existence problems. The ridge regression has been used in logistic regression to stabilize the estimates in cases of multicollinearity. However, neither solves the problems for each other. In this paper, we propose a double penalized maximum likelihood estimator combining Firth's penalized likelihood equation with a ridge parameter. We present a simulation study evaluating the empirical performance of the double penalized likelihood estimator in small to moderate sample sizes. We demonstrate the proposed approach using a current screening data from a community-based dementia study.
Three methods to construct predictive models using logistic regression and likelihood ratios to facilitate adjustment for pretest probability give similar results.

PubMed

Chan, Siew Foong; Deeks, Jonathan J; Macaskill, Petra; Irwig, Les

2008-01-01

To compare three predictive models based on logistic regression to estimate adjusted likelihood ratios allowing for interdependency between diagnostic variables (tests). This study was a review of the theoretical basis, assumptions, and limitations of published models; and a statistical extension of methods and application to a case study of the diagnosis of obstructive airways disease based on history and clinical examination. Albert's method includes an offset term to estimate an adjusted likelihood ratio for combinations of tests. Spiegelhalter and Knill-Jones method uses the unadjusted likelihood ratio for each test as a predictor and computes shrinkage factors to allow for interdependence. Knottnerus' method differs from the other methods because it requires sequencing of tests, which limits its application to situations where there are few tests and substantial data. Although parameter estimates differed between the models, predicted "posttest" probabilities were generally similar. Construction of predictive models using logistic regression is preferred to the independence Bayes' approach when it is important to adjust for dependency of tests errors. Methods to estimate adjusted likelihood ratios from predictive models should be considered in preference to a standard logistic regression model to facilitate ease of interpretation and application. Albert's method provides the most straightforward approach.
Length bias correction in gene ontology enrichment analysis using logistic regression.

PubMed

Mi, Gu; Di, Yanming; Emerson, Sarah; Cumbie, Jason S; Chang, Jeff H

2012-01-01

When assessing differential gene expression from RNA sequencing data, commonly used statistical tests tend to have greater power to detect differential expression of genes encoding longer transcripts. This phenomenon, called "length bias", will influence subsequent analyses such as Gene Ontology enrichment analysis. In the presence of length bias, Gene Ontology categories that include longer genes are more likely to be identified as enriched. These categories, however, are not necessarily biologically more relevant. We show that one can effectively adjust for length bias in Gene Ontology analysis by including transcript length as a covariate in a logistic regression model. The logistic regression model makes the statistical issue underlying length bias more transparent: transcript length becomes a confounding factor when it correlates with both the Gene Ontology membership and the significance of the differential expression test. The inclusion of the transcript length as a covariate allows one to investigate the direct correlation between the Gene Ontology membership and the significance of testing differential expression, conditional on the transcript length. We present both real and simulated data examples to show that the logistic regression approach is simple, effective, and flexible.
Controlling Type I Error Rates in Assessing DIF for Logistic Regression Method Combined with SIBTEST Regression Correction Procedure and DIF-Free-Then-DIF Strategy

ERIC Educational Resources Information Center

Shih, Ching-Lin; Liu, Tien-Hsiang; Wang, Wen-Chung

2014-01-01

The simultaneous item bias test (SIBTEST) method regression procedure and the differential item functioning (DIF)-free-then-DIF strategy are applied to the logistic regression (LR) method simultaneously in this study. These procedures are used to adjust the effects of matching true score on observed score and to better control the Type I error…
Predicting Student Success on the Texas Chemistry STAAR Test: A Logistic Regression Analysis

ERIC Educational Resources Information Center

Johnson, William L.; Johnson, Annabel M.; Johnson, Jared

2012-01-01

Background: The context is the new Texas STAAR end-of-course testing program. Purpose: The authors developed a logistic regression model to predict who would pass-or-fail the new Texas chemistry STAAR end-of-course exam. Setting: Robert E. Lee High School (5A) with an enrollment of 2700 students, Tyler, Texas. Date of the study was the 2011-2012…
Modification of the Mantel-Haenszel and Logistic Regression DIF Procedures to Incorporate the SIBTEST Regression Correction

ERIC Educational Resources Information Center

DeMars, Christine E.

2009-01-01

The Mantel-Haenszel (MH) and logistic regression (LR) differential item functioning (DIF) procedures have inflated Type I error rates when there are large mean group differences, short tests, and large sample sizes.When there are large group differences in mean score, groups matched on the observed number-correct score differ on true score,…
Multinomial logistic regression modelling of obesity and overweight among primary school students in a rural area of Negeri Sembilan

DOE Office of Scientific and Technical Information (OSTI.GOV)

Ghazali, Amirul Syafiq Mohd; Ali, Zalila; Noor, Norlida Mohd

Multinomial logistic regression is widely used to model the outcomes of a polytomous response variable, a categorical dependent variable with more than two categories. The model assumes that the conditional mean of the dependent categorical variables is the logistic function of an affine combination of predictor variables. Its procedure gives a number of logistic regression models that make specific comparisons of the response categories. When there are q categories of the response variable, the model consists of q-1 logit equations which are fitted simultaneously. The model is validated by variable selection procedures, tests of regression coefficients, a significant test ofmore » the overall model, goodness-of-fit measures, and validation of predicted probabilities using odds ratio. This study used the multinomial logistic regression model to investigate obesity and overweight among primary school students in a rural area on the basis of their demographic profiles, lifestyles and on the diet and food intake. The results indicated that obesity and overweight of students are related to gender, religion, sleep duration, time spent on electronic games, breakfast intake in a week, with whom meals are taken, protein intake, and also, the interaction between breakfast intake in a week with sleep duration, and the interaction between gender and protein intake.« less
Multinomial logistic regression modelling of obesity and overweight among primary school students in a rural area of Negeri Sembilan

NASA Astrophysics Data System (ADS)

Ghazali, Amirul Syafiq Mohd; Ali, Zalila; Noor, Norlida Mohd; Baharum, Adam

2015-10-01

Multinomial logistic regression is widely used to model the outcomes of a polytomous response variable, a categorical dependent variable with more than two categories. The model assumes that the conditional mean of the dependent categorical variables is the logistic function of an affine combination of predictor variables. Its procedure gives a number of logistic regression models that make specific comparisons of the response categories. When there are q categories of the response variable, the model consists of q-1 logit equations which are fitted simultaneously. The model is validated by variable selection procedures, tests of regression coefficients, a significant test of the overall model, goodness-of-fit measures, and validation of predicted probabilities using odds ratio. This study used the multinomial logistic regression model to investigate obesity and overweight among primary school students in a rural area on the basis of their demographic profiles, lifestyles and on the diet and food intake. The results indicated that obesity and overweight of students are related to gender, religion, sleep duration, time spent on electronic games, breakfast intake in a week, with whom meals are taken, protein intake, and also, the interaction between breakfast intake in a week with sleep duration, and the interaction between gender and protein intake.

Model building strategy for logistic regression: purposeful selection.

PubMed

Zhang, Zhongheng

2016-03-01

Logistic regression is one of the most commonly used models to account for confounders in medical literature. The article introduces how to perform purposeful selection model building strategy with R. I stress on the use of likelihood ratio test to see whether deleting a variable will have significant impact on model fit. A deleted variable should also be checked for whether it is an important adjustment of remaining covariates. Interaction should be checked to disentangle complex relationship between covariates and their synergistic effect on response variable. Model should be checked for the goodness-of-fit (GOF). In other words, how the fitted model reflects the real data. Hosmer-Lemeshow GOF test is the most widely used for logistic regression model.
A novel hybrid method of beta-turn identification in protein using binary logistic regression and neural network

PubMed Central

Asghari, Mehdi Poursheikhali; Hayatshahi, Sayyed Hamed Sadat; Abdolmaleki, Parviz

2012-01-01

From both the structural and functional points of view, β-turns play important biological roles in proteins. In the present study, a novel two-stage hybrid procedure has been developed to identify β-turns in proteins. Binary logistic regression was initially used for the first time to select significant sequence parameters in identification of β-turns due to a re-substitution test procedure. Sequence parameters were consisted of 80 amino acid positional occurrences and 20 amino acid percentages in sequence. Among these parameters, the most significant ones which were selected by binary logistic regression model, were percentages of Gly, Ser and the occurrence of Asn in position i+2, respectively, in sequence. These significant parameters have the highest effect on the constitution of a β-turn sequence. A neural network model was then constructed and fed by the parameters selected by binary logistic regression to build a hybrid predictor. The networks have been trained and tested on a non-homologous dataset of 565 protein chains. With applying a nine fold cross-validation test on the dataset, the network reached an overall accuracy (Qtotal) of 74, which is comparable with results of the other β-turn prediction methods. In conclusion, this study proves that the parameter selection ability of binary logistic regression together with the prediction capability of neural networks lead to the development of more precise models for identifying β-turns in proteins. PMID:27418910
A novel hybrid method of beta-turn identification in protein using binary logistic regression and neural network.

PubMed

Asghari, Mehdi Poursheikhali; Hayatshahi, Sayyed Hamed Sadat; Abdolmaleki, Parviz

2012-01-01

From both the structural and functional points of view, β-turns play important biological roles in proteins. In the present study, a novel two-stage hybrid procedure has been developed to identify β-turns in proteins. Binary logistic regression was initially used for the first time to select significant sequence parameters in identification of β-turns due to a re-substitution test procedure. Sequence parameters were consisted of 80 amino acid positional occurrences and 20 amino acid percentages in sequence. Among these parameters, the most significant ones which were selected by binary logistic regression model, were percentages of Gly, Ser and the occurrence of Asn in position i+2, respectively, in sequence. These significant parameters have the highest effect on the constitution of a β-turn sequence. A neural network model was then constructed and fed by the parameters selected by binary logistic regression to build a hybrid predictor. The networks have been trained and tested on a non-homologous dataset of 565 protein chains. With applying a nine fold cross-validation test on the dataset, the network reached an overall accuracy (Qtotal) of 74, which is comparable with results of the other β-turn prediction methods. In conclusion, this study proves that the parameter selection ability of binary logistic regression together with the prediction capability of neural networks lead to the development of more precise models for identifying β-turns in proteins.
Differential item functioning analysis with ordinal logistic regression techniques. DIFdetect and difwithpar.

PubMed

Crane, Paul K; Gibbons, Laura E; Jolley, Lance; van Belle, Gerald

2006-11-01

We present an ordinal logistic regression model for identification of items with differential item functioning (DIF) and apply this model to a Mini-Mental State Examination (MMSE) dataset. We employ item response theory ability estimation in our models. Three nested ordinal logistic regression models are applied to each item. Model testing begins with examination of the statistical significance of the interaction term between ability and the group indicator, consistent with nonuniform DIF. Then we turn our attention to the coefficient of the ability term in models with and without the group term. If including the group term has a marked effect on that coefficient, we declare that it has uniform DIF. We examined DIF related to language of test administration in addition to self-reported race, Hispanic ethnicity, age, years of education, and sex. We used PARSCALE for IRT analyses and STATA for ordinal logistic regression approaches. We used an iterative technique for adjusting IRT ability estimates on the basis of DIF findings. Five items were found to have DIF related to language. These same items also had DIF related to other covariates. The ordinal logistic regression approach to DIF detection, when combined with IRT ability estimates, provides a reasonable alternative for DIF detection. There appear to be several items with significant DIF related to language of test administration in the MMSE. More attention needs to be paid to the specific criteria used to determine whether an item has DIF, not just the technique used to identify DIF.
Testing Gene-Gene Interactions in the Case-Parents Design

PubMed Central

Yu, Zhaoxia

2011-01-01

The case-parents design has been widely used to detect genetic associations as it can prevent spurious association that could occur in population-based designs. When examining the effect of an individual genetic locus on a disease, logistic regressions developed by conditioning on parental genotypes provide complete protection from spurious association caused by population stratification. However, when testing gene-gene interactions, it is unknown whether conditional logistic regressions are still robust. Here we evaluate the robustness and efficiency of several gene-gene interaction tests that are derived from conditional logistic regressions. We found that in the presence of SNP genotype correlation due to population stratification or linkage disequilibrium, tests with incorrectly specified main-genetic-effect models can lead to inflated type I error rates. We also found that a test with fully flexible main genetic effects always maintains correct test size and its robustness can be achieved with negligible sacrifice of its power. When testing gene-gene interactions is the focus, the test allowing fully flexible main effects is recommended to be used. PMID:21778736
Satellite rainfall retrieval by logistic regression

NASA Technical Reports Server (NTRS)

Chiu, Long S.

1986-01-01

The potential use of logistic regression in rainfall estimation from satellite measurements is investigated. Satellite measurements provide covariate information in terms of radiances from different remote sensors.The logistic regression technique can effectively accommodate many covariates and test their significance in the estimation. The outcome from the logistical model is the probability that the rainrate of a satellite pixel is above a certain threshold. By varying the thresholds, a rainrate histogram can be obtained, from which the mean and the variant can be estimated. A logistical model is developed and applied to rainfall data collected during GATE, using as covariates the fractional rain area and a radiance measurement which is deduced from a microwave temperature-rainrate relation. It is demonstrated that the fractional rain area is an important covariate in the model, consistent with the use of the so-called Area Time Integral in estimating total rain volume in other studies. To calibrate the logistical model, simulated rain fields generated by rainfield models with prescribed parameters are needed. A stringent test of the logistical model is its ability to recover the prescribed parameters of simulated rain fields. A rain field simulation model which preserves the fractional rain area and lognormality of rainrates as found in GATE is developed. A stochastic regression model of branching and immigration whose solutions are lognormally distributed in some asymptotic limits has also been developed.
London Measure of Unplanned Pregnancy: guidance for its use as an outcome measure

PubMed Central

Hall, Jennifer A; Barrett, Geraldine; Copas, Andrew; Stephenson, Judith

2017-01-01

Background The London Measure of Unplanned Pregnancy (LMUP) is a psychometrically validated measure of the degree of intention of a current or recent pregnancy. The LMUP is increasingly being used worldwide, and can be used to evaluate family planning or preconception care programs. However, beyond recommending the use of the full LMUP scale, there is no published guidance on how to use the LMUP as an outcome measure. Ordinal logistic regression has been recommended informally, but studies published to date have all used binary logistic regression and dichotomized the scale at different cut points. There is thus a need for evidence-based guidance to provide a standardized methodology for multivariate analysis and to enable comparison of results. This paper makes recommendations for the regression method for analysis of the LMUP as an outcome measure. Materials and methods Data collected from 4,244 pregnant women in Malawi were used to compare five regression methods: linear, logistic with two cut points, and ordinal logistic with either the full or grouped LMUP score. The recommendations were then tested on the original UK LMUP data. Results There were small but no important differences in the findings across the regression models. Logistic regression resulted in the largest loss of information, and assumptions were violated for the linear and ordinal logistic regression. Consequently, robust standard errors were used for linear regression and a partial proportional odds ordinal logistic regression model attempted. The latter could only be fitted for grouped LMUP score. Conclusion We recommend the linear regression model with robust standard errors to make full use of the LMUP score when analyzed as an outcome measure. Ordinal logistic regression could be considered, but a partial proportional odds model with grouped LMUP score may be required. Logistic regression is the least-favored option, due to the loss of information. For logistic regression, the cut point for un/planned pregnancy should be between nine and ten. These recommendations will standardize the analysis of LMUP data and enhance comparability of results across studies. PMID:28435343
4D-Fingerprint Categorical QSAR Models for Skin Sensitization Based on Classification Local Lymph Node Assay Measures

PubMed Central

Li, Yi; Tseng, Yufeng J.; Pan, Dahua; Liu, Jianzhong; Kern, Petra S.; Gerberick, G. Frank; Hopfinger, Anton J.

2008-01-01

Currently, the only validated methods to identify skin sensitization effects are in vivo models, such as the Local Lymph Node Assay (LLNA) and guinea pig studies. There is a tremendous need, in particular due to novel legislation, to develop animal alternatives, eg. Quantitative Structure-Activity Relationship (QSAR) models. Here, QSAR models for skin sensitization using LLNA data have been constructed. The descriptors used to generate these models are derived from the 4D-molecular similarity paradigm and are referred to as universal 4D-fingerprints. A training set of 132 structurally diverse compounds and a test set of 15 structurally diverse compounds were used in this study. The statistical methodologies used to build the models are logistic regression (LR), and partial least square coupled logistic regression (PLS-LR), which prove to be effective tools for studying skin sensitization measures expressed in the two categorical terms of sensitizer and non-sensitizer. QSAR models with low values of the Hosmer-Lemeshow goodness-of-fit statistic, χHL2, are significant and predictive. For the training set, the cross-validated prediction accuracy of the logistic regression models ranges from 77.3% to 78.0%, while that of PLS-logistic regression models ranges from 87.1% to 89.4%. For the test set, the prediction accuracy of logistic regression models ranges from 80.0%-86.7%, while that of PLS-logistic regression models ranges from 73.3%-80.0%. The QSAR models are made up of 4D-fingerprints related to aromatic atoms, hydrogen bond acceptors and negatively partially charged atoms. PMID:17226934
Strategies for Testing Statistical and Practical Significance in Detecting DIF with Logistic Regression Models

ERIC Educational Resources Information Center

Fidalgo, Angel M.; Alavi, Seyed Mohammad; Amirian, Seyed Mohammad Reza

2014-01-01

This study examines three controversial aspects in differential item functioning (DIF) detection by logistic regression (LR) models: first, the relative effectiveness of different analytical strategies for detecting DIF; second, the suitability of the Wald statistic for determining the statistical significance of the parameters of interest; and…
Modeling Polytomous Item Responses Using Simultaneously Estimated Multinomial Logistic Regression Models

ERIC Educational Resources Information Center

Anderson, Carolyn J.; Verkuilen, Jay; Peyton, Buddy L.

2010-01-01

Survey items with multiple response categories and multiple-choice test questions are ubiquitous in psychological and educational research. We illustrate the use of log-multiplicative association (LMA) models that are extensions of the well-known multinomial logistic regression model for multiple dependent outcome variables to reanalyze a set of…
Label-noise resistant logistic regression for functional data classification with an application to Alzheimer's disease study.

PubMed

Lee, Seokho; Shin, Hyejin; Lee, Sang Han

2016-12-01

Alzheimer's disease (AD) is usually diagnosed by clinicians through cognitive and functional performance test with a potential risk of misdiagnosis. Since the progression of AD is known to cause structural changes in the corpus callosum (CC), the CC thickness can be used as a functional covariate in AD classification problem for a diagnosis. However, misclassified class labels negatively impact the classification performance. Motivated by AD-CC association studies, we propose a logistic regression for functional data classification that is robust to misdiagnosis or label noise. Specifically, our logistic regression model is constructed by adopting individual intercepts to functional logistic regression model. This approach enables to indicate which observations are possibly mislabeled and also lead to a robust and efficient classifier. An effective algorithm using MM algorithm provides simple closed-form update formulas. We test our method using synthetic datasets to demonstrate its superiority over an existing method, and apply it to differentiating patients with AD from healthy normals based on CC from MRI. © 2016, The International Biometric Society.
Remote sensing and GIS-based landslide hazard analysis and cross-validation using multivariate logistic regression model on three test areas in Malaysia

NASA Astrophysics Data System (ADS)

Pradhan, Biswajeet

2010-05-01

This paper presents the results of the cross-validation of a multivariate logistic regression model using remote sensing data and GIS for landslide hazard analysis on the Penang, Cameron, and Selangor areas in Malaysia. Landslide locations in the study areas were identified by interpreting aerial photographs and satellite images, supported by field surveys. SPOT 5 and Landsat TM satellite imagery were used to map landcover and vegetation index, respectively. Maps of topography, soil type, lineaments and land cover were constructed from the spatial datasets. Ten factors which influence landslide occurrence, i.e., slope, aspect, curvature, distance from drainage, lithology, distance from lineaments, soil type, landcover, rainfall precipitation, and normalized difference vegetation index (ndvi), were extracted from the spatial database and the logistic regression coefficient of each factor was computed. Then the landslide hazard was analysed using the multivariate logistic regression coefficients derived not only from the data for the respective area but also using the logistic regression coefficients calculated from each of the other two areas (nine hazard maps in all) as a cross-validation of the model. For verification of the model, the results of the analyses were then compared with the field-verified landslide locations. Among the three cases of the application of logistic regression coefficient in the same study area, the case of Selangor based on the Selangor logistic regression coefficients showed the highest accuracy (94%), where as Penang based on the Penang coefficients showed the lowest accuracy (86%). Similarly, among the six cases from the cross application of logistic regression coefficient in other two areas, the case of Selangor based on logistic coefficient of Cameron showed highest (90%) prediction accuracy where as the case of Penang based on the Selangor logistic regression coefficients showed the lowest accuracy (79%). Qualitatively, the cross application model yields reasonable results which can be used for preliminary landslide hazard mapping.
Discrete post-processing of total cloud cover ensemble forecasts

NASA Astrophysics Data System (ADS)

Hemri, Stephan; Haiden, Thomas; Pappenberger, Florian

2017-04-01

This contribution presents an approach to post-process ensemble forecasts for the discrete and bounded weather variable of total cloud cover. Two methods for discrete statistical post-processing of ensemble predictions are tested. The first approach is based on multinomial logistic regression, the second involves a proportional odds logistic regression model. Applying them to total cloud cover raw ensemble forecasts from the European Centre for Medium-Range Weather Forecasts improves forecast skill significantly. Based on station-wise post-processing of raw ensemble total cloud cover forecasts for a global set of 3330 stations over the period from 2007 to early 2014, the more parsimonious proportional odds logistic regression model proved to slightly outperform the multinomial logistic regression model. Reference Hemri, S., Haiden, T., & Pappenberger, F. (2016). Discrete post-processing of total cloud cover ensemble forecasts. Monthly Weather Review 144, 2565-2577.
Analysis of a database to predict the result of allergy testing in vivo in patients with chronic nasal symptoms.

PubMed

Lacagnina, Valerio; Leto-Barone, Maria S; La Piana, Simona; Seidita, Aurelio; Pingitore, Giuseppe; Di Lorenzo, Gabriele

2014-01-01

This article uses the logistic regression model for diagnostic decision making in patients with chronic nasal symptoms. We studied the ability of the logistic regression model, obtained by the evaluation of a database, to detect patients with positive allergy skin-prick test (SPT) and patients with negative SPT. The model developed was validated using the data set obtained from another medical institution. The analysis was performed using a database obtained from a questionnaire administered to the patients with nasal symptoms containing personal data, clinical data, and results of allergy testing (SPT). All variables found to be significantly different between patients with positive and negative SPT (p < 0.05) were selected for the logistic regression models and were analyzed with backward stepwise logistic regression, evaluated with area under the curve of the receiver operating characteristic curve. A second set of patients from another institution was used to prove the model. The accuracy of the model in identifying, over the second set, both patients whose SPT will be positive and negative was high. The model detected 96% of patients with nasal symptoms and positive SPT and classified 94% of those with negative SPT. This study is preliminary to the creation of a software that could help the primary care doctors in a diagnostic decision making process (need of allergy testing) in patients complaining of chronic nasal symptoms.
Accuracy of Bayes and Logistic Regression Subscale Probabilities for Educational and Certification Tests

ERIC Educational Resources Information Center

Rudner, Lawrence

2016-01-01

In the machine learning literature, it is commonly accepted as fact that as calibration sample sizes increase, Naïve Bayes classifiers initially outperform Logistic Regression classifiers in terms of classification accuracy. Applied to subtests from an on-line final examination and from a highly regarded certification examination, this study shows…
School Exits in the Milwaukee Parental Choice Program: Evidence of a Marketplace?

ERIC Educational Resources Information Center

Ford, Michael

2011-01-01

This article examines whether the large number of school exits from the Milwaukee school voucher program is evidence of a marketplace. Two logistic regression and multinomial logistic regression models tested the relation between the inability to draw large numbers of voucher students and the ability for a private school to remain viable. Data on…
A Method for Calculating the Probability of Successfully Completing a Rocket Propulsion Ground Test

NASA Technical Reports Server (NTRS)

Messer, Bradley P.

2004-01-01

Propulsion ground test facilities face the daily challenges of scheduling multiple customers into limited facility space and successfully completing their propulsion test projects. Due to budgetary and schedule constraints, NASA and industry customers are pushing to test more components, for less money, in a shorter period of time. As these new rocket engine component test programs are undertaken, the lack of technology maturity in the test articles, combined with pushing the test facilities capabilities to their limits, tends to lead to an increase in facility breakdowns and unsuccessful tests. Over the last five years Stennis Space Center's propulsion test facilities have performed hundreds of tests, collected thousands of seconds of test data, and broken numerous test facility and test article parts. While various initiatives have been implemented to provide better propulsion test techniques and improve the quality, reliability, and maintainability of goods and parts used in the propulsion test facilities, unexpected failures during testing still occur quite regularly due to the harsh environment in which the propulsion test facilities operate. Previous attempts at modeling the lifecycle of a propulsion component test project have met with little success. Each of the attempts suffered form incomplete or inconsistent data on which to base the models. By focusing on the actual test phase of the tests project rather than the formulation, design or construction phases of the test project, the quality and quantity of available data increases dramatically. A logistic regression model has been developed form the data collected over the last five years, allowing the probability of successfully completing a rocket propulsion component test to be calculated. A logistic regression model is a mathematical modeling approach that can be used to describe the relationship of several independent predictor variables X(sub 1), X(sub 2),..,X(sub k) to a binary or dichotomous dependent variable Y, where Y can only be one of two possible outcomes, in this case Success or Failure. Logistic regression has primarily been used in the fields of epidemiology and biomedical research, but lends itself to many other applications. As indicated the use of logistic regression is not new, however, modeling propulsion ground test facilities using logistic regression is both a new and unique application of the statistical technique. Results from the models provide project managers with insight and confidence into the affectivity of rocket engine component ground test projects. The initial success in modeling rocket propulsion ground test projects clears the way for more complex models to be developed in this area.
New robust statistical procedures for the polytomous logistic regression models.

PubMed

Castilla, Elena; Ghosh, Abhik; Martin, Nirian; Pardo, Leandro

2018-05-17

This article derives a new family of estimators, namely the minimum density power divergence estimators, as a robust generalization of the maximum likelihood estimator for the polytomous logistic regression model. Based on these estimators, a family of Wald-type test statistics for linear hypotheses is introduced. Robustness properties of both the proposed estimators and the test statistics are theoretically studied through the classical influence function analysis. Appropriate real life examples are presented to justify the requirement of suitable robust statistical procedures in place of the likelihood based inference for the polytomous logistic regression model. The validity of the theoretical results established in the article are further confirmed empirically through suitable simulation studies. Finally, an approach for the data-driven selection of the robustness tuning parameter is proposed with empirical justifications. © 2018, The International Biometric Society.
A Generalized Logistic Regression Procedure to Detect Differential Item Functioning among Multiple Groups

ERIC Educational Resources Information Center

Magis, David; Raiche, Gilles; Beland, Sebastien; Gerard, Paul

2011-01-01

We present an extension of the logistic regression procedure to identify dichotomous differential item functioning (DIF) in the presence of more than two groups of respondents. Starting from the usual framework of a single focal group, we propose a general approach to estimate the item response functions in each group and to test for the presence…
A general equation to obtain multiple cut-off scores on a test from multinomial logistic regression.

PubMed

Bersabé, Rosa; Rivas, Teresa

2010-05-01

The authors derive a general equation to compute multiple cut-offs on a total test score in order to classify individuals into more than two ordinal categories. The equation is derived from the multinomial logistic regression (MLR) model, which is an extension of the binary logistic regression (BLR) model to accommodate polytomous outcome variables. From this analytical procedure, cut-off scores are established at the test score (the predictor variable) at which an individual is as likely to be in category j as in category j+1 of an ordinal outcome variable. The application of the complete procedure is illustrated by an example with data from an actual study on eating disorders. In this example, two cut-off scores on the Eating Attitudes Test (EAT-26) scores are obtained in order to classify individuals into three ordinal categories: asymptomatic, symptomatic and eating disorder. Diagnoses were made from the responses to a self-report (Q-EDD) that operationalises DSM-IV criteria for eating disorders. Alternatives to the MLR model to set multiple cut-off scores are discussed.

Comparison of Logistic Regression and Artificial Neural Network in Low Back Pain Prediction: Second National Health Survey

PubMed Central

Parsaeian, M; Mohammad, K; Mahmoudi, M; Zeraati, H

2012-01-01

Background: The purpose of this investigation was to compare empirically predictive ability of an artificial neural network with a logistic regression in prediction of low back pain. Methods: Data from the second national health survey were considered in this investigation. This data includes the information of low back pain and its associated risk factors among Iranian people aged 15 years and older. Artificial neural network and logistic regression models were developed using a set of 17294 data and they were validated in a test set of 17295 data. Hosmer and Lemeshow recommendation for model selection was used in fitting the logistic regression. A three-layer perceptron with 9 inputs, 3 hidden and 1 output neurons was employed. The efficiency of two models was compared by receiver operating characteristic analysis, root mean square and -2 Loglikelihood criteria. Results: The area under the ROC curve (SE), root mean square and -2Loglikelihood of the logistic regression was 0.752 (0.004), 0.3832 and 14769.2, respectively. The area under the ROC curve (SE), root mean square and -2Loglikelihood of the artificial neural network was 0.754 (0.004), 0.3770 and 14757.6, respectively. Conclusions: Based on these three criteria, artificial neural network would give better performance than logistic regression. Although, the difference is statistically significant, it does not seem to be clinically significant. PMID:23113198
Comparison of logistic regression and artificial neural network in low back pain prediction: second national health survey.

PubMed

Parsaeian, M; Mohammad, K; Mahmoudi, M; Zeraati, H

2012-01-01

The purpose of this investigation was to compare empirically predictive ability of an artificial neural network with a logistic regression in prediction of low back pain. Data from the second national health survey were considered in this investigation. This data includes the information of low back pain and its associated risk factors among Iranian people aged 15 years and older. Artificial neural network and logistic regression models were developed using a set of 17294 data and they were validated in a test set of 17295 data. Hosmer and Lemeshow recommendation for model selection was used in fitting the logistic regression. A three-layer perceptron with 9 inputs, 3 hidden and 1 output neurons was employed. The efficiency of two models was compared by receiver operating characteristic analysis, root mean square and -2 Loglikelihood criteria. The area under the ROC curve (SE), root mean square and -2Loglikelihood of the logistic regression was 0.752 (0.004), 0.3832 and 14769.2, respectively. The area under the ROC curve (SE), root mean square and -2Loglikelihood of the artificial neural network was 0.754 (0.004), 0.3770 and 14757.6, respectively. Based on these three criteria, artificial neural network would give better performance than logistic regression. Although, the difference is statistically significant, it does not seem to be clinically significant.
Modelling of binary logistic regression for obesity among secondary students in a rural area of Kedah

NASA Astrophysics Data System (ADS)

Kamaruddin, Ainur Amira; Ali, Zalila; Noor, Norlida Mohd.; Baharum, Adam; Ahmad, Wan Muhamad Amir W.

2014-07-01

Logistic regression analysis examines the influence of various factors on a dichotomous outcome by estimating the probability of the event's occurrence. Logistic regression, also called a logit model, is a statistical procedure used to model dichotomous outcomes. In the logit model the log odds of the dichotomous outcome is modeled as a linear combination of the predictor variables. The log odds ratio in logistic regression provides a description of the probabilistic relationship of the variables and the outcome. In conducting logistic regression, selection procedures are used in selecting important predictor variables, diagnostics are used to check that assumptions are valid which include independence of errors, linearity in the logit for continuous variables, absence of multicollinearity, and lack of strongly influential outliers and a test statistic is calculated to determine the aptness of the model. This study used the binary logistic regression model to investigate overweight and obesity among rural secondary school students on the basis of their demographics profile, medical history, diet and lifestyle. The results indicate that overweight and obesity of students are influenced by obesity in family and the interaction between a student's ethnicity and routine meals intake. The odds of a student being overweight and obese are higher for a student having a family history of obesity and for a non-Malay student who frequently takes routine meals as compared to a Malay student.
Genetic prediction of type 2 diabetes using deep neural network.

PubMed

Kim, J; Kim, J; Kwak, M J; Bajaj, M

2018-04-01

Type 2 diabetes (T2DM) has strong heritability but genetic models to explain heritability have been challenging. We tested deep neural network (DNN) to predict T2DM using the nested case-control study of Nurses' Health Study (3326 females, 45.6% T2DM) and Health Professionals Follow-up Study (2502 males, 46.5% T2DM). We selected 96, 214, 399, and 678 single-nucleotide polymorphism (SNPs) through Fisher's exact test and L1-penalized logistic regression. We split each dataset randomly in 4:1 to train prediction models and test their performance. DNN and logistic regressions showed better area under the curve (AUC) of ROC curves than the clinical model when 399 or more SNPs included. DNN was superior than logistic regressions in AUC with 399 or more SNPs in male and 678 SNPs in female. Addition of clinical factors consistently increased AUC of DNN but failed to improve logistic regressions with 214 or more SNPs. In conclusion, we show that DNN can be a versatile tool to predict T2DM incorporating large numbers of SNPs and clinical information. Limitations include a relatively small number of the subjects mostly of European ethnicity. Further studies are warranted to confirm and improve performance of genetic prediction models using DNN in different ethnic groups. © 2017 John Wiley & Sons A/S. Published by John Wiley & Sons Ltd.
Use of generalized ordered logistic regression for the analysis of multidrug resistance data.

PubMed

Agga, Getahun E; Scott, H Morgan

2015-10-01

Statistical analysis of antimicrobial resistance data largely focuses on individual antimicrobial's binary outcome (susceptible or resistant). However, bacteria are becoming increasingly multidrug resistant (MDR). Statistical analysis of MDR data is mostly descriptive often with tabular or graphical presentations. Here we report the applicability of generalized ordinal logistic regression model for the analysis of MDR data. A total of 1,152 Escherichia coli, isolated from the feces of weaned pigs experimentally supplemented with chlortetracycline (CTC) and copper, were tested for susceptibilities against 15 antimicrobials and were binary classified into resistant or susceptible. The 15 antimicrobial agents tested were grouped into eight different antimicrobial classes. We defined MDR as the number of antimicrobial classes to which E. coli isolates were resistant ranging from 0 to 8. Proportionality of the odds assumption of the ordinal logistic regression model was violated only for the effect of treatment period (pre-treatment, during-treatment and post-treatment); but not for the effect of CTC or copper supplementation. Subsequently, a partially constrained generalized ordinal logistic model was built that allows for the effect of treatment period to vary while constraining the effects of treatment (CTC and copper supplementation) to be constant across the levels of MDR classes. Copper (Proportional Odds Ratio [Prop OR]=1.03; 95% CI=0.73-1.47) and CTC (Prop OR=1.1; 95% CI=0.78-1.56) supplementation were not significantly associated with the level of MDR adjusted for the effect of treatment period. MDR generally declined over the trial period. In conclusion, generalized ordered logistic regression can be used for the analysis of ordinal data such as MDR data when the proportionality assumptions for ordered logistic regression are violated. Published by Elsevier B.V.
Science of Test Research Consortium: Year Two Final Report

DTIC Science & Technology

2012-10-02

July 2012. Analysis of an Intervention for Small Unmanned Aerial System ( SUAS ) Accidents, submitted to Quality Engineering, LQEN-2012-0056. Stone... Systems Engineering. Wolf, S. E., R. R. Hill, and J. J. Pignatiello. June 2012. Using Neural Networks and Logistic Regression to Model Small Unmanned ...Human Retina. 6. Wolf, S. E. March 2012. Modeling Small Unmanned Aerial System Mishaps using Logistic Regression and Artificial Neural Networks. 7
HRCT findings of collagen vascular disease-related interstitial pneumonia (CVD-IP): a comparative study among individual underlying diseases.

PubMed

Tanaka, N; Kunihiro, Y; Kubo, M; Kawano, R; Oishi, K; Ueda, K; Gondo, T

2018-05-29

To identify characteristic high-resolution computed tomography (CT) findings for individual collagen vascular disease (CVD)-related interstitial pneumonias (IPs). The HRCT findings of 187 patients with CVD, including 55 patients with rheumatoid arthritis (RA), 50 with systemic sclerosis (SSc), 46 with polymyositis/dermatomyositis (PM/DM), 15 with mixed connective tissue disease, 11 with primary Sjögren's syndrome, and 10 with systemic lupus erythematosus, were evaluated. Lung parenchymal abnormalities were compared among CVDs using χ 2 test, Kruskal-Wallis test, and multiple logistic regression analysis. A CT-pathology correlation was performed in 23 patients. In RA-IP, honeycombing was identified as the significant indicator based on multiple logistic regression analyses. Traction bronchiectasis (81.8%) was further identified as the most frequent finding based on χ 2 test. In SSc IP, lymph node enlargement and oesophageal dilatation were identified as the indicators based on multiple logistic regression analyses, and ground-glass opacity (GGO) was the most extensive based on Kruskal-Wallis test, which reflects the higher frequency of the pathological nonspecific interstitial pneumonia (NSIP) pattern present in the CT-pathology correlation. In PM/DM IP, airspace consolidation and the absence of honeycombing were identified as the indicators based on multiple logistic regression analyses, and predominance of consolidation over GGO (32.6%) and predominant subpleural distribution of GGO/consolidation (41.3%) were further identified as the most frequent findings based on χ 2 test, which reflects the higher frequency of the pathological NSIP and/or the organising pneumonia patterns present in the CT-pathology correlation. Several characteristic high-resolution CT findings with utility for estimating underlying CVD were identified. Copyright © 2018 The Royal College of Radiologists. Published by Elsevier Ltd. All rights reserved.
Logistic regression function for detection of suspicious performance during baseline evaluations using concussion vital signs.

PubMed

Hill, Benjamin David; Womble, Melissa N; Rohling, Martin L

2015-01-01

This study utilized logistic regression to determine whether performance patterns on Concussion Vital Signs (CVS) could differentiate known groups with either genuine or feigned performance. For the embedded measure development group (n = 174), clinical patients and undergraduate students categorized as feigning obtained significantly lower scores on the overall test battery mean for the CVS, Shipley-2 composite score, and California Verbal Learning Test-Second Edition subtests than did genuinely performing individuals. The final full model of 3 predictor variables (Verbal Memory immediate hits, Verbal Memory immediate correct passes, and Stroop Test complex reaction time correct) was significant and correctly classified individuals in their known group 83% of the time (sensitivity = .65; specificity = .97) in a mixed sample of young-adult clinical cases and simulators. The CVS logistic regression function was applied to a separate undergraduate college group (n = 378) that was asked to perform genuinely and identified 5% as having possibly feigned performance indicating a low false-positive rate. The failure rate was 11% and 16% at baseline cognitive testing in samples of high school and college athletes, respectively. These findings have particular relevance given the increasing use of computerized test batteries for baseline cognitive testing and return-to-play decisions after concussion.
Statistical power analyses using G*Power 3.1: tests for correlation and regression analyses.

PubMed

Faul, Franz; Erdfelder, Edgar; Buchner, Axel; Lang, Albert-Georg

2009-11-01

G*Power is a free power analysis program for a variety of statistical tests. We present extensions and improvements of the version introduced by Faul, Erdfelder, Lang, and Buchner (2007) in the domain of correlation and regression analyses. In the new version, we have added procedures to analyze the power of tests based on (1) single-sample tetrachoric correlations, (2) comparisons of dependent correlations, (3) bivariate linear regression, (4) multiple linear regression based on the random predictor model, (5) logistic regression, and (6) Poisson regression. We describe these new features and provide a brief introduction to their scope and handling.
A general framework for the use of logistic regression models in meta-analysis.

PubMed

Simmonds, Mark C; Higgins, Julian Pt

2016-12-01

Where individual participant data are available for every randomised trial in a meta-analysis of dichotomous event outcomes, "one-stage" random-effects logistic regression models have been proposed as a way to analyse these data. Such models can also be used even when individual participant data are not available and we have only summary contingency table data. One benefit of this one-stage regression model over conventional meta-analysis methods is that it maximises the correct binomial likelihood for the data and so does not require the common assumption that effect estimates are normally distributed. A second benefit of using this model is that it may be applied, with only minor modification, in a range of meta-analytic scenarios, including meta-regression, network meta-analyses and meta-analyses of diagnostic test accuracy. This single model can potentially replace the variety of often complex methods used in these areas. This paper considers, with a range of meta-analysis examples, how random-effects logistic regression models may be used in a number of different types of meta-analyses. This one-stage approach is compared with widely used meta-analysis methods including Bayesian network meta-analysis and the bivariate and hierarchical summary receiver operating characteristic (ROC) models for meta-analyses of diagnostic test accuracy. © The Author(s) 2014.
Radiomorphometric analysis of frontal sinus for sex determination.

PubMed

Verma, Saumya; Mahima, V G; Patil, Karthikeya

2014-09-01

Sex determination of unknown individuals carries crucial significance in forensic research, in cases where fragments of skull persist with no likelihood of identification based on dental arch. In these instances sex determination becomes important to rule out certain number of possibilities instantly and helps in establishing a biological profile of human remains. The aim of the study is to evaluate a mathematical method based on logistic regression analysis capable of ascertaining the sex of individuals in the South Indian population. The study was conducted in the department of Oral Medicine and Radiology. The right and left areas, maximum height, width of frontal sinus were determined in 100 Caldwell views of 50 women and 50 men aged 20 years and above, with the help of Vernier callipers and a square grid with 1 square measuring 1mm(2) in area. Student's t-test, logistic regression analysis. The mean values of variables were greater in men, based on Student's t-test at 5% level of significance. The mathematical model based on logistic regression analysis gave percentage agreement of total area to correctly predict the female gender as 55.2%, of right area as 60.9% and of left area as 55.2%. The areas of the frontal sinus and the logistic regression proved to be unreliable in sex determination. (Logit = 0.924 - 0.00217 × right area).
Unconditional or Conditional Logistic Regression Model for Age-Matched Case-Control Data?

PubMed

Kuo, Chia-Ling; Duan, Yinghui; Grady, James

2018-01-01

Matching on demographic variables is commonly used in case-control studies to adjust for confounding at the design stage. There is a presumption that matched data need to be analyzed by matched methods. Conditional logistic regression has become a standard for matched case-control data to tackle the sparse data problem. The sparse data problem, however, may not be a concern for loose-matching data when the matching between cases and controls is not unique, and one case can be matched to other controls without substantially changing the association. Data matched on a few demographic variables are clearly loose-matching data, and we hypothesize that unconditional logistic regression is a proper method to perform. To address the hypothesis, we compare unconditional and conditional logistic regression models by precision in estimates and hypothesis testing using simulated matched case-control data. Our results support our hypothesis; however, the unconditional model is not as robust as the conditional model to the matching distortion that the matching process not only makes cases and controls similar for matching variables but also for the exposure status. When the study design involves other complex features or the computational burden is high, matching in loose-matching data can be ignored for negligible loss in testing and estimation if the distributions of matching variables are not extremely different between cases and controls.
Unconditional or Conditional Logistic Regression Model for Age-Matched Case–Control Data?

PubMed Central

Kuo, Chia-Ling; Duan, Yinghui; Grady, James

2018-01-01

Matching on demographic variables is commonly used in case–control studies to adjust for confounding at the design stage. There is a presumption that matched data need to be analyzed by matched methods. Conditional logistic regression has become a standard for matched case–control data to tackle the sparse data problem. The sparse data problem, however, may not be a concern for loose-matching data when the matching between cases and controls is not unique, and one case can be matched to other controls without substantially changing the association. Data matched on a few demographic variables are clearly loose-matching data, and we hypothesize that unconditional logistic regression is a proper method to perform. To address the hypothesis, we compare unconditional and conditional logistic regression models by precision in estimates and hypothesis testing using simulated matched case–control data. Our results support our hypothesis; however, the unconditional model is not as robust as the conditional model to the matching distortion that the matching process not only makes cases and controls similar for matching variables but also for the exposure status. When the study design involves other complex features or the computational burden is high, matching in loose-matching data can be ignored for negligible loss in testing and estimation if the distributions of matching variables are not extremely different between cases and controls. PMID:29552553
Polymorphism Thr160Thr in SRD5A1, involved in the progesterone metabolism, modifies postmenopausal breast cancer risk associated with menopausal hormone therapy.

PubMed

Hein, R; Abbas, S; Seibold, P; Salazar, R; Flesch-Janys, D; Chang-Claude, J

2012-01-01

Menopausal hormone therapy (MHT) is associated with an increased breast cancer risk in postmenopausal women, with combined estrogen-progestagen therapy posing a greater risk than estrogen monotherapy. However, few studies focused on potential effect modification of MHT-associated breast cancer risk by genetic polymorphisms in the progesterone metabolism. We assessed effect modification of MHT use by five coding single nucleotide polymorphisms (SNPs) in the progesterone metabolizing enzymes AKR1C3 (rs7741), AKR1C4 (rs3829125, rs17134592), and SRD5A1 (rs248793, rs3736316) using a two-center population-based case-control study from Germany with 2,502 postmenopausal breast cancer patients and 4,833 matched controls. An empirical-Bayes procedure that tests for interaction using a weighted combination of the prospective and the retrospective case-control estimators as well as standard prospective logistic regression were applied to assess multiplicative statistical interaction between polymorphisms and duration of MHT use with regard to breast cancer risk assuming a log-additive mode of inheritance. No genetic marginal effects were observed. Breast cancer risk associated with duration of combined therapy was significantly modified by SRD5A1_rs3736316, showing a reduced risk elevation in carriers of the minor allele (p (interaction,empirical-Bayes) = 0.006 using the empirical-Bayes method, p (interaction,logistic regression) = 0.013 using logistic regression). The risk associated with duration of use of monotherapy was increased by AKR1C3_rs7741 in minor allele carriers (p (interaction,empirical-Bayes) = 0.083, p (interaction,logistic regression) = 0.029) and decreased in minor allele carriers of two SNPs in AKR1C4 (rs3829125: p (interaction,empirical-Bayes) = 0.07, p (interaction,logistic regression) = 0.021; rs17134592: p (interaction,empirical-Bayes) = 0.101, p (interaction,logistic regression) = 0.038). After Bonferroni correction for multiple testing only SRD5A1_rs3736316 assessed using the empirical-Bayes method remained significant. Postmenopausal breast cancer risk associated with combined therapy may be modified by genetic variation in SRD5A1. Further well-powered studies are, however, required to replicate our finding.
Neuropsychological tests for predicting cognitive decline in older adults

PubMed Central

Baerresen, Kimberly M; Miller, Karen J; Hanson, Eric R; Miller, Justin S; Dye, Richelin V; Hartman, Richard E; Vermeersch, David; Small, Gary W

2015-01-01

Summary Aim To determine neuropsychological tests likely to predict cognitive decline. Methods A sample of nonconverters (n = 106) was compared with those who declined in cognitive status (n = 24). Significant univariate logistic regression prediction models were used to create multivariate logistic regression models to predict decline based on initial neuropsychological testing. Results Rey–Osterrieth Complex Figure Test (RCFT) Retention predicted conversion to mild cognitive impairment (MCI) while baseline Buschke Delay predicted conversion to Alzheimer’s disease (AD). Due to group sample size differences, additional analyses were conducted using a subsample of demographically matched nonconverters. Analyses indicated RCFT Retention predicted conversion to MCI and AD, and Buschke Delay predicted conversion to AD. Conclusion Results suggest RCFT Retention and Buschke Delay may be useful in predicting cognitive decline. PMID:26107318
Nowcasting of Low-Visibility Procedure States with Ordered Logistic Regression at Vienna International Airport

NASA Astrophysics Data System (ADS)

Kneringer, Philipp; Dietz, Sebastian; Mayr, Georg J.; Zeileis, Achim

2017-04-01

Low-visibility conditions have a large impact on aviation safety and economic efficiency of airports and airlines. To support decision makers, we develop a statistical probabilistic nowcasting tool for the occurrence of capacity-reducing operations related to low visibility. The probabilities of four different low visibility classes are predicted with an ordered logistic regression model based on time series of meteorological point measurements. Potential predictor variables for the statistical models are visibility, humidity, temperature and wind measurements at several measurement sites. A stepwise variable selection method indicates that visibility and humidity measurements are the most important model inputs. The forecasts are tested with a 30 minute forecast interval up to two hours, which is a sufficient time span for tactical planning at Vienna Airport. The ordered logistic regression models outperform persistence and are competitive with human forecasters.
Measurement of faculty anesthesiologists' quality of clinical supervision has greater reliability when controlling for the leniency of the rating anesthesia resident: a retrospective cohort study.

PubMed

Dexter, Franklin; Ledolter, Johannes; Hindman, Bradley J

2017-06-01

Our department monitors the quality of anesthesiologists' clinical supervision and provides each anesthesiologist with periodic feedback. We hypothesized that greater differentiation among anesthesiologists' supervision scores could be obtained by adjusting for leniency of the rating resident. From July 1, 2013 to December 31, 2015, our department has utilized the de Oliveira Filho unidimensional nine-item supervision scale to assess the quality of clinical supervision provided by faculty as rated by residents. We examined all 13,664 ratings of the 97 anesthesiologists (ratees) by the 65 residents (raters). Testing for internal consistency among answers to questions (large Cronbach's alpha > 0.90) was performed to rule out that one or two questions accounted for leniency. Mixed-effects logistic regression was used to compare ratees while controlling for rater leniency vs using Student t tests without rater leniency. The mean supervision scale score was calculated for each combination of the 65 raters and nine questions. The Cronbach's alpha was very large (0.977). The mean score was calculated for each of the 3,421 observed combinations of resident and anesthesiologist. The logits of the percentage of scores equal to the maximum value of 4.00 were normally distributed (residents, P = 0.24; anesthesiologists, P = 0.50). There were 20/97 anesthesiologists identified as significant outliers (13 with below average supervision scores and seven with better than average) using the mixed-effects logistic regression with rater leniency entered as a fixed effect but not by Student's t test. In contrast, there were three of 97 anesthesiologists identified as outliers (all three above average) using Student's t tests but not by logistic regression with leniency. The 20 vs 3 was significant (P < 0.001). Use of logistic regression with leniency results in greater detection of anesthesiologists with significantly better (or worse) clinical supervision scores than use of Student's t tests (i.e., without adjustment for rater leniency).
The effect of high leverage points on the logistic ridge regression estimator having multicollinearity

NASA Astrophysics Data System (ADS)

Ariffin, Syaiba Balqish; Midi, Habshah

2014-06-01

This article is concerned with the performance of logistic ridge regression estimation technique in the presence of multicollinearity and high leverage points. In logistic regression, multicollinearity exists among predictors and in the information matrix. The maximum likelihood estimator suffers a huge setback in the presence of multicollinearity which cause regression estimates to have unduly large standard errors. To remedy this problem, a logistic ridge regression estimator is put forward. It is evident that the logistic ridge regression estimator outperforms the maximum likelihood approach for handling multicollinearity. The effect of high leverage points are then investigated on the performance of the logistic ridge regression estimator through real data set and simulation study. The findings signify that logistic ridge regression estimator fails to provide better parameter estimates in the presence of both high leverage points and multicollinearity.
Sample size determination for logistic regression on a logit-normal distribution.

PubMed

Kim, Seongho; Heath, Elisabeth; Heilbrun, Lance

2017-06-01

Although the sample size for simple logistic regression can be readily determined using currently available methods, the sample size calculation for multiple logistic regression requires some additional information, such as the coefficient of determination ([Formula: see text]) of a covariate of interest with other covariates, which is often unavailable in practice. The response variable of logistic regression follows a logit-normal distribution which can be generated from a logistic transformation of a normal distribution. Using this property of logistic regression, we propose new methods of determining the sample size for simple and multiple logistic regressions using a normal transformation of outcome measures. Simulation studies and a motivating example show several advantages of the proposed methods over the existing methods: (i) no need for [Formula: see text] for multiple logistic regression, (ii) available interim or group-sequential designs, and (iii) much smaller required sample size.
A comparison of Cox and logistic regression for use in genome-wide association studies of cohort and case-cohort design.

PubMed

Staley, James R; Jones, Edmund; Kaptoge, Stephen; Butterworth, Adam S; Sweeting, Michael J; Wood, Angela M; Howson, Joanna M M

2017-06-01

Logistic regression is often used instead of Cox regression to analyse genome-wide association studies (GWAS) of single-nucleotide polymorphisms (SNPs) and disease outcomes with cohort and case-cohort designs, as it is less computationally expensive. Although Cox and logistic regression models have been compared previously in cohort studies, this work does not completely cover the GWAS setting nor extend to the case-cohort study design. Here, we evaluated Cox and logistic regression applied to cohort and case-cohort genetic association studies using simulated data and genetic data from the EPIC-CVD study. In the cohort setting, there was a modest improvement in power to detect SNP-disease associations using Cox regression compared with logistic regression, which increased as the disease incidence increased. In contrast, logistic regression had more power than (Prentice weighted) Cox regression in the case-cohort setting. Logistic regression yielded inflated effect estimates (assuming the hazard ratio is the underlying measure of association) for both study designs, especially for SNPs with greater effect on disease. Given logistic regression is substantially more computationally efficient than Cox regression in both settings, we propose a two-step approach to GWAS in cohort and case-cohort studies. First to analyse all SNPs with logistic regression to identify associated variants below a pre-defined P-value threshold, and second to fit Cox regression (appropriately weighted in case-cohort studies) to those identified SNPs to ensure accurate estimation of association with disease.

Estimation of the Regression Effect Using a Latent Trait Model.

ERIC Educational Resources Information Center

Quinn, Jimmy L.

A logistic model was used to generate data to serve as a proxy for an immediate retest from item responses to a fourth grade standardized reading comprehension test of 45 items. Assuming that the actual test may be considered a pretest and the proxy data may be considered a retest, the effect of regression was investigated using a percentage of…
A comparative analysis of predictive models of morbidity in intensive care unit after cardiac surgery - part II: an illustrative example.

PubMed

Cevenini, Gabriele; Barbini, Emanuela; Scolletta, Sabino; Biagioli, Bonizella; Giomarelli, Pierpaolo; Barbini, Paolo

2007-11-22

Popular predictive models for estimating morbidity probability after heart surgery are compared critically in a unitary framework. The study is divided into two parts. In the first part modelling techniques and intrinsic strengths and weaknesses of different approaches were discussed from a theoretical point of view. In this second part the performances of the same models are evaluated in an illustrative example. Eight models were developed: Bayes linear and quadratic models, k-nearest neighbour model, logistic regression model, Higgins and direct scoring systems and two feed-forward artificial neural networks with one and two layers. Cardiovascular, respiratory, neurological, renal, infectious and hemorrhagic complications were defined as morbidity. Training and testing sets each of 545 cases were used. The optimal set of predictors was chosen among a collection of 78 preoperative, intraoperative and postoperative variables by a stepwise procedure. Discrimination and calibration were evaluated by the area under the receiver operating characteristic curve and Hosmer-Lemeshow goodness-of-fit test, respectively. Scoring systems and the logistic regression model required the largest set of predictors, while Bayesian and k-nearest neighbour models were much more parsimonious. In testing data, all models showed acceptable discrimination capacities, however the Bayes quadratic model, using only three predictors, provided the best performance. All models showed satisfactory generalization ability: again the Bayes quadratic model exhibited the best generalization, while artificial neural networks and scoring systems gave the worst results. Finally, poor calibration was obtained when using scoring systems, k-nearest neighbour model and artificial neural networks, while Bayes (after recalibration) and logistic regression models gave adequate results. Although all the predictive models showed acceptable discrimination performance in the example considered, the Bayes and logistic regression models seemed better than the others, because they also had good generalization and calibration. The Bayes quadratic model seemed to be a convincing alternative to the much more usual Bayes linear and logistic regression models. It showed its capacity to identify a minimum core of predictors generally recognized as essential to pragmatically evaluate the risk of developing morbidity after heart surgery.
The crux of the method: assumptions in ordinary least squares and logistic regression.

PubMed

Long, Rebecca G

2008-10-01

Logistic regression has increasingly become the tool of choice when analyzing data with a binary dependent variable. While resources relating to the technique are widely available, clear discussions of why logistic regression should be used in place of ordinary least squares regression are difficult to find. The current paper compares and contrasts the assumptions of ordinary least squares with those of logistic regression and explains why logistic regression's looser assumptions make it adept at handling violations of the more important assumptions in ordinary least squares.
Using Dominance Analysis to Determine Predictor Importance in Logistic Regression

ERIC Educational Resources Information Center

Azen, Razia; Traxel, Nicole

2009-01-01

This article proposes an extension of dominance analysis that allows researchers to determine the relative importance of predictors in logistic regression models. Criteria for choosing logistic regression R[superscript 2] analogues were determined and measures were selected that can be used to perform dominance analysis in logistic regression. A…
Prediction model for the return to work of workers with injuries in Hong Kong.

PubMed

Xu, Yanwen; Chan, Chetwyn C H; Lo, Karen Hui Yu-Ling; Tang, Dan

2008-01-01

This study attempts to formulate a prediction model of return to work for a group of workers who have been suffering from chronic pain and physical injury while also being out of work in Hong Kong. The study used Case-based Reasoning (CBR) method, and compared the result with the statistical method of logistic regression model. The database of the algorithm of CBR was composed of 67 cases who were also used in the logistic regression model. The testing cases were 32 participants who had a similar background and characteristics to those in the database. The methods of setting constraints and Euclidean distance metric were used in CBR to search the closest cases to the trial case based on the matrix. The usefulness of the algorithm was tested on 32 new participants, and the accuracy of predicting return to work outcomes was 62.5%, which was no better than the 71.2% accuracy derived from the logistic regression model. The results of the study would enable us to have a better understanding of the CBR applied in the field of occupational rehabilitation by comparing with the conventional regression analysis. The findings would also shed light on the development of relevant interventions for the return-to-work process of these workers.
Forecasting Air Force Logistics Command Second Destination Transportation: An Application of Multiple Regression Analysis and Neural Networks

DTIC Science & Technology

1990-09-01

without the help from the DSXR staff. William Lyons, Charles Ramsey , and Martin Meeks went above and beyond to help complete this research. Special...develop a valid forecasting model that is significantly more accurate than the one presently used by DSXR and suggested the development and testing of a...method, Strom tested DSXR’s iterative linear regression forecasting technique by examining P1 in the simple regression equation to determine whether
Applying Kaplan-Meier to Item Response Data

ERIC Educational Resources Information Center

McNeish, Daniel

2018-01-01

Some IRT models can be equivalently modeled in alternative frameworks such as logistic regression. Logistic regression can also model time-to-event data, which concerns the probability of an event occurring over time. Using the relation between time-to-event models and logistic regression and the relation between logistic regression and IRT, this…
Testing for gene-environment interaction under exposure misspecification.

PubMed

Sun, Ryan; Carroll, Raymond J; Christiani, David C; Lin, Xihong

2017-11-09

Complex interplay between genetic and environmental factors characterizes the etiology of many diseases. Modeling gene-environment (GxE) interactions is often challenged by the unknown functional form of the environment term in the true data-generating mechanism. We study the impact of misspecification of the environmental exposure effect on inference for the GxE interaction term in linear and logistic regression models. We first examine the asymptotic bias of the GxE interaction regression coefficient, allowing for confounders as well as arbitrary misspecification of the exposure and confounder effects. For linear regression, we show that under gene-environment independence and some confounder-dependent conditions, when the environment effect is misspecified, the regression coefficient of the GxE interaction can be unbiased. However, inference on the GxE interaction is still often incorrect. In logistic regression, we show that the regression coefficient is generally biased if the genetic factor is associated with the outcome directly or indirectly. Further, we show that the standard robust sandwich variance estimator for the GxE interaction does not perform well in practical GxE studies, and we provide an alternative testing procedure that has better finite sample properties. © 2017, The International Biometric Society.
Modeling recall memory for emotional objects in Alzheimer's disease.

PubMed

Sundstrøm, Martin

2011-07-01

To examine whether emotional memory (EM) of objects with self-reference in Alzheimer's disease (AD) can be modeled with binomial logistic regression in a free recall and an object recognition test to predict EM enhancement. Twenty patients with AD and twenty healthy controls were studied. Six objects (three presented as gifts) were shown to each participant. Ten minutes later, a free recall and a recognition test were applied. The recognition test had target-objects mixed with six similar distracter objects. Participants were asked to name any object in the recall test and identify each object in the recognition test as known or unknown. The total of gift objects recalled in AD patients (41.6%) was larger than neutral objects (13.3%) and a significant EM recall effect for gifts was found (Wilcoxon: p < .003). EM was not found for recognition in AD patients due to a ceiling effect. Healthy older adults scored overall higher in recall and recognition but showed no EM enhancement due to a ceiling effect. A logistic regression showed that likelihood of emotional recall memory can be modeled as a function of MMSE score (p < .014) and object status (p < .0001) as gift or non-gift. Recall memory was enhanced in AD patients for emotional objects indicating that EM in mild to moderate AD although impaired can be provoked with strong emotional load. The logistic regression model suggests that EM declines with the progression of AD rather than disrupts and may be a useful tool for evaluating magnitude of emotional load.
A Powerful Test for Comparing Multiple Regression Functions.

PubMed

Maity, Arnab

2012-09-01

In this article, we address the important problem of comparison of two or more population regression functions. Recently, Pardo-Fernández, Van Keilegom and González-Manteiga (2007) developed test statistics for simple nonparametric regression models: Y(ij) = θ(j)(Z(ij)) + σ(j)(Z(ij))∊(ij), based on empirical distributions of the errors in each population j = 1, … , J. In this paper, we propose a test for equality of the θ(j)(·) based on the concept of generalized likelihood ratio type statistics. We also generalize our test for other nonparametric regression setups, e.g, nonparametric logistic regression, where the loglikelihood for population j is any general smooth function [Formula: see text]. We describe a resampling procedure to obtain the critical values of the test. In addition, we present a simulation study to evaluate the performance of the proposed test and compare our results to those in Pardo-Fernández et al. (2007).
Comparative study of biodegradability prediction of chemicals using decision trees, functional trees, and logistic regression.

PubMed

Chen, Guangchao; Li, Xuehua; Chen, Jingwen; Zhang, Ya-Nan; Peijnenburg, Willie J G M

2014-12-01

Biodegradation is the principal environmental dissipation process of chemicals. As such, it is a dominant factor determining the persistence and fate of organic chemicals in the environment, and is therefore of critical importance to chemical management and regulation. In the present study, the authors developed in silico methods assessing biodegradability based on a large heterogeneous set of 825 organic compounds, using the techniques of the C4.5 decision tree, the functional inner regression tree, and logistic regression. External validation was subsequently carried out by 2 independent test sets of 777 and 27 chemicals. As a result, the functional inner regression tree exhibited the best predictability with predictive accuracies of 81.5% and 81.0%, respectively, on the training set (825 chemicals) and test set I (777 chemicals). Performance of the developed models on the 2 test sets was subsequently compared with that of the Estimation Program Interface (EPI) Suite Biowin 5 and Biowin 6 models, which also showed a better predictability of the functional inner regression tree model. The model built in the present study exhibits a reasonable predictability compared with existing models while possessing a transparent algorithm. Interpretation of the mechanisms of biodegradation was also carried out based on the models developed. © 2014 SETAC.
Improving virtual screening predictive accuracy of Human kallikrein 5 inhibitors using machine learning models.

PubMed

Fang, Xingang; Bagui, Sikha; Bagui, Subhash

2017-08-01

The readily available high throughput screening (HTS) data from the PubChem database provides an opportunity for mining of small molecules in a variety of biological systems using machine learning techniques. From the thousands of available molecular descriptors developed to encode useful chemical information representing the characteristics of molecules, descriptor selection is an essential step in building an optimal quantitative structural-activity relationship (QSAR) model. For the development of a systematic descriptor selection strategy, we need the understanding of the relationship between: (i) the descriptor selection; (ii) the choice of the machine learning model; and (iii) the characteristics of the target bio-molecule. In this work, we employed the Signature descriptor to generate a dataset on the Human kallikrein 5 (hK 5) inhibition confirmatory assay data and compared multiple classification models including logistic regression, support vector machine, random forest and k-nearest neighbor. Under optimal conditions, the logistic regression model provided extremely high overall accuracy (98%) and precision (90%), with good sensitivity (65%) in the cross validation test. In testing the primary HTS screening data with more than 200K molecular structures, the logistic regression model exhibited the capability of eliminating more than 99.9% of the inactive structures. As part of our exploration of the descriptor-model-target relationship, the excellent predictive performance of the combination of the Signature descriptor and the logistic regression model on the assay data of the Human kallikrein 5 (hK 5) target suggested a feasible descriptor/model selection strategy on similar targets. Copyright © 2017 Elsevier Ltd. All rights reserved.
HIV testing among MSM in Bogotá, Colombia: The role of structural and individual characteristics

PubMed Central

Reisen, Carol A.; Zea, Maria Cecilia; Bianchi, Fernanda T.; Poppen, Paul J.; del Río González, Ana Maria; Romero, Rodrigo A. Aguayo; Pérez, Carolin

2014-01-01

This study used mixed methods to examine characteristics related to HIV testing among men who have sex with men (MSM) in Bogotá, Colombia. A sample of 890 MSM responded to a computerized quantitative survey. Follow-up qualitative data included 20 in-depth interviews with MSM and 12 key informant interviews. Hierarchical logistic set regression indicated that sequential sets of variables reflecting demographic characteristics, insurance coverage, risk appraisal, and social context each added to the explanation of HIV testing. Follow-up logistic regression showed that individuals who were older, had higher income, paid for their own insurance, had had a sexually transmitted infection, knew more people living with HIV, and had greater social support were more likely to have been tested for HIV at least once. Qualitative findings provided details of personal and structural barriers to testing, as well as interrelationships among these factors. Recommendations to increase HIV testing among Colombian MSM are offered. PMID:25068180
Efficient logistic regression designs under an imperfect population identifier.

PubMed

Albert, Paul S; Liu, Aiyi; Nansel, Tonja

2014-03-01

Motivated by actual study designs, this article considers efficient logistic regression designs where the population is identified with a binary test that is subject to diagnostic error. We consider the case where the imperfect test is obtained on all participants, while the gold standard test is measured on a small chosen subsample. Under maximum-likelihood estimation, we evaluate the optimal design in terms of sample selection as well as verification. We show that there may be substantial efficiency gains by choosing a small percentage of individuals who test negative on the imperfect test for inclusion in the sample (e.g., verifying 90% test-positive cases). We also show that a two-stage design may be a good practical alternative to a fixed design in some situations. Under optimal and nearly optimal designs, we compare maximum-likelihood and semi-parametric efficient estimators under correct and misspecified models with simulations. The methodology is illustrated with an analysis from a diabetes behavioral intervention trial. © 2013, The International Biometric Society.
Predictors of postoperative outcomes of cubital tunnel syndrome treatments using multiple logistic regression analysis.

PubMed

Suzuki, Taku; Iwamoto, Takuji; Shizu, Kanae; Suzuki, Katsuji; Yamada, Harumoto; Sato, Kazuki

2017-05-01

This retrospective study was designed to investigate prognostic factors for postoperative outcomes for cubital tunnel syndrome (CubTS) using multiple logistic regression analysis with a large number of patients. Eighty-three patients with CubTS who underwent surgeries were enrolled. The following potential prognostic factors for disease severity were selected according to previous reports: sex, age, type of surgery, disease duration, body mass index, cervical lesion, presence of diabetes mellitus, Workers' Compensation status, preoperative severity, and preoperative electrodiagnostic testing. Postoperative severity of disease was assessed 2 years after surgery by Messina's criteria which is an outcome measure specifically for CubTS. Bivariate analysis was performed to select candidate prognostic factors for multiple linear regression analyses. Multiple logistic regression analysis was conducted to identify the association between postoperative severity and selected prognostic factors. Both bivariate and multiple linear regression analysis revealed only preoperative severity as an independent risk factor for poor prognosis, while other factors did not show any significant association. Although conflicting results exist regarding prognosis of CubTS, this study supports evidence from previous studies and concludes early surgical intervention portends the most favorable prognosis. Copyright © 2017 The Japanese Orthopaedic Association. Published by Elsevier B.V. All rights reserved.
Inferring microhabitat preferences of Lilium catesbaei (Liliaceae).

PubMed

Sommers, Kristen Penney; Elswick, Michael; Herrick, Gabriel I; Fox, Gordon A

2011-05-01

Microhabitat studies use varied statistical methods, some treating site occupancy as a dependent and others as an independent variable. Using the rare Lilium catesbaei as an example, we show why approaches to testing hypotheses of differences between occupied and unoccupied sites can lead to erroneous conclusions about habitat preferences. Predictive approaches like logistic regression can better lead to understanding of habitat requirements. Using 32 lily locations and 30 random locations >2 m from a lily (complete data: 31 lily and 28 random spots), we measured physical conditions--photosynthetically active radiation (PAR), canopy cover, litter depth, distance to and height of nearest shrub, and soil moisture--and number and identity of neighboring plants. Twelve lilies were used to estimate a photosynthetic assimilation curve. Analyses used logistic regression, discriminant function analysis (DFA), (multivariate) analysis of variance, and resampled Wilcoxon tests. Logistic regression and DFA found identical predictors of presence (PAR, canopy cover, distance to shrub, litter), but hypothesis tests pointed to a different set (PAR, litter, canopy cover, height of nearest shrub). Lilies are mainly in high-PAR spots, often close to light saturation. By contrast, PAR in random spots was often near the lily light compensation point. Lilies were near Serenoa repens less than at random; otherwise, neighbor identity had no significant effect. Predictive methods are more useful in this context than the hypothesis tests. Light availability plays a big role in lily presence, which may help to explain increases in flowering and emergence after fire and roller-chopping.
Standards for Standardized Logistic Regression Coefficients

ERIC Educational Resources Information Center

Menard, Scott

2011-01-01

Standardized coefficients in logistic regression analysis have the same utility as standardized coefficients in linear regression analysis. Although there has been no consensus on the best way to construct standardized logistic regression coefficients, there is now sufficient evidence to suggest a single best approach to the construction of a…
A trend analysis of laboratory positive propoxyphene workplace urine drug screens before and after the product recall.

PubMed

Price, James

2015-01-01

Propoxyphene was withdrawn from the US market in November 2010. This drug is still tested for in the workplace as part of expanded panel nonregulated testing. A convenience sample of urine specimens (n = 7838) were provided by workers from various industries. The percentage of positive specimens with 95% confidence intervals was calculated for each year of the study. Logistic regression was used to assess the impact of the year upon the propoxyphene result. The prevalence of positive propoxyphene tests was much higher before the product's withdrawal from the market. Logistic regression provided evidence of a decreasing linear trend (P < 0.000; β = -0.71). The odds ratio signifies that for every additional year the urine specimens were 0.49 times less likely to be positive for propoxyphene. This favors the determination that the change in propoxyphene positive drug test over the years is not by chance. The conclusion supports no longer performing nonregulated workplace propoxyphene urine drug testing for this population.
Further investigations of the W-test for pairwise epistasis testing.

PubMed

Howey, Richard; Cordell, Heather J

2017-01-01

Background: In a recent paper, a novel W-test for pairwise epistasis testing was proposed that appeared, in computer simulations, to have higher power than competing alternatives. Application to genome-wide bipolar data detected significant epistasis between SNPs in genes of relevant biological function. Network analysis indicated that the implicated genes formed two separate interaction networks, each containing genes highly related to autism and neurodegenerative disorders. Methods: Here we investigate further the properties and performance of the W-test via theoretical evaluation, computer simulations and application to real data. Results: We demonstrate that, for common variants, the W-test is closely related to several existing tests of association allowing for interaction, including logistic regression on 8 degrees of freedom, although logistic regression can show inflated type I error for low minor allele frequencies, whereas the W-test shows good/conservative type I error control. Although in some situations the W-test can show higher power, logistic regression is not limited to tests on 8 degrees of freedom but can instead be tailored to impose greater structure on the assumed alternative hypothesis, offering a power advantage when the imposed structure matches the true structure. Conclusions: The W-test is a potentially useful method for testing for association - without necessarily implying interaction - between genetic variants disease, particularly when one or more of the genetic variants are rare. For common variants, the advantages of the W-test are less clear, and, indeed, there are situations where existing methods perform better. In our investigations, we further uncover a number of problems with the practical implementation and application of the W-test (to bipolar disorder) previously described, apparently due to inadequate use of standard data quality-control procedures. This observation leads us to urge caution in interpretation of the previously-presented results, most of which we consider are highly likely to be artefacts.
[Calculating Pearson residual in logistic regressions: a comparison between SPSS and SAS].

PubMed

Xu, Hao; Zhang, Tao; Li, Xiao-song; Liu, Yuan-yuan

2015-01-01

To compare the results of Pearson residual calculations in logistic regression models using SPSS and SAS. We reviewed Pearson residual calculation methods, and used two sets of data to test logistic models constructed by SPSS and STATA. One model contained a small number of covariates compared to the number of observed. The other contained a similar number of covariates as the number of observed. The two software packages produced similar Pearson residual estimates when the models contained a similar number of covariates as the number of observed, but the results differed when the number of observed was much greater than the number of covariates. The two software packages produce different results of Pearson residuals, especially when the models contain a small number of covariates. Further studies are warranted.

Propensity score estimation: machine learning and classification methods as alternatives to logistic regression

PubMed Central

Westreich, Daniel; Lessler, Justin; Funk, Michele Jonsson

2010-01-01

Summary Objective Propensity scores for the analysis of observational data are typically estimated using logistic regression. Our objective in this Review was to assess machine learning alternatives to logistic regression which may accomplish the same goals but with fewer assumptions or greater accuracy. Study Design and Setting We identified alternative methods for propensity score estimation and/or classification from the public health, biostatistics, discrete mathematics, and computer science literature, and evaluated these algorithms for applicability to the problem of propensity score estimation, potential advantages over logistic regression, and ease of use. Results We identified four techniques as alternatives to logistic regression: neural networks, support vector machines, decision trees (CART), and meta-classifiers (in particular, boosting). Conclusion While the assumptions of logistic regression are well understood, those assumptions are frequently ignored. All four alternatives have advantages and disadvantages compared with logistic regression. Boosting (meta-classifiers) and to a lesser extent decision trees (particularly CART) appear to be most promising for use in the context of propensity score analysis, but extensive simulation studies are needed to establish their utility in practice. PMID:20630332
Robust mislabel logistic regression without modeling mislabel probabilities.

PubMed

Hung, Hung; Jou, Zhi-Yu; Huang, Su-Yun

2018-03-01

Logistic regression is among the most widely used statistical methods for linear discriminant analysis. In many applications, we only observe possibly mislabeled responses. Fitting a conventional logistic regression can then lead to biased estimation. One common resolution is to fit a mislabel logistic regression model, which takes into consideration of mislabeled responses. Another common method is to adopt a robust M-estimation by down-weighting suspected instances. In this work, we propose a new robust mislabel logistic regression based on γ-divergence. Our proposal possesses two advantageous features: (1) It does not need to model the mislabel probabilities. (2) The minimum γ-divergence estimation leads to a weighted estimating equation without the need to include any bias correction term, that is, it is automatically bias-corrected. These features make the proposed γ-logistic regression more robust in model fitting and more intuitive for model interpretation through a simple weighting scheme. Our method is also easy to implement, and two types of algorithms are included. Simulation studies and the Pima data application are presented to demonstrate the performance of γ-logistic regression. © 2017, The International Biometric Society.
Fungible weights in logistic regression.

PubMed

Jones, Jeff A; Waller, Niels G

2016-06-01

In this article we develop methods for assessing parameter sensitivity in logistic regression models. To set the stage for this work, we first review Waller's (2008) equations for computing fungible weights in linear regression. Next, we describe 2 methods for computing fungible weights in logistic regression. To demonstrate the utility of these methods, we compute fungible logistic regression weights using data from the Centers for Disease Control and Prevention's (2010) Youth Risk Behavior Surveillance Survey, and we illustrate how these alternate weights can be used to evaluate parameter sensitivity. To make our work accessible to the research community, we provide R code (R Core Team, 2015) that will generate both kinds of fungible logistic regression weights. (PsycINFO Database Record (c) 2016 APA, all rights reserved).
Forecasting the probability of future groundwater levels declining below specified low thresholds in the conterminous U.S.

USGS Publications Warehouse

Dudley, Robert W.; Hodgkins, Glenn A.; Dickinson, Jesse

2017-01-01

We present a logistic regression approach for forecasting the probability of future groundwater levels declining or maintaining below specific groundwater-level thresholds. We tested our approach on 102 groundwater wells in different climatic regions and aquifers of the United States that are part of the U.S. Geological Survey Groundwater Climate Response Network. We evaluated the importance of current groundwater levels, precipitation, streamflow, seasonal variability, Palmer Drought Severity Index, and atmosphere/ocean indices for developing the logistic regression equations. Several diagnostics of model fit were used to evaluate the regression equations, including testing of autocorrelation of residuals, goodness-of-fit metrics, and bootstrap validation testing. The probabilistic predictions were most successful at wells with high persistence (low month-to-month variability) in their groundwater records and at wells where the groundwater level remained below the defined low threshold for sustained periods (generally three months or longer). The model fit was weakest at wells with strong seasonal variability in levels and with shorter duration low-threshold events. We identified challenges in deriving probabilistic-forecasting models and possible approaches for addressing those challenges.
Propensity score estimation: neural networks, support vector machines, decision trees (CART), and meta-classifiers as alternatives to logistic regression.

PubMed

Westreich, Daniel; Lessler, Justin; Funk, Michele Jonsson

2010-08-01

Propensity scores for the analysis of observational data are typically estimated using logistic regression. Our objective in this review was to assess machine learning alternatives to logistic regression, which may accomplish the same goals but with fewer assumptions or greater accuracy. We identified alternative methods for propensity score estimation and/or classification from the public health, biostatistics, discrete mathematics, and computer science literature, and evaluated these algorithms for applicability to the problem of propensity score estimation, potential advantages over logistic regression, and ease of use. We identified four techniques as alternatives to logistic regression: neural networks, support vector machines, decision trees (classification and regression trees [CART]), and meta-classifiers (in particular, boosting). Although the assumptions of logistic regression are well understood, those assumptions are frequently ignored. All four alternatives have advantages and disadvantages compared with logistic regression. Boosting (meta-classifiers) and, to a lesser extent, decision trees (particularly CART), appear to be most promising for use in the context of propensity score analysis, but extensive simulation studies are needed to establish their utility in practice. Copyright (c) 2010 Elsevier Inc. All rights reserved.
An ultra low power feature extraction and classification system for wearable seizure detection.

PubMed

Page, Adam; Pramod Tim Oates, Siddharth; Mohsenin, Tinoosh

2015-01-01

In this paper we explore the use of a variety of machine learning algorithms for designing a reliable and low-power, multi-channel EEG feature extractor and classifier for predicting seizures from electroencephalographic data (scalp EEG). Different machine learning classifiers including k-nearest neighbor, support vector machines, naïve Bayes, logistic regression, and neural networks are explored with the goal of maximizing detection accuracy while minimizing power, area, and latency. The input to each machine learning classifier is a 198 feature vector containing 9 features for each of the 22 EEG channels obtained over 1-second windows. All classifiers were able to obtain F1 scores over 80% and onset sensitivity of 100% when tested on 10 patients. Among five different classifiers that were explored, logistic regression (LR) proved to have minimum hardware complexity while providing average F-1 score of 91%. Both ASIC and FPGA implementations of logistic regression are presented and show the smallest area, power consumption, and the lowest latency when compared to the previous work.
Sexual Orientation Differences in HIV Testing Motivation among College Men

ERIC Educational Resources Information Center

Kort, Daniel N.; Samsa, Gregory P.; McKellar, Mehri S.

2017-01-01

Objective: To investigate sexual orientation differences in college men's motivations for HIV testing. Participants: 665 male college students in the Southeastern United States from 2006 to 2014. Methods: Students completed a survey on HIV risk factors and testing motivations. Logistic regressions were conducted to determine the differences…
Should metacognition be measured by logistic regression?

PubMed

Rausch, Manuel; Zehetleitner, Michael

2017-03-01

Are logistic regression slopes suitable to quantify metacognitive sensitivity, i.e. the efficiency with which subjective reports differentiate between correct and incorrect task responses? We analytically show that logistic regression slopes are independent from rating criteria in one specific model of metacognition, which assumes (i) that rating decisions are based on sensory evidence generated independently of the sensory evidence used for primary task responses and (ii) that the distributions of evidence are logistic. Given a hierarchical model of metacognition, logistic regression slopes depend on rating criteria. According to all considered models, regression slopes depend on the primary task criterion. A reanalysis of previous data revealed that massive numbers of trials are required to distinguish between hierarchical and independent models with tolerable accuracy. It is argued that researchers who wish to use logistic regression as measure of metacognitive sensitivity need to control the primary task criterion and rating criteria. Copyright © 2017 Elsevier Inc. All rights reserved.
Methodologic considerations in the design and analysis of nested case-control studies: association between cytokines and postoperative delirium.

PubMed

Ngo, Long H; Inouye, Sharon K; Jones, Richard N; Travison, Thomas G; Libermann, Towia A; Dillon, Simon T; Kuchel, George A; Vasunilashorn, Sarinnapha M; Alsop, David C; Marcantonio, Edward R

2017-06-06

The nested case-control study (NCC) design within a prospective cohort study is used when outcome data are available for all subjects, but the exposure of interest has not been collected, and is difficult or prohibitively expensive to obtain for all subjects. A NCC analysis with good matching procedures yields estimates that are as efficient and unbiased as estimates from the full cohort study. We present methodological considerations in a matched NCC design and analysis, which include the choice of match algorithms, analysis methods to evaluate the association of exposures of interest with outcomes, and consideration of overmatching. Matched, NCC design within a longitudinal observational prospective cohort study in the setting of two academic hospitals. Study participants are patients aged over 70 years who underwent scheduled major non-cardiac surgery. The primary outcome was postoperative delirium from in-hospital interviews and medical record review. The main exposure was IL-6 concentration (pg/ml) from blood sampled at three time points before delirium occurred. We used nonparametric signed ranked test to test for the median of the paired differences. We used conditional logistic regression to model the risk of IL-6 on delirium incidence. Simulation was used to generate a sample of cohort data on which unconditional multivariable logistic regression was used, and the results were compared to those of the conditional logistic regression. Partial R-square was used to assess the level of overmatching. We found that the optimal match algorithm yielded more matched pairs than the greedy algorithm. The choice of analytic strategy-whether to consider measured cytokine levels as the predictor or outcome-- yielded inferences that have different clinical interpretations but similar levels of statistical significance. Estimation results from NCC design using conditional logistic regression, and from simulated cohort design using unconditional logistic regression, were similar. We found minimal evidence for overmatching. Using a matched NCC approach introduces methodological challenges into the study design and data analysis. Nonetheless, with careful selection of the match algorithm, match factors, and analysis methods, this design is cost effective and, for our study, yields estimates that are similar to those from a prospective cohort study design.
Filtering data from the collaborative initial glaucoma treatment study for improved identification of glaucoma progression.

PubMed

Schell, Greggory J; Lavieri, Mariel S; Stein, Joshua D; Musch, David C

2013-12-21

Open-angle glaucoma (OAG) is a prevalent, degenerate ocular disease which can lead to blindness without proper clinical management. The tests used to assess disease progression are susceptible to process and measurement noise. The aim of this study was to develop a methodology which accounts for the inherent noise in the data and improve significant disease progression identification. Longitudinal observations from the Collaborative Initial Glaucoma Treatment Study (CIGTS) were used to parameterize and validate a Kalman filter model and logistic regression function. The Kalman filter estimates the true value of biomarkers associated with OAG and forecasts future values of these variables. We develop two logistic regression models via generalized estimating equations (GEE) for calculating the probability of experiencing significant OAG progression: one model based on the raw measurements from CIGTS and another model based on the Kalman filter estimates of the CIGTS data. Receiver operating characteristic (ROC) curves and associated area under the ROC curve (AUC) estimates are calculated using cross-fold validation. The logistic regression model developed using Kalman filter estimates as data input achieves higher sensitivity and specificity than the model developed using raw measurements. The mean AUC for the Kalman filter-based model is 0.961 while the mean AUC for the raw measurements model is 0.889. Hence, using the probability function generated via Kalman filter estimates and GEE for logistic regression, we are able to more accurately classify patients and instances as experiencing significant OAG progression. A Kalman filter approach for estimating the true value of OAG biomarkers resulted in data input which improved the accuracy of a logistic regression classification model compared to a model using raw measurements as input. This methodology accounts for process and measurement noise to enable improved discrimination between progression and nonprogression in chronic diseases.
Use of genetic programming, logistic regression, and artificial neural nets to predict readmission after coronary artery bypass surgery.

PubMed

Engoren, Milo; Habib, Robert H; Dooner, John J; Schwann, Thomas A

2013-08-01

As many as 14 % of patients undergoing coronary artery bypass surgery are readmitted within 30 days. Readmission is usually the result of morbidity and may lead to death. The purpose of this study is to develop and compare statistical and genetic programming models to predict readmission. Patients were divided into separate Construction and Validation populations. Using 88 variables, logistic regression, genetic programs, and artificial neural nets were used to develop predictive models. Models were first constructed and tested on the Construction populations, then validated on the Validation population. Areas under the receiver operator characteristic curves (AU ROC) were used to compare the models. Two hundred and two patients (7.6 %) in the 2,644 patient Construction group and 216 (8.0 %) of the 2,711 patient Validation group were re-admitted within 30 days of CABG surgery. Logistic regression predicted readmission with AU ROC = .675 ± .021 in the Construction group. Genetic programs significantly improved the accuracy, AU ROC = .767 ± .001, p < .001). Artificial neural nets were less accurate with AU ROC = 0.597 ± .001 in the Construction group. Predictive accuracy of all three techniques fell in the Validation group. However, the accuracy of genetic programming (AU ROC = .654 ± .001) was still trivially but statistically non-significantly better than that of the logistic regression (AU ROC = .644 ± .020, p = .61). Genetic programming and logistic regression provide alternative methods to predict readmission that are similarly accurate.
Logistic models--an odd(s) kind of regression.

PubMed

Jupiter, Daniel C

2013-01-01

The logistic regression model bears some similarity to the multivariable linear regression with which we are familiar. However, the differences are great enough to warrant a discussion of the need for and interpretation of logistic regression. Copyright © 2013 American College of Foot and Ankle Surgeons. Published by Elsevier Inc. All rights reserved.
The contribution of culture to Korean American women's cervical cancer screening behavior: the critical role of prevention orientation.

PubMed

Lee, Hee Yun; Roh, Soonhee; Vang, Suzanne; Jin, Seok Won

2011-01-01

Despite the proven benefits of Pap testing, Korean American women have one of the lowest cervical cancer screening rates in the United States. This study examined how cultural factors are associated with Pap test utilization among Korean American women participants. Quota sampling was used to recruit 202 Korean American women participants residing in New York City. Hierarchical logistic regression was used to assess the association of cultural variables with Pap test receipt. Overall, participants in our study reported significantly lower Pap test utilization; only 58% reported lifetime receipt of this screening test. Logistic regression analysis revealed one of the cultural variables--prevention orientation--was the strongest correlate of recent Pap test use. Older age and married status were also found to be significant predictors of Pap test use. Findings suggest cultural factors should be considered in interventions promoting cervical cancer screening among Korean American women. Furthermore, younger Korean American women and those not living with a spouse/partner should be targeted in cervical cancer screening efforts.
Modeling brook trout presence and absence from landscape variables using four different analytical methods

USGS Publications Warehouse

Steen, Paul J.; Passino-Reader, Dora R.; Wiley, Michael J.

2006-01-01

As a part of the Great Lakes Regional Aquatic Gap Analysis Project, we evaluated methodologies for modeling associations between fish species and habitat characteristics at a landscape scale. To do this, we created brook trout Salvelinus fontinalis presence and absence models based on four different techniques: multiple linear regression, logistic regression, neural networks, and classification trees. The models were tested in two ways: by application to an independent validation database and cross-validation using the training data, and by visual comparison of statewide distribution maps with historically recorded occurrences from the Michigan Fish Atlas. Although differences in the accuracy of our models were slight, the logistic regression model predicted with the least error, followed by multiple regression, then classification trees, then the neural networks. These models will provide natural resource managers a way to identify habitats requiring protection for the conservation of fish species.
Building a Decision Support System for Inpatient Admission Prediction With the Manchester Triage System and Administrative Check-in Variables.

PubMed

Zlotnik, Alexander; Alfaro, Miguel Cuchí; Pérez, María Carmen Pérez; Gallardo-Antolín, Ascensión; Martínez, Juan Manuel Montero

2016-05-01

The usage of decision support tools in emergency departments, based on predictive models, capable of estimating the probability of admission for patients in the emergency department may give nursing staff the possibility of allocating resources in advance. We present a methodology for developing and building one such system for a large specialized care hospital using a logistic regression and an artificial neural network model using nine routinely collected variables available right at the end of the triage process.A database of 255.668 triaged nonobstetric emergency department presentations from the Ramon y Cajal University Hospital of Madrid, from January 2011 to December 2012, was used to develop and test the models, with 66% of the data used for derivation and 34% for validation, with an ordered nonrandom partition. On the validation dataset areas under the receiver operating characteristic curve were 0.8568 (95% confidence interval, 0.8508-0.8583) for the logistic regression model and 0.8575 (95% confidence interval, 0.8540-0. 8610) for the artificial neural network model. χ Values for Hosmer-Lemeshow fixed "deciles of risk" were 65.32 for the logistic regression model and 17.28 for the artificial neural network model. A nomogram was generated upon the logistic regression model and an automated software decision support system with a Web interface was built based on the artificial neural network model.
Detecting a Gender-Related DIF Using Logistic Regression and Transformed Item Difficulty

ERIC Educational Resources Information Center

Abedlaziz, Nabeel; Ismail, Wail; Hussin, Zaharah

2011-01-01

Test items are designed to provide information about the examinees. Difficult items are designed to be more demanding and easy items are less so. However, sometimes, test items carry with their demands other than those intended by the test developer (Scheuneman & Gerritz, 1990). When personal attributes such as gender systematically affect…
Parameters Estimation of Geographically Weighted Ordinal Logistic Regression (GWOLR) Model

NASA Astrophysics Data System (ADS)

Zuhdi, Shaifudin; Retno Sari Saputro, Dewi; Widyaningsih, Purnami

2017-06-01

A regression model is the representation of relationship between independent variable and dependent variable. The dependent variable has categories used in the logistic regression model to calculate odds on. The logistic regression model for dependent variable has levels in the logistics regression model is ordinal. GWOLR model is an ordinal logistic regression model influenced the geographical location of the observation site. Parameters estimation in the model needed to determine the value of a population based on sample. The purpose of this research is to parameters estimation of GWOLR model using R software. Parameter estimation uses the data amount of dengue fever patients in Semarang City. Observation units used are 144 villages in Semarang City. The results of research get GWOLR model locally for each village and to know probability of number dengue fever patient categories.
PARAMETRIC AND NON PARAMETRIC (MARS: MULTIVARIATE ADDITIVE REGRESSION SPLINES) LOGISTIC REGRESSIONS FOR PREDICTION OF A DICHOTOMOUS RESPONSE VARIABLE WITH AN EXAMPLE FOR PRESENCE/ABSENCE OF AMPHIBIANS

EPA Science Inventory

The purpose of this report is to provide a reference manual that could be used by investigators for making informed use of logistic regression using two methods (standard logistic regression and MARS). The details for analyses of relationships between a dependent binary response ...
Predicting U.S. Army Reserve Unit Manning Using Market Demographics

DTIC Science & Technology

2015-06-01

develops linear regression , classification tree, and logistic regression models to determine the ability of the location to support manning requirements... logistic regression model delivers predictive results that allow decision-makers to identify locations with a high probability of meeting unit...manning requirements. The recommendation of this thesis is that the USAR implement the logistic regression model. 14. SUBJECT TERMS U.S
Analyzing Student Learning Outcomes: Usefulness of Logistic and Cox Regression Models. IR Applications, Volume 5

ERIC Educational Resources Information Center

Chen, Chau-Kuang

2005-01-01

Logistic and Cox regression methods are practical tools used to model the relationships between certain student learning outcomes and their relevant explanatory variables. The logistic regression model fits an S-shaped curve into a binary outcome with data points of zero and one. The Cox regression model allows investigators to study the duration…

An appraisal of convergence failures in the application of logistic regression model in published manuscripts.

PubMed

Yusuf, O B; Bamgboye, E A; Afolabi, R F; Shodimu, M A

2014-09-01

Logistic regression model is widely used in health research for description and predictive purposes. Unfortunately, most researchers are sometimes not aware that the underlying principles of the techniques have failed when the algorithm for maximum likelihood does not converge. Young researchers particularly postgraduate students may not know why separation problem whether quasi or complete occurs, how to identify it and how to fix it. This study was designed to critically evaluate convergence issues in articles that employed logistic regression analysis published in an African Journal of Medicine and medical sciences between 2004 and 2013. Problems of quasi or complete separation were described and were illustrated with the National Demographic and Health Survey dataset. A critical evaluation of articles that employed logistic regression was conducted. A total of 581 articles was reviewed, of which 40 (6.9%) used binary logistic regression. Twenty-four (60.0%) stated the use of logistic regression model in the methodology while none of the articles assessed model fit. Only 3 (12.5%) properly described the procedures. Of the 40 that used the logistic regression model, the problem of convergence occurred in 6 (15.0%) of the articles. Logistic regression tends to be poorly reported in studies published between 2004 and 2013. Our findings showed that the procedure may not be well understood by researchers since very few described the process in their reports and may be totally unaware of the problem of convergence or how to deal with it.
Determination of osteoporosis risk factors using a multiple logistic regression model in postmenopausal Turkish women.

PubMed

Akkus, Zeki; Camdeviren, Handan; Celik, Fatma; Gur, Ali; Nas, Kemal

2005-09-01

To determine the risk factors of osteoporosis using a multiple binary logistic regression method and to assess the risk variables for osteoporosis, which is a major and growing health problem in many countries. We presented a case-control study, consisting of 126 postmenopausal healthy women as control group and 225 postmenopausal osteoporotic women as the case group. The study was carried out in the Department of Physical Medicine and Rehabilitation, Dicle University, Diyarbakir, Turkey between 1999-2002. The data from the 351 participants were collected using a standard questionnaire that contains 43 variables. A multiple logistic regression model was then used to evaluate the data and to find the best regression model. We classified 80.1% (281/351) of the participants using the regression model. Furthermore, the specificity value of the model was 67% (84/126) of the control group while the sensitivity value was 88% (197/225) of the case group. We found the distribution of residual values standardized for final model to be exponential using the Kolmogorow-Smirnow test (p=0.193). The receiver operating characteristic curve was found successful to predict patients with risk for osteoporosis. This study suggests that low levels of dietary calcium intake, physical activity, education, and longer duration of menopause are independent predictors of the risk of low bone density in our population. Adequate dietary calcium intake in combination with maintaining a daily physical activity, increasing educational level, decreasing birth rate, and duration of breast-feeding may contribute to healthy bones and play a role in practical prevention of osteoporosis in Southeast Anatolia. In addition, the findings of the present study indicate that the use of multivariate statistical method as a multiple logistic regression in osteoporosis, which maybe influenced by many variables, is better than univariate statistical evaluation.
Classification and regression tree analysis of acute-on-chronic hepatitis B liver failure: Seeing the forest for the trees.

PubMed

Shi, K-Q; Zhou, Y-Y; Yan, H-D; Li, H; Wu, F-L; Xie, Y-Y; Braddock, M; Lin, X-Y; Zheng, M-H

2017-02-01

At present, there is no ideal model for predicting the short-term outcome of patients with acute-on-chronic hepatitis B liver failure (ACHBLF). This study aimed to establish and validate a prognostic model by using the classification and regression tree (CART) analysis. A total of 1047 patients from two separate medical centres with suspected ACHBLF were screened in the study, which were recognized as derivation cohort and validation cohort, respectively. CART analysis was applied to predict the 3-month mortality of patients with ACHBLF. The accuracy of the CART model was tested using the area under the receiver operating characteristic curve, which was compared with the model for end-stage liver disease (MELD) score and a new logistic regression model. CART analysis identified four variables as prognostic factors of ACHBLF: total bilirubin, age, serum sodium and INR, and three distinct risk groups: low risk (4.2%), intermediate risk (30.2%-53.2%) and high risk (81.4%-96.9%). The new logistic regression model was constructed with four independent factors, including age, total bilirubin, serum sodium and prothrombin activity by multivariate logistic regression analysis. The performances of the CART model (0.896), similar to the logistic regression model (0.914, P=.382), exceeded that of MELD score (0.667, P<.001). The results were confirmed in the validation cohort. We have developed and validated a novel CART model superior to MELD for predicting three-month mortality of patients with ACHBLF. Thus, the CART model could facilitate medical decision-making and provide clinicians with a validated practical bedside tool for ACHBLF risk stratification. © 2016 John Wiley & Sons Ltd.
Artificial Neural Network for the Prediction of Chromosomal Abnormalities in Azoospermic Males.

PubMed

Akinsal, Emre Can; Haznedar, Bulent; Baydilli, Numan; Kalinli, Adem; Ozturk, Ahmet; Ekmekçioğlu, Oğuz

2018-02-04

To evaluate whether an artifical neural network helps to diagnose any chromosomal abnormalities in azoospermic males. The data of azoospermic males attending to a tertiary academic referral center were evaluated retrospectively. Height, total testicular volume, follicle stimulating hormone, luteinising hormone, total testosterone and ejaculate volume of the patients were used for the analyses. In artificial neural network, the data of 310 azoospermics were used as the education and 115 as the test set. Logistic regression analyses and discriminant analyses were performed for statistical analyses. The tests were re-analysed with a neural network. Both logistic regression analyses and artificial neural network predicted the presence or absence of chromosomal abnormalities with more than 95% accuracy. The use of artificial neural network model has yielded satisfactory results in terms of distinguishing patients whether they have any chromosomal abnormality or not.
Organizational Justice and Physiological Coronary Heart Disease Risk Factors in Japanese Employees: a Cross-Sectional Study.

PubMed

Inoue, Akiomi; Kawakami, Norito; Eguchi, Hisashi; Miyaki, Koichi; Tsutsumi, Akizumi

2015-12-01

Growing evidence has shown that lack of organizational justice (i.e., procedural justice and interactional justice) is associated with coronary heart disease (CHD) while biological mechanisms underlying this association have not yet been fully clarified. The purpose of the present study was to investigate the cross-sectional association of organizational justice with physiological CHD risk factors (i.e., blood pressure, high-density lipoprotein [HDL] cholesterol, low-density lipoprotein [LDL] cholesterol, and triglyceride) in Japanese employees. Overall, 3598 male and 901 female employees from two manufacturing companies in Japan completed self-administered questionnaires measuring organizational justice, demographic characteristics, and lifestyle factors. They completed health checkup, which included blood pressure and serum lipid measurements. Multiple logistic regression analyses and trend tests were conducted. Among male employees, multiple logistic regression analyses and trend tests showed significant associations of low procedural justice and low interactional justice with high triglyceride (defined as 150 mg/dL or greater) after adjusting for demographic characteristics and lifestyle factors. Among female employees, trend tests showed significant dose-response relationship between low interactional justice and high LDL cholesterol (defined as 140 mg/dL or greater) while multiple logistic regression analysis showed only marginally significant or insignificant odds ratio of high LDL cholesterol among the low interactional justice group. Neither procedural justice nor interactional justice was associated with blood pressure or HDL cholesterol. Organizational justice may be an important psychosocial factor associated with increased triglyceride at least among Japanese male employees.
Validation of use of the International Consultation on Incontinence Questionnaire-Urinary Incontinence-Short Form (ICIQ-UI-SF) for impairment rating: a transversal retrospective study of 120 patients.

PubMed

Timmermans, Luc; Falez, Freddy; Mélot, Christian; Wespes, Eric

2013-09-01

A urinary incontinence impairment rating must be a highly accurate, non-invasive exploration of the condition using International Classification of Functioning (ICF)-based assessment tools. The objective of this study was to identify the best evaluation test and to determine an impairment rating model of urinary incontinence. In performing a cross-sectional study comparing successive urodynamic tests using both the International Consultation on Incontinence Questionnaire-Urinary Incontinence-Short Form (ICIQ-UI-SF) and the 1-hr pad-weighing test in 120 patients, we performed statistical likelihood ratio analysis and used logistic regression to calculate the probability of urodynamic incontinence using the most significant independent predictors. Subsequently, we created a template that was based on the significant predictors and the probability of urodynamic incontinence. The mean ICIQ-UI-SF score was 13.5 ± 4.6, and the median pad test value was 8 g. The discrimination statistic (receiver operating characteristic) described how well the urodynamic observations matched the ICIQ-UI-SF scores (under curve area (UDA):0.689) and the pad test data (UDA: 0.693). Using logistic regression analysis, we demonstrated that the best independent predictors of urodynamic incontinence were the patient's age and the ICIQ-UI-SF score. The logistic regression model permitted us to construct an equation to determine the probability of urodynamic incontinence. Using these tools, we created a template to generate a probability index of urodynamic urinary incontinence. Using this probability index, relative to the patient and to the maximum impairment of the whole person (MIWP) relative to urinary incontinence, we were able to calculate a patient's permanent impairment. Copyright © 2012 Wiley Periodicals, Inc.
Logistic Regression: Concept and Application

ERIC Educational Resources Information Center

Cokluk, Omay

2010-01-01

The main focus of logistic regression analysis is classification of individuals in different groups. The aim of the present study is to explain basic concepts and processes of binary logistic regression analysis intended to determine the combination of independent variables which best explain the membership in certain groups called dichotomous…
Regression analysis for solving diagnosis problem of children's health

NASA Astrophysics Data System (ADS)

Cherkashina, Yu A.; Gerget, O. M.

2016-04-01

The paper includes results of scientific researches. These researches are devoted to the application of statistical techniques, namely, regression analysis, to assess the health status of children in the neonatal period based on medical data (hemostatic parameters, parameters of blood tests, the gestational age, vascular-endothelial growth factor) measured at 3-5 days of children's life. In this paper a detailed description of the studied medical data is given. A binary logistic regression procedure is discussed in the paper. Basic results of the research are presented. A classification table of predicted values and factual observed values is shown, the overall percentage of correct recognition is determined. Regression equation coefficients are calculated, the general regression equation is written based on them. Based on the results of logistic regression, ROC analysis was performed, sensitivity and specificity of the model are calculated and ROC curves are constructed. These mathematical techniques allow carrying out diagnostics of health of children providing a high quality of recognition. The results make a significant contribution to the development of evidence-based medicine and have a high practical importance in the professional activity of the author.
An Entropy-Based Measure for Assessing Fuzziness in Logistic Regression

PubMed Central

Weiss, Brandi A.; Dardick, William

2015-01-01

This article introduces an entropy-based measure of data–model fit that can be used to assess the quality of logistic regression models. Entropy has previously been used in mixture-modeling to quantify how well individuals are classified into latent classes. The current study proposes the use of entropy for logistic regression models to quantify the quality of classification and separation of group membership. Entropy complements preexisting measures of data–model fit and provides unique information not contained in other measures. Hypothetical data scenarios, an applied example, and Monte Carlo simulation results are used to demonstrate the application of entropy in logistic regression. Entropy should be used in conjunction with other measures of data–model fit to assess how well logistic regression models classify cases into observed categories. PMID:29795897
Logistic regression applied to natural hazards: rare event logistic regression with replications

NASA Astrophysics Data System (ADS)

Guns, M.; Vanacker, V.

2012-06-01

Statistical analysis of natural hazards needs particular attention, as most of these phenomena are rare events. This study shows that the ordinary rare event logistic regression, as it is now commonly used in geomorphologic studies, does not always lead to a robust detection of controlling factors, as the results can be strongly sample-dependent. In this paper, we introduce some concepts of Monte Carlo simulations in rare event logistic regression. This technique, so-called rare event logistic regression with replications, combines the strength of probabilistic and statistical methods, and allows overcoming some of the limitations of previous developments through robust variable selection. This technique was here developed for the analyses of landslide controlling factors, but the concept is widely applicable for statistical analyses of natural hazards.
Large unbalanced credit scoring using Lasso-logistic regression ensemble.

PubMed

Wang, Hong; Xu, Qingsong; Zhou, Lifeng

2015-01-01

Recently, various ensemble learning methods with different base classifiers have been proposed for credit scoring problems. However, for various reasons, there has been little research using logistic regression as the base classifier. In this paper, given large unbalanced data, we consider the plausibility of ensemble learning using regularized logistic regression as the base classifier to deal with credit scoring problems. In this research, the data is first balanced and diversified by clustering and bagging algorithms. Then we apply a Lasso-logistic regression learning ensemble to evaluate the credit risks. We show that the proposed algorithm outperforms popular credit scoring models such as decision tree, Lasso-logistic regression and random forests in terms of AUC and F-measure. We also provide two importance measures for the proposed model to identify important variables in the data.
An Entropy-Based Measure for Assessing Fuzziness in Logistic Regression.

PubMed

Weiss, Brandi A; Dardick, William

2016-12-01

This article introduces an entropy-based measure of data-model fit that can be used to assess the quality of logistic regression models. Entropy has previously been used in mixture-modeling to quantify how well individuals are classified into latent classes. The current study proposes the use of entropy for logistic regression models to quantify the quality of classification and separation of group membership. Entropy complements preexisting measures of data-model fit and provides unique information not contained in other measures. Hypothetical data scenarios, an applied example, and Monte Carlo simulation results are used to demonstrate the application of entropy in logistic regression. Entropy should be used in conjunction with other measures of data-model fit to assess how well logistic regression models classify cases into observed categories.
New machine-learning algorithms for prediction of Parkinson's disease

NASA Astrophysics Data System (ADS)

Mandal, Indrajit; Sairam, N.

2014-03-01

This article presents an enhanced prediction accuracy of diagnosis of Parkinson's disease (PD) to prevent the delay and misdiagnosis of patients using the proposed robust inference system. New machine-learning methods are proposed and performance comparisons are based on specificity, sensitivity, accuracy and other measurable parameters. The robust methods of treating Parkinson's disease (PD) includes sparse multinomial logistic regression, rotation forest ensemble with support vector machines and principal components analysis, artificial neural networks, boosting methods. A new ensemble method comprising of the Bayesian network optimised by Tabu search algorithm as classifier and Haar wavelets as projection filter is used for relevant feature selection and ranking. The highest accuracy obtained by linear logistic regression and sparse multinomial logistic regression is 100% and sensitivity, specificity of 0.983 and 0.996, respectively. All the experiments are conducted over 95% and 99% confidence levels and establish the results with corrected t-tests. This work shows a high degree of advancement in software reliability and quality of the computer-aided diagnosis system and experimentally shows best results with supportive statistical inference.
Landslide Hazard Mapping in Rwanda Using Logistic Regression

NASA Astrophysics Data System (ADS)

Piller, A.; Anderson, E.; Ballard, H.

2015-12-01

Landslides in the United States cause more than $1 billion in damages and 50 deaths per year (USGS 2014). Globally, figures are much more grave, yet monitoring, mapping and forecasting of these hazards are less than adequate. Seventy-five percent of the population of Rwanda earns a living from farming, mostly subsistence. Loss of farmland, housing, or life, to landslides is a very real hazard. Landslides in Rwanda have an impact at the economic, social, and environmental level. In a developing nation that faces challenges in tracking, cataloging, and predicting the numerous landslides that occur each year, satellite imagery and spatial analysis allow for remote study. We have focused on the development of a landslide inventory and a statistical methodology for assessing landslide hazards. Using logistic regression on approximately 30 test variables (i.e. slope, soil type, land cover, etc.) and a sample of over 200 landslides, we determine which variables are statistically most relevant to landslide occurrence in Rwanda. A preliminary predictive hazard map for Rwanda has been produced, using the variables selected from the logistic regression analysis.
A logistic regression analysis of factors related to the treatment compliance of infertile patients with polycystic ovary syndrome.

PubMed

Li, Saijiao; He, Aiyan; Yang, Jing; Yin, TaiLang; Xu, Wangming

2011-01-01

To investigate factors that can affect compliance with treatment of polycystic ovary syndrome (PCOS) in infertile patients and to provide a basis for clinical treatment, specialist consultation and health education. Patient compliance was assessed via a questionnaire based on the Morisky-Green test and the treatment principles of PCOS. Then interviews were conducted with 99 infertile patients diagnosed with PCOS at Renmin Hospital of Wuhan University in China, from March to September 2009. Finally, these data were analyzed using logistic regression analysis. Logistic regression analysis revealed that a total of 23 (25.6%) of the participants showed good compliance. Factors that significantly (p < 0.05) affected compliance with treatment were the patient's body mass index, convenience of medical treatment and concerns about adverse drug reactions. Patients who are obese, experience inconvenient medical treatment or are concerned about adverse drug reactions are more likely to exhibit noncompliance. Treatment education and intervention aimed at these patients should be strengthened in the clinic to improve treatment compliance. Further research is needed to better elucidate the compliance behavior of patients with PCOS.
A Methodology for Generating Placement Rules that Utilizes Logistic Regression

ERIC Educational Resources Information Center

Wurtz, Keith

2008-01-01

The purpose of this article is to provide the necessary tools for institutional researchers to conduct a logistic regression analysis and interpret the results. Aspects of the logistic regression procedure that are necessary to evaluate models are presented and discussed with an emphasis on cutoff values and choosing the appropriate number of…
Comparison of standard maximum likelihood classification and polytomous logistic regression used in remote sensing

Treesearch

John Hogland; Nedret Billor; Nathaniel Anderson

2013-01-01

Discriminant analysis, referred to as maximum likelihood classification within popular remote sensing software packages, is a common supervised technique used by analysts. Polytomous logistic regression (PLR), also referred to as multinomial logistic regression, is an alternative classification approach that is less restrictive, more flexible, and easy to interpret. To...
Patient acceptance of non-invasive testing for fetal aneuploidy via cell-free fetal DNA.

PubMed

Vahanian, Sevan A; Baraa Allaf, M; Yeh, Corinne; Chavez, Martin R; Kinzler, Wendy L; Vintzileos, Anthony M

2014-01-01

To evaluate factors associated with patient acceptance of noninvasive prenatal testing for trisomy 21, 18 and 13 via cell-free fetal DNA. This was a retrospective study of all patients who were offered noninvasive prenatal testing at a single institution from 1 March 2012 to 2 July 2012. Patients were identified through our perinatal ultrasound database; demographic information, testing indication and insurance coverage were compared between patients who accepted the test and those who declined. Parametric and nonparametric tests were used as appropriate. Significant variables were assessed using multivariate logistic regression. The value p < 0.05 was considered significant. Two hundred thirty-five patients were offered noninvasive prenatal testing. Ninety-three patients (40%) accepted testing and 142 (60%) declined. Women who accepted noninvasive prenatal testing were more commonly white, had private insurance and had more than one testing indication. There was no statistical difference in the number or the type of testing indications. Multivariable logistic regression analysis was then used to assess individual variables. After controlling for race, patients with public insurance were 83% less likely to accept noninvasive prenatal testing than those with private insurance (3% vs. 97%, adjusted RR 0.17, 95% CI 0.05-0.62). In our population, having public insurance was the factor most strongly associated with declining noninvasive prenatal testing.
Large Unbalanced Credit Scoring Using Lasso-Logistic Regression Ensemble

PubMed Central

Wang, Hong; Xu, Qingsong; Zhou, Lifeng

2015-01-01

Recently, various ensemble learning methods with different base classifiers have been proposed for credit scoring problems. However, for various reasons, there has been little research using logistic regression as the base classifier. In this paper, given large unbalanced data, we consider the plausibility of ensemble learning using regularized logistic regression as the base classifier to deal with credit scoring problems. In this research, the data is first balanced and diversified by clustering and bagging algorithms. Then we apply a Lasso-logistic regression learning ensemble to evaluate the credit risks. We show that the proposed algorithm outperforms popular credit scoring models such as decision tree, Lasso-logistic regression and random forests in terms of AUC and F-measure. We also provide two importance measures for the proposed model to identify important variables in the data. PMID:25706988
Association Between Socio-Demographic Background and Self-Esteem of University Students.

PubMed

Haq, Muhammad Ahsan Ul

2016-12-01

The purpose of this study was to scrutinize self-esteem of university students and explore association of self-esteem with academic achievement, gender and other factors. A sample of 346 students was selected from Punjab University, Lahore Pakistan. Rosenberg self-esteem scale with demographic variables was used for data collection. Besides descriptive statistics, binary logistic regression and t test were used for analysing the data. Significant gender difference was observed, self-esteem was significantly higher in males than females. Logistic regression indicates that age, medium of instruction, family income, student monthly expenditures, GPA and area of residence has direct effect on self-esteem; while number of siblings showed an inverse effect.

Higher direct bilirubin levels during mid-pregnancy are associated with lower risk of gestational diabetes mellitus.

PubMed

Liu, Chaoqun; Zhong, Chunrong; Zhou, Xuezhen; Chen, Renjuan; Wu, Jiangyue; Wang, Weiye; Li, Xiating; Ding, Huisi; Guo, Yanfang; Gao, Qin; Hu, Xingwen; Xiong, Guoping; Yang, Xuefeng; Hao, Liping; Xiao, Mei; Yang, Nianhong

2017-01-01

Bilirubin concentrations have been recently reported to be negatively associated with type 2 diabetes mellitus. We examined the association between bilirubin concentrations and gestational diabetes mellitus. In a prospective cohort study, 2969 pregnant women were recruited prior to 16 weeks of gestation and were followed up until delivery. The value of bilirubin was tested and oral glucose tolerance test was conducted to screen gestational diabetes mellitus. The relationship between serum bilirubin concentration and gestational weeks was studied by two-piecewise linear regression. A subsample of 1135 participants with serum bilirubin test during 16-18 weeks gestation was conducted to research the association between serum bilirubin levels and risk of gestational diabetes mellitus by logistic regression. Gestational diabetes mellitus developed in 8.5 % of the participants (223 of 2969). Two-piecewise linear regression analyses demonstrated that the levels of bilirubin decreased with gestational week up to the turning point 23 and after that point, levels of bilirubin were increased slightly. In multiple logistic regression analysis, the relative risk of developing gestational diabetes mellitus was lower in the highest tertile of direct bilirubin than that in the lowest tertile (RR 0.60; 95 % CI, 0.35-0.89). The results suggested that women with higher serum direct bilirubin levels during the second trimester of pregnancy have lower risk for development of gestational diabetes mellitus.
Accounting for informatively missing data in logistic regression by means of reassessment sampling.

PubMed

Lin, Ji; Lyles, Robert H

2015-05-20

We explore the 'reassessment' design in a logistic regression setting, where a second wave of sampling is applied to recover a portion of the missing data on a binary exposure and/or outcome variable. We construct a joint likelihood function based on the original model of interest and a model for the missing data mechanism, with emphasis on non-ignorable missingness. The estimation is carried out by numerical maximization of the joint likelihood function with close approximation of the accompanying Hessian matrix, using sharable programs that take advantage of general optimization routines in standard software. We show how likelihood ratio tests can be used for model selection and how they facilitate direct hypothesis testing for whether missingness is at random. Examples and simulations are presented to demonstrate the performance of the proposed method. Copyright © 2015 John Wiley & Sons, Ltd.
Testing a model of research intention among U.K. clinical psychologists: a logistic regression analysis.

PubMed

Eke, Gemma; Holttum, Sue; Hayward, Mark

2012-03-01

Previous research highlights barriers to clinical psychologists conducting research, but has rarely examined U.K. clinical psychologists. The study investigated U.K. clinical psychologists' self-reported research output and tested part of a theoretical model of factors influencing their intention to conduct research. Questionnaires were mailed to 1,300 U.K. clinical psychologists. Three hundred and seventy-four questionnaires were returned (29% response-rate). This study replicated in a U.K. sample the finding that the modal number of publications was zero, highlighted in a number of U.K. and U.S. studies. Research intention was bimodally distributed, and logistic regression classified 78% of cases successfully. Outcome expectations, perceived behavioral control and normative beliefs mediated between research training environment and intention. Further research should explore how research is negotiated in clinical roles, and this issue should be incorporated into prequalification training. © 2012 Wiley Periodicals, Inc.
Individual and community risk factors and sexually transmitted diseases among arrested youths: a two level analysis.

PubMed

Dembo, Richard; Belenko, Steven; Childs, Kristina; Wareham, Jennifer; Schmeidler, James

2009-08-01

High rates of infection for chlamydia and gonorrhea have been noted among youths involved in the juvenile justice system. Although both individual and community-level factors have been found to be associated with sexually transmitted disease (STD) risk, their relative importance has not been tested in this population. A two-level logistic regression analysis was completed to assess the influence of individual-level and community-level predictors on STD test results among arrested youths processed at a centralized intake facility. Results from weighted two level logistic regression analyses (n = 1,368) indicated individual-level factors of gender (being female), age, race (being African American), and criminal history predicted the youths' positive STD status. For the community-level predictors, concentrated disadvantage significantly and positively predicted the youths' STD status. Implications of these findings for future research and public health policy are discussed.
An Entropy-Based Measure for Assessing Fuzziness in Logistic Regression

ERIC Educational Resources Information Center

Weiss, Brandi A.; Dardick, William

2016-01-01

This article introduces an entropy-based measure of data-model fit that can be used to assess the quality of logistic regression models. Entropy has previously been used in mixture-modeling to quantify how well individuals are classified into latent classes. The current study proposes the use of entropy for logistic regression models to quantify…
What Are the Odds of that? A Primer on Understanding Logistic Regression

ERIC Educational Resources Information Center

Huang, Francis L.; Moon, Tonya R.

2013-01-01

The purpose of this Methodological Brief is to present a brief primer on logistic regression, a commonly used technique when modeling dichotomous outcomes. Using data from the National Education Longitudinal Study of 1988 (NELS:88), logistic regression techniques were used to investigate student-level variables in eighth grade (i.e., enrolled in a…
On the Usefulness of a Multilevel Logistic Regression Approach to Person-Fit Analysis

ERIC Educational Resources Information Center

Conijn, Judith M.; Emons, Wilco H. M.; van Assen, Marcel A. L. M.; Sijtsma, Klaas

2011-01-01

The logistic person response function (PRF) models the probability of a correct response as a function of the item locations. Reise (2000) proposed to use the slope parameter of the logistic PRF as a person-fit measure. He reformulated the logistic PRF model as a multilevel logistic regression model and estimated the PRF parameters from this…
Determination of riverbank erosion probability using Locally Weighted Logistic Regression

NASA Astrophysics Data System (ADS)

Ioannidou, Elena; Flori, Aikaterini; Varouchakis, Emmanouil A.; Giannakis, Georgios; Vozinaki, Anthi Eirini K.; Karatzas, George P.; Nikolaidis, Nikolaos

2015-04-01

Riverbank erosion is a natural geomorphologic process that affects the fluvial environment. The most important issue concerning riverbank erosion is the identification of the vulnerable locations. An alternative to the usual hydrodynamic models to predict vulnerable locations is to quantify the probability of erosion occurrence. This can be achieved by identifying the underlying relations between riverbank erosion and the geomorphological or hydrological variables that prevent or stimulate erosion. Thus, riverbank erosion can be determined by a regression model using independent variables that are considered to affect the erosion process. The impact of such variables may vary spatially, therefore, a non-stationary regression model is preferred instead of a stationary equivalent. Locally Weighted Regression (LWR) is proposed as a suitable choice. This method can be extended to predict the binary presence or absence of erosion based on a series of independent local variables by using the logistic regression model. It is referred to as Locally Weighted Logistic Regression (LWLR). Logistic regression is a type of regression analysis used for predicting the outcome of a categorical dependent variable (e.g. binary response) based on one or more predictor variables. The method can be combined with LWR to assign weights to local independent variables of the dependent one. LWR allows model parameters to vary over space in order to reflect spatial heterogeneity. The probabilities of the possible outcomes are modelled as a function of the independent variables using a logistic function. Logistic regression measures the relationship between a categorical dependent variable and, usually, one or several continuous independent variables by converting the dependent variable to probability scores. Then, a logistic regression is formed, which predicts success or failure of a given binary variable (e.g. erosion presence or absence) for any value of the independent variables. The erosion occurrence probability can be calculated in conjunction with the model deviance regarding the independent variables tested. The most straightforward measure for goodness of fit is the G statistic. It is a simple and effective way to study and evaluate the Logistic Regression model efficiency and the reliability of each independent variable. The developed statistical model is applied to the Koiliaris River Basin on the island of Crete, Greece. Two datasets of river bank slope, river cross-section width and indications of erosion were available for the analysis (12 and 8 locations). Two different types of spatial dependence functions, exponential and tricubic, were examined to determine the local spatial dependence of the independent variables at the measurement locations. The results show a significant improvement when the tricubic function is applied as the erosion probability is accurately predicted at all eight validation locations. Results for the model deviance show that cross-section width is more important than bank slope in the estimation of erosion probability along the Koiliaris riverbanks. The proposed statistical model is a useful tool that quantifies the erosion probability along the riverbanks and can be used to assist managing erosion and flooding events. Acknowledgements This work is part of an on-going THALES project (CYBERSENSORS - High Frequency Monitoring System for Integrated Water Resources Management of Rivers). The project has been co-financed by the European Union (European Social Fund - ESF) and Greek national funds through the Operational Program "Education and Lifelong Learning" of the National Strategic Reference Framework (NSRF) - Research Funding Program: THALES. Investing in knowledge society through the European Social Fund.
Mortality risk prediction in burn injury: Comparison of logistic regression with machine learning approaches.

PubMed

Stylianou, Neophytos; Akbarov, Artur; Kontopantelis, Evangelos; Buchan, Iain; Dunn, Ken W

2015-08-01

Predicting mortality from burn injury has traditionally employed logistic regression models. Alternative machine learning methods have been introduced in some areas of clinical prediction as the necessary software and computational facilities have become accessible. Here we compare logistic regression and machine learning predictions of mortality from burn. An established logistic mortality model was compared to machine learning methods (artificial neural network, support vector machine, random forests and naïve Bayes) using a population-based (England & Wales) case-cohort registry. Predictive evaluation used: area under the receiver operating characteristic curve; sensitivity; specificity; positive predictive value and Youden's index. All methods had comparable discriminatory abilities, similar sensitivities, specificities and positive predictive values. Although some machine learning methods performed marginally better than logistic regression the differences were seldom statistically significant and clinically insubstantial. Random forests were marginally better for high positive predictive value and reasonable sensitivity. Neural networks yielded slightly better prediction overall. Logistic regression gives an optimal mix of performance and interpretability. The established logistic regression model of burn mortality performs well against more complex alternatives. Clinical prediction with a small set of strong, stable, independent predictors is unlikely to gain much from machine learning outside specialist research contexts. Copyright © 2015 Elsevier Ltd and ISBI. All rights reserved.
Assessing risk factors for periodontitis using regression

NASA Astrophysics Data System (ADS)

Lobo Pereira, J. A.; Ferreira, Maria Cristina; Oliveira, Teresa

2013-10-01

Multivariate statistical analysis is indispensable to assess the associations and interactions between different factors and the risk of periodontitis. Among others, regression analysis is a statistical technique widely used in healthcare to investigate and model the relationship between variables. In our work we study the impact of socio-demographic, medical and behavioral factors on periodontal health. Using regression, linear and logistic models, we can assess the relevance, as risk factors for periodontitis disease, of the following independent variables (IVs): Age, Gender, Diabetic Status, Education, Smoking status and Plaque Index. The multiple linear regression analysis model was built to evaluate the influence of IVs on mean Attachment Loss (AL). Thus, the regression coefficients along with respective p-values will be obtained as well as the respective p-values from the significance tests. The classification of a case (individual) adopted in the logistic model was the extent of the destruction of periodontal tissues defined by an Attachment Loss greater than or equal to 4 mm in 25% (AL≥4mm/≥25%) of sites surveyed. The association measures include the Odds Ratios together with the correspondent 95% confidence intervals.
Predictors of Assessment Accommodations Use for Students Who Are Deaf or Hard of Hearing

ERIC Educational Resources Information Center

Cawthon, Stephanie W.; Wurtz, Keith A.

2010-01-01

Current accountability reform requires annual assessment for all students, including students with disabilities. Testing accommodations are one way to increase access to assessments while maintaining the validity of test scores. This paper provides findings from an exploratory logistic regression analysis of predictors of four accommodations used…
Prevalence of abortion and stillbirth in a beef cattle system in Southeastern Mexico.

PubMed

Segura-Correa, José C; Segura-Correa, Victor M

2009-12-01

Prenatal mortality is an important cause of production losses in the livestock industry. This study estimates the prevalences of abortion and stillbirth in a beef cattle system and determines the significance of some risk factors, in the tropics of Mexico. Data were obtained from a Zebu cattle herd and their crosses with Bos taurus breeds, in Yucatan, Mexico. The logit of the probability of an abortion or stillbirth was modeled using binary logistic regression. The risk factors tested were: year of abortion (or calving), season of abortion (or calving), parity number and dam breed group. The effect of twins on stillbirth was tested using Fisher exact test. Of the 4175 calvings studied 49 were abortions (1.17%). Significant factors in the logistic regression analysis for abortions were season of abortion and parity number. The risk of abortion was lower in the dry seasons compared to the rainy and windy seasons (P = 0.009). The risk of abortion was higher in second parity cows followed by the third and first parity cows, as compared to older cows (P = 0.015). Of the 4126 births, 87 were stillbirths (2.11%). Significant factors in the logistic regression analysis for stillbirth were year of calving (P = 0.0001) and parity number (P < 0.001). The risk of stillbirth in first parity cows was 2.6 times that of old cows. Of the total births, 15 were twins (0.36%) of which 7 were born dead calves. Herd owners must focus on the significant risk factors under their control to reduce the prevalence of prenatal mortality.
Knowledge, attitudes and practices survey on organ donation among a selected adult population of Pakistan

PubMed Central

Saleem, Taimur; Ishaque, Sidra; Habib, Nida; Hussain, Syedda Saadia; Jawed, Areeba; Khan, Aamir Ali; Ahmad, Muhammad Imran; Iftikhar, Mian Omer; Mughal, Hamza Pervez; Jehan, Imtiaz

2009-01-01

Background To determine the knowledge, attitudes and practices regarding organ donation in a selected adult population in Pakistan. Methods Convenience sampling was used to generate a sample of 440; 408 interviews were successfully completed and used for analysis. Data collection was carried out via a face to face interview based on a pre-tested questionnaire in selected public areas of Karachi, Pakistan. Data was analyzed using SPSS v.15 and associations were tested using the Pearson's Chi square test. Multiple logistic regression was used to find independent predictors of knowledge status and motivation of organ donation. Results Knowledge about organ donation was significantly associated with education (p = 0.000) and socioeconomic status (p = 0.038). 70/198 (35.3%) people expressed a high motivation to donate. Allowance of organ donation in religion was significantly associated with the motivation to donate (p = 0.000). Multiple logistic regression analysis revealed that higher level of education and higher socioeconomic status were significant (p < 0.05) independent predictors of knowledge status of organ donation. For motivation, multiple logistic regression revealed that higher socioeconomic status, adequate knowledge score and belief that organ donation is allowed in religion were significant (p < 0.05) independent predictors. Television emerged as the major source of information. Only 3.5% had themselves donated an organ; with only one person being an actual kidney donor. Conclusion Better knowledge may ultimately translate into the act of donation. Effective measures should be taken to educate people with relevant information with the involvement of media, doctors and religious scholars. PMID:19534793
Attrition in Developmental Psychology: A Review of Modern Missing Data Reporting and Practices

ERIC Educational Resources Information Center

Nicholson, Jody S.; Deboeck, Pascal R.; Howard, Waylon

2017-01-01

Inherent in applied developmental sciences is the threat to validity and generalizability due to missing data as a result of participant drop-out. The current paper provides an overview of how attrition should be reported, which tests can examine the potential of bias due to attrition (e.g., t-tests, logistic regression, Little's MCAR test,…
Utility of an Abbreviated Dizziness Questionnaire to Differentiate between Causes of Vertigo and Guide Appropriate Referral: A Multicenter Prospective Blinded Study

PubMed Central

Roland, Lauren T.; Kallogjeri, Dorina; Sinks, Belinda C.; Rauch, Steven D.; Shepard, Neil T.; White, Judith A.; Goebel, Joel A.

2015-01-01

Objective Test performance of a focused dizziness questionnaire’s ability to discriminate between peripheral and non-peripheral causes of vertigo. Study Design Prospective multi-center Setting Four academic centers with experienced balance specialists Patients New dizzy patients Interventions A 32-question survey was given to participants. Balance specialists were blinded and a diagnosis was established for all participating patients within 6 months. Main outcomes Multinomial logistic regression was used to evaluate questionnaire performance in predicting final diagnosis and differentiating between peripheral and non-peripheral vertigo. Univariate and multivariable stepwise logistic regression were used to identify questions as significant predictors of the ultimate diagnosis. C-index was used to evaluate performance and discriminative power of the multivariable models. Results 437 patients participated in the study. Eight participants without confirmed diagnoses were excluded and 429 were included in the analysis. Multinomial regression revealed that the model had good overall predictive accuracy of 78.5% for the final diagnosis and 75.5% for differentiating between peripheral and non-peripheral vertigo. Univariate logistic regression identified significant predictors of three main categories of vertigo: peripheral, central and other. Predictors were entered into forward stepwise multivariable logistic regression. The discriminative power of the final models for peripheral, central and other causes were considered good as measured by c-indices of 0.75, 0.7 and 0.78, respectively. Conclusions This multicenter study demonstrates a focused dizziness questionnaire can accurately predict diagnosis for patients with chronic/relapsing dizziness referred to outpatient clinics. Additionally, this survey has significant capability to differentiate peripheral from non-peripheral causes of vertigo and may, in the future, serve as a screening tool for specialty referral. Clinical utility of this questionnaire to guide specialty referral is discussed. PMID:26485598
Utility of an Abbreviated Dizziness Questionnaire to Differentiate Between Causes of Vertigo and Guide Appropriate Referral: A Multicenter Prospective Blinded Study.

PubMed

Roland, Lauren T; Kallogjeri, Dorina; Sinks, Belinda C; Rauch, Steven D; Shepard, Neil T; White, Judith A; Goebel, Joel A

2015-12-01

Test performance of a focused dizziness questionnaire's ability to discriminate between peripheral and nonperipheral causes of vertigo. Prospective multicenter. Four academic centers with experienced balance specialists. New dizzy patients. A 32-question survey was given to participants. Balance specialists were blinded and a diagnosis was established for all participating patients within 6 months. Multinomial logistic regression was used to evaluate questionnaire performance in predicting final diagnosis and differentiating between peripheral and nonperipheral vertigo. Univariate and multivariable stepwise logistic regression were used to identify questions as significant predictors of the ultimate diagnosis. C-index was used to evaluate performance and discriminative power of the multivariable models. In total, 437 patients participated in the study. Eight participants without confirmed diagnoses were excluded and 429 were included in the analysis. Multinomial regression revealed that the model had good overall predictive accuracy of 78.5% for the final diagnosis and 75.5% for differentiating between peripheral and nonperipheral vertigo. Univariate logistic regression identified significant predictors of three main categories of vertigo: peripheral, central, and other. Predictors were entered into forward stepwise multivariable logistic regression. The discriminative power of the final models for peripheral, central, and other causes was considered good as measured by c-indices of 0.75, 0.7, and 0.78, respectively. This multicenter study demonstrates a focused dizziness questionnaire can accurately predict diagnosis for patients with chronic/relapsing dizziness referred to outpatient clinics. Additionally, this survey has significant capability to differentiate peripheral from nonperipheral causes of vertigo and may, in the future, serve as a screening tool for specialty referral. Clinical utility of this questionnaire to guide specialty referral is discussed.
SU-E-J-256: Predicting Metastasis-Free Survival of Rectal Cancer Patients Treated with Neoadjuvant Chemo-Radiotherapy by Data-Mining of CT Texture Features of Primary Lesions

DOE Office of Scientific and Technical Information (OSTI.GOV)

Zhong, H; Wang, J; Shen, L

Purpose: The purpose of this study is to investigate the relationship between computed tomographic (CT) texture features of primary lesions and metastasis-free survival for rectal cancer patients; and to develop a datamining prediction model using texture features. Methods: A total of 220 rectal cancer patients treated with neoadjuvant chemo-radiotherapy (CRT) were enrolled in this study. All patients underwent CT scans before CRT. The primary lesions on the CT images were delineated by two experienced oncologists. The CT images were filtered by Laplacian of Gaussian (LoG) filters with different filter values (1.0–2.5: from fine to coarse). Both filtered and unfiltered imagesmore » were analyzed using Gray-level Co-occurrence Matrix (GLCM) texture analysis with different directions (transversal, sagittal, and coronal). Totally, 270 texture features with different species, directions and filter values were extracted. Texture features were examined with Student’s t-test for selecting predictive features. Principal Component Analysis (PCA) was performed upon the selected features to reduce the feature collinearity. Artificial neural network (ANN) and logistic regression were applied to establish metastasis prediction models. Results: Forty-six of 220 patients developed metastasis with a follow-up time of more than 2 years. Sixtyseven texture features were significantly different in t-test (p<0.05) between patients with and without metastasis, and 12 of them were extremely significant (p<0.001). The Area-under-the-curve (AUC) of ANN was 0.72, and the concordance index (CI) of logistic regression was 0.71. The predictability of ANN was slightly better than logistic regression. Conclusion: CT texture features of primary lesions are related to metastasisfree survival of rectal cancer patients. Both ANN and logistic regression based models can be developed for prediction.« less
Logistic regression for risk factor modelling in stuttering research.

PubMed

Reed, Phil; Wu, Yaqionq

2013-06-01

To outline the uses of logistic regression and other statistical methods for risk factor analysis in the context of research on stuttering. The principles underlying the application of a logistic regression are illustrated, and the types of questions to which such a technique has been applied in the stuttering field are outlined. The assumptions and limitations of the technique are discussed with respect to existing stuttering research, and with respect to formulating appropriate research strategies to accommodate these considerations. Finally, some alternatives to the approach are briefly discussed. The way the statistical procedures are employed are demonstrated with some hypothetical data. Research into several practical issues concerning stuttering could benefit if risk factor modelling were used. Important examples are early diagnosis, prognosis (whether a child will recover or persist) and assessment of treatment outcome. After reading this article you will: (a) Summarize the situations in which logistic regression can be applied to a range of issues about stuttering; (b) Follow the steps in performing a logistic regression analysis; (c) Describe the assumptions of the logistic regression technique and the precautions that need to be checked when it is employed; (d) Be able to summarize its advantages over other techniques like estimation of group differences and simple regression. Copyright © 2012 Elsevier Inc. All rights reserved.
Dynamic Dimensionality Selection for Bayesian Classifier Ensembles

DTIC Science & Technology

2015-03-19

learning of weights in an otherwise generatively learned naive Bayes classifier. WANBIA-C is very cometitive to Logistic Regression but much more...classifier, Generative learning, Discriminative learning, Naïve Bayes, Feature selection, Logistic regression , higher order attribute independence 16...discriminative learning of weights in an otherwise generatively learned naive Bayes classifier. WANBIA-C is very cometitive to Logistic Regression but
A review of logistic regression models used to predict post-fire tree mortality of western North American conifers

Treesearch

Travis Woolley; David C. Shaw; Lisa M. Ganio; Stephen Fitzgerald

2012-01-01

Logistic regression models used to predict tree mortality are critical to post-fire management, planning prescribed bums and understanding disturbance ecology. We review literature concerning post-fire mortality prediction using logistic regression models for coniferous tree species in the western USA. We include synthesis and review of: methods to develop, evaluate...

Covariate Imbalance and Adjustment for Logistic Regression Analysis of Clinical Trial Data

PubMed Central

Ciolino, Jody D.; Martin, Reneé H.; Zhao, Wenle; Jauch, Edward C.; Hill, Michael D.; Palesch, Yuko Y.

2014-01-01

In logistic regression analysis for binary clinical trial data, adjusted treatment effect estimates are often not equivalent to unadjusted estimates in the presence of influential covariates. This paper uses simulation to quantify the benefit of covariate adjustment in logistic regression. However, International Conference on Harmonization guidelines suggest that covariate adjustment be pre-specified. Unplanned adjusted analyses should be considered secondary. Results suggest that that if adjustment is not possible or unplanned in a logistic setting, balance in continuous covariates can alleviate some (but never all) of the shortcomings of unadjusted analyses. The case of log binomial regression is also explored. PMID:24138438
Differentially private distributed logistic regression using private and public data.

PubMed

Ji, Zhanglong; Jiang, Xiaoqian; Wang, Shuang; Xiong, Li; Ohno-Machado, Lucila

2014-01-01

Privacy protecting is an important issue in medical informatics and differential privacy is a state-of-the-art framework for data privacy research. Differential privacy offers provable privacy against attackers who have auxiliary information, and can be applied to data mining models (for example, logistic regression). However, differentially private methods sometimes introduce too much noise and make outputs less useful. Given available public data in medical research (e.g. from patients who sign open-consent agreements), we can design algorithms that use both public and private data sets to decrease the amount of noise that is introduced. In this paper, we modify the update step in Newton-Raphson method to propose a differentially private distributed logistic regression model based on both public and private data. We try our algorithm on three different data sets, and show its advantage over: (1) a logistic regression model based solely on public data, and (2) a differentially private distributed logistic regression model based on private data under various scenarios. Logistic regression models built with our new algorithm based on both private and public datasets demonstrate better utility than models that trained on private or public datasets alone without sacrificing the rigorous privacy guarantee.
Logistic regression analysis of conventional ultrasonography, strain elastosonography, and contrast-enhanced ultrasound characteristics for the differentiation of benign and malignant thyroid nodules

PubMed Central

Deng, Yingyuan; Wang, Tianfu; Chen, Siping; Liu, Weixiang

2017-01-01

The aim of the study is to screen the significant sonographic features by logistic regression analysis and fit a model to diagnose thyroid nodules. A total of 525 pathological thyroid nodules were retrospectively analyzed. All the nodules underwent conventional ultrasonography (US), strain elastosonography (SE), and contrast -enhanced ultrasound (CEUS). Those nodules’ 12 suspicious sonographic features were used to assess thyroid nodules. The significant features of diagnosing thyroid nodules were picked out by logistic regression analysis. All variables that were statistically related to diagnosis of thyroid nodules, at a level of p < 0.05 were embodied in a logistic regression analysis model. The significant features in the logistic regression model of diagnosing thyroid nodules were calcification, suspected cervical lymph node metastasis, hypoenhancement pattern, margin, shape, vascularity, posterior acoustic, echogenicity, and elastography score. According to the results of logistic regression analysis, the formula that could predict whether or not thyroid nodules are malignant was established. The area under the receiver operating curve (ROC) was 0.930 and the sensitivity, specificity, accuracy, positive predictive value, and negative predictive value were 83.77%, 89.56%, 87.05%, 86.04%, and 87.79% respectively. PMID:29228030
Logistic regression analysis of conventional ultrasonography, strain elastosonography, and contrast-enhanced ultrasound characteristics for the differentiation of benign and malignant thyroid nodules.

PubMed

Pang, Tiantian; Huang, Leidan; Deng, Yingyuan; Wang, Tianfu; Chen, Siping; Gong, Xuehao; Liu, Weixiang

2017-01-01

The aim of the study is to screen the significant sonographic features by logistic regression analysis and fit a model to diagnose thyroid nodules. A total of 525 pathological thyroid nodules were retrospectively analyzed. All the nodules underwent conventional ultrasonography (US), strain elastosonography (SE), and contrast -enhanced ultrasound (CEUS). Those nodules' 12 suspicious sonographic features were used to assess thyroid nodules. The significant features of diagnosing thyroid nodules were picked out by logistic regression analysis. All variables that were statistically related to diagnosis of thyroid nodules, at a level of p < 0.05 were embodied in a logistic regression analysis model. The significant features in the logistic regression model of diagnosing thyroid nodules were calcification, suspected cervical lymph node metastasis, hypoenhancement pattern, margin, shape, vascularity, posterior acoustic, echogenicity, and elastography score. According to the results of logistic regression analysis, the formula that could predict whether or not thyroid nodules are malignant was established. The area under the receiver operating curve (ROC) was 0.930 and the sensitivity, specificity, accuracy, positive predictive value, and negative predictive value were 83.77%, 89.56%, 87.05%, 86.04%, and 87.79% respectively.
Prevalence and Determinants of Preterm Birth in Tehran, Iran: A Comparison between Logistic Regression and Decision Tree Methods.

PubMed

Amini, Payam; Maroufizadeh, Saman; Samani, Reza Omani; Hamidi, Omid; Sepidarkish, Mahdi

2017-06-01

Preterm birth (PTB) is a leading cause of neonatal death and the second biggest cause of death in children under five years of age. The objective of this study was to determine the prevalence of PTB and its associated factors using logistic regression and decision tree classification methods. This cross-sectional study was conducted on 4,415 pregnant women in Tehran, Iran, from July 6-21, 2015. Data were collected by a researcher-developed questionnaire through interviews with mothers and review of their medical records. To evaluate the accuracy of the logistic regression and decision tree methods, several indices such as sensitivity, specificity, and the area under the curve were used. The PTB rate was 5.5% in this study. The logistic regression outperformed the decision tree for the classification of PTB based on risk factors. Logistic regression showed that multiple pregnancies, mothers with preeclampsia, and those who conceived with assisted reproductive technology had an increased risk for PTB ( p < 0.05). Identifying and training mothers at risk as well as improving prenatal care may reduce the PTB rate. We also recommend that statisticians utilize the logistic regression model for the classification of risk groups for PTB.
Constitution of traditional chinese medicine and related factors in women of childbearing age.

PubMed

Jiang, Qiao-Yu; Li, Jue; Zheng, Liang; Wang, Guang-Hua; Wang, Jing

2018-04-01

This study investigates the constitution of traditional Chinese medicine (TCM) among women who want to be pregnant in one year and explores factors related to TCM constitution. This study was conducted on women who participated in free preconception check-ups provided by the Zhabei District Maternity and Child Care Center in Shanghai, China. The information regarding the female demographic characteristics, physical condition, history of pregnancy and childbearing, diet and behavior, and social psychological factors was collected, and TCM constitution assessment was performed. The Chi-square test, t-test, logistic regression analysis, and multinomial logistic regression analysis were used to explore the related factors of TCM constitution. The participants in this study were aged 28.3 ± 3.0 years. Approximately fifty-five women in this study had Unbalanced Constitution. Logistic regression analysis showed that Shanghai residence, dysmenorrhea, gum bleeding, aversion to vegetables, preference for raw meat, job stress, and economic stress were significantly and negatively associated with Balanced Constitution. Multinomial logistic analysis showed that Shanghai residence was significantly associated with Yang-deficiency, Yin-deficiency, and Stagnant Qi Constitutions; gum bleeding was significantly associated with Yin-deficiency, Stagnant Blood, Stagnant Qi, and Inherited Special Constitutions; aversion to vegetables was significantly associated with Damp-heat Constitution; job stress was significantly associated with Yang-deficiency, Phlegm-dampness, Damp-heat, Stagnant Blood, and Stagnant Qi Constitutions; and economic stress was significantly associated with Yang-deficiency, and Stagnant Qi Constitutions. The application of TCM constitution to preconception care would be beneficial for early identification of potential TCM constitution risks and be beneficial for early intervention (e.g., health education, and dietary education), especially during the women who do not have a medical condition and those who have related factors found in this study. Copyright © 2018. Published by Elsevier Taiwan LLC.
Methodology for constructing a colour-difference acceptability scale.

PubMed

Laborie, Baptiste; Viénot, Françoise; Langlois, Sabine

2010-09-01

Observers were invited to report their degree of satisfaction on a 6-point semantic scale with respect to the conformity of a test colour with a white reference colour, simultaneously presented on a PDP display. Eight test patches were chosen along each of the +a*, -a*, +b*, -b* axes of the CIELAB chromaticity plane, at Y = 80 ± 2 cd.m(-2) . Experimental conditions reliably represented the automotive environment (patch size, angular distance between patches) and observers could move their head and eyes freely. We have compared several methods of category scaling, the Torgerson-DMT method (Torgerson, W. S. (1958). Theory and methods of scaling. Wiley, New York, USA); two versions of the regression method i.e. Bonnet's (Bonnet, C. (1986). Manuel pratique de psychophysique. Armand Colin, Paris, France) and logistic regression; and the medians method. We describe in detail a case where all methods yield substantial but slightly different results. The solution proposed by the regression method which works with incomplete matrices and yields results directly on a colorimetric scale is probably the most useful in this industrial context. Finally we summarize the implementation of the logistic regression method over four hues and for one experimental condition. © 2010 The Authors, Ophthalmic and Physiological Optics © 2010 The College of Optometrists.
Independent Prognostic Factors for Acute Organophosphorus Pesticide Poisoning.

PubMed

Tang, Weidong; Ruan, Feng; Chen, Qi; Chen, Suping; Shao, Xuebo; Gao, Jianbo; Zhang, Mao

2016-07-01

Acute organophosphorus pesticide poisoning (AOPP) is becoming a significant problem and a potential cause of human mortality because of the abuse of organophosphate compounds. This study aims to determine the independent prognostic factors of AOPP by using multivariate logistic regression analysis. The clinical data for 71 subjects with AOPP admitted to our hospital were retrospectively analyzed. This information included the Acute Physiology and Chronic Health Evaluation II (APACHE II) scores, 6-h post-admission blood lactate levels, post-admission 6-h lactate clearance rates, admission blood cholinesterase levels, 6-h post-admission blood cholinesterase levels, cholinesterase activity, blood pH, and other factors. Univariate analysis and multivariate logistic regression analyses were conducted to identify all prognostic factors and independent prognostic factors, respectively. A receiver operating characteristic curve was plotted to analyze the testing power of independent prognostic factors. Twelve of 71 subjects died. Admission blood lactate levels, 6-h post-admission blood lactate levels, post-admission 6-h lactate clearance rates, blood pH, and APACHE II scores were identified as prognostic factors for AOPP according to the univariate analysis, whereas only 6-h post-admission blood lactate levels, post-admission 6-h lactate clearance rates, and blood pH were independent prognostic factors identified by multivariate logistic regression analysis. The receiver operating characteristic analysis suggested that post-admission 6-h lactate clearance rates were of moderate diagnostic value. High 6-h post-admission blood lactate levels, low blood pH, and low post-admission 6-h lactate clearance rates were independent prognostic factors identified by multivariate logistic regression analysis. Copyright © 2016 by Daedalus Enterprises.
Selenium in irrigated agricultural areas of the western United States

USGS Publications Warehouse

Nolan, B.T.; Clark, M.L.

1997-01-01

A logistic regression model was developed to predict the likelihood that Se exceeds the USEPA chronic criterion for aquatic life (5 ??g/L) in irrigated agricultural areas of the western USA. Preliminary analysis of explanatory variables used in the model indicated that surface-water Se concentration increased with increasing dissolved solids (DS) concentration and with the presence of Upper Cretaceous, mainly marine sediment. The presence or absence of Cretaceous sediment was the major variable affecting Se concentration in surface-water samples from the National Irrigation Water Quality Program. Median Se concentration was 14 ??g/L in samples from areas underlain by Cretaceous sediments and < 1 ??g/L in samples from areas underlain by non-Cretaceous sediments. Wilcoxon rank sum tests indicated that elevated Se concentrations in samples from areas with Cretaceous sediments, irrigated areas, and from closed lakes and ponds were statistically significant. Spearman correlations indicated that Se was positively correlated with a binary geology variable (0.64) and DS (0.45). Logistic regression models indicated that the concentration of Se in surface water was almost certain to exceed the Environmental Protection Agency aquatic-life chronic criterion of 5 ??g/L when DS was greater than 3000 mg/L in areas with Cretaceous sediments. The 'best' logistic regression model correctly predicted Se exceedances and nonexceedances 84.4% of the time, and model sensitivity was 80.7%. A regional map of Cretaceous sediment showed the location of potential problem areas. The map and logistic regression model are tools that can be used to determine the potential for Se contamination of irrigated agricultural areas in the western USA.
Logistic regression for dichotomized counts.

PubMed

Preisser, John S; Das, Kalyan; Benecha, Habtamu; Stamm, John W

2016-12-01

Sometimes there is interest in a dichotomized outcome indicating whether a count variable is positive or zero. Under this scenario, the application of ordinary logistic regression may result in efficiency loss, which is quantifiable under an assumed model for the counts. In such situations, a shared-parameter hurdle model is investigated for more efficient estimation of regression parameters relating to overall effects of covariates on the dichotomous outcome, while handling count data with many zeroes. One model part provides a logistic regression containing marginal log odds ratio effects of primary interest, while an ancillary model part describes the mean count of a Poisson or negative binomial process in terms of nuisance regression parameters. Asymptotic efficiency of the logistic model parameter estimators of the two-part models is evaluated with respect to ordinary logistic regression. Simulations are used to assess the properties of the models with respect to power and Type I error, the latter investigated under both misspecified and correctly specified models. The methods are applied to data from a randomized clinical trial of three toothpaste formulations to prevent incident dental caries in a large population of Scottish schoolchildren. © The Author(s) 2014.
Predicting 30-day Hospital Readmission with Publicly Available Administrative Database. A Conditional Logistic Regression Modeling Approach.

PubMed

Zhu, K; Lou, Z; Zhou, J; Ballester, N; Kong, N; Parikh, P

2015-01-01

This article is part of the Focus Theme of Methods of Information in Medicine on "Big Data and Analytics in Healthcare". Hospital readmissions raise healthcare costs and cause significant distress to providers and patients. It is, therefore, of great interest to healthcare organizations to predict what patients are at risk to be readmitted to their hospitals. However, current logistic regression based risk prediction models have limited prediction power when applied to hospital administrative data. Meanwhile, although decision trees and random forests have been applied, they tend to be too complex to understand among the hospital practitioners. Explore the use of conditional logistic regression to increase the prediction accuracy. We analyzed an HCUP statewide inpatient discharge record dataset, which includes patient demographics, clinical and care utilization data from California. We extracted records of heart failure Medicare beneficiaries who had inpatient experience during an 11-month period. We corrected the data imbalance issue with under-sampling. In our study, we first applied standard logistic regression and decision tree to obtain influential variables and derive practically meaning decision rules. We then stratified the original data set accordingly and applied logistic regression on each data stratum. We further explored the effect of interacting variables in the logistic regression modeling. We conducted cross validation to assess the overall prediction performance of conditional logistic regression (CLR) and compared it with standard classification models. The developed CLR models outperformed several standard classification models (e.g., straightforward logistic regression, stepwise logistic regression, random forest, support vector machine). For example, the best CLR model improved the classification accuracy by nearly 20% over the straightforward logistic regression model. Furthermore, the developed CLR models tend to achieve better sensitivity of more than 10% over the standard classification models, which can be translated to correct labeling of additional 400 - 500 readmissions for heart failure patients in the state of California over a year. Lastly, several key predictor identified from the HCUP data include the disposition location from discharge, the number of chronic conditions, and the number of acute procedures. It would be beneficial to apply simple decision rules obtained from the decision tree in an ad-hoc manner to guide the cohort stratification. It could be potentially beneficial to explore the effect of pairwise interactions between influential predictors when building the logistic regression models for different data strata. Judicious use of the ad-hoc CLR models developed offers insights into future development of prediction models for hospital readmissions, which can lead to better intuition in identifying high-risk patients and developing effective post-discharge care strategies. Lastly, this paper is expected to raise the awareness of collecting data on additional markers and developing necessary database infrastructure for larger-scale exploratory studies on readmission risk prediction.
Characteristics and Psychosocial Predictors of Adolescent Nonsuicidal Self-Injury in Residential Care

ERIC Educational Resources Information Center

Gallant, Jason; Snyder, Gregory S.; von der Embse, Nathaniel P.

2014-01-01

This study examined characteristics and biopsychosocial predictors of nonsuicidal self-injury in a sample (N = 753) of youth in residential care admitted between 2005 and 2010. To model the data, the authors used t-tests, chi-square tests, and multiple logistic regressions stratified by gender. Results suggested that 12% of youth engaged in…
HIV Testing Behavior among Pacific Islanders in Southern California: Exploring the Importance of Race/Ethnicity, Knowledge, and Domestic Violence

ERIC Educational Resources Information Center

Takahashi, Lois M.; Kim, Anna J.; Sablan-Santos, Lola; Quitugua, Lourdes Flores; Lepule, Jonathan; Maguadog, Tony; Perez, Rose; Young, Steve; Young, Louise

2011-01-01

This article presents an analysis of a 2008 community needs assessment survey of a convenience sample of 179 Pacific Islander respondents in southern California; the needs assessment focused on HIV knowledge, HIV testing behavior, and experience with intimate partner/relationship violence. Multivariate logistic regression results indicated that…
An Empirical Investigation of the Potential Impact of Item Misfit on Test Scores. Research Report. ETS RR-17-60

ERIC Educational Resources Information Center

Kim, Sooyeon; Robin, Frederic

2017-01-01

In this study, we examined the potential impact of item misfit on the reported scores of an admission test from the subpopulation invariance perspective. The target population of the test consisted of 3 major subgroups with different geographic regions. We used the logistic regression function to estimate item parameters of the operational items…
Achievement Gap Projection for Standardized Testing through Logistic Regression within a Large Arizona School District

ERIC Educational Resources Information Center

Kellermeyer, Steven Bruce

2011-01-01

In the last few decades high-stakes testing has become more political than educational. The Districts within Arizona are bound by the mandates of both AZ LEARNS and the No Child Left Behind Act of 2001. At the time of this writing, both legislative mandates relied on the Arizona Instrument for Measuring Standards (AIMS) as State Tests for gauging…
HIV-Related Risk Behaviors, Perceptions of Risk, HIV Testing, and Exposure to Prevention Messages and Methods among Urban American Indians and Alaska Natives

ERIC Educational Resources Information Center

Lapidus, Jodi A.; Bertolli, Jeanne; McGowan, Karen; Sullivan, Patrick

2006-01-01

The goal of this study was to describe HIV risk behaviors, perceptions, testing, and prevention exposure among urban American Indians and Alaska Natives (AI/AN). Interviewers administered a questionnaire to participants recruited through anonymous peer-referral sampling. Chi-square tests and multiple logistic regression were used to compare HIV…
Interpretation of commonly used statistical regression models.

PubMed

Kasza, Jessica; Wolfe, Rory

2014-01-01

A review of some regression models commonly used in respiratory health applications is provided in this article. Simple linear regression, multiple linear regression, logistic regression and ordinal logistic regression are considered. The focus of this article is on the interpretation of the regression coefficients of each model, which are illustrated through the application of these models to a respiratory health research study. © 2013 The Authors. Respirology © 2013 Asian Pacific Society of Respirology.
Evaluation of logistic regression models and effect of covariates for case-control study in RNA-Seq analysis.

PubMed

Choi, Seung Hoan; Labadorf, Adam T; Myers, Richard H; Lunetta, Kathryn L; Dupuis, Josée; DeStefano, Anita L

2017-02-06

Next generation sequencing provides a count of RNA molecules in the form of short reads, yielding discrete, often highly non-normally distributed gene expression measurements. Although Negative Binomial (NB) regression has been generally accepted in the analysis of RNA sequencing (RNA-Seq) data, its appropriateness has not been exhaustively evaluated. We explore logistic regression as an alternative method for RNA-Seq studies designed to compare cases and controls, where disease status is modeled as a function of RNA-Seq reads using simulated and Huntington disease data. We evaluate the effect of adjusting for covariates that have an unknown relationship with gene expression. Finally, we incorporate the data adaptive method in order to compare false positive rates. When the sample size is small or the expression levels of a gene are highly dispersed, the NB regression shows inflated Type-I error rates but the Classical logistic and Bayes logistic (BL) regressions are conservative. Firth's logistic (FL) regression performs well or is slightly conservative. Large sample size and low dispersion generally make Type-I error rates of all methods close to nominal alpha levels of 0.05 and 0.01. However, Type-I error rates are controlled after applying the data adaptive method. The NB, BL, and FL regressions gain increased power with large sample size, large log2 fold-change, and low dispersion. The FL regression has comparable power to NB regression. We conclude that implementing the data adaptive method appropriately controls Type-I error rates in RNA-Seq analysis. Firth's logistic regression provides a concise statistical inference process and reduces spurious associations from inaccurately estimated dispersion parameters in the negative binomial framework.
ADCYAP1R1 and asthma in Puerto Rican children.

PubMed

Chen, Wei; Boutaoui, Nadia; Brehm, John M; Han, Yueh-Ying; Schmitz, Cassandra; Cressley, Alex; Acosta-Pérez, Edna; Alvarez, María; Colón-Semidey, Angel; Baccarelli, Andrea A; Weeks, Daniel E; Kolls, Jay K; Canino, Glorisa; Celedón, Juan C

2013-03-15

Epigenetic and/or genetic variation in the gene encoding the receptor for adenylate-cyclase activating polypeptide 1 (ADCYAP1R1) has been linked to post-traumatic stress disorder in adults and anxiety in children. Psychosocial stress has been linked to asthma morbidity in Puerto Rican children. To examine whether epigenetic or genetic variation in ADCYAP1R1 is associated with childhood asthma in Puerto Ricans. We conducted a case-control study of 516 children ages 6-14 years living in San Juan, Puerto Rico. We assessed methylation at a CpG site in the promoter of ADCYAP1R1 (cg11218385) using a pyrosequencing assay in DNA from white blood cells. We tested whether cg11218385 methylation (range, 0.4-6.1%) is associated with asthma using logistic regression. We also examined whether exposure to violence (assessed by the Exposure to Violence [ETV] Scale in children 9 yr and older) is associated with cg11218385 methylation (using linear regression) or asthma (using logistic regression). Logistic regression was used to test for association between a single nucleotide polymorphism in ADCYAP1R1 (rs2267735) and asthma under an additive model. All multivariate models were adjusted for age, sex, household income, and principal components. EACH 1% increment in cg11218385 methylation was associated with increased odds of asthma (adjusted odds ratio, 1.3; 95% confidence interval, 1.0-1.6; P = 0.03). Among children 9 years and older, exposure to violence was associated with cg11218385 methylation. The C allele of single nucleotide polymorphism rs2267735 was significantly associated with increased odds of asthma (adjusted odds ratio, 1.3; 95% confidence interval, 1.02-1.67; P = 0.03). Epigenetic and genetic variants in ADCYAP1R1 are associated with asthma in Puerto Rican children.
[Use of data display screens and ocular hypertension in local public sector workers].

PubMed

Abellán Torró, Rosana; Merelles Tormo, Antoni

2014-01-01

The main objective of this study is to examine the association between work with data display screens (DDS) and ocular hypertension (OHT). A cross-sectional study among local public sector workers (Diputación Provincial de Valencia). Data from 620 people were collected over 25 months, from periodic medical examinations performed at an occupational health unit. Intraocular pressure (IOP) was obtained with a portable puff tonometer validated for screening, establishing the cut-off point for OHT at 22 mmHg. Both biological characteristics and other work-related variables were taken into account as covariates. Descriptive statistics of the data were obtained, together with nonparametric tests with a level of significance of 95% and logistic regression with p 〈0.1 as the level of significance of the likelihood test. The average age of the study population is 52.8 years. The prevalence of OHT was 3.5% (5.1% among men and 1.2% among women; p=0.012). No significant associations were found between hours of DDS-related work and OHT were found (p=0.395). Logistic regression corroborated the association between gender and OHT, with women less affected (OR = 0.234; 95%CI: 0.068 - 0.799; p=0.020). In our study, no associations were found between time of exposure to data display screens and ocular hypertension. Logistic regression points to a certain association between ocular hypertension and gender, with men being more predisposed. Copyright belongs to the Societat Catalana de Salut Laboral.

Differentially private distributed logistic regression using private and public data

PubMed Central

2014-01-01

Background Privacy protecting is an important issue in medical informatics and differential privacy is a state-of-the-art framework for data privacy research. Differential privacy offers provable privacy against attackers who have auxiliary information, and can be applied to data mining models (for example, logistic regression). However, differentially private methods sometimes introduce too much noise and make outputs less useful. Given available public data in medical research (e.g. from patients who sign open-consent agreements), we can design algorithms that use both public and private data sets to decrease the amount of noise that is introduced. Methodology In this paper, we modify the update step in Newton-Raphson method to propose a differentially private distributed logistic regression model based on both public and private data. Experiments and results We try our algorithm on three different data sets, and show its advantage over: (1) a logistic regression model based solely on public data, and (2) a differentially private distributed logistic regression model based on private data under various scenarios. Conclusion Logistic regression models built with our new algorithm based on both private and public datasets demonstrate better utility than models that trained on private or public datasets alone without sacrificing the rigorous privacy guarantee. PMID:25079786
A retrospective analysis to identify the factors affecting infection in patients undergoing chemotherapy.

PubMed

Park, Ji Hyun; Kim, Hyeon-Young; Lee, Hanna; Yun, Eun Kyoung

2015-12-01

This study compares the performance of the logistic regression and decision tree analysis methods for assessing the risk factors for infection in cancer patients undergoing chemotherapy. The subjects were 732 cancer patients who were receiving chemotherapy at K university hospital in Seoul, Korea. The data were collected between March 2011 and February 2013 and were processed for descriptive analysis, logistic regression and decision tree analysis using the IBM SPSS Statistics 19 and Modeler 15.1 programs. The most common risk factors for infection in cancer patients receiving chemotherapy were identified as alkylating agents, vinca alkaloid and underlying diabetes mellitus. The logistic regression explained 66.7% of the variation in the data in terms of sensitivity and 88.9% in terms of specificity. The decision tree analysis accounted for 55.0% of the variation in the data in terms of sensitivity and 89.0% in terms of specificity. As for the overall classification accuracy, the logistic regression explained 88.0% and the decision tree analysis explained 87.2%. The logistic regression analysis showed a higher degree of sensitivity and classification accuracy. Therefore, logistic regression analysis is concluded to be the more effective and useful method for establishing an infection prediction model for patients undergoing chemotherapy. Copyright © 2015 Elsevier Ltd. All rights reserved.
Performance and strategy comparisons of human listeners and logistic regression in discriminating underwater targets.

PubMed

Yang, Lixue; Chen, Kean

2015-11-01

To improve the design of underwater target recognition systems based on auditory perception, this study compared human listeners with automatic classifiers. Performances measures and strategies in three discrimination experiments, including discriminations between man-made and natural targets, between ships and submarines, and among three types of ships, were used. In the experiments, the subjects were asked to assign a score to each sound based on how confident they were about the category to which it belonged, and logistic regression, which represents linear discriminative models, also completed three similar tasks by utilizing many auditory features. The results indicated that the performances of logistic regression improved as the ratio between inter- and intra-class differences became larger, whereas the performances of the human subjects were limited by their unfamiliarity with the targets. Logistic regression performed better than the human subjects in all tasks but the discrimination between man-made and natural targets, and the strategies employed by excellent human subjects were similar to that of logistic regression. Logistic regression and several human subjects demonstrated similar performances when discriminating man-made and natural targets, but in this case, their strategies were not similar. An appropriate fusion of their strategies led to further improvement in recognition accuracy.
Simulating land-use changes by incorporating spatial autocorrelation and self-organization in CLUE-S modeling: a case study in Zengcheng District, Guangzhou, China

NASA Astrophysics Data System (ADS)

Mei, Zhixiong; Wu, Hao; Li, Shiyun

2018-06-01

The Conversion of Land Use and its Effects at Small regional extent (CLUE-S), which is a widely used model for land-use simulation, utilizes logistic regression to estimate the relationships between land use and its drivers, and thus, predict land-use change probabilities. However, logistic regression disregards possible spatial autocorrelation and self-organization in land-use data. Autologistic regression can depict spatial autocorrelation but cannot address self-organization, while logistic regression by considering only self-organization (NElogistic regression) fails to capture spatial autocorrelation. Therefore, this study developed a regression (NE-autologistic regression) method, which incorporated both spatial autocorrelation and self-organization, to improve CLUE-S. The Zengcheng District of Guangzhou, China was selected as the study area. The land-use data of 2001, 2005, and 2009, as well as 10 typical driving factors, were used to validate the proposed regression method and the improved CLUE-S model. Then, three future land-use scenarios in 2020: the natural growth scenario, ecological protection scenario, and economic development scenario, were simulated using the improved model. Validation results showed that NE-autologistic regression performed better than logistic regression, autologistic regression, and NE-logistic regression in predicting land-use change probabilities. The spatial allocation accuracy and kappa values of NE-autologistic-CLUE-S were higher than those of logistic-CLUE-S, autologistic-CLUE-S, and NE-logistic-CLUE-S for the simulations of two periods, 2001-2009 and 2005-2009, which proved that the improved CLUE-S model achieved the best simulation and was thereby effective to a certain extent. The scenario simulation results indicated that under all three scenarios, traffic land and residential/industrial land would increase, whereas arable land and unused land would decrease during 2009-2020. Apparent differences also existed in the simulated change sizes and locations of each land-use type under different scenarios. The results not only demonstrate the validity of the improved model but also provide a valuable reference for relevant policy-makers.
[Analysis of risk factors for dry eye syndrome in visual display terminal workers].

PubMed

Zhu, Yong; Yu, Wen-lan; Xu, Ming; Han, Lei; Cao, Wen-dong; Zhang, Hong-bing; Zhang, Heng-dong

2013-08-01

To analyze the risk factors for dry eye syndrome in visual display terminal (VDT) workers and to provide a scientific basis for protecting the eye health of VDT workers. Questionnaire survey, Schirmer I test, tear break-up time test, and workshop microenvironment evaluation were performed in 185 VDT workers. Multivariate logistic regression analysis was performed to determine the risk factors for dry eye syndrome in VDT workers after adjustment for confounding factors. In the logistic regression model, the regression coefficients of daily mean time of exposure to screen, daily mean time of watching TV, parallel screen-eye angle, upward screen-eye angle, eye-screen distance of less than 20 cm, irregular breaks during screen-exposed work, age, and female gender on the results of Schirmer I test were 0.153, 0.548, 0.400, 0.796, 0.234, 0.516, 0.559, and -0.685, respectively; the regression coefficients of daily mean time of exposure to screen, parallel screen-eye angle, upward screen-eye angle, age, working years, and female gender on tear break-up time were 0.021, 0.625, 2.652, 0.749, 0.403, and 1.481, respectively. Daily mean time of exposure to screen, daily mean time of watching TV, parallel screen-eye angle, upward screen-eye angle, eye-screen distance of less than 20 cm, irregular breaks during screen-exposed work, age, and working years are risk factors for dry eye syndrome in VDT workers.
Peak oxygen consumption measured during the stair-climbing test in lung resection candidates.

PubMed

Brunelli, Alessandro; Xiumé, Francesco; Refai, Majed; Salati, Michele; Di Nunzio, Luca; Pompili, Cecilia; Sabbatini, Armando

2010-01-01

The stair-climbing test is commonly used in the preoperative evaluation of lung resection candidates, but it is difficult to standardize and provides little physiologic information on the performance. To verify the association between the altitude and the V(O2peak) measured during the stair-climbing test. 109 consecutive candidates for lung resection performed a symptom-limited stair-climbing test with direct breath-by-breath measurement of V(O2peak) by a portable gas analyzer. Stepwise logistic regression and bootstrap analyses were used to verify the association of several perioperative variables with a V(O2peak) <15 ml/kg/min. Subsequently, multiple regression analysis was also performed to develop an equation to estimate V(O2peak) from stair-climbing parameters and other patient-related variables. 56% of patients climbing <14 m had a V(O2peak) <15 ml/kg/min, whereas 98% of those climbing >22 m had a V(O2peak) >15 ml/kg/min. The altitude reached at stair-climbing test resulted in the only significant predictor of a V(O2peak) <15 ml/kg/min after logistic regression analysis. Multiple regression analysis yielded an equation to estimate V(O2peak) factoring altitude (p < 0.0001), speed of ascent (p = 0.005) and body mass index (p = 0.0008). There was an association between altitude and V(O2peak) measured during the stair-climbing test. Most of the patients climbing more than 22 m are able to generate high values of V(O2peak) and can proceed to surgery without any additional tests. All others need to be referred for a formal cardiopulmonary exercise test. In addition, we were able to generate an equation to estimate V(O2peak), which could assist in streamlining the preoperative workup and could be used across different settings to standardize this test. Copyright (c) 2010 S. Karger AG, Basel.
Unitary Response Regression Models

ERIC Educational Resources Information Center

Lipovetsky, S.

2007-01-01

The dependent variable in a regular linear regression is a numerical variable, and in a logistic regression it is a binary or categorical variable. In these models the dependent variable has varying values. However, there are problems yielding an identity output of a constant value which can also be modelled in a linear or logistic regression with…
Binary logistic regression-Instrument for assessing museum indoor air impact on exhibits.

PubMed

Bucur, Elena; Danet, Andrei Florin; Lehr, Carol Blaziu; Lehr, Elena; Nita-Lazar, Mihai

2017-04-01

This paper presents a new way to assess the environmental impact on historical artifacts using binary logistic regression. The prediction of the impact on the exhibits during certain pollution scenarios (environmental impact) was calculated by a mathematical model based on the binary logistic regression; it allows the identification of those environmental parameters from a multitude of possible parameters with a significant impact on exhibitions and ranks them according to their severity effect. Air quality (NO 2 , SO 2 , O 3 and PM 2.5 ) and microclimate parameters (temperature, humidity) monitoring data from a case study conducted within exhibition and storage spaces of the Romanian National Aviation Museum Bucharest have been used for developing and validating the binary logistic regression method and the mathematical model. The logistic regression analysis was used on 794 data combinations (715 to develop of the model and 79 to validate it) by a Statistical Package for Social Sciences (SPSS 20.0). The results from the binary logistic regression analysis demonstrated that from six parameters taken into consideration, four of them present a significant effect upon exhibits in the following order: O 3 >PM 2.5 >NO 2 >humidity followed at a significant distance by the effects of SO 2 and temperature. The mathematical model, developed in this study, correctly predicted 95.1 % of the cumulated effect of the environmental parameters upon the exhibits. Moreover, this model could also be used in the decisional process regarding the preventive preservation measures that should be implemented within the exhibition space. The paper presents a new way to assess the environmental impact on historical artifacts using binary logistic regression. The mathematical model developed on the environmental parameters analyzed by the binary logistic regression method could be useful in a decision-making process establishing the best measures for pollution reduction and preventive preservation of exhibits.
Determining factors influencing survival of breast cancer by fuzzy logistic regression model.

PubMed

Nikbakht, Roya; Bahrampour, Abbas

2017-01-01

Fuzzy logistic regression model can be used for determining influential factors of disease. This study explores the important factors of actual predictive survival factors of breast cancer's patients. We used breast cancer data which collected by cancer registry of Kerman University of Medical Sciences during the period of 2000-2007. The variables such as morphology, grade, age, and treatments (surgery, radiotherapy, and chemotherapy) were applied in the fuzzy logistic regression model. Performance of model was determined in terms of mean degree of membership (MDM). The study results showed that almost 41% of patients were in neoplasm and malignant group and more than two-third of them were still alive after 5-year follow-up. Based on the fuzzy logistic model, the most important factors influencing survival were chemotherapy, morphology, and radiotherapy, respectively. Furthermore, the MDM criteria show that the fuzzy logistic regression have a good fit on the data (MDM = 0.86). Fuzzy logistic regression model showed that chemotherapy is more important than radiotherapy in survival of patients with breast cancer. In addition, another ability of this model is calculating possibilistic odds of survival in cancer patients. The results of this study can be applied in clinical research. Furthermore, there are few studies which applied the fuzzy logistic models. Furthermore, we recommend using this model in various research areas.
Seroprevalence of human hydatidosis using ELISA method in qom province, central iran.

PubMed

Rakhshanpour, A; Harandi, M Fasihi; Moazezi, Ss; Rahimi, Mt; Mohebali, M; Mowlavi, Ghh; Babaei, Z; Ariaeipour, M; Heidari, Z; Rokni, Mb

2012-01-01

The objective of this study was to determine the prevalence of cystic echinococcosis (CE) in Qom Province, central Iran using ELISA test. Overall, 1564 serum samples (800 males and 764 females) were collected from selected subjects by randomized cluster sampling in 2011-2012. Sera were analyzed by ELISA test using AgB. Before sampling, a questionnaire was filled out for each case. Data were analyzed using Chi-square test and multivariate logistic regression for risk factors analysis. Seropositivity was 1.6% (25 cases). Males (2.2%) showed significantly more positivity than females (0.9%) (P= 0.03). There was no significant association between CE seropositivity and age group, occupation, and region. Age group of 30-60 years encompassed the highest rate of positivity. The seropositivity of CE was 2.1% and 1.2% for urban and rural cases respectively. Binary logistic regression showed that males were 2.5 times at higher risk for infection than females. Although seroprevalence of CE is relatively low in Qom Province, yet due to the importance of the disease, all preventive measures should be taken into consideration.
Effort test failure: toward a predictive model.

PubMed

Webb, James W; Batchelor, Jennifer; Meares, Susanne; Taylor, Alan; Marsh, Nigel V

2012-01-01

Predictors of effort test failure were examined in an archival sample of 555 traumatically brain-injured (TBI) adults. Logistic regression models were used to examine whether compensation-seeking, injury-related, psychological, demographic, and cultural factors predicted effort test failure (ETF). ETF was significantly associated with compensation-seeking (OR = 3.51, 95% CI [1.25, 9.79]), low education (OR:. 83 [.74, . 94]), self-reported mood disorder (OR: 5.53 [3.10, 9.85]), exaggerated displays of behavior (OR: 5.84 [2.15, 15.84]), psychotic illness (OR: 12.86 [3.21, 51.44]), being foreign-born (OR: 5.10 [2.35, 11.06]), having sustained a workplace accident (OR: 4.60 [2.40, 8.81]), and mild traumatic brain injury severity compared with very severe traumatic brain injury severity (OR: 0.37 [0.13, 0.995]). ETF was associated with a broader range of statistical predictors than has previously been identified and the relative importance of psychological and behavioral predictors of ETF was evident in the logistic regression model. Variables that might potentially extend the model of ETF are identified for future research efforts.
Prevalence of consistent condom use with various types of sex partners and associated factors among money boys in Changsha, China.

PubMed

Wang, Lian-Hong; Yan, Jin; Yang, Guo-Li; Long, Shuo; Yu, Yong; Wu, Xi-Lin

2015-04-01

Money boys with inconsistent condom use (less than 100% of the time) are at high risk of infection by human immunodeficiency virus (HIV) or sexually transmitted infection (STI), but relatively little research has examined their risk behaviors. We investigated the prevalence of consistent condom use (100% of the time) and associated factors among money boys. A cross-sectional study using a structured questionnaire was conducted among money boys in Changsha, China, between July 2012 and January 2013. Independent variables included socio-demographic data, substance abuse history, work characteristics, and self-reported HIV and STI history. Dependent variables included the consistent condom use with different types of sex partners. Among the participants, 82.4% used condoms consistently with male clients, 80.2% with male sex partners, and 77.1% with female sex partners in the past 3 months. A multiple stepwise logistic regression model identified four statistically significant factors associated with lower likelihoods of consistent condom use with male clients: age group, substance abuse, lack of an "employment" arrangement, and having no HIV test within the prior 6 months. In a similar model, only one factor associated significantly with lower likelihoods of consistent condom use with male sex partners was identified in multiple stepwise logistic regression analyses: having no HIV test within the prior six months. As for female sex partners, two significant variables were statistically significant in the multiple stepwise logistic regression analysis: having no HIV test within the prior 6 months and having STI history. Interventions which are linked with more realistic and acceptable HIV prevention methods are greatly warranted and should increase risk awareness and the behavior of consistent condom use in both commercial and personal relationship. © 2015 International Society for Sexual Medicine.
C-reactive protein, platelets, and patent ductus arteriosus.

PubMed

Meinarde, Leonardo; Hillman, Macarena; Rizzotti, Alina; Basquiera, Ana Lisa; Tabares, Aldo; Cuestas, Eduardo

2016-12-01

The association between inflammation, platelets, and patent ductus arteriosus (PDA) has not been studied so far. The purpose of this study was to evaluate whether C-reactive protein (CRP) is related to low platelet count and PDA. This was a retrospective study of 88 infants with a birth weight ≤1500 g and a gestational age ≤30 weeks. Platelet count, CRP, and an echocardiogram were assessed in all infants. The subjects were matched by sex, gestational age, and birth weight. Differences were compared using the χ 2 , t-test, or Mann-Whitney U-test, as appropriate. Significant variables were entered into a logistic regression model. The association between CRP and platelets was evaluated by correlation and regression analysis. Platelet count (167 000 vs. 213 000 µl -1 , p = 0.015) was lower and the CRP (0.45 vs. 0.20 mg/dl, p = 0.002) was higher, and the platelet count correlated inversely with CRP (r = -0.145, p = 0.049) in the infants with vs. without PDA. Only CRP was independently associated with PDA in a logistic regression model (OR 64.1, 95% confidence interval 1.4-2941, p = 0.033).
Predicting No-Shows in Radiology Using Regression Modeling of Data Available in the Electronic Medical Record.

PubMed

Harvey, H Benjamin; Liu, Catherine; Ai, Jing; Jaworsky, Cristina; Guerrier, Claude Emmanuel; Flores, Efren; Pianykh, Oleg

2017-10-01

To test whether data elements available in the electronic medical record (EMR) can be effectively leveraged to predict failure to attend a scheduled radiology examination. Using data from a large academic medical center, we identified all patients with a diagnostic imaging examination scheduled from January 1, 2016, to April 1, 2016, and determined whether the patient successfully attended the examination. Demographic, clinical, and health services utilization variables available in the EMR potentially relevant to examination attendance were recorded for each patient. We used descriptive statistics and logistic regression models to test whether these data elements could predict failure to attend a scheduled radiology examination. The predictive accuracy of the regression models were determined by calculating the area under the receiver operator curve. Among the 54,652 patient appointments with radiology examinations scheduled during the study period, 6.5% were no-shows. No-show rates were highest for the modalities of mammography and CT and lowest for PET and MRI. Logistic regression indicated that 16 of the 27 demographic, clinical, and health services utilization factors were significantly associated with failure to attend a scheduled radiology examination (P ≤ .05). Stepwise logistic regression analysis demonstrated that previous no-shows, days between scheduling and appointments, modality type, and insurance type were most strongly predictive of no-show. A model considering all 16 data elements had good ability to predict radiology no-shows (area under the receiver operator curve = 0.753). The predictive ability was similar or improved when these models were analyzed by modality. Patient and examination information readily available in the EMR can be successfully used to predict radiology no-shows. Moving forward, this information can be proactively leveraged to identify patients who might benefit from additional patient engagement through appointment reminders or other targeted interventions to avoid no-shows. Copyright © 2017 American College of Radiology. Published by Elsevier Inc. All rights reserved.
Mapping of the DLQI scores to EQ-5D utility values using ordinal logistic regression.

PubMed

Ali, Faraz Mahmood; Kay, Richard; Finlay, Andrew Y; Piguet, Vincent; Kupfer, Joerg; Dalgard, Florence; Salek, M Sam

2017-11-01

The Dermatology Life Quality Index (DLQI) and the European Quality of Life-5 Dimension (EQ-5D) are separate measures that may be used to gather health-related quality of life (HRQoL) information from patients. The EQ-5D is a generic measure from which health utility estimates can be derived, whereas the DLQI is a specialty-specific measure to assess HRQoL. To reduce the burden of multiple measures being administered and to enable a more disease-specific calculation of health utility estimates, we explored an established mathematical technique known as ordinal logistic regression (OLR) to develop an appropriate model to map DLQI data to EQ-5D-based health utility estimates. Retrospective data from 4010 patients were randomly divided five times into two groups for the derivation and testing of the mapping model. Split-half cross-validation was utilized resulting in a total of ten ordinal logistic regression models for each of the five EQ-5D dimensions against age, sex, and all ten items of the DLQI. Using Monte Carlo simulation, predicted health utility estimates were derived and compared against those observed. This method was repeated for both OLR and a previously tested mapping methodology based on linear regression. The model was shown to be highly predictive and its repeated fitting demonstrated a stable model using OLR as well as linear regression. The mean differences between OLR-predicted health utility estimates and observed health utility estimates ranged from 0.0024 to 0.0239 across the ten modeling exercises, with an average overall difference of 0.0120 (a 1.6% underestimate, not of clinical importance). This modeling framework developed in this study will enable researchers to calculate EQ-5D health utility estimates from a specialty-specific study population, reducing patient and economic burden.
Mixed conditional logistic regression for habitat selection studies.

PubMed

Duchesne, Thierry; Fortin, Daniel; Courbin, Nicolas

2010-05-01

1. Resource selection functions (RSFs) are becoming a dominant tool in habitat selection studies. RSF coefficients can be estimated with unconditional (standard) and conditional logistic regressions. While the advantage of mixed-effects models is recognized for standard logistic regression, mixed conditional logistic regression remains largely overlooked in ecological studies. 2. We demonstrate the significance of mixed conditional logistic regression for habitat selection studies. First, we use spatially explicit models to illustrate how mixed-effects RSFs can be useful in the presence of inter-individual heterogeneity in selection and when the assumption of independence from irrelevant alternatives (IIA) is violated. The IIA hypothesis states that the strength of preference for habitat type A over habitat type B does not depend on the other habitat types also available. Secondly, we demonstrate the significance of mixed-effects models to evaluate habitat selection of free-ranging bison Bison bison. 3. When movement rules were homogeneous among individuals and the IIA assumption was respected, fixed-effects RSFs adequately described habitat selection by simulated animals. In situations violating the inter-individual homogeneity and IIA assumptions, however, RSFs were best estimated with mixed-effects regressions, and fixed-effects models could even provide faulty conclusions. 4. Mixed-effects models indicate that bison did not select farmlands, but exhibited strong inter-individual variations in their response to farmlands. Less than half of the bison preferred farmlands over forests. Conversely, the fixed-effect model simply suggested an overall selection for farmlands. 5. Conditional logistic regression is recognized as a powerful approach to evaluate habitat selection when resource availability changes. This regression is increasingly used in ecological studies, but almost exclusively in the context of fixed-effects models. Fitness maximization can imply differences in trade-offs among individuals, which can yield inter-individual differences in selection and lead to departure from IIA. These situations are best modelled with mixed-effects models. Mixed-effects conditional logistic regression should become a valuable tool for ecological research.
Advanced colorectal neoplasia risk stratification by penalized logistic regression.

PubMed

Lin, Yunzhi; Yu, Menggang; Wang, Sijian; Chappell, Richard; Imperiale, Thomas F

2016-08-01

Colorectal cancer is the second leading cause of death from cancer in the United States. To facilitate the efficiency of colorectal cancer screening, there is a need to stratify risk for colorectal cancer among the 90% of US residents who are considered "average risk." In this article, we investigate such risk stratification rules for advanced colorectal neoplasia (colorectal cancer and advanced, precancerous polyps). We use a recently completed large cohort study of subjects who underwent a first screening colonoscopy. Logistic regression models have been used in the literature to estimate the risk of advanced colorectal neoplasia based on quantifiable risk factors. However, logistic regression may be prone to overfitting and instability in variable selection. Since most of the risk factors in our study have several categories, it was tempting to collapse these categories into fewer risk groups. We propose a penalized logistic regression method that automatically and simultaneously selects variables, groups categories, and estimates their coefficients by penalizing the [Formula: see text]-norm of both the coefficients and their differences. Hence, it encourages sparsity in the categories, i.e. grouping of the categories, and sparsity in the variables, i.e. variable selection. We apply the penalized logistic regression method to our data. The important variables are selected, with close categories simultaneously grouped, by penalized regression models with and without the interactions terms. The models are validated with 10-fold cross-validation. The receiver operating characteristic curves of the penalized regression models dominate the receiver operating characteristic curve of naive logistic regressions, indicating a superior discriminative performance. © The Author(s) 2013.
Using Logistic Regression To Predict the Probability of Debris Flows Occurring in Areas Recently Burned By Wildland Fires

USGS Publications Warehouse

Rupert, Michael G.; Cannon, Susan H.; Gartner, Joseph E.

2003-01-01

Logistic regression was used to predict the probability of debris flows occurring in areas recently burned by wildland fires. Multiple logistic regression is conceptually similar to multiple linear regression because statistical relations between one dependent variable and several independent variables are evaluated. In logistic regression, however, the dependent variable is transformed to a binary variable (debris flow did or did not occur), and the actual probability of the debris flow occurring is statistically modeled. Data from 399 basins located within 15 wildland fires that burned during 2000-2002 in Colorado, Idaho, Montana, and New Mexico were evaluated. More than 35 independent variables describing the burn severity, geology, land surface gradient, rainfall, and soil properties were evaluated. The models were developed as follows: (1) Basins that did and did not produce debris flows were delineated from National Elevation Data using a Geographic Information System (GIS). (2) Data describing the burn severity, geology, land surface gradient, rainfall, and soil properties were determined for each basin. These data were then downloaded to a statistics software package for analysis using logistic regression. (3) Relations between the occurrence/non-occurrence of debris flows and burn severity, geology, land surface gradient, rainfall, and soil properties were evaluated and several preliminary multivariate logistic regression models were constructed. All possible combinations of independent variables were evaluated to determine which combination produced the most effective model. The multivariate model that best predicted the occurrence of debris flows was selected. (4) The multivariate logistic regression model was entered into a GIS, and a map showing the probability of debris flows was constructed. The most effective model incorporates the percentage of each basin with slope greater than 30 percent, percentage of land burned at medium and high burn severity in each basin, particle size sorting, average storm intensity (millimeters per hour), soil organic matter content, soil permeability, and soil drainage. The results of this study demonstrate that logistic regression is a valuable tool for predicting the probability of debris flows occurring in recently-burned landscapes.
Prediction of unwanted pregnancies using logistic regression, probit regression and discriminant analysis

PubMed Central

Ebrahimzadeh, Farzad; Hajizadeh, Ebrahim; Vahabi, Nasim; Almasian, Mohammad; Bakhteyar, Katayoon

2015-01-01

Background: Unwanted pregnancy not intended by at least one of the parents has undesirable consequences for the family and the society. In the present study, three classification models were used and compared to predict unwanted pregnancies in an urban population. Methods: In this cross-sectional study, 887 pregnant mothers referring to health centers in Khorramabad, Iran, in 2012 were selected by the stratified and cluster sampling; relevant variables were measured and for prediction of unwanted pregnancy, logistic regression, discriminant analysis, and probit regression models and SPSS software version 21 were used. To compare these models, indicators such as sensitivity, specificity, the area under the ROC curve, and the percentage of correct predictions were used. Results: The prevalence of unwanted pregnancies was 25.3%. The logistic and probit regression models indicated that parity and pregnancy spacing, contraceptive methods, household income and number of living male children were related to unwanted pregnancy. The performance of the models based on the area under the ROC curve was 0.735, 0.733, and 0.680 for logistic regression, probit regression, and linear discriminant analysis, respectively. Conclusion: Given the relatively high prevalence of unwanted pregnancies in Khorramabad, it seems necessary to revise family planning programs. Despite the similar accuracy of the models, if the researcher is interested in the interpretability of the results, the use of the logistic regression model is recommended. PMID:26793655
Prediction of unwanted pregnancies using logistic regression, probit regression and discriminant analysis.

PubMed

Ebrahimzadeh, Farzad; Hajizadeh, Ebrahim; Vahabi, Nasim; Almasian, Mohammad; Bakhteyar, Katayoon

2015-01-01

Unwanted pregnancy not intended by at least one of the parents has undesirable consequences for the family and the society. In the present study, three classification models were used and compared to predict unwanted pregnancies in an urban population. In this cross-sectional study, 887 pregnant mothers referring to health centers in Khorramabad, Iran, in 2012 were selected by the stratified and cluster sampling; relevant variables were measured and for prediction of unwanted pregnancy, logistic regression, discriminant analysis, and probit regression models and SPSS software version 21 were used. To compare these models, indicators such as sensitivity, specificity, the area under the ROC curve, and the percentage of correct predictions were used. The prevalence of unwanted pregnancies was 25.3%. The logistic and probit regression models indicated that parity and pregnancy spacing, contraceptive methods, household income and number of living male children were related to unwanted pregnancy. The performance of the models based on the area under the ROC curve was 0.735, 0.733, and 0.680 for logistic regression, probit regression, and linear discriminant analysis, respectively. Given the relatively high prevalence of unwanted pregnancies in Khorramabad, it seems necessary to revise family planning programs. Despite the similar accuracy of the models, if the researcher is interested in the interpretability of the results, the use of the logistic regression model is recommended.

Obsessional personality features in employed Japanese adults with a lifetime history of depression: assessment by the Munich Personality Test (MPT).

PubMed

Sakado, K; Sakado, M; Seki, T; Kuwabara, H; Kojima, M; Sato, T; Someya, T

2001-06-01

Although a number of studies have reported on the association between obsessional personality features as measured by the Munich Personality Test (MPT) "Rigidity" scale and depression, there has been no examination of these relationships in a non-clinical sample. The dimensional scores on the MPT were compared between subjects with and without lifetime depression, using a sample of employed Japanese adults. The odds ratio for suffering from lifetime depression was estimated by multiple logistic regression analysis. To diagnose a lifetime history of depression, the Inventory to Diagnose Depression, Lifetime version (IDDL) was used. The subjects with lifetime depression scored significantly higher on the "Rigidity" scale than the subjects without lifetime depression. In our logistic regression analysis, three risk factors were identified as each independently increasing a person's risk for suffering from lifetime depression: higher levels of "Rigidity", being of the female gender, and suffering from current depressive symptoms. The MPT "Rigidity" scale is a sensitive measure of personality features that occur with depression.
Hypomagnesemia predicts postoperative biochemical hypocalcemia after thyroidectomy.

PubMed

Luo, Han; Yang, Hongliu; Zhao, Wanjun; Wei, Tao; Su, Anping; Wang, Bin; Zhu, Jingqiang

2017-05-25

To investigate the role of magnesium in biochemical and symptomatic hypocalcemia, a retrospective study was conducted. Less-than-total thyroidectomy patients were excluded from the final analysis. Identified the risk factors of biochemical and symptomatic hypocalcemia, and investigated the correlation by logistic regression and correlation test respectively. A total of 304 patients were included in the final analysis. General incidence of hypomagnesemia was 23.36%. Logistic regression showed that gender (female) (OR = 2.238, p = 0.015) and postoperative hypomagnesemia (OR = 2.010, p = 0.017) were independent risk factors for biochemical hypocalcemia. Both Pearson and partial correlation tests indicated there was indeed significant relation between calcium and magnesium. However, relative decreasing of iPTH (>70%) (6.691, p < 0.001) and hypocalcemia (2.222, p = 0.046) were identified as risk factors of symptomatic hypocalcemia. The difference remained significant even in normoparathyroidism patients. Postoperative hypomagnesemia was independent risk factor of biochemical hypocalcemia. Relative decline of iPTH was predominating in predicting symptomatic hypocalcemia.
Avoiding overstating the strength of forensic evidence: Shrunk likelihood ratios/Bayes factors.

PubMed

Morrison, Geoffrey Stewart; Poh, Norman

2018-05-01

When strength of forensic evidence is quantified using sample data and statistical models, a concern may be raised as to whether the output of a model overestimates the strength of evidence. This is particularly the case when the amount of sample data is small, and hence sampling variability is high. This concern is related to concern about precision. This paper describes, explores, and tests three procedures which shrink the value of the likelihood ratio or Bayes factor toward the neutral value of one. The procedures are: (1) a Bayesian procedure with uninformative priors, (2) use of empirical lower and upper bounds (ELUB), and (3) a novel form of regularized logistic regression. As a benchmark, they are compared with linear discriminant analysis, and in some instances with non-regularized logistic regression. The behaviours of the procedures are explored using Monte Carlo simulated data, and tested on real data from comparisons of voice recordings, face images, and glass fragments. Copyright © 2018 The Authors. Published by Elsevier B.V. All rights reserved.
Gender, Work, and HIV Risk: Determinants of Risky Sexual Behavior among Female Entertainment Workers in China

ERIC Educational Resources Information Center

Yang, Xiushi; Xia, Guomei

2006-01-01

We proposed to integrate cognitive and social factors in the study of unprotected commercial sex. Data from 159 female entertainment workers from 15 establishments in Shanghai who reported commercial sex in the month prior to interview were used to test the approach. Two-sample t tests and multivariate logistic regression were conducted to examine…
Predictors of course in obsessive-compulsive disorder: logistic regression versus Cox regression for recurrent events.

PubMed

Kempe, P T; van Oppen, P; de Haan, E; Twisk, J W R; Sluis, A; Smit, J H; van Dyck, R; van Balkom, A J L M

2007-09-01

Two methods for predicting remissions in obsessive-compulsive disorder (OCD) treatment are evaluated. Y-BOCS measurements of 88 patients with a primary OCD (DSM-III-R) diagnosis were performed over a 16-week treatment period, and during three follow-ups. Remission at any measurement was defined as a Y-BOCS score lower than thirteen combined with a reduction of seven points when compared with baseline. Logistic regression models were compared with a Cox regression for recurrent events model. Logistic regression yielded different models at different evaluation times. The recurrent events model remained stable when fewer measurements were used. Higher baseline levels of neuroticism and more severe OCD symptoms were associated with a lower chance of remission, early age of onset and more depressive symptoms with a higher chance. Choice of outcome time affects logistic regression prediction models. Recurrent events analysis uses all information on remissions and relapses. Short- and long-term predictors for OCD remission show overlap.
Use of logistic regression for modelling risk factors: with application to non-melanoma skin cancer

DOE Office of Scientific and Technical Information (OSTI.GOV)

Vitaliano, P.P.

Logistic regression was used to estimate the relative risk of basal and squamous skin cancer for such factors as cumulative lifetime solar exposure, age, complexion, and tannability. In previous reports, a subject's exposure was estimated indirectly, by latitude, or by the number of sun days in a subject's habitat. In contrast, these results are based on interview data gathered for each subject. A relatively new technique was used to estimate relative risk by controlling for confounding and testing for effect modification. A linear effect for the relative risk of cancer versus exposure was found. Tannability was shown to be amore » more important risk factor than complexion. This result is consistent with the work of Silverstone and Searle.« less
Estimating the exceedance probability of rain rate by logistic regression

NASA Technical Reports Server (NTRS)

Chiu, Long S.; Kedem, Benjamin

1990-01-01

Recent studies have shown that the fraction of an area with rain intensity above a fixed threshold is highly correlated with the area-averaged rain rate. To estimate the fractional rainy area, a logistic regression model, which estimates the conditional probability that rain rate over an area exceeds a fixed threshold given the values of related covariates, is developed. The problem of dependency in the data in the estimation procedure is bypassed by the method of partial likelihood. Analyses of simulated scanning multichannel microwave radiometer and observed electrically scanning microwave radiometer data during the Global Atlantic Tropical Experiment period show that the use of logistic regression in pixel classification is superior to multiple regression in predicting whether rain rate at each pixel exceeds a given threshold, even in the presence of noisy data. The potential of the logistic regression technique in satellite rain rate estimation is discussed.
Comparison of naïve Bayes and logistic regression for computer-aided diagnosis of breast masses using ultrasound imaging

NASA Astrophysics Data System (ADS)

Cary, Theodore W.; Cwanger, Alyssa; Venkatesh, Santosh S.; Conant, Emily F.; Sehgal, Chandra M.

2012-03-01

This study compares the performance of two proven but very different machine learners, Naïve Bayes and logistic regression, for differentiating malignant and benign breast masses using ultrasound imaging. Ultrasound images of 266 masses were analyzed quantitatively for shape, echogenicity, margin characteristics, and texture features. These features along with patient age, race, and mammographic BI-RADS category were used to train Naïve Bayes and logistic regression classifiers to diagnose lesions as malignant or benign. ROC analysis was performed using all of the features and using only a subset that maximized information gain. Performance was determined by the area under the ROC curve, Az, obtained from leave-one-out cross validation. Naïve Bayes showed significant variation (Az 0.733 +/- 0.035 to 0.840 +/- 0.029, P < 0.002) with the choice of features, but the performance of logistic regression was relatively unchanged under feature selection (Az 0.839 +/- 0.029 to 0.859 +/- 0.028, P = 0.605). Out of 34 features, a subset of 6 gave the highest information gain: brightness difference, margin sharpness, depth-to-width, mammographic BI-RADs, age, and race. The probabilities of malignancy determined by Naïve Bayes and logistic regression after feature selection showed significant correlation (R2= 0.87, P < 0.0001). The diagnostic performance of Naïve Bayes and logistic regression can be comparable, but logistic regression is more robust. Since probability of malignancy cannot be measured directly, high correlation between the probabilities derived from two basic but dissimilar models increases confidence in the predictive power of machine learning models for characterizing solid breast masses on ultrasound.
[Logistic regression model of noninvasive prediction for portal hypertensive gastropathy in patients with hepatitis B associated cirrhosis].

PubMed

Wang, Qingliang; Li, Xiaojie; Hu, Kunpeng; Zhao, Kun; Yang, Peisheng; Liu, Bo

2015-05-12

To explore the risk factors of portal hypertensive gastropathy (PHG) in patients with hepatitis B associated cirrhosis and establish a Logistic regression model of noninvasive prediction. The clinical data of 234 hospitalized patients with hepatitis B associated cirrhosis from March 2012 to March 2014 were analyzed retrospectively. The dependent variable was the occurrence of PHG while the independent variables were screened by binary Logistic analysis. Multivariate Logistic regression was used for further analysis of significant noninvasive independent variables. Logistic regression model was established and odds ratio was calculated for each factor. The accuracy, sensitivity and specificity of model were evaluated by the curve of receiver operating characteristic (ROC). According to univariate Logistic regression, the risk factors included hepatic dysfunction, albumin (ALB), bilirubin (TB), prothrombin time (PT), platelet (PLT), white blood cell (WBC), portal vein diameter, spleen index, splenic vein diameter, diameter ratio, PLT to spleen volume ratio, esophageal varices (EV) and gastric varices (GV). Multivariate analysis showed that hepatic dysfunction (X1), TB (X2), PLT (X3) and splenic vein diameter (X4) were the major occurring factors for PHG. The established regression model was Logit P=-2.667+2.186X1-2.167X2+0.725X3+0.976X4. The accuracy of model for PHG was 79.1% with a sensitivity of 77.2% and a specificity of 80.8%. Hepatic dysfunction, TB, PLT and splenic vein diameter are risk factors for PHG and the noninvasive predicted Logistic regression model was Logit P=-2.667+2.186X1-2.167X2+0.725X3+0.976X4.
Variable Selection in Logistic Regression.

DTIC Science & Technology

1987-06-01

23 %. AUTIOR(.) S. CONTRACT OR GRANT NUMBE Rf.i %Z. D. Bai, P. R. Krishnaiah and . C. Zhao F49620-85- C-0008 " PERFORMING ORGANIZATION NAME AND AOORESS...d I7 IOK-TK- d 7 -I0 7’ VARIABLE SELECTION IN LOGISTIC REGRESSION Z. D. Bai, P. R. Krishnaiah and L. C. Zhao Center for Multivariate Analysis...University of Pittsburgh Center for Multivariate Analysis University of Pittsburgh Y !I VARIABLE SELECTION IN LOGISTIC REGRESSION Z- 0. Bai, P. R. Krishnaiah
Multinomial Logistic Regression Predicted Probability Map To Visualize The Influence Of Socio-Economic Factors On Breast Cancer Occurrence in Southern Karnataka

NASA Astrophysics Data System (ADS)

Madhu, B.; Ashok, N. C.; Balasubramanian, S.

2014-11-01

Multinomial logistic regression analysis was used to develop statistical model that can predict the probability of breast cancer in Southern Karnataka using the breast cancer occurrence data during 2007-2011. Independent socio-economic variables describing the breast cancer occurrence like age, education, occupation, parity, type of family, health insurance coverage, residential locality and socioeconomic status of each case was obtained. The models were developed as follows: i) Spatial visualization of the Urban- rural distribution of breast cancer cases that were obtained from the Bharat Hospital and Institute of Oncology. ii) Socio-economic risk factors describing the breast cancer occurrences were complied for each case. These data were then analysed using multinomial logistic regression analysis in a SPSS statistical software and relations between the occurrence of breast cancer across the socio-economic status and the influence of other socio-economic variables were evaluated and multinomial logistic regression models were constructed. iii) the model that best predicted the occurrence of breast cancer were identified. This multivariate logistic regression model has been entered into a geographic information system and maps showing the predicted probability of breast cancer occurrence in Southern Karnataka was created. This study demonstrates that Multinomial logistic regression is a valuable tool for developing models that predict the probability of breast cancer Occurrence in Southern Karnataka.
Understanding logistic regression analysis.

PubMed

Sperandei, Sandro

2014-01-01

Logistic regression is used to obtain odds ratio in the presence of more than one explanatory variable. The procedure is quite similar to multiple linear regression, with the exception that the response variable is binomial. The result is the impact of each variable on the odds ratio of the observed event of interest. The main advantage is to avoid confounding effects by analyzing the association of all variables together. In this article, we explain the logistic regression procedure using examples to make it as simple as possible. After definition of the technique, the basic interpretation of the results is highlighted and then some special issues are discussed.
Discriminating between adaptive and carcinogenic liver hypertrophy in rat studies using logistic ridge regression analysis of toxicogenomic data: The mode of action and predictive models

DOE Office of Scientific and Technical Information (OSTI.GOV)

Liu, Shujie; Kawamoto, Taisuke; Morita, Osamu

Chemical exposure often results in liver hypertrophy in animal tests, characterized by increased liver weight, hepatocellular hypertrophy, and/or cell proliferation. While most of these changes are considered adaptive responses, there is concern that they may be associated with carcinogenesis. In this study, we have employed a toxicogenomic approach using a logistic ridge regression model to identify genes responsible for liver hypertrophy and hypertrophic hepatocarcinogenesis and to develop a predictive model for assessing hypertrophy-inducing compounds. Logistic regression models have previously been used in the quantification of epidemiological risk factors. DNA microarray data from the Toxicogenomics Project-Genomics Assisted Toxicity Evaluation System weremore » used to identify hypertrophy-related genes that are expressed differently in hypertrophy induced by carcinogens and non-carcinogens. Data were collected for 134 chemicals (72 non-hypertrophy-inducing chemicals, 27 hypertrophy-inducing non-carcinogenic chemicals, and 15 hypertrophy-inducing carcinogenic compounds). After applying logistic ridge regression analysis, 35 genes for liver hypertrophy (e.g., Acot1 and Abcc3) and 13 genes for hypertrophic hepatocarcinogenesis (e.g., Asns and Gpx2) were selected. The predictive models built using these genes were 94.8% and 82.7% accurate, respectively. Pathway analysis of the genes indicates that, aside from a xenobiotic metabolism-related pathway as an adaptive response for liver hypertrophy, amino acid biosynthesis and oxidative responses appear to be involved in hypertrophic hepatocarcinogenesis. Early detection and toxicogenomic characterization of liver hypertrophy using our models may be useful for predicting carcinogenesis. In addition, the identified genes provide novel insight into discrimination between adverse hypertrophy associated with carcinogenesis and adaptive hypertrophy in risk assessment. - Highlights: • Hypertrophy (H) and hypertrophic carcinogenesis (C) were studied by toxicogenomics. • Important genes for H and C were selected by logistic ridge regression analysis. • Amino acid biosynthesis and oxidative responses may be involved in C. • Predictive models for H and C provided 94.8% and 82.7% accuracy, respectively. • The identified genes could be useful for assessment of liver hypertrophy.« less
Prediction of Emergency Department Hospital Admission Based on Natural Language Processing and Neural Networks.

PubMed

Zhang, Xingyu; Kim, Joyce; Patzer, Rachel E; Pitts, Stephen R; Patzer, Aaron; Schrager, Justin D

2017-10-26

To describe and compare logistic regression and neural network modeling strategies to predict hospital admission or transfer following initial presentation to Emergency Department (ED) triage with and without the addition of natural language processing elements. Using data from the National Hospital Ambulatory Medical Care Survey (NHAMCS), a cross-sectional probability sample of United States EDs from 2012 and 2013 survey years, we developed several predictive models with the outcome being admission to the hospital or transfer vs. discharge home. We included patient characteristics immediately available after the patient has presented to the ED and undergone a triage process. We used this information to construct logistic regression (LR) and multilayer neural network models (MLNN) which included natural language processing (NLP) and principal component analysis from the patient's reason for visit. Ten-fold cross validation was used to test the predictive capacity of each model and receiver operating curves (AUC) were then calculated for each model. Of the 47,200 ED visits from 642 hospitals, 6,335 (13.42%) resulted in hospital admission (or transfer). A total of 48 principal components were extracted by NLP from the reason for visit fields, which explained 75% of the overall variance for hospitalization. In the model including only structured variables, the AUC was 0.824 (95% CI 0.818-0.830) for logistic regression and 0.823 (95% CI 0.817-0.829) for MLNN. Models including only free-text information generated AUC of 0.742 (95% CI 0.731- 0.753) for logistic regression and 0.753 (95% CI 0.742-0.764) for MLNN. When both structured variables and free text variables were included, the AUC reached 0.846 (95% CI 0.839-0.853) for logistic regression and 0.844 (95% CI 0.836-0.852) for MLNN. The predictive accuracy of hospital admission or transfer for patients who presented to ED triage overall was good, and was improved with the inclusion of free text data from a patient's reason for visit regardless of modeling approach. Natural language processing and neural networks that incorporate patient-reported outcome free text may increase predictive accuracy for hospital admission.
Using a binary logistic regression method and GIS for evaluating and mapping the groundwater spring potential in the Sultan Mountains (Aksehir, Turkey)

NASA Astrophysics Data System (ADS)

Ozdemir, Adnan

2011-07-01

SummaryThe purpose of this study is to produce a groundwater spring potential map of the Sultan Mountains in central Turkey, based on a logistic regression method within a Geographic Information System (GIS) environment. Using field surveys, the locations of the springs (440 springs) were determined in the study area. In this study, 17 spring-related factors were used in the analysis: geology, relative permeability, land use/land cover, precipitation, elevation, slope, aspect, total curvature, plan curvature, profile curvature, wetness index, stream power index, sediment transport capacity index, distance to drainage, distance to fault, drainage density, and fault density map. The coefficients of the predictor variables were estimated using binary logistic regression analysis and were used to calculate the groundwater spring potential for the entire study area. The accuracy of the final spring potential map was evaluated based on the observed springs. The accuracy of the model was evaluated by calculating the relative operating characteristics. The area value of the relative operating characteristic curve model was found to be 0.82. These results indicate that the model is a good estimator of the spring potential in the study area. The spring potential map shows that the areas of very low, low, moderate and high groundwater spring potential classes are 105.586 km 2 (28.99%), 74.271 km 2 (19.906%), 101.203 km 2 (27.14%), and 90.05 km 2 (24.671%), respectively. The interpretations of the potential map showed that stream power index, relative permeability of lithologies, geology, elevation, aspect, wetness index, plan curvature, and drainage density play major roles in spring occurrence and distribution in the Sultan Mountains. The logistic regression approach has not yet been used to delineate groundwater potential zones. In this study, the logistic regression method was used to locate potential zones for groundwater springs in the Sultan Mountains. The evolved model was found to be in strong agreement with the available groundwater spring test data. Hence, this method can be used routinely in groundwater exploration under favourable conditions.
An Alternative Flight Software Trigger Paradigm: Applying Multivariate Logistic Regression to Sense Trigger Conditions Using Inaccurate or Scarce Information

NASA Technical Reports Server (NTRS)

Smith, Kelly M.; Gay, Robert S.; Stachowiak, Susan J.

2013-01-01

In late 2014, NASA will fly the Orion capsule on a Delta IV-Heavy rocket for the Exploration Flight Test-1 (EFT-1) mission. For EFT-1, the Orion capsule will be flying with a new GPS receiver and new navigation software. Given the experimental nature of the flight, the flight software must be robust to the loss of GPS measurements. Once the high-speed entry is complete, the drogue parachutes must be deployed within the proper conditions to stabilize the vehicle prior to deploying the main parachutes. When GPS is available in nominal operations, the vehicle will deploy the drogue parachutes based on an altitude trigger. However, when GPS is unavailable, the navigated altitude errors become excessively large, driving the need for a backup barometric altimeter to improve altitude knowledge. In order to increase overall robustness, the vehicle also has an alternate method of triggering the parachute deployment sequence based on planet-relative velocity if both the GPS and the barometric altimeter fail. However, this backup trigger results in large altitude errors relative to the targeted altitude. Motivated by this challenge, this paper demonstrates how logistic regression may be employed to semi-automatically generate robust triggers based on statistical analysis. Logistic regression is used as a ground processor pre-flight to develop a statistical classifier. The classifier would then be implemented in flight software and executed in real-time. This technique offers improved performance even in the face of highly inaccurate measurements. Although the logistic regression-based trigger approach will not be implemented within EFT-1 flight software, the methodology can be carried forward for future missions and vehicles.
An Alternative Flight Software Paradigm: Applying Multivariate Logistic Regression to Sense Trigger Conditions using Inaccurate or Scarce Information

NASA Technical Reports Server (NTRS)

Smith, Kelly; Gay, Robert; Stachowiak, Susan

2013-01-01

In late 2014, NASA will fly the Orion capsule on a Delta IV-Heavy rocket for the Exploration Flight Test-1 (EFT-1) mission. For EFT-1, the Orion capsule will be flying with a new GPS receiver and new navigation software. Given the experimental nature of the flight, the flight software must be robust to the loss of GPS measurements. Once the high-speed entry is complete, the drogue parachutes must be deployed within the proper conditions to stabilize the vehicle prior to deploying the main parachutes. When GPS is available in nominal operations, the vehicle will deploy the drogue parachutes based on an altitude trigger. However, when GPS is unavailable, the navigated altitude errors become excessively large, driving the need for a backup barometric altimeter to improve altitude knowledge. In order to increase overall robustness, the vehicle also has an alternate method of triggering the parachute deployment sequence based on planet-relative velocity if both the GPS and the barometric altimeter fail. However, this backup trigger results in large altitude errors relative to the targeted altitude. Motivated by this challenge, this paper demonstrates how logistic regression may be employed to semi-automatically generate robust triggers based on statistical analysis. Logistic regression is used as a ground processor pre-flight to develop a statistical classifier. The classifier would then be implemented in flight software and executed in real-time. This technique offers improved performance even in the face of highly inaccurate measurements. Although the logistic regression-based trigger approach will not be implemented within EFT-1 flight software, the methodology can be carried forward for future missions and vehicles
An Alternative Flight Software Trigger Paradigm: Applying Multivariate Logistic Regression to Sense Trigger Conditions using Inaccurate or Scarce Information

NASA Technical Reports Server (NTRS)

Smith, Kelly M.; Gay, Robert S.; Stachowiak, Susan J.

2013-01-01

In late 2014, NASA will fly the Orion capsule on a Delta IV-Heavy rocket for the Exploration Flight Test-1 (EFT-1) mission. For EFT-1, the Orion capsule will be flying with a new GPS receiver and new navigation software. Given the experimental nature of the flight, the flight software must be robust to the loss of GPS measurements. Once the high-speed entry is complete, the drogue parachutes must be deployed within the proper conditions to stabilize the vehicle prior to deploying the main parachutes. When GPS is available in nominal operations, the vehicle will deploy the drogue parachutes based on an altitude trigger. However, when GPS is unavailable, the navigated altitude errors become excessively large, driving the need for a backup barometric altimeter. In order to increase overall robustness, the vehicle also has an alternate method of triggering the drogue parachute deployment based on planet-relative velocity if both the GPS and the barometric altimeter fail. However, this velocity-based trigger results in large altitude errors relative to the targeted altitude. Motivated by this challenge, this paper demonstrates how logistic regression may be employed to automatically generate robust triggers based on statistical analysis. Logistic regression is used as a ground processor pre-flight to develop a classifier. The classifier would then be implemented in flight software and executed in real-time. This technique offers excellent performance even in the face of highly inaccurate measurements. Although the logistic regression-based trigger approach will not be implemented within EFT-1 flight software, the methodology can be carried forward for future missions and vehicles.
Discriminating between adaptive and carcinogenic liver hypertrophy in rat studies using logistic ridge regression analysis of toxicogenomic data: The mode of action and predictive models.

PubMed

Liu, Shujie; Kawamoto, Taisuke; Morita, Osamu; Yoshinari, Kouichi; Honda, Hiroshi

2017-03-01

Chemical exposure often results in liver hypertrophy in animal tests, characterized by increased liver weight, hepatocellular hypertrophy, and/or cell proliferation. While most of these changes are considered adaptive responses, there is concern that they may be associated with carcinogenesis. In this study, we have employed a toxicogenomic approach using a logistic ridge regression model to identify genes responsible for liver hypertrophy and hypertrophic hepatocarcinogenesis and to develop a predictive model for assessing hypertrophy-inducing compounds. Logistic regression models have previously been used in the quantification of epidemiological risk factors. DNA microarray data from the Toxicogenomics Project-Genomics Assisted Toxicity Evaluation System were used to identify hypertrophy-related genes that are expressed differently in hypertrophy induced by carcinogens and non-carcinogens. Data were collected for 134 chemicals (72 non-hypertrophy-inducing chemicals, 27 hypertrophy-inducing non-carcinogenic chemicals, and 15 hypertrophy-inducing carcinogenic compounds). After applying logistic ridge regression analysis, 35 genes for liver hypertrophy (e.g., Acot1 and Abcc3) and 13 genes for hypertrophic hepatocarcinogenesis (e.g., Asns and Gpx2) were selected. The predictive models built using these genes were 94.8% and 82.7% accurate, respectively. Pathway analysis of the genes indicates that, aside from a xenobiotic metabolism-related pathway as an adaptive response for liver hypertrophy, amino acid biosynthesis and oxidative responses appear to be involved in hypertrophic hepatocarcinogenesis. Early detection and toxicogenomic characterization of liver hypertrophy using our models may be useful for predicting carcinogenesis. In addition, the identified genes provide novel insight into discrimination between adverse hypertrophy associated with carcinogenesis and adaptive hypertrophy in risk assessment. Copyright © 2017 Elsevier Inc. All rights reserved.
Comparing Methodologies for Developing an Early Warning System: Classification and Regression Tree Model versus Logistic Regression. REL 2015-077

ERIC Educational Resources Information Center

Koon, Sharon; Petscher, Yaacov

2015-01-01

The purpose of this report was to explicate the use of logistic regression and classification and regression tree (CART) analysis in the development of early warning systems. It was motivated by state education leaders' interest in maintaining high classification accuracy while simultaneously improving practitioner understanding of the rules by…

Understanding Civic Identity in College

ERIC Educational Resources Information Center

Weerts, David J.; Cabrera, Alberto F.

2015-01-01

Past literature has examined ways in which college students adopt civic identities. However, little is known about characteristics of students that vary in their expression of these identities. Drawing on data from American College Testing (ACT), this study employs multinomial logistic regression to understand attributes of students who vary in…
Secure Logistic Regression Based on Homomorphic Encryption: Design and Evaluation

PubMed Central

Song, Yongsoo; Wang, Shuang; Xia, Yuhou; Jiang, Xiaoqian

2018-01-01

Background Learning a model without accessing raw data has been an intriguing idea to security and machine learning researchers for years. In an ideal setting, we want to encrypt sensitive data to store them on a commercial cloud and run certain analyses without ever decrypting the data to preserve privacy. Homomorphic encryption technique is a promising candidate for secure data outsourcing, but it is a very challenging task to support real-world machine learning tasks. Existing frameworks can only handle simplified cases with low-degree polynomials such as linear means classifier and linear discriminative analysis. Objective The goal of this study is to provide a practical support to the mainstream learning models (eg, logistic regression). Methods We adapted a novel homomorphic encryption scheme optimized for real numbers computation. We devised (1) the least squares approximation of the logistic function for accuracy and efficiency (ie, reduce computation cost) and (2) new packing and parallelization techniques. Results Using real-world datasets, we evaluated the performance of our model and demonstrated its feasibility in speed and memory consumption. For example, it took approximately 116 minutes to obtain the training model from the homomorphically encrypted Edinburgh dataset. In addition, it gives fairly accurate predictions on the testing dataset. Conclusions We present the first homomorphically encrypted logistic regression outsourcing model based on the critical observation that the precision loss of classification models is sufficiently small so that the decision plan stays still. PMID:29666041
Dynamic Network Logistic Regression: A Logistic Choice Analysis of Inter- and Intra-Group Blog Citation Dynamics in the 2004 US Presidential Election

PubMed Central

2013-01-01

Methods for analysis of network dynamics have seen great progress in the past decade. This article shows how Dynamic Network Logistic Regression techniques (a special case of the Temporal Exponential Random Graph Models) can be used to implement decision theoretic models for network dynamics in a panel data context. We also provide practical heuristics for model building and assessment. We illustrate the power of these techniques by applying them to a dynamic blog network sampled during the 2004 US presidential election cycle. This is a particularly interesting case because it marks the debut of Internet-based media such as blogs and social networking web sites as institutionally recognized features of the American political landscape. Using a longitudinal sample of all Democratic National Convention/Republican National Convention–designated blog citation networks, we are able to test the influence of various strategic, institutional, and balance-theoretic mechanisms as well as exogenous factors such as seasonality and political events on the propensity of blogs to cite one another over time. Using a combination of deviance-based model selection criteria and simulation-based model adequacy tests, we identify the combination of processes that best characterizes the choice behavior of the contending blogs. PMID:24143060
Item Response Theory Modeling of the Philadelphia Naming Test.

PubMed

Fergadiotis, Gerasimos; Kellough, Stacey; Hula, William D

2015-06-01

In this study, we investigated the fit of the Philadelphia Naming Test (PNT; Roach, Schwartz, Martin, Grewal, & Brecher, 1996) to an item-response-theory measurement model, estimated the precision of the resulting scores and item parameters, and provided a theoretical rationale for the interpretation of PNT overall scores by relating explanatory variables to item difficulty. This article describes the statistical model underlying the computer adaptive PNT presented in a companion article (Hula, Kellough, & Fergadiotis, 2015). Using archival data, we evaluated the fit of the PNT to 1- and 2-parameter logistic models and examined the precision of the resulting parameter estimates. We regressed the item difficulty estimates on three predictor variables: word length, age of acquisition, and contextual diversity. The 2-parameter logistic model demonstrated marginally better fit, but the fit of the 1-parameter logistic model was adequate. Precision was excellent for both person ability and item difficulty estimates. Word length, age of acquisition, and contextual diversity all independently contributed to variance in item difficulty. Item-response-theory methods can be productively used to analyze and quantify anomia severity in aphasia. Regression of item difficulty on lexical variables supported the validity of the PNT and interpretation of anomia severity scores in the context of current word-finding models.
Logistic and Multiple Regression: A Two-Pronged Approach to Accurately Estimate Cost Growth in Major DoD Weapon Systems

DTIC Science & Technology

2004-03-01

Breusch - Pagan test for constant variance of the residuals. Using Microsoft Excel® we calculate a p-value of 0.841237. This high p-value, which is above...our alpha of 0.05, indicates that our residuals indeed pass the Breusch - Pagan test for constant variance. In addition to the assumption tests , we...Wilk Test for Normality – Support (Reduced) Model (OLS) Finally, we perform a Breusch - Pagan test for constant variance of the residuals. Using
HIV-related stigma, social norms, and HIV testing in Soweto and Vulindlela, South Africa: National Institutes of Mental Health Project Accept (HPTN 043).

PubMed

Young, Sean D; Hlavka, Zdenek; Modiba, Precious; Gray, Glenda; Van Rooyen, Heidi; Richter, Linda; Szekeres, Greg; Coates, Thomas

2010-12-15

HIV testing is necessary to curb the increasing epidemic. However, HIV-related stigma and perceptions of low likelihood of societal HIV testing may reduce testing rates. This study aimed to explore this association in South Africa, where HIV rates are extraordinarily high. Data were taken from the Soweto and Vulindlela, South African sites of Project Accept, a multinational HIV prevention trial. Self-reported HIV testing, stigma, and social norms items were used to study the relationship between HIV testing, stigma, and perceptions about societal testing rates. The stigma items were broken into 3 factors: negative attitudes, negative perceptions about people living with HIV, and perceptions of fair treatment for people living with HIV (equity). Results from a univariate logistic regression suggest that history of HIV testing was associated with decreased negative attitudes about people living with HIV/AIDS, increased perceptions that people living with HIV/AIDS experience discrimination, and increased perceptions that people with HIV should be treated equitably. Results from a multivariate logistic regression confirm these effects and suggest that these differences vary according to sex and age. Compared with people who had never tested for HIV, those who had previously tested were more likely to believe that the majority of people have tested for HIV. Data suggest that interventions designed to increase HIV testing in South Africa should address stigma and perceptions of societal testing.
A simple measure of cognitive reserve is relevant for cognitive performance in MS patients.

PubMed

Della Corte, Marida; Santangelo, Gabriella; Bisecco, Alvino; Sacco, Rosaria; Siciliano, Mattia; d'Ambrosio, Alessandro; Docimo, Renato; Cuomo, Teresa; Lavorgna, Luigi; Bonavita, Simona; Tedeschi, Gioacchino; Gallo, Antonio

2018-05-04

Cognitive reserve (CR) contributes to preserve cognition despite brain damage. This theory has been applied to multiple sclerosis (MS) to explain the partial relationship between cognition and MRI markers of brain pathology. Our aim was to determine the relationship between two measures of CR and cognition in MS. One hundred and forty-seven MS patients were enrolled. Cognition was assessed using the Rao's Brief Repeatable Battery and the Stroop Test. CR was measured as the vocabulary subtest of the WAIS-R score (VOC) and the number of years of formal education (EDU). Regression analysis included raw score data on each neuropsychological (NP) test as dependent variables and demographic/clinical parameters, VOC, and EDU as independent predictors. A binary logistic regression analysis including clinical/CR parameters as covariates and absence/presence of cognitive deficits as dependent variables was performed too. VOC, but not EDU, was strongly correlated with performances at all ten NP tests. EDU was correlated with executive performances. The binary logistic regression showed that only the Expanded Disability Status Scale (EDSS) and VOC were independently correlated with the presence/absence of CD. The lower the VOC and/or the higher the EDSS, the higher the frequency of CD. In conclusion, our study supports the relevance of CR in subtending cognitive performances and the presence of CD in MS patients.
Using Multiple and Logistic Regression to Estimate the Median WillCost and Probability of Cost and Schedule Overrun for Program Managers

DTIC Science & Technology

2017-03-23

PUBLIC RELEASE; DISTRIBUTION UNLIMITED Using Multiple and Logistic Regression to Estimate the Median Will- Cost and Probability of Cost and... Cost and Probability of Cost and Schedule Overrun for Program Managers Ryan C. Trudelle Follow this and additional works at: https://scholar.afit.edu...afit.edu. Recommended Citation Trudelle, Ryan C., "Using Multiple and Logistic Regression to Estimate the Median Will- Cost and Probability of Cost and
Expression of Proteins Involved in Epithelial-Mesenchymal Transition as Predictors of Metastasis and Survival in Breast Cancer Patients

DTIC Science & Technology

2013-11-01

Ptrend 0.78 0.62 0.75 Unconditional logistic regression was used to estimate odds ratios (OR) and 95 % confidence intervals (CI) for risk of node...Ptrend 0.71 0.67 Unconditional logistic regression was used to estimate odds ratios (OR) and 95 % confidence intervals (CI) for risk of high-grade tumors... logistic regression was used to estimate odds ratios (OR) and 95 % confidence intervals (CI) for the associations between each of the seven SNPs and
Risk Factors for Venous Thromboembolism After Spine Surgery

PubMed Central

Tominaga, Hiroyuki; Setoguchi, Takao; Tanabe, Fumito; Kawamura, Ichiro; Tsuneyoshi, Yasuhiro; Kawabata, Naoya; Nagano, Satoshi; Abematsu, Masahiko; Yamamoto, Takuya; Yone, Kazunori; Komiya, Setsuro

2015-01-01

Abstract The efficacy and safety of chemical prophylaxis to prevent the development of deep venous thrombosis (DVT) or pulmonary embolism (PE) following spine surgery are controversial because of the possibility of epidural hematoma formation. Postoperative venous thromboembolism (VTE) after spine surgery occurs at a frequency similar to that seen after joint operations, so it is important to identify the risk factors for VTE formation following spine surgery. We therefore retrospectively studied data from patients who had undergone spinal surgery and developed postoperative VTE to identify those risk factors. We conducted a retrospective clinical study with logistic regression analysis of a group of 80 patients who had undergone spine surgery at our institution from June 2012 to August 2013. All patients had been screened by ultrasonography for DVT in the lower extremities. Parameters of the patients with VTE were compared with those without VTE using the Mann–Whitney U-test and Fisher exact probability test. Logistic regression analysis was used to analyze the risk factors associated with VTE. A value of P < 0.05 was used to denote statistical significance. The prevalence of VTE was 25.0% (20/80 patients). One patient had sensed some incongruity in the chest area, but the vital signs of all patients were stable. VTEs had developed in the pulmonary artery in one patient, in the superficial femoral vein in one patient, in the popliteal vein in two patients, and in the soleal vein in 18 patients. The Mann–Whitney U-test and Fisher exact probability test showed that, except for preoperative walking disability, none of the parameters showed a significant difference between patients with and without VTE. Risk factors identified in the multivariate logistic regression analysis were preoperative walking disability and age. The prevalence of VTE after spine surgery was relatively high. The most important risk factor for developing postoperative VTE was preoperative walking disability. Gait training during the early postoperative period is required to prevent VTE. PMID:25654385
Comparison of Xenon-Enhanced Area-Detector CT and Krypton Ventilation SPECT/CT for Assessment of Pulmonary Functional Loss and Disease Severity in Smokers.

PubMed

Ohno, Yoshiharu; Fujisawa, Yasuko; Takenaka, Daisuke; Kaminaga, Shigeo; Seki, Shinichiro; Sugihara, Naoki; Yoshikawa, Takeshi

2018-02-01

The objective of this study was to compare the capability of xenon-enhanced area-detector CT (ADCT) performed with a subtraction technique and coregistered 81m Kr-ventilation SPECT/CT for the assessment of pulmonary functional loss and disease severity in smokers. Forty-six consecutive smokers (32 men and 14 women; mean age, 67.0 years) underwent prospective unenhanced and xenon-enhanced ADCT, 81m Kr-ventilation SPECT/CT, and pulmonary function tests. Disease severity was evaluated according to the Global Initiative for Chronic Obstructive Lung Disease (GOLD) classification. CT-based functional lung volume (FLV), the percentage of wall area to total airway area (WA%), and ventilated FLV on xenon-enhanced ADCT and SPECT/CT were calculated for each smoker. All indexes were correlated with percentage of forced expiratory volume in 1 second (%FEV 1 ) using step-wise regression analyses, and univariate and multivariate logistic regression analyses were performed. In addition, the diagnostic accuracy of the proposed model was compared with that of each radiologic index by means of McNemar analysis. Multivariate logistic regression showed that %FEV 1 was significantly affected (r = 0.77, r 2 = 0.59) by two factors: the first factor, ventilated FLV on xenon-enhanced ADCT (p < 0.0001); and the second factor, WA% (p = 0.004). Univariate logistic regression analyses indicated that all indexes significantly affected GOLD classification (p < 0.05). Multivariate logistic regression analyses revealed that ventilated FLV on xenon-enhanced ADCT and CT-based FLV significantly influenced GOLD classification (p < 0.0001). The diagnostic accuracy of the proposed model was significantly higher than that of ventilated FLV on SPECT/CT (p = 0.03) and WA% (p = 0.008). Xenon-enhanced ADCT is more effective than 81m Kr-ventilation SPECT/CT for the assessment of pulmonary functional loss and disease severity.
The alarming problems of confounding equivalence using logistic regression models in the perspective of causal diagrams.

PubMed

Yu, Yuanyuan; Li, Hongkai; Sun, Xiaoru; Su, Ping; Wang, Tingting; Liu, Yi; Yuan, Zhongshang; Liu, Yanxun; Xue, Fuzhong

2017-12-28

Confounders can produce spurious associations between exposure and outcome in observational studies. For majority of epidemiologists, adjusting for confounders using logistic regression model is their habitual method, though it has some problems in accuracy and precision. It is, therefore, important to highlight the problems of logistic regression and search the alternative method. Four causal diagram models were defined to summarize confounding equivalence. Both theoretical proofs and simulation studies were performed to verify whether conditioning on different confounding equivalence sets had the same bias-reducing potential and then to select the optimum adjusting strategy, in which logistic regression model and inverse probability weighting based marginal structural model (IPW-based-MSM) were compared. The "do-calculus" was used to calculate the true causal effect of exposure on outcome, then the bias and standard error were used to evaluate the performances of different strategies. Adjusting for different sets of confounding equivalence, as judged by identical Markov boundaries, produced different bias-reducing potential in the logistic regression model. For the sets satisfied G-admissibility, adjusting for the set including all the confounders reduced the equivalent bias to the one containing the parent nodes of the outcome, while the bias after adjusting for the parent nodes of exposure was not equivalent to them. In addition, all causal effect estimations through logistic regression were biased, although the estimation after adjusting for the parent nodes of exposure was nearest to the true causal effect. However, conditioning on different confounding equivalence sets had the same bias-reducing potential under IPW-based-MSM. Compared with logistic regression, the IPW-based-MSM could obtain unbiased causal effect estimation when the adjusted confounders satisfied G-admissibility and the optimal strategy was to adjust for the parent nodes of outcome, which obtained the highest precision. All adjustment strategies through logistic regression were biased for causal effect estimation, while IPW-based-MSM could always obtain unbiased estimation when the adjusted set satisfied G-admissibility. Thus, IPW-based-MSM was recommended to adjust for confounders set.
Use and interpretation of logistic regression in habitat-selection studies

USGS Publications Warehouse

Keating, Kim A.; Cherry, Steve

2004-01-01

Logistic regression is an important tool for wildlife habitat-selection studies, but the method frequently has been misapplied due to an inadequate understanding of the logistic model, its interpretation, and the influence of sampling design. To promote better use of this method, we review its application and interpretation under 3 sampling designs: random, case-control, and use-availability. Logistic regression is appropriate for habitat use-nonuse studies employing random sampling and can be used to directly model the conditional probability of use in such cases. Logistic regression also is appropriate for studies employing case-control sampling designs, but careful attention is required to interpret results correctly. Unless bias can be estimated or probability of use is small for all habitats, results of case-control studies should be interpreted as odds ratios, rather than probability of use or relative probability of use. When data are gathered under a use-availability design, logistic regression can be used to estimate approximate odds ratios if probability of use is small, at least on average. More generally, however, logistic regression is inappropriate for modeling habitat selection in use-availability studies. In particular, using logistic regression to fit the exponential model of Manly et al. (2002:100) does not guarantee maximum-likelihood estimates, valid probabilities, or valid likelihoods. We show that the resource selection function (RSF) commonly used for the exponential model is proportional to a logistic discriminant function. Thus, it may be used to rank habitats with respect to probability of use and to identify important habitat characteristics or their surrogates, but it is not guaranteed to be proportional to probability of use. Other problems associated with the exponential model also are discussed. We describe an alternative model based on Lancaster and Imbens (1996) that offers a method for estimating conditional probability of use in use-availability studies. Although promising, this model fails to converge to a unique solution in some important situations. Further work is needed to obtain a robust method that is broadly applicable to use-availability studies.
An Attempt at Quantifying Factors that Affect Efficiency in the Management of Solid Waste Produced by Commercial Businesses in the City of Tshwane, South Africa

PubMed Central

Worku, Yohannes; Muchie, Mammo

2012-01-01

Objective. The objective was to investigate factors that affect the efficient management of solid waste produced by commercial businesses operating in the city of Pretoria, South Africa. Methods. Data was gathered from 1,034 businesses. Efficiency in solid waste management was assessed by using a structural time-based model designed for evaluating efficiency as a function of the length of time required to manage waste. Data analysis was performed using statistical procedures such as frequency tables, Pearson's chi-square tests of association, and binary logistic regression analysis. Odds ratios estimated from logistic regression analysis were used for identifying key factors that affect efficiency in the proper disposal of waste. Results. The study showed that 857 of the 1,034 businesses selected for the study (83%) were found to be efficient enough with regards to the proper collection and disposal of solid waste. Based on odds ratios estimated from binary logistic regression analysis, efficiency in the proper management of solid waste was significantly influenced by 4 predictor variables. These 4 influential predictor variables are lack of adherence to waste management regulations, wrong perception, failure to provide customers with enough trash cans, and operation of businesses by employed managers, in a decreasing order of importance. PMID:23209483
Comparison of the Relationship between Women' Empowerment and Fertility between Single-child and Multi-child Families

PubMed Central

Saberi, Tahereh; Ehsanpour, Soheila; Mahaki, Behzad; Kohan, Shahnaz

2018-01-01

Background: The reduction in fertility and increase in the number of single-child families in Iran will result in an increased risk of population aging. One of the factors affecting fertility is women's empowerment. This study aimed to evaluate the relationship between women's empowerment and fertility in single-child and multi-child families. Materials and Methods: This case-control study was conducted among 350 women (120 who had only 1 child as case group and 230 who had 2 or more children as control group) of 15–49 years of age in Isfahan, Iran, in 2016. For data collection, a 2-part questionnaire was designed. Data were analyzed using independent t-test, Chi-square test, and logistic regression analysis. Results: The difference between average scores of women's empowerment in the case group 54.08 (9.88) and control group 51.47 (8.57) was significant (p = 0.002). Simple logistic regression analysis showed that under diploma education, compared to postgraduate education, (OR = 0.21, p = 0.001) and being a housewife, compared to being employed, (OR = 0.45, p = 0.004) decreased the odds of having only 1 child. Multiple logistic regression results showed that the relationship between women's empowerment and fertility was not significant (p = 0.265). Conclusions: Although women in single-child families were more empowered, this was not the main reason for their preference to have only 1 child. In fact, educated and employed women postpone marriage and childbearing and limit fertility to only 1 child despite their desire. PMID:29628961
Modeling Governance KB with CATPCA to Overcome Multicollinearity in the Logistic Regression

NASA Astrophysics Data System (ADS)

Khikmah, L.; Wijayanto, H.; Syafitri, U. D.

2017-04-01

The problem often encounters in logistic regression modeling are multicollinearity problems. Data that have multicollinearity between explanatory variables with the result in the estimation of parameters to be bias. Besides, the multicollinearity will result in error in the classification. In general, to overcome multicollinearity in regression used stepwise regression. They are also another method to overcome multicollinearity which involves all variable for prediction. That is Principal Component Analysis (PCA). However, classical PCA in only for numeric data. Its data are categorical, one method to solve the problems is Categorical Principal Component Analysis (CATPCA). Data were used in this research were a part of data Demographic and Population Survey Indonesia (IDHS) 2012. This research focuses on the characteristic of women of using the contraceptive methods. Classification results evaluated using Area Under Curve (AUC) values. The higher the AUC value, the better. Based on AUC values, the classification of the contraceptive method using stepwise method (58.66%) is better than the logistic regression model (57.39%) and CATPCA (57.39%). Evaluation of the results of logistic regression using sensitivity, shows the opposite where CATPCA method (99.79%) is better than logistic regression method (92.43%) and stepwise (92.05%). Therefore in this study focuses on major class classification (using a contraceptive method), then the selected model is CATPCA because it can raise the level of the major class model accuracy.
Utilisation of cancer screening services by disabled women in Chile

PubMed Central

Rotarou, Elena S.

2017-01-01

Background Research has shown that women with disabilities face additional challenges in accessing and using healthcare services compared to non-disabled women. However, relatively little is known about the utilisation of cancer screening services for women with disabilities. This study addresses this gap by examining the utilisation of the Papanicolaou test and mammography for disabled women in Chile. Methods We used cross-sectional data, taken from a 2015 nationally-representative survey. Initially, we employed logistic regressions to test for differences in utilisation rates for the Papanicolaou test (66,281 observations) and the mammogram (35,294 observations) between disabled and non-disabled women. Next, logistic regressions were used to investigate the demographic, socioeconomic, and health-related factors affecting utilisation rates for cancer screening services for disabled women (sample sizes: 5,823 observations for the Papanicolaou test and 5,731 observations for the mammogram). Results Disabled women were less likely to undergo screening tests than non-disabled women. For the Papanicolaou test and mammography, the multivariable regression models showed that living in rural areas, having higher education, being affiliated with a private health insurance company, giving a good health self-assessment score, and being under medical treatment for other illnesses were associated with higher utilisation rates. On the other hand, being single, inactive with regard to employment, and having a better income were linked with lower utilisation. While utilisation rates for both disabled and non-disabled women have increased since 2006, the utilisation disparity has slightly increased. Conclusions This study shows the influence of various factors in the utilisation rates of preventive cancer screening services for disabled women. To develop effective initiatives targeting inequalities in the utilisation of cancer screening tests, it is important to move beyond an exclusively single-disease approach and acknowledge the complexity of the patient population. PMID:28459874
Utilisation of cancer screening services by disabled women in Chile.

PubMed

Sakellariou, Dikaios; Rotarou, Elena S

2017-01-01

Research has shown that women with disabilities face additional challenges in accessing and using healthcare services compared to non-disabled women. However, relatively little is known about the utilisation of cancer screening services for women with disabilities. This study addresses this gap by examining the utilisation of the Papanicolaou test and mammography for disabled women in Chile. We used cross-sectional data, taken from a 2015 nationally-representative survey. Initially, we employed logistic regressions to test for differences in utilisation rates for the Papanicolaou test (66,281 observations) and the mammogram (35,294 observations) between disabled and non-disabled women. Next, logistic regressions were used to investigate the demographic, socioeconomic, and health-related factors affecting utilisation rates for cancer screening services for disabled women (sample sizes: 5,823 observations for the Papanicolaou test and 5,731 observations for the mammogram). Disabled women were less likely to undergo screening tests than non-disabled women. For the Papanicolaou test and mammography, the multivariable regression models showed that living in rural areas, having higher education, being affiliated with a private health insurance company, giving a good health self-assessment score, and being under medical treatment for other illnesses were associated with higher utilisation rates. On the other hand, being single, inactive with regard to employment, and having a better income were linked with lower utilisation. While utilisation rates for both disabled and non-disabled women have increased since 2006, the utilisation disparity has slightly increased. This study shows the influence of various factors in the utilisation rates of preventive cancer screening services for disabled women. To develop effective initiatives targeting inequalities in the utilisation of cancer screening tests, it is important to move beyond an exclusively single-disease approach and acknowledge the complexity of the patient population.
A Pilot Test of Indicator Species to Assess Uniqueness of Oak-Dominated Ecoregions in Central Tennessee

Treesearch

W. Henry McNab; David L. Loftis; Callie J. Schweitzer; Raymond Sheffield

2004-01-01

We used tree indicator species occurring on 438 plots in the Plateau counties of Tennessee to test the uniqueness of four conterminous ecoregions. Multinomial logistic regression indicated that the presence of 14 tree species allowed classification of sample plots according to ecoregion with an average overall accuracy of 75 percent (range 45 to 94 percent). Additional...
Using Logistic Regression for Validating or Invalidating Initial Statewide Cut-Off Scores on Basic Skills Placement Tests at the Community College Level

ERIC Educational Resources Information Center

Secolsky, Charles; Krishnan, Sathasivam; Judd, Thomas P.

2013-01-01

The community colleges in the state of New Jersey went through a process of establishing statewide cut-off scores for English and mathematics placement tests. The colleges wanted to communicate to secondary schools a consistent preparation that would be necessary for enrolling in Freshman Composition and College Algebra at the community college…

An Examination of Master's Student Retention & Completion

ERIC Educational Resources Information Center

Barry, Melissa; Mathies, Charles

2011-01-01

This study was conducted at a research-extensive public university in the southeastern United States. It examined the retention and completion of master's degree students across numerous disciplines. Results were derived from a series of descriptive statistics, T-tests, and a series of binary logistic regression models. The findings from binary…
Support vector machines classifiers of physical activities in preschoolers

USDA-ARS?s Scientific Manuscript database

The goal of this study is to develop, test, and compare multinomial logistic regression (MLR) and support vector machines (SVM) in classifying preschool-aged children physical activity data acquired from an accelerometer. In this study, 69 children aged 3-5 years old were asked to participate in a s...
The Effect of Religiosity and Campus Alcohol Culture on Collegiate Alcohol Consumption

ERIC Educational Resources Information Center

Wells, Gayle M.

2010-01-01

Religiosity and campus culture were examined in relationship to alcohol consumption among college students using reference group theory. Participants and Methods: College students (N = 530) at a religious college and at a state university complete questionnaires on alcohol use and religiosity. Statistical tests and logistic regression were…
Logistic regression models of factors influencing the location of bioenergy and biofuels plants

Treesearch

T.M. Young; R.L. Zaretzki; J.H. Perdue; F.M. Guess; X. Liu

2011-01-01

Logistic regression models were developed to identify significant factors that influence the location of existing wood-using bioenergy/biofuels plants and traditional wood-using facilities. Logistic models provided quantitative insight for variables influencing the location of woody biomass-using facilities. Availability of "thinnings to a basal area of 31.7m2/ha...
Fuzzy multinomial logistic regression analysis: A multi-objective programming approach

NASA Astrophysics Data System (ADS)

Abdalla, Hesham A.; El-Sayed, Amany A.; Hamed, Ramadan

2017-05-01

Parameter estimation for multinomial logistic regression is usually based on maximizing the likelihood function. For large well-balanced datasets, Maximum Likelihood (ML) estimation is a satisfactory approach. Unfortunately, ML can fail completely or at least produce poor results in terms of estimated probabilities and confidence intervals of parameters, specially for small datasets. In this study, a new approach based on fuzzy concepts is proposed to estimate parameters of the multinomial logistic regression. The study assumes that the parameters of multinomial logistic regression are fuzzy. Based on the extension principle stated by Zadeh and Bárdossy's proposition, a multi-objective programming approach is suggested to estimate these fuzzy parameters. A simulation study is used to evaluate the performance of the new approach versus Maximum likelihood (ML) approach. Results show that the new proposed model outperforms ML in cases of small datasets.
A Primer on Logistic Regression.

ERIC Educational Resources Information Center

Woldbeck, Tanya

This paper introduces logistic regression as a viable alternative when the researcher is faced with variables that are not continuous. If one is to use simple regression, the dependent variable must be measured on a continuous scale. In the behavioral sciences, it may not always be appropriate or possible to have a measured dependent variable on a…
Study of relationship between clinical factors and velopharyngeal closure in cleft palate patients

PubMed Central

Chen, Qi; Zheng, Qian; Shi, Bing; Yin, Heng; Meng, Tian; Zheng, Guang-ning

2011-01-01

BACKGROUND: This study was carried out to analyze the relationship between clinical factors and velopharyngeal closure (VPC) in cleft palate patients. METHODS: Chi-square test was used to compare the postoperative velopharyngeal closure rate. Logistic regression model was used to analyze independent variables associated with velopharyngeal closure. RESULTS: Difference of postoperative VPC rate in different cleft types, operative ages and surgical techniques was significant (P=0.000). Results of logistic regression analysis suggested that when operative age was beyond deciduous dentition stage, or cleft palate type was complete, or just had undergone a simple palatoplasty without levator veli palatini retropositioning, patients would suffer a higher velopharyngeal insufficiency rate after primary palatal repair. CONCLUSIONS: Cleft type, operative age and surgical technique were the contributing factors influencing VPC rate after primary palatal repair of cleft palate patients. PMID:22279464
Interest in Genetic Testing in Ashkenazi Jewish Parkinson’s Disease Patients and Their Unaffected Relatives

PubMed Central

Gupte, Manisha; Alcalay, Roy N.; Mejia-Santana, Helen; Raymond, Deborah; Saunders-Pullman, Rachel; Roos, Ernest; Orbe-Reily, Martha; Tang, Ming-X; Mirelman, Anat; Ozelius, Laurie; Orr-Urtreger, Avi; Clark, Lorraine; Giladi, Nir; Bressman, Susan

2014-01-01

Our objective was to explore interest in genetic testing among Ashkenazi Jewish (AJ) Parkinson’s Disease (PD) cases and first-degree relatives, as genetic testing for LRRK2 G2019S is widely available. Approximately 18 % of AJ PD cases carry G2019S mutations; penetrance estimations vary between 24 and 100 % by age 80. A Genetic Attitude Questionnaire (GAQ) was administered at two New York sites to PD families unaware of LRRK2 G2019S mutation status. The association of G2019S, age, education, gender and family history of PD with desire for genetic testing (outcome) was modeled using logistic regression. One-hundred eleven PD cases and 77 relatives completed the GAQ. Both PD cases and relatives had excellent PD-specific genetic knowledge. Among PD, 32.6 % “definitely” and 41.1 % “probably” wanted testing, if offered “now.” Among relatives, 23.6 % “definitely” and 36.1 % “probably” wanted testing “now.” Desire for testing in relatives increased incrementally based on hypothetical risk of PD. The most important reasons for testing in probands and relatives were: if it influenced medication response, identifying no mutation, and early prevention and treatment. In logistic regression, older age was associated with less desire for testing in probands OR=0.921 95%CI 0.868–0.977, p=0.009. Both probands and relatives express interest in genetic testing, despite no link to current treatment or prevention. PMID:25127731
Predicting space telerobotic operator training performance from human spatial ability assessment

NASA Astrophysics Data System (ADS)

Liu, Andrew M.; Oman, Charles M.; Galvan, Raquel; Natapoff, Alan

2013-11-01

Our goal was to determine whether existing tests of spatial ability can predict an astronaut's qualification test performance after robotic training. Because training astronauts to be qualified robotics operators is so long and expensive, NASA is interested in tools that can predict robotics performance before training begins. Currently, the Astronaut Office does not have a validated tool to predict robotics ability as part of its astronaut selection or training process. Commonly used tests of human spatial ability may provide such a tool to predict robotics ability. We tested the spatial ability of 50 active astronauts who had completed at least one robotics training course, then used logistic regression models to analyze the correlation between spatial ability test scores and the astronauts' performance in their evaluation test at the end of the training course. The fit of the logistic function to our data is statistically significant for several spatial tests. However, the prediction performance of the logistic model depends on the criterion threshold assumed. To clarify the critical selection issues, we show how the probability of correct classification vs. misclassification varies as a function of the mental rotation test criterion level. Since the costs of misclassification are low, the logistic models of spatial ability and robotic performance are reliable enough only to be used to customize regular and remedial training. We suggest several changes in tracking performance throughout robotics training that could improve the range and reliability of predictive models.
[Influences of environmental factors and interaction of several chemokines gene-environmental on systemic lupus erythematosus].

PubMed

Ye, Dong-qing; Hu, Yi-song; Li, Xiang-pei; Huang, Fen; Yang, Shi-gui; Hao, Jia-hu; Yin, Jing; Zhang, Guo-qing; Liu, Hui-hui

2004-11-01

To explore the impact of environmental factors, daily lifestyle, psycho-social factors and the interactions between environmental factors and chemokines genes on systemic lupus erythematosus (SLE). Case-control study was carried out and environmental factors for SLE were analyzed by univariate and multivariate unconditional logistic regression. Interactions between environmental factors and chemokines polymorphism contributing to systemic lupus erythematosus were also analyzed by logistic regression model. There were nineteen factors associated with SLE when univariate unconditional logistic regression was used. However, when multivariate unconditional logistic regression was used, only five factors showed having impacts on the disease, in which drinking well water (OR=0.099) was protective factor for SLE, and multiple drug allergy (OR=8.174), over-exposure to sunshine (OR=18.339), taking antibiotics (OR=9.630) and oral contraceptives were risk factors for SLE. When unconditional logistic regression model was used, results showed that there was interaction between eating irritable food and -2518MCP-1G/G genotype (OR=4.387). No interaction between environmental factors was found that contributing to SLE in this study. Many environmental factors were related to SLE, and there was an interaction between -2518MCP-1G/G genotype and eating irritable food.
A deeper look at two concepts of measuring gene-gene interactions: logistic regression and interaction information revisited.

PubMed

Mielniczuk, Jan; Teisseyre, Paweł

2018-03-01

Detection of gene-gene interactions is one of the most important challenges in genome-wide case-control studies. Besides traditional logistic regression analysis, recently the entropy-based methods attracted a significant attention. Among entropy-based methods, interaction information is one of the most promising measures having many desirable properties. Although both logistic regression and interaction information have been used in several genome-wide association studies, the relationship between them has not been thoroughly investigated theoretically. The present paper attempts to fill this gap. We show that although certain connections between the two methods exist, in general they refer two different concepts of dependence and looking for interactions in those two senses leads to different approaches to interaction detection. We introduce ordering between interaction measures and specify conditions for independent and dependent genes under which interaction information is more discriminative measure than logistic regression. Moreover, we show that for so-called perfect distributions those measures are equivalent. The numerical experiments illustrate the theoretical findings indicating that interaction information and its modified version are more universal tools for detecting various types of interaction than logistic regression and linkage disequilibrium measures. © 2017 WILEY PERIODICALS, INC.
Access disparities to Magnet hospitals for patients undergoing neurosurgical operations

PubMed Central

Missios, Symeon; Bekelis, Kimon

2017-01-01

Background Centers of excellence focusing on quality improvement have demonstrated superior outcomes for a variety of surgical interventions. We investigated the presence of access disparities to hospitals recognized by the Magnet Recognition Program of the American Nurses Credentialing Center (ANCC) for patients undergoing neurosurgical operations. Methods We performed a cohort study of all neurosurgery patients who were registered in the New York Statewide Planning and Research Cooperative System (SPARCS) database from 2009–2013. We examined the association of African-American race and lack of insurance with Magnet status hospitalization for neurosurgical procedures. A mixed effects propensity adjusted multivariable regression analysis was used to control for confounding. Results During the study period, 190,535 neurosurgical patients met the inclusion criteria. Using a multivariable logistic regression, we demonstrate that African-Americans had lower admission rates to Magnet institutions (OR 0.62; 95% CI, 0.58–0.67). This persisted in a mixed effects logistic regression model (OR 0.77; 95% CI, 0.70–0.83) to adjust for clustering at the patient county level, and a propensity score adjusted logistic regression model (OR 0.75; 95% CI, 0.69–0.82). Additionally, lack of insurance was associated with lower admission rates to Magnet institutions (OR 0.71; 95% CI, 0.68–0.73), in a multivariable logistic regression model. This persisted in a mixed effects logistic regression model (OR 0.72; 95% CI, 0.69–0.74), and a propensity score adjusted logistic regression model (OR 0.72; 95% CI, 0.69–0.75). Conclusions Using a comprehensive all-payer cohort of neurosurgery patients in New York State we identified an association of African-American race and lack of insurance with lower rates of admission to Magnet hospitals. PMID:28684152
Adjusting for Confounding in Early Postlaunch Settings: Going Beyond Logistic Regression Models.

PubMed

Schmidt, Amand F; Klungel, Olaf H; Groenwold, Rolf H H

2016-01-01

Postlaunch data on medical treatments can be analyzed to explore adverse events or relative effectiveness in real-life settings. These analyses are often complicated by the number of potential confounders and the possibility of model misspecification. We conducted a simulation study to compare the performance of logistic regression, propensity score, disease risk score, and stabilized inverse probability weighting methods to adjust for confounding. Model misspecification was induced in the independent derivation dataset. We evaluated performance using relative bias confidence interval coverage of the true effect, among other metrics. At low events per coefficient (1.0 and 0.5), the logistic regression estimates had a large relative bias (greater than -100%). Bias of the disease risk score estimates was at most 13.48% and 18.83%. For the propensity score model, this was 8.74% and >100%, respectively. At events per coefficient of 1.0 and 0.5, inverse probability weighting frequently failed or reduced to a crude regression, resulting in biases of -8.49% and 24.55%. Coverage of logistic regression estimates became less than the nominal level at events per coefficient ≤5. For the disease risk score, inverse probability weighting, and propensity score, coverage became less than nominal at events per coefficient ≤2.5, ≤1.0, and ≤1.0, respectively. Bias of misspecified disease risk score models was 16.55%. In settings with low events/exposed subjects per coefficient, disease risk score methods can be useful alternatives to logistic regression models, especially when propensity score models cannot be used. Despite better performance of disease risk score methods than logistic regression and propensity score models in small events per coefficient settings, bias, and coverage still deviated from nominal.
Creating Cost Growth Models for the Engineering and Manufacturing Development Phase of Acquisition Using Logistic and Multiple Regression

DTIC Science & Technology

2004-03-01

constant variance via an analysis of the residuals, as well as the Breusch - Pagan test (see Figure 3 below). As a result, we follow the footsteps of...reasonably normal, which ensures that our residuals meet the assumption of constant variance by passing the Breusch - Pagan test (see Figure 4 below...sections for Research and Development, Test and Evaluation (RDT&E), procurement and military construction (Jarvaise, 1996:3). While differing
A 3-Year Study of Predictive Factors for Positive and Negative Appendicectomies.

PubMed

Chang, Dwayne T S; Maluda, Melissa; Lee, Lisa; Premaratne, Chandrasiri; Khamhing, Srisongham

2018-03-06

Early and accurate identification or exclusion of acute appendicitis is the key to avoid the morbidity of delayed treatment for true appendicitis or unnecessary appendicectomy, respectively. We aim (i) to identify potential predictive factors for positive and negative appendicectomies; and (ii) to analyse the use of ultrasound scans (US) and computed tomography (CT) scans for acute appendicitis. All appendicectomies that took place at our hospital from the 1st of January 2013 to the 31st of December 2015 were retrospectively recorded. Test results of potential predictive factors of acute appendicitis were recorded. Statistical analysis was performed using Fisher exact test, logistic regression analysis, sensitivity, specificity, and positive and negative predictive values calculation. 208 patients were included in this study. 184 patients had histologically proven acute appendicitis. The other 24 patients had either nonappendicitis pathology or normal appendix. Logistic regression analysis showed statistically significant associations between appendicitis and white cell count, neutrophil count, C-reactive protein, and bilirubin. Neutrophil count was the test with the highest sensitivity and negative predictive values, whereas bilirubin was the test with the highest specificity and positive predictive values (PPV). US and CT scans had high sensitivity and PPV for diagnosing appendicitis. No single test was sufficient to diagnose or exclude acute appendicitis by itself. Combining tests with high sensitivity (abnormal neutrophil count, and US and CT scans) and high specificity (raised bilirubin) may predict acute appendicitis more accurately.
Screening for physical inactivity among adults: the value of distance walked in the six-minute walk test. A cross-sectional diagnostic study.

PubMed

Sperandio, Evandro Fornias; Arantes, Rodolfo Leite; da Silva, Rodrigo Pereira; Matheus, Agatha Caveda; Lauria, Vinícius Tonon; Bianchim, Mayara Silveira; Romiti, Marcello; Gagliardi, Antônio Ricardo de Toledo; Dourado, Victor Zuniga

2016-01-01

Accelerometry provides objective measurement of physical activity levels, but is unfeasible in clinical practice. Thus, we aimed to identify physical fitness tests capable of predicting physical inactivity among adults. Diagnostic test study developed at a university laboratory and a diagnostic clinic. 188 asymptomatic subjects underwent assessment of physical activity levels through accelerometry, ergospirometry on treadmill, body composition from bioelectrical impedance, isokinetic muscle function, postural balance on a force platform and six-minute walk test. We conducted descriptive analysis and multiple logistic regression including age, sex, oxygen uptake, body fat, center of pressure, quadriceps peak torque, distance covered in six-minute walk test and steps/day in the model, as predictors of physical inactivity. We also determined sensitivity (S), specificity (Sp) and area under the curve of the main predictors by means of receiver operating characteristic curves. The prevalence of physical inactivity was 14%. The mean number of steps/day (≤ 5357) was the best predictor of physical inactivity (S = 99%; Sp = 82%). The best physical fitness test was a distance in the six-minute walk test and ≤ 96% of predicted values (S = 70%; Sp = 80%). Body fat > 25% was also significant (S = 83%; Sp = 51%). After logistic regression, steps/day and distance in the six-minute walk test remained predictors of physical inactivity. The six-minute walk test should be included in epidemiological studies as a simple and cheap tool for screening for physical inactivity.
Stop the drama Downunder: a social marketing campaign increases HIV/sexually transmitted infection knowledge and testing in Australian gay men.

PubMed

Pedrana, Alisa; Hellard, Margaret; Guy, Rebecca; El-Hayek, Carol; Gouillou, Maelenn; Asselin, Jason; Batrouney, Colin; Nguyen, Phuong; Stoovè, Mark

2012-08-01

Since 2000, notifications of HIV and other sexually transmitted infections (STIs) have increased significantly in Australian gay men. We evaluated the impact of a social marketing campaign in 2008-2009 aimed to increase health-seeking behavior and STI testing and enhance HIV/STI knowledge in gay men. A convenience sample of 295 gay men (18-66 years of age) was surveyed to evaluate the effectiveness of the campaign. Participants were asked about campaign awareness, HIV/STI knowledge, health-seeking behavior, and HIV/STI testing. We examined associations between recent STI testing and campaign awareness. Trends in HIV/STI monthly tests at 3 clinics with a high case load of gay men were also assessed. Logistic and Poisson regressions and χ tests were used. Both unaided (43%) and aided (86%) campaign awareness was high. In a multivariable logistic regression, awareness of the campaign (aided) was independently associated with having had any STI test within the past 6 months (prevalence ratio = 1.5; 95% confidence interval = 1.0-2.4. Compared with the 13 months before the campaign, clinic data showed significant increasing testing rates for HIV, syphilis, and chlamydia among HIV-negative gay men during the initial and continued campaign periods. These findings suggest that the campaign was successful in achieving its aims of increasing health-seeking behavior, STI testing, and HIV/STI knowledge among gay men in Victoria.
Evaluation of select neurophysiological, clinical and psychological tests for burning mouth syndrome.

PubMed

Mendak-Ziółko, Magdalena; Konopka, Tomasz; Bogucki, Zdzisław Artur

2012-09-01

The objective of this study was to identify, among an array of potential risk factors for burning mouth syndrome (BMS), those that are potentially the most significant in the development of the disease. Sixty-three participants, divided into group I (with BMS: 33 patients ages 41 to 82 years [mean age: 61.5 ± 9.4]) and group II (without BMS: 30 healthy volunteers ages 42-83 years [mean age: 60.5 ± 10.5]) were studied. All underwent a dental examination and psychological tests. Neurological tests (neurophysiological test, electroneurography, and tests of the autonomic nervous system) were performed. Mean parameters were analyzed by Student t test, Kruskal-Wallis test, and χ(2) test, and multifactor analysis was performed with logistic regression and by calculating the odds ratio. In the logistic regression test, 3 factors were significant in the etiopathogenesis of BMS: a value more than 39 μV for the amplitude of the positive peak of the potential induced by stimulating the trigeminal nerve on the left side (P2-L); a value above 5.96 ms for the latency of wave V of the brainstem auditory evoked potentials on the right side (V-R); and a value over 2.35 ms for the latency of the sensory ulnar nerve response. The BMS sufferer was characterized as having mild sensory and autonomic small fiber neuropathy with concomitant central disorders. Copyright © 2012 Elsevier Inc. All rights reserved.
On the use and misuse of scalar scores of confounders in design and analysis of observational studies.

PubMed

Pfeiffer, R M; Riedl, R

2015-08-15

We assess the asymptotic bias of estimates of exposure effects conditional on covariates when summary scores of confounders, instead of the confounders themselves, are used to analyze observational data. First, we study regression models for cohort data that are adjusted for summary scores. Second, we derive the asymptotic bias for case-control studies when cases and controls are matched on a summary score, and then analyzed either using conditional logistic regression or by unconditional logistic regression adjusted for the summary score. Two scores, the propensity score (PS) and the disease risk score (DRS) are studied in detail. For cohort analysis, when regression models are adjusted for the PS, the estimated conditional treatment effect is unbiased only for linear models, or at the null for non-linear models. Adjustment of cohort data for DRS yields unbiased estimates only for linear regression; all other estimates of exposure effects are biased. Matching cases and controls on DRS and analyzing them using conditional logistic regression yields unbiased estimates of exposure effect, whereas adjusting for the DRS in unconditional logistic regression yields biased estimates, even under the null hypothesis of no association. Matching cases and controls on the PS yield unbiased estimates only under the null for both conditional and unconditional logistic regression, adjusted for the PS. We study the bias for various confounding scenarios and compare our asymptotic results with those from simulations with limited sample sizes. To create realistic correlations among multiple confounders, we also based simulations on a real dataset. Copyright © 2015 John Wiley & Sons, Ltd.
[Application of SAS macro to evaluated multiplicative and additive interaction in logistic and Cox regression in clinical practices].

PubMed

Nie, Z Q; Ou, Y Q; Zhuang, J; Qu, Y J; Mai, J Z; Chen, J M; Liu, X Q

2016-05-01

Conditional logistic regression analysis and unconditional logistic regression analysis are commonly used in case control study, but Cox proportional hazard model is often used in survival data analysis. Most literature only refer to main effect model, however, generalized linear model differs from general linear model, and the interaction was composed of multiplicative interaction and additive interaction. The former is only statistical significant, but the latter has biological significance. In this paper, macros was written by using SAS 9.4 and the contrast ratio, attributable proportion due to interaction and synergy index were calculated while calculating the items of logistic and Cox regression interactions, and the confidence intervals of Wald, delta and profile likelihood were used to evaluate additive interaction for the reference in big data analysis in clinical epidemiology and in analysis of genetic multiplicative and additive interactions.

Comparison of V50 Shot Placement on Final Outcome

DTIC Science & Technology

2014-11-01

molecular- weight polyethylene (UHMWPE). In V50 testing of those types of materials, large delaminations may occur that influence the results. This...placement, a proper evaluation of materials may not be possible. 15. SUBJECT TERMS ballistics, V50 test, logistic regression , statistical inference...from an impact. While this may work with ceramics or metal armor, it is inappropriate for use on composite armors like ultra-high-molecular- weight
Has there been a change in the knowledge of GP registrars between 2011 and 2016 as measured by performance on common items in the Applied Knowledge Test?

PubMed

Neden, Catherine A; Parkin, Claire; Blow, Carol; Siriwardena, Aloysius Niroshan

2018-05-08

The aim of this study was to assess whether the absolute standard of candidates sitting the MRCGP Applied Knowledge Test (AKT) between 2011 and 2016 had changed. It is a descriptive study comparing the performance on marker questions of a reference group of UK graduates taking the AKT for the first time between 2011 and 2016. Using aggregated examination data, the performance of individual 'marker' questions was compared using Pearson's chi-squared tests and trend-line analysis. Binary logistic regression was used to analyse changes in performance over the study period. Changes in performance of individual marker questions using Pearson's chi-squared test showed statistically significant differences in 32 of the 49 questions included in the study. Trend line analysis showed a positive trend in 29 questions and a negative trend in the remaining 23. The magnitude of change was small. Logistic regression did not demonstrate any evidence for a change in the performance of the question set over the study period. However, candidates were more likely to get items on administration wrong compared with clinical medicine or research. There was no evidence of a change in performance of the question set as a whole.
Binary Logistic Regression Versus Boosted Regression Trees in Assessing Landslide Susceptibility for Multiple-Occurring Regional Landslide Events: Application to the 2009 Storm Event in Messina (Sicily, southern Italy).

NASA Astrophysics Data System (ADS)

Lombardo, L.; Cama, M.; Maerker, M.; Parisi, L.; Rotigliano, E.

2014-12-01

This study aims at comparing the performances of Binary Logistic Regression (BLR) and Boosted Regression Trees (BRT) methods in assessing landslide susceptibility for multiple-occurrence regional landslide events within the Mediterranean region. A test area was selected in the north-eastern sector of Sicily (southern Italy), corresponding to the catchments of the Briga and the Giampilieri streams both stretching for few kilometres from the Peloritan ridge (eastern Sicily, Italy) to the Ionian sea. This area was struck on the 1st October 2009 by an extreme climatic event resulting in thousands of rapid shallow landslides, mainly of debris flows and debris avalanches types involving the weathered layer of a low to high grade metamorphic bedrock. Exploiting the same set of predictors and the 2009 landslide archive, BLR- and BRT-based susceptibility models were obtained for the two catchments separately, adopting a random partition (RP) technique for validation; besides, the models trained in one of the two catchments (Briga) were tested in predicting the landslide distribution in the other (Giampilieri), adopting a spatial partition (SP) based validation procedure. All the validation procedures were based on multi-folds tests so to evaluate and compare the reliability of the fitting, the prediction skill, the coherence in the predictor selection and the precision of the susceptibility estimates. All the obtained models for the two methods produced very high predictive performances, with a general congruence between BLR and BRT in the predictor importance. In particular, the research highlighted that BRT-models reached a higher prediction performance with respect to BLR-models, for RP based modelling, whilst for the SP-based models the difference in predictive skills between the two methods dropped drastically, converging to an analogous excellent performance. However, when looking at the precision of the probability estimates, BLR demonstrated to produce more robust models in terms of selected predictors and coefficients, as well as of dispersion of the estimated probabilities around the mean value for each mapped pixel. The difference in the behaviour could be interpreted as the result of overfitting effects, which heavily affect decision tree classification more than logistic regression techniques.
Applicability of the Ricketts' posteroanterior cephalometry for sex determination using logistic regression analysis in Hispano American Peruvians.

PubMed

Perez, Ivan; Chavez, Allison K; Ponce, Dario

2016-01-01

The Ricketts' posteroanterior (PA) cephalometry seems to be the most widely used and it has not been tested by multivariate statistics for sex determination. The objective was to determine the applicability of Ricketts' PA cephalometry for sex determination using the logistic regression analysis. The logistic models were estimated at distinct age cutoffs (all ages, 11 years, 13 years, and 15 years) in a database from 1,296 Hispano American Peruvians between 5 years and 44 years of age. The logistic models were composed by six cephalometric measurements; the accuracy achieved by resubstitution varied between 60% and 70% and all the variables, with one exception, exhibited a direct relationship with the probability of being classified as male; the nasal width exhibited an indirect relationship. The maxillary and facial widths were present in all models and may represent a sexual dimorphism indicator. The accuracy found was lower than the literature and the Ricketts' PA cephalometry may not be adequate for sex determination. The indirect relationship of the nasal width in models with data from patients of 12 years of age or less may be a trait related to age or a characteristic in the studied population, which could be better studied and confirmed.
No rationale for 1 variable per 10 events criterion for binary logistic regression analysis.

PubMed

van Smeden, Maarten; de Groot, Joris A H; Moons, Karel G M; Collins, Gary S; Altman, Douglas G; Eijkemans, Marinus J C; Reitsma, Johannes B

2016-11-24

Ten events per variable (EPV) is a widely advocated minimal criterion for sample size considerations in logistic regression analysis. Of three previous simulation studies that examined this minimal EPV criterion only one supports the use of a minimum of 10 EPV. In this paper, we examine the reasons for substantial differences between these extensive simulation studies. The current study uses Monte Carlo simulations to evaluate small sample bias, coverage of confidence intervals and mean square error of logit coefficients. Logistic regression models fitted by maximum likelihood and a modified estimation procedure, known as Firth's correction, are compared. The results show that besides EPV, the problems associated with low EPV depend on other factors such as the total sample size. It is also demonstrated that simulation results can be dominated by even a few simulated data sets for which the prediction of the outcome by the covariates is perfect ('separation'). We reveal that different approaches for identifying and handling separation leads to substantially different simulation results. We further show that Firth's correction can be used to improve the accuracy of regression coefficients and alleviate the problems associated with separation. The current evidence supporting EPV rules for binary logistic regression is weak. Given our findings, there is an urgent need for new research to provide guidance for supporting sample size considerations for binary logistic regression analysis.
Does clinical pretest probability influence image quality and diagnostic accuracy in dual-source coronary CT angiography?

PubMed

Thomas, Christoph; Brodoefel, Harald; Tsiflikas, Ilias; Bruckner, Friederike; Reimann, Anja; Ketelsen, Dominik; Drosch, Tanja; Claussen, Claus D; Kopp, Andreas; Heuschmid, Martin; Burgstahler, Christof

2010-02-01

To prospectively evaluate the influence of the clinical pretest probability assessed by the Morise score onto image quality and diagnostic accuracy in coronary dual-source computed tomography angiography (DSCTA). In 61 patients, DSCTA and invasive coronary angiography were performed. Subjective image quality and accuracy for stenosis detection (>50%) of DSCTA with invasive coronary angiography as gold standard were evaluated. The influence of pretest probability onto image quality and accuracy was assessed by logistic regression and chi-square testing. Correlations of image quality and accuracy with the Morise score were determined using linear regression. Thirty-eight patients were categorized into the high, 21 into the intermediate, and 2 into the low probability group. Accuracies for the detection of significant stenoses were 0.94, 0.97, and 1.00, respectively. Logistic regressions and chi-square tests showed statistically significant correlations between Morise score and image quality (P < .0001 and P < .001) and accuracy (P = .0049 and P = .027). Linear regression revealed a cutoff Morise score for a good image quality of 16 and a cutoff for a barely diagnostic image quality beyond the upper Morise scale. Pretest probability is a weak predictor of image quality and diagnostic accuracy in coronary DSCTA. A sufficient image quality for diagnostic images can be reached with all pretest probabilities. Therefore, coronary DSCTA might be suitable also for patients with a high pretest probability. Copyright 2010 AUR. Published by Elsevier Inc. All rights reserved.
Secure Logistic Regression Based on Homomorphic Encryption: Design and Evaluation.

PubMed

Kim, Miran; Song, Yongsoo; Wang, Shuang; Xia, Yuhou; Jiang, Xiaoqian

2018-04-17

Learning a model without accessing raw data has been an intriguing idea to security and machine learning researchers for years. In an ideal setting, we want to encrypt sensitive data to store them on a commercial cloud and run certain analyses without ever decrypting the data to preserve privacy. Homomorphic encryption technique is a promising candidate for secure data outsourcing, but it is a very challenging task to support real-world machine learning tasks. Existing frameworks can only handle simplified cases with low-degree polynomials such as linear means classifier and linear discriminative analysis. The goal of this study is to provide a practical support to the mainstream learning models (eg, logistic regression). We adapted a novel homomorphic encryption scheme optimized for real numbers computation. We devised (1) the least squares approximation of the logistic function for accuracy and efficiency (ie, reduce computation cost) and (2) new packing and parallelization techniques. Using real-world datasets, we evaluated the performance of our model and demonstrated its feasibility in speed and memory consumption. For example, it took approximately 116 minutes to obtain the training model from the homomorphically encrypted Edinburgh dataset. In addition, it gives fairly accurate predictions on the testing dataset. We present the first homomorphically encrypted logistic regression outsourcing model based on the critical observation that the precision loss of classification models is sufficiently small so that the decision plan stays still. ©Miran Kim, Yongsoo Song, Shuang Wang, Yuhou Xia, Xiaoqian Jiang. Originally published in JMIR Medical Informatics (http://medinform.jmir.org), 17.04.2018.
MODELING SNAKE MICROHABITAT FROM RADIOTELEMETRY STUDIES USING POLYTOMOUS LOGISTIC REGRESSION

EPA Science Inventory

Multivariate analysis of snake microhabitat has historically used techniques that were derived under assumptions of normality and common covariance structure (e.g., discriminant function analysis, MANOVA). In this study, polytomous logistic regression (PLR which does not require ...
Automated particle identification through regression analysis of size, shape and colour

NASA Astrophysics Data System (ADS)

Rodriguez Luna, J. C.; Cooper, J. M.; Neale, S. L.

2016-04-01

Rapid point of care diagnostic tests and tests to provide therapeutic information are now available for a range of specific conditions from the measurement of blood glucose levels for diabetes to card agglutination tests for parasitic infections. Due to a lack of specificity these test are often then backed up by more conventional lab based diagnostic methods for example a card agglutination test may be carried out for a suspected parasitic infection in the field and if positive a blood sample can then be sent to a lab for confirmation. The eventual diagnosis is often achieved by microscopic examination of the sample. In this paper we propose a computerized vision system for aiding in the diagnostic process; this system used a novel particle recognition algorithm to improve specificity and speed during the diagnostic process. We will show the detection and classification of different types of cells in a diluted blood sample using regression analysis of their size, shape and colour. The first step is to define the objects to be tracked by a Gaussian Mixture Model for background subtraction and binary opening and closing for noise suppression. After subtracting the objects of interest from the background the next challenge is to predict if a given object belongs to a certain category or not. This is a classification problem, and the output of the algorithm is a Boolean value (true/false). As such the computer program should be able to "predict" with reasonable level of confidence if a given particle belongs to the kind we are looking for or not. We show the use of a binary logistic regression analysis with three continuous predictors: size, shape and color histogram. The results suggest this variables could be very useful in a logistic regression equation as they proved to have a relatively high predictive value on their own.
Motivations and Benefits for Attaining HR Certifications

ERIC Educational Resources Information Center

Lester, Scott W.; Dwyer, Dale J.

2012-01-01

Purpose: The aim of this paper is to examine the motivations and benefits for pursuing or not pursuing the PHR and SPHR. Design/methodology/approach: Using a sample of 1,862 participants, the study used multinomial logistic and hierarchical linear regression to test six hypotheses. Findings: Participants pursuing SPHR were more likely to report…
Item and Testlet Position Effects in Computer-Based Alternate Assessments for Students with Disabilities

ERIC Educational Resources Information Center

Bulut, Okan; Lei, Ming; Guo, Qi

2018-01-01

Item positions in educational assessments are often randomized across students to prevent cheating. However, if altering item positions results in any significant impact on students' performance, it may threaten the validity of test scores. Two widely used approaches for detecting position effects -- logistic regression and hierarchical…
Racial Threat and White Opposition to Bilingual Education in Texas

ERIC Educational Resources Information Center

Hempel, Lynn M.; Dowling, Julie A.; Boardman, Jason D.; Ellison, Christopher G.

2013-01-01

This study examines local contextual conditions that influence opposition to bilingual education among non-Hispanic Whites, net of individual-level characteristics. Data from the Texas Poll (N = 615) are used in conjunction with U.S. Census data to test five competing hypotheses using binomial and multinomial logistic regression models. Our…
Factors Contributing to the Upward Transfer of Baccalaureate Aspirants Beginning at Community Colleges. WISCAPE Working Paper

ERIC Educational Resources Information Center

Wang, Xueli

2010-01-01

Incorporating the psychological perspective, this study examines factors associated with the upward transfer of baccalaureate aspirants beginning at community colleges. Based on data from the National Education Longitudinal Study of 1988 and the Postsecondary Education Transcript Study, the study tests a logistic regression model to predict…
College Student Characteristics and Experiences as Predictors of Interracial Dating

ERIC Educational Resources Information Center

Harper, Casandra E.; Yeung, Fanny P.

2015-01-01

This study utilized logistic regression to test whether students' personal characteristics and experiences significantly predict their likelihood of dating interracially in college. The data were drawn from the Campus Life in America Student Survey (CLASS), which was administered to freshmen who were then resurveyed as juniors (n = 513). The most…
Predictors of Success in Accelerated and Enrichment Summer Mathematics Courses for Academically Talented Adolescents

ERIC Educational Resources Information Center

Young, Adena E.; Worrell, Frank C.; Gabelko, Nina H.

2011-01-01

In this study, we used logistic regression to examine how well student background and prior achievement variables predicted success among students attending accelerated and enrichment mathematics courses at a summer program (N = 459). Socioeconomic status, grade point average (GPA), and mathematics diagnostic test scores significantly predicted…
Predictors of Child Molestation: Adult Attachment, Cognitive Distortions, and Empathy

ERIC Educational Resources Information Center

Wood, Eric; Riggs, Shelley

2008-01-01

A conceptual model derived from attachment theory was tested by examining adult attachment style, cognitive distortions, and both general and victim empathy in a sample of 61 paroled child molesters and 51 community controls. Results of logistic multiple regression showed that attachment anxiety, cognitive distortions, high general empathy but low…
Predicting Teacher Value-Added Results in Non-Tested Subjects Based on Confounding Variables: A Multinomial Logistic Regression

ERIC Educational Resources Information Center

Street, Nathan Lee

2017-01-01

Teacher value-added measures (VAM) are designed to provide information regarding teachers' causal impact on the academic growth of students while controlling for exogenous variables. While some researchers contend VAMs successfully and authentically measure teacher causality on learning, others suggest VAMs cannot adequately control for exogenous…
Use of the Child Behavior Checklist as a Diagnostic Screening Tool in Community Mental Health

ERIC Educational Resources Information Center

Rishel, Carrie W.; Greeno, Catherine; Marcus, Steven C.; Shear, M. Katherine; Anderson, Carol

2005-01-01

Objective: This study examines whether the Child Behavior Checklist (CBCL) can be used as an accurate psychiatric screening tool for children in community mental health settings. Method: Associations, logistic regression models, and receiver operating characteristic (ROC) analysis were used to test the predictive relationship between the CBCL and…
[The effect of self-foot reflexology on the relief of premenstrual syndrome and dysmenorrhea in high school girls].

PubMed

Kim, Yi-Soon; Kim, Min-Za; Jeong, Ihn-Sook

2004-08-01

This study was aimed to identify the effect of self-foot reflexology on the relief of premenstrual syndrome and dysmenorrhea in high school girls. Study subjects was 236 women residing in the community, teachers and nurses who were older than 45 were recruited. Data was collected with self administered questionnaires from July 1st to August 31st, 2003 and analysed using SPSS/WIN 10.0 with Xtest, t-test, and stepwise multiple logistic regression at a significant level of =.05. The breast cancer screening rate was 57.2%, and repeat screening rate was 15.3%. With the multiple logistic regression analysis, factors associated with mammography screening were age and perceived barriers of action, and factors related to the repeat mammography screening were education level and other cancer screening experience. Based on the results, we recommend the development of an intervention program to decrease the perceived barrier of action, to regard mammography as an essential test in regular check-up, and to give active advertisement and education to the public to improve the rates of breast cancer screening and repeat screening.
[Transmission disequilibrium test for nonsyndromic cleft lip and palate and segment homeobox gene-1 gene].

PubMed

Wu, Ping-An; Li, Yun-Liang; Wu, Han-Jiang; Wang, Kai; Fan, Guo-Zheng

2007-09-01

To investigate the relationship between muscle segment homeobox gene-1 (MSX1) and the genetic susceptibility of nonsyndromic cleft lip and palate (NSCLP) in Hunan Hans. One microsatellite DNA marker CA repeat in MSX1 intron region was used as genetic marker. The genotypes of 387 members in 129 NSCLP nuclear family trios were analyzed by polymerase chain reaction (PCR) and denaturing polyacrylamide gel electrophoresis. Then transmission disequilibrium test (TDT) and Logistic regression analysis were used to conduct association analysis. TDT analysis confirmed that CA4 allele in CL/P and CPO groups preferentially transmitted to the affected offspring (P = 0.018, P = 0.041). Logistic regression analysis indicated that the recessive model of inheritance was supported, and CA4 itself or CA4 acting as a marker for a disease allele or haplotype was inherited in a recessive fashion (P = 0.009). MSX1 gene is associated with NSCLP, and MSX1 gene may be directly involved either in the etiology of NSCLP or in linkage disequilibrium with disease-predisposing sites.

Influence of child rearing by grandparent on the development of children aged six to twelve years.

PubMed

Nanthamongkolchai, Sutham; Munsawaengsub, Chokchai; Nanthamongkolchai, Chantira

2009-03-01

To investigate the influence of child rearing by grandparent on the development of children aged six to twelve years. A cross-sectional study was conducted in 320 children that were cared for by a parent and grandparent selected by cluster sampling. The data were collected between March 10 and April 8, 2006 by questionnaire about child and family factors. The TONI-III test was used to test the child development. Data were analyzed by frequency distribution, logistic regression, and multiple logistic regression. Child caregiver had a significant influence on child development (p-value < 0.05). Children reared by a grandparent had 2.0 times higher chance of having delayed development compared with those who were reared by the parent. In addition, significant family factors that had impact on the child development were child rearing and family income. Child rearing by a grandparent had 2.0 times higher chance of having delayed development than those reared by the parent. Therefore, family and health personnel should plan to ensure the development and learning process of children that are cared by the grandparent.
Risk factor for preterm labor in Haji Adam Malik General Hospital, Pirngadi General hospital and satellite hospitals in Medan from January 2014 to December 2016

NASA Astrophysics Data System (ADS)

Sukatendel, K.; Hasibuan, C. L.; Pasaribu, H. P.; Sihite, H.; Ardyansah, E.; Situmorang, M. F.

2018-03-01

In 2010, Indonesia was ranked fifth in the world for the number of premature birth. Prematurity is a multifactorial problem. Preterm Labor (PTL) can occur spontaneously without a clear cause. Preventing PTL, its associated risk factors must be recognized first. To analyze risk factors associated with the incidence of PTL. It is a cross sectional study using secondary data obtained from medical records in Haji Adam Malik general hospital, Pirngadi general hospital and satellite hospitals in Medan from January 2014 to December 2016. Data were analyzed using chi-square method and logistic regression test. 148 cases for each group of preterm labor and obtained term laborin this study. Using the logistic regression test, three factors with astrong association to the incidence of identifiedpreterm labor. Antenatal Care frequency (OR 2,326; CI 95%), leucorrhea (OR 6,291; 95%), and premature rupture of membrane (OR 9,755; CI 95%). In conclusion, antenatal care frequency, leucorrhea, and history of premature rupture of themembrane may increase the incidence of Preterm Labor (PTL).
Relationship between risk factors and activities of daily living using modified Shah Barthel Index in stroke patients

NASA Astrophysics Data System (ADS)

Kusumaningsih, W.; Rachmayanti, S.; Werdhani, R. A.

2017-08-01

Hypertension and diabetes mellitus are the most common risk factors of stroke. The study aimed to determine the relationship between hypertension and diabetes mellitus risk factors and dependence on assistance with activities of daily living in chronic stroke patients. The study used an analytical observational cross-sectional design. The study’s sample included 44 stroke patients selected using the quota sampling method. The relationship between the variables was analyzed using the bivariate chi-squared test and multivariate logistic regression. Based on the chi-squared test, the relationship between the Modified Shah Barthel Index (MSBI) score and hypertension and diabetes mellitus as stroke risk factors, were p = 0.122 and p = 0.002, respectively. The logistic regression results suggest that hypertension and diabetes mellitus are stroke risk factors related to the MSBI score: p = 0.076 (OR 4.076; CI 95% 0.861-19.297) and p = 0.007 (OR 22.690; CI 95% 2.332-220.722), respectively. Diabetes mellitus is the most prominent risk factor of severe dependency on assistance with activities of daily living in chronic stroke patients.
Evaluating construct validity of the second version of the Copenhagen Psychosocial Questionnaire through analysis of differential item functioning and differential item effect.

PubMed

Bjorner, Jakob Bue; Pejtersen, Jan Hyld

2010-02-01

To evaluate the construct validity of the Copenhagen Psychosocial Questionnaire II (COPSOQ II) by means of tests for differential item functioning (DIF) and differential item effect (DIE). We used a Danish general population postal survey (n = 4,732 with 3,517 wage earners) with a one-year register based follow up for long-term sickness absence. DIF was evaluated against age, gender, education, social class, public/private sector employment, and job type using ordinal logistic regression. DIE was evaluated against job satisfaction and self-rated health (using ordinal logistic regression), against depressive symptoms, burnout, and stress (using multiple linear regression), and against long-term sick leave (using a proportional hazards model). We used a cross-validation approach to counter the risk of significant results due to multiple testing. Out of 1,052 tests, we found 599 significant instances of DIF/DIE, 69 of which showed both practical and statistical significance across two independent samples. Most DIF occurred for job type (in 20 cases), while we found little DIF for age, gender, education, social class and sector. DIE seemed to pertain to particular items, which showed DIE in the same direction for several outcome variables. The results allowed a preliminary identification of items that have a positive impact on construct validity and items that have negative impact on construct validity. These results can be used to develop better shortform measures and to improve the conceptual framework, items and scales of the COPSOQ II. We conclude that tests of DIF and DIE are useful for evaluating construct validity.
Selecting risk factors: a comparison of discriminant analysis, logistic regression and Cox's regression model using data from the Tromsø Heart Study.

PubMed

Brenn, T; Arnesen, E

1985-01-01

For comparative evaluation, discriminant analysis, logistic regression and Cox's model were used to select risk factors for total and coronary deaths among 6595 men aged 20-49 followed for 9 years. Groups with mortality between 5 and 93 per 1000 were considered. Discriminant analysis selected variable sets only marginally different from the logistic and Cox methods which always selected the same sets. A time-saving option, offered for both the logistic and Cox selection, showed no advantage compared with discriminant analysis. Analysing more than 3800 subjects, the logistic and Cox methods consumed, respectively, 80 and 10 times more computer time than discriminant analysis. When including the same set of variables in non-stepwise analyses, all methods estimated coefficients that in most cases were almost identical. In conclusion, discriminant analysis is advocated for preliminary or stepwise analysis, otherwise Cox's method should be used.
Sex differences in the effect of aging on dry eye disease.

PubMed

Ahn, Jong Ho; Choi, Yoon-Hyeong; Paik, Hae Jung; Kim, Mee Kum; Wee, Won Ryang; Kim, Dong Hyun

2017-01-01

Aging is a major risk factor in dry eye disease (DED), and understanding sexual differences is very important in biomedical research. However, there is little information about sex differences in the effect of aging on DED. We investigated sex differences in the effect of aging and other risk factors for DED. This study included data of 16,824 adults from the Korea National Health and Nutrition Examination Survey (2010-2012), which is a population-based cross-sectional survey. DED was defined as the presence of frequent ocular dryness or a previous diagnosis by an ophthalmologist. Basic sociodemographic factors and previously known risk factors for DED were included in the analyses. Linear regression modeling and multivariate logistic regression modeling were used to compare the sex differences in the effect of risk factors for DED; we additionally performed tests for interactions between sex and other risk factors for DED in logistic regression models. In our linear regression models, the prevalence of DED symptoms in men increased with age ( R =0.311, P =0.012); however, there was no association between aging and DED in women ( P >0.05). Multivariate logistic regression analyses showed that aging in men was not associated with DED (DED symptoms/diagnosis: odds ratio [OR] =1.01/1.04, each P >0.05), while aging in women was protectively associated with DED (DED symptoms/diagnosis: OR =0.94/0.91, P =0.011/0.003). Previous ocular surgery was significantly associated with DED in both men and women (men/women: OR =2.45/1.77 [DED symptoms] and 3.17/2.05 [DED diagnosis], each P <0.001). Tests for interactions of sex revealed significantly different aging × sex and previous ocular surgery × sex interactions ( P for interaction of sex: DED symptoms/diagnosis - 0.044/0.011 [age] and 0.012/0.006 [previous ocular surgery]). There were distinct sex differences in the effect of aging on DED in the Korean population. DED following ocular surgery also showed sexually different patterns. Age matching and sex matching are strongly recommended in further studies about DED, especially DED following ocular surgery.
Mindfulness, Physical Activity and Avoidance of Secondhand Smoke: A Study of College Students in Shanghai.

PubMed

Gao, Yu; Shi, Lu

2015-08-21

To better understand the documented link between mindfulness and longevity, we examine the association between mindfulness and conscious avoidance of secondhand smoke (SHS), as well as the association between mindfulness and physical activity. In Shanghai University of Finance and Economics (SUFE) we surveyed a convenience sample of 1516 college freshmen. We measured mindfulness, weekly physical activity, and conscious avoidance of secondhand smoke, along with demographic and behavioral covariates. We used a multilevel logistic regression to test the association between mindfulness and conscious avoidance of secondhand smoke, and used a Tobit regression model to test the association between mindfulness and metabolic equivalent hours per week. In both models the home province of the student respondent was used as the cluster variable, and demographic and behavioral covariates, such as age, gender, smoking history, household registration status (urban vs. rural), the perceived smog frequency in their home towns, and the asthma diagnosis. The logistic regression of consciously avoiding SHS shows that a higher level of mindfulness was associated with an increase in the odds ratio of conscious SHS avoidance (logged odds: 0.22, standard error: 0.07, p < 0.01). The Tobit regression shows that a higher level of mindfulness was associated with more metabolic equivalent hours per week (Tobit coefficient: 4.09, standard error: 1.13, p < 0.001). This study is an innovative attempt to study the behavioral issue of secondhand smoke from the perspective of the potential victim, rather than the active smoker. The observed associational patterns here are consistent with previous findings that mindfulness is associated with healthier behaviors in obesity prevention and substance use. Research designs with interventions are needed to test the causal link between mindfulness and these healthy behaviors.
Mindfulness, Physical Activity and Avoidance of Secondhand Smoke: A Study of College Students in Shanghai

PubMed Central

Gao, Yu; Shi, Lu

2015-01-01

Introduction: To better understand the documented link between mindfulness and longevity, we examine the association between mindfulness and conscious avoidance of secondhand smoke (SHS), as well as the association between mindfulness and physical activity. Method: In Shanghai University of Finance and Economics (SUFE) we surveyed a convenience sample of 1516 college freshmen. We measured mindfulness, weekly physical activity, and conscious avoidance of secondhand smoke, along with demographic and behavioral covariates. We used a multilevel logistic regression to test the association between mindfulness and conscious avoidance of secondhand smoke, and used a Tobit regression model to test the association between mindfulness and metabolic equivalent hours per week. In both models the home province of the student respondent was used as the cluster variable, and demographic and behavioral covariates, such as age, gender, smoking history, household registration status (urban vs. rural), the perceived smog frequency in their home towns, and the asthma diagnosis. Results: The logistic regression of consciously avoiding SHS shows that a higher level of mindfulness was associated with an increase in the odds ratio of conscious SHS avoidance (logged odds: 0.22, standard error: 0.07, p < 0.01). The Tobit regression shows that a higher level of mindfulness was associated with more metabolic equivalent hours per week (Tobit coefficient: 4.09, standard error: 1.13, p < 0.001). Discussion: This study is an innovative attempt to study the behavioral issue of secondhand smoke from the perspective of the potential victim, rather than the active smoker. The observed associational patterns here are consistent with previous findings that mindfulness is associated with healthier behaviors in obesity prevention and substance use. Research designs with interventions are needed to test the causal link between mindfulness and these healthy behaviors. PMID:26308029
Practical Session: Logistic Regression

NASA Astrophysics Data System (ADS)

Clausel, M.; Grégoire, G.

2014-12-01

An exercise is proposed to illustrate the logistic regression. One investigates the different risk factors in the apparition of coronary heart disease. It has been proposed in Chapter 5 of the book of D.G. Kleinbaum and M. Klein, "Logistic Regression", Statistics for Biology and Health, Springer Science Business Media, LLC (2010) and also by D. Chessel and A.B. Dufour in Lyon 1 (see Sect. 6 of http://pbil.univ-lyon1.fr/R/pdf/tdr341.pdf). This example is based on data given in the file evans.txt coming from http://www.sph.emory.edu/dkleinb/logreg3.htm#data.
The cross-validated AUC for MCP-logistic regression with high-dimensional data.

PubMed

Jiang, Dingfeng; Huang, Jian; Zhang, Ying

2013-10-01

We propose a cross-validated area under the receiving operator characteristic (ROC) curve (CV-AUC) criterion for tuning parameter selection for penalized methods in sparse, high-dimensional logistic regression models. We use this criterion in combination with the minimax concave penalty (MCP) method for variable selection. The CV-AUC criterion is specifically designed for optimizing the classification performance for binary outcome data. To implement the proposed approach, we derive an efficient coordinate descent algorithm to compute the MCP-logistic regression solution surface. Simulation studies are conducted to evaluate the finite sample performance of the proposed method and its comparison with the existing methods including the Akaike information criterion (AIC), Bayesian information criterion (BIC) or Extended BIC (EBIC). The model selected based on the CV-AUC criterion tends to have a larger predictive AUC and smaller classification error than those with tuning parameters selected using the AIC, BIC or EBIC. We illustrate the application of the MCP-logistic regression with the CV-AUC criterion on three microarray datasets from the studies that attempt to identify genes related to cancers. Our simulation studies and data examples demonstrate that the CV-AUC is an attractive method for tuning parameter selection for penalized methods in high-dimensional logistic regression models.
Varicella infection is not associated with increasing prevalence of eczema: a U.S. population-based study.

PubMed

Li, J C; Silverberg, J I

2015-11-01

Chickenpox infection early in childhood has previously been shown to protect against the development of childhood eczema in line with the hygiene hypothesis. In 1995, the American Academy of Pediatrics recommended routine vaccination against varicella zoster virus in the United States. Subsequently, rates of chickenpox infection have dramatically decreased in childhood. We sought to understand the impact of declining rates of chickenpox infection on the prevalence of eczema. We analysed data from 207 007 children in the 1997-2013 National Health Interview Survey. One-year prevalence of eczema and 'ever had' history of chickenpox were analysed. Associations between chickenpox infection and eczema were tested using survey-weighted logistic regression. The impact of chickenpox on trends of eczema prevalence was tested using survey logistic regression and generalized linear models. Children with a history of chickenpox compared with those without chickenpox had a lower prevalence [survey-weighted logistic regression (95% confidence interval, CI)] of eczema [8·8% (8·5-9·0%) vs. 10·6% (10·4-10·8%)]. In pooled multivariate models controlling for age, sex, race/ethnicity, household income, highest level of household education, insurance coverage, U.S. birthplace and family size, eczema was inversely associated with chickenpox [adjusted odds ratio (95% CI), 0·90 (0·86-0·94), P < 0·001]. The prevalence of eczema significantly increased over time (Tukey post-hoc test, P < 0·001 for comparisons of survey years 2001-13 vs. 1997-2000, 2008-13 vs. 2001-04 and 2008-13 vs. 2005-07). In multivariate generalized linear models, the odds of eczema was not associated with chickenpox in 2001-13 (P ≥ 0·06). These findings suggest that lower rates of chickenpox infection secondary to widespread vaccination against varicella zoster virus are not contributing to higher rates of childhood eczema in the U.S. © 2015 British Association of Dermatologists.
Estimating Procurement Cost Growth Using Logistic and Multiple Regression

DTIC Science & Technology

2003-03-01

Figure 4). The plots fail to pass the visual inspection for constant variance as well as the Breusch - Pagan test (Neter, 1996: 112) at an alpha level...plots fail to pass the visual inspection for constant variance as well as the Breusch - Pagan test at an alpha level of 0.05. Based on these findings...amount of cost growth a program will have 13 once model A deems that the program will incur cost growth. Sipple conducts validation testing on
Odontological approach to sexual dimorphism in southeastern France.

PubMed

Lladeres, Emilie; Saliba-Serre, Bérengère; Sastre, Julien; Foti, Bruno; Tardivo, Delphine; Adalian, Pascal

2013-01-01

The aim of this study was to establish a prediction formula to allow for the determination of sex among the southeastern French population using dental measurements. The sample consisted of 105 individuals (57 males and 48 females, aged between 18 and 25 years). Dental measurements were calculated using Euclidean distances, in three-dimensional space, from point coordinates obtained by a Microscribe. A multiple logistic regression analysis was performed to establish the prediction formula. Among 12 selected dental distances, a stepwise logistic regression analysis highlighted the two most significant discriminate predictors of sex: one located at the mandible and the other at the maxilla. A cutpoint was proposed to prediction of true sex. The prediction formula was then tested on a validation sample (20 males and 34 females, aged between 18 and 62 years and with a history of orthodontics or restorative care) to evaluate the accuracy of the method. © 2012 American Academy of Forensic Sciences.
Association of school, family, and mental health characteristics with suicidal ideation among Korean adolescents.

PubMed

Lee, Gyu-Young; Choi, Yun-Jung

2015-08-01

In a cross-sectional research design, we investigated factors related to suicidal ideation in adolescents using data from the 2013 Online Survey of Youth Health Behavior in Korea. This self-report questionnaire was administered to 72,435 adolescents aged 13-18 years in middle and high school. School characteristics, family characteristics, and mental health variables were analyzed using descriptive statistics, χ(2) tests, and logistic regression. Both suicidal ideation and behavior were more common in girls. Suicidal ideation was most common in 11th grade for boys and 8th grade for girls. Across the sample, in logistic regression, suicidal ideation was predicted by low socioeconomic status, high stress, inadequate sleep, substance use, alcohol use, and smoking. Living apart from family predicted suicidal ideation in boys but not in girls. Gender- and school-grade-specific intervention programs may be useful for reducing suicidal ideation in students. © 2015 Wiley Periodicals, Inc.
Minimal intervention dentistry for early childhood caries and child dental anxiety: a randomized controlled trial.

PubMed

Arrow, P; Klobas, E

2017-06-01

To compare changes in child dental anxiety after treatment for early childhood caries (ECC) using two treatment approaches. Children with ECC were randomized to test (atraumatic restorative treatment (ART)-based approach) or control (standard care approach) groups. Children aged 3 years or older completed a dental anxiety scale at baseline and follow up. Changes in child dental anxiety from baseline to follow up were tested using the chi-squared statistic, Wilcoxon rank sum test, McNemar's test and multinomial logistic regression. Two hundred and fifty-four children were randomized (N = 127 test, N = 127 control). At baseline, 193 children completed the dental anxiety scale, 211 at follow up and 170 completed the scale on both occasions. Children who were anxious at baseline (11%) were no longer anxious at follow up, and 11% non-anxious children became anxious. Multinomial logistic regression found each increment in the number of visits increased the odds of worsening dental anxiety (odds ratio (OR), 2.2; P < 0.05), whereas each increment in the number of treatments lowered the odds of worsening anxiety (OR, 0.50; P = 0.05). The ART-based approach to managing ECC resulted in similar levels of dental anxiety to the standard treatment approach and provides a valuable alternative approach to the management of ECC in a primary dental care setting. © 2016 Australian Dental Association.
The effect of service satisfaction and spiritual well-being on the quality of life of patients with schizophrenia.

PubMed

Lanfredi, Mariangela; Candini, Valentina; Buizza, Chiara; Ferrari, Clarissa; Boero, Maria E; Giobbio, Gian M; Goldschmidt, Nicoletta; Greppo, Stefania; Iozzino, Laura; Maggi, Paolo; Melegari, Anna; Pasqualetti, Patrizio; Rossi, Giuseppe; de Girolamo, Giovanni

2014-05-15

Quality of life (QOL) has been considered an important outcome measure in psychiatric research and determinants of QOL have been widely investigated. We aimed at detecting predictors of QOL at baseline and at testing the longitudinal interrelations of the baseline predictors with QOL scores at a 1-year follow-up in a sample of patients living in Residential Facilities (RFs). Logistic regression models were adopted to evaluate the association between WHOQoL-Bref scores and potential determinants of QOL. In addition, all variables significantly associated with QOL domains in the final logistic regression model were included by using the Structural Equation Modeling (SEM). We included 139 patients with a diagnosis of schizophrenia spectrum. In the final logistic regression model level of activity, social support, age, service satisfaction, spiritual well-being and symptoms' severity were identified as predictors of QOL scores at baseline. Longitudinal analyses carried out by SEM showed that 40% of QOL follow-up variability was explained by QOL at baseline, and significant indirect effects toward QOL at follow-up were found for satisfaction with services and for social support. Rehabilitation plans for people with schizophrenia living in RFs should also consider mediators of change in subjective QOL such as satisfaction with mental health services. Copyright © 2014 Elsevier Ireland Ltd. All rights reserved.
Can shoulder dystocia be reliably predicted?

PubMed

Dodd, Jodie M; Catcheside, Britt; Scheil, Wendy

2012-06-01

To evaluate factors reported to increase the risk of shoulder dystocia, and to evaluate their predictive value at a population level. The South Australian Pregnancy Outcome Unit's population database from 2005 to 2010 was accessed to determine the occurrence of shoulder dystocia in addition to reported risk factors, including age, parity, self-reported ethnicity, presence of diabetes and infant birth weight. Odds ratios (and 95% confidence interval) of shoulder dystocia was calculated for each risk factor, which were then incorporated into a logistic regression model. Test characteristics for each variable in predicting shoulder dystocia were calculated. As a proportion of all births, the reported rate of shoulder dystocia increased significantly from 0.95% in 2005 to 1.38% in 2010 (P = 0.0002). Using a logistic regression model, induction of labour and infant birth weight greater than both 4000 and 4500 g were identified as significant independent predictors of shoulder dystocia. The value of risk factors alone and when incorporated into the logistic regression model was poorly predictive of the occurrence of shoulder dystocia. While there are a number of factors associated with an increased risk of shoulder dystocia, none are of sufficient sensitivity or positive predictive value to allow their use clinically to reliably and accurately identify the occurrence of shoulder dystocia. © 2012 The Authors ANZJOG © 2012 The Royal Australian and New Zealand College of Obstetricians and Gynaecologists.
Assessing the potential for improving S2S forecast skill through multimodel ensembling

NASA Astrophysics Data System (ADS)

Vigaud, N.; Robertson, A. W.; Tippett, M. K.; Wang, L.; Bell, M. J.

2016-12-01

Non-linear logistic regression is well suited to probability forecasting and has been successfully applied in the past to ensemble weather and climate predictions, providing access to the full probabilities distribution without any Gaussian assumption. However, little work has been done at sub-monthly lead times where relatively small re-forecast ensembles and lengths represent new challenges for which post-processing avenues have yet to be investigated. A promising approach consists in extending the definition of non-linear logistic regression by including the quantile of the forecast distribution as one of the predictors. So-called Extended Logistic Regression (ELR), which enables mutually consistent individual threshold probabilities, is here applied to ECMWF, CFSv2 and CMA re-forecasts from the S2S database in order to produce rainfall probabilities at weekly resolution. The ELR model is trained on seasonally-varying tercile categories computed for lead times of 1 to 4 weeks. It is then tested in a cross-validated manner, i.e. allowing real-time predictability applications, to produce rainfall tercile probabilities from individual weekly hindcasts that are finally combined by equal pooling. Results will be discussed over a broader North American region, where individual and MME forecasts generated out to 4 weeks lead are characterized by good probabilistic reliability but low sharpness, exhibiting systematically more skill in winter than summer.
Sample size estimation for alternating logistic regressions analysis of multilevel randomized community trials of under-age drinking.

PubMed

Reboussin, Beth A; Preisser, John S; Song, Eun-Young; Wolfson, Mark

2012-07-01

Under-age drinking is an enormous public health issue in the USA. Evidence that community level structures may impact on under-age drinking has led to a proliferation of efforts to change the environment surrounding the use of alcohol. Although the focus of these efforts is to reduce drinking by individual youths, environmental interventions are typically implemented at the community level with entire communities randomized to the same intervention condition. A distinct feature of these trials is the tendency of the behaviours of individuals residing in the same community to be more alike than that of others residing in different communities, which is herein called 'clustering'. Statistical analyses and sample size calculations must account for this clustering to avoid type I errors and to ensure an appropriately powered trial. Clustering itself may also be of scientific interest. We consider the alternating logistic regressions procedure within the population-averaged modelling framework to estimate the effect of a law enforcement intervention on the prevalence of under-age drinking behaviours while modelling the clustering at multiple levels, e.g. within communities and within neighbourhoods nested within communities, by using pairwise odds ratios. We then derive sample size formulae for estimating intervention effects when planning a post-test-only or repeated cross-sectional community-randomized trial using the alternating logistic regressions procedure.
Neck-focused panic attacks among Cambodian refugees; a logistic and linear regression analysis.

PubMed

Hinton, Devon E; Chhean, Dara; Pich, Vuth; Um, Khin; Fama, Jeanne M; Pollack, Mark H

2006-01-01

Consecutive Cambodian refugees attending a psychiatric clinic were assessed for the presence and severity of current--i.e., at least one episode in the last month--neck-focused panic. Among the whole sample (N=130), in a logistic regression analysis, the Anxiety Sensitivity Index (ASI; odds ratio=3.70) and the Clinician-Administered PTSD Scale (CAPS; odds ratio=2.61) significantly predicted the presence of current neck panic (NP). Among the neck panic patients (N=60), in the linear regression analysis, NP severity was significantly predicted by NP-associated flashbacks (beta=.42), NP-associated catastrophic cognitions (beta=.22), and CAPS score (beta=.28). Further analysis revealed the effect of the CAPS score to be significantly mediated (Sobel test [Baron, R. M., & Kenny, D. A. (1986). The moderator-mediator variable distinction in social psychological research: conceptual, strategic, and statistical considerations. Journal of Personality and Social Psychology, 51, 1173-1182]) by both NP-associated flashbacks and catastrophic cognitions. In the care of traumatized Cambodian refugees, NP severity, as well as NP-associated flashbacks and catastrophic cognitions, should be specifically assessed and treated.

[Analysis on willingness to pay for HIV antibody saliva rapid test and related factors].

PubMed

Li, Junjie; Huo, Junli; Cui, Wenqing; Zhang, Xiujie; Hu, Yi; Su, Xingfang; Zhang, Wanyue; Li, Youfang; Shi, Yuhua; Jia, Manhong

2015-02-01

To understand the willingness to pay for HIV antibody saliva rapid test and its influential factors among people seeking counsel and HIV test, STD clinic patients, university students, migrant people, female sex workers (FSWs), men who have sex with men (MSM) and injecting drug users (IDUs). An anonymous questionnaire survey was conducted among 511 subjects in the 7 groups selected by different sampling methods, and 509 valid questionnaires were collected. The majority of subjects were males (54.8%) and aged 20-29 years (41.5%). Among the subjects, 60.3% had education level of high school or above, 55.4% were unmarried, 37.3% were unemployed, 73.3% had monthly expenditure <2 000 Yuan RMB, 44.2% had received HIV test, 28.3% knew HIV saliva test, 21.0% were willing to receive HIV saliva test, 2.0% had received HIV saliva test, only 1.0% had bought HIV test kit for self-test, and 84.1% were willing to pay for HIV antibody saliva rapid test. Univariate logistic regression analysis indicated that subject group, age, education level, employment status, monthly expenditure level, HIV test experience and willingness to receive HIV saliva test were correlated statistically with willingness to pay for HIV antibody saliva rapid test. Multivariate logistic regression analysis showed that subject group and monthly expenditure level were statistically correlated with willingness to pay for HIV antibody saliva rapid test. The willingness to pay for HIV antibody saliva rapid test and acceptable price of HIV antibody saliva rapid test varied in different areas and populations. Different populations may have different willingness to pay for HIV antibody saliva rapid test;the affordability of the test could influence the willingness to pay for the test.
Support vector machine learning model for the prediction of sentinel node status in patients with cutaneous melanoma.

PubMed

Mocellin, Simone; Ambrosi, Alessandro; Montesco, Maria Cristina; Foletto, Mirto; Zavagno, Giorgio; Nitti, Donato; Lise, Mario; Rossi, Carlo Riccardo

2006-08-01

Currently, approximately 80% of melanoma patients undergoing sentinel node biopsy (SNB) have negative sentinel lymph nodes (SLNs), and no prediction system is reliable enough to be implemented in the clinical setting to reduce the number of SNB procedures. In this study, the predictive power of support vector machine (SVM)-based statistical analysis was tested. The clinical records of 246 patients who underwent SNB at our institution were used for this analysis. The following clinicopathologic variables were considered: the patient's age and sex and the tumor's histological subtype, Breslow thickness, Clark level, ulceration, mitotic index, lymphocyte infiltration, regression, angiolymphatic invasion, microsatellitosis, and growth phase. The results of SVM-based prediction of SLN status were compared with those achieved with logistic regression. The SLN positivity rate was 22% (52 of 234). When the accuracy was > or = 80%, the negative predictive value, positive predictive value, specificity, and sensitivity were 98%, 54%, 94%, and 77% and 82%, 41%, 69%, and 93% by using SVM and logistic regression, respectively. Moreover, SVM and logistic regression were associated with a diagnostic error and an SNB percentage reduction of (1) 1% and 60% and (2) 15% and 73%, respectively. The results from this pilot study suggest that SVM-based prediction of SLN status might be evaluated as a prognostic method to avoid the SNB procedure in 60% of patients currently eligible, with a very low error rate. If validated in larger series, this strategy would lead to obvious advantages in terms of both patient quality of life and costs for the health care system.
Missed opportunities for concurrent HIV-STD testing in an academic emergency department.

PubMed

Klein, Pamela W; Martin, Ian B K; Quinlivan, Evelyn B; Gay, Cynthia L; Leone, Peter A

2014-01-01

We evaluated emergency department (ED) provider adherence to guidelines for concurrent HIV-sexually transmitted disease (STD) testing within an expanded HIV testing program and assessed demographic and clinical factors associated with concurrent HIV-STD testing. We examined concurrent HIV-STD testing in a suburban academic ED with a targeted, expanded HIV testing program. Patients aged 18-64 years who were tested for syphilis, gonorrhea, or chlamydia in 2009 were evaluated for concurrent HIV testing. We analyzed demographic and clinical factors associated with concurrent HIV-STD testing using multivariate logistic regression with a robust variance estimator or, where applicable, exact logistic regression. Only 28.3% of patients tested for syphilis, 3.8% tested for gonorrhea, and 3.8% tested for chlamydia were concurrently tested for HIV during an ED visit. Concurrent HIV-syphilis testing was more likely among younger patients aged 25-34 years (adjusted odds ratio [AOR] = 0.36, 95% confidence interval [CI] 0.78, 2.10) and patients with STD-related chief complaints at triage (AOR=11.47, 95% CI 5.49, 25.06). Concurrent HIV-gonorrhea/chlamydia testing was more likely among men (gonorrhea: AOR=3.98, 95% CI 2.25, 7.02; chlamydia: AOR=3.25, 95% CI 1.80, 5.86) and less likely among patients with STD-related chief complaints at triage (gonorrhea: AOR=0.31, 95% CI 0.13, 0.82; chlamydia: AOR=0.21, 95% CI 0.09, 0.50). Concurrent HIV-STD testing in an academic ED remains low. Systematic interventions that remove the decision-making burden of ordering an HIV test from providers may increase HIV testing in this high-risk population of suspected STD patients.
A simple approach to power and sample size calculations in logistic regression and Cox regression models.

PubMed

Vaeth, Michael; Skovlund, Eva

2004-06-15

For a given regression problem it is possible to identify a suitably defined equivalent two-sample problem such that the power or sample size obtained for the two-sample problem also applies to the regression problem. For a standard linear regression model the equivalent two-sample problem is easily identified, but for generalized linear models and for Cox regression models the situation is more complicated. An approximately equivalent two-sample problem may, however, also be identified here. In particular, we show that for logistic regression and Cox regression models the equivalent two-sample problem is obtained by selecting two equally sized samples for which the parameters differ by a value equal to the slope times twice the standard deviation of the independent variable and further requiring that the overall expected number of events is unchanged. In a simulation study we examine the validity of this approach to power calculations in logistic regression and Cox regression models. Several different covariate distributions are considered for selected values of the overall response probability and a range of alternatives. For the Cox regression model we consider both constant and non-constant hazard rates. The results show that in general the approach is remarkably accurate even in relatively small samples. Some discrepancies are, however, found in small samples with few events and a highly skewed covariate distribution. Comparison with results based on alternative methods for logistic regression models with a single continuous covariate indicates that the proposed method is at least as good as its competitors. The method is easy to implement and therefore provides a simple way to extend the range of problems that can be covered by the usual formulas for power and sample size determination. Copyright 2004 John Wiley & Sons, Ltd.
Robust logistic regression to narrow down the winner's curse for rare and recessive susceptibility variants.

PubMed

Kesselmeier, Miriam; Lorenzo Bermejo, Justo

2017-11-01

Logistic regression is the most common technique used for genetic case-control association studies. A disadvantage of standard maximum likelihood estimators of the genotype relative risk (GRR) is their strong dependence on outlier subjects, for example, patients diagnosed at unusually young age. Robust methods are available to constrain outlier influence, but they are scarcely used in genetic studies. This article provides a non-intimidating introduction to robust logistic regression, and investigates its benefits and limitations in genetic association studies. We applied the bounded Huber and extended the R package 'robustbase' with the re-descending Hampel functions to down-weight outlier influence. Computer simulations were carried out to assess the type I error rate, mean squared error (MSE) and statistical power according to major characteristics of the genetic study and investigated markers. Simulations were complemented with the analysis of real data. Both standard and robust estimation controlled type I error rates. Standard logistic regression showed the highest power but standard GRR estimates also showed the largest bias and MSE, in particular for associated rare and recessive variants. For illustration, a recessive variant with a true GRR=6.32 and a minor allele frequency=0.05 investigated in a 1000 case/1000 control study by standard logistic regression resulted in power=0.60 and MSE=16.5. The corresponding figures for Huber-based estimation were power=0.51 and MSE=0.53. Overall, Hampel- and Huber-based GRR estimates did not differ much. Robust logistic regression may represent a valuable alternative to standard maximum likelihood estimation when the focus lies on risk prediction rather than identification of susceptibility variants. © The Author 2016. Published by Oxford University Press. All rights reserved. For Permissions, please email: journals.permissions@oup.com.
Analysis of training sample selection strategies for regression-based quantitative landslide susceptibility mapping methods

NASA Astrophysics Data System (ADS)

Erener, Arzu; Sivas, A. Abdullah; Selcuk-Kestel, A. Sevtap; Düzgün, H. Sebnem

2017-07-01

All of the quantitative landslide susceptibility mapping (QLSM) methods requires two basic data types, namely, landslide inventory and factors that influence landslide occurrence (landslide influencing factors, LIF). Depending on type of landslides, nature of triggers and LIF, accuracy of the QLSM methods differs. Moreover, how to balance the number of 0 (nonoccurrence) and 1 (occurrence) in the training set obtained from the landslide inventory and how to select which one of the 1's and 0's to be included in QLSM models play critical role in the accuracy of the QLSM. Although performance of various QLSM methods is largely investigated in the literature, the challenge of training set construction is not adequately investigated for the QLSM methods. In order to tackle this challenge, in this study three different training set selection strategies along with the original data set is used for testing the performance of three different regression methods namely Logistic Regression (LR), Bayesian Logistic Regression (BLR) and Fuzzy Logistic Regression (FLR). The first sampling strategy is proportional random sampling (PRS), which takes into account a weighted selection of landslide occurrences in the sample set. The second method, namely non-selective nearby sampling (NNS), includes randomly selected sites and their surrounding neighboring points at certain preselected distances to include the impact of clustering. Selective nearby sampling (SNS) is the third method, which concentrates on the group of 1's and their surrounding neighborhood. A randomly selected group of landslide sites and their neighborhood are considered in the analyses similar to NNS parameters. It is found that LR-PRS, FLR-PRS and BLR-Whole Data set-ups, with order, yield the best fits among the other alternatives. The results indicate that in QLSM based on regression models, avoidance of spatial correlation in the data set is critical for the model's performance.
Sperm Retrieval in Patients with Klinefelter Syndrome: A Skewed Regression Model Analysis.

PubMed

Chehrazi, Mohammad; Rahimiforoushani, Abbas; Sabbaghian, Marjan; Nourijelyani, Keramat; Sadighi Gilani, Mohammad Ali; Hoseini, Mostafa; Vesali, Samira; Yaseri, Mehdi; Alizadeh, Ahad; Mohammad, Kazem; Samani, Reza Omani

2017-01-01

The most common chromosomal abnormality due to non-obstructive azoospermia (NOA) is Klinefelter syndrome (KS) which occurs in 1-1.72 out of 500-1000 male infants. The probability of retrieving sperm as the outcome could be asymmetrically different between patients with and without KS, therefore logistic regression analysis is not a well-qualified test for this type of data. This study has been designed to evaluate skewed regression model analysis for data collected from microsurgical testicular sperm extraction (micro-TESE) among azoospermic patients with and without non-mosaic KS syndrome. This cohort study compared the micro-TESE outcome between 134 men with classic KS and 537 men with NOA and normal karyotype who were referred to Royan Institute between 2009 and 2011. In addition to our main outcome, which was sperm retrieval, we also used logistic and skewed regression analyses to compare the following demographic and hormonal factors: age, level of follicle stimulating hormone (FSH), luteinizing hormone (LH), and testosterone between the two groups. A comparison of the micro-TESE between the KS and control groups showed a success rate of 28.4% (38/134) for the KS group and 22.2% (119/537) for the control group. In the KS group, a significantly difference (P<0.001) existed between testosterone levels for the successful sperm retrieval group (3.4 ± 0.48 mg/mL) compared to the unsuccessful sperm retrieval group (2.33 ± 0.23 mg/mL). The index for quasi Akaike information criterion (QAIC) had a goodness of fit of 74 for the skewed model which was lower than logistic regression (QAIC=85). According to the results, skewed regression is more efficient in estimating sperm retrieval success when the data from patients with KS are analyzed. This finding should be investigated by conducting additional studies with different data structures.
Development and Validation of a Deep Neural Network Model for Prediction of Postoperative In-hospital Mortality.

PubMed

Lee, Christine K; Hofer, Ira; Gabel, Eilon; Baldi, Pierre; Cannesson, Maxime

2018-04-17

The authors tested the hypothesis that deep neural networks trained on intraoperative features can predict postoperative in-hospital mortality. The data used to train and validate the algorithm consists of 59,985 patients with 87 features extracted at the end of surgery. Feed-forward networks with a logistic output were trained using stochastic gradient descent with momentum. The deep neural networks were trained on 80% of the data, with 20% reserved for testing. The authors assessed improvement of the deep neural network by adding American Society of Anesthesiologists (ASA) Physical Status Classification and robustness of the deep neural network to a reduced feature set. The networks were then compared to ASA Physical Status, logistic regression, and other published clinical scores including the Surgical Apgar, Preoperative Score to Predict Postoperative Mortality, Risk Quantification Index, and the Risk Stratification Index. In-hospital mortality in the training and test sets were 0.81% and 0.73%. The deep neural network with a reduced feature set and ASA Physical Status classification had the highest area under the receiver operating characteristics curve, 0.91 (95% CI, 0.88 to 0.93). The highest logistic regression area under the curve was found with a reduced feature set and ASA Physical Status (0.90, 95% CI, 0.87 to 0.93). The Risk Stratification Index had the highest area under the receiver operating characteristics curve, at 0.97 (95% CI, 0.94 to 0.99). Deep neural networks can predict in-hospital mortality based on automatically extractable intraoperative data, but are not (yet) superior to existing methods.
Predicting diabetes mellitus using SMOTE and ensemble machine learning approach: The Henry Ford ExercIse Testing (FIT) project.

PubMed

Alghamdi, Manal; Al-Mallah, Mouaz; Keteyian, Steven; Brawner, Clinton; Ehrman, Jonathan; Sakr, Sherif

2017-01-01

Machine learning is becoming a popular and important approach in the field of medical research. In this study, we investigate the relative performance of various machine learning methods such as Decision Tree, Naïve Bayes, Logistic Regression, Logistic Model Tree and Random Forests for predicting incident diabetes using medical records of cardiorespiratory fitness. In addition, we apply different techniques to uncover potential predictors of diabetes. This FIT project study used data of 32,555 patients who are free of any known coronary artery disease or heart failure who underwent clinician-referred exercise treadmill stress testing at Henry Ford Health Systems between 1991 and 2009 and had a complete 5-year follow-up. At the completion of the fifth year, 5,099 of those patients have developed diabetes. The dataset contained 62 attributes classified into four categories: demographic characteristics, disease history, medication use history, and stress test vital signs. We developed an Ensembling-based predictive model using 13 attributes that were selected based on their clinical importance, Multiple Linear Regression, and Information Gain Ranking methods. The negative effect of the imbalance class of the constructed model was handled by Synthetic Minority Oversampling Technique (SMOTE). The overall performance of the predictive model classifier was improved by the Ensemble machine learning approach using the Vote method with three Decision Trees (Naïve Bayes Tree, Random Forest, and Logistic Model Tree) and achieved high accuracy of prediction (AUC = 0.92). The study shows the potential of ensembling and SMOTE approaches for predicting incident diabetes using cardiorespiratory fitness data.
A logistic regression equation for estimating the probability of a stream in Vermont having intermittent flow

USGS Publications Warehouse

Olson, Scott A.; Brouillette, Michael C.

2006-01-01

A logistic regression equation was developed for estimating the probability of a stream flowing intermittently at unregulated, rural stream sites in Vermont. These determinations can be used for a wide variety of regulatory and planning efforts at the Federal, State, regional, county and town levels, including such applications as assessing fish and wildlife habitats, wetlands classifications, recreational opportunities, water-supply potential, waste-assimilation capacities, and sediment transport. The equation will be used to create a derived product for the Vermont Hydrography Dataset having the streamflow characteristic of 'intermittent' or 'perennial.' The Vermont Hydrography Dataset is Vermont's implementation of the National Hydrography Dataset and was created at a scale of 1:5,000 based on statewide digital orthophotos. The equation was developed by relating field-verified perennial or intermittent status of a stream site during normal summer low-streamflow conditions in the summer of 2005 to selected basin characteristics of naturally flowing streams in Vermont. The database used to develop the equation included 682 stream sites with drainage areas ranging from 0.05 to 5.0 square miles. When the 682 sites were observed, 126 were intermittent (had no flow at the time of the observation) and 556 were perennial (had flowing water at the time of the observation). The results of the logistic regression analysis indicate that the probability of a stream having intermittent flow in Vermont is a function of drainage area, elevation of the site, the ratio of basin relief to basin perimeter, and the areal percentage of well- and moderately well-drained soils in the basin. Using a probability cutpoint (a lower probability indicates the site has perennial flow and a higher probability indicates the site has intermittent flow) of 0.5, the logistic regression equation correctly predicted the perennial or intermittent status of 116 test sites 85 percent of the time.
Nonconvex Sparse Logistic Regression With Weakly Convex Regularization

NASA Astrophysics Data System (ADS)

Shen, Xinyue; Gu, Yuantao

2018-06-01

In this work we propose to fit a sparse logistic regression model by a weakly convex regularized nonconvex optimization problem. The idea is based on the finding that a weakly convex function as an approximation of the $\\ell_0$ pseudo norm is able to better induce sparsity than the commonly used $\\ell_1$ norm. For a class of weakly convex sparsity inducing functions, we prove the nonconvexity of the corresponding sparse logistic regression problem, and study its local optimality conditions and the choice of the regularization parameter to exclude trivial solutions. Despite the nonconvexity, a method based on proximal gradient descent is used to solve the general weakly convex sparse logistic regression, and its convergence behavior is studied theoretically. Then the general framework is applied to a specific weakly convex function, and a necessary and sufficient local optimality condition is provided. The solution method is instantiated in this case as an iterative firm-shrinkage algorithm, and its effectiveness is demonstrated in numerical experiments by both randomly generated and real datasets.
A comparative study on entrepreneurial attitudes modeled with logistic regression and Bayes nets.

PubMed

López Puga, Jorge; García García, Juan

2012-11-01

Entrepreneurship research is receiving increasing attention in our context, as entrepreneurs are key social agents involved in economic development. We compare the success of the dichotomic logistic regression model and the Bayes simple classifier to predict entrepreneurship, after manipulating the percentage of missing data and the level of categorization in predictors. A sample of undergraduate university students (N = 1230) completed five scales (motivation, attitude towards business creation, obstacles, deficiencies, and training needs) and we found that each of them predicted different aspects of the tendency to business creation. Additionally, our results show that the receiver operating characteristic (ROC) curve is affected by the rate of missing data in both techniques, but logistic regression seems to be more vulnerable when faced with missing data, whereas Bayes nets underperform slightly when categorization has been manipulated. Our study sheds light on the potential entrepreneur profile and we propose to use Bayesian networks as an additional alternative to overcome the weaknesses of logistic regression when missing data are present in applied research.
Epidemiologic programs for computers and calculators. A microcomputer program for multiple logistic regression by unconditional and conditional maximum likelihood methods.

PubMed

Campos-Filho, N; Franco, E L

1989-02-01

A frequent procedure in matched case-control studies is to report results from the multivariate unmatched analyses if they do not differ substantially from the ones obtained after conditioning on the matching variables. Although conceptually simple, this rule requires that an extensive series of logistic regression models be evaluated by both the conditional and unconditional maximum likelihood methods. Most computer programs for logistic regression employ only one maximum likelihood method, which requires that the analyses be performed in separate steps. This paper describes a Pascal microcomputer (IBM PC) program that performs multiple logistic regression by both maximum likelihood estimation methods, which obviates the need for switching between programs to obtain relative risk estimates from both matched and unmatched analyses. The program calculates most standard statistics and allows factoring of categorical or continuous variables by two distinct methods of contrast. A built-in, descriptive statistics option allows the user to inspect the distribution of cases and controls across categories of any given variable.
Comparison of cranial sex determination by discriminant analysis and logistic regression.

PubMed

Amores-Ampuero, Anabel; Alemán, Inmaculada

2016-04-05

Various methods have been proposed for estimating dimorphism. The objective of this study was to compare sex determination results from cranial measurements using discriminant analysis or logistic regression. The study sample comprised 130 individuals (70 males) of known sex, age, and cause of death from San José cemetery in Granada (Spain). Measurements of 19 neurocranial dimensions and 11 splanchnocranial dimensions were subjected to discriminant analysis and logistic regression, and the percentages of correct classification were compared between the sex functions obtained with each method. The discriminant capacity of the selected variables was evaluated with a cross-validation procedure. The percentage accuracy with discriminant analysis was 78.2% for the neurocranium (82.4% in females and 74.6% in males) and 73.7% for the splanchnocranium (79.6% in females and 68.8% in males). These percentages were higher with logistic regression analysis: 85.7% for the neurocranium (in both sexes) and 94.1% for the splanchnocranium (100% in females and 91.7% in males).
Accounting for center in the Early External Cephalic Version trials: an empirical comparison of statistical methods to adjust for center in a multicenter trial with binary outcomes.

PubMed

Reitsma, Angela; Chu, Rong; Thorpe, Julia; McDonald, Sarah; Thabane, Lehana; Hutton, Eileen

2014-09-26

Clustering of outcomes at centers involved in multicenter trials is a type of center effect. The Consolidated Standards of Reporting Trials Statement recommends that multicenter randomized controlled trials (RCTs) should account for center effects in their analysis, however most do not. The Early External Cephalic Version (EECV) trials published in 2003 and 2011 stratified by center at randomization, but did not account for center in the analyses, and due to the nature of the intervention and number of centers, may have been prone to center effects. Using data from the EECV trials, we undertook an empirical study to compare various statistical approaches to account for center effect while estimating the impact of external cephalic version timing (early or delayed) on the outcomes of cesarean section, preterm birth, and non-cephalic presentation at the time of birth. The data from the EECV pilot trial and the EECV2 trial were merged into one dataset. Fisher's exact method was used to test the overall effect of external cephalic version timing unadjusted for center effects. Seven statistical models that accounted for center effects were applied to the data. The models included: i) the Mantel-Haenszel test, ii) logistic regression with fixed center effect and fixed treatment effect, iii) center-size weighted and iv) un-weighted logistic regression with fixed center effect and fixed treatment-by-center interaction, iv) logistic regression with random center effect and fixed treatment effect, v) logistic regression with random center effect and random treatment-by-center interaction, and vi) generalized estimating equations. For each of the three outcomes of interest approaches to account for center effect did not alter the overall findings of the trial. The results were similar for the majority of the methods used to adjust for center, illustrating the robustness of the findings. Despite literature that suggests center effect can change the estimate of effect in multicenter trials, this empirical study does not show a difference in the outcomes of the EECV trials when accounting for center effect. The EECV2 trial was registered on 30 July 30 2005 with Current Controlled Trials: ISRCTN 56498577.
Statistical analysis and interpretation of prenatal diagnostic imaging studies, Part 2: descriptive and inferential statistical methods.

PubMed

Tuuli, Methodius G; Odibo, Anthony O

2011-08-01

The objective of this article is to discuss the rationale for common statistical tests used for the analysis and interpretation of prenatal diagnostic imaging studies. Examples from the literature are used to illustrate descriptive and inferential statistics. The uses and limitations of linear and logistic regression analyses are discussed in detail.
Motivation towards Medical Career Choice and Future Career Plans of Polish Medical Students

ERIC Educational Resources Information Center

Gasiorowski, Jakub; Rudowicz, Elzbieta; Safranow, Krzysztof

2015-01-01

This longitudinal study aimed at investigating Polish medical students' career choice motivation, factors influencing specialty choices, professional plans and expectations. The same cohort of students responded to the same questionnaire, at the end of Year 1 and Year 6. The Chi-square, Mann-Whitney U tests and logistic regression were used in…
Educational Subculture and Dropping out in Higher Education: A Longitudinal Case Study

ERIC Educational Resources Information Center

Venuleo, C.; Mossi, P.; Salvatore, S.

2016-01-01

The paper tests longitudinally the hypothesis that educational subcultures in terms of which students interpret their role and their educational setting affect the probability of dropping out of higher education. A logistic regression model was performed to predict drop out at the beginning of the second academic year for the 823 freshmen of a…
Unmet Dental Needs and Barriers to Dental Care among Children with Autism Spectrum Disorders

ERIC Educational Resources Information Center

Lai, Bien; Milano, Michael; Roberts, Michael W.; Hooper, Stephen R.

2012-01-01

Mail-in pilot-tested questionnaires were sent to a stratified random sample of 1,500 families from the North Carolina Autism Registry. Multivariate logistic regression analysis was used to determine the significance of unmet dental needs and other predictors. Of 568 surveys returned (Response Rate = 38%), 555 were complete and usable. Sixty-five…
Exploring Crossing Differential Item Functioning by Gender in Mathematics Assessment

ERIC Educational Resources Information Center

Ong, Yoke Mooi; Williams, Julian; Lamprianou, Iasonas

2015-01-01

The purpose of this article is to explore crossing differential item functioning (DIF) in a test drawn from a national examination of mathematics for 11-year-old pupils in England. An empirical dataset was analyzed to explore DIF by gender in a mathematics assessment. A two-step process involving the logistic regression (LR) procedure for…

Testing an Online English Course: Lessons Learned from an Analysis of Postcourse Proficiency Change Scores

ERIC Educational Resources Information Center

Jee, Rebecca Y.

2015-01-01

Voxy, an English-language-learning company, has developed a custom, in-house proficiency exam, the Voxy Proficiency Assessment (VPA), which is given to all learners at the beginning and end of their courses. Using Multinomial Logistic Regression (MLR), the impact of covariates, such as total learning activities completed and total number of…
Exploring Person Fit with an Approach Based on Multilevel Logistic Regression

ERIC Educational Resources Information Center

Walker, A. Adrienne; Engelhard, George, Jr.

2015-01-01

The idea that test scores may not be valid representations of what students know, can do, and should learn next is well known. Person fit provides an important aspect of validity evidence. Person fit analyses at the individual student level are not typically conducted and person fit information is not communicated to educational stakeholders. In…
Stepwise Distributed Open Innovation Contests for Software Development: Acceleration of Genome-Wide Association Analysis

PubMed Central

Hill, Andrew; Loh, Po-Ru; Bharadwaj, Ragu B.; Pons, Pascal; Shang, Jingbo; Guinan, Eva; Lakhani, Karim; Kilty, Iain

2017-01-01

Abstract Background: The association of differing genotypes with disease-related phenotypic traits offers great potential to both help identify new therapeutic targets and support stratification of patients who would gain the greatest benefit from specific drug classes. Development of low-cost genotyping and sequencing has made collecting large-scale genotyping data routine in population and therapeutic intervention studies. In addition, a range of new technologies is being used to capture numerous new and complex phenotypic descriptors. As a result, genotype and phenotype datasets have grown exponentially. Genome-wide association studies associate genotypes and phenotypes using methods such as logistic regression. As existing tools for association analysis limit the efficiency by which value can be extracted from increasing volumes of data, there is a pressing need for new software tools that can accelerate association analyses on large genotype-phenotype datasets. Results: Using open innovation (OI) and contest-based crowdsourcing, the logistic regression analysis in a leading, community-standard genetics software package (PLINK 1.07) was substantially accelerated. OI allowed us to do this in <6 months by providing rapid access to highly skilled programmers with specialized, difficult-to-find skill sets. Through a crowd-based contest a combination of computational, numeric, and algorithmic approaches was identified that accelerated the logistic regression in PLINK 1.07 by 18- to 45-fold. Combining contest-derived logistic regression code with coarse-grained parallelization, multithreading, and associated changes to data initialization code further developed through distributed innovation, we achieved an end-to-end speedup of 591-fold for a data set size of 6678 subjects by 645 863 variants, compared to PLINK 1.07's logistic regression. This represents a reduction in run time from 4.8 hours to 29 seconds. Accelerated logistic regression code developed in this project has been incorporated into the PLINK2 project. Conclusions: Using iterative competition-based OI, we have developed a new, faster implementation of logistic regression for genome-wide association studies analysis. We present lessons learned and recommendations on running a successful OI process for bioinformatics. PMID:28327993
Stepwise Distributed Open Innovation Contests for Software Development: Acceleration of Genome-Wide Association Analysis.

PubMed

Hill, Andrew; Loh, Po-Ru; Bharadwaj, Ragu B; Pons, Pascal; Shang, Jingbo; Guinan, Eva; Lakhani, Karim; Kilty, Iain; Jelinsky, Scott A

2017-05-01

The association of differing genotypes with disease-related phenotypic traits offers great potential to both help identify new therapeutic targets and support stratification of patients who would gain the greatest benefit from specific drug classes. Development of low-cost genotyping and sequencing has made collecting large-scale genotyping data routine in population and therapeutic intervention studies. In addition, a range of new technologies is being used to capture numerous new and complex phenotypic descriptors. As a result, genotype and phenotype datasets have grown exponentially. Genome-wide association studies associate genotypes and phenotypes using methods such as logistic regression. As existing tools for association analysis limit the efficiency by which value can be extracted from increasing volumes of data, there is a pressing need for new software tools that can accelerate association analyses on large genotype-phenotype datasets. Using open innovation (OI) and contest-based crowdsourcing, the logistic regression analysis in a leading, community-standard genetics software package (PLINK 1.07) was substantially accelerated. OI allowed us to do this in <6 months by providing rapid access to highly skilled programmers with specialized, difficult-to-find skill sets. Through a crowd-based contest a combination of computational, numeric, and algorithmic approaches was identified that accelerated the logistic regression in PLINK 1.07 by 18- to 45-fold. Combining contest-derived logistic regression code with coarse-grained parallelization, multithreading, and associated changes to data initialization code further developed through distributed innovation, we achieved an end-to-end speedup of 591-fold for a data set size of 6678 subjects by 645 863 variants, compared to PLINK 1.07's logistic regression. This represents a reduction in run time from 4.8 hours to 29 seconds. Accelerated logistic regression code developed in this project has been incorporated into the PLINK2 project. Using iterative competition-based OI, we have developed a new, faster implementation of logistic regression for genome-wide association studies analysis. We present lessons learned and recommendations on running a successful OI process for bioinformatics. © The Author 2017. Published by Oxford University Press.
Easy and low-cost identification of metabolic syndrome in patients treated with second-generation antipsychotics: artificial neural network and logistic regression models.

PubMed

Lin, Chao-Cheng; Bai, Ya-Mei; Chen, Jen-Yeu; Hwang, Tzung-Jeng; Chen, Tzu-Ting; Chiu, Hung-Wen; Li, Yu-Chuan

2010-03-01

Metabolic syndrome (MetS) is an important side effect of second-generation antipsychotics (SGAs). However, many SGA-treated patients with MetS remain undetected. In this study, we trained and validated artificial neural network (ANN) and multiple logistic regression models without biochemical parameters to rapidly identify MetS in patients with SGA treatment. A total of 383 patients with a diagnosis of schizophrenia or schizoaffective disorder (DSM-IV criteria) with SGA treatment for more than 6 months were investigated to determine whether they met the MetS criteria according to the International Diabetes Federation. The data for these patients were collected between March 2005 and September 2005. The input variables of ANN and logistic regression were limited to demographic and anthropometric data only. All models were trained by randomly selecting two-thirds of the patient data and were internally validated with the remaining one-third of the data. The models were then externally validated with data from 69 patients from another hospital, collected between March 2008 and June 2008. The area under the receiver operating characteristic curve (AUC) was used to measure the performance of all models. Both the final ANN and logistic regression models had high accuracy (88.3% vs 83.6%), sensitivity (93.1% vs 86.2%), and specificity (86.9% vs 83.8%) to identify MetS in the internal validation set. The mean +/- SD AUC was high for both the ANN and logistic regression models (0.934 +/- 0.033 vs 0.922 +/- 0.035, P = .63). During external validation, high AUC was still obtained for both models. Waist circumference and diastolic blood pressure were the common variables that were left in the final ANN and logistic regression models. Our study developed accurate ANN and logistic regression models to detect MetS in patients with SGA treatment. The models are likely to provide a noninvasive tool for large-scale screening of MetS in this group of patients. (c) 2010 Physicians Postgraduate Press, Inc.
Bayesian logistic regression in detection of gene-steroid interaction for cancer at PDLIM5 locus.

PubMed

Wang, Ke-Sheng; Owusu, Daniel; Pan, Yue; Xie, Changchun

2016-06-01

The PDZ and LIM domain 5 (PDLIM5) gene may play a role in cancer, bipolar disorder, major depression, alcohol dependence and schizophrenia; however, little is known about the interaction effect of steroid and PDLIM5 gene on cancer. This study examined 47 single-nucleotide polymorphisms (SNPs) within the PDLIM5 gene in the Marshfield sample with 716 cancer patients (any diagnosed cancer, excluding minor skin cancer) and 2848 noncancer controls. Multiple logistic regression model in PLINK software was used to examine the association of each SNP with cancer. Bayesian logistic regression in PROC GENMOD in SAS statistical software, ver. 9.4 was used to detect gene- steroid interactions influencing cancer. Single marker analysis using PLINK identified 12 SNPs associated with cancer (P< 0.05); especially, SNP rs6532496 revealed the strongest association with cancer (P = 6.84 × 10⁻³); while the next best signal was rs951613 (P = 7.46 × 10⁻³). Classic logistic regression in PROC GENMOD showed that both rs6532496 and rs951613 revealed strong gene-steroid interaction effects (OR=2.18, 95% CI=1.31-3.63 with P = 2.9 × 10⁻³ for rs6532496 and OR=2.07, 95% CI=1.24-3.45 with P = 5.43 × 10⁻³ for rs951613, respectively). Results from Bayesian logistic regression showed stronger interaction effects (OR=2.26, 95% CI=1.2-3.38 for rs6532496 and OR=2.14, 95% CI=1.14-3.2 for rs951613, respectively). All the 12 SNPs associated with cancer revealed significant gene-steroid interaction effects (P < 0.05); whereas 13 SNPs showed gene-steroid interaction effects without main effect on cancer. SNP rs4634230 revealed the strongest gene-steroid interaction effect (OR=2.49, 95% CI=1.5-4.13 with P = 4.0 × 10⁻⁴ based on the classic logistic regression and OR=2.59, 95% CI=1.4-3.97 from Bayesian logistic regression; respectively). This study provides evidence of common genetic variants within the PDLIM5 gene and interactions between PLDIM5 gene polymorphisms and steroid use influencing cancer.
Diagnostic Algorithm to Reflect Regressive Changes of Human Papilloma Virus in Tissue Biopsies

PubMed Central

Lhee, Min Jin; Cha, Youn Jin; Bae, Jong Man; Kim, Young Tae

2014-01-01

Purpose Landmark indicators have not yet to be developed to detect the regression of cervical intraepithelial neoplasia (CIN). We propose that quantitative viral load and indicative histological criteria can be used to differentiate between atypical squamous cells of undetermined significance (ASCUS) and a CIN of grade 1. Materials and Methods We collected 115 tissue biopsies from women who tested positive for the human papilloma virus (HPV). Nine morphological parameters including nuclear size, perinuclear halo, hyperchromasia, typical koilocyte (TK), abortive koilocyte (AK), bi-/multi-nucleation, keratohyaline granules, inflammation, and dyskeratosis were examined for each case. Correlation analyses, cumulative logistic regression, and binary logistic regression were used to determine optimal cut-off values of HPV copy numbers. The parameters TK, perinuclear halo, multi-nucleation, and nuclear size were significantly correlated quantitatively to HPV copy number. Results An HPV loading number of 58.9 and AK number of 20 were optimal to discriminate between negative and subtle findings in biopsies. An HPV loading number of 271.49 and AK of 20 were optimal for discriminating between equivocal changes and obvious koilocytosis. Conclusion We propose that a squamous epithelial lesion with AK of >20 and quantitative HPV copy number between 58.9-271.49 represents a new spectrum of subtle pathological findings, characterized by AK in ASCUS. This can be described as a distinct entity and called "regressing koilocytosis". PMID:24532500
Deletion Diagnostics for Alternating Logistic Regressions

PubMed Central

Preisser, John S.; By, Kunthel; Perin, Jamie; Qaqish, Bahjat F.

2013-01-01

Deletion diagnostics are introduced for the regression analysis of clustered binary outcomes estimated with alternating logistic regressions, an implementation of generalized estimating equations (GEE) that estimates regression coefficients in a marginal mean model and in a model for the intracluster association given by the log odds ratio. The diagnostics are developed within an estimating equations framework that recasts the estimating functions for association parameters based upon conditional residuals into equivalent functions based upon marginal residuals. Extensions of earlier work on GEE diagnostics follow directly, including computational formulae for one-step deletion diagnostics that measure the influence of a cluster of observations on the estimated regression parameters and on the overall marginal mean or association model fit. The diagnostic formulae are evaluated with simulations studies and with an application concerning an assessment of factors associated with health maintenance visits in primary care medical practices. The application and the simulations demonstrate that the proposed cluster-deletion diagnostics for alternating logistic regressions are good approximations of their exact fully iterated counterparts. PMID:22777960
Knowledge of HIV testing and attitudes towards blood donation at three blood centres in Brazil

PubMed Central

Miranda, C.; Moreno, E.; Bruhn, R.; Larsen, N. M.; Wright, D. J.; Oliveira, C. D. L.; Carneiro-Proietti, A. B. F.; Loureiro, P.; de Almeida-Neto, C.; Custer, B.; Sabino, E. C.; Gonçalez, T. T.

2015-01-01

Background Reducing risk of HIV window period transmission requires understanding of donor knowledge and attitudes related to HIV and risk factors. Study Design and Methods We conducted a survey of 7635 presenting blood donors at three Brazilian blood centres from 15 October through 20 November 2009. Participants completed a questionnaire on HIV knowledge and attitudes about blood donation. Six questions about blood testing and HIV were evaluated using maximum likelihood chi-square and logistic regression. Test seeking was classified in non-overlapping categories according to answers to one direct and two indirect questions. Results Overall, respondents were male (64%) repeat donors (67%) between 18 and 49 years old (91%). Nearly 60% believed blood centres use better HIV tests than other places; however, 42% were unaware of the HIV window period. Approximately 50% believed it was appropriate to donate to be tested for HIV, but 67% said it was not acceptable to donate with risk factors even if blood is tested. Logistic regression found that less education, Hemope-Recife blood centre, replacement, potential and self-disclosed test-seeking were associated with less HIV knowledge. Conclusion HIV knowledge related to blood safety remains low among Brazilian blood donors. A subset finds it appropriate to be tested at blood centres and may be unaware of the HIV window period. These donations may impose a significant risk to the safety of the blood supply. Decreasing test-seeking and changing beliefs about the appropriateness of individuals with behavioural risk factors donating blood could reduce the risk of transfusing an infectious unit. PMID:24313562
Risk factors for lesions of the knee menisci among workers in South Korea's national parks.

PubMed

Shin, Donghee; Youn, Kanwoo; Lee, Eunja; Lee, Myeongjun; Chung, Hweemin; Kim, Deokweon

2016-01-01

This study was designed to investigate the prevalence of the menisci lesions in national park workers and work factors affecting this prevalence. The study subjects were 698 workers who worked in 20 Korean national parks in 2014. An orthopedist visited each national park and performed physical examinations. Knee MRI was performed if the McMurray test or Apley test was positive and there was a complaint of pain in knee area. An orthopedist and a radiologist respectively read these images of the menisci using a grading system based on the MRI signals. To calculate the cumulative intensity of trekking of the workers, the mean trail distance, the difficulty of the trail, the tenure at each national parks, and the number of treks per month for each worker from the start of work until the present were investigated. Chi-square tests was performed to see if there were differences in the menisci lesions grade according to the variables. The variables used in the Chi-square test were evaluated using simple logistic regression analysis to get crude odds ratios, and adjusted odds ratios and 95 % confidence intervals were calculated using multivariate logistic regression analysis after establishing three different models according to the adjusted variables. According to the MRI signal grades of menisci, 29 % were grade 0, 11.3 % were grade 1, 46.0 % were grade 2, and 13.7 % were grade 3. The differences in the MRI signal grades of menisci according to age and the intensity of trekking as calculated by the three different methods were statistically significant. Multiple logistic regression analysis was performed for three models. In model 1, there was no statistically significant factor affecting the menisci lesions. In model 2, among the factors affecting the menisci lesions, the OR of a high cumulative intensity of trekking was 4.08 (95 % CI 1.00-16.61), and in model 3, the OR of a high cumulative intensity of trekking was 5.84 (95 % CI 1.09-31.26). The factor that most affected the menisci lesions among the workers in Korean national park was a high cumulative intensity of trekking.
Estimating interaction on an additive scale between continuous determinants in a logistic regression model.

PubMed

Knol, Mirjam J; van der Tweel, Ingeborg; Grobbee, Diederick E; Numans, Mattijs E; Geerlings, Mirjam I

2007-10-01

To determine the presence of interaction in epidemiologic research, typically a product term is added to the regression model. In linear regression, the regression coefficient of the product term reflects interaction as departure from additivity. However, in logistic regression it refers to interaction as departure from multiplicativity. Rothman has argued that interaction estimated as departure from additivity better reflects biologic interaction. So far, literature on estimating interaction on an additive scale using logistic regression only focused on dichotomous determinants. The objective of the present study was to provide the methods to estimate interaction between continuous determinants and to illustrate these methods with a clinical example. and results From the existing literature we derived the formulas to quantify interaction as departure from additivity between one continuous and one dichotomous determinant and between two continuous determinants using logistic regression. Bootstrapping was used to calculate the corresponding confidence intervals. To illustrate the theory with an empirical example, data from the Utrecht Health Project were used, with age and body mass index as risk factors for elevated diastolic blood pressure. The methods and formulas presented in this article are intended to assist epidemiologists to calculate interaction on an additive scale between two variables on a certain outcome. The proposed methods are included in a spreadsheet which is freely available at: http://www.juliuscenter.nl/additive-interaction.xls.
Mother-Son Communication about Sex and Routine HIV Testing among Younger Men of Color Who Have Sex with Men

PubMed Central

Bouris, Alida; Hill, Brandon J.; Fisher, Kimberly; Erickson, Greg; Schneider, John A.

2015-01-01

Purpose To document the HIV testing behaviors and serostatus of younger men of color who have sex with men (YMSM), and to explore sociodemographic, behavioral, and maternal correlates of HIV testing in the past six months. Methods 135 YMSM aged 16–19 completed a close-ended survey on HIV testing and risk behaviors, mother-son communication, and sociodemographic characteristics. Youth were offered point-of-care HIV testing, with results provided at survey end. Multivariate logistic regression analyzed the sociodemographic, behavioral, and maternal factors associated with routine HIV testing. Results 90.3% of YMSM had previously tested for HIV and 70.9 % had tested in the past six months. In total, 11.7% of youth reported being HIV-positive and 3.3% reported unknown serostatus. When offered an HIV test, 97.8% accepted. Of these, 14.7% had a positive oral test result and 31.58% of HIV-positive YMSM (n=6) were seropositive unaware. Logistic regression results indicated that maternal communication about sex with males was positively associated with routine testing (OR=2.36; 95% CI=1.13–4.94). Conversely, communication about puberty and general human sexuality was negatively associated (OR=0.45; 95% CI=0.24–0.86). Condomless anal intercourse and positive STI history were negatively associated with routine testing; however, frequency of alcohol use was positively associated. Conclusions Despite high rates of testing, we found high rates of HIV infection, with 31.58% of HIV-positive YMSM being seropositive unaware. Mother-son communication about sex needs to address same-sex behavior, as this appears to be more important than other topics. YMSM with known risk factors for HIV are not testing at the recommended time intervals. PMID:26321527
An artificial neural network prediction model of congenital heart disease based on risk factors: A hospital-based case-control study.

PubMed

Li, Huixia; Luo, Miyang; Zheng, Jianfei; Luo, Jiayou; Zeng, Rong; Feng, Na; Du, Qiyun; Fang, Junqun

2017-02-01

An artificial neural network (ANN) model was developed to predict the risks of congenital heart disease (CHD) in pregnant women.This hospital-based case-control study involved 119 CHD cases and 239 controls all recruited from birth defect surveillance hospitals in Hunan Province between July 2013 and June 2014. All subjects were interviewed face-to-face to fill in a questionnaire that covered 36 CHD-related variables. The 358 subjects were randomly divided into a training set and a testing set at the ratio of 85:15. The training set was used to identify the significant predictors of CHD by univariate logistic regression analyses and develop a standard feed-forward back-propagation neural network (BPNN) model for the prediction of CHD. The testing set was used to test and evaluate the performance of the ANN model. Univariate logistic regression analyses were performed on SPSS 18.0. The ANN models were developed on Matlab 7.1.The univariate logistic regression identified 15 predictors that were significantly associated with CHD, including education level (odds ratio = 0.55), gravidity (1.95), parity (2.01), history of abnormal reproduction (2.49), family history of CHD (5.23), maternal chronic disease (4.19), maternal upper respiratory tract infection (2.08), environmental pollution around maternal dwelling place (3.63), maternal exposure to occupational hazards (3.53), maternal mental stress (2.48), paternal chronic disease (4.87), paternal exposure to occupational hazards (2.51), intake of vegetable/fruit (0.45), intake of fish/shrimp/meat/egg (0.59), and intake of milk/soymilk (0.55). After many trials, we selected a 3-layer BPNN model with 15, 12, and 1 neuron in the input, hidden, and output layers, respectively, as the best prediction model. The prediction model has accuracies of 0.91 and 0.86 on the training and testing sets, respectively. The sensitivity, specificity, and Yuden Index on the testing set (training set) are 0.78 (0.83), 0.90 (0.95), and 0.68 (0.78), respectively. The areas under the receiver operating curve on the testing and training sets are 0.87 and 0.97, respectively.This study suggests that the BPNN model could be used to predict the risk of CHD in individuals. This model should be further improved by large-sample-size research.
Logits and Tigers and Bears, Oh My! A Brief Look at the Simple Math of Logistic Regression and How It Can Improve Dissemination of Results

ERIC Educational Resources Information Center

Osborne, Jason W.

2012-01-01

Logistic regression is slowly gaining acceptance in the social sciences, and fills an important niche in the researcher's toolkit: being able to predict important outcomes that are not continuous in nature. While OLS regression is a valuable tool, it cannot routinely be used to predict outcomes that are binary or categorical in nature. These…
Large scale landslide susceptibility assessment using the statistical methods of logistic regression and BSA - study case: the sub-basin of the small Niraj (Transylvania Depression, Romania)

NASA Astrophysics Data System (ADS)

Roşca, S.; Bilaşco, Ş.; Petrea, D.; Fodorean, I.; Vescan, I.; Filip, S.; Măguţ, F.-L.

2015-11-01

The existence of a large number of GIS models for the identification of landslide occurrence probability makes difficult the selection of a specific one. The present study focuses on the application of two quantitative models: the logistic and the BSA models. The comparative analysis of the results aims at identifying the most suitable model. The territory corresponding to the Niraj Mic Basin (87 km2) is an area characterised by a wide variety of the landforms with their morphometric, morphographical and geological characteristics as well as by a high complexity of the land use types where active landslides exist. This is the reason why it represents the test area for applying the two models and for the comparison of the results. The large complexity of input variables is illustrated by 16 factors which were represented as 72 dummy variables, analysed on the basis of their importance within the model structures. The testing of the statistical significance corresponding to each variable reduced the number of dummy variables to 12 which were considered significant for the test area within the logistic model, whereas for the BSA model all the variables were employed. The predictability degree of the models was tested through the identification of the area under the ROC curve which indicated a good accuracy (AUROC = 0.86 for the testing area) and predictability of the logistic model (AUROC = 0.63 for the validation area).
Ensemble habitat mapping of invasive plant species

USGS Publications Warehouse

Stohlgren, T.J.; Ma, P.; Kumar, S.; Rocca, M.; Morisette, J.T.; Jarnevich, C.S.; Benson, N.

2010-01-01

Ensemble species distribution models combine the strengths of several species environmental matching models, while minimizing the weakness of any one model. Ensemble models may be particularly useful in risk analysis of recently arrived, harmful invasive species because species may not yet have spread to all suitable habitats, leaving species-environment relationships difficult to determine. We tested five individual models (logistic regression, boosted regression trees, random forest, multivariate adaptive regression splines (MARS), and maximum entropy model or Maxent) and ensemble modeling for selected nonnative plant species in Yellowstone and Grand Teton National Parks, Wyoming; Sequoia and Kings Canyon National Parks, California, and areas of interior Alaska. The models are based on field data provided by the park staffs, combined with topographic, climatic, and vegetation predictors derived from satellite data. For the four invasive plant species tested, ensemble models were the only models that ranked in the top three models for both field validation and test data. Ensemble models may be more robust than individual species-environment matching models for risk analysis. ?? 2010 Society for Risk Analysis.
The correlations of psychological status, quality of life, self-esteem, social support and body image disturbance in Chinese patients with Systemic Lupus Erythematosus.

PubMed

Zhao, Qian; Chen, Haoyang; Yan, Hongyan; He, Yan; Zhu, Li; Fu, WenTing; Shen, Biyu

2018-01-31

This study aimed (i) to complement existing research by focusing on body image disturbance issues in Chinese Systemic Lupus Erythematosus (SLE) patients; (ii) to investigate how Chinese patients make sense of disease diagnosis and perceived cultural influences within the context of their SLE. A total of 118 SLE patients underwent standardized laboratory examinations and completed several questionnaires. Independent sample t-test, Mann-Whitney U-test, Chi-square test, and multivariate analysis using backward stepwise logistic regression model were used to analyze these data. We found 18.3% SLE patients had BID, which were significantly higher than the control group (.8%). SLE patients are more concerned about their physical changes caused by disease. There were significant correlations among personal health insurance, complication of diabetes, appearance of new rash, depression, anxiety, self-esteem and BID in patients with SLE. Meanwhile, logistic regression analysis revealed that appearance of new rash and high anxiety were significantly associated with BID in SLE patients. In conclusion, it is beneficial to pay attention to the physical and mental health of patients with rheumatic disease from the perspective of body image, to understand their needs and to provide effective and effective service for them.
Social Impact of Stigma Regarding Tuberculosis Hindering Adherence to Treatment: A Cross Sectional Study Involving Tuberculosis Patients in Rajshahi City, Bangladesh.

PubMed

Chowdhury, Md Rocky Khan; Rahman, Md Shafiur; Mondal, Md Nazrul Islam; Sayem, Abu; Billah, Baki

2015-01-01

Stigma, considered a social disease, is more apparent in developing societies which are driven by various social affairs, and influences adherence to treatment. The aim of the present study was to examine levels of social stigma related to tuberculosis (TB) in sociodemographic context and identify the effects of sociodemographic factors on stigma. The study sample consisted of 372 TB patients. Data were collected using stratified sampling with simple random sampling techniques. T tests, chi-square tests, and binary logistic regression analysis were performed to examine correlations between stigma and sociodemographic variables. Approximately 85.9% of patients had experienced stigma. The most frequent indicator of the stigma experienced by patients involved problems taking part in social programs (79.5%). Mean levels of stigma were significantly higher in women (55.5%), illiterate individuals (60.8%), and villagers (60.8%) relative to those of other groups. Chi-square tests revealed that education, monthly family income, and type of patient (pulmonary and extrapulmonary) were significantly associated with stigma. Binary logistic regression analysis demonstrated that stigma was influenced by sex, education, and type of patient. Stigma is one of the most important barriers to treatment adherence. Therefore, in interventions that aim to reduce stigma, strong collaboration between various institutions is essential.
Embedded measures of performance validity using verbal fluency tests in a clinical sample.

PubMed

Sugarman, Michael A; Axelrod, Bradley N

2015-01-01

The objective of this study was to determine to what extent verbal fluency measures can be used as performance validity indicators during neuropsychological evaluation. Participants were clinically referred for neuropsychological evaluation in an urban-based Veteran's Affairs hospital. Participants were placed into 2 groups based on their objectively evaluated effort on performance validity tests (PVTs). Individuals who exhibited credible performance (n = 431) failed 0 PVTs, and those with poor effort (n = 192) failed 2 or more PVTs. All participants completed the Controlled Oral Word Association Test (COWAT) and Animals verbal fluency measures. We evaluated how well verbal fluency scores could discriminate between the 2 groups. Raw scores and T scores for Animals discriminated between the credible performance and poor-effort groups with 90% specificity and greater than 40% sensitivity. COWAT scores had lower sensitivity for detecting poor effort. A combination of FAS and Animals scores into logistic regression models yielded acceptable group classification, with 90% specificity and greater than 44% sensitivity. Verbal fluency measures can yield adequate detection of poor effort during neuropsychological evaluation. We provide suggested cut points and logistic regression models for predicting the probability of poor effort in our clinical setting and offer suggested cutoff scores to optimize sensitivity and specificity.
Classification of sodium MRI data of cartilage using machine learning.

PubMed

Madelin, Guillaume; Poidevin, Frederick; Makrymallis, Antonios; Regatte, Ravinder R

2015-11-01

To assess the possible utility of machine learning for classifying subjects with and subjects without osteoarthritis using sodium magnetic resonance imaging data. Theory: Support vector machine, k-nearest neighbors, naïve Bayes, discriminant analysis, linear regression, logistic regression, neural networks, decision tree, and tree bagging were tested. Sodium magnetic resonance imaging with and without fluid suppression by inversion recovery was acquired on the knee cartilage of 19 controls and 28 osteoarthritis patients. Sodium concentrations were measured in regions of interests in the knee for both acquisitions. Mean (MEAN) and standard deviation (STD) of these concentrations were measured in each regions of interest, and the minimum, maximum, and mean of these two measurements were calculated over all regions of interests for each subject. The resulting 12 variables per subject were used as predictors for classification. Either Min [STD] alone, or in combination with Mean [MEAN] or Min [MEAN], all from fluid suppressed data, were the best predictors with an accuracy >74%, mainly with linear logistic regression and linear support vector machine. Other good classifiers include discriminant analysis, linear regression, and naïve Bayes. Machine learning is a promising technique for classifying osteoarthritis patients and controls from sodium magnetic resonance imaging data. © 2014 Wiley Periodicals, Inc.

Examining the Link Between Public Transit Use and Active Commuting

PubMed Central

Bopp, Melissa; Gayah, Vikash V.; Campbell, Matthew E.

2015-01-01

Background: An established relationship exists between public transportation (PT) use and physical activity. However, there is limited literature that examines the link between PT use and active commuting (AC) behavior. This study examines this link to determine if PT users commute more by active modes. Methods: A volunteer, convenience sample of adults (n = 748) completed an online survey about AC/PT patterns, demographic, psychosocial, community and environmental factors. t-test compared differences between PT riders and non-PT riders. Binary logistic regression analyses examined the effect of multiple factors on AC and a full logistic regression model was conducted to examine AC. Results: Non-PT riders (n = 596) reported less AC than PT riders. There were several significant relationships with AC for demographic, interpersonal, worksite, community and environmental factors when considering PT use. The logistic multivariate analysis for included age, number of children and perceived distance to work as negative predictors and PT use, feelings of bad weather and lack of on-street bike lanes as a barrier to AC, perceived behavioral control and spouse AC were positive predictors. Conclusions: This study revealed the complex relationship between AC and PT use. Further research should investigate how AC and public transit use are related. PMID:25898405
Examining the link between public transit use and active commuting.

PubMed

Bopp, Melissa; Gayah, Vikash V; Campbell, Matthew E

2015-04-17

An established relationship exists between public transportation (PT) use and physical activity. However, there is limited literature that examines the link between PT use and active commuting (AC) behavior. This study examines this link to determine if PT users commute more by active modes. A volunteer, convenience sample of adults (n = 748) completed an online survey about AC/PT patterns, demographic, psychosocial, community and environmental factors. t-test compared differences between PT riders and non-PT riders. Binary logistic regression analyses examined the effect of multiple factors on AC and a full logistic regression model was conducted to examine AC. Non-PT riders (n = 596) reported less AC than PT riders. There were several significant relationships with AC for demographic, interpersonal, worksite, community and environmental factors when considering PT use. The logistic multivariate analysis for included age, number of children and perceived distance to work as negative predictors and PT use, feelings of bad weather and lack of on-street bike lanes as a barrier to AC, perceived behavioral control and spouse AC were positive predictors. This study revealed the complex relationship between AC and PT use. Further research should investigate how AC and public transit use are related.
Intermediate and advanced topics in multilevel logistic regression analysis

PubMed Central

Merlo, Juan

2017-01-01

Multilevel data occur frequently in health services, population and public health, and epidemiologic research. In such research, binary outcomes are common. Multilevel logistic regression models allow one to account for the clustering of subjects within clusters of higher‐level units when estimating the effect of subject and cluster characteristics on subject outcomes. A search of the PubMed database demonstrated that the use of multilevel or hierarchical regression models is increasing rapidly. However, our impression is that many analysts simply use multilevel regression models to account for the nuisance of within‐cluster homogeneity that is induced by clustering. In this article, we describe a suite of analyses that can complement the fitting of multilevel logistic regression models. These ancillary analyses permit analysts to estimate the marginal or population‐average effect of covariates measured at the subject and cluster level, in contrast to the within‐cluster or cluster‐specific effects arising from the original multilevel logistic regression model. We describe the interval odds ratio and the proportion of opposed odds ratios, which are summary measures of effect for cluster‐level covariates. We describe the variance partition coefficient and the median odds ratio which are measures of components of variance and heterogeneity in outcomes. These measures allow one to quantify the magnitude of the general contextual effect. We describe an R 2 measure that allows analysts to quantify the proportion of variation explained by different multilevel logistic regression models. We illustrate the application and interpretation of these measures by analyzing mortality in patients hospitalized with a diagnosis of acute myocardial infarction. © 2017 The Authors. Statistics in Medicine published by John Wiley & Sons Ltd. PMID:28543517
Intermediate and advanced topics in multilevel logistic regression analysis.

PubMed

Austin, Peter C; Merlo, Juan

2017-09-10

Multilevel data occur frequently in health services, population and public health, and epidemiologic research. In such research, binary outcomes are common. Multilevel logistic regression models allow one to account for the clustering of subjects within clusters of higher-level units when estimating the effect of subject and cluster characteristics on subject outcomes. A search of the PubMed database demonstrated that the use of multilevel or hierarchical regression models is increasing rapidly. However, our impression is that many analysts simply use multilevel regression models to account for the nuisance of within-cluster homogeneity that is induced by clustering. In this article, we describe a suite of analyses that can complement the fitting of multilevel logistic regression models. These ancillary analyses permit analysts to estimate the marginal or population-average effect of covariates measured at the subject and cluster level, in contrast to the within-cluster or cluster-specific effects arising from the original multilevel logistic regression model. We describe the interval odds ratio and the proportion of opposed odds ratios, which are summary measures of effect for cluster-level covariates. We describe the variance partition coefficient and the median odds ratio which are measures of components of variance and heterogeneity in outcomes. These measures allow one to quantify the magnitude of the general contextual effect. We describe an R 2 measure that allows analysts to quantify the proportion of variation explained by different multilevel logistic regression models. We illustrate the application and interpretation of these measures by analyzing mortality in patients hospitalized with a diagnosis of acute myocardial infarction. © 2017 The Authors. Statistics in Medicine published by John Wiley & Sons Ltd. © 2017 The Authors. Statistics in Medicine published by John Wiley & Sons Ltd.
Role of social support in adolescent suicidal ideation and suicide attempts.

PubMed

Miller, Adam Bryant; Esposito-Smythers, Christianne; Leichtweis, Richard N

2015-03-01

The present study examined the relative contributions of perceptions of social support from parents, close friends, and school on current suicidal ideation (SI) and suicide attempt (SA) history in a clinical sample of adolescents. Participants were 143 adolescents (64% female; 81% white; range, 12-18 years; M = 15.38; standard deviation = 1.43) admitted to a partial hospitalization program. Data were collected with well-validated assessments and a structured clinical interview. Main and interactive effects of perceptions of social support on SI were tested with linear regression. Main and interactive effects of social support on the odds of SA were tested with logistic regression. Results from the linear regression analysis revealed that perceptions of lower school support independently predicted greater severity of SI, accounting for parent and close friend support. Further, the relationship between lower perceived school support and SI was the strongest among those who perceived lower versus higher parental support. Results from the logistic regression analysis revealed that perceptions of lower parental support independently predicted SA history, accounting for school and close friend support. Further, those who perceived lower support from school and close friends reported the greatest odds of an SA history. Results address a significant gap in the social support and suicide literature by demonstrating that perceptions of parent and school support are relatively more important than peer support in understanding suicidal thoughts and history of suicidal behavior. Results suggest that improving social support across these domains may be important in suicide prevention efforts. Copyright © 2015 Society for Adolescent Health and Medicine. Published by Elsevier Inc. All rights reserved.
Factors associated with abnormal eating attitudes among Greek adolescents.

PubMed

Bilali, Aggeliki; Galanis, Petros; Velonakis, Emmanuel; Katostaras, Theofanis

2010-01-01

To estimate the prevalence of abnormal eating attitudes among Greek adolescents and identify possible risk factors associated with these attitudes. Cross-sectional, school-based study. Six randomly selected schools in Patras, southern Greece. The study population consisted of 540 Greek students aged 13-18 years, and the response rate was 97%. The dependent variable was scores on the Eating Attitudes Test-26, with scores > or = 20 indicating abnormal eating attitudes. Bivariate analysis included independent Student t test, chi-square test, and Fisher's exact test. Multivariate logistic regression analysis was applied for the identification of the predictive factors, which were associated independently with abnormal eating attitudes. A 2-sided P value of less than .05 was considered statistically significant. The prevalence of abnormal eating attitudes was 16.7%. Multivariate logistic regression analysis demonstrated that females, urban residents, and those with a body mass index outside normal range, a perception of being overweight, body dissatisfaction, and a family member on a diet were independently related to abnormal eating attitudes. The results indicate that a proportion of Greek adolescents report abnormal eating attitudes and suggest that multiple factors contribute to the development of these attitudes. These findings are useful for further research into this topic and would be valuable in designing preventive interventions. Copyright 2010 Society for Nutrition Education. Published by Elsevier Inc. All rights reserved.
Multivariate Models for Prediction of Human Skin Sensitization ...

EPA Pesticide Factsheets

One of the lnteragency Coordinating Committee on the Validation of Alternative Method's (ICCVAM) top priorities is the development and evaluation of non-animal approaches to identify potential skin sensitizers. The complexity of biological events necessary to produce skin sensitization suggests that no single alternative method will replace the currently accepted animal tests. ICCVAM is evaluating an integrated approach to testing and assessment based on the adverse outcome pathway for skin sensitization that uses machine learning approaches to predict human skin sensitization hazard. We combined data from three in chemico or in vitro assays - the direct peptide reactivity assay (DPRA), human cell line activation test (h-CLAT) and KeratinoSens TM assay - six physicochemical properties and an in silico read-across prediction of skin sensitization hazard into 12 variable groups. The variable groups were evaluated using two machine learning approaches , logistic regression and support vector machine, to predict human skin sensitization hazard. Models were trained on 72 substances and tested on an external set of 24 substances. The six models (three logistic regression and three support vector machine) with the highest accuracy (92%) used: (1) DPRA, h-CLAT and read-across; (2) DPRA, h-CLAT, read-across and KeratinoSens; or (3) DPRA, h-CLAT, read-across, KeratinoSens and log P. The models performed better at predicting human skin sensitization hazard than the murine
Predicting Social Trust with Binary Logistic Regression

ERIC Educational Resources Information Center

Adwere-Boamah, Joseph; Hufstedler, Shirley

2015-01-01

This study used binary logistic regression to predict social trust with five demographic variables from a national sample of adult individuals who participated in The General Social Survey (GSS) in 2012. The five predictor variables were respondents' highest degree earned, race, sex, general happiness and the importance of personally assisting…
Effect of folic acid on appetite in children: ordinal logistic and fuzzy logistic regressions.

PubMed

Namdari, Mahshid; Abadi, Alireza; Taheri, S Mahmoud; Rezaei, Mansour; Kalantari, Naser; Omidvar, Nasrin

2014-03-01

Reduced appetite and low food intake are often a concern in preschool children, since it can lead to malnutrition, a leading cause of impaired growth and mortality in childhood. It is occasionally considered that folic acid has a positive effect on appetite enhancement and consequently growth in children. The aim of this study was to assess the effect of folic acid on the appetite of preschool children 3 to 6 y old. The study sample included 127 children ages 3 to 6 who were randomly selected from 20 preschools in the city of Tehran in 2011. Since appetite was measured by linguistic terms, a fuzzy logistic regression was applied for modeling. The obtained results were compared with a statistical ordinal logistic model. After controlling for the potential confounders, in a statistical ordinal logistic model, serum folate showed a significantly positive effect on appetite. A small but positive effect of folate was detected by fuzzy logistic regression. Based on fuzzy regression, the risk for poor appetite in preschool children was related to the employment status of their mothers. In this study, a positive association was detected between the levels of serum folate and improved appetite. For further investigation, a randomized controlled, double-blind clinical trial could be helpful to address causality. Copyright © 2014 Elsevier Inc. All rights reserved.
Differential Item Functioning (DIF) among Spanish-Speaking English Language Learners (ELLs) in State Science Tests

NASA Astrophysics Data System (ADS)

Ilich, Maria O.

Psychometricians and test developers evaluate standardized tests for potential bias against groups of test-takers by using differential item functioning (DIF). English language learners (ELLs) are a diverse group of students whose native language is not English. While they are still learning the English language, they must take their standardized tests for their school subjects, including science, in English. In this study, linguistic complexity was examined as a possible source of DIF that may result in test scores that confound science knowledge with a lack of English proficiency among ELLs. Two years of fifth-grade state science tests were analyzed for evidence of DIF using two DIF methods, Simultaneous Item Bias Test (SIBTest) and logistic regression. The tests presented a unique challenge in that the test items were grouped together into testlets---groups of items referring to a scientific scenario to measure knowledge of different science content or skills. Very large samples of 10, 256 students in 2006 and 13,571 students in 2007 were examined. Half of each sample was composed of Spanish-speaking ELLs; the balance was comprised of native English speakers. The two DIF methods were in agreement about the items that favored non-ELLs and the items that favored ELLs. Logistic regression effect sizes were all negligible, while SIBTest flagged items with low to high DIF. A decrease in socioeconomic status and Spanish-speaking ELL diversity may have led to inconsistent SIBTest effect sizes for items used in both testing years. The DIF results for the testlets suggested that ELLs lacked sufficient opportunity to learn science content. The DIF results further suggest that those constructed response test items requiring the student to draw a conclusion about a scientific investigation or to plan a new investigation tended to favor ELLs.
Extension of the Peters–Belson method to estimate health disparities among multiple groups using logistic regression with survey data

PubMed Central

Li, Y.; Graubard, B. I.; Huang, P.; Gastwirth, J. L.

2015-01-01

Determining the extent of a disparity, if any, between groups of people, for example, race or gender, is of interest in many fields, including public health for medical treatment and prevention of disease. An observed difference in the mean outcome between an advantaged group (AG) and disadvantaged group (DG) can be due to differences in the distribution of relevant covariates. The Peters–Belson (PB) method fits a regression model with covariates to the AG to predict, for each DG member, their outcome measure as if they had been from the AG. The difference between the mean predicted and the mean observed outcomes of DG members is the (unexplained) disparity of interest. We focus on applying the PB method to estimate the disparity based on binary/multinomial/proportional odds logistic regression models using data collected from complex surveys with more than one DG. Estimators of the unexplained disparity, an analytic variance–covariance estimator that is based on the Taylor linearization variance–covariance estimation method, as well as a Wald test for testing a joint null hypothesis of zero for unexplained disparities between two or more minority groups and a majority group, are provided. Simulation studies with data selected from simple random sampling and cluster sampling, as well as the analyses of disparity in body mass index in the National Health and Nutrition Examination Survey 1999–2004, are conducted. Empirical results indicate that the Taylor linearization variance–covariance estimation is accurate and that the proposed Wald test maintains the nominal level. PMID:25382235
Determinants and prevalence of late HIV testing in Tijuana, Mexico.

PubMed

Carrizosa, Claudia M; Blumberg, Elaine J; Hovell, Melbourne F; Martinez-Donate, Ana P; Garcia-Gonzalez, Gregorio; Lozada, Remedios; Kelley, Norma J; Hofstetter, C Richard; Sipan, Carol L

2010-05-01

Timely diagnosis of HIV is essential to improve survival rates and reduce transmission of the virus. Insufficient progress has been made in effecting earlier HIV diagnoses. The Mexican border city of Tijuana has one of the highest AIDS incidence and mortality rates in all of Mexico. This study examined the prevalence and potential correlates of late HIV testing in Tijuana, Mexico. Late testers were defined as participants who had at least one of: (1) an AIDS-defining illness within 1 year of first positive HIV test; (2) a date of AIDS diagnosis within 1 year of first positive HIV test; or (3) an initial CD4 cell count below 200 cells per microliter within 1 year of first positive HIV test. Medical charts of 670 HIV-positive patients from two HIV/AIDS public clinics in Tijuana were reviewed and abstracted; 362 of these patients were interviewed using a cross-sectional survey. Using multivariate logistic regression, we explored potential correlates of late HIV testing based on the Behavioral Ecological Model. From 342 participants for whom late testing could be determined, the prevalence of late testing was 43.2%. Multivariate logistic regression results (n = 275) revealed five significant correlates of late testing: "I preferred not to know I had HIV" (adjusted odds ratio [AOR] = 2.78, 1.46-5.31); clinic (AOR = 1.90, 1.06-3.41); exposure to peers engaging in high-risk sexual behavior (AOR = 1.14, 1.02-1.27); stigma regarding HIV-infected individuals (AOR = 0.65, 0.47-0.92); and stigma regarding HIV testing (AOR = 0.66, 0.45-0.97). These findings may inform the design of interventions to increase timely HIV testing and help reduce HIV transmission in the community at large.
Molecular markers of neuropsychological functioning and Alzheimer's disease.

PubMed

Edwards, Melissa; Balldin, Valerie Hobson; Hall, James; O'Bryant, Sid

2015-03-01

The current project sought to examine molecular markers of neuropsychological functioning among elders with and without Alzheimer's disease (AD) and determine the predictive ability of combined molecular markers and select neuropsychological tests in detecting disease presence. Data were analyzed from 300 participants (n = 150, AD and n = 150, controls) enrolled in the Texas Alzheimer's Research and Care Consortium. Linear regression models were created to examine the link between the top five molecular markers from our AD blood profile and neuropsychological test scores. Logistical regressions were used to predict AD presence using serum biomarkers in combination with select neuropsychological measures. Using the neuropsychological test with the least amount of variance overlap with the molecular markers, the combined neuropsychological test and molecular markers was highly accurate in detecting AD presence. This work provides the foundation for the generation of a point-of-care device that can be used to screen for AD.
Clustering performance comparison using K-means and expectation maximization algorithms.

PubMed

Jung, Yong Gyu; Kang, Min Soo; Heo, Jun

2014-11-14

Clustering is an important means of data mining based on separating data categories by similar features. Unlike the classification algorithm, clustering belongs to the unsupervised type of algorithms. Two representatives of the clustering algorithms are the K -means and the expectation maximization (EM) algorithm. Linear regression analysis was extended to the category-type dependent variable, while logistic regression was achieved using a linear combination of independent variables. To predict the possibility of occurrence of an event, a statistical approach is used. However, the classification of all data by means of logistic regression analysis cannot guarantee the accuracy of the results. In this paper, the logistic regression analysis is applied to EM clusters and the K -means clustering method for quality assessment of red wine, and a method is proposed for ensuring the accuracy of the classification results.
Racial/ethnic and educational differences in the estimated odds of recent nitrite use among adult household residents in the United States: an illustration of matching and conditional logistic regression.

PubMed

Delva, J; Spencer, M S; Lin, J K

2000-01-01

This article compares estimates of the relative odds of nitrite use obtained from weighted unconditional logistic regression with estimates obtained from conditional logistic regression after post-stratification and matching of cases with controls by neighborhood of residence. We illustrate these methods by comparing the odds associated with nitrite use among adults of four racial/ethnic groups, with and without a high school education. We used aggregated data from the 1994-B through 1996 National Household Survey on Drug Abuse (NHSDA). Difference between the methods and implications for analysis and inference are discussed.
Investigating Community Factors as Predictors of Rural 11th-Grade Agricultural Science Students' Choice of Careers in Agriculture

ERIC Educational Resources Information Center

Adedokun, Omolola A.; Balschweid, Mark A.

2008-01-01

This study investigates the links between community contexts/factors and rural 11th-grade agricultural science students' choice of careers in agriculture. A logistic regression model was developed and tested to examine the extent to which nine measures of community contexts (i.e., membership in FFA, membership in 4-H, community attachment,…
Wildland recreation in the rural South: an examination of marginality and ethnicity theory

Treesearch

Cassandra Y. Johnson; J. Michael Bowker; Donald B.K. English; Dreamal Worthen

1998-01-01

The ethnicity and marginality explanations of minority recreation participation provide the conceptual basis for the authorsâ inquiry. These theories are examined for a sample of rural African-Americans and whites. Using logistic regression, the researchers test for black and while differences in: 1) visitation to wildland areas in general; 2) visitation to national...
The Associations between Health Literacy, Reasons for Seeking Health Information, and Information Sources Utilized by Taiwanese Adults

ERIC Educational Resources Information Center

Wei, Mi-Hsiu

2014-01-01

Objective: To determine the associations between health literacy, the reasons for seeking health information, and the information sources utilized by Taiwanese adults. Method: A cross-sectional survey of 752 adults residing in rural and urban areas of Taiwan was conducted via questionnaires. Chi-squared tests and logistic regression were used for…
Landscape evaluation of female black bear habitat effectiveness and capability in the North Cascades, Washington.

Treesearch

William L. Gaines; Andrea L. Lyons; John F. Lehmkuhl; Kenneth J. Raedeke

2005-01-01

We used logistic regression to derive scaled resource selection functions (RSFs) for female black bears at two study areas in the North Cascades Mountains. We tested the hypothesis that the influence of roads would result in potential habitat effectiveness (RSFs without the influence of roads) being greater than realized habitat effectiveness (RSFs with roads). Roads...
Application of Bayesian methods to habitat selection modeling of the northern spotted owl in California: new statistical methods for wildlife research

Treesearch

Howard B. Stauffer; Cynthia J. Zabel; Jeffrey R. Dunk

2005-01-01

We compared a set of competing logistic regression habitat selection models for Northern Spotted Owls (Strix occidentalis caurina) in California. The habitat selection models were estimated, compared, evaluated, and tested using multiple sample datasets collected on federal forestlands in northern California. We used Bayesian methods in interpreting...

Combining biological and psychosocial baseline variables did not improve prediction of outcome of a very-low-energy diet in a clinic referral population.

PubMed

Sumithran, P; Purcell, K; Kuyruk, S; Proietto, J; Prendergast, L A

2018-02-01

Consistent, strong predictors of obesity treatment outcomes have not been identified. It has been suggested that broadening the range of predictor variables examined may be valuable. We explored methods to predict outcomes of a very-low-energy diet (VLED)-based programme in a clinically comparable setting, using a wide array of pre-intervention biological and psychosocial participant data. A total of 61 women and 39 men (mean ± standard deviation [SD] body mass index: 39.8 ± 7.3 kg/m 2 ) underwent an 8-week VLED and 12-month follow-up. At baseline, participants underwent a blood test and assessment of psychological, social and behavioural factors previously associated with treatment outcomes. Logistic regression, linear discriminant analysis, decision trees and random forests were used to model outcomes from baseline variables. Of the 100 participants, 88 completed the VLED and 42 attended the Week 60 visit. Overall prediction rates for weight loss of ≥10% at weeks 8 and 60, and attrition at Week 60, using combined data were between 77.8 and 87.6% for logistic regression, and lower for other methods. When logistic regression analyses included only baseline demographic and anthropometric variables, prediction rates were 76.2-86.1%. In this population, considering a wide range of biological and psychosocial data did not improve outcome prediction compared to simply-obtained baseline characteristics. © 2017 World Obesity Federation.
Host Polymorphisms in TLR9 and IL10 Are Associated With the Outcomes of Experimental Haemophilus ducreyi Infection in Human Volunteers.

PubMed

Singer, Martin; Li, Wei; Morré, Servaas A; Ouburg, Sander; Spinola, Stanley M

2016-08-01

In humans inoculated with Haemophilus ducreyi, there are host effects on the possible clinical outcomes-pustule formation versus spontaneous resolution of infection. However, the immunogenetic factors that influence these outcomes are unknown. Here we examined the role of 14 single-nucleotide polymorphisms (SNPs) in 7 selected pathogen-recognition pathways and cytokine genes on the gradated outcomes of experimental infection. DNAs from 105 volunteers infected with H. ducreyi at 3 sites were genotyped for SNPs, using real-time polymerase chain reaction. The participants were classified into 2 cohorts, by race, and into 4 groups, based on whether they formed 0, 1, 2, or 3 pustules. χ(2) tests for trend and logistic regression analyses were performed on the data. In European Americans, the most significant findings were a protective association of the TLR9 +2848 GG genotype and a risk-enhancing association of the TLR9 TA haplotype with pustule formation; logistic regression showed a trend toward protection for the TLR9 +2848 GG genotype. In African Americans, logistic regression showed a protective effect for the IL10 -2849 AA genotype and a risk-enhancing effect for the IL10 AAC haplotype. Variations in TLR9 and IL10 are associated with the outcome of H. ducreyi infection. © The Author 2016. Published by Oxford University Press for the Infectious Diseases Society of America. All rights reserved. For permissions, e-mail journals.permissions@oup.com.
Logistic regression analysis of psychosocial correlates associated with recovery from schizophrenia in a Chinese community.

PubMed

Tse, Samson; Davidson, Larry; Chung, Ka-Fai; Yu, Chong Ho; Ng, King Lam; Tsoi, Emily

2015-02-01

More mental health services are adopting the recovery paradigm. This study adds to prior research by (a) using measures of stages of recovery and elements of recovery that were designed and validated in a non-Western, Chinese culture and (b) testing which demographic factors predict advanced recovery and whether placing importance on certain elements predicts advanced recovery. We examined recovery and factors associated with recovery among 75 Hong Kong adults who were diagnosed with schizophrenia and assessed to be in clinical remission. Data were collected on socio-demographic factors, recovery stages and elements associated with recovery. Logistic regression analysis was used to identify variables that could best predict stages of recovery. Receiver operating characteristic curves were used to detect the classification accuracy of the model (i.e. rates of correct classification of stages of recovery). Logistic regression results indicated that stages of recovery could be distinguished with reasonable accuracy for Stage 3 ('living with disability', classification accuracy = 75.45%) and Stage 4 ('living beyond disability', classification accuracy = 75.50%). However, there was no sufficient information to predict Combined Stages 1 and 2 ('overwhelmed by disability' and 'struggling with disability'). It was found that having a meaningful role and age were the most important differentiators of recovery stage. Preliminary findings suggest that adopting salient life roles personally is important to recovery and that this component should be incorporated into mental health services. © The Author(s) 2014.
Smoking media literacy in Vietnamese adolescents.

PubMed

Page, Randy M; Huong, Nguyen T; Chi, Hoang K; Tien, Truong Q

2011-01-01

Smoking media literacy (SML) has been found to be independently associated with reduced current smoking and reduced susceptibility to future smoking in a sample of American adolescents, but not in other populations of adolescents. Thus, the purpose of this study was to assess SML in Vietnamese adolescents and to determine the association with smoking behavior and susceptibility to future smoking. A cross-sectional survey of 2000 high school students completed the SML scale, which is based on an integrated theoretical framework of media literacy, and items assessing cigarette use. Ordinal logistic regression was used to determine the association of SML with smoking and susceptibility to future smoking. Ordinal logistic regression was also to determine whether smoking in the past 30 days was associated with the 8 domains/core concepts of media literacy which comprise the SML. Smoking media literacy was lower among the Vietnamese adolescents than what has been previously reported in American adolescents. Ordinal logistic regression analysis results showed that in the total sample SML was associated with reduced smoking, but there was no association with susceptibility to future smoking. Further analysis showed that results differed according to school and grade level. There did not appear to be association of smoking with the specific domains/concepts that comprise the SML. The association of SML with reduced smoking suggests the need for further research involving SML, including the testing of media literacy training interventions, in Vietnamese adolescents and also other populations of adolescents. © 2011, American School Health Association.
Analysis of occlusal variables, dental attrition, and age for distinguishing healthy controls from female patients with intracapsular temporomandibular disorders.

PubMed

Seligman, D A; Pullinger, A G

2000-01-01

Confusion about the relationship of occlusion to temporomandibular disorders (TMD) persists. This study attempted to identify occlusal and attrition factors plus age that would characterize asymptomatic normal female subjects. A total of 124 female patients with intracapsular TMD were compared with 47 asymptomatic female controls for associations to 9 occlusal factors, 3 attrition severity measures, and age using classification tree, multiple stepwise logistic regression, and univariate analyses. Models were tested for accuracy (sensitivity and specificity) and total contribution to the variance. The classification tree model had 4 terminal nodes that used only anterior attrition and age. "Normals" were mainly characterized by low attrition levels, whereas patients had higher attrition and tended to be younger. The tree model was only moderately useful (sensitivity 63%, specificity 94%) in predicting normals. The logistic regression model incorporated unilateral posterior crossbite and mediotrusive attrition severity in addition to the 2 factors in the tree, but was slightly less accurate than the tree (sensitivity 51%, specificity 90%). When only occlusal factors were considered in the analysis, normals were additionally characterized by a lack of anterior open bite, smaller overjet, and smaller RCP-ICP slides. The log likelihood accounted for was similar for both the tree (pseudo R(2) = 29.38%; mean deviance = 0.95) and the multiple logistic regression (Cox Snell R(2) = 30.3%, mean deviance = 0.84) models. The occlusal and attrition factors studied were only moderately useful in differentiating normals from TMD patients.
Novel solutions for an old disease: diagnosis of acute appendicitis with random forest, support vector machines, and artificial neural networks.

PubMed

Hsieh, Chung-Ho; Lu, Ruey-Hwa; Lee, Nai-Hsin; Chiu, Wen-Ta; Hsu, Min-Huei; Li, Yu-Chuan Jack

2011-01-01

Diagnosing acute appendicitis clinically is still difficult. We developed random forests, support vector machines, and artificial neural network models to diagnose acute appendicitis. Between January 2006 and December 2008, patients who had a consultation session with surgeons for suspected acute appendicitis were enrolled. Seventy-five percent of the data set was used to construct models including random forest, support vector machines, artificial neural networks, and logistic regression. Twenty-five percent of the data set was withheld to evaluate model performance. The area under the receiver operating characteristic curve (AUC) was used to evaluate performance, which was compared with that of the Alvarado score. Data from a total of 180 patients were collected, 135 used for training and 45 for testing. The mean age of patients was 39.4 years (range, 16-85). Final diagnosis revealed 115 patients with and 65 without appendicitis. The AUC of random forest, support vector machines, artificial neural networks, logistic regression, and Alvarado was 0.98, 0.96, 0.91, 0.87, and 0.77, respectively. The sensitivity, specificity, positive, and negative predictive values of random forest were 94%, 100%, 100%, and 87%, respectively. Random forest performed better than artificial neural networks, logistic regression, and Alvarado. We demonstrated that random forest can predict acute appendicitis with good accuracy and, deployed appropriately, can be an effective tool in clinical decision making. Copyright © 2011 Mosby, Inc. All rights reserved.
Does substance misuse moderate the relationship between criminal thinking and recidivism?

PubMed Central

Caudy, Michael S.; Folk, Johanna B.; Stuewig, Jeffrey B.; Wooditch, Alese; Martinez, Andres; Maass, Stephanie; Tangney, June P.; Taxman, Faye S.

2014-01-01

Purpose Some differential intervention frameworks contend that substance use is less robustly related to recidivism outcomes than other criminogenic needs such as criminal thinking. The current study tested the hypothesis that substance use disorder severity moderates the relationship between criminal thinking and recidivism. Methods The study utilized two independent criminal justice samples. Study 1 included 226 drug-involved probationers. Study 2 included 337 jail inmates with varying levels of substance use disorder severity. Logistic regression was employed to test the main and interactive effects of criminal thinking and substance use on multiple dichotomous indicators of recidivism. Results Bivariate analyses revealed a significant correlation between criminal thinking and recidivism in the jail sample (r = .18, p < .05) but no significant relationship in the probation sample. Logistic regressions revealed that SUD symptoms moderated the relationship between criminal thinking and recidivism in the jail-based sample (B = −.58, p < .05). A significant moderation effect was not observed in the probation sample. Conclusions Study findings indicate that substance use disorder symptoms moderate the strength of the association between criminal thinking and recidivism. These findings demonstrate the need for further research into the interaction between various dynamic risk factors. PMID:25598559
Multidrug-resistant pulmonary tuberculosis in Los Altos, Selva and Norte regions, Chiapas, Mexico.

PubMed

Sánchez-Pérez, H J; Díaz-Vázquez, A; Nájera-Ortiz, J C; Balandrano, S; Martín-Mateo, M

2010-01-01

To analyse the proportion of multidrug-resistant tuberculosis (MDR-TB) in cultures performed during the period 2000-2002 in Los Altos, Selva and Norte regions, Chiapas, Mexico, and to analyse MDR-TB in terms of clinical and sociodemographic indicators. Cross-sectional study of patients with pulmonary tuberculosis (PTB) from the above regions. Drug susceptibility testing results from two research projects were analysed, as were those of routine sputum samples sent in by health personnel for processing (n = 114). MDR-TB was analysed in terms of the various variables of interest using bivariate tests of association and logistic regression. The proportion of primary MDR-TB was 4.6% (2 of 43), that of secondary MDR-TB was 29.2% (7/24), while among those whose history of treatment was unknown the proportion was 14.3% (3/21). According to the logistic regression model, the variables most highly associated with MDR-TB were as follows: having received anti-tuberculosis treatment previously, cough of >3 years' duration and not being indigenous. The high proportion of MDR cases found in the regions studied shows that it is necessary to significantly improve the control and surveillance of PTB.
Sex-related perceptions associated with sexual activity status among Japanese adolescents who heavily use text messaging.

PubMed

Kawamura, Yoko

2012-01-01

This study examines the relationship between sex-related perceptions and engagement in sexual intercourse among adolescents in Japan who were heavy users of text massaging. Using the data from the 6th National Survey on Youth Sexual Behavior of 548 high school students who heavily use text messaging, multinomial logistic regression analyses on variables constructing sexual norms and gender-role attitudes were conducted to assess the relationship with sexual activity status as the first step. A backward stepwise elimination method of multinomial logistic regression was used as the second step at which variables for each set of two factors were tested, and as the third step at which variables of two factors were simultaneously tested. The study results showed that perceptions were related to engagement in sexual intercourse among adolescents who heavily used text messaging. In particular, those who perceived that sex is an act to be engaged in at an earlier stage of a relationship and that men have a stronger sex drive tended to be sexually active or have experienced sexual intercourse. These findings could be utilized to design more effective sexual health education messages for Japanese adolescents who are at an elevated risk.
Prediction of Return-to-original-work after an Industrial Accident Using Machine Learning and Comparison of Techniques

PubMed Central

2018-01-01

Background Many studies have tried to develop predictors for return-to-work (RTW). However, since complex factors have been demonstrated to predict RTW, it is difficult to use them practically. This study investigated whether factors used in previous studies could predict whether an individual had returned to his/her original work by four years after termination of the worker's recovery period. Methods An initial logistic regression analysis of 1,567 participants of the fourth Panel Study of Worker's Compensation Insurance yielded odds ratios. The participants were divided into two subsets, a training dataset and a test dataset. Using the training dataset, logistic regression, decision tree, random forest, and support vector machine models were established, and important variables of each model were identified. The predictive abilities of the different models were compared. Results The analysis showed that only earned income and company-related factors significantly affected return-to-original-work (RTOW). The random forest model showed the best accuracy among the tested machine learning models; however, the difference was not prominent. Conclusion It is possible to predict a worker's probability of RTOW using machine learning techniques with moderate accuracy. PMID:29736160
Camelus dromedarius brucellosis and its public health associated risks in the Afar National Regional State in northeastern Ethiopia

PubMed Central

2013-01-01

Background A cross-sectional study was carried out in four districts of the Afar region in Ethiopia to determine the prevalence of brucellosis in camels, and to identify risky practices that would facilitate the transmission of zoonoses to humans. This study involved testing 461 camels and interviewing 120 livestock owners. The modified Rose Bengal plate test (mRBPT) and complement fixation test (CFT) were used as screening and confirmatory tests, respectively. SPSS 16 was used to analyze the overall prevalence and potential risk factors for seropositivity, using a multivariable logistic regression analysis. Results In the camel herds tested, 5.4% had antibodies against Brucella species, and the district level seroprevalence ranged from 11.7% to 15.5% in camels. The logistic regression model for camels in a herd size > 20 animals (OR = 2.8; 95% CI: 1.16-6.62) and greater than four years of age (OR = 4.9; 95% CI: 1.45-16.82) showed a higher risk of infection when compared to small herds and those ≤ 4 years old. The questionnaire survey revealed that most respondents did not know about the transmission of zoonotic diseases, and that their practices could potentially facilitate the transmission of zoonotic pathogens. Conclusions The results of this study revealed that camel brucellosis is prevalent in the study areas. Therefore, there is a need for implementing control measures and increasing public awareness in the prevention methods of brucellosis. PMID:24344729
Assessing contaminant sensitivity of endangered and threatened aquatic species: part II. Chronic toxicity of copper and pentachlorophenol to two endangered species and two surrogate species.

PubMed

Besser, J M; Wang, N; Dwyer, F J; Mayer, F L; Ingersoll, C G

2005-02-01

Early life-stage toxicity tests with copper and pentachlorophenol (PCP) were conducted with two species listed under the United States Endangered Species Act (the endangered fountain darter, Etheostoma fonticola, and the threatened spotfin chub, Cyprinella monacha) and two commonly tested species (fathead minnow, Pimephales promelas, and rainbow trout, Oncorhynchus mykiss). Results were compared using lowest-observed effect concentrations (LOECs) based on statistical hypothesis tests and by point estimates derived by linear interpolation and logistic regression. Sublethal end points, growth (mean individual dry weight) and biomass (total dry weight per replicate) were usually more sensitive than survival. The biomass end point was equally sensitive as growth and had less among-test variation. Effect concentrations based on linear interpolation were less variable than LOECs, which corresponded to effects ranging from 9% to 76% relative to controls and were consistent with thresholds based on logistic regression. Fountain darter was the most sensitive species for both chemicals tested, with effect concentrations for biomass at < or = 11 microg/L (LOEC and 25% inhibition concentration [IC25]) for copper and at 21 microg/L (IC25) for PCP, but spotfin chub was no more sensitive than the commonly tested species. Effect concentrations for fountain darter were lower than current chronic water quality criteria for both copper and PCP. Protectiveness of chronic water-quality criteria for threatened and endangered species could be improved by the use of safety factors or by conducting additional chronic toxicity tests with species and chemicals of concern.
Assessing contaminant sensitivity of endangered and threatened aquatic species: Part II. chronic toxicity of copper and pentachlorophenol to two endangered species and two surrogate species

USGS Publications Warehouse

Besser, J.M.; Wang, N.; Dwyer, F.J.; Mayer, F.L.; Ingersoll, C.G.

2005-01-01

Early life-stage toxicity tests with copper and pentachlorophenol (PCP) were conducted with two species listed under the United States Endangered Species Act (the endangered fountain darter, Etheostoma fonticola, and the threatened spotfin chub, Cyprinella monacha) and two commonly tested species (fathead minnow, Pimephales promelas, and rainbow trout, Oncorhynchus mykiss). Results were compared using lowest-observed effect concentrations (LOECs) based on statistical hypothesis tests and by point estimates derived by linear interpolation and logistic regression. Sublethal end points, growth (mean individual dry weight) and biomass (total dry weight per replicate) were usually more sensitive than survival. The biomass end point was equally sensitive as growth and had less among-test variation. Effect concentrations based on linear interpolation were less variable than LOECs, which corresponded to effects ranging from 9% to 76% relative to controls and were consistent with thresholds based on logistic regression. Fountain darter was the most sensitive species for both chemicals tested, with effect concentrations for biomass at ??? 11 ??g/L (LOEC and 25% inhibition concentration [IC25]) for copper and at 21 ??g/L (IC25) for PCP, but spotfin chub was no more sensitive than the commonly tested species. Effect concentrations for fountain darter were lower than current chronic water quality criteria for both copper and PCP. Protectiveness of chronic water-quality criteria for threatened and endangered species could be improved by the use of safety factors or by conducting additional chronic toxicity tests with species and chemicals of concern. ?? 2005 Springer Science+Business Media, Inc.
Regression trees for predicting mortality in patients with cardiovascular disease: What improvement is achieved by using ensemble-based methods?

PubMed Central

Austin, Peter C; Lee, Douglas S; Steyerberg, Ewout W; Tu, Jack V

2012-01-01

In biomedical research, the logistic regression model is the most commonly used method for predicting the probability of a binary outcome. While many clinical researchers have expressed an enthusiasm for regression trees, this method may have limited accuracy for predicting health outcomes. We aimed to evaluate the improvement that is achieved by using ensemble-based methods, including bootstrap aggregation (bagging) of regression trees, random forests, and boosted regression trees. We analyzed 30-day mortality in two large cohorts of patients hospitalized with either acute myocardial infarction (N = 16,230) or congestive heart failure (N = 15,848) in two distinct eras (1999–2001 and 2004–2005). We found that both the in-sample and out-of-sample prediction of ensemble methods offered substantial improvement in predicting cardiovascular mortality compared to conventional regression trees. However, conventional logistic regression models that incorporated restricted cubic smoothing splines had even better performance. We conclude that ensemble methods from the data mining and machine learning literature increase the predictive performance of regression trees, but may not lead to clear advantages over conventional logistic regression models for predicting short-term mortality in population-based samples of subjects with cardiovascular disease. PMID:22777999
Iterative Purification and Effect Size Use with Logistic Regression for Differential Item Functioning Detection

ERIC Educational Resources Information Center

French, Brian F.; Maller, Susan J.

2007-01-01

Two unresolved implementation issues with logistic regression (LR) for differential item functioning (DIF) detection include ability purification and effect size use. Purification is suggested to control inaccuracies in DIF detection as a result of DIF items in the ability estimate. Additionally, effect size use may be beneficial in controlling…
"Let Me Count the Ways:" Fostering Reasons for Living among Low-Income, Suicidal, African American Women

ERIC Educational Resources Information Center

West, Lindsey M.; Davis, Telsie A.; Thompson, Martie P.; Kaslow, Nadine J.

2011-01-01

Protective factors for fostering reasons for living were examined among low-income, suicidal, African American women. Bivariate logistic regressions revealed that higher levels of optimism, spiritual well-being, and family social support predicted reasons for living. Multivariate logistic regressions indicated that spiritual well-being showed…
Comparison of Two Approaches for Handling Missing Covariates in Logistic Regression

ERIC Educational Resources Information Center

Peng, Chao-Ying Joanne; Zhu, Jin

2008-01-01

For the past 25 years, methodological advances have been made in missing data treatment. Most published work has focused on missing data in dependent variables under various conditions. The present study seeks to fill the void by comparing two approaches for handling missing data in categorical covariates in logistic regression: the…
Multiple Logistic Regression Analysis of Cigarette Use among High School Students

ERIC Educational Resources Information Center

Adwere-Boamah, Joseph

2011-01-01

A binary logistic regression analysis was performed to predict high school students' cigarette smoking behavior from selected predictors from 2009 CDC Youth Risk Behavior Surveillance Survey. The specific target student behavior of interest was frequent cigarette use. Five predictor variables included in the model were: a) race, b) frequency of…
Propensity Score Estimation with Data Mining Techniques: Alternatives to Logistic Regression

ERIC Educational Resources Information Center

Keller, Bryan S. B.; Kim, Jee-Seon; Steiner, Peter M.

2013-01-01

Propensity score analysis (PSA) is a methodological technique which may correct for selection bias in a quasi-experiment by modeling the selection process using observed covariates. Because logistic regression is well understood by researchers in a variety of fields and easy to implement in a number of popular software packages, it has…
Two-factor logistic regression in pediatric liver transplantation

NASA Astrophysics Data System (ADS)

Uzunova, Yordanka; Prodanova, Krasimira; Spasov, Lyubomir

2017-12-01

Using a two-factor logistic regression analysis an estimate is derived for the probability of absence of infections in the early postoperative period after pediatric liver transplantation. The influence of both the bilirubin level and the international normalized ratio of prothrombin time of blood coagulation at the 5th postoperative day is studied.

Predictors of Placement Stability at the State Level: The Use of Logistic Regression to Inform Practice

ERIC Educational Resources Information Center

Courtney, Jon R.; Prophet, Retta

2011-01-01

Placement instability is often associated with a number of negative outcomes for children. To gain state level contextual knowledge of factors associated with placement stability/instability, logistic regression was applied to selected variables from the New Mexico Adoption and Foster Care Administrative Reporting System dataset. Predictors…
Classifying machinery condition using oil samples and binary logistic regression

NASA Astrophysics Data System (ADS)

Phillips, J.; Cripps, E.; Lau, John W.; Hodkiewicz, M. R.

2015-08-01

The era of big data has resulted in an explosion of condition monitoring information. The result is an increasing motivation to automate the costly and time consuming human elements involved in the classification of machine health. When working with industry it is important to build an understanding and hence some trust in the classification scheme for those who use the analysis to initiate maintenance tasks. Typically "black box" approaches such as artificial neural networks (ANN) and support vector machines (SVM) can be difficult to provide ease of interpretability. In contrast, this paper argues that logistic regression offers easy interpretability to industry experts, providing insight to the drivers of the human classification process and to the ramifications of potential misclassification. Of course, accuracy is of foremost importance in any automated classification scheme, so we also provide a comparative study based on predictive performance of logistic regression, ANN and SVM. A real world oil analysis data set from engines on mining trucks is presented and using cross-validation we demonstrate that logistic regression out-performs the ANN and SVM approaches in terms of prediction for healthy/not healthy engines.
Matched samples logistic regression in case-control studies with missing values: when to break the matches.

PubMed

Hansson, Lisbeth; Khamis, Harry J

2008-12-01

Simulated data sets are used to evaluate conditional and unconditional maximum likelihood estimation in an individual case-control design with continuous covariates when there are different rates of excluded cases and different levels of other design parameters. The effectiveness of the estimation procedures is measured by method bias, variance of the estimators, root mean square error (RMSE) for logistic regression and the percentage of explained variation. Conditional estimation leads to higher RMSE than unconditional estimation in the presence of missing observations, especially for 1:1 matching. The RMSE is higher for the smaller stratum size, especially for the 1:1 matching. The percentage of explained variation appears to be insensitive to missing data, but is generally higher for the conditional estimation than for the unconditional estimation. It is particularly good for the 1:2 matching design. For minimizing RMSE, a high matching ratio is recommended; in this case, conditional and unconditional logistic regression models yield comparable levels of effectiveness. For maximizing the percentage of explained variation, the 1:2 matching design with the conditional logistic regression model is recommended.
The Effect of Latent Binary Variables on the Uncertainty of the Prediction of a Dichotomous Outcome Using Logistic Regression Based Propensity Score Matching.

PubMed

Szekér, Szabolcs; Vathy-Fogarassy, Ágnes

2018-01-01

Logistic regression based propensity score matching is a widely used method in case-control studies to select the individuals of the control group. This method creates a suitable control group if all factors affecting the output variable are known. However, if relevant latent variables exist as well, which are not taken into account during the calculations, the quality of the control group is uncertain. In this paper, we present a statistics-based research in which we try to determine the relationship between the accuracy of the logistic regression model and the uncertainty of the dependent variable of the control group defined by propensity score matching. Our analyses show that there is a linear correlation between the fit of the logistic regression model and the uncertainty of the output variable. In certain cases, a latent binary explanatory variable can result in a relative error of up to 70% in the prediction of the outcome variable. The observed phenomenon calls the attention of analysts to an important point, which must be taken into account when deducting conclusions.
History of falls, gait, balance, and fall risks in older cancer survivors living in the community.

PubMed

Huang, Min H; Shilling, Tracy; Miller, Kara A; Smith, Kristin; LaVictoire, Kayle

2015-01-01

Older cancer survivors may be predisposed to falls because cancer-related sequelae affect virtually all body systems. The use of a history of falls, gait speed, and balance tests to assess fall risks remains to be investigated in this population. This study examined the relationship of previous falls, gait, and balance with falls in community-dwelling older cancer survivors. At the baseline, demographics, health information, and the history of falls in the past year were obtained through interviewing. Participants performed tests including gait speed, Balance Evaluation Systems Test, and short-version of Activities-specific Balance Confidence scale. Falls were tracked by mailing of monthly reports for 6 months. A "faller" was a person with ≥1 fall during follow-up. Univariate analyses, including independent sample t-tests and Fisher's exact tests, compared baseline demographics, gait speed, and balance between fallers and non-fallers. For univariate analyses, Bonferroni correction was applied for multiple comparisons. Baseline variables with P<0.15 were included in a forward logistic regression model to identify factors predictive of falls with age as covariate. Sensitivity and specificity of each predictor of falls in the model were calculated. Significance level for the regression analysis was P<0.05. During follow-up, 59% of participants had one or more falls. Baseline demographics, health information, history of falls, gaits speed, and balance tests did not differ significantly between fallers and non-fallers. Forward logistic regression revealed that a history of falls was a significant predictor of falls in the final model (odds ratio =6.81; 95% confidence interval =1.594-29.074) (P<0.05). Sensitivity and specificity for correctly identifying a faller using the positive history of falls were 74% and 69%, respectively. Current findings suggested that for community-dwelling older cancer survivors with mixed diagnoses, asking about the history of falls may help detect individuals at risk of falling.
History of falls, gait, balance, and fall risks in older cancer survivors living in the community

PubMed Central

Huang, Min H; Shilling, Tracy; Miller, Kara A; Smith, Kristin; LaVictoire, Kayle

2015-01-01

Older cancer survivors may be predisposed to falls because cancer-related sequelae affect virtually all body systems. The use of a history of falls, gait speed, and balance tests to assess fall risks remains to be investigated in this population. This study examined the relationship of previous falls, gait, and balance with falls in community-dwelling older cancer survivors. At the baseline, demographics, health information, and the history of falls in the past year were obtained through interviewing. Participants performed tests including gait speed, Balance Evaluation Systems Test, and short-version of Activities-specific Balance Confidence scale. Falls were tracked by mailing of monthly reports for 6 months. A “faller” was a person with ≥1 fall during follow-up. Univariate analyses, including independent sample t-tests and Fisher’s exact tests, compared baseline demographics, gait speed, and balance between fallers and non-fallers. For univariate analyses, Bonferroni correction was applied for multiple comparisons. Baseline variables with P<0.15 were included in a forward logistic regression model to identify factors predictive of falls with age as covariate. Sensitivity and specificity of each predictor of falls in the model were calculated. Significance level for the regression analysis was P<0.05. During follow-up, 59% of participants had one or more falls. Baseline demographics, health information, history of falls, gaits speed, and balance tests did not differ significantly between fallers and non-fallers. Forward logistic regression revealed that a history of falls was a significant predictor of falls in the final model (odds ratio =6.81; 95% confidence interval =1.594–29.074) (P<0.05). Sensitivity and specificity for correctly identifying a faller using the positive history of falls were 74% and 69%, respectively. Current findings suggested that for community-dwelling older cancer survivors with mixed diagnoses, asking about the history of falls may help detect individuals at risk of falling. PMID:26425079
Presenilin E318G variant and Alzheimer's disease risk: the Cache County study.

PubMed

Hippen, Ariel A; Ebbert, Mark T W; Norton, Maria C; Tschanz, JoAnn T; Munger, Ronald G; Corcoran, Christopher D; Kauwe, John S K

2016-06-29

Alzheimer's disease is the leading cause of dementia in the elderly and the third most common cause of death in the United States. A vast number of genes regulate Alzheimer's disease, including Presenilin 1 (PSEN1). Multiple studies have attempted to locate novel variants in the PSEN1 gene that affect Alzheimer's disease status. A recent study suggested that one of these variants, PSEN1 E318G (rs17125721), significantly affects Alzheimer's disease status in a large case-control dataset, particularly in connection with the APOEε4 allele. Our study looks at the same variant in the Cache County Study on Memory and Aging, a large population-based dataset. We tested for association between E318G genotype and Alzheimer's disease status by running a series of Fisher's exact tests. We also performed logistic regression to test for an additive effect of E318G genotype on Alzheimer's disease status and for the existence of an interaction between E318G and APOEε4. In our Fisher's exact test, it appeared that APOEε4 carriers with an E318G allele have slightly higher risk for AD than those without the allele (3.3 vs. 3.8); however, the 95 % confidence intervals of those estimates overlapped completely, indicating non-significance. Our logistic regression model found a positive but non-significant main effect for E318G (p = 0.895). The interaction term between E318G and APOEε4 was also non-significant (p = 0.689). Our findings do not provide significant support for E318G as a risk factor for AD in APOEε4 carriers. Our calculations indicated that the overall sample used in the logistic regression models was adequately powered to detect the sort of effect sizes observed previously. However, the power analyses of our Fisher's exact tests indicate that our partitioned data was underpowered, particularly in regards to the low number of E318G carriers, both AD cases and controls, in the Cache county dataset. Thus, the differences in types of datasets used may help to explain the difference in effect magnitudes seen. Analyses in additional case-control datasets will be required to understand fully the effect of E318G on Alzheimer's disease status.
Logistic regression for circular data

NASA Astrophysics Data System (ADS)

Al-Daffaie, Kadhem; Khan, Shahjahan

2017-05-01

This paper considers the relationship between a binary response and a circular predictor. It develops the logistic regression model by employing the linear-circular regression approach. The maximum likelihood method is used to estimate the parameters. The Newton-Raphson numerical method is used to find the estimated values of the parameters. A data set from weather records of Toowoomba city is analysed by the proposed methods. Moreover, a simulation study is considered. The R software is used for all computations and simulations.
Naval Research Logistics Quarterly. Volume 28. Number 3,

DTIC Science & Technology

1981-09-01

denotes component-wise maximum. f has antone (isotone) differences on C x D if for cl < c2 and d, < d2, NAVAL RESEARCH LOGISTICS QUARTERLY VOL. 28...or negative correlations and linear or nonlinear regressions. Given are the mo- ments to order two and, for special cases, (he regression function and...data sets. We designate this bnb distribution as G - B - N(a, 0, v). The distribution admits only of positive correlation and linear regressions
Conditional Poisson models: a flexible alternative to conditional logistic case cross-over analysis.

PubMed

Armstrong, Ben G; Gasparrini, Antonio; Tobias, Aurelio

2014-11-24

The time stratified case cross-over approach is a popular alternative to conventional time series regression for analysing associations between time series of environmental exposures (air pollution, weather) and counts of health outcomes. These are almost always analyzed using conditional logistic regression on data expanded to case-control (case crossover) format, but this has some limitations. In particular adjusting for overdispersion and auto-correlation in the counts is not possible. It has been established that a Poisson model for counts with stratum indicators gives identical estimates to those from conditional logistic regression and does not have these limitations, but it is little used, probably because of the overheads in estimating many stratum parameters. The conditional Poisson model avoids estimating stratum parameters by conditioning on the total event count in each stratum, thus simplifying the computing and increasing the number of strata for which fitting is feasible compared with the standard unconditional Poisson model. Unlike the conditional logistic model, the conditional Poisson model does not require expanding the data, and can adjust for overdispersion and auto-correlation. It is available in Stata, R, and other packages. By applying to some real data and using simulations, we demonstrate that conditional Poisson models were simpler to code and shorter to run than are conditional logistic analyses and can be fitted to larger data sets than possible with standard Poisson models. Allowing for overdispersion or autocorrelation was possible with the conditional Poisson model but when not required this model gave identical estimates to those from conditional logistic regression. Conditional Poisson regression models provide an alternative to case crossover analysis of stratified time series data with some advantages. The conditional Poisson model can also be used in other contexts in which primary control for confounding is by fine stratification.
Artificial neural networks predict the incidence of portosplenomesenteric venous thrombosis in patients with acute pancreatitis.

PubMed

Fei, Y; Hu, J; Li, W-Q; Wang, W; Zong, G-Q

2017-03-01

Essentials Predicting the occurrence of portosplenomesenteric vein thrombosis (PSMVT) is difficult. We studied 72 patients with acute pancreatitis. Artificial neural networks modeling was more accurate than logistic regression in predicting PSMVT. Additional predictive factors may be incorporated into artificial neural networks. Objective To construct and validate artificial neural networks (ANNs) for predicting the occurrence of portosplenomesenteric venous thrombosis (PSMVT) and compare the predictive ability of the ANNs with that of logistic regression. Methods The ANNs and logistic regression modeling were constructed using simple clinical and laboratory data of 72 acute pancreatitis (AP) patients. The ANNs and logistic modeling were first trained on 48 randomly chosen patients and validated on the remaining 24 patients. The accuracy and the performance characteristics were compared between these two approaches by SPSS17.0 software. Results The training set and validation set did not differ on any of the 11 variables. After training, the back propagation network training error converged to 1 × 10 -20 , and it retained excellent pattern recognition ability. When the ANNs model was applied to the validation set, it revealed a sensitivity of 80%, specificity of 85.7%, a positive predictive value of 77.6% and negative predictive value of 90.7%. The accuracy was 83.3%. Differences could be found between ANNs modeling and logistic regression modeling in these parameters (10.0% [95% CI, -14.3 to 34.3%], 14.3% [95% CI, -8.6 to 37.2%], 15.7% [95% CI, -9.9 to 41.3%], 11.8% [95% CI, -8.2 to 31.8%], 22.6% [95% CI, -1.9 to 47.1%], respectively). When ANNs modeling was used to identify PSMVT, the area under receiver operating characteristic curve was 0.849 (95% CI, 0.807-0.901), which demonstrated better overall properties than logistic regression modeling (AUC = 0.716) (95% CI, 0.679-0.761). Conclusions ANNs modeling was a more accurate tool than logistic regression in predicting the occurrence of PSMVT following AP. More clinical factors or biomarkers may be incorporated into ANNs modeling to improve its predictive ability. © 2016 International Society on Thrombosis and Haemostasis.
PREDICTION OF MALIGNANT BREAST LESIONS FROM MRI FEATURES: A COMPARISON OF ARTIFICIAL NEURAL NETWORK AND LOGISTIC REGRESSION TECHNIQUES

PubMed Central

McLaren, Christine E.; Chen, Wen-Pin; Nie, Ke; Su, Min-Ying

2009-01-01

Rationale and Objectives Dynamic contrast enhanced MRI (DCE-MRI) is a clinical imaging modality for detection and diagnosis of breast lesions. Analytical methods were compared for diagnostic feature selection and performance of lesion classification to differentiate between malignant and benign lesions in patients. Materials and Methods The study included 43 malignant and 28 benign histologically-proven lesions. Eight morphological parameters, ten gray level co-occurrence matrices (GLCM) texture features, and fourteen Laws’ texture features were obtained using automated lesion segmentation and quantitative feature extraction. Artificial neural network (ANN) and logistic regression analysis were compared for selection of the best predictors of malignant lesions among the normalized features. Results Using ANN, the final four selected features were compactness, energy, homogeneity, and Law_LS, with area under the receiver operating characteristic curve (AUC) = 0.82, and accuracy = 0.76. The diagnostic performance of these 4-features computed on the basis of logistic regression yielded AUC = 0.80 (95% CI, 0.688 to 0.905), similar to that of ANN. The analysis also shows that the odds of a malignant lesion decreased by 48% (95% CI, 25% to 92%) for every increase of 1 SD in the Law_LS feature, adjusted for differences in compactness, energy, and homogeneity. Using logistic regression with z-score transformation, a model comprised of compactness, NRL entropy, and gray level sum average was selected, and it had the highest overall accuracy of 0.75 among all models, with AUC = 0.77 (95% CI, 0.660 to 0.880). When logistic modeling of transformations using the Box-Cox method was performed, the most parsimonious model with predictors, compactness and Law_LS, had an AUC of 0.79 (95% CI, 0.672 to 0.898). Conclusion The diagnostic performance of models selected by ANN and logistic regression was similar. The analytic methods were found to be roughly equivalent in terms of predictive ability when a small number of variables were chosen. The robust ANN methodology utilizes a sophisticated non-linear model, while logistic regression analysis provides insightful information to enhance interpretation of the model features. PMID:19409817
Logistic regression analysis of factors associated with avascular necrosis of the femoral head following femoral neck fractures in middle-aged and elderly patients.

PubMed

Ai, Zi-Sheng; Gao, You-Shui; Sun, Yuan; Liu, Yue; Zhang, Chang-Qing; Jiang, Cheng-Hua

2013-03-01

Risk factors for femoral neck fracture-induced avascular necrosis of the femoral head have not been elucidated clearly in middle-aged and elderly patients. Moreover, the high incidence of screw removal in China and its effect on the fate of the involved femoral head require statistical methods to reflect their intrinsic relationship. Ninety-nine patients older than 45 years with femoral neck fracture were treated by internal fixation between May 1999 and April 2004. Descriptive analysis, interaction analysis between associated factors, single factor logistic regression, multivariate logistic regression, and detailed interaction analysis were employed to explore potential relationships among associated factors. Avascular necrosis of the femoral head was found in 15 cases (15.2 %). Age × the status of implants (removal vs. maintenance) and gender × the timing of reduction were interactive according to two-factor interactive analysis. Age, the displacement of fractures, the quality of reduction, and the status of implants were found to be significant factors in single factor logistic regression analysis. Age, age × the status of implants, and the quality of reduction were found to be significant factors in multivariate logistic regression analysis. In fine interaction analysis after multivariate logistic regression analysis, implant removal was the most important risk factor for avascular necrosis in 56-to-85-year-old patients, with a risk ratio of 26.00 (95 % CI = 3.076-219.747). The middle-aged and elderly have less incidence of avascular necrosis of the femoral head following femoral neck fractures treated by cannulated screws. The removal of cannulated screws can induce a significantly high incidence of avascular necrosis of the femoral head in elderly patients, while a high-quality reduction is helpful to reduce avascular necrosis.
Multivariate logistic regression analysis of postoperative complications and risk model establishment of gastrectomy for gastric cancer: A single-center cohort report.

PubMed

Zhou, Jinzhe; Zhou, Yanbing; Cao, Shougen; Li, Shikuan; Wang, Hao; Niu, Zhaojian; Chen, Dong; Wang, Dongsheng; Lv, Liang; Zhang, Jian; Li, Yu; Jiao, Xuelong; Tan, Xiaojie; Zhang, Jianli; Wang, Haibo; Zhang, Bingyuan; Lu, Yun; Sun, Zhenqing

2016-01-01

Reporting of surgical complications is common, but few provide information about the severity and estimate risk factors of complications. If have, but lack of specificity. We retrospectively analyzed data on 2795 gastric cancer patients underwent surgical procedure at the Affiliated Hospital of Qingdao University between June 2007 and June 2012, established multivariate logistic regression model to predictive risk factors related to the postoperative complications according to the Clavien-Dindo classification system. Twenty-four out of 86 variables were identified statistically significant in univariate logistic regression analysis, 11 significant variables entered multivariate analysis were employed to produce the risk model. Liver cirrhosis, diabetes mellitus, Child classification, invasion of neighboring organs, combined resection, introperative transfusion, Billroth II anastomosis of reconstruction, malnutrition, surgical volume of surgeons, operating time and age were independent risk factors for postoperative complications after gastrectomy. Based on logistic regression equation, p=Exp∑BiXi / (1+Exp∑BiXi), multivariate logistic regression predictive model that calculated the risk of postoperative morbidity was developed, p = 1/(1 + e((4.810-1.287X1-0.504X2-0.500X3-0.474X4-0.405X5-0.318X6-0.316X7-0.305X8-0.278X9-0.255X10-0.138X11))). The accuracy, sensitivity and specificity of the model to predict the postoperative complications were 86.7%, 76.2% and 88.6%, respectively. This risk model based on Clavien-Dindo grading severity of complications system and logistic regression analysis can predict severe morbidity specific to an individual patient's risk factors, estimate patients' risks and benefits of gastric surgery as an accurate decision-making tool and may serve as a template for the development of risk models for other surgical groups.
Building and verifying a severity prediction model of acute pancreatitis (AP) based on BISAP, MEWS and routine test indexes.

PubMed

Ye, Jiang-Feng; Zhao, Yu-Xin; Ju, Jian; Wang, Wei

2017-10-01

To discuss the value of the Bedside Index for Severity in Acute Pancreatitis (BISAP), Modified Early Warning Score (MEWS), serum Ca2+, similarly hereinafter, and red cell distribution width (RDW) for predicting the severity grade of acute pancreatitis and to develop and verify a more accurate scoring system to predict the severity of AP. In 302 patients with AP, we calculated BISAP and MEWS scores and conducted regression analyses on the relationships of BISAP scoring, RDW, MEWS, and serum Ca2+ with the severity of AP using single-factor logistics. The variables with statistical significance in the single-factor logistic regression were used in a multi-factor logistic regression model; forward stepwise regression was used to screen variables and build a multi-factor prediction model. A receiver operating characteristic curve (ROC curve) was constructed, and the significance of multi- and single-factor prediction models in predicting the severity of AP using the area under the ROC curve (AUC) was evaluated. The internal validity of the model was verified through bootstrapping. Among 302 patients with AP, 209 had mild acute pancreatitis (MAP) and 93 had severe acute pancreatitis (SAP). According to single-factor logistic regression analysis, we found that BISAP, MEWS and serum Ca2+ are prediction indexes of the severity of AP (P-value<0.001), whereas RDW is not a prediction index of AP severity (P-value>0.05). The multi-factor logistic regression analysis showed that BISAP and serum Ca2+ are independent prediction indexes of AP severity (P-value<0.001), and MEWS is not an independent prediction index of AP severity (P-value>0.05); BISAP is negatively related to serum Ca2+ (r=-0.330, P-value<0.001). The constructed model is as follows: ln()=7.306+1.151*BISAP-4.516*serum Ca2+. The predictive ability of each model for SAP follows the order of the combined BISAP and serum Ca2+ prediction model>Ca2+>BISAP. There is no statistical significance for the predictive ability of BISAP and serum Ca2+ (P-value>0.05); however, there is remarkable statistical significance for the predictive ability using the newly built prediction model as well as BISAP and serum Ca2+ individually (P-value<0.01). Verification of the internal validity of the models by bootstrapping is favorable. BISAP and serum Ca2+ have high predictive value for the severity of AP. However, the model built by combining BISAP and serum Ca2+ is remarkably superior to those of BISAP and serum Ca2+ individually. Furthermore, this model is simple, practical and appropriate for clinical use. Copyright © 2016. Published by Elsevier Masson SAS.
Use of geographically weighted logistic regression to quantify spatial variation in the environmental and sociodemographic drivers of leptospirosis in Fiji: a modelling study.

PubMed

Mayfield, Helen J; Lowry, John H; Watson, Conall H; Kama, Mike; Nilles, Eric J; Lau, Colleen L

2018-05-01

Leptospirosis is a globally important zoonotic disease, with complex exposure pathways that depend on interactions between human beings, animals, and the environment. Major drivers of outbreaks include flooding, urbanisation, poverty, and agricultural intensification. The intensity of these drivers and their relative importance vary between geographical areas; however, non-spatial regression methods are incapable of capturing the spatial variations. This study aimed to explore the use of geographically weighted logistic regression (GWLR) to provide insights into the ecoepidemiology of human leptospirosis in Fiji. We obtained field data from a cross-sectional community survey done in 2013 in the three main islands of Fiji. A blood sample obtained from each participant (aged 1-90 years) was tested for anti-Leptospira antibodies and household locations were recorded using GPS receivers. We used GWLR to quantify the spatial variation in the relative importance of five environmental and sociodemographic covariates (cattle density, distance to river, poverty rate, residential setting [urban or rural], and maximum rainfall in the wettest month) on leptospirosis transmission in Fiji. We developed two models, one using GWLR and one with standard logistic regression; for each model, the dependent variable was the presence or absence of anti-Leptospira antibodies. GWLR results were compared with results obtained with standard logistic regression, and used to produce a predictive risk map and maps showing the spatial variation in odds ratios (OR) for each covariate. The dataset contained location information for 2046 participants from 1922 households representing 81 communities. The Aikaike information criterion value of the GWLR model was 1935·2 compared with 1254·2 for the standard logistic regression model, indicating that the GWLR model was more efficient. Both models produced similar OR for the covariates, but GWLR also detected spatial variation in the effect of each covariate. Maximum rainfall had the least variation across space (median OR 1·30, IQR 1·27-1·35), and distance to river varied the most (1·45, 1·35-2·05). The predictive risk map indicated that the highest risk was in the interior of Viti Levu, and the agricultural region and southern end of Vanua Levu. GWLR provided a valuable method for modelling spatial heterogeneity of covariates for leptospirosis infection and their relative importance over space. Results of GWLR could be used to inform more place-specific interventions, particularly for diseases with strong environmental or sociodemographic drivers of transmission. WHO, Australian National Health & Medical Research Council, University of Queensland, UK Medical Research Council, Chadwick Trust. Copyright © 2018 The Author(s). Published by Elsevier Ltd. This is an Open Access article under the CC BY 4.0 license. Published by Elsevier Ltd.. All rights reserved.
Factors Associated with HIV Testing Among Participants from Substance Use Disorder Treatment Programs in the US: A Machine Learning Approach.

PubMed

Pan, Yue; Liu, Hongmei; Metsch, Lisa R; Feaster, Daniel J

2017-02-01

HIV testing is the foundation for consolidated HIV treatment and prevention. In this study, we aim to discover the most relevant variables for predicting HIV testing uptake among substance users in substance use disorder treatment programs by applying random forest (RF), a robust multivariate statistical learning method. We also provide a descriptive introduction to this method for those who are unfamiliar with it. We used data from the National Institute on Drug Abuse Clinical Trials Network HIV testing and counseling study (CTN-0032). A total of 1281 HIV-negative or status unknown participants from 12 US community-based substance use disorder treatment programs were included and were randomized into three HIV testing and counseling treatment groups. The a priori primary outcome was self-reported receipt of HIV test results. Classification accuracy of RF was compared to logistic regression, a standard statistical approach for binary outcomes. Variable importance measures for the RF model were used to select the most relevant variables. RF based models produced much higher classification accuracy than those based on logistic regression. Treatment group is the most important predictor among all covariates, with a variable importance index of 12.9%. RF variable importance revealed that several types of condomless sex behaviors, condom use self-efficacy and attitudes towards condom use, and level of depression are the most important predictors of receipt of HIV testing results. There is a non-linear negative relationship between count of condomless sex acts and the receipt of HIV testing. In conclusion, RF seems promising in discovering important factors related to HIV testing uptake among large numbers of predictors and should be encouraged in future HIV prevention and treatment research and intervention program evaluations.
A population study of the contribution of medical comorbidity to the risk of prematurity in blacks.

PubMed

Ehrenthal, Deborah B; Jurkovitz, Claudine; Hoffman, Matthew; Kroelinger, Charlan; Weintraub, William

2007-10-01

The purpose of this study was to test the hypothesis that the higher prevalence of medical comorbidities among black women accounts for their increased risk of prematurity. A population-based regional cohort of women receiving obstetric care for singleton pregnancies at a large community hospital between 2003 and 2006 were analyzed using univariate and multivariable logistic regression. Data for 18,624 consecutive births found increased odds of adverse outcomes for black compared to white women: prematurity OR = 1.6 (1.4-1.8), extreme prematurity OR = 2.5 (2.0-3.2). Logistic regression modeling identified black race, age < 20, preconception diabetes and hypertension, smoking, underweight, and gestational hypertension as the greatest risks for adverse outcomes. Controlling for these risks did not attenuate the higher risk for prematurity among blacks. Though there is a greater burden of health risk among black women, this did not account for the higher rates of low birthweight and prematurity.
Quantitative appraisal of the Amyloid Imaging Taskforce appropriate use criteria for amyloid-PET.

PubMed

Altomare, Daniele; Ferrari, Clarissa; Festari, Cristina; Guerra, Ugo Paolo; Muscio, Cristina; Padovani, Alessandro; Frisoni, Giovanni B; Boccardi, Marina

2018-04-18

We test the hypothesis that amyloid-PET prescriptions, considered appropriate based on the Amyloid Imaging Taskforce (AIT) criteria, lead to greater clinical utility than AIT-inappropriate prescriptions. We compared the clinical utility between patients who underwent amyloid-PET appropriately or inappropriately and among the subgroups of patients defined by the AIT criteria. Finally, we performed logistic regressions to identify variables associated with clinical utility. We identified 171 AIT-appropriate and 67 AIT-inappropriate patients. AIT-appropriate and AIT-inappropriate cases did not differ in any outcomes of clinical utility (P > .05). Subgroup analysis denoted both expected and unexpected results. The logistic regressions outlined the primary role of clinical picture and clinical or neuropsychological profile in identifying patients benefitting from amyloid-PET. Contrary to our hypothesis, also AIT-inappropriate prescriptions were associated with clinical utility. Clinical or neuropsychological variables, not taken into account by the AIT criteria, may help further refine criteria for appropriateness. Copyright © 2018. Published by Elsevier Inc.
Gluten-free is not enough--perception and suggestions of celiac consumers.

PubMed

do Nascimento, Amanda Bagolin; Fiates, Giovanna Medeiros Rataichesck; dos Anjos, Adilson; Teixeira, Evanilda

2014-06-01

The present study investigated the perceptions of individuals with celiac disease about gluten-free (GF) products, their consumer behavior and which product is the most desired. A survey was used to collect information. Descriptive analysis, χ² tests and Multiple Logistic Regressions were conducted. Ninety-one questionnaires were analyzed. Limited variety and availability, the high price of products and the social restrictions imposed by the diet were the factors that caused the most dissatisfaction and difficulty. A total of 71% of the participants confirmed having moderate to high difficulty finding GF products. The logistic regression identified a significant relationship between dissatisfaction, texture and variety (p < 0.05) and between variety and difficulty of finding GF products (p < 0.05). The sensory characteristics were the most important variables considered for actual purchases. Bread was the most desired product. The participants were dissatisfaction with GF products. The desire for bread with better sensory characteristics reinforces the challenge to develop higher quality baking products.

An Exploratory Factor Analysis of Coping Styles and Relationship to Depression Among a Sample of Homeless Youth.

PubMed

Brown, Samantha M; Begun, Stephanie; Bender, Kimberly; Ferguson, Kristin M; Thompson, Sanna J

2015-10-01

The extent to which measures of coping adequately capture the ways that homeless youth cope with challenges, and the influence these coping styles have on mental health outcomes, is largely absent from the literature. This study tests the factor structure of the Coping Scale using Exploratory Factor Analysis (EFA) and then investigates the relationship between coping styles and depression using hierarchical logistic regression with data from 201 homeless youth. Results of the EFA indicate a 3-factor structure of coping, which includes active, avoidant, and social coping styles. Results of the hierarchical logistic regression show that homeless youth who engage in greater avoidant coping are at increased risk of meeting criteria for major depressive disorder. Findings provide insight into the utility of a preliminary tool for assessing homeless youths' coping styles. Such assessment may identify malleable risk factors that could be addressed by service providers to help prevent mental health problems.
Cross-sectional study on risk factors of HIV among female commercial sex workers in Cambodia.

PubMed Central

Ohshige, K.; Morio, S.; Mizushima, S.; Kitamura, K.; Tajima, K.; Ito, A.; Suyama, A.; Usuku, S.; Saphonn, V.; Heng, S.; Hor, L. B.; Tia, P.; Soda, K.

2000-01-01

To describe epidemiological features on HIV prevalence among female commercial sex workers (CSWs), a cross-sectional study on sexual behaviour and serological prevalence was carried out in Cambodia. The CSWs were interviewed on their demographic characters and behaviour and their blood samples were taken for testing on sexually transmitted diseases, including HIV, Chlamydia trachomatis, syphilis, and hepatitis B. Associations between risk factors and HIV seropositivity were analysed. High seroprevalence of HIV and Chlamydia trachomatis IgG antibody (CT-IgG-Ab) was shown among the CSWs (54 and 81.7%, respectively). Univariate logistic regression analyses showed an association between HIV seropositivity and age, duration of prostitution, the number of clients per day and CT-IgG-Ab. Especially, high-titre chlamydial seropositivity showed a strong significant association with HIV prevalence. In multiple logistic regression analyses, CT-IgG-Ab with higher titre was significantly independently related to HIV infection. These suggest that existence of Chlamydia trachomatis is highly related to HIV prevalence. PMID:10722142
A Pilot Study of Reasons and Risk Factors for "No-Shows" in a Pediatric Neurology Clinic.

PubMed

Guzek, Lindsay M; Fadel, William F; Golomb, Meredith R

2015-09-01

Missed clinic appointments lead to decreased patient access, worse patient outcomes, and increased healthcare costs. The goal of this pilot study was to identify reasons for and risk factors associated with missed pediatric neurology outpatient appointments ("no-shows"). This was a prospective cohort study of patients scheduled for 1 week of clinic. Data on patient clinical and demographic information were collected by record review; data on reasons for missed appointments were collected by phone interviews. Univariate and multivariate analyses were conducted using chi-square tests and multiple logistic regression to assess risk factors for missed appointments. Fifty-nine (25%) of 236 scheduled patients were no-shows. Scheduling conflicts (25.9%) and forgetting (20.4%) were the most common reasons for missed appointments. When controlling for confounding factors in the logistic regression, Medicaid (odds ratio 2.36), distance from clinic, and time since appointment was scheduled were associated with missed appointments. Further work in this area is needed. © The Author(s) 2014.
A modified approach to estimating sample size for simple logistic regression with one continuous covariate.

PubMed

Novikov, I; Fund, N; Freedman, L S

2010-01-15

Different methods for the calculation of sample size for simple logistic regression (LR) with one normally distributed continuous covariate give different results. Sometimes the difference can be large. Furthermore, some methods require the user to specify the prevalence of cases when the covariate equals its population mean, rather than the more natural population prevalence. We focus on two commonly used methods and show through simulations that the power for a given sample size may differ substantially from the nominal value for one method, especially when the covariate effect is large, while the other method performs poorly if the user provides the population prevalence instead of the required parameter. We propose a modification of the method of Hsieh et al. that requires specification of the population prevalence and that employs Schouten's sample size formula for a t-test with unequal variances and group sizes. This approach appears to increase the accuracy of the sample size estimates for LR with one continuous covariate.
Predictors of adherence with self-care guidelines among persons with type 2 diabetes: results from a logistic regression tree analysis.

PubMed

Yamashita, Takashi; Kart, Cary S; Noe, Douglas A

2012-12-01

Type 2 diabetes is known to contribute to health disparities in the U.S. and failure to adhere to recommended self-care behaviors is a contributing factor. Intervention programs face difficulties as a result of patient diversity and limited resources. With data from the 2005 Behavioral Risk Factor Surveillance System, this study employs a logistic regression tree algorithm to identify characteristics of sub-populations with type 2 diabetes according to their reported frequency of adherence to four recommended diabetes self-care behaviors including blood glucose monitoring, foot examination, eye examination and HbA1c testing. Using Andersen's health behavior model, need factors appear to dominate the definition of which sub-groups were at greatest risk for low as well as high adherence. Findings demonstrate the utility of easily interpreted tree diagrams to design specific culturally appropriate intervention programs targeting sub-populations of diabetes patients who need to improve their self-care behaviors. Limitations and contributions of the study are discussed.
Risk Factors for Developing Scoliosis in Cerebral Palsy: A Cross-Sectional Descriptive Study.

PubMed

Bertoncelli, Carlo M; Solla, Federico; Loughenbury, Peter R; Tsirikos, Athanasios I; Bertoncelli, Domenico; Rampal, Virginie

2017-06-01

This study aims to identify the risk factors leading to the development of severe scoliosis among children with cerebral palsy. A cross-sectional descriptive study of 70 children (aged 12-18 years) with severe spastic and/or dystonic cerebral palsy treated in a single specialist unit is described. Statistical analysis included Fisher exact test and logistic regression analysis to identify risk factors. Severe scoliosis is more likely to occur in patients with intractable epilepsy ( P = .008), poor gross motor functional assessment scores ( P = .018), limb spasticity ( P = .045), a history of previous hip surgery ( P = .048), and nonambulatory patients ( P = .013). Logistic regression model confirms the major risk factors are previous hip surgery ( P = .001), moderate to severe epilepsy ( P = .007), and female gender ( P = .03). History of previous hip surgery, intractable epilepsy, and female gender are predictors of developing severe scoliosis in children with cerebral palsy. This knowledge should aid in the early diagnosis of scoliosis and timely referral to specialist services.
Rank-Optimized Logistic Matrix Regression toward Improved Matrix Data Classification.

PubMed

Zhang, Jianguang; Jiang, Jianmin

2018-02-01

While existing logistic regression suffers from overfitting and often fails in considering structural information, we propose a novel matrix-based logistic regression to overcome the weakness. In the proposed method, 2D matrices are directly used to learn two groups of parameter vectors along each dimension without vectorization, which allows the proposed method to fully exploit the underlying structural information embedded inside the 2D matrices. Further, we add a joint [Formula: see text]-norm on two parameter matrices, which are organized by aligning each group of parameter vectors in columns. This added co-regularization term has two roles-enhancing the effect of regularization and optimizing the rank during the learning process. With our proposed fast iterative solution, we carried out extensive experiments. The results show that in comparison to both the traditional tensor-based methods and the vector-based regression methods, our proposed solution achieves better performance for matrix data classifications.
Association between molecular markers and behavioral phenotypes in the immatures of a butterfly.

PubMed

De Nardin, Janaína; Buffon, Vanessa; Revers, Luís Fernando; de Araújo, Aldo Mellender

2018-01-01

Newly hatched caterpillars of the butterfly Heliconius erato phyllis routinely cannibalize eggs. In a manifestation of kin recognition they cannibalize sibling eggs less frequently than unrelated eggs. Previous work has estimated the heritability of kin recognition in H. erato phyllis to lie between 14 and 48%. It has furthermore been shown that the inheritance of kin recognition is compatible with a quantitative model with a threshold. Here we present the results of a preliminary study, in which we tested for associations between behavioral kin recognition phenotypes and AFLP and SSR markers. We implemented two experimental approaches: (1) a cannibalism test using sibling eggs only, which allowed for only two behavioral outcomes (cannibal and non-cannibal), and (2) a cannibalism test using two sibling eggs and one unrelated egg, which allowed four outcomes [cannibal who does not recognize siblings, cannibal who recognizes siblings, "super-cannibal" (cannibal of both eggs), and "super non-cannibal" (does not cannibalize eggs at all)]. Single-marker analyses were performed using χ2 tests and logistic regression with null markers as covariates. Results of the χ2 tests identified 72 associations for experimental design 1 and 73 associations for design 2. Logistic regression analysis of the markers found to be significant in the χ2 test resulted in 20 associations for design 1 and 11 associations for design 2. Experiment 2 identified markers that were more frequently present or absent in cannibals who recognize siblings and super non-cannibals; i.e. in both phenotypes capable of kin recognition.
Association between molecular markers and behavioral phenotypes in the immatures of a butterfly

PubMed Central

De Nardin, Janaína; Buffon, Vanessa; Revers, Luís Fernando; de Araújo, Aldo Mellender

2018-01-01

Abstract Newly hatched caterpillars of the butterfly Heliconius erato phyllis routinely cannibalize eggs. In a manifestation of kin recognition they cannibalize sibling eggs less frequently than unrelated eggs. Previous work has estimated the heritability of kin recognition in H. erato phyllis to lie between 14 and 48%. It has furthermore been shown that the inheritance of kin recognition is compatible with a quantitative model with a threshold. Here we present the results of a preliminary study, in which we tested for associations between behavioral kin recognition phenotypes and AFLP and SSR markers. We implemented two experimental approaches: (1) a cannibalism test using sibling eggs only, which allowed for only two behavioral outcomes (cannibal and non-cannibal), and (2) a cannibalism test using two sibling eggs and one unrelated egg, which allowed four outcomes [cannibal who does not recognize siblings, cannibal who recognizes siblings, “super-cannibal” (cannibal of both eggs), and “super non-cannibal” (does not cannibalize eggs at all)]. Single-marker analyses were performed using χ2 tests and logistic regression with null markers as covariates. Results of the χ2 tests identified 72 associations for experimental design 1 and 73 associations for design 2. Logistic regression analysis of the markers found to be significant in the χ2 test resulted in 20 associations for design 1 and 11 associations for design 2. Experiment 2 identified markers that were more frequently present or absent in cannibals who recognize siblings and super non-cannibals; i.e. in both phenotypes capable of kin recognition. PMID:29583155
Self-reported HIV antibody testing among Latino urban day laborers.

PubMed

Solorio, Maria Rosa; Galvan, Frank H

2009-12-01

To identify the characteristics of male Latino urban day laborers who self-report having tested for human immunodeficiency virus (HIV). A cross-sectional survey was conducted with 356 Latino day laborers, aged 18 to 40 years, who had been sexually active in the previous 12 months, from 6 day labor sites in the City of Los Angeles. Most of the men were single, mainly from Mexico and Guatemala, and had been employed as a day laborer for fewer than 3 years; 38% had an annual income of $4000 or less. Ninety-two percent of the men reported having sex with women only, and 8% reported a history of having sex with men and women. Forty-six percent had received an HIV test in the previous 12 months and 1 person tested positive. In univariate logistic regression analyses, day laborers who were aged 26 years or older, had more than 3 years in the United States, had more than 1 year but fewer than 5 years employed as a day laborer, and had annual incomes greater than $4000 were significantly more likely to self-report HIV testing in the previous 12 months. In a multivariate logistic regression analysis, only higher annual income was found to be significantly associated with self-reported HIV testing. Interventions that target lower-income Latino day laborers are needed to promote early HIV detection. HIV detection offers individual benefits through treatment, with decreased morbidity and mortality, as well as public health benefits through decreased rates of HIV transmission in the community.
Testing concordance of instrumental variable effects in generalized linear models with application to Mendelian randomization

PubMed Central

Dai, James Y.; Chan, Kwun Chuen Gary; Hsu, Li

2014-01-01

Instrumental variable regression is one way to overcome unmeasured confounding and estimate causal effect in observational studies. Built on structural mean models, there has been considerale work recently developed for consistent estimation of causal relative risk and causal odds ratio. Such models can sometimes suffer from identification issues for weak instruments. This hampered the applicability of Mendelian randomization analysis in genetic epidemiology. When there are multiple genetic variants available as instrumental variables, and causal effect is defined in a generalized linear model in the presence of unmeasured confounders, we propose to test concordance between instrumental variable effects on the intermediate exposure and instrumental variable effects on the disease outcome, as a means to test the causal effect. We show that a class of generalized least squares estimators provide valid and consistent tests of causality. For causal effect of a continuous exposure on a dichotomous outcome in logistic models, the proposed estimators are shown to be asymptotically conservative. When the disease outcome is rare, such estimators are consistent due to the log-linear approximation of the logistic function. Optimality of such estimators relative to the well-known two-stage least squares estimator and the double-logistic structural mean model is further discussed. PMID:24863158
Detecting DIF in Polytomous Items Using MACS, IRT and Ordinal Logistic Regression

ERIC Educational Resources Information Center

Elosua, Paula; Wells, Craig

2013-01-01

The purpose of the present study was to compare the Type I error rate and power of two model-based procedures, the mean and covariance structure model (MACS) and the item response theory (IRT), and an observed-score based procedure, ordinal logistic regression, for detecting differential item functioning (DIF) in polytomous items. A simulation…
Comparing Linear Discriminant Function with Logistic Regression for the Two-Group Classification Problem.

ERIC Educational Resources Information Center

Fan, Xitao; Wang, Lin

The Monte Carlo study compared the performance of predictive discriminant analysis (PDA) and that of logistic regression (LR) for the two-group classification problem. Prior probabilities were used for classification, but the cost of misclassification was assumed to be equal. The study used a fully crossed three-factor experimental design (with…
Effects of Social Class and School Conditions on Educational Enrollment and Achievement of Boys and Girls in Rural Viet Nam

ERIC Educational Resources Information Center

Nguyen, Phuong L.

2006-01-01

This study examines the effects of parental SES, school quality, and community factors on children's enrollment and achievement in rural areas in Viet Nam, using logistic regression and ordered logistic regression. Multivariate analysis reveals significant differences in educational enrollment and outcomes by level of household expenditures and…
Hierarchical Bayesian Logistic Regression to forecast metabolic control in type 2 DM patients.

PubMed

Dagliati, Arianna; Malovini, Alberto; Decata, Pasquale; Cogni, Giulia; Teliti, Marsida; Sacchi, Lucia; Cerra, Carlo; Chiovato, Luca; Bellazzi, Riccardo

2016-01-01

In this work we present our efforts in building a model able to forecast patients' changes in clinical conditions when repeated measurements are available. In this case the available risk calculators are typically not applicable. We propose a Hierarchical Bayesian Logistic Regression model, which allows taking into account individual and population variability in model parameters estimate. The model is used to predict metabolic control and its variation in type 2 diabetes mellitus. In particular we have analyzed a population of more than 1000 Italian type 2 diabetic patients, collected within the European project Mosaic. The results obtained in terms of Matthews Correlation Coefficient are significantly better than the ones gathered with standard logistic regression model, based on data pooling.
An empirical study of statistical properties of variance partition coefficients for multi-level logistic regression models

USGS Publications Warehouse

Li, Ji; Gray, B.R.; Bates, D.M.

2008-01-01

Partitioning the variance of a response by design levels is challenging for binomial and other discrete outcomes. Goldstein (2003) proposed four definitions for variance partitioning coefficients (VPC) under a two-level logistic regression model. In this study, we explicitly derived formulae for multi-level logistic regression model and subsequently studied the distributional properties of the calculated VPCs. Using simulations and a vegetation dataset, we demonstrated associations between different VPC definitions, the importance of methods for estimating VPCs (by comparing VPC obtained using Laplace and penalized quasilikehood methods), and bivariate dependence between VPCs calculated at different levels. Such an empirical study lends an immediate support to wider applications of VPC in scientific data analysis.
Maximal bite force, facial morphology and sucking habits in young children with functional posterior crossbite.

PubMed

Castelo, Paula Midori; Gavião, Maria Beatriz Duarte; Pereira, Luciano José; Bonjardim, Leonardo Rigoldi

2010-01-01

The maintenance of normal conditions of the masticatory function is determinant for the correct growth and development of its structures. Thus, the aims of this study were to evaluate the influence of sucking habits on the presence of crossbite and its relationship with maximal bite force, facial morphology and body variables in 67 children of both genders (3.5-7 years) with primary or early mixed dentition. The children were divided in four groups: primary-normocclusion (PN, n=19), primary-crossbite (PC, n=19), mixed-normocclusion (MN, n=13), and mixed-crossbite (MC, n=16). Bite force was measured with a pressurized tube, and facial morphology was determined by standardized frontal photographs: AFH (anterior face height) and BFW (bizygomatic facial width). It was observed that MC group showed lower bite force than MN, and AFH/BFW was significantly smaller in PN than PC (t-test). Weight and height were only significantly correlated with bite force in PC group (Pearson's correlation test). In the primary dentition, AFH/BFW and breast-feeding (at least six months) were positive and negatively associated with crossbite, respectively (multiple logistic regression). In the mixed dentition, breast-feeding and bite force showed negative associations with crossbite (univariate regression), while nonnutritive sucking (up to 3 years) associated significantly with crossbite in all groups (multiple logistic regression). In the studied sample, sucking habits played an important role in the etiology of crossbite, which was associated with lower bite force and long-face tendency.
The South African English Smartphone Digits-in-Noise Hearing Test: Effect of Age, Hearing Loss, and Speaking Competence.

PubMed

Potgieter, Jenni-Marí; Swanepoel, De Wet; Myburgh, Hermanus Carel; Smits, Cas

2017-11-20

This study determined the effect of hearing loss and English-speaking competency on the South African English digits-in-noise hearing test to evaluate its suitability for use across native (N) and non-native (NN) speakers. A prospective cross-sectional cohort study of N and NN English adults with and without sensorineural hearing loss compared pure-tone air conduction thresholds to the speech reception threshold (SRT) recorded with the smartphone digits-in-noise hearing test. A rating scale was used for NN English listeners' self-reported competence in speaking English. This study consisted of 454 adult listeners (164 male, 290 female; range 16 to 90 years), of whom 337 listeners had a best ear four-frequency pure-tone average (4FPTA; 0.5, 1, 2, and 4 kHz) of ≤25 dB HL. A linear regression model identified three predictors of the digits-in-noise SRT, namely, 4FPTA, age, and self-reported English-speaking competence. The NN group with poor self-reported English-speaking competence (≤5/10) performed significantly (p < 0.01) poorer than the N and NN (≥6/10) groups on the digits-in-noise test. Screening characteristics of the test improved with separate cutoff values depending on English-speaking competence for the N and NN groups (≥6/10) and NN group alone (≤5/10). Logistic regression models, which include age in the analysis, showed a further improvement in sensitivity and specificity for both groups (area under the receiver operating characteristic curve, 0.962 and 0.903, respectively). Self-reported English-speaking competence had a significant influence on the SRT obtained with the smartphone digits-in-noise test. A logistic regression approach considering SRT, self-reported English-speaking competence, and age as predictors of best ear 4FPTA >25 dB HL showed that the test can be used as an accurate hearing screening tool for N and NN English speakers. The smartphone digits-in-noise test, therefore, allows testing in a multilingual population familiar with English digits using dynamic cutoff values that can be chosen according to self-reported English-speaking competence and age.
Religiosity and decreased risk of substance use disorders: is the effect mediated by social support or mental health status?

PubMed Central

Harris, Katherine M.; Koenig, Harold G.; Han, Xiaotong; Sullivan, Greer; Mattox, Rhonda; Tang, Lingqi

2009-01-01

Objective The negative association between religiosity (religious beliefs and church attendance) and the likelihood of substance use disorders is well established, but the mechanism(s) remain poorly understood. We investigated whether this association was mediated by social support or mental health status. Method We utilized cross-sectional data from the 2002 National Survey on Drug Use and Health (n = 36,370). We first used logistic regression to regress any alcohol use in the past year on sociodemographic and religiosity variables. Then, among individuals who drank in the past year, we regressed past year alcohol abuse/dependence on sociodemographic and religiosity variables. To investigate whether social support mediated the association between religiosity and alcohol use and alcohol abuse/dependence we repeated the above models, adding the social support variables. To the extent that these added predictors modified the magnitude of the effect of the religiosity variables, we interpreted social support as a possible mediator. We also formally tested for mediation using path analysis. We investigated the possible mediating role of mental health status analogously. Parallel sets of analyses were conducted for any drug use, and drug abuse/dependence among those using any drugs as the dependent variables. Results The addition of social support and mental health status variables to logistic regression models had little effect on the magnitude of the religiosity coefficients in any of the models. While some of the tests of mediation were significant in the path analyses, the results were not always in the expected direction, and the magnitude of the effects was small. Conclusions The association between religiosity and decreased likelihood of a substance use disorder does not appear to be substantively mediated by either social support or mental health status. PMID:19714282
Acculturation and healthy lifestyle habits among Hispanics in United States-Mexico border communities.

PubMed

Ghaddar, Suad; Brown, Cynthia J; Pagán, José A; Díaz, Violeta

2010-09-01

To explore the relationship between acculturation and healthy lifestyle habits in the largely Hispanic populations living in underserved communities in the United States of America along the U.S.-Mexico border. A cross-sectional study was conducted from April 2006 to June 2008 using survey data from the Alliance for a Healthy Border, a program designed to reduce health disparities in the U.S.-Mexico border region by funding nutrition and physical activity education programs at 12 federally qualified community health centers in Arizona, California, New Mexico, and Texas. The survey included questions on acculturation, diet, exercise, and demographic factors and was completed by 2,381 Alliance program participants, of whom 95.3% were Hispanic and 45.4% were under the U.S. poverty level for 2007. Chi-square (χ2) and Student's t tests were used for bivariate comparisons between acculturation and dietary and physical activity measures. Linear regression and binary logistic regression were used to control for factors associated with nutrition and exercise. Based on univariate tests and confirmed by regression analysis controlling for sociodemographic and health variables, less acculturated survey respondents reported a significantly higher frequency of fruit and vegetable consumption and healthier dietary habits than those who were more acculturated. Adjusted binary logistic regression confirmed that individuals with low language acculturation were less likely to engage in physical activity than those with moderate to high acculturation (odds ratio 0.75, 95% confidence interval 0.59-0.95). Findings confirmed an association between acculturation and healthy lifestyle habits and supported the hypothesis that acculturation in border community populations tends to decrease the practice of some healthy dietary habits while increasing exposure to and awareness of the importance of other healthy behaviors.

Gender Differential Item Functioning on a National Field-Specific Test: The Case of PhD Entrance Exam of TEFL in Iran

ERIC Educational Resources Information Center

Ahmadi, Alireza; Bazvand, Ali Darabi

2016-01-01

Differential Item Functioning (DIF) exists when examinees of equal ability from different groups have different probabilities of successful performance in a certain item. This study examined gender differential item functioning across the PhD Entrance Exam of TEFL (PEET) in Iran, using both logistic regression (LR) and one-parameter item response…
Bioelectrical impedance analysis-derived phase angle at admission as a predictor of 90-day mortality in intensive care patients.

PubMed

Stapel, Sandra N; Looijaard, Wilhelmus G P M; Dekker, Ingeborg M; Girbes, Armand R J; Weijs, Peter J M; Oudemans-van Straaten, Heleen M

2018-05-11

A low bioelectrical impedance analysis (BIA)-derived phase angle (PA) predicts morbidity and mortality in different patient groups. An association between PA and long-term mortality in ICU patients has not been demonstrated before. The purpose of the present study was to determine whether PA on ICU admission independently predicts 90-day mortality. This prospective observational study was performed in a mixed university ICU. BIA was performed in 196 patients within 24 h of ICU admission. To test the independent association between PA and 90-day mortality, logistic regression analysis was performed using the APACHE IV predicted mortality as confounder. The optimal cutoff value of PA for mortality prediction was determined by ROC curve analysis. Using this cutoff value, patients were categorized into low or normal PA group and the association with 90-day mortality was tested again. The PA of survivors was higher than of the non-survivors (5.0° ± 1.3° vs. 4.1° ± 1.2°, p < 0.001). The area under the ROC curve of PA for 90-day mortality was 0.70 (CI 0.59-0.80). PA was associated with 90-day mortality (OR = 0.56, CI: 0.38-0.77, p = 0.001) on univariate logistic regression analysis and also after adjusting for BMI, gender, age, and APACHE IV on multivariable logistic regression (OR = 0.65, CI: 0.44-0.96, p = 0.031). A PA < 4.8° was an independent predictor of 90-day mortality (adjusted OR = 3.65, CI: 1.34-9.93, p = 0.011). Phase angle at ICU admission is an independent predictor of 90-day mortality. This biological marker can aid in long-term mortality risk assessment of critically ill patients.
[Study on the correlation among adolescents' family function, negative life events stress amount and suicide ideation].

PubMed

Zhang, Dongdong; Chen, Ling; Yin, Dan; Miao, Jinping; Sun, Yehuan

2014-07-01

To explore the correlation between suicide ideation and family function & negative life events, as well as other influential factors in adolescents, thus present a theoretical base for clinicians and school staff to develop intervention for those problems. By adopting current situation random sampling method, Self-Rating Idea of Suicide Scale, Adolescent Self-Rating Life Events Check List and Family APGAR Index were used to assess adolescents at random in a hygiene vocational school in Changzhou City, Jiangsu Province and a collage in Wuhu City, Anhui Province. 3700 questionnaires were granted, 3675 questionnaires were collected, among which 3620 were valid. Chi-square test, t-test, and univariate logistic regression were employed in univariate analysis, multivariate logistic regression was used in multivariate analysis. The detection rate of suicide ideation is 7.0%, and the top five suicide ideation characteristics were: poor academic performance (33.6%), serious family functional impairment (25.8%), lower-middle academic performance (11.7%), bad economic conditions (10.8%) and study in Grade Three (9.9%). Multiple logistic regression showed that the following three high-level stress amount in negative life events are most crucial for suicide ideation. They are "relationships" (OR = 1.135, 95% CI 1.071 - 1. 202), "academic pressure" (OR = 1.169, 95% CI 1.101 - 1.241), and "external events" (OR = 1.278, 95% CI 1.187 - 1.376). What' s more, the stress of attending higher grades (OR = 1.980, 95% CI 1.302 - 3.008), poor academic performance (OR = 7.206, 95% CI 1.745 - 9.789), moderate family functional impairment (OR = 2.562, 95% CI 1.527 - 2.892) and its serious level (OR = 8.287, 95% CI 3.154 - 6.917) are also influential factors for suicide ideation. Severe family functional impairment and high-level stress amount of negative life events produced the main factors of suicide ideation. Therefore, necessary and sufficient support should be given to adolescents by families and schools.
Artificial neural network in predicting craniocervical junction injury: an alternative approach to trauma patients.

PubMed

Bektaş, Frat; Eken, Cenker; Soyuncu, Secgin; Kilicaslan, Isa; Cete, Yildiray

2008-12-01

The aim of this study is to determine the efficiency of artificial intelligence in detecting craniocervical junction injuries by using an artificial neural network (ANN) that may be applicable in future studies of different traumatic injuries. Major head trauma patients with Glasgow Coma Scale
Latent profile analysis of regression-based norms demonstrates relationship of compounding MS symptom burden and negative work events.

PubMed

Frndak, Seth E; Smerbeck, Audrey M; Irwin, Lauren N; Drake, Allison S; Kordovski, Victoria M; Kunker, Katrina A; Khan, Anjum L; Benedict, Ralph H B

2016-10-01

We endeavored to clarify how distinct co-occurring symptoms relate to the presence of negative work events in employed multiple sclerosis (MS) patients. Latent profile analysis (LPA) was utilized to elucidate common disability patterns by isolating patient subpopulations. Samples of 272 employed MS patients and 209 healthy controls (HC) were administered neuroperformance tests of ambulation, hand dexterity, processing speed, and memory. Regression-based norms were created from the HC sample. LPA identified latent profiles using the regression-based z-scores. Finally, multinomial logistic regression tested for negative work event differences among the latent profiles. Four profiles were identified via LPA: a common profile (55%) characterized by slightly below average performance in all domains, a broadly low-performing profile (18%), a poor motor abilities profile with average cognition (17%), and a generally high-functioning profile (9%). Multinomial regression analysis revealed that the uniformly low-performing profile demonstrated a higher likelihood of reported negative work events. Employed MS patients with co-occurring motor, memory and processing speed impairments were most likely to report a negative work event, classifying them as uniquely at risk for job loss.
Assessing landslide susceptibility by statistical data analysis and GIS: the case of Daunia (Apulian Apennines, Italy)

NASA Astrophysics Data System (ADS)

Ceppi, C.; Mancini, F.; Ritrovato, G.

2009-04-01

This study aim at the landslide susceptibility mapping within an area of the Daunia (Apulian Apennines, Italy) by a multivariate statistical method and data manipulation in a Geographical Information System (GIS) environment. Among the variety of existing statistical data analysis techniques, the logistic regression was chosen to produce a susceptibility map all over an area where small settlements are historically threatened by landslide phenomena. By logistic regression a best fitting between the presence or absence of landslide (dependent variable) and the set of independent variables is performed on the basis of a maximum likelihood criterion, bringing to the estimation of regression coefficients. The reliability of such analysis is therefore due to the ability to quantify the proneness to landslide occurrences by the probability level produced by the analysis. The inventory of dependent and independent variables were managed in a GIS, where geometric properties and attributes have been translated into raster cells in order to proceed with the logistic regression by means of SPSS (Statistical Package for the Social Sciences) package. A landslide inventory was used to produce the bivariate dependent variable whereas the independent set of variable concerned with slope, aspect, elevation, curvature, drained area, lithology and land use after their reductions to dummy variables. The effect of independent parameters on landslide occurrence was assessed by the corresponding coefficient in the logistic regression function, highlighting a major role played by the land use variable in determining occurrence and distribution of phenomena. Once the outcomes of the logistic regression are determined, data are re-introduced in the GIS to produce a map reporting the proneness to landslide as predicted level of probability. As validation of results and regression model a cell-by-cell comparison between the susceptibility map and the initial inventory of landslide events was performed and an agreement at 75% level achieved.
[Epidemiological study on HIV/AIDS in Cambodia seroprevalence of HIV/STD among commercial sex workers].

PubMed

Ohshige, K; Morio, S; Mizushima, S; Kitamura, K; Tajima, K; Ito, A; Suyama, A; Usuku, S; Phalla, T; Leng, H B; Sopheab, H; Eab, B; Soda, K

1999-01-01

To describe epidemiological features of HIV prevalence among female commercial sex workers (CSWs) in Cambodia, a cross-sectional study using a questionnaire study and serological tests was carried out from December 1997 to January 1998. We report the main results of the analyses of serological tests in this article. Two hundred ninety six CSWs working in Sisophon and Poi Pet, located in northwest Cambodia, Bantey Mean Chey province, were recruited for interview based on a questionnaire on sexual behavior, and serological tests. The blood samples were examined for HIV antibody, Chlamydia trachomatis IgG antibody, TPHA, Hepatitis B surface antigen, and Hepatitis B surface antibody. The relationship between HIV and the other STD's was analyzed by using logistic regression analysis. The HIV seroprevalence rate was 43.9% (130 out of 296). The seropositive rate of Chlamydia trachomatis IgG antibody (C.T.-IgG-Ab) was 73.3% (217 out of 296). Logistic regression analysis showed a significant association between C.T.-IgG-Ab positive and HIV prevalence. (Odds Ratio: 5.33; 95% Confidence Interval, 2.82-10.07). This study suggests that the existence of Chlamydia trachomatis is closely related with HIV prevalence among CSWs in Cambodia. Other STDs may also increase susceptibility to male-to-female sexual transmission of HIV. This suggests that appropriate prevention against STDs will be needed for the control of HIV prevalence in Cambodia.
Risk factors for indications of intraoperative blood transfusion among patients undergoing surgical treatment for colorectal adenocarcinoma.

PubMed

Gonçalves, Iara; Linhares, Marcelo; Bordin, Jose; Matos, Delcio

2009-01-01

Identification of risk factors for requiring transfusions during surgery for colorectal cancer may lead to preventive actions or alternative measures, towards decreasing the use of blood components in these procedures, and also rationalization of resources use in hemotherapy services. This was a retrospective case-control study using data from 383 patients who were treated surgically for colorectal adenocarcinoma at 'Fundação Pio XII', in Barretos-SP, Brazil, between 1999 and 2003. To recognize significant risk factors for requiring intraoperative blood transfusion in colorectal cancer surgical procedures. Univariate analyses were performed using Fisher's exact test or the chi-squared test for dichotomous variables and Student's t test for continuous variables, followed by multivariate analysis using multiple logistic regression. In the univariate analyses, height (P = 0.06), glycemia (P = 0.05), previous abdominal or pelvic surgery (P = 0.031), abdominoperineal surgery (P<0.001), extended surgery (P<0.001) and intervention with radical intent (P<0.001) were considered significant. In the multivariate analysis using logistic regression, intervention with radical intent (OR = 10.249, P<0.001, 95% CI = 3.071-34.212) and abdominoperineal amputation (OR = 3.096, P = 0.04, 95% CI = 1.445-6.623) were considered to be independently significant. This investigation allows the conclusion that radical intervention and the abdominoperineal procedure in the surgical treatment of colorectal adenocarcinoma are risk factors for requiring intraoperative blood transfusion.
Calibration power of the Braden scale in predicting pressure ulcer development.

PubMed

Chen, Hong-Lin; Cao, Ying-Juan; Wang, Jing; Huai, Bao-Sha

2016-11-02

Calibration is the degree of correspondence between the estimated probability produced by a model and the actual observed probability. The aim of this study was to investigate the calibration power of the Braden scale in predicting pressure ulcer development (PU). A retrospective analysis was performed among consecutive patients in 2013. The patients were separated into training a group and a validation group. The predicted incidence was calculated using a logistic regression model in the training group and the Hosmer-Lemeshow test was used for assessing the goodness of fit. In the validation cohort, the observed and the predicted incidence were compared by the Chi-square (χ 2 ) goodness of fit test for calibration power. We included 2585 patients in the study, of these 78 patients (3.0%) developed a PU. Between the training and validation groups the patient characteristics were non-significant (p>0.05). In the training group, the logistic regression model for predicting pressure ulcer was Logit(P) = -0.433*Braden score+2.616. The Hosmer-Lemeshow test showed no goodness fit (χ 2 =13.472; p=0.019). In the validation group, the predicted pressure ulcer incidence also did not fit well with the observed incidence (χ 2 =42.154, p=0.000 by Braden scores; and χ 2 =17.223, p=0.001 by Braden scale risk classification). The Braden scale has low calibration power in predicting PU formation.
Validation of Metrics as Error Predictors

NASA Astrophysics Data System (ADS)

Mendling, Jan

In this chapter, we test the validity of metrics that were defined in the previous chapter for predicting errors in EPC business process models. In Section 5.1, we provide an overview of how the analysis data is generated. Section 5.2 describes the sample of EPCs from practice that we use for the analysis. Here we discuss a disaggregation by the EPC model group and by error as well as a correlation analysis between metrics and error. Based on this sample, we calculate a logistic regression model for predicting error probability with the metrics as input variables in Section 5.3. In Section 5.4, we then test the regression function for an independent sample of EPC models from textbooks as a cross-validation. Section 5.5 summarizes the findings.
Epidemiological characteristics and deaths of premature infants in a referral hospital for high-risk pregnancies

PubMed Central

de Freitas, Brunnella Alcantara Chagas; Sant'Ana, Luciana Ferreira da Rocha; Longo, Giana Zarbato; Siqueira-Batista, Rodrigo; Priore, Silvia Eloiza; Franceschin, Sylvia do Carmo Castro

2012-01-01

Objective To analyze the process of care provided to premature infants in a neonatal intensive care unit and the factors associated with their mortality. Methods Cross-sectional retrospective study of premature infants in an intensive care unit between 2008 and 2010. The characteristics of the mothers and premature infants were described, and a bivariate analysis was performed on the following characteristics: the study period and the "death" outcome (hospital, neonatal and early) using Pearson's chi-square test, Fisher's exact test or a chi-square test for linear trends. Bivariate and multivariable logistic regression analyses were performed using a stepwise backward logistic regression method between the variables with p<0.20 and the "death" outcome. A p value <0.05 was considered to be significant. Results In total, 293 preterm infants were studied. Increased access to complementary tests (transfontanellar ultrasound and Doppler echocardiogram) and breastfeeding rates were indicators of improving care. Mortality was concentrated in the neonatal period, especially in the early neonatal period, and was associated with extreme prematurity, small size for gestational age and an Apgar score <7 at 5 minutes after birth. The late-onset sepsis was also associated with a greater chance of neonatal death, and antenatal corticosteroids were protective against neonatal and early deaths. Conclusions Although these results are comparable to previous findings regarding mortality among premature infants in Brazil, the study emphasizes the need to implement strategies that promote breastfeeding and reduce neonatal mortality and its early component. PMID:23917938
The Mantel-Haenszel procedure revisited: models and generalizations.

PubMed

Fidler, Vaclav; Nagelkerke, Nico

2013-01-01

Several statistical methods have been developed for adjusting the Odds Ratio of the relation between two dichotomous variables X and Y for some confounders Z. With the exception of the Mantel-Haenszel method, commonly used methods, notably binary logistic regression, are not symmetrical in X and Y. The classical Mantel-Haenszel method however only works for confounders with a limited number of discrete strata, which limits its utility, and appears to have no basis in statistical models. Here we revisit the Mantel-Haenszel method and propose an extension to continuous and vector valued Z. The idea is to replace the observed cell entries in strata of the Mantel-Haenszel procedure by subject specific classification probabilities for the four possible values of (X,Y) predicted by a suitable statistical model. For situations where X and Y can be treated symmetrically we propose and explore the multinomial logistic model. Under the homogeneity hypothesis, which states that the odds ratio does not depend on Z, the logarithm of the odds ratio estimator can be expressed as a simple linear combination of three parameters of this model. Methods for testing the homogeneity hypothesis are proposed. The relationship between this method and binary logistic regression is explored. A numerical example using survey data is presented.
The Mantel-Haenszel Procedure Revisited: Models and Generalizations

PubMed Central

Fidler, Vaclav; Nagelkerke, Nico

2013-01-01

Several statistical methods have been developed for adjusting the Odds Ratio of the relation between two dichotomous variables X and Y for some confounders Z. With the exception of the Mantel-Haenszel method, commonly used methods, notably binary logistic regression, are not symmetrical in X and Y. The classical Mantel-Haenszel method however only works for confounders with a limited number of discrete strata, which limits its utility, and appears to have no basis in statistical models. Here we revisit the Mantel-Haenszel method and propose an extension to continuous and vector valued Z. The idea is to replace the observed cell entries in strata of the Mantel-Haenszel procedure by subject specific classification probabilities for the four possible values of (X,Y) predicted by a suitable statistical model. For situations where X and Y can be treated symmetrically we propose and explore the multinomial logistic model. Under the homogeneity hypothesis, which states that the odds ratio does not depend on Z, the logarithm of the odds ratio estimator can be expressed as a simple linear combination of three parameters of this model. Methods for testing the homogeneity hypothesis are proposed. The relationship between this method and binary logistic regression is explored. A numerical example using survey data is presented. PMID:23516463
Landslide susceptibility mapping using frequency ratio, logistic regression, artificial neural networks and their comparison: A case study from Kat landslides (Tokat—Turkey)

NASA Astrophysics Data System (ADS)

Yilmaz, Işık

2009-06-01

The purpose of this study is to compare the landslide susceptibility mapping methods of frequency ratio (FR), logistic regression and artificial neural networks (ANN) applied in the Kat County (Tokat—Turkey). Digital elevation model (DEM) was first constructed using GIS software. Landslide-related factors such as geology, faults, drainage system, topographical elevation, slope angle, slope aspect, topographic wetness index (TWI) and stream power index (SPI) were used in the landslide susceptibility analyses. Landslide susceptibility maps were produced from the frequency ratio, logistic regression and neural networks models, and they were then compared by means of their validations. The higher accuracies of the susceptibility maps for all three models were obtained from the comparison of the landslide susceptibility maps with the known landslide locations. However, respective area under curve (AUC) values of 0.826, 0.842 and 0.852 for frequency ratio, logistic regression and artificial neural networks showed that the map obtained from ANN model is more accurate than the other models, accuracies of all models can be evaluated relatively similar. The results obtained in this study also showed that the frequency ratio model can be used as a simple tool in assessment of landslide susceptibility when a sufficient number of data were obtained. Input process, calculations and output process are very simple and can be readily understood in the frequency ratio model, however logistic regression and neural networks require the conversion of data to ASCII or other formats. Moreover, it is also very hard to process the large amount of data in the statistical package.
A Comparison of Logistic Regression, Neural Networks, and Classification Trees Predicting Success of Actuarial Students

ERIC Educational Resources Information Center

Schumacher, Phyllis; Olinsky, Alan; Quinn, John; Smith, Richard

2010-01-01

The authors extended previous research by 2 of the authors who conducted a study designed to predict the successful completion of students enrolled in an actuarial program. They used logistic regression to determine the probability of an actuarial student graduating in the major or dropping out. They compared the results of this study with those…
Logistic regression accuracy across different spatial and temporal scales for a wide-ranging species, the marbled murrelet

Treesearch

Carolyn B. Meyer; Sherri L. Miller; C. John Ralph

2004-01-01

The scale at which habitat variables are measured affects the accuracy of resource selection functions in predicting animal use of sites. We used logistic regression models for a wide-ranging species, the marbled murrelet, (Brachyramphus marmoratus) in a large region in California to address how much changing the spatial or temporal scale of...
Odds Ratio, Delta, ETS Classification, and Standardization Measures of DIF Magnitude for Binary Logistic Regression

ERIC Educational Resources Information Center

Monahan, Patrick O.; McHorney, Colleen A.; Stump, Timothy E.; Perkins, Anthony J.

2007-01-01

Previous methodological and applied studies that used binary logistic regression (LR) for detection of differential item functioning (DIF) in dichotomously scored items either did not report an effect size or did not employ several useful measures of DIF magnitude derived from the LR model. Equations are provided for these effect size indices.…
Risk Factors of Falls in Community-Dwelling Older Adults: Logistic Regression Tree Analysis

ERIC Educational Resources Information Center

Yamashita, Takashi; Noe, Douglas A.; Bailer, A. John

2012-01-01

Purpose of the Study: A novel logistic regression tree-based method was applied to identify fall risk factors and possible interaction effects of those risk factors. Design and Methods: A nationally representative sample of American older adults aged 65 years and older (N = 9,592) in the Health and Retirement Study 2004 and 2006 modules was used.…
Estimation of Logistic Regression Models in Small Samples. A Simulation Study Using a Weakly Informative Default Prior Distribution

ERIC Educational Resources Information Center

Gordovil-Merino, Amalia; Guardia-Olmos, Joan; Pero-Cebollero, Maribel

2012-01-01

In this paper, we used simulations to compare the performance of classical and Bayesian estimations in logistic regression models using small samples. In the performed simulations, conditions were varied, including the type of relationship between independent and dependent variable values (i.e., unrelated and related values), the type of variable…
Using multiple logistic regression and GIS technology to predict landslide hazard in northeast Kansas, USA

USGS Publications Warehouse

Ohlmacher, G.C.; Davis, J.C.

2003-01-01

Landslides in the hilly terrain along the Kansas and Missouri rivers in northeastern Kansas have caused millions of dollars in property damage during the last decade. To address this problem, a statistical method called multiple logistic regression has been used to create a landslide-hazard map for Atchison, Kansas, and surrounding areas. Data included digitized geology, slopes, and landslides, manipulated using ArcView GIS. Logistic regression relates predictor variables to the occurrence or nonoccurrence of landslides within geographic cells and uses the relationship to produce a map showing the probability of future landslides, given local slopes and geologic units. Results indicated that slope is the most important variable for estimating landslide hazard in the study area. Geologic units consisting mostly of shale, siltstone, and sandstone were most susceptible to landslides. Soil type and aspect ratio were considered but excluded from the final analysis because these variables did not significantly add to the predictive power of the logistic regression. Soil types were highly correlated with the geologic units, and no significant relationships existed between landslides and slope aspect. ?? 2003 Elsevier Science B.V. All rights reserved.

Predicting risk for portal vein thrombosis in acute pancreatitis patients: A comparison of radical basis function artificial neural network and logistic regression models.

PubMed

Fei, Yang; Hu, Jian; Gao, Kun; Tu, Jianfeng; Li, Wei-Qin; Wang, Wei

2017-06-01

To construct a radical basis function (RBF) artificial neural networks (ANNs) model to predict the incidence of acute pancreatitis (AP)-induced portal vein thrombosis. The analysis included 353 patients with AP who had admitted between January 2011 and December 2015. RBF ANNs model and logistic regression model were constructed based on eleven factors relevant to AP respectively. Statistical indexes were used to evaluate the value of the prediction in two models. The predict sensitivity, specificity, positive predictive value, negative predictive value and accuracy by RBF ANNs model for PVT were 73.3%, 91.4%, 68.8%, 93.0% and 87.7%, respectively. There were significant differences between the RBF ANNs and logistic regression models in these parameters (P<0.05). In addition, a comparison of the area under receiver operating characteristic curves of the two models showed a statistically significant difference (P<0.05). The RBF ANNs model is more likely to predict the occurrence of PVT induced by AP than logistic regression model. D-dimer, AMY, Hct and PT were important prediction factors of approval for AP-induced PVT. Copyright © 2017 Elsevier Inc. All rights reserved.
Risk Factors for 30-Day Complications After Thumb CMC Joint Arthroplasty: An American College of Surgeons National Surgery Quality Improvement Program Study.

PubMed

Shah, Kalpit N; Defroda, Steven F; Wang, Bo; Weiss, Arnold-Peter C

2017-12-01

The first carpometacarpal (CMC) joint is a common site of osteoarthritis, with arthroplasty being a common procedure to provide pain relief and improve function with low complications. However, little is known about risk factors that may predispose a patient for postoperative complications. All CMC joint arthroplasty from 2005 to 2015 in the prospectively collected American College of Surgeons National Surgical Quality Improvement Program (ACS-NSQIP) database were identified. Bivariate testing and multiple logistic regressions were performed to determine which patient demographics, surgical variables and medical comorbidities were significant predictors for complications. These included wound related, cardiopulmonary, neurological and renal complications, return to the operating room (OR) and readmission. A total of 3344 patients were identified from the database. Of those, 45 patients (1.3%) experienced a complication including wound issues (0.66%), return to the OR (0.15%) and readmission (0.27%) amongst others. When performing bivariate analysis, age over 65, American Society of Anesthesiologists (ASA) Class, diabetes and renal dialysis were significant risk factors. Multiple logistic regression after adjusting for confounding factors demonstrated that insulin-dependent diabetes and ASA Class 4 had a strong trend while renal dialysis was a significant risk factor. CMC arthroplasty has a very low overall complication rate of 1.3% and wound complication rate of 0.66%. Diabetes requiring insulin and ASA Class 4 trended towards significance while renal dialysis was found to be a significant risk factors in logistic regression. This information may be useful for preoperative counseling and discussion with patients who have these risk factors.
EXpectation Propagation LOgistic REgRession (EXPLORER): Distributed Privacy-Preserving Online Model Learning

PubMed Central

Wang, Shuang; Jiang, Xiaoqian; Wu, Yuan; Cui, Lijuan; Cheng, Samuel; Ohno-Machado, Lucila

2013-01-01

We developed an EXpectation Propagation LOgistic REgRession (EXPLORER) model for distributed privacy-preserving online learning. The proposed framework provides a high level guarantee for protecting sensitive information, since the information exchanged between the server and the client is the encrypted posterior distribution of coefficients. Through experimental results, EXPLORER shows the same performance (e.g., discrimination, calibration, feature selection etc.) as the traditional frequentist Logistic Regression model, but provides more flexibility in model updating. That is, EXPLORER can be updated one point at a time rather than having to retrain the entire data set when new observations are recorded. The proposed EXPLORER supports asynchronized communication, which relieves the participants from coordinating with one another, and prevents service breakdown from the absence of participants or interrupted communications. PMID:23562651
Late HIV Testing in a Cohort of HIV-Infected Patients in Puerto Rico.

PubMed

Tossas-Milligan, Katherine Y; Hunter-Mellado, Robert F; Mayor, Angel M; Fernández-Santos, Diana M; Dworkin, Mark S

2015-09-01

Late HIV testing (LT), defined as receiving an AIDS diagnosis within a year of one's first positive HIV test, is associated with higher HIV transmission, lower HAART effectiveness, and worse outcomes. Latinos represent 36% of LT in the US, yet research concerning LT among HIV cases in Puerto Rico is scarce. Multivariable logistic regression analysis was used to identify factors associated with LT, and a Cochran‒Armitage test was used to determine LT trends in an HIV-infected cohort followed at a clinic in Puerto Rico specialized in the management and treatment of HIV. From 2000 to 2011, 47% of eligible patients were late testers, with lower median CD4 counts (54 vs. 420 cells/mm3) and higher median HIV viral load counts (253,680 vs. 23,700 copies/mL) than non-LT patients. LT prevalence decreased significantly, from 47% in 2000 to 37% in 2011. In a mutually adjusted logistic regression model, males, older age at enrollment and past history of IDU significantly increased LT odds, whereas having a history of amphetamine use decreased LT odds. When the data were stratified by mode of transmission, it became apparent that only the category men who have sex with men (MSM) saw a significant reduction in the proportion of LT, falling from 67% in 2000 to 33% in 2011. These results suggest a gap in early HIV detection in Puerto Rico, a gap that decreased only among MSM. An evaluation of the manner in which current HIV-testing guidelines are implemented on the island is needed.
[Predictors of hospitalization for alcohol use disorder in Korean men].

PubMed

Hong, Hae-Sook; Park, Jeong-Eun; Park, Wan-Ju

2014-10-01

This study was done to identify the patterns and significant predictors influencing hospitalization of Korean men for alcohol use disorder. A descriptive study design was utilized. Data were collected using self-report questionnaires from 143 inpatients who met the DSM-5 alcohol use disorder criteria and were receiving treatment and 157 social drinkers living in the community. The questionnaires included Alcohol Use Disorders Identification Test (AUDIT), Alcohol Problems, Alcohol Expectancy Questionnaire (AEQ), Life Position, and The Korean version of the Children of Alcoholics Screening Test (CAST-K). Data were analyzed using descriptive statistics, t-test, χ²-test, F-test, Pearson correlation coefficients, and logistic regression with forward stepwise. AUDIT had significant correlations with alcohol problems, alcohol expectancy, and parents' alcoholism. In logistic regression, factors significantly affecting hospitalization were divorced (OR=4.18, 95% CI: 1.28-13.71), graduation from elementary school (OR=28.50, 95% CI: 8.07-100.69), middle school (OR=6.66, 95% CI: 2.21-20.09), high school (OR=6.31, 95% CI: 2.59-15.36), drinking alone (OR=9.07, 95% CI: 1.78-46.17), family history of alcoholism (OR=2.41, 95% CI: 1.11-5.25), interpersonal relationship problems (OR=1.28, 95% CI:1.17-1.41), and sexual enhancement of alcohol expectancy (OR=0.83, 95% CI: 0.72-0.94), which accounted for 53% of the variance. Results suggest that interpersonal relationship programs and customized cognitive programs for social drinkers in the community are needed to decreased alcohol related hospitalization in Korean men.
Fatigue design of a cellular phone folder using regression model-based multi-objective optimization

NASA Astrophysics Data System (ADS)

Kim, Young Gyun; Lee, Jongsoo

2016-08-01

In a folding cellular phone, the folding device is repeatedly opened and closed by the user, which eventually results in fatigue damage, particularly to the front of the folder. Hence, it is important to improve the safety and endurance of the folder while also reducing its weight. This article presents an optimal design for the folder front that maximizes its fatigue endurance while minimizing its thickness. Design data for analysis and optimization were obtained experimentally using a test jig. Multi-objective optimization was carried out using a nonlinear regression model. Three regression methods were employed: back-propagation neural networks, logistic regression and support vector machines. The AdaBoost ensemble technique was also used to improve the approximation. Two-objective Pareto-optimal solutions were identified using the non-dominated sorting genetic algorithm (NSGA-II). Finally, a numerically optimized solution was validated against experimental product data, in terms of both fatigue endurance and thickness index.
Dietary consumption patterns and laryngeal cancer risk.

PubMed

Vlastarakos, Petros V; Vassileiou, Andrianna; Delicha, Evie; Kikidis, Dimitrios; Protopapas, Dimosthenis; Nikolopoulos, Thomas P

2016-06-01

We conducted a case-control study to investigate the effect of diet on laryngeal carcinogenesis. Our study population was made up of 140 participants-70 patients with laryngeal cancer (LC) and 70 controls with a non-neoplastic condition that was unrelated to diet, smoking, or alcohol. A food-frequency questionnaire determined the mean consumption of 113 different items during the 3 years prior to symptom onset. Total energy intake and cooking mode were also noted. The relative risk, odds ratio (OR), and 95% confidence interval (CI) were estimated by multiple logistic regression analysis. We found that the total energy intake was significantly higher in the LC group (p < 0.001), and that the difference remained statistically significant after logistic regression analysis (p < 0.001; OR: 118.70). Notably, meat consumption was higher in the LC group (p < 0.001), and the difference remained significant after logistic regression analysis (p = 0.029; OR: 1.16). LC patients also consumed significantly more fried food (p = 0.036); this difference also remained significant in the logistic regression model (p = 0.026; OR: 5.45). The LC group also consumed significantly more seafood (p = 0.012); the difference persisted after logistic regression analysis (p = 0.009; OR: 2.48), with the consumption of shrimp proving detrimental (p = 0.049; OR: 2.18). Finally, the intake of zinc was significantly higher in the LC group before and after logistic regression analysis (p = 0.034 and p = 0.011; OR: 30.15, respectively). Cereal consumption (including pastas) was also higher among the LC patients (p = 0.043), with logistic regression analysis showing that their negative effect was possibly associated with the sauces and dressings that traditionally accompany pasta dishes (p = 0.006; OR: 4.78). Conversely, a higher consumption of dairy products was found in controls (p < 0.05); logistic regression analysis showed that calcium appeared to be protective at the micronutrient level (p < 0.001; OR: 0.27). We found no difference in the overall consumption of fruits and vegetables between the LC patients and controls; however, the LC patients did have a greater consumption of cooked tomatoes and cooked root vegetables (p = 0.039 for both), and the controls had more consumption of leeks (p = 0.042) and, among controls younger than 65 years, cooked beans (p = 0.037). Lemon (p = 0.037), squeezed fruit juice (p = 0.032), and watermelon (p = 0.018) were also more frequently consumed by the controls. Other differences at the micronutrient level included greater consumption by the LC patients of retinol (p = 0.044), polyunsaturated fats (p = 0.041), and linoleic acid (p = 0.008); LC patients younger than 65 years also had greater intake of riboflavin (p = 0.045). We conclude that the differences in dietary consumption patterns between LC patients and controls indicate a possible role for lifestyle modifications involving nutritional factors as a means of decreasing the risk of laryngeal cancer.
Mother-Son Communication About Sex and Routine Human Immunodeficiency Virus Testing Among Younger Men of Color Who Have Sex With Men.

PubMed

Bouris, Alida; Hill, Brandon J; Fisher, Kimberly; Erickson, Greg; Schneider, John A

2015-11-01

The purposes of this study were to document the HIV testing behaviors and serostatus of younger men of color who have sex with men (YMSM) and to explore sociodemographic, behavioral, and maternal correlates of HIV testing in the past 6 months. A total of 135 YMSM aged 16-19 years completed a close-ended survey on HIV testing and risk behaviors, mother-son communication, and sociodemographic characteristics. Youth were offered point-of-care HIV testing, with results provided at survey end. Multivariate logistic regression analyzed the sociodemographic, behavioral, and maternal factors associated with routine HIV testing. A total of 90.3% of YMSM had previously tested for HIV, and 70.9% had tested in the past 6 months. In total, 11.7% of youth reported being HIV positive, and 3.3% reported unknown serostatus. When offered an HIV test, 97.8% accepted. Of these, 14.7% had a positive oral test result, and 31.58% of HIV-positive YMSM (n = 6) were seropositive unaware. Logistic regression results indicated that maternal communication about sex with males was positively associated with routine testing (odds ratio = 2.36; 95% confidence interval = 1.13-4.94). Conversely, communication about puberty and general human sexuality was negatively associated (odds ratio = .45; 95% confidence interval = .24-.86). Condomless anal intercourse and positive sexually transmitted infection history were negatively associated with routine testing; however, frequency of alcohol use was positively associated. Despite high rates of testing, we found high rates of HIV infection, with 31.58% of HIV-positive YMSM being seropositive unaware. Mother-son communication about sex needs to address same-sex behavior as this appears to be more important than other topics. YMSM with known risk factors for HIV are not testing at the recommended time intervals. Copyright © 2015 Society for Adolescent Health and Medicine. Published by Elsevier Inc. All rights reserved.
Genetic Modeling of Radiation Injury in Prostate Cancer Patients Treated with Radiotherapy

DTIC Science & Technology

2017-10-01

approaches in the GWAS meta-analysis: 1) logistic regression to test association of each SNP with grade 1 or worse toxicity at 2 years post ...Annual PREPARED FOR: U.S. Army Medical Research and Materiel Command Fort Detrick, Maryland 21702-5012 DISTRIBUTION STATEMENT: Approved for...Army Medical Research and Materiel Command Fort Detrick, Maryland 21702-5012 11. SPONSOR/MONITOR’S REPORT NUMBER(S) 12. DISTRIBUTION / AVAILABILITY
Association between kyphosis and subacromial impingement syndrome: LOHAS study.

PubMed

Otoshi, Kenichi; Takegami, Misa; Sekiguchi, Miho; Onishi, Yoshihiro; Yamazaki, Shin; Otani, Koji; Shishido, Hiroaki; Kikuchi, Shinichi; Konno, Shinichi

2014-12-01

Kyphosis is a cause of scapular dyskinesis, which can induce various shoulder disorders, including subacromial impingement syndrome (SIS). This study aimed to clarify the impact of kyphosis on SIS with use of cross-sectional data from the Locomotive Syndrome and Health Outcome in Aizu Cohort Study (LOHAS). The study enrolled 2144 participants who were older than 40 years and participated in health checkups in 2010. Kyphosis was assessed by the wall-occiput test (WOT) for thoracic kyphosis and the rib-pelvic distance test (RPDT) for lumbar kyphosis. The associations between kyphosis, SIS, and reduction in shoulder elevation (RSE) were investigated. Age- and gender-adjusted logistic regression analysis demonstrated significant association between SIS and WOT (odds ratio, 1.65; 95% confidence interval, 1.02, 2.64; P < .05), whereas there was no significant association between SIS and RPDT. Multivariable logistic regression analysis demonstrated no significant association between SIS and both WOT and RPDT, whereas there was significant association between SIS and RSE. RSE plays a key role in the development of SIS, and thoracic kyphosis might influence the development of SIS indirectly by reducing shoulder elevation induced by the restriction of the thoracic spine extension and scapular dyskinesis. Copyright © 2014 Journal of Shoulder and Elbow Surgery Board of Trustees. Published by Elsevier Inc. All rights reserved.
Ventilator-associated pneumonia: the influence of bacterial resistance, prescription errors, and de-escalation of antimicrobial therapy on mortality rates.

PubMed

Souza-Oliveira, Ana Carolina; Cunha, Thúlio Marquez; Passos, Liliane Barbosa da Silva; Lopes, Gustavo Camargo; Gomes, Fabiola Alves; Röder, Denise Von Dolinger de Brito

2016-01-01

Ventilator-associated pneumonia is the most prevalent nosocomial infection in intensive care units and is associated with high mortality rates (14-70%). This study evaluated factors influencing mortality of patients with Ventilator-associated pneumonia (VAP), including bacterial resistance, prescription errors, and de-escalation of antibiotic therapy. This retrospective study included 120 cases of Ventilator-associated pneumonia admitted to the adult adult intensive care unit of the Federal University of Uberlândia. The chi-square test was used to compare qualitative variables. Student's t-test was used for quantitative variables and multiple logistic regression analysis to identify independent predictors of mortality. De-escalation of antibiotic therapy and resistant bacteria did not influence mortality. Mortality was 4 times and 3 times higher, respectively, in patients who received an inappropriate antibiotic loading dose and in patients whose antibiotic dose was not adjusted for renal function. Multiple logistic regression analysis revealed the incorrect adjustment for renal function was the only independent factor associated with increased mortality. Prescription errors influenced mortality of patients with Ventilator-associated pneumonia, underscoring the challenge of proper Ventilator-associated pneumonia treatment, which requires continuous reevaluation to ensure that clinical response to therapy meets expectations. Copyright © 2016. Published by Elsevier Editora Ltda.
Risk factors for the breakdown of perineal laceration repair after vaginal delivery.

PubMed

Williams, Meredith K; Chames, Mark C

2006-09-01

The purpose of this study was to identify risk factors that are associated with the breakdown of perineal laceration repair in the postpartum period. We conducted a retrospective, case-control study to review perineal laceration repair breakdown in patients who were delivered between September 1995 and February 2005 at the University of Michigan. Bivariate analysis with chi-square test and t-test and stepwise logistic regression analysis were performed. Fifty-nine cases and 118 control deliveries were identified from a total of 14,124 vaginal deliveries. Risk factors were longer second stage of labor (142 vs 87 minutes; P = .001), operative vaginal delivery (odds ratio, 3.6; 95% CI, 1.8-7.3), mediolateral episiotomy (odds ratio, 6.9; 95% CI, 2.6-18.7), third- or fourth-degree laceration (odds ratio, 3.1; 95% CI, 1.5-6.4), and meconium-stained amniotic fluid (odds ratio, 3.0; 95% CI, 1.1-7.9). Previous vaginal delivery was protective (odds ratio, 0.38; 95% CI, 0.18-0.84). Logistic regression showed the most significant factor to be an interaction between operative vaginal delivery and mediolateral episiotomy (odd ratio, 6.36; 95% CI, 2.18-18.57). The most significant events were mediolateral episiotomy, especially in conjunction with operative vaginal delivery, third- and fourth-degree lacerations, and meconium.
Robust experimental design for optimizing the microbial inhibitor test for penicillin detection in milk.

PubMed

Nagel, O G; Molina, M P; Basílico, J C; Zapata, M L; Althaus, R L

2009-06-01

To use experimental design techniques and a multiple logistic regression model to optimize a microbiological inhibition test with dichotomous response for the detection of Penicillin G in milk. A 2(3) x 2(2) robust experimental design with two replications was used. The effects of three control factors (V: culture medium volume, S: spore concentration of Geobacillus stearothermophilus, I: indicator concentration), two noise factors (Dt: diffusion time, Ip: incubation period) and their interactions were studied. The V, S, Dt, Ip factors and V x S, V x Ip, S x Ip interactions showed significant effects. The use of 100 microl culture medium volume, 2 x 10(5) spores ml(-1), 60 min diffusion time and 3 h incubation period is recommended. In these elaboration conditions, the penicillin detection limit was of 3.9 microg l(-1), similar to the maximum residue limit (MRL). Of the two noise factors studied, the incubation period can be controlled by means of the culture medium volume and spore concentration. We were able to optimize bioassays of dichotomous response using an experimental design and logistic regression model for the detection of residues at the level of MRL, aiding in the avoidance of health problems in the consumer.
Constructive thinking, rational intelligence and irritable bowel syndrome.

PubMed

Rey, Enrique; Moreno Ortega, Marta; Garcia Alonso, Monica-Olga; Diaz-Rubio, Manuel

2009-07-07

To evaluate rational and experiential intelligence in irritable bowel syndrome (IBS) sufferers. We recruited 100 subjects with IBS as per Rome II criteria (50 consulters and 50 non-consulters) and 100 healthy controls, matched by age, sex and educational level. Cases and controls completed a clinical questionnaire (including symptom characteristics and medical consultation) and the following tests: rational-intelligence (Wechsler Adult Intelligence Scale, 3rd edition); experiential-intelligence (Constructive Thinking Inventory); personality (NEO personality inventory); psychopathology (MMPI-2), anxiety (state-trait anxiety inventory) and life events (social readjustment rating scale). Analysis of variance was used to compare the test results of IBS-sufferers and controls, and a logistic regression model was then constructed and adjusted for age, sex and educational level to evaluate any possible association with IBS. No differences were found between IBS cases and controls in terms of IQ (102.0 +/- 10.8 vs 102.8 +/- 12.6), but IBS sufferers scored significantly lower in global constructive thinking (43.7 +/- 9.4 vs 49.6 +/- 9.7). In the logistic regression model, global constructive thinking score was independently linked to suffering from IBS [OR 0.92 (0.87-0.97)], without significant OR for total IQ. IBS subjects do not show lower rational intelligence than controls, but lower experiential intelligence is nevertheless associated with IBS.
Subjective vs objective evaluations of smile esthetics.

PubMed

Schabel, Brian J; Franchi, Lorenzo; Baccetti, Tiziano; McNamara, James A

2009-04-01

The aim of this study was to analyze the relationships between subjective evaluations of posttreatment smiles captured with clinical photography and rated by a panel of orthodontists and parents of orthodontic patients, and objective evaluations of the same smiles from the Smile Mesh program (TDG Computing, Philadelphia, Pa). The clinical photographs of 48 orthodontically treated patients were rated by a panel of 25 experienced orthodontists and 20 parents of patients. Independent samples t tests were used to test whether objective measurements were significantly different between subjects with "attractive" and "unattractive" smiles, and those with the "most attractive" and "least attractive" smiles. Additionally, logistic regression was performed to evaluate whether the measurements could predict whether a smile captured with clinical photography would be attractive or unattractive. The comparison between groups showed no significant differences for any measurement. Subjects with the "most unattractive" smiles had a significantly greater distance between the incisal edge of the maxillary central incisors and the lower lip during smiling, and a significantly smaller smile index than did those with the "most attractive" smiles. As shown by the coefficients of logistic regression, smile attractiveness could not be predicted by any objectively gathered measurement. No objective measure of the smile could predict attractive or unattractive smiles as judged subjectively.
[Diabetic Foot Neuropathy and Related Factors in Patients With Type 2 Diabetes Mellitus].

PubMed

Chen, Tzu-Yu; Lin, Chia-Huei; Chang, Yue-Cune; Wang, Chih-Hsin; Hung, Yi-Jen; Tzeng, Wen-Chii

2018-06-01

Patients with type 2 diabetes mellitus (T2DM) face a higher risk of diabetic foot neuropathy, which increases the risk of death. The early detection of factors that influence diabetic neuropathy reduces the risk of foot lesions, including foot ulcerations, lower extremity amputation, and mortality. To explore the demographic, disease-characteristic, health-literacy, and foot-self-care-behavior factors that affect diabetic foot neuropathy in patients with T2DM. A case-control study design was employed in which cases (Michigan Neuropathy Screening Instrument, MNSI) ≥ 2 were matched to controls based on age and gender in a medical center. A total of 114 patients diagnosed with T2DM in a medical center were recruited as participants. Data were collected using a structured questionnaire. The collected data were analyzed using Fisher's exact test, Mann-Whitney U test, and logistic regression. The results of multiple logistic regression showed that glycated hemoglobin (B = 1.696, p = .041) and communication and critical health literacy (B = -0.082, p = .034) were significant factors of diabetic foot neuropathy. The findings of this study suggest that nurses should assess the health literacy of patients with T2DM before providing health education and should develop a specific foot-care intervention for individuals with poor glycemic control.
Plasma Homocysteine and Asymmetrical Dimethyl-l-Arginine (ADMA) and Whole Blood DNA Methylation in Early and Neovascular Age-Related Macular Degeneration: A Pilot Study.

PubMed

Pinna, Antonio; Zinellu, Angelo; Tendas, Donatella; Blasetti, Francesco; Carru, Ciriaco; Castiglia, Paolo

2016-01-01

To compare the plasma levels of homocysteine and asymmetrical dimethyl-l-arginine (ADMA) and the degree of whole blood DNA methylation in patients with early and neovascular age-related macular degeneration (AMD) and in controls without maculopathy of any sort. This observational case-control pilot study included 39 early AMD patients, 27 neovascular AMD patients and 132 sex- and age-matched controls without maculopathy. Plasma homocysteine and ADMA concentrations and the degree of whole blood DNA methylation were measured. Quantitative variables were compared by Student's t-test or Mann-Whitney test. Logistic regression models were used to investigate the significance of the association between early or wet AMD and some variables. There were no significant differences in mean plasma homocysteine and ADMA concentrations and in the degree of whole blood DNA methylation between patients with early or neovascular AMD and their controls. Similarly, logistic regression analysis disclosed that plasma homocysteine and ADMA levels were not associated with an increased risk for early or neovascular AMD. We failed to demonstrate an association between early or neovascular AMD and increased plasma homocysteine and/or ADMA. Results also suggest that the degree of whole blood DNA methylation is not a marker of AMD.
Prediction of outcome in internet-delivered cognitive behaviour therapy for paediatric obsessive-compulsive disorder: A machine learning approach.

PubMed

Lenhard, Fabian; Sauer, Sebastian; Andersson, Erik; Månsson, Kristoffer Nt; Mataix-Cols, David; Rück, Christian; Serlachius, Eva

2018-03-01

There are no consistent predictors of treatment outcome in paediatric obsessive-compulsive disorder (OCD). One reason for this might be the use of suboptimal statistical methodology. Machine learning is an approach to efficiently analyse complex data. Machine learning has been widely used within other fields, but has rarely been tested in the prediction of paediatric mental health treatment outcomes. To test four different machine learning methods in the prediction of treatment response in a sample of paediatric OCD patients who had received Internet-delivered cognitive behaviour therapy (ICBT). Participants were 61 adolescents (12-17 years) who enrolled in a randomized controlled trial and received ICBT. All clinical baseline variables were used to predict strictly defined treatment response status three months after ICBT. Four machine learning algorithms were implemented. For comparison, we also employed a traditional logistic regression approach. Multivariate logistic regression could not detect any significant predictors. In contrast, all four machine learning algorithms performed well in the prediction of treatment response, with 75 to 83% accuracy. The results suggest that machine learning algorithms can successfully be applied to predict paediatric OCD treatment outcome. Validation studies and studies in other disorders are warranted. Copyright © 2017 John Wiley & Sons, Ltd.
A Comparison of the Logistic Regression and Contingency Table Methods for Simultaneous Detection of Uniform and Nonuniform DIF

ERIC Educational Resources Information Center

Guler, Nese; Penfield, Randall D.

2009-01-01

In this study, we investigate the logistic regression (LR), Mantel-Haenszel (MH), and Breslow-Day (BD) procedures for the simultaneous detection of both uniform and nonuniform differential item functioning (DIF). A simulation study was used to assess and compare the Type I error rate and power of a combined decision rule (CDR), which assesses DIF…
The Overall Odds Ratio as an Intuitive Effect Size Index for Multiple Logistic Regression: Examination of Further Refinements

ERIC Educational Resources Information Center

Le, Huy; Marcus, Justin

2012-01-01

This study used Monte Carlo simulation to examine the properties of the overall odds ratio (OOR), which was recently introduced as an index for overall effect size in multiple logistic regression. It was found that the OOR was relatively independent of study base rate and performed better than most commonly used R-square analogs in indexing model…

Using ROC curves to compare neural networks and logistic regression for modeling individual noncatastrophic tree mortality

Treesearch

Susan L. King

2003-01-01

The performance of two classifiers, logistic regression and neural networks, are compared for modeling noncatastrophic individual tree mortality for 21 species of trees in West Virginia. The output of the classifier is usually a continuous number between 0 and 1. A threshold is selected between 0 and 1 and all of the trees below the threshold are classified as...
Logistic regression trees for initial selection of interesting loci in case-control studies

PubMed Central

Nickolov, Radoslav Z; Milanov, Valentin B

2007-01-01

Modern genetic epidemiology faces the challenge of dealing with hundreds of thousands of genetic markers. The selection of a small initial subset of interesting markers for further investigation can greatly facilitate genetic studies. In this contribution we suggest the use of a logistic regression tree algorithm known as logistic tree with unbiased selection. Using the simulated data provided for Genetic Analysis Workshop 15, we show how this algorithm, with incorporation of multifactor dimensionality reduction method, can reduce an initial large pool of markers to a small set that includes the interesting markers with high probability. PMID:18466557
Using Logistic Regression to Predict the Probability of Debris Flows in Areas Burned by Wildfires, Southern California, 2003-2006

USGS Publications Warehouse

Rupert, Michael G.; Cannon, Susan H.; Gartner, Joseph E.; Michael, John A.; Helsel, Dennis R.

2008-01-01

Logistic regression was used to develop statistical models that can be used to predict the probability of debris flows in areas recently burned by wildfires by using data from 14 wildfires that burned in southern California during 2003-2006. Twenty-eight independent variables describing the basin morphology, burn severity, rainfall, and soil properties of 306 drainage basins located within those burned areas were evaluated. The models were developed as follows: (1) Basins that did and did not produce debris flows soon after the 2003 to 2006 fires were delineated from data in the National Elevation Dataset using a geographic information system; (2) Data describing the basin morphology, burn severity, rainfall, and soil properties were compiled for each basin. These data were then input to a statistics software package for analysis using logistic regression; and (3) Relations between the occurrence or absence of debris flows and the basin morphology, burn severity, rainfall, and soil properties were evaluated, and five multivariate logistic regression models were constructed. All possible combinations of independent variables were evaluated to determine which combinations produced the most effective models, and the multivariate models that best predicted the occurrence of debris flows were identified. Percentage of high burn severity and 3-hour peak rainfall intensity were significant variables in all models. Soil organic matter content and soil clay content were significant variables in all models except Model 5. Soil slope was a significant variable in all models except Model 4. The most suitable model can be selected from these five models on the basis of the availability of independent variables in the particular area of interest and field checking of probability maps. The multivariate logistic regression models can be entered into a geographic information system, and maps showing the probability of debris flows can be constructed in recently burned areas of southern California. This study demonstrates that logistic regression is a valuable tool for developing models that predict the probability of debris flows occurring in recently burned landscapes.
A Cross-border Comparison of Hepatitis B Testing Among Chinese Residing in Canada and the United States

PubMed Central

Tu, Shin-Ping; Li, Lin; Tsai, Jenny Hsin-Chun; Yip, Mei-Po; Terasaki, Genji; Teh, Chong; Yasui, Yutaka; Hislop, T Gregory; Taylor, Vicky

2013-01-01

Background The Western Pacific region has the highest level of endemic hepatitis B virus (HBV) infection in the world, with the Chinese representing nearly one-third of infected persons globally. HBV carriers are potentially infectious to others and have an increased risk of chronic active hepatitis, cirrhosis, and hepatocellular carcinoma. Studies from the U.S. and Canada demonstrate that immigrants, particularly from Asia, are disproportionately affected by liver cancer. Purpose Given the different health care systems in Seattle and Vancouver, two geographically proximate cities, we examined HBV testing levels and factors associated with testing among Chinese residents of these cities. Methods We surveyed Chinese living in areas of Seattle and Vancouver with relatively high proportions of Chinese residents. In-person interviews were conducted in Cantonese, Mandarin, or English. Our bivariate analyses consisted of the chi-square test, with Fisher’s Exact test as necessary. We then performed unconditional logistic regression, first examining only the city effect as the sole explanatory variable of the model, then assessing the adjusted city effect in a final main-effects model that was constructed through backward selection to select statistically significant variables at alpha = 0.05. Results Survey cooperation rates for Seattle and Vancouver were 58% and 59%, respectively. In Seattle, 48% reported HBV testing, whereas in Vancouver, 55% reported testing. HBV testing in Seattle was lower than in Vancouver, with a crude odds ratio of 0.73 (95% CI = 0.56, 0.94). However after adjusting for demographic, health care access, knowledge, and social support variables, we found no significant differences in HBV testing between the two cities. In our logistic regression model, the odds of HBV testing were greatest when the doctor recommended the test, followed by when the employer asked for the test. Discussion Findings from this study support the need for additional research to examine the effectiveness of clinic-based and workplace interventions to promote HBV testing among immigrants to North America. PMID:19640196
A cross-border comparison of hepatitis B testing among chinese residing in Canada and the United States.

PubMed

Tu R, Shin-Ping; Li, Lin; Tsai, Jenny Hsin-Chun; Yip, Mei-Po; Terasaki, Genji; Teh, Chong; Yasui, Yutaka; Hislop, T Gregory; Taylor, Vicky

2009-01-01

The Western Pacific region has the highest level of endemic hepatitis B virus (HBV) infection in the world, with the Chinese representing nearly one-third of infected persons globally. HBV carriers are potentially infectious to others and have an increased risk of chronic active hepatitis, cirrhosis, and hepatocellular carcinoma. Studies from the U.S. and Canada demonstrate that immigrants, particularly from Asia, are disproportionately affected by liver cancer. Given the different health care systems in Seattle and Vancouver, two geographically proximate cities, we examined HBV testing levels and factors associated with testing among Chinese residents of these cities. We surveyed Chinese living in areas of Seattle and Vancouver with relatively high proportions of Chinese residents. In-person interviews were conducted in Cantonese, Mandarin, or English. Our bivariate analyses consisted of the chi-square test, with Fisher's Exact test as necessary. We then performed unconditional logistic regression, first examining only the city effect as the sole explanatory variable of the model, then assessing the adjusted city effect in a final main-effects model that was constructed through backward selection to select statistically significant variables at alpha=0.05. Survey cooperation rates for Seattle and Vancouver were 58% and 59%, respectively. In Seattle, 48% reported HBV testing, whereas in Vancouver, 55% reported testing. HBV testing in Seattle was lower than in Vancouver, with a crude odds ratio of 0.73 (95% CI = 0.56, 0.94). However after adjusting for demographic, health care access, knowledge, and social support variables, we found no significant differences in HBV testing between the two cities. In our logistic regression model, the odds of HBV testing were greatest when the doctor recommended the test, followed by when the employer asked for the test. Findings from this study support the need for additional research to examine the effectiveness of clinic-based and workplace interventions to promote HBV testing among immigrants to North America.
[Acceptability of the opportunistic search for human immunodeficiency virus infection by serology in patients recruited in Primary Care Centres in Spain].

PubMed

Puentes Torres, Rafael Carlos; Aguado Taberné, Cristina; Pérula de Torres, Luis Angel; Espejo Espejo, José; Castro Fernández, Cristina; Fransi Galiana, Luís

2016-01-01

To assess the acceptability of opportunistic search for human immunodeficiency virus (HIV). Cross-sectional, observational study. Primary Care Centres (PCC) of the Spanish National Health Care System. patients aged 18 to 65 years who had never been tested for HIV, and were having a blood test for other reasons. RECORDED VARIABLES: age, gender, stable partner, educational level, tobacco/alcohol use, reason for blood testing, acceptability of taking the HIV test, reasons for refusing to take the HIV test, and reasons for not having taken an HIV test previously. A descriptive, bivariate, multivariate (logistic regression) statistical analysis was performed. A total of 208 general practitioners (GPs) from 150 health care centres recruited 3,314 patients. Most (93.1%) of patients agreed to take the HIV test (95%CI: 92.2-93.9). Of these patients, 56.9% reported never having had an HIV test before because they considered not to be at risk of infection, whereas 34.8% reported never having been tested for HIV because their doctor had never offered it to them. Of the 6.9% who refused to take the HIV test, 73.9% considered that they were not at risk. According to the logistic regression analysis, acceptability was positively associated to age (higher among between 26 and 35 year olds, OR=1.79; 95%CI: 1.10-2.91) and non-smokers (OR=1.39; 95%CI: 1.01-1.93). Those living in towns with between 10,000 and 50,000 inhabitants showed less acceptance to the test (OR=0.57; 95%CI: 0.40-0.80). The HIV prevalence detected was 0.24% Acceptability of HIV testing is very high among patients having a blood test in primary care settings in Spain. Opportunistic search is cost-effective. Copyright © 2015 Elsevier España, S.L.U. All rights reserved.
Demographic variations in HIV testing history among emergency department patients: implications for HIV screening in US emergency departments

PubMed Central

Merchant, Roland C; Catanzaro, Bethany M; Seage, George R; Mayer, Kenneth H; Clark, Melissa A; DeGruttola, Victor G; Becker, Bruce M

2011-01-01

Objective To determine the proportion of emergency department (ED) patients who have been tested for human immunodeficiency virus (HIV) infection and assess if patient history of HIV testing varies according to patient demographic characteristics. Design From July 2005–July 2006, a random sample of 18–55-year-old English-speaking patients being treated for sub-critical injury or illness at a northeastern US ED were interviewed on their history of HIV testing. Logistic regression models were created to compare patients by their history of being tested for HIV according to their demography. Odds ratios (ORs) with 95% confidence intervals (CIs) were estimated. Results Of 2107 patients surveyed who were not known to be HIV-infected, the median age was 32 years; 54% were male, 71% were white, and 45% were single/never married; 49% had private health-care insurance and 45% had never been tested for HIV. Of the 946 never previously tested for HIV, 56.1% did not consider themselves at risk for HIV. In multivariable logistic regression analyses, those less likely to have been HIV tested were male (OR: 1.32 [1.37–2.73]), white (OR: 1.93 [1.37–2.73]), married (OR: 1.53 [1.12–2.08]), and had private health-care insurance (OR: 2.10 [1.69–2.61]). There was a U-shaped relationship between age and history of being tested for HIV; younger and older patients were less likely to have been tested. History of HIV testing and years of formal education were not related. Conclusion Almost half of ED patients surveyed had never been tested for HIV. Certain demographic groups are being missed though HIV diagnostic testing and screening programmes in other settings. These groups could potentially be reached through universal screening. PMID:19564517
Relationship of lead, mercury, mirex, dichlorodiphenyldichloroethylene, hexachlorobenzene, and polychlorinated biphenyls to timing of menarche among Akwesasne Mohawk girls.

PubMed

Denham, Melinda; Schell, Lawrence M; Deane, Glenn; Gallo, Mia V; Ravenscroft, Julia; DeCaprio, Anthony P

2005-02-01

Children are commonly exposed at background levels to several ubiquitous environmental pollutants, such as lead and persistent organic pollutants, that have been linked to neurologic and endocrine effects. These effects have prompted concern about alterations in human reproductive development. Few studies have examined the effects of these toxicants on human sexual maturation at levels commonly found in the general population, and none has been able to examine multiple toxicant exposures. The aim of the current investigation was to examine the relationship between attainment of menarche and levels of 6 environmental pollutants to which children are commonly exposed at low levels, ie, dichlorodiphenyldichloroethylene (p,p'-DDE), hexachlorobenzene (HCB), polychlorinated biphenyls (PCBs), mirex, lead, and mercury. This study was conducted with residents of the Akwesasne Mohawk Nation, a sovereign territory that spans the St Lawrence River and the boundaries of New York State and Ontario and Quebec, Canada. Since the 1950s, the St Lawrence River has been a site of substantial industrial development, and the Nation is currently adjacent to a US National Priority Superfund site. PCB, p,p'-DDE, HCB, and mirex levels exceeding the US Food and Drug Administration recommended tolerance limits for human consumption have been found in local animal species. The present analysis included 138 Akwesasne Mohawk Nation girls 10 to 16.9 years of age. Blood samples and sociodemographic data were collected by Akwesasne community members, without prior knowledge of participants' exposure status. Attainment of menses (menarche) was assessed as present or absent at the time of the interview. Congener-specific PCB analysis was available, and all 16 PCB congeners detected in >50% of the sample were included in analyses (International Union of Pure and Applied Chemistry numbers 52, 70, 74, 84, 87, 95, 99, 101 [+90], 105, 110, 118, 138 [+163 and 164], 149 [+123], 153, 180, and 187). Probit analysis was used to determine the median age at menarche for the sample. Binary logistic regression analysis was used to determine predictors of menarcheal status. Six toxicants (p,p'-DDE, HCB, PCBs, mirex, lead, and mercury) were entered into the logistic regression model. Age, socioeconomic status (SES), and BMI were tested as potential cofounders and were included in the model at P < .05. Interactions among toxicants were also evaluated. Toxicant levels were measured in blood for this sample and were consistent with long-term exposure to a variety of toxicants in multiple media. Mercury levels were at or below background levels, all lead levels were well below the Centers for Disease Control and Prevention action limit of 10 microg/dL, and PCB levels were consistent with a cumulative, continuing exposure pattern. The median age at menarche for the total sample was 12.2 years. The predicted age at menarche for girls with lead levels above the median (1.2 microg/dL) was 10.5 months later than that for girls with lead levels below the median. In the logistic regression analysis, age was the strongest predictor of menarcheal status and SES was also a significant predictor but BMI was not. The logistic regression analysis that corrected for age, SES, and other pollutants (p,p'-DDE, HCB, mirex, and mercury) indicated that, at their respective geometric means, lead (geometric mean: 0.49 microg/dL) was associated with a significantly lower probability of having reached menarche (beta = -1.29) and a group of 4 potentially estrogenic PCB congeners (E-PCB) (geometric mean: 0.12 ppb; International Union of Pure and Applied Chemistry numbers 52, 70, 101 [+90], and 187) was associated with a significantly greater probability of having reached menarche (beta = 2.13). Predicted probabilities at different levels of lead and PCBs were calculated on the basis of the logistic regression model. At the respective means of all toxicants and SES, 69% of 12-year-old girls were predicted to have reached menarche. However, at the 75th percentile of lead levels, only 10% of 12-year-old Mohawk girls were predicted to have reached menarche; at the 75th percentile of E-PCB levels, 86% of 12-year-old Mohawk girls were predicted to have reached menarche. No association was observed between mirex, p,p'-DDE, or HCB and menarcheal status. Although BMI was not a significant predictor, we tested BMI in the logistic regression model; it had little effect on the relationships between menarcheal status and either lead or E-PCB. In models testing toxicant interactions, age, SES, lead levels, and PCB levels continued to be significant predictors of menarcheal status. When each toxicant was tested in a logistic regression model correcting only for age and SES, we observed little change in the effects of lead or E-PCB on menarcheal status. The analysis of multichemical exposure among Akwesasne Mohawk Nation adolescent girls suggests that the attainment of menarche may be sensitive to relatively low levels of lead and certain PCB congeners. This study is distinguished by the ability to test many toxicants simultaneously and thus to exclude effects from unmeasured but coexisting exposures. By testing several PCB congener groupings, we were able to determine that specifically a group of potentially estrogenic PCB congeners affected the odds of reaching menarche. The lead and PCB findings are consistent with the literature and are biologically plausible. The sample size, cross-sectional study design, and possible occurrence of confounders beyond those tested suggest that results should be interpreted cautiously. Additional investigation to determine whether such low toxicant levels may affect reproduction and disorders of the reproductive system is warranted.
Logistic Regression Likelihood Ratio Test Analysis for Detecting Signals of Adverse Events in Post-market Safety Surveillance.

PubMed

Nam, Kijoeng; Henderson, Nicholas C; Rohan, Patricia; Woo, Emily Jane; Russek-Cohen, Estelle

2017-01-01

The Vaccine Adverse Event Reporting System (VAERS) and other product surveillance systems compile reports of product-associated adverse events (AEs), and these reports may include a wide range of information including age, gender, and concomitant vaccines. Controlling for possible confounding variables such as these is an important task when utilizing surveillance systems to monitor post-market product safety. A common method for handling possible confounders is to compare observed product-AE combinations with adjusted baseline frequencies where the adjustments are made by stratifying on observable characteristics. Though approaches such as these have proven to be useful, in this article we propose a more flexible logistic regression approach which allows for covariates of all types rather than relying solely on stratification. Indeed, a main advantage of our approach is that the general regression framework provides flexibility to incorporate additional information such as demographic factors and concomitant vaccines. As part of our covariate-adjusted method, we outline a procedure for signal detection that accounts for multiple comparisons and controls the overall Type 1 error rate. To demonstrate the effectiveness of our approach, we illustrate our method with an example involving febrile convulsion, and we further evaluate its performance in a series of simulation studies.
[Prediction of histological liver damage in asymptomatic alcoholic patients by means of clinical and laboratory data].

PubMed

Iturriaga, H; Hirsch, S; Bunout, D; Díaz, M; Kelly, M; Silva, G; de la Maza, M P; Petermann, M; Ugarte, G

1993-04-01

Looking for a noninvasive method to predict liver histologic alterations in alcoholic patients without clinical signs of liver failure, we studied 187 chronic alcoholics recently abstinent, divided in 2 series. In the model series (n = 94) several clinical variables and results of common laboratory tests were confronted to the findings of liver biopsies. These were classified in 3 groups: 1. Normal liver; 2. Moderate alterations; 3. Marked alterations, including alcoholic hepatitis and cirrhosis. Multivariate methods used were logistic regression analysis and a classification and regression tree (CART). Both methods entered gamma-glutamyltransferase (GGT), aspartate-aminotransferase (AST), weight and age as significant and independent variables. Univariate analysis with GGT and AST at different cutoffs were also performed. To predict the presence of any kind of damage (Groups 2 and 3), CART and AST > 30 IU showed the higher sensitivity, specificity and correct prediction, both in the model and validation series. For prediction of marked liver damage, a score based on logistic regression and GGT > 110 IU had the higher efficiencies. It is concluded that GGT and AST are good markers of alcoholic liver damage and that, using sample cutoffs, histologic diagnosis can be correctly predicted in 80% of recently abstinent asymptomatic alcoholics.
Applications of statistics to medical science, III. Correlation and regression.

PubMed

Watanabe, Hiroshi

2012-01-01

In this third part of a series surveying medical statistics, the concepts of correlation and regression are reviewed. In particular, methods of linear regression and logistic regression are discussed. Arguments related to survival analysis will be made in a subsequent paper.
Computing group cardinality constraint solutions for logistic regression problems.

PubMed

Zhang, Yong; Kwon, Dongjin; Pohl, Kilian M

2017-01-01

We derive an algorithm to directly solve logistic regression based on cardinality constraint, group sparsity and use it to classify intra-subject MRI sequences (e.g. cine MRIs) of healthy from diseased subjects. Group cardinality constraint models are often applied to medical images in order to avoid overfitting of the classifier to the training data. Solutions within these models are generally determined by relaxing the cardinality constraint to a weighted feature selection scheme. However, these solutions relate to the original sparse problem only under specific assumptions, which generally do not hold for medical image applications. In addition, inferring clinical meaning from features weighted by a classifier is an ongoing topic of discussion. Avoiding weighing features, we propose to directly solve the group cardinality constraint logistic regression problem by generalizing the Penalty Decomposition method. To do so, we assume that an intra-subject series of images represents repeated samples of the same disease patterns. We model this assumption by combining series of measurements created by a feature across time into a single group. Our algorithm then derives a solution within that model by decoupling the minimization of the logistic regression function from enforcing the group sparsity constraint. The minimum to the smooth and convex logistic regression problem is determined via gradient descent while we derive a closed form solution for finding a sparse approximation of that minimum. We apply our method to cine MRI of 38 healthy controls and 44 adult patients that received reconstructive surgery of Tetralogy of Fallot (TOF) during infancy. Our method correctly identifies regions impacted by TOF and generally obtains statistically significant higher classification accuracy than alternative solutions to this model, i.e., ones relaxing group cardinality constraints. Copyright © 2016 Elsevier B.V. All rights reserved.
Influential factors of red-light running at signalized intersection and prediction using a rare events logistic regression model.

PubMed

Ren, Yilong; Wang, Yunpeng; Wu, Xinkai; Yu, Guizhen; Ding, Chuan

2016-10-01

Red light running (RLR) has become a major safety concern at signalized intersection. To prevent RLR related crashes, it is critical to identify the factors that significantly impact the drivers' behaviors of RLR, and to predict potential RLR in real time. In this research, 9-month's RLR events extracted from high-resolution traffic data collected by loop detectors from three signalized intersections were applied to identify the factors that significantly affect RLR behaviors. The data analysis indicated that occupancy time, time gap, used yellow time, time left to yellow start, whether the preceding vehicle runs through the intersection during yellow, and whether there is a vehicle passing through the intersection on the adjacent lane were significantly factors for RLR behaviors. Furthermore, due to the rare events nature of RLR, a modified rare events logistic regression model was developed for RLR prediction. The rare events logistic regression method has been applied in many fields for rare events studies and shows impressive performance, but so far none of previous research has applied this method to study RLR. The results showed that the rare events logistic regression model performed significantly better than the standard logistic regression model. More importantly, the proposed RLR prediction method is purely based on loop detector data collected from a single advance loop detector located 400 feet away from stop-bar. This brings great potential for future field applications of the proposed method since loops have been widely implemented in many intersections and can collect data in real time. This research is expected to contribute to the improvement of intersection safety significantly. Copyright © 2016 Elsevier Ltd. All rights reserved.
Artificial neural network, genetic algorithm, and logistic regression applications for predicting renal colic in emergency settings.

PubMed

Eken, Cenker; Bilge, Ugur; Kartal, Mutlu; Eray, Oktay

2009-06-03

Logistic regression is the most common statistical model for processing multivariate data in the medical literature. Artificial intelligence models like an artificial neural network (ANN) and genetic algorithm (GA) may also be useful to interpret medical data. The purpose of this study was to perform artificial intelligence models on a medical data sheet and compare to logistic regression. ANN, GA, and logistic regression analysis were carried out on a data sheet of a previously published article regarding patients presenting to an emergency department with flank pain suspicious for renal colic. The study population was composed of 227 patients: 176 patients had a diagnosis of urinary stone, while 51 ultimately had no calculus. The GA found two decision rules in predicting urinary stones. Rule 1 consisted of being male, pain not spreading to back, and no fever. In rule 2, pelvicaliceal dilatation on bedside ultrasonography replaced no fever. ANN, GA rule 1, GA rule 2, and logistic regression had a sensitivity of 94.9, 67.6, 56.8, and 95.5%, a specificity of 78.4, 76.47, 86.3, and 47.1%, a positive likelihood ratio of 4.4, 2.9, 4.1, and 1.8, and a negative likelihood ratio of 0.06, 0.42, 0.5, and 0.09, respectively. The area under the curve was found to be 0.867, 0.720, 0.715, and 0.713 for all applications, respectively. Data mining techniques such as ANN and GA can be used for predicting renal colic in emergency settings and to constitute clinical decision rules. They may be an alternative to conventional multivariate analysis applications used in biostatistics.
Application of logistic regression for landslide susceptibility zoning of Cekmece Area, Istanbul, Turkey

NASA Astrophysics Data System (ADS)

Duman, T. Y.; Can, T.; Gokceoglu, C.; Nefeslioglu, H. A.; Sonmez, H.

2006-11-01

As a result of industrialization, throughout the world, cities have been growing rapidly for the last century. One typical example of these growing cities is Istanbul, the population of which is over 10 million. Due to rapid urbanization, new areas suitable for settlement and engineering structures are necessary. The Cekmece area located west of the Istanbul metropolitan area is studied, because the landslide activity is extensive in this area. The purpose of this study is to develop a model that can be used to characterize landslide susceptibility in map form using logistic regression analysis of an extensive landslide database. A database of landslide activity was constructed using both aerial-photography and field studies. About 19.2% of the selected study area is covered by deep-seated landslides. The landslides that occur in the area are primarily located in sandstones with interbedded permeable and impermeable layers such as claystone, siltstone and mudstone. About 31.95% of the total landslide area is located at this unit. To apply logistic regression analyses, a data matrix including 37 variables was constructed. The variables used in the forwards stepwise analyses are different measures of slope, aspect, elevation, stream power index (SPI), plan curvature, profile curvature, geology, geomorphology and relative permeability of lithological units. A total of 25 variables were identified as exerting strong influence on landslide occurrence, and included by the logistic regression equation. Wald statistics values indicate that lithology, SPI and slope are more important than the other parameters in the equation. Beta coefficients of the 25 variables included the logistic regression equation provide a model for landslide susceptibility in the Cekmece area. This model is used to generate a landslide susceptibility map that correctly classified 83.8% of the landslide-prone areas.
Predicting on-road assessment pass and fail outcomes in older drivers with cognitive impairment using a battery of computerized sensory-motor and cognitive tests.

PubMed

Hoggarth, Petra A; Innes, Carrie R H; Dalrymple-Alford, John C; Jones, Richard D

2013-12-01

To generate a robust model of computerized sensory-motor and cognitive test performance to predict on-road driving assessment outcomes in older persons with diagnosed or suspected cognitive impairment. A logistic regression model classified pass–fail outcomes of a blinded on-road driving assessment. Generalizability of the model was tested using leave-one-out cross-validation. Three specialist clinics in New Zealand. Drivers (n=279; mean age 78.4, 65% male) with diagnosed or suspected dementia, mild cognitive impairment, unspecified cognitive impairment, or memory problems referred for a medical driving assessment. A computerized battery of sensory-motor and cognitive tests and an on-road medical driving assessment. One hundred fifty-five participants (55.5%) received an on-road fail score. Binary logistic regression correctly classified 75.6% of the sample into on-road pass and fail groups. The cross-validation indicated accuracy of the model of 72.0% with sensitivity for detecting on-road fails of 73.5%, specificity of 70.2%, positive predictive value of 75.5%, and negative predictive value of 68%. The off-road assessment prediction model resulted in a substantial number of people who were assessed as likely to fail despite passing an on-road assessment and vice versa. Thus, despite a large multicenter sample, the use of off-road tests previously found to be useful in other older populations, and a carefully constructed and tested prediction model, off-road measures have yet to be found that are sufficiently accurate to allow acceptable determination of on-road driving safety of cognitively impaired older drivers. © 2013, Copyright the Authors Journal compilation © 2013, The American Geriatrics Society.
Updated logistic regression equations for the calculation of post-fire debris-flow likelihood in the western United States

USGS Publications Warehouse

Staley, Dennis M.; Negri, Jacquelyn A.; Kean, Jason W.; Laber, Jayme L.; Tillery, Anne C.; Youberg, Ann M.

2016-06-30

Wildfire can significantly alter the hydrologic response of a watershed to the extent that even modest rainstorms can generate dangerous flash floods and debris flows. To reduce public exposure to hazard, the U.S. Geological Survey produces post-fire debris-flow hazard assessments for select fires in the western United States. We use publicly available geospatial data describing basin morphology, burn severity, soil properties, and rainfall characteristics to estimate the statistical likelihood that debris flows will occur in response to a storm of a given rainfall intensity. Using an empirical database and refined geospatial analysis methods, we defined new equations for the prediction of debris-flow likelihood using logistic regression methods. We showed that the new logistic regression model outperformed previous models used to predict debris-flow likelihood.
EXpectation Propagation LOgistic REgRession (EXPLORER): distributed privacy-preserving online model learning.

PubMed

Wang, Shuang; Jiang, Xiaoqian; Wu, Yuan; Cui, Lijuan; Cheng, Samuel; Ohno-Machado, Lucila

2013-06-01

We developed an EXpectation Propagation LOgistic REgRession (EXPLORER) model for distributed privacy-preserving online learning. The proposed framework provides a high level guarantee for protecting sensitive information, since the information exchanged between the server and the client is the encrypted posterior distribution of coefficients. Through experimental results, EXPLORER shows the same performance (e.g., discrimination, calibration, feature selection, etc.) as the traditional frequentist logistic regression model, but provides more flexibility in model updating. That is, EXPLORER can be updated one point at a time rather than having to retrain the entire data set when new observations are recorded. The proposed EXPLORER supports asynchronized communication, which relieves the participants from coordinating with one another, and prevents service breakdown from the absence of participants or interrupted communications. Copyright © 2013 Elsevier Inc. All rights reserved.
A computational approach to compare regression modelling strategies in prediction research.

PubMed

Pajouheshnia, Romin; Pestman, Wiebe R; Teerenstra, Steven; Groenwold, Rolf H H

2016-08-25

It is often unclear which approach to fit, assess and adjust a model will yield the most accurate prediction model. We present an extension of an approach for comparing modelling strategies in linear regression to the setting of logistic regression and demonstrate its application in clinical prediction research. A framework for comparing logistic regression modelling strategies by their likelihoods was formulated using a wrapper approach. Five different strategies for modelling, including simple shrinkage methods, were compared in four empirical data sets to illustrate the concept of a priori strategy comparison. Simulations were performed in both randomly generated data and empirical data to investigate the influence of data characteristics on strategy performance. We applied the comparison framework in a case study setting. Optimal strategies were selected based on the results of a priori comparisons in a clinical data set and the performance of models built according to each strategy was assessed using the Brier score and calibration plots. The performance of modelling strategies was highly dependent on the characteristics of the development data in both linear and logistic regression settings. A priori comparisons in four empirical data sets found that no strategy consistently outperformed the others. The percentage of times that a model adjustment strategy outperformed a logistic model ranged from 3.9 to 94.9 %, depending on the strategy and data set. However, in our case study setting the a priori selection of optimal methods did not result in detectable improvement in model performance when assessed in an external data set. The performance of prediction modelling strategies is a data-dependent process and can be highly variable between data sets within the same clinical domain. A priori strategy comparison can be used to determine an optimal logistic regression modelling strategy for a given data set before selecting a final modelling approach.
Factors Infuencing Women in Pap Smear Uptake

NASA Astrophysics Data System (ADS)

Wijayanti, K. E.; Alam, I. G.

2017-03-01

Objective: Pap smear has proven can decrease death caused by cervical cancer. However, in Indonesia, only few woman who already did pap smear. The aim of this study was to investigate women’s knowledge about pap smear cervical cancer, and to investigate factors influence women to do pap smear test. Methods: Quantitative data colected through questionairre towards 31 women who did pap smear and 55 women who did not do pap smear. Questionairre was made using Health Belief model as a guideline to examine percieved susceptibility, perceived serioussnes, perceived benefits and perceived barriers. Chi square and multiple logistic regresion were used to investigate difference in knowledge and what the most factor that influence women to take pap smear test. Results: There’s significance knowledge difference betweeen women who did and did not do pap smear. But furthermore, by using Multiple Logistic Regression test, appearantly knowledge was not a strong predictor factor for women to take pap smear test (koefisiensi β = -0,164) Conclusion: Perceived barriers were factors that affected pap smear uptake in women in Indonesia. Few respondents get the wrong informations about pap smear, cevical cancer and its symptoms

Femoral neck shaft angle width is associated with hip-fracture risk in males but not independently of femoral neck bone density.

PubMed

Ripamonti, C; Lisi, L; Avella, M

2014-05-01

To investigate the specificity of the neck shaft angle (NSA) to predict hip fracture in males. We consecutively studied 228 males without fracture and 38 with hip fracture. A further 49 males with spine fracture were studied to evaluate the specificity of NSA for hip-fracture prediction. Femoral neck (FN) bone mineral density (FN-BMD), NSA, hip axis length and FN diameter (FND) were measured in each subject by dual X-ray absorptiometry. Between-mean differences in the studied variables were tested by the unpaired t-test. The ability of NSA to predict hip fracture was tested by logistic regression. Compared with controls, FN-BMD (p < 0.01) was significantly lower in both groups of males with fractures, whereas FND (p < 0.01) and NSA (p = 0.05) were higher only in the hip-fracture group. A significant inverse correlation (p < 0.01) was found between NSA and FN-BMD. By age-, height- and weight-corrected logistic regression, none of the tested geometric parameters, separately considered from FN-BMD, entered the best model to predict spine fracture, whereas NSA (p < 0.03) predicted hip fracture together with age (p < 0.001). When forced into the regression, FN-BMD (p < 0.001) became the only fracture predictor to enter the best model to predict both fracture types. NSA is associated with hip-fracture risk in males but is not independent of FN-BMD. The lack of ability of NSA to predict hip fracture in males independent of FN-BMD should depend on its inverse correlation with FN-BMD by capturing, as the strongest fracture predictor, some of the effects of NSA on the hip fracture. Conversely, NSA in females does not correlate with FN-BMD but independently predicts hip fractures.
Femoral neck shaft angle width is associated with hip-fracture risk in males but not independently of femoral neck bone density

PubMed Central

Lisi, L; Avella, M

2014-01-01

Objective: To investigate the specificity of the neck shaft angle (NSA) to predict hip fracture in males. Methods: We consecutively studied 228 males without fracture and 38 with hip fracture. A further 49 males with spine fracture were studied to evaluate the specificity of NSA for hip-fracture prediction. Femoral neck (FN) bone mineral density (FN-BMD), NSA, hip axis length and FN diameter (FND) were measured in each subject by dual X-ray absorptiometry. Between-mean differences in the studied variables were tested by the unpaired t-test. The ability of NSA to predict hip fracture was tested by logistic regression. Results: Compared with controls, FN-BMD (p < 0.01) was significantly lower in both groups of males with fractures, whereas FND (p < 0.01) and NSA (p = 0.05) were higher only in the hip-fracture group. A significant inverse correlation (p < 0.01) was found between NSA and FN-BMD. By age-, height- and weight-corrected logistic regression, none of the tested geometric parameters, separately considered from FN-BMD, entered the best model to predict spine fracture, whereas NSA (p < 0.03) predicted hip fracture together with age (p < 0.001). When forced into the regression, FN-BMD (p < 0.001) became the only fracture predictor to enter the best model to predict both fracture types. Conclusion: NSA is associated with hip-fracture risk in males but is not independent of FN-BMD. Advances in knowledge: The lack of ability of NSA to predict hip fracture in males independent of FN-BMD should depend on its inverse correlation with FN-BMD by capturing, as the strongest fracture predictor, some of the effects of NSA on the hip fracture. Conversely, NSA in females does not correlate with FN-BMD but independently predicts hip fractures. PMID:24678889
Comparison of statistical tests for association between rare variants and binary traits.

PubMed

Bacanu, Silviu-Alin; Nelson, Matthew R; Whittaker, John C

2012-01-01

Genome-wide association studies have found thousands of common genetic variants associated with a wide variety of diseases and other complex traits. However, a large portion of the predicted genetic contribution to many traits remains unknown. One plausible explanation is that some of the missing variation is due to the effects of rare variants. Nonetheless, the statistical analysis of rare variants is challenging. A commonly used method is to contrast, within the same region (gene), the frequency of minor alleles at rare variants between cases and controls. However, this strategy is most useful under the assumption that the tested variants have similar effects. We previously proposed a method that can accommodate heterogeneous effects in the analysis of quantitative traits. Here we extend this method to include binary traits that can accommodate covariates. We use simulations for a variety of causal and covariate impact scenarios to compare the performance of the proposed method to standard logistic regression, C-alpha, SKAT, and EREC. We found that i) logistic regression methods perform well when the heterogeneity of the effects is not extreme and ii) SKAT and EREC have good performance under all tested scenarios but they can be computationally intensive. Consequently, it would be more computationally desirable to use a two-step strategy by (i) selecting promising genes by faster methods and ii) analyzing selected genes using SKAT/EREC. To select promising genes one can use (1) regression methods when effect heterogeneity is assumed to be low and the covariates explain a non-negligible part of trait variability, (2) C-alpha when heterogeneity is assumed to be large and covariates explain a small fraction of trait's variability and (3) the proposed trend and heterogeneity test when the heterogeneity is assumed to be non-trivial and the covariates explain a large fraction of trait variability.
Cytopathologic differential diagnosis of low-grade urothelial carcinoma and reactive urothelial proliferation in bladder washings: a logistic regression analysis.

PubMed

Cakir, Ebru; Kucuk, Ulku; Pala, Emel Ebru; Sezer, Ozlem; Ekin, Rahmi Gokhan; Cakmak, Ozgur

2017-05-01

Conventional cytomorphologic assessment is the first step to establish an accurate diagnosis in urinary cytology. In cytologic preparations, the separation of low-grade urothelial carcinoma (LGUC) from reactive urothelial proliferation (RUP) can be exceedingly difficult. The bladder washing cytologies of 32 LGUC and 29 RUP were reviewed. The cytologic slides were examined for the presence or absence of the 28 cytologic features. The cytologic criteria showing statistical significance in LGUC were increased numbers of monotonous single (non-umbrella) cells, three-dimensional cellular papillary clusters without fibrovascular cores, irregular bordered clusters, atypical single cells, irregular nuclear overlap, cytoplasmic homogeneity, increased N/C ratio, pleomorphism, nuclear border irregularity, nuclear eccentricity, elongated nuclei, and hyperchromasia (p ˂ 0.05), and the cytologic criteria showing statistical significance in RUP were inflammatory background, mixture of small and large urothelial cells, loose monolayer aggregates, and vacuolated cytoplasm (p ˂ 0.05). When these variables were subjected to a stepwise logistic regression analysis, four features were selected to distinguish LGUC from RUP: increased numbers of monotonous single (non-umbrella) cells, increased nuclear cytoplasmic ratio, hyperchromasia, and presence of small and large urothelial cells (p = 0.0001). By this logistic model of the 32 cases with proven LGUC, the stepwise logistic regression analysis correctly predicted 31 (96.9%) patients with this diagnosis, and of the 29 patients with RUP, the logistic model correctly predicted 26 (89.7%) patients as having this disease. There are several cytologic features to separate LGUC from RUP. Stepwise logistic regression analysis is a valuable tool for determining the most useful cytologic criteria to distinguish these entities. © 2017 APMIS. Published by John Wiley & Sons Ltd.
Maximal bite force, facial morphology and sucking habits in young children with functional posterior crossbite

PubMed Central

CASTELO, Paula Midori; GAVIÃO, Maria Beatriz Duarte; PEREIRA, Luciano José; BONJARDIM, Leonardo Rigoldi

2010-01-01

Objective The maintenance of normal conditions of the masticatory function is determinant for the correct growth and development of its structures. Thus, the aims of this study were to evaluate the influence of sucking habits on the presence of crossbite and its relationship with maximal bite force, facial morphology and body variables in 67 children of both genders (3.5-7 years) with primary or early mixed dentition. Material and methods The children were divided in four groups: primary-normocclusion (PN, n=19), primary-crossbite (PC, n=19), mixed-normocclusion (MN, n=13), and mixed-crossbite (MC, n=16). Bite force was measured with a pressurized tube, and facial morphology was determined by standardized frontal photographs: AFH (anterior face height) and BFW (bizygomatic facial width). Results It was observed that MC group showed lower bite force than MN, and AFH/ BFW was significantly smaller in PN than PC (t-test). Weight and height were only significantly correlated with bite force in PC group (Pearson’s correlation test). In the primary dentition, AFH/BFW and breast-feeding (at least six months) were positive and negatively associated with crossbite, respectively (multiple logistic regression). In the mixed dentition, breastfeeding and bite force showed negative associations with crossbite (univariate regression), while nonnutritive sucking (up to 3 years) associated significantly with crossbite in all groups (multiple logistic regression). Conclusions In the studied sample, sucking habits played an important role in the etiology of crossbite, which was associated with lower bite force and long-face tendency. PMID:20485925
Binary Logistic Regression Analysis for Detecting Differential Item Functioning: Effectiveness of R[superscript 2] and Delta Log Odds Ratio Effect Size Measures

ERIC Educational Resources Information Center

Hidalgo, Mª Dolores; Gómez-Benito, Juana; Zumbo, Bruno D.

2014-01-01

The authors analyze the effectiveness of the R[superscript 2] and delta log odds ratio effect size measures when using logistic regression analysis to detect differential item functioning (DIF) in dichotomous items. A simulation study was carried out, and the Type I error rate and power estimates under conditions in which only statistical testing…
Logistic quantile regression provides improved estimates for bounded avian counts: a case study of California Spotted Owl fledgling production

Treesearch

Brian S. Cade; Barry R. Noon; Rick D. Scherer; John J. Keane

2017-01-01

Counts of avian fledglings, nestlings, or clutch size that are bounded below by zero and above by some small integer form a discrete random variable distribution that is not approximated well by conventional parametric count distributions such as the Poisson or negative binomial. We developed a logistic quantile regression model to provide estimates of the empirical...
Comparison of four methods for deriving hospital standardised mortality ratios from a single hierarchical logistic regression model.

PubMed

Mohammed, Mohammed A; Manktelow, Bradley N; Hofer, Timothy P

2016-04-01

There is interest in deriving case-mix adjusted standardised mortality ratios so that comparisons between healthcare providers, such as hospitals, can be undertaken in the controversial belief that variability in standardised mortality ratios reflects quality of care. Typically standardised mortality ratios are derived using a fixed effects logistic regression model, without a hospital term in the model. This fails to account for the hierarchical structure of the data - patients nested within hospitals - and so a hierarchical logistic regression model is more appropriate. However, four methods have been advocated for deriving standardised mortality ratios from a hierarchical logistic regression model, but their agreement is not known and neither do we know which is to be preferred. We found significant differences between the four types of standardised mortality ratios because they reflect a range of underlying conceptual issues. The most subtle issue is the distinction between asking how an average patient fares in different hospitals versus how patients at a given hospital fare at an average hospital. Since the answers to these questions are not the same and since the choice between these two approaches is not obvious, the extent to which profiling hospitals on mortality can be undertaken safely and reliably, without resolving these methodological issues, remains questionable. © The Author(s) 2012.
A comparison of three methods of assessing differential item functioning (DIF) in the Hospital Anxiety Depression Scale: ordinal logistic regression, Rasch analysis and the Mantel chi-square procedure.

PubMed

Cameron, Isobel M; Scott, Neil W; Adler, Mats; Reid, Ian C

2014-12-01

It is important for clinical practice and research that measurement scales of well-being and quality of life exhibit only minimal differential item functioning (DIF). DIF occurs where different groups of people endorse items in a scale to different extents after being matched by the intended scale attribute. We investigate the equivalence or otherwise of common methods of assessing DIF. Three methods of measuring age- and sex-related DIF (ordinal logistic regression, Rasch analysis and Mantel χ(2) procedure) were applied to Hospital Anxiety Depression Scale (HADS) data pertaining to a sample of 1,068 patients consulting primary care practitioners. Three items were flagged by all three approaches as having either age- or sex-related DIF with a consistent direction of effect; a further three items identified did not meet stricter criteria for important DIF using at least one method. When applying strict criteria for significant DIF, ordinal logistic regression was slightly less sensitive. Ordinal logistic regression, Rasch analysis and contingency table methods yielded consistent results when identifying DIF in the HADS depression and HADS anxiety scales. Regardless of methods applied, investigators should use a combination of statistical significance, magnitude of the DIF effect and investigator judgement when interpreting the results.
Extreme Sparse Multinomial Logistic Regression: A Fast and Robust Framework for Hyperspectral Image Classification

NASA Astrophysics Data System (ADS)

Cao, Faxian; Yang, Zhijing; Ren, Jinchang; Ling, Wing-Kuen; Zhao, Huimin; Marshall, Stephen

2017-12-01

Although the sparse multinomial logistic regression (SMLR) has provided a useful tool for sparse classification, it suffers from inefficacy in dealing with high dimensional features and manually set initial regressor values. This has significantly constrained its applications for hyperspectral image (HSI) classification. In order to tackle these two drawbacks, an extreme sparse multinomial logistic regression (ESMLR) is proposed for effective classification of HSI. First, the HSI dataset is projected to a new feature space with randomly generated weight and bias. Second, an optimization model is established by the Lagrange multiplier method and the dual principle to automatically determine a good initial regressor for SMLR via minimizing the training error and the regressor value. Furthermore, the extended multi-attribute profiles (EMAPs) are utilized for extracting both the spectral and spatial features. A combinational linear multiple features learning (MFL) method is proposed to further enhance the features extracted by ESMLR and EMAPs. Finally, the logistic regression via the variable splitting and the augmented Lagrangian (LORSAL) is adopted in the proposed framework for reducing the computational time. Experiments are conducted on two well-known HSI datasets, namely the Indian Pines dataset and the Pavia University dataset, which have shown the fast and robust performance of the proposed ESMLR framework.
Latin hypercube approach to estimate uncertainty in ground water vulnerability

USGS Publications Warehouse

Gurdak, J.J.; McCray, J.E.; Thyne, G.; Qi, S.L.

2007-01-01

A methodology is proposed to quantify prediction uncertainty associated with ground water vulnerability models that were developed through an approach that coupled multivariate logistic regression with a geographic information system (GIS). This method uses Latin hypercube sampling (LHS) to illustrate the propagation of input error and estimate uncertainty associated with the logistic regression predictions of ground water vulnerability. Central to the proposed method is the assumption that prediction uncertainty in ground water vulnerability models is a function of input error propagation from uncertainty in the estimated logistic regression model coefficients (model error) and the values of explanatory variables represented in the GIS (data error). Input probability distributions that represent both model and data error sources of uncertainty were simultaneously sampled using a Latin hypercube approach with logistic regression calculations of probability of elevated nonpoint source contaminants in ground water. The resulting probability distribution represents the prediction intervals and associated uncertainty of the ground water vulnerability predictions. The method is illustrated through a ground water vulnerability assessment of the High Plains regional aquifer. Results of the LHS simulations reveal significant prediction uncertainties that vary spatially across the regional aquifer. Additionally, the proposed method enables a spatial deconstruction of the prediction uncertainty that can lead to improved prediction of ground water vulnerability. ?? 2007 National Ground Water Association.
The weighted priors approach for combining expert opinions in logistic regression experiments

DOE PAGES

Quinlan, Kevin R.; Anderson-Cook, Christine M.; Myers, Kary L.

2017-04-24

When modeling the reliability of a system or component, it is not uncommon for more than one expert to provide very different prior estimates of the expected reliability as a function of an explanatory variable such as age or temperature. Our goal in this paper is to incorporate all information from the experts when choosing a design about which units to test. Bayesian design of experiments has been shown to be very successful for generalized linear models, including logistic regression models. We use this approach to develop methodology for the case where there are several potentially non-overlapping priors under consideration.more » While multiple priors have been used for analysis in the past, they have never been used in a design context. The Weighted Priors method performs well for a broad range of true underlying model parameter choices and is more robust when compared to other reasonable design choices. Finally, we illustrate the method through multiple scenarios and a motivating example. Additional figures for this article are available in the online supplementary information.« less
Analyzing thresholds and efficiency with hierarchical Bayesian logistic regression.

PubMed

Houpt, Joseph W; Bittner, Jennifer L

2018-07-01

Ideal observer analysis is a fundamental tool used widely in vision science for analyzing the efficiency with which a cognitive or perceptual system uses available information. The performance of an ideal observer provides a formal measure of the amount of information in a given experiment. The ratio of human to ideal performance is then used to compute efficiency, a construct that can be directly compared across experimental conditions while controlling for the differences due to the stimuli and/or task specific demands. In previous research using ideal observer analysis, the effects of varying experimental conditions on efficiency have been tested using ANOVAs and pairwise comparisons. In this work, we present a model that combines Bayesian estimates of psychometric functions with hierarchical logistic regression for inference about both unadjusted human performance metrics and efficiencies. Our approach improves upon the existing methods by constraining the statistical analysis using a standard model connecting stimulus intensity to human observer accuracy and by accounting for variability in the estimates of human and ideal observer performance scores. This allows for both individual and group level inferences. Copyright © 2018 Elsevier Ltd. All rights reserved.
Profiles of internalizing and externalizing symptoms associated with bullying victimization.

PubMed

Eastman, Meridith; Foshee, Vangie; Ennett, Susan; Sotres-Alvarez, Daniela; Reyes, H Luz McNaughton; Faris, Robert; North, Kari

2018-06-01

This study identified profiles of internalizing (anxiety and depression) and externalizing (delinquency and violence against peers) symptoms among bullying victims and examined associations between bullying victimization characteristics and profile membership. The sample consisted of 1196 bullying victims in grades 8-10 (M age  = 14.4, SD = 1.01) who participated in The Context Study in three North Carolina counties in Fall 2003. Five profiles were identified using latent profile analysis: an asymptomatic profile and four profiles capturing combinations of internalizing and externalizing symptoms. Associations between bullying characteristics and membership in symptom profiles were tested using multinomial logistic regression. More frequent victimization increased odds of membership in the two high internalizing profiles compared to the asymptomatic profile. Across all multinomial logistic regression models, when the high internalizing, high externalizing profile was the reference category, adolescents who received any type of bullying (direct, indirect, or dual) were more likely to be in this category than any others. Copyright © 2018 The Foundation for Professionals in Services for Adolescents. Published by Elsevier Ltd. All rights reserved.
Cannabis, tobacco and domestic fumes intake are associated with nasopharyngeal carcinoma in North Africa.

PubMed

Feng, B-J; Khyatti, M; Ben-Ayoub, W; Dahmoul, S; Ayad, M; Maachi, F; Bedadra, W; Abdoun, M; Mesli, S; Bakkali, H; Jalbout, M; Hamdi-Cherif, M; Boualga, K; Bouaouina, N; Chouchane, L; Benider, A; Ben-Ayed, F; Goldgar, D E; Corbex, M

2009-10-06

The lifestyle risk factors for nasopharyngeal carcinoma (NPC) in North Africa are not known. From 2002 to 2005, we interviewed 636 patients and 615 controls from Algeria, Morocco and Tunisia, frequency-matched by centre, age, sex, and childhood household type (urban/rural). Conditional logistic regression was used to evaluate the association of lifestyles with NPC risk, controlling for socioeconomic status and dietary risk factors. Cigarette smoking and snuff (tobacco powder with additives) intake were significantly associated with differentiated NPC but not with undifferentiated carcinoma (UCNT), which is the major histological type of NPC in these populations. As demonstrated by a stratified permutation test and by conditional logistic regression, marijuana smoking significantly elevated NPC risk independently of cigarette smoking, suggesting dissimilar carcinogenic mechanisms between cannabis and tobacco. Domestic cooking fumes intake by using kanoun (compact charcoal oven) during childhood increased NPC risk, whereas exposure during adulthood had less effect. Neither alcohol nor shisha (water pipe) was associated with risk. Tobacco, cannabis and domestic cooking fumes intake are risk factors for NPC in western North Africa.
Cannabis, tobacco and domestic fumes intake are associated with nasopharyngeal carcinoma in North Africa

PubMed Central

Feng, B-J; Khyatti, M; Ben-Ayoub, W; Dahmoul, S; Ayad, M; Maachi, F; Bedadra, W; Abdoun, M; Mesli, S; Bakkali, H; Jalbout, M; Hamdi-Cherif, M; Boualga, K; Bouaouina, N; Chouchane, L; Benider, A; Ben-Ayed, F; Goldgar, D E; Corbex, M

2009-01-01

Background: The lifestyle risk factors for nasopharyngeal carcinoma (NPC) in North Africa are not known. Methods: From 2002 to 2005, we interviewed 636 patients and 615 controls from Algeria, Morocco and Tunisia, frequency-matched by centre, age, sex, and childhood household type (urban/rural). Conditional logistic regression was used to evaluate the association of lifestyles with NPC risk, controlling for socioeconomic status and dietary risk factors. Results: Cigarette smoking and snuff (tobacco powder with additives) intake were significantly associated with differentiated NPC but not with undifferentiated carcinoma (UCNT), which is the major histological type of NPC in these populations. As demonstrated by a stratified permutation test and by conditional logistic regression, marijuana smoking significantly elevated NPC risk independently of cigarette smoking, suggesting dissimilar carcinogenic mechanisms between cannabis and tobacco. Domestic cooking fumes intake by using kanoun (compact charcoal oven) during childhood increased NPC risk, whereas exposure during adulthood had less effect. Neither alcohol nor shisha (water pipe) was associated with risk. Conclusion: Tobacco, cannabis and domestic cooking fumes intake are risk factors for NPC in western North Africa. PMID:19724280
Seroprevalence and Risk Factors of Chlamydia abortus Infection in Tibetan Sheep in Gansu Province, Northwest China

PubMed Central

Qin, Si-Yuan; Yin, Ming-Yang; Cong, Wei; Zhou, Dong-Hui; Zhang, Xiao-Xuan; Zhao, Quan; Zhu, Xing-Quan; Zhou, Ji-Zhang; Qian, Ai-Dong

2014-01-01

Chlamydia abortus, an important pathogen in a variety of animals, is associated with abortion in sheep. In the present study, 1732 blood samples, collected from Tibetan sheep between June 2013 and April 2014, were examined by the indirect hemagglutination (IHA) test, aiming to evaluate the seroprevalence and risk factors of C. abortus infection in Tibetan sheep. 323 of 1732 (18.65%) samples were seropositive for C. abortus antibodies at the cut-off of 1 : 16. A multivariate logistic regression analysis was used to evaluate the risk factors associated with seroprevalence, which could provide foundation to prevent and control C. abortus infection in Tibetan sheep. Gender of Tibetan sheep was left out of the final model because it is not significant in the logistic regression analysis (P > 0.05). Region, season, and age were considered as major risk factors associated with C. abortus infection in Tibetan sheep. Our study revealed a widespread and high prevalence of C. abortus infection in Tibetan sheep in Gansu province, northwest China, with higher exposure risk in different seasons and ages and distinct geographical distribution. PMID:25401129
The weighted priors approach for combining expert opinions in logistic regression experiments

DOE Office of Scientific and Technical Information (OSTI.GOV)

Quinlan, Kevin R.; Anderson-Cook, Christine M.; Myers, Kary L.

When modeling the reliability of a system or component, it is not uncommon for more than one expert to provide very different prior estimates of the expected reliability as a function of an explanatory variable such as age or temperature. Our goal in this paper is to incorporate all information from the experts when choosing a design about which units to test. Bayesian design of experiments has been shown to be very successful for generalized linear models, including logistic regression models. We use this approach to develop methodology for the case where there are several potentially non-overlapping priors under consideration.more » While multiple priors have been used for analysis in the past, they have never been used in a design context. The Weighted Priors method performs well for a broad range of true underlying model parameter choices and is more robust when compared to other reasonable design choices. Finally, we illustrate the method through multiple scenarios and a motivating example. Additional figures for this article are available in the online supplementary information.« less
Beyond logistic regression: structural equations modelling for binary variables and its application to investigating unobserved confounders.

PubMed

Kupek, Emil

2006-03-15

Structural equation modelling (SEM) has been increasingly used in medical statistics for solving a system of related regression equations. However, a great obstacle for its wider use has been its difficulty in handling categorical variables within the framework of generalised linear models. A large data set with a known structure among two related outcomes and three independent variables was generated to investigate the use of Yule's transformation of odds ratio (OR) into Q-metric by (OR-1)/(OR+1) to approximate Pearson's correlation coefficients between binary variables whose covariance structure can be further analysed by SEM. Percent of correctly classified events and non-events was compared with the classification obtained by logistic regression. The performance of SEM based on Q-metric was also checked on a small (N = 100) random sample of the data generated and on a real data set. SEM successfully recovered the generated model structure. SEM of real data suggested a significant influence of a latent confounding variable which would have not been detectable by standard logistic regression. SEM classification performance was broadly similar to that of the logistic regression. The analysis of binary data can be greatly enhanced by Yule's transformation of odds ratios into estimated correlation matrix that can be further analysed by SEM. The interpretation of results is aided by expressing them as odds ratios which are the most frequently used measure of effect in medical statistics.
Joint and fascial chronic graft-vs-host disease: correlations with clinical and laboratory parameters

PubMed Central

Vukić, Tamara; Smith, Sean Robinson; Ljubas Kelečić, Dina; Desnica, Lana; Prenc, Ema; Pulanić, Dražen; Vrhovac, Radovan; Nemet, Damir; Pavletic, Steven Z.

2016-01-01

Aim To determine if there are correlations between joint and fascial chronic graft-vs-host disease (cGVHD) with clinical findings, laboratory parameters, and measures of functional capacity. Methods 29 patients were diagnosed with cGVHD based on National Institutes of Health (NIH) Consensus Criteria at the University Hospital Centre Zagreb from October 2013 to October 2015. Physical examination, including functional measures such as 2-minute walk test and hand grip strength, as well as laboratory tests were performed. The relationship between these evaluations and the severity of joint and fascial cGVHD was tested by logistical regression analysis. Results 12 of 29 patients (41.3%) had joint and fascial cGVHD diagnosed according to NIH Consensus Criteria. There was a significant positive correlation of joint and fascial cGVHD and skin cGVHD (P < 0.001), serum C3 complement level (P = 0.045), and leukocytes (P = 0.032). There was a significant negative correlation between 2-minute walk test (P = 0.016), percentage of cytotoxic T cells CD3+/CD8+ (P = 0.022), serum albumin (P = 0.047), and Karnofsky score (P < 0.001). Binary logistic regression model found that a significant predictor for joint and fascial cGVHD was cGVHD skin involvement (odds ratio, 7.79; 95 confidence interval 1.87-32.56; P = 0.005). Conclusion Joint and fascial cGVHD manifestations correlated with multiple laboratory measurements, clinical features, and cGVHD skin involvement, which was a significant predictor for joint and fascial cGVHD. PMID:27374828

A case-control study of the risk factors for obstetric fistula in Tigray, Ethiopia.

PubMed

Lewis Wall, L; Belay, Shewaye; Haregot, Tesfahun; Dukes, Jonathan; Berhan, Eyoel; Abreha, Melaku

2017-12-01

We tested the null hypothesis that there were no differences between patients with obstetric fistula and parous controls without fistula. A unmatched case-control study was carried out comparing 75 women with a history of obstetric fistula with 150 parous controls with no history of fistula. Height and weight were measured for each participant, along with basic socio-demographic and obstetric information. Descriptive statistics were calculated and differences between the groups were analyzed using Student's t test, Mann-Whitney U test where appropriate, and Chi-squared or Fisher's exact test, along with backward stepwise logistic regression analyses to detect predictors of obstetric fistula. Associations with a p value <0.05 were considered significant. Patients with fistulas married earlier and delivered their first pregnancies earlier than controls. They had significantly less education, a higher prevalence of divorce/separation, and lived in more impoverished circumstances than controls. Fistula patients had worse reproductive histories, with greater numbers of stillbirths/abortions and higher rates of assisted vaginal delivery and cesarean section. The final logistic regression model found four significant risk factors for developing an obstetric fistula: age at marriage (OR 1.23), history of assisted vaginal delivery (OR 3.44), lack of adequate antenatal care (OR 4.43), and a labor lasting longer than 1 day (OR 14.84). Our data indicate that obstetric fistula results from the lack of access to effective obstetrical services when labor is prolonged. Rural poverty and lack of adequate transportation infrastructure are probably important co-factors in inhibiting access to needed care.
Isomorphic red blood cells using automated urine flow cytometry is a reliable method in diagnosis of bladder cancer.

PubMed

Muto, Satoru; Sugiura, Syo-Ichiro; Nakajima, Akiko; Horiuchi, Akira; Inoue, Masahiro; Saito, Keisuke; Isotani, Shuji; Yamaguchi, Raizo; Ide, Hisamitsu; Horie, Shigeo

2014-10-01

We aimed to identify patients with a chief complaint of hematuria who could safely avoid unnecessary radiation and instrumentation in the diagnosis of bladder cancer (BC), using automated urine flow cytometry to detect isomorphic red blood cells (RBCs) in urine. We acquired urine samples from 134 patients over the age of 35 years with a chief complaint of hematuria and a positive urine occult blood test or microhematuria. The data were analyzed using the UF-1000i (®) (Sysmex Co., Ltd., Kobe, Japan) automated urine flow cytometer to determine RBC morphology, which was classified as isomorphic or dysmorphic. The patients were divided into two groups (BC versus non-BC) for statistical analysis. Multivariate logistic regression analysis was used to determine the predictive value of flow cytometry versus urine cytology, the bladder tumor antigen test, occult blood in urine test, and microhematuria test. BC was confirmed in 26 of 134 patients (19.4 %). The area under the curve for RBC count using the automated urine flow cytometer was 0.94, representing the highest reference value obtained in this study. Isomorphic RBCs were detected in all patients in the BC group. On multivariate logistic regression analysis, only isomorphic RBC morphology was significantly predictive for BC (p < 0.001). Analytical parameters such as sensitivity, specificity, positive predictive value, and negative predictive value of isomorphic RBCs in urine were 100.0, 91.7, 74.3, and 100.0 %, respectively. Detection of urinary isomorphic RBCs using automated urine flow cytometry is a reliable method in the diagnosis of BC with hematuria.
Prevalence of and Factors Associated with Negative Microscopic Diagnosis of Cutaneous Leishmaniasis in Rural Peru.

PubMed

Lamm, Ryan; Alves, Clark; Perrotta, Grace; Murphy, Meagan; Messina, Catherine; Sanchez, Juan F; Perez, Erika; Rosales, Luis Angel; Lescano, Andres G; Smith, Edward; Valdivia, Hugo; Fuhrer, Jack; Ballard, Sarah-Blythe

2018-06-04

Cutaneous leishmaniasis is endemic to South America where diagnosis is most commonly conducted via microscopy. Patients with suspected leishmaniasis were referred for enrollment by the Ministry of Health (MoH) in Lima, Iquitos, Puerto Maldonado, and several rural areas of Peru. A 43-question survey requesting age, gender, occupation, characterization of the lesion(s), history of leishmaniasis, and insect-deterrent behaviors was administered. Polymerase chain reaction (PCR) was conducted on lesion materials at the Naval Medical Research Unit No. 6 in Lima, and the results were compared with those obtained by the MoH using microscopy. Factors associated with negative microscopy and positive PCR results were identified using χ 2 test, t -test, and multivariate logistic regression analyses. Negative microscopy with positive PCR occurred in 31% (123/403) of the 403 cases. After adjusting for confounders, binary multivariate logistic regression analyses revealed that negative microscopy with positive PCR was associated with patients who were male (adjusted OR = 1.93 [1.06-3.53], P = 0.032), had previous leishmaniasis (adjusted OR = 2.93 [1.65-5.22], P < 0.0001), had larger lesions (adjusted OR = 1.02 [1.003-1.03], P = 0.016), and/or had a longer duration between lesion appearance and PCR testing (adjusted OR = 1.12 [1.02-1.22], P = 0.017). Future research should focus on further exploration of these underlying variables, discovery of other factors that may be associated with negative microscopy diagnosis, and the development and implementation of improved testing in endemic regions.
Analysis of Radiation Effects in Digital Subtraction Angiography of Intracranial Artery Stenosis.

PubMed

Guo, Chaoqun; Shi, Xiaolei; Ding, Xianhui; Zhou, Zhiming

2018-04-21

Intracranial artery stenosis (IAS) is the most common cause for acute cerebral accidents. Digital subtraction angiography (DSA) is the gold standard to detect IAS and usually brings excess radiation exposure to examinees and examiners. The artery pathology might influence the interventional procedure, causing prolonged radiation effects. However, no studies on the association between IAS pathology and operational parameters are available. A retrospective analysis was conducted on 93 patients with first-ever stroke/transient ischemic attack, who received DSA examination within 3 months from onset in this single center. Comparison of baseline characteristics was determined by 2-tailed Student's t-test or the chi-square test between subjects with and without IAS. A binary logistic regression analysis was performed to determine the association between IAS pathology and the items with a P value <0.05 in Student's t-test or chi-square test. There were 93 candidates (42 with IAS and 51 without IAS) in this study. The 2 groups shared no significance of the baseline characteristics (P > 0.05). We found a significantly higher total time, higher kerma area product, greater total dose, and greater DSA dose in the IAS group than in those without IAS (P < 0.05). A binary logistic regression analysis indicated the significant association between total time and IAS pathology (P < 0.05) but no significance in kerma area product, radiation dose, and DSA dose (P > 0.05). IAS pathology would indicate a prolonged total time of DSA procedure in clinical practice. However, the radiation effects would not change with pathologic changes. Copyright © 2018 Elsevier Inc. All rights reserved.
Specifying the Role of Exposure to Violence and Violent Behavior on Initiation of Gun Carrying: A Longitudinal Test of Three Models of Youth Gun Carrying

ERIC Educational Resources Information Center

Spano, Richard; Pridemore, William Alex; Bolland, John

2012-01-01

Two waves of longitudinal data from 1,049 African American youth living in extreme poverty are used to examine the impact of exposure to violence (Time 1) and violent behavior (Time 1) on first time gun carrying (Time 2). Multivariate logistic regression results indicate that (a) violent behavior (Time 1) increased the likelihood of initiation of…
Are Students of Color More Likely to Graduate from College if They Attend More Selective Institutions? Evidence from a Cohort of Recipients and Nonrecipients of the Gates Millennium Scholarship Program

ERIC Educational Resources Information Center

Melguizo, Tatiana

2010-01-01

The study takes advantage of the nontraditional selection process of the Gates Millennium Scholars (GMS) program to test the association between selectivity of 4-year institution attended as well as other noncognitive variables on the college completion rates of a sample of students of color. The results of logistic regression and propensity score…
Impulsivity, attention, memory, and decision-making among adolescent marijuana users.

PubMed

Dougherty, Donald M; Mathias, Charles W; Dawes, Michael A; Furr, R Michael; Charles, Nora E; Liguori, Anthony; Shannon, Erin E; Acheson, Ashley

2013-03-01

Marijuana is a popular drug of abuse among adolescents, and they may be uniquely vulnerable to resulting cognitive and behavioral impairments. Previous studies have found impairments among adolescent marijuana users. However, the majority of this research has examined measures individually rather than multiple domains in a single cohesive analysis. This study used a logistic regression model that combines performance on a range of tasks to identify which measures were most altered among adolescent marijuana users. The purpose of this research was to determine unique associations between adolescent marijuana use and performances on multiple cognitive and behavioral domains (attention, memory, decision-making, and impulsivity) in 14- to 17-year-olds while simultaneously controlling for performances across the measures to determine which measures most strongly distinguish marijuana users from nonusers. Marijuana-using adolescents (n = 45) and controls (n = 48) were tested. Logistic regression analyses were conducted to test for: (1) differences between marijuana users and nonusers on each measure, (2) associations between marijuana use and each measure after controlling for the other measures, and (3) the degree to which (1) and (2) together elucidated differences among marijuana users and nonusers. Of all the cognitive and behavioral domains tested, impaired short-term recall memory and consequence sensitivity impulsivity were associated with marijuana use after controlling for performances across all measures. This study extends previous findings by identifying cognitive and behavioral impairments most strongly associated with adolescent marijuana users. These specific deficits are potential targets of intervention for this at-risk population.
[The role of uric acid in the insulin resistance in children and adolescents with obesity].

PubMed

de Miranda, Josiane Aparecida; Almeida, Guilherme Gomide; Martins, Raissa Isabelle Leão; Cunha, Mariana Botrel; Belo, Vanessa Almeida; dos Santos, José Eduardo Tanus; Mourão-Júnior, Carlos Alberto; Lanna, Carla Márcia Moreira

2015-12-01

To investigate the association between serum uric acid levels and insulin resistance in children and adolescents with obesity. Cross-sectional study with 245 children and adolescents (134 obese and 111 controls), aged 8 to 18 years. The anthropometric variables (weight, height and waist circumference), blood pressure and biochemical parameters were collected. The clinical characteristics of the groups were analyzed by t-test or chi-square test. To evaluate the association between uric acid levels and insulin resistance the Pearson's test and logistic regression were applied. The prevalence of insulin resistance was 26.9%. The anthropometric variables, systolic and diastolic blood pressure and biochemical variables were significantly higher in the obese group (p<0.001), except for the high-density-lipoprotein cholesterol. There was a positive and significant correlation between anthropometric variables and uric acid with HOMA-IR in the obese and in the control groups, which was higher in the obese group and in the total sample. The logistic regression model that included age, gender and obesity, showed an odds ratio of uric acid as a variable associated with insulin resistance of 1.91 (95%CI 1.40 to 2.62; p<-0.001). The increase in serum uric acid showed a positive statistical correlation with insulin resistance and it is associated with and increased risk of insulin resistance in obese children and adolescents. Copyright © 2015 Sociedade de Pediatria de São Paulo. Publicado por Elsevier Editora Ltda. All rights reserved.
Serological Survey and Associated Risk Factors of Visceral Leish-maniasis in Qom Province, Central Iran.

PubMed

Rakhshanpour, Arash; Mohebali, Mehdi; Akhondi, Behnaz; Rahimi, Mohammad Taghi; Rokni, Mohammad Bagher

2014-01-01

Visceral leishmaniasis (VL) or kala-azar is considered as a parasitic disease caused by the species of Leishmania donovani complex which is intracellular parasites. This systemic disease is endemic in some parts of provinc-es of Iran. The aim of this study was to determine the seroprevalence of VL in Qom Province, central Iran using di-rect agglutination test (DAT). Overall, 1564 serum samples (800 males and 764 females) were collected from selected subjects by random-ized cluster sampling in 2011-2012. Sera were tested and analyzed by DAT. Before sampling; a questionnaire was filled out for each case. Data were analyzed using Chi-square test and multivariate logistic regression for risk factors analysis. Of 1564 individuals, 53 cases (3.38%) showed Leishmania specific antibodies as follows: with 1:400 titer 16 cases (1.02%), with 1:800 titer 20 cases (1.27%), with 1:1600 titer 16 cases (1.02%) whereas only one subject (0.06%) showed titers of ≥ 1:3200. There was no significant association between VL seropositivity and gender, age group and occupation. Binary logistic regression showed that rural areas was 0.44 times at higher risk of infection than urban areas (OR= 0.44; %95 CI= 0.25- 0.78). Although the seroprevalence of VL is relatively low in Qom Province, yet due to the importance of the disease, the surveillance system should be monitored by health authorities.
Serological Survey and Associated Risk Factors of Visceral Leish-maniasis in Qom Province, Central Iran

PubMed Central

RAKHSHANPOUR, Arash; MOHEBALI, Mehdi; AKHONDI, Behnaz; RAHIMI, Mohammad Taghi; ROKNI, Mohammad Bagher

2014-01-01

Abstract Background Visceral leishmaniasis (VL) or kala-azar is considered as a parasitic disease caused by the species of Leishmania donovani complex which is intracellular parasites. This systemic disease is endemic in some parts of provinc-es of Iran. The aim of this study was to determine the seroprevalence of VL in Qom Province, central Iran using di-rect agglutination test (DAT). Methods Overall, 1564 serum samples (800 males and 764 females) were collected from selected subjects by random-ized cluster sampling in 2011-2012. Sera were tested and analyzed by DAT. Before sampling; a questionnaire was filled out for each case. Data were analyzed using Chi-square test and multivariate logistic regression for risk factors analysis. Results Of 1564 individuals, 53 cases (3.38%) showed Leishmania specific antibodies as follows: with 1:400 titer 16 cases (1.02%), with 1:800 titer 20 cases (1.27%), with 1:1600 titer 16 cases (1.02%) whereas only one subject (0.06%) showed titers of ≥ 1:3200. There was no significant association between VL seropositivity and gender, age group and occupation. Binary logistic regression showed that rural areas was 0.44 times at higher risk of infection than urban areas (OR= 0.44; %95 CI= 0.25- 0.78). Conclusion Although the seroprevalence of VL is relatively low in Qom Province, yet due to the importance of the disease, the surveillance system should be monitored by health authorities. PMID:26060679
Sperm function and assisted reproduction technology

PubMed Central

MAAß, GESA; BÖDEKER, ROLF‐HASSO; SCHEIBELHUT, CHRISTINE; STALF, THOMAS; MEHNERT, CLAAS; SCHUPPE, HANS‐CHRISTIAN; JUNG, ANDREAS; SCHILL, WOLF‐BERNHARD

2005-01-01

The evaluation of different functional sperm parameters has become a tool in andrological diagnosis. These assays determine the sperm's capability to fertilize an oocyte. It also appears that sperm functions and semen parameters are interrelated and interdependent. Therefore, the question arose whether a given laboratory test or a battery of tests can predict the outcome in in vitro fertilization (IVF). One‐hundred and sixty‐one patients who underwent an IVF treatment were selected from a database of 4178 patients who had been examined for male infertility 3 months before or after IVF. Sperm concentration, motility, acrosin activity, acrosome reaction, sperm morphology, maternal age, number of transferred embryos, embryo score, fertilization rate and pregnancy rate were determined. In addition, logistic regression models to describe fertilization rate and pregnancy were developed. All the parameters in the models were dichotomized and intra‐ and interindividual variability of the parameters were assessed. Although the sperm parameters showed good correlations with IVF when correlated separately, the only essential parameter in the multivariate model was morphology. The enormous intra‐ and interindividual variability of the values was striking. In conclusion, our data indicate that the andrological status at the end of the respective treatment does not necessarily represent the status at the time of IVF. Despite a relatively low correlation coefficient in the logistic regression model, it appears that among the parameters tested, the most reliable parameter to predict fertilization is normal sperm morphology. (Reprod Med Biol 2005; 4: 7–30) PMID:29699207
Roadside sobriety tests and attitudes toward a regulated cannabis market

PubMed Central

Looby, Alison; Earleywine, Mitch; Gieringer, Dale

2007-01-01

Background Many argue that prohibition creates more troubles than alternative policies, but fewer than half of American voters support a taxed and regulated market for cannabis. Some oppose a regulated market because of concerns about driving after smoking cannabis. Although a roadside sobriety test for impairment exists, few voters know about it. The widespread use of a roadside sobriety test that could detect recent cannabis use might lead some voters who currently oppose a regulated market to support it. In contrast, a question that primes respondents about the potential for driving after cannabis use might lead respondents to be less likely to support a regulated market. Methods Phone interviews with a national sample of 1002 registered voters asked about support for a regulated cannabis market and support for such a market if a reliable roadside sobriety test were widely available. Results In this sample of registered voters, 36% supported a regulated cannabis market. Exploratory chi-square tests revealed significantly higher support among men and Caucasians but no link to age or education. These demographic variables covaried significantly. Logistic regression revealed that gender, ethnicity, and political party were significant when all predictors were included. Support increased significantly with a reliable roadside sobriety test to 44%, but some respondents who had agreed to the regulated market no longer agreed when the sobriety test was mentioned. Logistic regression revealed that ethnicity and political affiliation were again significant predictors of support with a reliable sobriety test, but gender was no longer significant. None of these demographic variables could identify who would change their votes in response to the reliable roadside test. Conclusion Increased awareness and use of roadside sobriety tests that detect recent cannabis use could increase support for a regulated cannabis market. Identifying concerns of voters who are not Caucasian or Democrats could help alter cannabis policy. PMID:17266759
Multiple balance tests improve the assessment of postural stability in subjects with Parkinson's disease

PubMed Central

Jacobs, J V; Horak, F B; Tran, V K; Nutt, J G

2006-01-01

Objectives Clinicians often base the implementation of therapies on the presence of postural instability in subjects with Parkinson's disease (PD). These decisions are frequently based on the pull test from the Unified Parkinson's Disease Rating Scale (UPDRS). We sought to determine whether combining the pull test, the one‐leg stance test, the functional reach test, and UPDRS items 27–29 (arise from chair, posture, and gait) predicts balance confidence and falling better than any test alone. Methods The study included 67 subjects with PD. Subjects performed the one‐leg stance test, the functional reach test, and the UPDRS motor exam. Subjects also responded to the Activities‐specific Balance Confidence (ABC) scale and reported how many times they fell during the previous year. Regression models determined the combination of tests that optimally predicted mean ABC scores or categorised fall frequency. Results When all tests were included in a stepwise linear regression, only gait (UPDRS item 29), the pull test (UPDRS item 30), and the one‐leg stance test, in combination, represented significant predictor variables for mean ABC scores (r2 = 0.51). A multinomial logistic regression model including the one‐leg stance test and gait represented the model with the fewest significant predictor variables that correctly identified the most subjects as fallers or non‐fallers (85% of subjects were correctly identified). Conclusions Multiple balance tests (including the one‐leg stance test, and the gait and pull test items of the UPDRS) that assess different types of postural stress provide an optimal assessment of postural stability in subjects with PD. PMID:16484639
A large-scale assessment of two-way SNP interactions in breast cancer susceptibility using 46 450 cases and 42 461 controls from the breast cancer association consortium

PubMed Central

Milne, Roger L.; Herranz, Jesús; Michailidou, Kyriaki; Dennis, Joe; Tyrer, Jonathan P.; Zamora, M. Pilar; Arias-Perez, José Ignacio; González-Neira, Anna; Pita, Guillermo; Alonso, M. Rosario; Wang, Qin; Bolla, Manjeet K.; Czene, Kamila; Eriksson, Mikael; Humphreys, Keith; Darabi, Hatef; Li, Jingmei; Anton-Culver, Hoda; Neuhausen, Susan L.; Ziogas, Argyrios; Clarke, Christina A.; Hopper, John L.; Dite, Gillian S.; Apicella, Carmel; Southey, Melissa C.; Chenevix-Trench, Georgia; Swerdlow, Anthony; Ashworth, Alan; Orr, Nicholas; Schoemaker, Minouk; Jakubowska, Anna; Lubinski, Jan; Jaworska-Bieniek, Katarzyna; Durda, Katarzyna; Andrulis, Irene L.; Knight, Julia A.; Glendon, Gord; Mulligan, Anna Marie; Bojesen, Stig E.; Nordestgaard, Børge G.; Flyger, Henrik; Nevanlinna, Heli; Muranen, Taru A.; Aittomäki, Kristiina; Blomqvist, Carl; Chang-Claude, Jenny; Rudolph, Anja; Seibold, Petra; Flesch-Janys, Dieter; Wang, Xianshu; Olson, Janet E.; Vachon, Celine; Purrington, Kristen; Winqvist, Robert; Pylkäs, Katri; Jukkola-Vuorinen, Arja; Grip, Mervi; Dunning, Alison M.; Shah, Mitul; Guénel, Pascal; Truong, Thérèse; Sanchez, Marie; Mulot, Claire; Brenner, Hermann; Dieffenbach, Aida Karina; Arndt, Volker; Stegmaier, Christa; Lindblom, Annika; Margolin, Sara; Hooning, Maartje J.; Hollestelle, Antoinette; Collée, J. Margriet; Jager, Agnes; Cox, Angela; Brock, Ian W.; Reed, Malcolm W.R.; Devilee, Peter; Tollenaar, Robert A.E.M.; Seynaeve, Caroline; Haiman, Christopher A.; Henderson, Brian E.; Schumacher, Fredrick; Le Marchand, Loic; Simard, Jacques; Dumont, Martine; Soucy, Penny; Dörk, Thilo; Bogdanova, Natalia V.; Hamann, Ute; Försti, Asta; Rüdiger, Thomas; Ulmer, Hans-Ulrich; Fasching, Peter A.; Häberle, Lothar; Ekici, Arif B.; Beckmann, Matthias W.; Fletcher, Olivia; Johnson, Nichola; dos Santos Silva, Isabel; Peto, Julian; Radice, Paolo; Peterlongo, Paolo; Peissel, Bernard; Mariani, Paolo; Giles, Graham G.; Severi, Gianluca; Baglietto, Laura; Sawyer, Elinor; Tomlinson, Ian; Kerin, Michael; Miller, Nicola; Marme, Federik; Burwinkel, Barbara; Mannermaa, Arto; Kataja, Vesa; Kosma, Veli-Matti; Hartikainen, Jaana M.; Lambrechts, Diether; Yesilyurt, Betul T.; Floris, Giuseppe; Leunen, Karin; Alnæs, Grethe Grenaker; Kristensen, Vessela; Børresen-Dale, Anne-Lise; García-Closas, Montserrat; Chanock, Stephen J.; Lissowska, Jolanta; Figueroa, Jonine D.; Schmidt, Marjanka K.; Broeks, Annegien; Verhoef, Senno; Rutgers, Emiel J.; Brauch, Hiltrud; Brüning, Thomas; Ko, Yon-Dschun; Couch, Fergus J.; Toland, Amanda E.; Yannoukakos, Drakoulis; Pharoah, Paul D.P.; Hall, Per; Benítez, Javier; Malats, Núria; Easton, Douglas F.

2014-01-01

Part of the substantial unexplained familial aggregation of breast cancer may be due to interactions between common variants, but few studies have had adequate statistical power to detect interactions of realistic magnitude. We aimed to assess all two-way interactions in breast cancer susceptibility between 70 917 single nucleotide polymorphisms (SNPs) selected primarily based on prior evidence of a marginal effect. Thirty-eight international studies contributed data for 46 450 breast cancer cases and 42 461 controls of European origin as part of a multi-consortium project (COGS). First, SNPs were preselected based on evidence (P < 0.01) of a per-allele main effect, and all two-way combinations of those were evaluated by a per-allele (1 d.f.) test for interaction using logistic regression. Second, all 2.5 billion possible two-SNP combinations were evaluated using Boolean operation-based screening and testing, and SNP pairs with the strongest evidence of interaction (P < 10−4) were selected for more careful assessment by logistic regression. Under the first approach, 3277 SNPs were preselected, but an evaluation of all possible two-SNP combinations (1 d.f.) identified no interactions at P < 10−8. Results from the second analytic approach were consistent with those from the first (P > 10−10). In summary, we observed little evidence of two-way SNP interactions in breast cancer susceptibility, despite the large number of SNPs with potential marginal effects considered and the very large sample size. This finding may have important implications for risk prediction, simplifying the modelling required. Further comprehensive, large-scale genome-wide interaction studies may identify novel interacting loci if the inherent logistic and computational challenges can be overcome. PMID:24242184
A large-scale assessment of two-way SNP interactions in breast cancer susceptibility using 46,450 cases and 42,461 controls from the breast cancer association consortium.

PubMed

Milne, Roger L; Herranz, Jesús; Michailidou, Kyriaki; Dennis, Joe; Tyrer, Jonathan P; Zamora, M Pilar; Arias-Perez, José Ignacio; González-Neira, Anna; Pita, Guillermo; Alonso, M Rosario; Wang, Qin; Bolla, Manjeet K; Czene, Kamila; Eriksson, Mikael; Humphreys, Keith; Darabi, Hatef; Li, Jingmei; Anton-Culver, Hoda; Neuhausen, Susan L; Ziogas, Argyrios; Clarke, Christina A; Hopper, John L; Dite, Gillian S; Apicella, Carmel; Southey, Melissa C; Chenevix-Trench, Georgia; Swerdlow, Anthony; Ashworth, Alan; Orr, Nicholas; Schoemaker, Minouk; Jakubowska, Anna; Lubinski, Jan; Jaworska-Bieniek, Katarzyna; Durda, Katarzyna; Andrulis, Irene L; Knight, Julia A; Glendon, Gord; Mulligan, Anna Marie; Bojesen, Stig E; Nordestgaard, Børge G; Flyger, Henrik; Nevanlinna, Heli; Muranen, Taru A; Aittomäki, Kristiina; Blomqvist, Carl; Chang-Claude, Jenny; Rudolph, Anja; Seibold, Petra; Flesch-Janys, Dieter; Wang, Xianshu; Olson, Janet E; Vachon, Celine; Purrington, Kristen; Winqvist, Robert; Pylkäs, Katri; Jukkola-Vuorinen, Arja; Grip, Mervi; Dunning, Alison M; Shah, Mitul; Guénel, Pascal; Truong, Thérèse; Sanchez, Marie; Mulot, Claire; Brenner, Hermann; Dieffenbach, Aida Karina; Arndt, Volker; Stegmaier, Christa; Lindblom, Annika; Margolin, Sara; Hooning, Maartje J; Hollestelle, Antoinette; Collée, J Margriet; Jager, Agnes; Cox, Angela; Brock, Ian W; Reed, Malcolm W R; Devilee, Peter; Tollenaar, Robert A E M; Seynaeve, Caroline; Haiman, Christopher A; Henderson, Brian E; Schumacher, Fredrick; Le Marchand, Loic; Simard, Jacques; Dumont, Martine; Soucy, Penny; Dörk, Thilo; Bogdanova, Natalia V; Hamann, Ute; Försti, Asta; Rüdiger, Thomas; Ulmer, Hans-Ulrich; Fasching, Peter A; Häberle, Lothar; Ekici, Arif B; Beckmann, Matthias W; Fletcher, Olivia; Johnson, Nichola; dos Santos Silva, Isabel; Peto, Julian; Radice, Paolo; Peterlongo, Paolo; Peissel, Bernard; Mariani, Paolo; Giles, Graham G; Severi, Gianluca; Baglietto, Laura; Sawyer, Elinor; Tomlinson, Ian; Kerin, Michael; Miller, Nicola; Marme, Federik; Burwinkel, Barbara; Mannermaa, Arto; Kataja, Vesa; Kosma, Veli-Matti; Hartikainen, Jaana M; Lambrechts, Diether; Yesilyurt, Betul T; Floris, Giuseppe; Leunen, Karin; Alnæs, Grethe Grenaker; Kristensen, Vessela; Børresen-Dale, Anne-Lise; García-Closas, Montserrat; Chanock, Stephen J; Lissowska, Jolanta; Figueroa, Jonine D; Schmidt, Marjanka K; Broeks, Annegien; Verhoef, Senno; Rutgers, Emiel J; Brauch, Hiltrud; Brüning, Thomas; Ko, Yon-Dschun; Couch, Fergus J; Toland, Amanda E; Yannoukakos, Drakoulis; Pharoah, Paul D P; Hall, Per; Benítez, Javier; Malats, Núria; Easton, Douglas F

2014-04-01

Part of the substantial unexplained familial aggregation of breast cancer may be due to interactions between common variants, but few studies have had adequate statistical power to detect interactions of realistic magnitude. We aimed to assess all two-way interactions in breast cancer susceptibility between 70,917 single nucleotide polymorphisms (SNPs) selected primarily based on prior evidence of a marginal effect. Thirty-eight international studies contributed data for 46,450 breast cancer cases and 42,461 controls of European origin as part of a multi-consortium project (COGS). First, SNPs were preselected based on evidence (P < 0.01) of a per-allele main effect, and all two-way combinations of those were evaluated by a per-allele (1 d.f.) test for interaction using logistic regression. Second, all 2.5 billion possible two-SNP combinations were evaluated using Boolean operation-based screening and testing, and SNP pairs with the strongest evidence of interaction (P < 10(-4)) were selected for more careful assessment by logistic regression. Under the first approach, 3277 SNPs were preselected, but an evaluation of all possible two-SNP combinations (1 d.f.) identified no interactions at P < 10(-8). Results from the second analytic approach were consistent with those from the first (P > 10(-10)). In summary, we observed little evidence of two-way SNP interactions in breast cancer susceptibility, despite the large number of SNPs with potential marginal effects considered and the very large sample size. This finding may have important implications for risk prediction, simplifying the modelling required. Further comprehensive, large-scale genome-wide interaction studies may identify novel interacting loci if the inherent logistic and computational challenges can be overcome.
Dysglycemia, Glycemic Variability, and Outcome After Cardiac Arrest and Temperature Management at 33°C and 36°C.

PubMed

Borgquist, Ola; Wise, Matt P; Nielsen, Niklas; Al-Subaie, Nawaf; Cranshaw, Julius; Cronberg, Tobias; Glover, Guy; Hassager, Christian; Kjaergaard, Jesper; Kuiper, Michael; Smid, Ondrej; Walden, Andrew; Friberg, Hans

2017-08-01

Dysglycemia and glycemic variability are associated with poor outcomes in critically ill patients. Targeted temperature management alters blood glucose homeostasis. We investigated the association between blood glucose concentrations and glycemic variability and the neurologic outcomes of patients randomized to targeted temperature management at 33°C or 36°C after cardiac arrest. Post hoc analysis of the multicenter TTM-trial. Primary outcome of this analysis was neurologic outcome after 6 months, referred to as "Cerebral Performance Category." Thirty-six sites in Europe and Australia. All 939 patients with out-of-hospital cardiac arrest of presumed cardiac cause that had been included in the TTM-trial. Targeted temperature management at 33°C or 36°C. Nonparametric tests as well as multiple logistic regression and mixed effects logistic regression models were used. Median glucose concentrations on hospital admission differed significantly between Cerebral Performance Category outcomes (p < 0.0001). Hyper- and hypoglycemia were associated with poor neurologic outcome (p = 0.001 and p = 0.054). In the multiple logistic regression models, the median glycemic level was an independent predictor of poor Cerebral Performance Category (Cerebral Performance Category, 3-5) with an odds ratio (OR) of 1.13 in the adjusted model (p = 0.008; 95% CI, 1.03-1.24). It was also a predictor in the mixed model, which served as a sensitivity analysis to adjust for the multiple time points. The proportion of hyperglycemia was higher in the 33°C group compared with the 36°C group. Higher blood glucose levels at admission and during the first 36 hours, and higher glycemic variability, were associated with poor neurologic outcome and death. More patients in the 33°C treatment arm had hyperglycemia.
Confirming the validity of the CONUT system for early detection and monitoring of clinical undernutrition: comparison with two logistic regression models developed using SGA as the gold standard.

PubMed

González-Madroño, A; Mancha, A; Rodríguez, F J; Culebras, J; de Ulibarri, J I

2012-01-01

To ratify previous validations of the CONUT nutritional screening tool by the development of two probabilistic models using the parameters included in the CONUT, to see if the CONUT´s effectiveness could be improved. It is a two step prospective study. In Step 1, 101 patients were randomly selected, and SGA and CONUT was made. With data obtained an unconditional logistic regression model was developed, and two variants of CONUT were constructed: Model 1 was made by a method of logistic regression. Model 2 was made by dividing the probabilities of undernutrition obtained in model 1 in seven regular intervals. In step 2, 60 patients were selected and underwent the SGA, the original CONUT and the new models developed. The diagnostic efficacy of the original CONUT and the new models was tested by means of ROC curves. Both samples 1 and 2 were put together to measure the agreement degree between the original CONUT and SGA, and diagnostic efficacy parameters were calculated. No statistically significant differences were found between sample 1 and 2, regarding age, sex and medical/surgical distribution and undernutrition rates were similar (over 40%). The AUC for the ROC curves were 0.862 for the original CONUT, and 0.839 and 0.874, for model 1 and 2 respectively. The kappa index for the CONUT and SGA was 0.680. The CONUT, with the original scores assigned by the authors is equally good than mathematical models and thus is a valuable tool, highly useful and efficient for the purpose of Clinical Undernutrition screening.
Development of the Sydney Falls Risk Screening Tool in brain injury rehabilitation: A multisite prospective cohort study.

PubMed

McKechnie, Duncan; Fisher, Murray J; Pryor, Julie; Bonser, Melissa; Jesus, Jhoven De

2018-03-01

To develop a falls risk screening tool (FRST) sensitive to the traumatic brain injury rehabilitation population. Falls are the most frequently recorded patient safety incident within the hospital context. The inpatient traumatic brain injury rehabilitation population is one particular population that has been identified as at high risk of falls. However, no FRST has been developed for this patient population. Consequently in the traumatic brain injury rehabilitation population, there is the real possibility that nurses are using falls risk screening tools that have a poor clinical utility. Multisite prospective cohort study. Univariate and multiple logistic regression modelling techniques (backward elimination, elastic net and hierarchical) were used to examine each variable's association with patients who fell. The resulting FRST's clinical validity was examined. Of the 140 patients in the study, 41 (29%) fell. Through multiple logistic regression modelling, 11 variables were identified as predictors for falls. Using hierarchical logistic regression, five of these were identified for inclusion in the resulting falls risk screening tool: prescribed mobility aid (such as, wheelchair or frame), a fall since admission to hospital, impulsive behaviour, impaired orientation and bladder and/or bowel incontinence. The resulting FRST has good clinical validity (sensitivity = 0.9; specificity = 0.62; area under the curve = 0.87; Youden index = 0.54). The tool was significantly more accurate (p = .037 on DeLong test) in discriminating fallers from nonfallers than the Ontario Modified STRATIFY FRST. A FRST has been developed using a comprehensive statistical framework, and evidence has been provided of this tool's clinical validity. The developed tool, the Sydney Falls Risk Screening Tool, should be considered for use in brain injury rehabilitation populations. © 2017 John Wiley & Sons Ltd.
Contraceptive use before first pregnancy by women in India (2005-2006): determinants and differentials.

PubMed

Pandey, Anjali; Singh, K K

2015-12-29

There exist ample of research literature investigating the various facet of contraceptive use behaviors in India but the use of contraception by married Indian women, prior to having their first pregnancy has been neglected so far. This study attempts to identify the socio demographic determinants and differentials of contraceptive use or non use by a woman in India, before she proceeds to have her first child. The analysis was done using data from the third National Family Health Survey (2005-2006), India. This study utilized information from 54,918 women who ever have been married and whose current age at the time of NFHS-3 survey was 15-34 years. To identify the crucial socio-demographic determinants governing this pioneering behavior, logistic regression technique has been used. Hosmer Lemeshow test and ROC curve analysis was also performed in order to check the fitting of logistic regression model to the data under consideration. Of all the considered explanatory variables religion, caste, education, current age, age at marriage, media exposure and zonal classifications were found to be significantly affecting the study behavior. Place of residence i.e. urban--rural locality came to be insignificant in multivariable logistic regression. In the light of sufficient evidences confirming the presence of early marriages and child bearing practices in India, conjunct efforts are required to address the socio demographic differentials in contraceptive use by the young married women prior to their first pregnancy. Encouraging women to opt for higher education, ensuring marriages only after legal minimum age at marriage and promoting the family planning programs via print and electronic media may address the existing socio economic barriers. Also, the family planning programs should be oriented to take care of the geographical variations in the study behavior.
Factors determining anti-poliovirus type 3 antibodies among orally immunised Indian infants.

PubMed

Kaliappan, Saravanakumar Puthupalayam; Venugopal, Srinivasan; Giri, Sidhartha; Praharaj, Ira; Karthikeyan, Arun S; Babji, Sudhir; John, Jacob; Muliyil, Jayaprakash; Grassly, Nicholas; Kang, Gagandeep

2016-09-22

Among the three poliovirus serotypes, the lowest responses after vaccination with trivalent oral polio vaccine (tOPV) are to serotype 3. Although improvements in routine immunisation and supplementary immunisation activities have greatly increased vaccine coverage, there are limited data on antibody prevalence in Indian infants. Children aged 5-11months with a history of not having received inactivated polio vaccine were screened for serum antibodies to poliovirus serotype 3 (PV3) by a micro-neutralisation assay according to a modified World Health Organization (WHO) protocol. Limited demographic information was collected to assess risk-factors for a lack of protective antibodies. Student's t-test, logistic regression and multilevel logistic regression (MLR) model were used to estimate model parameters. Of 8454 children screened at a mean age of 8.3 (standard deviation [SD]-1.8) months, 88.1% (95% confidence interval (CI): 87.4-88.8) had protective antibodies to PV3. The number of tOPV doses received was the main determinant of seroprevalence; the maximum likelihood estimate yields a 37.7% (95% CI: 36.2-38.3) increase in seroprevalence per dose of tOPV. In multivariable logistic regression analysis increasing age, male sex, and urban residence were also independently associated with seropositivity (Odds Ratios (OR): 1.17 (95% CI: 1.12-1.23) per month of age, 1.27 (1.11-1.46) and 1.24 (1.05-1.45) respectively). Seroprevalence of antibodies to PV3 is associated with age, gender and place of residence, in addition to the number of tOPV doses received. Ensuring high coverage and monitoring of response are essential as long as oral vaccines are used in polio eradication. Copyright © 2016 The Author(s). Published by Elsevier Ltd.. All rights reserved.

Some links on this page may take you to non-federal websites. Their policies may differ from this site.