regression analysis yielded: Topics by WorldWideScience.org

Sample records for regression analysis yielded

Principal component regression for crop yield estimation

CERN Document Server

Suryanarayana, T M V

2016-01-01

This book highlights the estimation of crop yield in Central Gujarat, especially with regard to the development of Multiple Regression Models and Principal Component Regression (PCR) models using climatological parameters as independent variables and crop yield as a dependent variable. It subsequently compares the multiple linear regression (MLR) and PCR results, and discusses the significance of PCR for crop yield estimation. In this context, the book also covers Principal Component Analysis (PCA), a statistical procedure used to reduce a number of correlated variables into a smaller number of uncorrelated variables called principal components (PC). This book will be helpful to the students and researchers, starting their works on climate and agriculture, mainly focussing on estimation models. The flow of chapters takes the readers in a smooth path, in understanding climate and weather and impact of climate change, and gradually proceeds towards downscaling techniques and then finally towards development of ...
[Milk yield and environmental factors: Multiple regression analysis of the association between milk yield and udder health, fertility data and replacement rate].

Science.gov (United States)

Fölsche, C; Staufenbiel, R

2014-01-01

The relationship between milk yield and both fertility and general animal health in dairy herds is discussed from opposing viewpoints. The hypothesis (1) that raising the herd milk yield would decrease fertility results, the number of milk cells as an indicator for udder health and the replacement rate as a global indicator for animal health as well as increasing the occurrence of specific diseases as a herd problem was compared to the opposing hypotheses that there is no relationship (2) or that there is a differentiated and changing relationship (3). A total of 743 herd examinations, considered independent, were performed in 489 herds between 1995 and 2010. The milk yield, fertility rate, milk cell count, replacement rate, categorized herd problems and management information were recorded. The relationship between the milk yield and both the fertility data and animal health was evaluated using simple and multiple regression analyses. The period between calving and the first service displayed no significant relationship to the herd milk yield. Simple regression analysis showed that the period between calving and gestation, the calving interval and the insemination number were significantly positively associated with the herd milk yield. This positive correlation was lost in multiple regression analysis. The milk cell count and replacement rate using both the simple and multiple regression analyses displayed a significant negative relationship to the milk yield. The alternative hypothesis (3) was confirmed. A higher milk yield has no negative influence on the milk cell count and the replacement rate in terms of the udder and general health. When parameterizing the fertility, the herd milk yield should be considered. Extending the resting time may increase the milk yield while preventing a decline in the insemination index.
Relationship between rice yield and climate variables in southwest Nigeria using multiple linear regression and support vector machine analysis

Science.gov (United States)

Oguntunde, Philip G.; Lischeid, Gunnar; Dietrich, Ottfried

2018-03-01

This study examines the variations of climate variables and rice yield and quantifies the relationships among them using multiple linear regression, principal component analysis, and support vector machine (SVM) analysis in southwest Nigeria. The climate and yield data used was for a period of 36 years between 1980 and 2015. Similar to the observed decrease ( P 1 and explained 83.1% of the total variance of predictor variables. The SVM regression function using the scores of the first principal component explained about 75% of the variance in rice yield data and linear regression about 64%. SVM regression between annual solar radiation values and yield explained 67% of the variance. Only the first component of the principal component analysis (PCA) exhibited a clear long-term trend and sometimes short-term variance similar to that of rice yield. Short-term fluctuations of the scores of the PC1 are closely coupled to those of rice yield during the 1986-1993 and the 2006-2013 periods thereby revealing the inter-annual sensitivity of rice production to climate variability. Solar radiation stands out as the climate variable of highest influence on rice yield, and the influence was especially strong during monsoon and post-monsoon periods, which correspond to the vegetative, booting, flowering, and grain filling stages in the study area. The outcome is expected to provide more in-depth regional-specific climate-rice linkage for screening of better cultivars that can positively respond to future climate fluctuations as well as providing information that may help optimized planting dates for improved radiation use efficiency in the study area.
Regression Association Analysis of Yield-Related Traits with RAPD Molecular Markers in Pistachio (Pistacia vera L.

Directory of Open Access Journals (Sweden)

Saeid Mirzaei

2017-10-01

Full Text Available Introduction: The pistachio (Pistacia vera, a member of the cashew family, is a small tree originating from Central Asia and the Middle East. The tree produces seeds that are widely consumed as food. Pistacia vera often is confused with other species in the genus Pistacia that are also known as pistachio. These other species can be distinguished by their geographic distributions and their seeds which are much smaller and have a soft shell. Continual advances in crop improvement through plant breeding are driven by the available genetic diversity. Therefore, the recognition and measurement of such diversity is crucial to breeding programs. In the past 20 years, the major effort in plant breeding has changed from quantitative to molecular genetics with emphasis on quantitative trait loci (QTL identification and marker assisted selection (MAS. The germplasm-regression-combined association studies not only allow mapping of genes/QTLs with higher level of confidence, but also allow detection of genes/QTLs, which will otherwise escape detection in linkage-based QTL studies based on the planned populations. The development of the marker-based technology offers a fast, reliable, and easy way to perform multiple regression analysis and comprise an alternative approach to breeding in diverse species of plants. The availability of many makers and morphological traits can help to regression analysis between these markers and morphological traits. Materials and Methods: In this study, 20 genotypes of Pistachio were studied and yield related traits were measured. Young well-expanded leaves were collected for DNA extraction and total genomic DNA was extracted. Genotyping was performed using 15 RAPD primers and PCR amplification products were visualized by gel electrophoresis. The reproducible RAPD fragments were scored on the basis of present (1 or absent (0 bands and a binary matrix constructed using each molecular marker. Association analysis between
Relationship between rice yield and climate variables in southwest Nigeria using multiple linear regression and support vector machine analysis.

Science.gov (United States)

Oguntunde, Philip G; Lischeid, Gunnar; Dietrich, Ottfried

2018-03-01

This study examines the variations of climate variables and rice yield and quantifies the relationships among them using multiple linear regression, principal component analysis, and support vector machine (SVM) analysis in southwest Nigeria. The climate and yield data used was for a period of 36 years between 1980 and 2015. Similar to the observed decrease (P 1 and explained 83.1% of the total variance of predictor variables. The SVM regression function using the scores of the first principal component explained about 75% of the variance in rice yield data and linear regression about 64%. SVM regression between annual solar radiation values and yield explained 67% of the variance. Only the first component of the principal component analysis (PCA) exhibited a clear long-term trend and sometimes short-term variance similar to that of rice yield. Short-term fluctuations of the scores of the PC1 are closely coupled to those of rice yield during the 1986-1993 and the 2006-2013 periods thereby revealing the inter-annual sensitivity of rice production to climate variability. Solar radiation stands out as the climate variable of highest influence on rice yield, and the influence was especially strong during monsoon and post-monsoon periods, which correspond to the vegetative, booting, flowering, and grain filling stages in the study area. The outcome is expected to provide more in-depth regional-specific climate-rice linkage for screening of better cultivars that can positively respond to future climate fluctuations as well as providing information that may help optimized planting dates for improved radiation use efficiency in the study area.
Geoelectrical parameter-based multivariate regression borehole yield model for predicting aquifer yield in managing groundwater resource sustainability

Directory of Open Access Journals (Sweden)

Kehinde Anthony Mogaji

2016-07-01

Full Text Available This study developed a GIS-based multivariate regression (MVR yield rate prediction model of groundwater resource sustainability in the hard-rock geology terrain of southwestern Nigeria. This model can economically manage the aquifer yield rate potential predictions that are often overlooked in groundwater resources development. The proposed model relates the borehole yield rate inventory of the area to geoelectrically derived parameters. Three sets of borehole yield rate conditioning geoelectrically derived parameters—aquifer unit resistivity (ρ, aquifer unit thickness (D and coefficient of anisotropy (λ—were determined from the acquired and interpreted geophysical data. The extracted borehole yield rate values and the geoelectrically derived parameter values were regressed to develop the MVR relationship model by applying linear regression and GIS techniques. The sensitivity analysis results of the MVR model evaluated at P ⩽ 0.05 for the predictors ρ, D and λ provided values of 2.68 × 10−05, 2 × 10−02 and 2.09 × 10−06, respectively. The accuracy and predictive power tests conducted on the MVR model using the Theil inequality coefficient measurement approach, coupled with the sensitivity analysis results, confirmed the model yield rate estimation and prediction capability. The MVR borehole yield prediction model estimates were processed in a GIS environment to model an aquifer yield potential prediction map of the area. The information on the prediction map can serve as a scientific basis for predicting aquifer yield potential rates relevant in groundwater resources sustainability management. The developed MVR borehole yield rate prediction mode provides a good alternative to other methods used for this purpose.
A Technique of Fuzzy C-Mean in Multiple Linear Regression Model toward Paddy Yield

Science.gov (United States)

Syazwan Wahab, Nur; Saifullah Rusiman, Mohd; Mohamad, Mahathir; Amira Azmi, Nur; Che Him, Norziha; Ghazali Kamardan, M.; Ali, Maselan

2018-04-01

In this paper, we propose a hybrid model which is a combination of multiple linear regression model and fuzzy c-means method. This research involved a relationship between 20 variates of the top soil that are analyzed prior to planting of paddy yields at standard fertilizer rates. Data used were from the multi-location trials for rice carried out by MARDI at major paddy granary in Peninsular Malaysia during the period from 2009 to 2012. Missing observations were estimated using mean estimation techniques. The data were analyzed using multiple linear regression model and a combination of multiple linear regression model and fuzzy c-means method. Analysis of normality and multicollinearity indicate that the data is normally scattered without multicollinearity among independent variables. Analysis of fuzzy c-means cluster the yield of paddy into two clusters before the multiple linear regression model can be used. The comparison between two method indicate that the hybrid of multiple linear regression model and fuzzy c-means method outperform the multiple linear regression model with lower value of mean square error.
Genetic Parameters for Body condition score, Body weigth, Milk yield and Fertility estimated using random regression models

NARCIS (Netherlands)

Berry, D.P.; Buckley, F.; Dillon, P.; Evans, R.D.; Rath, M.; Veerkamp, R.F.

2003-01-01

Genetic (co)variances between body condition score (BCS), body weight (BW), milk yield, and fertility were estimated using a random regression animal model extended to multivariate analysis. The data analyzed included 81,313 BCS observations, 91,937 BW observations, and 100,458 milk test-day yields
Regression and regression analysis time series prediction modeling on climate data of quetta, pakistan

International Nuclear Information System (INIS)

Jafri, Y.Z.; Kamal, L.

2007-01-01

Various statistical techniques was used on five-year data from 1998-2002 of average humidity, rainfall, maximum and minimum temperatures, respectively. The relationships to regression analysis time series (RATS) were developed for determining the overall trend of these climate parameters on the basis of which forecast models can be corrected and modified. We computed the coefficient of determination as a measure of goodness of fit, to our polynomial regression analysis time series (PRATS). The correlation to multiple linear regression (MLR) and multiple linear regression analysis time series (MLRATS) were also developed for deciphering the interdependence of weather parameters. Spearman's rand correlation and Goldfeld-Quandt test were used to check the uniformity or non-uniformity of variances in our fit to polynomial regression (PR). The Breusch-Pagan test was applied to MLR and MLRATS, respectively which yielded homoscedasticity. We also employed Bartlett's test for homogeneity of variances on a five-year data of rainfall and humidity, respectively which showed that the variances in rainfall data were not homogenous while in case of humidity, were homogenous. Our results on regression and regression analysis time series show the best fit to prediction modeling on climatic data of Quetta, Pakistan. (author)
Plateletpheresis efficiency and mathematical correction of software-derived platelet yield prediction: A linear regression and ROC modeling approach.

Science.gov (United States)

Jaime-Pérez, José Carlos; Jiménez-Castillo, Raúl Alberto; Vázquez-Hernández, Karina Elizabeth; Salazar-Riojas, Rosario; Méndez-Ramírez, Nereida; Gómez-Almaguer, David

2017-10-01

Advances in automated cell separators have improved the efficiency of plateletpheresis and the possibility of obtaining double products (DP). We assessed cell processor accuracy of predicted platelet (PLT) yields with the goal of a better prediction of DP collections. This retrospective proof-of-concept study included 302 plateletpheresis procedures performed on a Trima Accel v6.0 at the apheresis unit of a hematology department. Donor variables, software predicted yield and actual PLT yield were statistically evaluated. Software prediction was optimized by linear regression analysis and its optimal cut-off to obtain a DP assessed by receiver operating characteristic curve (ROC) modeling. Three hundred and two plateletpheresis procedures were performed; in 271 (89.7%) occasions, donors were men and in 31 (10.3%) women. Pre-donation PLT count had the best direct correlation with actual PLT yield (r = 0.486. P Simple correction derived from linear regression analysis accurately corrected this underestimation and ROC analysis identified a precise cut-off to reliably predict a DP. © 2016 Wiley Periodicals, Inc.
Screening for ketosis using multiple logistic regression based on milk yield and composition.

Science.gov (United States)

Kayano, Mitsunori; Kataoka, Tomoko

2015-11-01

Multiple logistic regression was applied to milk yield and composition data for 632 records of healthy cows and 61 records of ketotic cows in Hokkaido, Japan. The purpose was to diagnose ketosis based on milk yield and composition, simultaneously. The cows were divided into two groups: (1) multiparous, including 314 healthy cows and 45 ketotic cows and (2) primiparous, including 318 healthy cows and 16 ketotic cows, since nutritional status, milk yield and composition are affected by parity. Multiple logistic regression was applied to these groups separately. For multiparous cows, milk yield (kg/day/cow) and protein-to-fat (P/F) ratio in milk were significant factors (Pketosis. For primiparous cows, lactose content (%), solid not fat (SNF) content (%) and milk urea nitrogen (MUN) content (mg/dl) were significantly associated with ketosis (Pketosis, provided the sensitivity, specificity and AUC values of (1) 0.711, 0.726 and 0.781; and (2) 0.678, 0.767 and 0.738, respectively.
Climate Impacts on Chinese Corn Yields: A Fractional Polynomial Regression Model

NARCIS (Netherlands)

Kooten, van G.C.; Sun, Baojing

2012-01-01

In this study, we examine the effect of climate on corn yields in northern China using data from ten districts in Inner Mongolia and two in Shaanxi province. A regression model with a flexible functional form is specified, with explanatory variables that include seasonal growing degree days,
Genetic correlations among body condition score, yield, and fertility in first-parity cows estimated by random regression models.

Science.gov (United States)

Veerkamp, R F; Koenen, E P; De Jong, G

2001-10-01

Twenty type classifiers scored body condition (BCS) of 91,738 first-parity cows from 601 sires and 5518 maternal grandsires. Fertility data during first lactation were extracted for 177,220 cows, of which 67,278 also had a BCS observation, and first-lactation 305-d milk, fat, and protein yields were added for 180,631 cows. Heritabilities and genetic correlations were estimated using a sire-maternal grandsire model. Heritability of BCS was 0.38. Heritabilities for fertility traits were low (0.01 to 0.07), but genetic standard deviations were substantial, 9 d for days to first service and calving interval, 0.25 for number of services, and 5% for first-service conception. Phenotypic correlations between fertility and yield or BCS were small (-0.15 to 0.20). Genetic correlations between yield and all fertility traits were unfavorable (0.37 to 0.74). Genetic correlations with BCS were between -0.4 and -0.6 for calving interval and days to first service. Random regression analysis (RR) showed that correlations changed with days in milk for BCS. Little agreement was found between variances and correlations from RR, and analysis including a single month (mo 1 to 10) of data for BCS, especially during early and late lactation. However, this was due to excluding data from the conventional analysis, rather than due to the polynomials used. RR and a conventional five-traits model where BCS in mo 1, 4, 7, and 10 was treated as a separate traits (plus yield or fertility) gave similar results. Thus a parsimonious random regression model gave more realistic estimates for the (co)variances than a series of bivariate analysis on subsets of the data for BCS. A higher genetic merit for yield has unfavorable effects on fertility, but the genetic correlation suggests that BCS (at some stages of lactation) might help to alleviate the unfavorable effect of selection for higher yield on fertility.
Genetic parameters for body condition score, body weight, milk yield, and fertility estimated using random regression models.

Science.gov (United States)

Berry, D P; Buckley, F; Dillon, P; Evans, R D; Rath, M; Veerkamp, R F

2003-11-01

Genetic (co)variances between body condition score (BCS), body weight (BW), milk yield, and fertility were estimated using a random regression animal model extended to multivariate analysis. The data analyzed included 81,313 BCS observations, 91,937 BW observations, and 100,458 milk test-day yields from 8725 multiparous Holstein-Friesian cows. A cubic random regression was sufficient to model the changing genetic variances for BCS, BW, and milk across different days in milk. The genetic correlations between BCS and fertility changed little over the lactation; genetic correlations between BCS and interval to first service and between BCS and pregnancy rate to first service varied from -0.47 to -0.31, and from 0.15 to 0.38, respectively. This suggests that maximum genetic gain in fertility from indirect selection on BCS should be based on measurements taken in midlactation when the genetic variance for BCS is largest. Selection for increased BW resulted in shorter intervals to first service, but more services and poorer pregnancy rates; genetic correlations between BW and pregnancy rate to first service varied from -0.52 to -0.45. Genetic selection for higher lactation milk yield alone through selection on increased milk yield in early lactation is likely to have a more deleterious effect on genetic merit for fertility than selection on higher milk yield in late lactation.
Heterogeneous global crop yield response to biochar: a meta-regression analysis

International Nuclear Information System (INIS)

Crane-Droesch, Andrew; Torn, Margaret S; Abiven, Samuel; Jeffery, Simon

2013-01-01

Biochar may contribute to climate change mitigation at negative cost by sequestering photosynthetically fixed carbon in soil while increasing crop yields. The magnitude of biochar’s potential in this regard will depend on crop yield benefits, which have not been well-characterized across different soils and biochars. Using data from 84 studies, we employ meta-analytical, missing data, and semiparametric statistical methods to explain heterogeneity in crop yield responses across different soils, biochars, and agricultural management factors, and then estimate potential changes in yield across different soil environments globally. We find that soil cation exchange capacity and organic carbon were strong predictors of yield response, with low cation exchange and low carbon associated with positive response. We also find that yield response increases over time since initial application, compared to non-biochar controls. High reported soil clay content and low soil pH were weaker predictors of higher yield response. No biochar parameters in our dataset—biochar pH, percentage carbon content, or temperature of pyrolysis—were significant predictors of yield impacts. Projecting our fitted model onto a global soil database, we find the largest potential increases in areas with highly weathered soils, such as those characterizing much of the humid tropics. Richer soils characterizing much of the world’s important agricultural areas appear to be less likely to benefit from biochar. (letter)
Regression analysis by example

CERN Document Server

Chatterjee, Samprit

2012-01-01

Praise for the Fourth Edition: ""This book is . . . an excellent source of examples for regression analysis. It has been and still is readily readable and understandable."" -Journal of the American Statistical Association Regression analysis is a conceptually simple method for investigating relationships among variables. Carrying out a successful application of regression analysis, however, requires a balance of theoretical results, empirical rules, and subjective judgment. Regression Analysis by Example, Fifth Edition has been expanded
Comparison of Regression Techniques to Predict Response of Oilseed Rape Yield to Variation in Climatic Conditions in Denmark

DEFF Research Database (Denmark)

Sharif, Behzad; Makowski, David; Plauborg, Finn

2017-01-01

Statistical regression models represent alternatives to process-based dynamic models for predicting the response of crop yields to variation in climatic conditions. Regression models can be used to quantify the effect of change in temperature and precipitation on yields. However, it is difficult ...
Exploratory regression analysis: a tool for selecting models and determining predictor importance.

Science.gov (United States)

Braun, Michael T; Oswald, Frederick L

2011-06-01

Linear regression analysis is one of the most important tools in a researcher's toolbox for creating and testing predictive models. Although linear regression analysis indicates how strongly a set of predictor variables, taken together, will predict a relevant criterion (i.e., the multiple R), the analysis cannot indicate which predictors are the most important. Although there is no definitive or unambiguous method for establishing predictor variable importance, there are several accepted methods. This article reviews those methods for establishing predictor importance and provides a program (in Excel) for implementing them (available for direct download at http://dl.dropbox.com/u/2480715/ERA.xlsm?dl=1) . The program investigates all 2(p) - 1 submodels and produces several indices of predictor importance. This exploratory approach to linear regression, similar to other exploratory data analysis techniques, has the potential to yield both theoretical and practical benefits.
Correlation, Regression and Path Analyses of Seed Yield Components in Crambe abyssinica, a Promising Industrial Oil Crop

OpenAIRE

Huang, Banglian; Yang, Yiming; Luo, Tingting; Wu, S.; Du, Xuezhu; Cai, Detian; Loo, van, E.N.; Huang Bangquan

2013-01-01

In the present study correlation, regression and path analyses were carried out to decide correlations among the agro- nomic traits and their contributions to seed yield per plant in Crambe abyssinica. Partial correlation analysis indicated that plant height (X1) was significantly correlated with branching height and the number of first branches (P <0.01); Branching height (X2) was significantly correlated with pod number of primary inflorescence (P <0.01) and number of secondary branch...
Statistical methods in regression and calibration analysis of chromosome aberration data

International Nuclear Information System (INIS)

Merkle, W.

1983-01-01

The method of iteratively reweighted least squares for the regression analysis of Poisson distributed chromosome aberration data is reviewed in the context of other fit procedures used in the cytogenetic literature. As an application of the resulting regression curves methods for calculating confidence intervals on dose from aberration yield are described and compared, and, for the linear quadratic model a confidence interval is given. Emphasis is placed on the rational interpretation and the limitations of various methods from a statistical point of view. (orig./MG)

Principal component regression analysis with SPSS.

Science.gov (United States)

Liu, R X; Kuang, J; Gong, Q; Hou, X L

2003-06-01

The paper introduces all indices of multicollinearity diagnoses, the basic principle of principal component regression and determination of 'best' equation method. The paper uses an example to describe how to do principal component regression analysis with SPSS 10.0: including all calculating processes of the principal component regression and all operations of linear regression, factor analysis, descriptives, compute variable and bivariate correlations procedures in SPSS 10.0. The principal component regression analysis can be used to overcome disturbance of the multicollinearity. The simplified, speeded up and accurate statistical effect is reached through the principal component regression analysis with SPSS.
Predictions of biochar production and torrefaction performance from sugarcane bagasse using interpolation and regression analysis.

Science.gov (United States)

Chen, Wei-Hsin; Hsu, Hung-Jen; Kumar, Gopalakrishnan; Budzianowski, Wojciech M; Ong, Hwai Chyuan

2017-12-01

This study focuses on the biochar formation and torrefaction performance of sugarcane bagasse, and they are predicted using the bilinear interpolation (BLI), inverse distance weighting (IDW) interpolation, and regression analysis. It is found that the biomass torrefied at 275°C for 60min or at 300°C for 30min or longer is appropriate to produce biochar as alternative fuel to coal with low carbon footprint, but the energy yield from the torrefaction at 300°C is too low. From the biochar yield, enhancement factor of HHV, and energy yield, the results suggest that the three methods are all feasible for predicting the performance, especially for the enhancement factor. The power parameter of unity in the IDW method provides the best predictions and the error is below 5%. The second order in regression analysis gives a more reasonable approach than the first order, and is recommended for the predictions. Copyright © 2017 Elsevier Ltd. All rights reserved.
A Mixed Application of Geographically Weighted Regression and Unsupervised Classification for Analyzing Latex Yield Variability in Yunnan, China

Directory of Open Access Journals (Sweden)

Oh Seok Kim

2017-05-01

Full Text Available This paper introduces a mixed method approach for analyzing the determinants of natural latex yields and the associated spatial variations and identifying the most suitable regions for producing latex. Geographically Weighted Regressions (GWR and Iterative Self-Organizing Data Analysis Technique (ISODATA are jointly applied to the georeferenced data points collected from the rubber plantations in Xishuangbanna (in Yunnan province, south China and other remotely-sensed spatial data. According to the GWR models, Age of rubber tree, Percent of clay in soil, Elevation, Solar radiation, Population, Distance from road, Distance from stream, Precipitation, and Mean temperature turn out statistically significant, indicating that these are the major determinants shaping latex yields at the prefecture level. However, the signs and magnitudes of the parameter estimates at the aggregate level are different from those at the lower spatial level, and the differences are due to diverse reasons. The ISODATA classifies the landscape into three categories: high, medium, and low potential yields. The map reveals that Mengla County has the majority of land with high potential yield, while Jinghong City and Menghai County show lower potential yield. In short, the mixed method can offer a means of providing greater insights in the prediction of agricultural production.
Polynomial regression analysis and significance test of the regression function

International Nuclear Information System (INIS)

Gao Zhengming; Zhao Juan; He Shengping

2012-01-01

In order to analyze the decay heating power of a certain radioactive isotope per kilogram with polynomial regression method, the paper firstly demonstrated the broad usage of polynomial function and deduced its parameters with ordinary least squares estimate. Then significance test method of polynomial regression function is derived considering the similarity between the polynomial regression model and the multivariable linear regression model. Finally, polynomial regression analysis and significance test of the polynomial function are done to the decay heating power of the iso tope per kilogram in accord with the authors' real work. (authors)
Applied regression analysis a research tool

CERN Document Server

Pantula, Sastry; Dickey, David

1998-01-01

Least squares estimation, when used appropriately, is a powerful research tool. A deeper understanding of the regression concepts is essential for achieving optimal benefits from a least squares analysis. This book builds on the fundamentals of statistical methods and provides appropriate concepts that will allow a scientist to use least squares as an effective research tool. Applied Regression Analysis is aimed at the scientist who wishes to gain a working knowledge of regression analysis. The basic purpose of this book is to develop an understanding of least squares and related statistical methods without becoming excessively mathematical. It is the outgrowth of more than 30 years of consulting experience with scientists and many years of teaching an applied regression course to graduate students. Applied Regression Analysis serves as an excellent text for a service course on regression for non-statisticians and as a reference for researchers. It also provides a bridge between a two-semester introduction to...
Trochanteric entry femoral nails yield better femoral version and lower revision rates-A large cohort multivariate regression analysis.

Science.gov (United States)

Yoon, Richard S; Gage, Mark J; Galos, David K; Donegan, Derek J; Liporace, Frank A

2017-06-01

Intramedullary nailing (IMN) has become the standard of care for the treatment of most femoral shaft fractures. Different IMN options include trochanteric and piriformis entry as well as retrograde nails, which may result in varying degrees of femoral rotation. The objective of this study was to analyze postoperative femoral version between three types of nails and to delineate any significant differences in femoral version (DFV) and revision rates. Over a 10-year period, 417 patients underwent IMN of a diaphyseal femur fracture (AO/OTA 32A-C). Of these patients, 316 met inclusion criteria and obtained postoperative computed tomography (CT) scanograms to calculate femoral version and were thus included in the study. In this study, our main outcome measure was the difference in femoral version (DFV) between the uninjured limb and the injured limb. The effect of the following variables on DFV and revision rates were determined via univariate, multivariate, and ordinal regression analyses: gender, age, BMI, ethnicity, mechanism of injury, operative side, open fracture, and table type/position. Statistical significance was set at pregression analysis revealed that a lower BMI was significantly associated with a lower DFV (p=0.006). Controlling for possible covariables, multivariate analysis yielded a significantly lower DFV for trochanteric entry nails than piriformis or retrograde nails (7.9±6.10 vs. 9.5±7.4 vs. 9.4±7.8°, pregression analysis. However, this is not to state that the other nail types exhibited abnormal DFV. Translation to the clinical impact of a few degrees of DFV is also unknown. Future studies to more in-depth study the intricacies of femoral version may lead to improved technology in addition to potentially improved clinical outcomes. Copyright © 2017 Elsevier Ltd. All rights reserved.
Genetic variability, partial regression, Co-heritability studies and their implication in selection of high yielding potato gen

International Nuclear Information System (INIS)

Iqbal, Z.M.; Khan, S.A.

2003-01-01

Partial regression coefficient, genotypic and phenotypic variabilities, heritability co-heritability and genetic advance were studied in 15 Potato varieties of exotic and local origin. Both genotypic and phenotypic coefficients of variations were high for scab and rhizoctonia incidence percentage. Significant partial regression coefficient for emergence percentage indicated its relative importance in tuber yield. High heritability (broadsense) estimates coupled with high genetic advance for plant height, number of stems per plant and scab percentage revealed substantial contribution of additive genetic variance in the expression of these traits. Hence, the selection based on these characters could play a significant role in their improvement the dominance and epistatic variance was more important for character expression of yield ha/sup -1/, emergence and rhizoctonia percentage. This phenomenon is mainly due to the accumulative effects of low heritability and low to moderate genetic advance. The high co-heritability coupled with negative genotypic and phenotypic covariance revealed that selection of varieties having low scab and rhizoctonia percentage resulted in more potato yield. (author)
Regression Analysis by Example. 5th Edition

Science.gov (United States)

Chatterjee, Samprit; Hadi, Ali S.

2012-01-01

Regression analysis is a conceptually simple method for investigating relationships among variables. Carrying out a successful application of regression analysis, however, requires a balance of theoretical results, empirical rules, and subjective judgment. "Regression Analysis by Example, Fifth Edition" has been expanded and thoroughly…
Regression analysis with categorized regression calibrated exposure: some interesting findings

Directory of Open Access Journals (Sweden)

Hjartåker Anette

2006-07-01

Full Text Available Abstract Background Regression calibration as a method for handling measurement error is becoming increasingly well-known and used in epidemiologic research. However, the standard version of the method is not appropriate for exposure analyzed on a categorical (e.g. quintile scale, an approach commonly used in epidemiologic studies. A tempting solution could then be to use the predicted continuous exposure obtained through the regression calibration method and treat it as an approximation to the true exposure, that is, include the categorized calibrated exposure in the main regression analysis. Methods We use semi-analytical calculations and simulations to evaluate the performance of the proposed approach compared to the naive approach of not correcting for measurement error, in situations where analyses are performed on quintile scale and when incorporating the original scale into the categorical variables, respectively. We also present analyses of real data, containing measures of folate intake and depression, from the Norwegian Women and Cancer study (NOWAC. Results In cases where extra information is available through replicated measurements and not validation data, regression calibration does not maintain important qualities of the true exposure distribution, thus estimates of variance and percentiles can be severely biased. We show that the outlined approach maintains much, in some cases all, of the misclassification found in the observed exposure. For that reason, regression analysis with the corrected variable included on a categorical scale is still biased. In some cases the corrected estimates are analytically equal to those obtained by the naive approach. Regression calibration is however vastly superior to the naive method when applying the medians of each category in the analysis. Conclusion Regression calibration in its most well-known form is not appropriate for measurement error correction when the exposure is analyzed on a
Vectors, a tool in statistical regression theory

NARCIS (Netherlands)

Corsten, L.C.A.

1958-01-01

Using linear algebra this thesis developed linear regression analysis including analysis of variance, covariance analysis, special experimental designs, linear and fertility adjustments, analysis of experiments at different places and times. The determination of the orthogonal projection, yielding
Gaussian process regression analysis for functional data

CERN Document Server

Shi, Jian Qing

2011-01-01

Gaussian Process Regression Analysis for Functional Data presents nonparametric statistical methods for functional regression analysis, specifically the methods based on a Gaussian process prior in a functional space. The authors focus on problems involving functional response variables and mixed covariates of functional and scalar variables.Covering the basics of Gaussian process regression, the first several chapters discuss functional data analysis, theoretical aspects based on the asymptotic properties of Gaussian process regression models, and new methodological developments for high dime
Determining Balıkesir’s Energy Potential Using a Regression Analysis Computer Program

Directory of Open Access Journals (Sweden)

Bedri Yüksel

2014-01-01

Full Text Available Solar power and wind energy are used concurrently during specific periods, while at other times only the more efficient is used, and hybrid systems make this possible. When establishing a hybrid system, the extent to which these two energy sources support each other needs to be taken into account. This paper is a study of the effects of wind speed, insolation levels, and the meteorological parameters of temperature and humidity on the energy potential in Balıkesir, in the Marmara region of Turkey. The relationship between the parameters was studied using a multiple linear regression method. Using a designed-for-purpose computer program, two different regression equations were derived, with wind speed being the dependent variable in the first and insolation levels in the second. The regression equations yielded accurate results. The computer program allowed for the rapid calculation of different acceptance rates. The results of the statistical analysis proved the reliability of the equations. An estimate of identified meteorological parameters and unknown parameters could be produced with a specified precision by using the regression analysis method. The regression equations also worked for the evaluation of energy potential.
Path and ridge regression analysis of seed yield and seed yield components of Russian wildrye (Psathyrostachys juncea Nevski) under field conditions

DEFF Research Database (Denmark)

Wang, Quanzhen; Zhang, Tiejun; Cui, Jian

2011-01-01

The correlations among seed yield components, and their direct and indirect effects on the seed yield (Z) of Russina wildrye (Psathyrostachys juncea Nevski) were investigated. The seed yield components: fertile tillers m-2 (Y1), spikelets per fertile tillers (Y2), florets per spikelet- (Y3), seed...
Regression Analysis and the Sociological Imagination

Science.gov (United States)

De Maio, Fernando

2014-01-01

Regression analysis is an important aspect of most introductory statistics courses in sociology but is often presented in contexts divorced from the central concerns that bring students into the discipline. Consequently, we present five lesson ideas that emerge from a regression analysis of income inequality and mortality in the USA and Canada.
Climate-based statistical regression models for crop yield forecasting of coffee in humid tropical Kerala, India

Science.gov (United States)

Jayakumar, M.; Rajavel, M.; Surendran, U.

2016-12-01

A study on the variability of coffee yield of both Coffea arabica and Coffea canephora as influenced by climate parameters (rainfall (RF), maximum temperature (Tmax), minimum temperature (Tmin), and mean relative humidity (RH)) was undertaken at Regional Coffee Research Station, Chundale, Wayanad, Kerala State, India. The result on the coffee yield data of 30 years (1980 to 2009) revealed that the yield of coffee is fluctuating with the variations in climatic parameters. Among the species, productivity was higher for C. canephora coffee than C. arabica in most of the years. Maximum yield of C. canephora (2040 kg ha-1) was recorded in 2003-2004 and there was declining trend of yield noticed in the recent years. Similarly, the maximum yield of C. arabica (1745 kg ha-1) was recorded in 1988-1989 and decreased yield was noticed in the subsequent years till 1997-1998 due to year to year variability in climate. The highest correlation coefficient was found between the yield of C. arabica coffee and maximum temperature during January (0.7) and between C. arabica coffee yield and RH during July (0.4). Yield of C. canephora coffee had highest correlation with maximum temperature, RH and rainfall during February. Statistical regression model between selected climatic parameters and yield of C. arabica and C. canephora coffee was developed to forecast the yield of coffee in Wayanad district in Kerala. The model was validated for years 2010, 2011, and 2012 with the coffee yield data obtained during the years and the prediction was found to be good.
Polylinear regression analysis in radiochemistry

International Nuclear Information System (INIS)

Kopyrin, A.A.; Terent'eva, T.N.; Khramov, N.N.

1995-01-01

A number of radiochemical problems have been formulated in the framework of polylinear regression analysis, which permits the use of conventional mathematical methods for their solution. The authors have considered features of the use of polylinear regression analysis for estimating the contributions of various sources to the atmospheric pollution, for studying irradiated nuclear fuel, for estimating concentrations from spectral data, for measuring neutron fields of a nuclear reactor, for estimating crystal lattice parameters from X-ray diffraction patterns, for interpreting data of X-ray fluorescence analysis, for estimating complex formation constants, and for analyzing results of radiometric measurements. The problem of estimating the target parameters can be incorrect at certain properties of the system under study. The authors showed the possibility of regularization by adding a fictitious set of data open-quotes obtainedclose quotes from the orthogonal design. To estimate only a part of the parameters under consideration, the authors used incomplete rank models. In this case, it is necessary to take into account the possibility of confounding estimates. An algorithm for evaluating the degree of confounding is presented which is realized using standard software or regression analysis
Genetic Analysis of Milk Yield Using Random Regression Test Day Model in Tehran Province Holstein Dairy Cow

Directory of Open Access Journals (Sweden)

A. Seyeddokht

2012-09-01

Full Text Available In this research a random regression test day model was used to estimate heritability values and calculation genetic correlations between test day milk records. a total of 140357 monthly test day milk records belonging to 28292 first lactation Holstein cattle(trice time a day milking distributed in 165 herd and calved from 2001 to 2010 belonging to the herds of Tehran province were used. The fixed effects of herd-year-month of calving as contemporary group and age at calving and Holstein gene percentage as covariate were fitted. Orthogonal legendre polynomial with a 4th-order was implemented to take account of genetic and environmental aspects of milk production over the course of lactation. RRM using Legendre polynomials as base functions appears to be the most adequate to describe the covariance structure of the data. The results showed that the average of heritability for the second half of lactation period was higher than that of the first half. The heritability value for the first month was lowest (0.117 and for the eighth month of the lactation was highest (0.230 compared to the other months of lactation. Because of genetic variation was increased gradually, and residual variance was high in the first months of lactation, heritabilities were different over the course of lactation. The RRMs with a higher number of parameters were more useful to describe the genetic variation of test-day milk yield throughout the lactation. In this research estimation of genetic parameters, and calculation genetic correlations were implemented by random regression test day model, therefore using this method is the exact way to take account of parameters rather than the other ways.
Linear Regression Analysis

CERN Document Server

Seber, George A F

2012-01-01

Concise, mathematically clear, and comprehensive treatment of the subject.* Expanded coverage of diagnostics and methods of model fitting.* Requires no specialized knowledge beyond a good grasp of matrix algebra and some acquaintance with straight-line regression and simple analysis of variance models.* More than 200 problems throughout the book plus outline solutions for the exercises.* This revision has been extensively class-tested.
Wheat yield dynamics: a structural econometric analysis.

Science.gov (United States)

Sahin, Afsin; Akdi, Yilmaz; Arslan, Fahrettin

2007-10-15

In this study we initially have tried to explore the wheat situation in Turkey, which has a small-open economy and in the member countries of European Union (EU). We have observed that increasing the wheat yield is fundamental to obtain comparative advantage among countries by depressing domestic prices. Also the changing structure of supporting schemes in Turkey makes it necessary to increase its wheat yield level. For this purpose, we have used available data to determine the dynamics of wheat yield by Ordinary Least Square Regression methods. In order to find out whether there is a linear relationship among these series we have checked each series whether they are integrated at the same order or not. Consequently, we have pointed out that fertilizer usage and precipitation level are substantial inputs for producing high wheat yield. Furthermore, in respect for our model, fertilizer usage affects wheat yield more than precipitation level.
Diagnosis of cranial hemangioma: Comparison between logistic regression analysis and neuronal network

International Nuclear Information System (INIS)

Arana, E.; Marti-Bonmati, L.; Bautista, D.; Paredes, R.

1998-01-01

To study the utility of logistic regression and the neuronal network in the diagnosis of cranial hemangiomas. Fifteen patients presenting hemangiomas were selected form a total of 167 patients with cranial lesions. All were evaluated by plain radiography and computed tomography (CT). Nineteen variables in their medical records were reviewed. Logistic regression and neuronal network models were constructed and validated by the jackknife (leave-one-out) approach. The yields of the two models were compared by means of ROC curves, using the area under the curve as parameter. Seven men and 8 women presented hemangiomas. The mean age of these patients was 38.4 (15.4 years (mea ± standard deviation). Logistic regression identified as significant variables the shape, soft tissue mass and periosteal reaction. The neuronal network lent more importance to the existence of ossified matrix, ruptured cortical vein and the mixed calcified-blastic (trabeculated) pattern. The neuronal network showed a greater yield than logistic regression (Az, 0.9409) (0.004 versus 0.7211± 0.075; p<0.001). The neuronal network discloses hidden interactions among the variables, providing a higher yield in the characterization of cranial hemangiomas and constituting a medical diagnostic acid. (Author)29 refs

A hydrologic regression sediment-yield model for two ungaged watershed outlet stations in Africa

International Nuclear Information System (INIS)

Moussa, O.M.; Smith, S.E.; Shrestha, R.L.

1991-01-01

A hydrologic regression sediment-yield model was established to determine the relationship between water discharge and suspended sediment discharge at the Blue Nile and the Atbara River outlet stations during the flood season. The model consisted of two main submodels: (1) a suspended sediment discharge model, which was used to determine suspended sediment discharge for each basin outlet; and (2) a sediment rating model, which related water discharge and suspended sediment discharge for each outlet station. Due to the absence of suspended sediment concentration measurements at or near the outlet stations, a minimum norm solution, which is based on the minimization of the unknowns rather than the residuals, was used to determine the suspended sediment discharges at the stations. In addition, the sediment rating submodel was regressed by using an observation equations procedure. Verification analyses on the model were carried out and the mean percentage errors were found to be +12.59 and -12.39, respectively, for the Blue Nile and Atbara. The hydrologic regression model was found to be most sensitive to the relative weight matrix, moderately sensitive to the mean water discharge ratio, and slightly sensitive to the concentration variation along the River Nile's course
Preface to Berk's "Regression Analysis: A Constructive Critique"

OpenAIRE

de Leeuw, Jan

2003-01-01

It is pleasure to write a preface for the book ”Regression Analysis” of my fellow series editor Dick Berk. And it is a pleasure in particular because the book is about regression analysis, the most popular and the most fundamental technique in applied statistics. And because it is critical of the way regression analysis is used in the sciences, in particular in the social and behavioral sciences. Although the book can be read as an introduction to regression analysis, it can also be read as a...
Multicollinearity and Regression Analysis

Science.gov (United States)

Daoud, Jamal I.

2017-12-01

In regression analysis it is obvious to have a correlation between the response and predictor(s), but having correlation among predictors is something undesired. The number of predictors included in the regression model depends on many factors among which, historical data, experience, etc. At the end selection of most important predictors is something objective due to the researcher. Multicollinearity is a phenomena when two or more predictors are correlated, if this happens, the standard error of the coefficients will increase [8]. Increased standard errors means that the coefficients for some or all independent variables may be found to be significantly different from In other words, by overinflating the standard errors, multicollinearity makes some variables statistically insignificant when they should be significant. In this paper we focus on the multicollinearity, reasons and consequences on the reliability of the regression model.
Hierarchical regression analysis in structural Equation Modeling

NARCIS (Netherlands)

de Jong, P.F.

1999-01-01

In a hierarchical or fixed-order regression analysis, the independent variables are entered into the regression equation in a prespecified order. Such an analysis is often performed when the extra amount of variance accounted for in a dependent variable by a specific independent variable is the main
Multivariate Regression Analysis and Slaughter Livestock,

Science.gov (United States)

AGRICULTURE, *ECONOMICS), (*MEAT, PRODUCTION), MULTIVARIATE ANALYSIS, REGRESSION ANALYSIS , ANIMALS, WEIGHT, COSTS, PREDICTIONS, STABILITY, MATHEMATICAL MODELS, STORAGE, BEEF, PORK, FOOD, STATISTICAL DATA, ACCURACY
Multitrait, Random Regression, or Simple Repeatability Model in High-Throughput Phenotyping Data Improve Genomic Prediction for Wheat Grain Yield.

Science.gov (United States)

Sun, Jin; Rutkoski, Jessica E; Poland, Jesse A; Crossa, José; Jannink, Jean-Luc; Sorrells, Mark E

2017-07-01

High-throughput phenotyping (HTP) platforms can be used to measure traits that are genetically correlated with wheat ( L.) grain yield across time. Incorporating such secondary traits in the multivariate pedigree and genomic prediction models would be desirable to improve indirect selection for grain yield. In this study, we evaluated three statistical models, simple repeatability (SR), multitrait (MT), and random regression (RR), for the longitudinal data of secondary traits and compared the impact of the proposed models for secondary traits on their predictive abilities for grain yield. Grain yield and secondary traits, canopy temperature (CT) and normalized difference vegetation index (NDVI), were collected in five diverse environments for 557 wheat lines with available pedigree and genomic information. A two-stage analysis was applied for pedigree and genomic selection (GS). First, secondary traits were fitted by SR, MT, or RR models, separately, within each environment. Then, best linear unbiased predictions (BLUPs) of secondary traits from the above models were used in the multivariate prediction models to compare predictive abilities for grain yield. Predictive ability was substantially improved by 70%, on average, from multivariate pedigree and genomic models when including secondary traits in both training and test populations. Additionally, (i) predictive abilities slightly varied for MT, RR, or SR models in this data set, (ii) results indicated that including BLUPs of secondary traits from the MT model was the best in severe drought, and (iii) the RR model was slightly better than SR and MT models under drought environment. Copyright © 2017 Crop Science Society of America.
Prediction of unwanted pregnancies using logistic regression, probit regression and discriminant analysis.

Science.gov (United States)

Ebrahimzadeh, Farzad; Hajizadeh, Ebrahim; Vahabi, Nasim; Almasian, Mohammad; Bakhteyar, Katayoon

2015-01-01

Unwanted pregnancy not intended by at least one of the parents has undesirable consequences for the family and the society. In the present study, three classification models were used and compared to predict unwanted pregnancies in an urban population. In this cross-sectional study, 887 pregnant mothers referring to health centers in Khorramabad, Iran, in 2012 were selected by the stratified and cluster sampling; relevant variables were measured and for prediction of unwanted pregnancy, logistic regression, discriminant analysis, and probit regression models and SPSS software version 21 were used. To compare these models, indicators such as sensitivity, specificity, the area under the ROC curve, and the percentage of correct predictions were used. The prevalence of unwanted pregnancies was 25.3%. The logistic and probit regression models indicated that parity and pregnancy spacing, contraceptive methods, household income and number of living male children were related to unwanted pregnancy. The performance of the models based on the area under the ROC curve was 0.735, 0.733, and 0.680 for logistic regression, probit regression, and linear discriminant analysis, respectively. Given the relatively high prevalence of unwanted pregnancies in Khorramabad, it seems necessary to revise family planning programs. Despite the similar accuracy of the models, if the researcher is interested in the interpretability of the results, the use of the logistic regression model is recommended.
Bayesian logistic regression analysis

NARCIS (Netherlands)

Van Erp, H.R.N.; Van Gelder, P.H.A.J.M.

2012-01-01

In this paper we present a Bayesian logistic regression analysis. It is found that if one wishes to derive the posterior distribution of the probability of some event, then, together with the traditional Bayes Theorem and the integrating out of nuissance parameters, the Jacobian transformation is an
Regression analysis using dependent Polya trees.

Science.gov (United States)

Schörgendorfer, Angela; Branscum, Adam J

2013-11-30

Many commonly used models for linear regression analysis force overly simplistic shape and scale constraints on the residual structure of data. We propose a semiparametric Bayesian model for regression analysis that produces data-driven inference by using a new type of dependent Polya tree prior to model arbitrary residual distributions that are allowed to evolve across increasing levels of an ordinal covariate (e.g., time, in repeated measurement studies). By modeling residual distributions at consecutive covariate levels or time points using separate, but dependent Polya tree priors, distributional information is pooled while allowing for broad pliability to accommodate many types of changing residual distributions. We can use the proposed dependent residual structure in a wide range of regression settings, including fixed-effects and mixed-effects linear and nonlinear models for cross-sectional, prospective, and repeated measurement data. A simulation study illustrates the flexibility of our novel semiparametric regression model to accurately capture evolving residual distributions. In an application to immune development data on immunoglobulin G antibodies in children, our new model outperforms several contemporary semiparametric regression models based on a predictive model selection criterion. Copyright © 2013 John Wiley & Sons, Ltd.
RAWS II: A MULTIPLE REGRESSION ANALYSIS PROGRAM,

Science.gov (United States)

This memorandum gives instructions for the use and operation of a revised version of RAWS, a multiple regression analysis program. The program...of preprocessed data, the directed retention of variable, listing of the matrix of the normal equations and its inverse, and the bypassing of the regression analysis to provide the input variable statistics only. (Author)
Estimation of Genetic Parameters for First Lactation Monthly Test-day Milk Yields using Random Regression Test Day Model in Karan Fries Cattle

Directory of Open Access Journals (Sweden)

Ajay Singh

2016-06-01

Full Text Available A single trait linear mixed random regression test-day model was applied for the first time for analyzing the first lactation monthly test-day milk yield records in Karan Fries cattle. The test-day milk yield data was modeled using a random regression model (RRM considering different order of Legendre polynomial for the additive genetic effect (4th order and the permanent environmental effect (5th order. Data pertaining to 1,583 lactation records spread over a period of 30 years were recorded and analyzed in the study. The variance component, heritability and genetic correlations among test-day milk yields were estimated using RRM. RRM heritability estimates of test-day milk yield varied from 0.11 to 0.22 in different test-day records. The estimates of genetic correlations between different test-day milk yields ranged 0.01 (test-day 1 [TD-1] and TD-11 to 0.99 (TD-4 and TD-5. The magnitudes of genetic correlations between test-day milk yields decreased as the interval between test-days increased and adjacent test-day had higher correlations. Additive genetic and permanent environment variances were higher for test-day milk yields at both ends of lactation. The residual variance was observed to be lower than the permanent environment variance for all the test-day milk yields.
Advanced statistics: linear regression, part II: multiple linear regression.

Science.gov (United States)

Marill, Keith A

2004-01-01

The applications of simple linear regression in medical research are limited, because in most situations, there are multiple relevant predictor variables. Univariate statistical techniques such as simple linear regression use a single predictor variable, and they often may be mathematically correct but clinically misleading. Multiple linear regression is a mathematical technique used to model the relationship between multiple independent predictor variables and a single dependent outcome variable. It is used in medical research to model observational data, as well as in diagnostic and therapeutic studies in which the outcome is dependent on more than one factor. Although the technique generally is limited to data that can be expressed with a linear function, it benefits from a well-developed mathematical framework that yields unique solutions and exact confidence intervals for regression coefficients. Building on Part I of this series, this article acquaints the reader with some of the important concepts in multiple regression analysis. These include multicollinearity, interaction effects, and an expansion of the discussion of inference testing, leverage, and variable transformations to multivariate models. Examples from the first article in this series are expanded on using a primarily graphic, rather than mathematical, approach. The importance of the relationships among the predictor variables and the dependence of the multivariate model coefficients on the choice of these variables are stressed. Finally, concepts in regression model building are discussed.
Common pitfalls in statistical analysis: Linear regression analysis

Directory of Open Access Journals (Sweden)

Rakesh Aggarwal

2017-01-01

Full Text Available In a previous article in this series, we explained correlation analysis which describes the strength of relationship between two continuous variables. In this article, we deal with linear regression analysis which predicts the value of one continuous variable from another. We also discuss the assumptions and pitfalls associated with this analysis.
Estimation of genotype X environment interactions, in a grassbased system, for milk yield, body condition score,and body weight using random regression models

NARCIS (Netherlands)

Berry, D.P.; Buckley, F.; Dillon, P.; Evans, R.D.; Rath, M.; Veerkamp, R.F.

2003-01-01

(Co)variance components for milk yield, body condition score (BCS), body weight (BW), BCS change and BW change over different herd-year mean milk yields (HMY) and nutritional environments (concentrate feeding level, grazing severity and silage quality) were estimated using a random regression model.
Genotypic Correlation and Path Analysis of Some Traits related to Oil Yield and Grain Yield in Canola (Brassica napus L. under Non-stress and Water Deficit Stress Conditions

Directory of Open Access Journals (Sweden)

A Ismaili

2017-03-01

between seed yield and different characters were subjected to path coefficient analysis separately for partitioning these values into direct and indirect effects. Step-wise regression technique was used to determine the best model, which accounted for variation exist in plant seed and oil yield as dependent variables in separate analysis. Direct and indirect effects of traits entered to regression model were determined by using path coefficient analysis. Results and Discussion Results of this study showed significant differences among all genotypes performances, and also stress condition caused a significant decrease in performance of all studied traits. The highest seed yield obtained from Geronimo and Dante (with 3668 and 3505 kg.ha-1, respectively under non stress condition, and the highest seed yield obtained from Zarfam and Dante (with 2948 and 2860 kg ha-1, respectively under drought stress condition. Genotype Licord produced the highest oil content, which was significantly higher than that produced by other genotypes in either regime. Genotypic and phenotypic correlation coefficients were estimated between all traits and using stepwise regression, best model was introduced for two conditions. Under Non-stress condition, the average of genetic correlations between grain yield and silique length was high and positive (0.92**, suggesting that the selection of prolific plants resulted in a gain of selection for yield. Under water deficit stress condition, a negative average of genetic correlations (-0.28 was observed for grain yield and days to maturity. Path analysis based on the genotypic correlation under non-stress conditions between grain yield and other traits showed that number of pods per plant and pod length had direct effects on grain yield, while under drought conditions, pod length and plant height had important direct effects. Results of path analysis for oil yield under non-stress and stress conditions showed that grain yield had the most direct effect on
Genetic correlations among body condition score, yield and fertility in multiparous cows using random regression models

OpenAIRE

Bastin, Catherine; Gillon, Alain; Massart, Xavier; Bertozzi, Carlo; Vanderick, Sylvie; Gengler, Nicolas

2010-01-01

Genetic correlations between body condition score (BCS) in lactation 1 to 3 and four economically important traits (days open, 305-days milk, fat, and protein yields recorded in the first 3 lactations) were estimated on about 12,500 Walloon Holstein cows using 4-trait random regression models. Results indicated moderate favorable genetic correlations between BCS and days open (from -0.46 to -0.62) and suggested the use of BCS for indirect selection on fertility. However, unfavorable genetic c...
Improving Crop Yield and Nutrient Use Efficiency via Biofertilization-A Global Meta-analysis.

Science.gov (United States)

Schütz, Lukas; Gattinger, Andreas; Meier, Matthias; Müller, Adrian; Boller, Thomas; Mäder, Paul; Mathimaran, Natarajan

2017-01-01

The application of microbial inoculants (biofertilizers) is a promising technology for future sustainable farming systems in view of rapidly decreasing phosphorus stocks and the need to more efficiently use available nitrogen (N). Various microbial taxa are currently used as biofertilizers, based on their capacity to access nutrients from fertilizers and soil stocks, to fix atmospheric nitrogen, to improve water uptake or to act as biocontrol agents. Despite the existence of a considerable knowledge on effects of specific taxa of biofertilizers, a comprehensive quantitative assessment of the performance of biofertilizers with different traits such as phosphorus solubilization and N fixation applied to various crops at a global scale is missing. We conducted a meta-analysis to quantify benefits of biofertilizers in terms of yield increase, nitrogen and phosphorus use efficiency, based on 171 peer reviewed publications that met eligibility criteria. Major findings are: (i) the superiority of biofertilizer performance in dry climates over other climatic regions (yield response: dry climate +20.0 ± 1.7%, tropical climate +14.9 ± 1.2%, oceanic climate +10.0 ± 3.7%, continental climate +8.5 ± 2.4%); (ii) meta-regression analyses revealed that yield response due to biofertilizer application was generally small at low soil P levels; efficacy increased along higher soil P levels in the order arbuscular mycorrhizal fungi (AMF), P solubilizers, and N fixers; (iii) meta-regressions showed that the success of inoculation with AMF was greater at low organic matter content and at neutral pH. Our comprehensive analysis provides a basis and guidance for proper choice and application of biofertilizers.
Moderation analysis using a two-level regression model.

Science.gov (United States)

Yuan, Ke-Hai; Cheng, Ying; Maxwell, Scott

2014-10-01

Moderation analysis is widely used in social and behavioral research. The most commonly used model for moderation analysis is moderated multiple regression (MMR) in which the explanatory variables of the regression model include product terms, and the model is typically estimated by least squares (LS). This paper argues for a two-level regression model in which the regression coefficients of a criterion variable on predictors are further regressed on moderator variables. An algorithm for estimating the parameters of the two-level model by normal-distribution-based maximum likelihood (NML) is developed. Formulas for the standard errors (SEs) of the parameter estimates are provided and studied. Results indicate that, when heteroscedasticity exists, NML with the two-level model gives more efficient and more accurate parameter estimates than the LS analysis of the MMR model. When error variances are homoscedastic, NML with the two-level model leads to essentially the same results as LS with the MMR model. Most importantly, the two-level regression model permits estimating the percentage of variance of each regression coefficient that is due to moderator variables. When applied to data from General Social Surveys 1991, NML with the two-level model identified a significant moderation effect of race on the regression of job prestige on years of education while LS with the MMR model did not. An R package is also developed and documented to facilitate the application of the two-level model.
Hydrogeological Characteristics of Groundwater Yield in Shallow ...

African Journals Online (AJOL)

Hydrogeological Characteristics of Groundwater Yield in Shallow Wells of the ... of Water Resources and Lower Niger River Basin Development Authority in Ilorin. ... moment correlation, multiple and stepwise multiple regression analysis.
Improving Crop Yield and Nutrient Use Efficiency via Biofertilization—A Global Meta-analysis

Science.gov (United States)

Schütz, Lukas; Gattinger, Andreas; Meier, Matthias; Müller, Adrian; Boller, Thomas; Mäder, Paul; Mathimaran, Natarajan

2018-01-01

The application of microbial inoculants (biofertilizers) is a promising technology for future sustainable farming systems in view of rapidly decreasing phosphorus stocks and the need to more efficiently use available nitrogen (N). Various microbial taxa are currently used as biofertilizers, based on their capacity to access nutrients from fertilizers and soil stocks, to fix atmospheric nitrogen, to improve water uptake or to act as biocontrol agents. Despite the existence of a considerable knowledge on effects of specific taxa of biofertilizers, a comprehensive quantitative assessment of the performance of biofertilizers with different traits such as phosphorus solubilization and N fixation applied to various crops at a global scale is missing. We conducted a meta-analysis to quantify benefits of biofertilizers in terms of yield increase, nitrogen and phosphorus use efficiency, based on 171 peer reviewed publications that met eligibility criteria. Major findings are: (i) the superiority of biofertilizer performance in dry climates over other climatic regions (yield response: dry climate +20.0 ± 1.7%, tropical climate +14.9 ± 1.2%, oceanic climate +10.0 ± 3.7%, continental climate +8.5 ± 2.4%); (ii) meta-regression analyses revealed that yield response due to biofertilizer application was generally small at low soil P levels; efficacy increased along higher soil P levels in the order arbuscular mycorrhizal fungi (AMF), P solubilizers, and N fixers; (iii) meta-regressions showed that the success of inoculation with AMF was greater at low organic matter content and at neutral pH. Our comprehensive analysis provides a basis and guidance for proper choice and application of biofertilizers. PMID:29375594

Improving Crop Yield and Nutrient Use Efficiency via Biofertilization—A Global Meta-analysis

Directory of Open Access Journals (Sweden)

Lukas Schütz

2018-01-01

Full Text Available The application of microbial inoculants (biofertilizers is a promising technology for future sustainable farming systems in view of rapidly decreasing phosphorus stocks and the need to more efficiently use available nitrogen (N. Various microbial taxa are currently used as biofertilizers, based on their capacity to access nutrients from fertilizers and soil stocks, to fix atmospheric nitrogen, to improve water uptake or to act as biocontrol agents. Despite the existence of a considerable knowledge on effects of specific taxa of biofertilizers, a comprehensive quantitative assessment of the performance of biofertilizers with different traits such as phosphorus solubilization and N fixation applied to various crops at a global scale is missing. We conducted a meta-analysis to quantify benefits of biofertilizers in terms of yield increase, nitrogen and phosphorus use efficiency, based on 171 peer reviewed publications that met eligibility criteria. Major findings are: (i the superiority of biofertilizer performance in dry climates over other climatic regions (yield response: dry climate +20.0 ± 1.7%, tropical climate +14.9 ± 1.2%, oceanic climate +10.0 ± 3.7%, continental climate +8.5 ± 2.4%; (ii meta-regression analyses revealed that yield response due to biofertilizer application was generally small at low soil P levels; efficacy increased along higher soil P levels in the order arbuscular mycorrhizal fungi (AMF, P solubilizers, and N fixers; (iii meta-regressions showed that the success of inoculation with AMF was greater at low organic matter content and at neutral pH. Our comprehensive analysis provides a basis and guidance for proper choice and application of biofertilizers.
Genotype X Environment Interaction for Yield in Field Pea Pisum ...

African Journals Online (AJOL)

user

analysis of variance with individual stability regression co- efficient ... environmental score derived from a principal component ... Grain yield analysis was carried .... Analysis of variance for Additive Main effects and Multiple Interaction (AMMI).
Two Paradoxes in Linear Regression Analysis

Science.gov (United States)

FENG, Ge; PENG, Jing; TU, Dongke; ZHENG, Julia Z.; FENG, Changyong

2016-01-01

Summary Regression is one of the favorite tools in applied statistics. However, misuse and misinterpretation of results from regression analysis are common in biomedical research. In this paper we use statistical theory and simulation studies to clarify some paradoxes around this popular statistical method. In particular, we show that a widely used model selection procedure employed in many publications in top medical journals is wrong. Formal procedures based on solid statistical theory should be used in model selection. PMID:28638214
Using Dominance Analysis to Determine Predictor Importance in Logistic Regression

Science.gov (United States)

Azen, Razia; Traxel, Nicole

2009-01-01

This article proposes an extension of dominance analysis that allows researchers to determine the relative importance of predictors in logistic regression models. Criteria for choosing logistic regression R[superscript 2] analogues were determined and measures were selected that can be used to perform dominance analysis in logistic regression. A…
Linear regression and sensitivity analysis in nuclear reactor design

International Nuclear Information System (INIS)

Kumar, Akansha; Tsvetkov, Pavel V.; McClarren, Ryan G.

2015-01-01

Highlights: • Presented a benchmark for the applicability of linear regression to complex systems. • Applied linear regression to a nuclear reactor power system. • Performed neutronics, thermal–hydraulics, and energy conversion using Brayton’s cycle for the design of a GCFBR. • Performed detailed sensitivity analysis to a set of parameters in a nuclear reactor power system. • Modeled and developed reactor design using MCNP, regression using R, and thermal–hydraulics in Java. - Abstract: The paper presents a general strategy applicable for sensitivity analysis (SA), and uncertainity quantification analysis (UA) of parameters related to a nuclear reactor design. This work also validates the use of linear regression (LR) for predictive analysis in a nuclear reactor design. The analysis helps to determine the parameters on which a LR model can be fit for predictive analysis. For those parameters, a regression surface is created based on trial data and predictions are made using this surface. A general strategy of SA to determine and identify the influential parameters those affect the operation of the reactor is mentioned. Identification of design parameters and validation of linearity assumption for the application of LR of reactor design based on a set of tests is performed. The testing methods used to determine the behavior of the parameters can be used as a general strategy for UA, and SA of nuclear reactor models, and thermal hydraulics calculations. A design of a gas cooled fast breeder reactor (GCFBR), with thermal–hydraulics, and energy transfer has been used for the demonstration of this method. MCNP6 is used to simulate the GCFBR design, and perform the necessary criticality calculations. Java is used to build and run input samples, and to extract data from the output files of MCNP6, and R is used to perform regression analysis and other multivariate variance, and analysis of the collinearity of data
Least Squares Adjustment: Linear and Nonlinear Weighted Regression Analysis

DEFF Research Database (Denmark)

Nielsen, Allan Aasbjerg

2007-01-01

This note primarily describes the mathematics of least squares regression analysis as it is often used in geodesy including land surveying and satellite positioning applications. In these fields regression is often termed adjustment. The note also contains a couple of typical land surveying...... and satellite positioning application examples. In these application areas we are typically interested in the parameters in the model typically 2- or 3-D positions and not in predictive modelling which is often the main concern in other regression analysis applications. Adjustment is often used to obtain...... the clock error) and to obtain estimates of the uncertainty with which the position is determined. Regression analysis is used in many other fields of application both in the natural, the technical and the social sciences. Examples may be curve fitting, calibration, establishing relationships between...
Dose-Dependent Effects of Statins for Patients with Aneurysmal Subarachnoid Hemorrhage: Meta-Regression Analysis.

Science.gov (United States)

To, Minh-Son; Prakash, Shivesh; Poonnoose, Santosh I; Bihari, Shailesh

2018-05-01

The study uses meta-regression analysis to quantify the dose-dependent effects of statin pharmacotherapy on vasospasm, delayed ischemic neurologic deficits (DIND), and mortality in aneurysmal subarachnoid hemorrhage. Prospective, retrospective observational studies, and randomized controlled trials (RCTs) were retrieved by a systematic database search. Summary estimates were expressed as absolute risk (AR) for a given statin dose or control (placebo). Meta-regression using inverse variance weighting and robust variance estimation was performed to assess the effect of statin dose on transformed AR in a random effects model. Dose-dependence of predicted AR with 95% confidence interval (CI) was recovered by using Miller's Freeman-Tukey inverse. The database search and study selection criteria yielded 18 studies (2594 patients) for analysis. These included 12 RCTs, 4 retrospective observational studies, and 2 prospective observational studies. Twelve studies investigated simvastatin, whereas the remaining studies investigated atorvastatin, pravastatin, or pitavastatin, with simvastatin-equivalent doses ranging from 20 to 80 mg. Meta-regression revealed dose-dependent reductions in Freeman-Tukey-transformed AR of vasospasm (slope coefficient -0.00404, 95% CI -0.00720 to -0.00087; P = 0.0321), DIND (slope coefficient -0.00316, 95% CI -0.00586 to -0.00047; P = 0.0392), and mortality (slope coefficient -0.00345, 95% CI -0.00623 to -0.00067; P = 0.0352). The present meta-regression provides weak evidence for dose-dependent reductions in vasospasm, DIND and mortality associated with acute statin use after aneurysmal subarachnoid hemorrhage. However, the analysis was limited by substantial heterogeneity among individual studies. Greater dosing strategies are a potential consideration for future RCTs. Copyright © 2018 Elsevier Inc. All rights reserved.
Design and analysis of experiments classical and regression approaches with SAS

CERN Document Server

Onyiah, Leonard C

2008-01-01

Introductory Statistical Inference and Regression Analysis Elementary Statistical Inference Regression Analysis Experiments, the Completely Randomized Design (CRD)-Classical and Regression Approaches Experiments Experiments to Compare Treatments Some Basic Ideas Requirements of a Good Experiment One-Way Experimental Layout or the CRD: Design and Analysis Analysis of Experimental Data (Fixed Effects Model) Expected Values for the Sums of Squares The Analysis of Variance (ANOVA) Table Follow-Up Analysis to Check fo
Yield Stability of Sorghum Hybrids and Parental Lines | Kenga ...

African Journals Online (AJOL)

Seventy-five sorghum hybrids and twenty parental lines were evaluated for two consecutive years at two locations. Our objective was to compare relative stability of grain yields among hybrids and parental lines. Mean grain yields and stability analysis of variance, which included linear regression coefficient (bi) and ...
Sparse Regression by Projection and Sparse Discriminant Analysis

KAUST Repository

Qi, Xin

2015-04-03

© 2015, © American Statistical Association, Institute of Mathematical Statistics, and Interface Foundation of North America. Recent years have seen active developments of various penalized regression methods, such as LASSO and elastic net, to analyze high-dimensional data. In these approaches, the direction and length of the regression coefficients are determined simultaneously. Due to the introduction of penalties, the length of the estimates can be far from being optimal for accurate predictions. We introduce a new framework, regression by projection, and its sparse version to analyze high-dimensional data. The unique nature of this framework is that the directions of the regression coefficients are inferred first, and the lengths and the tuning parameters are determined by a cross-validation procedure to achieve the largest prediction accuracy. We provide a theoretical result for simultaneous model selection consistency and parameter estimation consistency of our method in high dimension. This new framework is then generalized such that it can be applied to principal components analysis, partial least squares, and canonical correlation analysis. We also adapt this framework for discriminant analysis. Compared with the existing methods, where there is relatively little control of the dependency among the sparse components, our method can control the relationships among the components. We present efficient algorithms and related theory for solving the sparse regression by projection problem. Based on extensive simulations and real data analysis, we demonstrate that our method achieves good predictive performance and variable selection in the regression setting, and the ability to control relationships between the sparse components leads to more accurate classification. In supplementary materials available online, the details of the algorithms and theoretical proofs, and R codes for all simulation studies are provided.
Genetic analysis of yield and yield components in Oryza sativa x ...

African Journals Online (AJOL)

... inheritance of yield and yield components and to estimate the heritabilities of important quantitative traits in rice (Oryza sativa L.). Six generations viz., P1, P2, F1, F2, BCP1 and BCP2 of a cross between IET6279 and IR70445-146-3-3 were used for the study. Generation mean analysis suggested that additive effects had a ...
Establishing a Mathematical Equations and Improving the Production of L-tert-Leucine by Uniform Design and Regression Analysis.

Science.gov (United States)

Jiang, Wei; Xu, Chao-Zhen; Jiang, Si-Zhi; Zhang, Tang-Duo; Wang, Shi-Zhen; Fang, Bai-Shan

2017-04-01

L-tert-Leucine (L-Tle) and its derivatives are extensively used as crucial building blocks for chiral auxiliaries, pharmaceutically active ingredients, and ligands. Combining with formate dehydrogenase (FDH) for regenerating the expensive coenzyme NADH, leucine dehydrogenase (LeuDH) is continually used for synthesizing L-Tle from α-keto acid. A multilevel factorial experimental design was executed for research of this system. In this work, an efficient optimization method for improving the productivity of L-Tle was developed. And the mathematical model between different fermentation conditions and L-Tle yield was also determined in the form of the equation by using uniform design and regression analysis. The multivariate regression equation was conveniently implemented in water, with a space time yield of 505.9 g L -1 day -1 and an enantiomeric excess value of >99 %. These results demonstrated that this method might become an ideal protocol for industrial production of chiral compounds and unnatural amino acids such as chiral drug intermediates.
The Use of Nonparametric Kernel Regression Methods in Econometric Production Analysis

DEFF Research Database (Denmark)

Czekaj, Tomasz Gerard

and nonparametric estimations of production functions in order to evaluate the optimal firm size. The second paper discusses the use of parametric and nonparametric regression methods to estimate panel data regression models. The third paper analyses production risk, price uncertainty, and farmers' risk preferences...... within a nonparametric panel data regression framework. The fourth paper analyses the technical efficiency of dairy farms with environmental output using nonparametric kernel regression in a semiparametric stochastic frontier analysis. The results provided in this PhD thesis show that nonparametric......This PhD thesis addresses one of the fundamental problems in applied econometric analysis, namely the econometric estimation of regression functions. The conventional approach to regression analysis is the parametric approach, which requires the researcher to specify the form of the regression...
Predictors of course in obsessive-compulsive disorder: logistic regression versus Cox regression for recurrent events.

Science.gov (United States)

Kempe, P T; van Oppen, P; de Haan, E; Twisk, J W R; Sluis, A; Smit, J H; van Dyck, R; van Balkom, A J L M

2007-09-01

Two methods for predicting remissions in obsessive-compulsive disorder (OCD) treatment are evaluated. Y-BOCS measurements of 88 patients with a primary OCD (DSM-III-R) diagnosis were performed over a 16-week treatment period, and during three follow-ups. Remission at any measurement was defined as a Y-BOCS score lower than thirteen combined with a reduction of seven points when compared with baseline. Logistic regression models were compared with a Cox regression for recurrent events model. Logistic regression yielded different models at different evaluation times. The recurrent events model remained stable when fewer measurements were used. Higher baseline levels of neuroticism and more severe OCD symptoms were associated with a lower chance of remission, early age of onset and more depressive symptoms with a higher chance. Choice of outcome time affects logistic regression prediction models. Recurrent events analysis uses all information on remissions and relapses. Short- and long-term predictors for OCD remission show overlap.
Simulation Experiments in Practice: Statistical Design and Regression Analysis

OpenAIRE

Kleijnen, J.P.C.

2007-01-01

In practice, simulation analysts often change only one factor at a time, and use graphical analysis of the resulting Input/Output (I/O) data. The goal of this article is to change these traditional, naïve methods of design and analysis, because statistical theory proves that more information is obtained when applying Design Of Experiments (DOE) and linear regression analysis. Unfortunately, classic DOE and regression analysis assume a single simulation response that is normally and independen...
Evaluation of logistic regression models and effect of covariates for case-control study in RNA-Seq analysis.

Science.gov (United States)

Choi, Seung Hoan; Labadorf, Adam T; Myers, Richard H; Lunetta, Kathryn L; Dupuis, Josée; DeStefano, Anita L

2017-02-06

Next generation sequencing provides a count of RNA molecules in the form of short reads, yielding discrete, often highly non-normally distributed gene expression measurements. Although Negative Binomial (NB) regression has been generally accepted in the analysis of RNA sequencing (RNA-Seq) data, its appropriateness has not been exhaustively evaluated. We explore logistic regression as an alternative method for RNA-Seq studies designed to compare cases and controls, where disease status is modeled as a function of RNA-Seq reads using simulated and Huntington disease data. We evaluate the effect of adjusting for covariates that have an unknown relationship with gene expression. Finally, we incorporate the data adaptive method in order to compare false positive rates. When the sample size is small or the expression levels of a gene are highly dispersed, the NB regression shows inflated Type-I error rates but the Classical logistic and Bayes logistic (BL) regressions are conservative. Firth's logistic (FL) regression performs well or is slightly conservative. Large sample size and low dispersion generally make Type-I error rates of all methods close to nominal alpha levels of 0.05 and 0.01. However, Type-I error rates are controlled after applying the data adaptive method. The NB, BL, and FL regressions gain increased power with large sample size, large log2 fold-change, and low dispersion. The FL regression has comparable power to NB regression. We conclude that implementing the data adaptive method appropriately controls Type-I error rates in RNA-Seq analysis. Firth's logistic regression provides a concise statistical inference process and reduces spurious associations from inaccurately estimated dispersion parameters in the negative binomial framework.
Multiple linear regression and artificial neural networks for delta-endotoxin and protease yields modelling of Bacillus thuringiensis.

Science.gov (United States)

Ennouri, Karim; Ben Ayed, Rayda; Triki, Mohamed Ali; Ottaviani, Ennio; Mazzarello, Maura; Hertelli, Fathi; Zouari, Nabil

2017-07-01

The aim of the present work was to develop a model that supplies accurate predictions of the yields of delta-endotoxins and proteases produced by B. thuringiensis var. kurstaki HD-1. Using available medium ingredients as variables, a mathematical method, based on Plackett-Burman design (PB), was employed to analyze and compare data generated by the Bootstrap method and processed by multiple linear regressions (MLR) and artificial neural networks (ANN) including multilayer perceptron (MLP) and radial basis function (RBF) models. The predictive ability of these models was evaluated by comparison of output data through the determination of coefficient (R 2 ) and mean square error (MSE) values. The results demonstrate that the prediction of the yields of delta-endotoxin and protease was more accurate by ANN technique (87 and 89% for delta-endotoxin and protease determination coefficients, respectively) when compared with MLR method (73.1 and 77.2% for delta-endotoxin and protease determination coefficients, respectively), suggesting that the proposed ANNs, especially MLP, is a suitable new approach for determining yields of bacterial products that allow us to make more appropriate predictions in a shorter time and with less engineering effort.
General Nature of Multicollinearity in Multiple Regression Analysis.

Science.gov (United States)

Liu, Richard

1981-01-01

Discusses multiple regression, a very popular statistical technique in the field of education. One of the basic assumptions in regression analysis requires that independent variables in the equation should not be highly correlated. The problem of multicollinearity and some of the solutions to it are discussed. (Author)
Mathematical and statistical analysis of the effect of boron on yield parameters of wheat

Energy Technology Data Exchange (ETDEWEB)

Rawashdeh, Hamzeh [Water Management and Environment Research Department, National Center for Agricultural Research and Extension, P.O. Box 639, Baqa 19381 (Jordan); Sala, Florin [Soil Science and Plant Nutrition, Faculty of Agriculture, Banat University of Agricultural Sciences and Veterinary Medicine “Regele Mihai I al României” from Timişoara, Timişoara, 300645 (Romania); Boldea, Marius [Mathematics and Statistics, Faculty of Agriculture, Banat University of Agricultural Sciences and Veterinary Medicine “Regele Mihai I al României” from Timisoara, Timişoara, 300645 (Romania)

2015-03-10

The main objective of this research is to investigate the effect of foliar applications of boron at different growth stages on yield and yield parameters of wheat. The contribution of boron in achieving yield parameters is described by second degree polynomial equations, with high statistical confidence (p<0.01; F theoretical < F calculated, according to ANOVA test, for Alfa = 0.05). Regression analysis, based on R{sup 2} values obtained, made it possible to evaluate the particular contribution of boron to the realization of yield parameters. This was lower for spike length (R{sup 2} = 0.812), thousand seeds weight (R{sup 2} = 0.850) and higher in the case of the number of spikelets (R{sup 2} = 0.936) and the number of seeds on a spike (R{sup 2} = 0.960). These results confirm that boron plays an important part in achieving the number of seeds on a spike in the case of wheat, as the contribution of this element to the process of flower fertilization is well-known. In regards to productivity elements, the contribution of macroelements to yield quantity is clear, the contribution of B alone being R{sup 2} = 0.868.
On logistic regression analysis of dichotomized responses.

Science.gov (United States)

Lu, Kaifeng

2017-01-01

We study the properties of treatment effect estimate in terms of odds ratio at the study end point from logistic regression model adjusting for the baseline value when the underlying continuous repeated measurements follow a multivariate normal distribution. Compared with the analysis that does not adjust for the baseline value, the adjusted analysis produces a larger treatment effect as well as a larger standard error. However, the increase in standard error is more than offset by the increase in treatment effect so that the adjusted analysis is more powerful than the unadjusted analysis for detecting the treatment effect. On the other hand, the true adjusted odds ratio implied by the normal distribution of the underlying continuous variable is a function of the baseline value and hence is unlikely to be able to be adequately represented by a single value of adjusted odds ratio from the logistic regression model. In contrast, the risk difference function derived from the logistic regression model provides a reasonable approximation to the true risk difference function implied by the normal distribution of the underlying continuous variable over the range of the baseline distribution. We show that different metrics of treatment effect have similar statistical power when evaluated at the baseline mean. Copyright © 2016 John Wiley & Sons, Ltd. Copyright © 2016 John Wiley & Sons, Ltd.

On two flexible methods of 2-dimensional regression analysis

Czech Academy of Sciences Publication Activity Database

Volf, Petr

2012-01-01

Roč. 18, č. 4 (2012), s. 154-164 ISSN 1803-9782 Grant - others:GA ČR(CZ) GAP209/10/2045 Institutional support: RVO:67985556 Keywords : regression analysis * Gordon surface * prediction error * projection pursuit Subject RIV: BB - Applied Statistics, Operational Research http://library.utia.cas.cz/separaty/2013/SI/volf-on two flexible methods of 2-dimensional regression analysis.pdf
Resting-state functional magnetic resonance imaging: the impact of regression analysis.

Science.gov (United States)

Yeh, Chia-Jung; Tseng, Yu-Sheng; Lin, Yi-Ru; Tsai, Shang-Yueh; Huang, Teng-Yi

2015-01-01

To investigate the impact of regression methods on resting-state functional magnetic resonance imaging (rsfMRI). During rsfMRI preprocessing, regression analysis is considered effective for reducing the interference of physiological noise on the signal time course. However, it is unclear whether the regression method benefits rsfMRI analysis. Twenty volunteers (10 men and 10 women; aged 23.4 ± 1.5 years) participated in the experiments. We used node analysis and functional connectivity mapping to assess the brain default mode network by using five combinations of regression methods. The results show that regressing the global mean plays a major role in the preprocessing steps. When a global regression method is applied, the values of functional connectivity are significantly lower (P ≤ .01) than those calculated without a global regression. This step increases inter-subject variation and produces anticorrelated brain areas. rsfMRI data processed using regression should be interpreted carefully. The significance of the anticorrelated brain areas produced by global signal removal is unclear. Copyright © 2014 by the American Society of Neuroimaging.
Development of a User Interface for a Regression Analysis Software Tool

Science.gov (United States)

Ulbrich, Norbert Manfred; Volden, Thomas R.

2010-01-01

An easy-to -use user interface was implemented in a highly automated regression analysis tool. The user interface was developed from the start to run on computers that use the Windows, Macintosh, Linux, or UNIX operating system. Many user interface features were specifically designed such that a novice or inexperienced user can apply the regression analysis tool with confidence. Therefore, the user interface s design minimizes interactive input from the user. In addition, reasonable default combinations are assigned to those analysis settings that influence the outcome of the regression analysis. These default combinations will lead to a successful regression analysis result for most experimental data sets. The user interface comes in two versions. The text user interface version is used for the ongoing development of the regression analysis tool. The official release of the regression analysis tool, on the other hand, has a graphical user interface that is more efficient to use. This graphical user interface displays all input file names, output file names, and analysis settings for a specific software application mode on a single screen which makes it easier to generate reliable analysis results and to perform input parameter studies. An object-oriented approach was used for the development of the graphical user interface. This choice keeps future software maintenance costs to a reasonable limit. Examples of both the text user interface and graphical user interface are discussed in order to illustrate the user interface s overall design approach.
Path Analysis of Grain Yield and Yield Components and Some Agronomic Traits in Bread Wheat

Directory of Open Access Journals (Sweden)

Mohsen Janmohammadi

2014-01-01

Full Text Available Development of new bread wheat cultivars needs efficient tools to monitor trait association in a breeding program. This investigation was aimed to characterize grain yield components and some agronomic traits related to bread wheat grain yield. The efficiency of a breeding program depends mainly on the direction of the correlation between different traits and the relative importance of each component involved in contributing to grain yield. Correlation and path analysis were carried out in 56 bread wheat genotypes grown under field conditions of Maragheh, Iran. Observations were recorded on 18 wheat traits and correlation coefficient analysis revealed grain yield was positively correlated with stem diameter, spike length, floret number, spikelet number, grain diameter, grain length and 1000 seed weight traits. According to the variance inflation factor (VIF and tolerance as multicollinearity statistics, there are inconsistent relationships among the variables and all traits could be considered as first-order variables (Model I with grain yield as the response variable due to low multicollinearity of all measured traits. In the path coefficient analysis, grain yield represented the dependent variable and the spikelet number and 1000 seed weight traits were the independent ones. Our results indicated that the number of spikelets per spikes and leaf width and 1000 seed weight traits followed by the grain length, grain diameter and grain number per spike were the traits related to higher grain yield. The above mentioned traits along with their indirect causal factors should be considered simultaneously as an effective selection criteria evolving high yielding genotype because of their direct positive contribution to grain yield.
Method for nonlinear exponential regression analysis

Science.gov (United States)

Junkin, B. G.

1972-01-01

Two computer programs developed according to two general types of exponential models for conducting nonlinear exponential regression analysis are described. Least squares procedure is used in which the nonlinear problem is linearized by expanding in a Taylor series. Program is written in FORTRAN 5 for the Univac 1108 computer.
Quantitative Genetic Analysis for Yield and Yield Components in Boro Rice (Oryza sativa L.

Directory of Open Access Journals (Sweden)

Supriyo CHAKRABORTY

2010-03-01

Full Text Available Twenty-nine genotypes of boro rice (Oryza sativa L. were grown in a randomized block design with three replications in plots of 4m x 1m with a crop geometry of 20 cm x 20 cm between November-April, in Regional Agricultural Research Station, Nagaon, India. Quantitative data were collected on five randomly selected plants of each genotype per replication for yield/plant, and six other yield components, namely plant height, panicles/plant, panicle length, effective grains/panicle, 100 grain weight and harvest index. Mean values of the characters for each genotype were used for analysis of variance and covariance to obtain information on genotypic and phenotypic correlation along with coheritability between two characters. Path analyses were carried out to estimate the direct and indirect effects of boro rice�s yield components. The objective of the study was to identify the characters that mostly influence the yield for increasing boro rice productivity through breeding program. Correlation analysis revealed significant positive genotypic correlation of yield/plant with plant height (0.21, panicles/plant (0.53, panicle length (0.53, effective grains/panicle (0.57 and harvest index (0.86. Path analysis based on genotypic correlation coefficients elucidated high positive direct effect of harvest index (0.8631, panicle length (0.2560 and 100 grain weight (0.1632 on yield/plant with a residual effect of 0.33. Plant height and panicles/plant recorded high positive indirect effect on yield/plant via harvest index whereas effective grains/panicle on yield/plant via harvest index and panicle length. Results of the present study suggested that five component characters, namely harvest index, effective grains/plant, panicle length, panicles/plant and plant height influenced the yield of boro rice. A genotype with higher magnitude of these component characters could be either selected from the existing genotypes or evolved by breeding program for genetic
Regression of uveal malignant melanomas following cobalt-60 plaque. Correlates between acoustic spectrum analysis and tumor regression

International Nuclear Information System (INIS)

Coleman, D.J.; Lizzi, F.L.; Silverman, R.H.; Ellsworth, R.M.; Haik, B.G.; Abramson, D.H.; Smith, M.E.; Rondeau, M.J.

1985-01-01

Parameters derived from computer analysis of digital radio-frequency (rf) ultrasound scan data of untreated uveal malignant melanomas were examined for correlations with tumor regression following cobalt-60 plaque. Parameters included tumor height, normalized power spectrum and acoustic tissue type (ATT). Acoustic tissue type was based upon discriminant analysis of tumor power spectra, with spectra of tumors of known pathology serving as a model. Results showed ATT to be correlated with tumor regression during the first 18 months following treatment. Tumors with ATT associated with spindle cell malignant melanoma showed over twice the percentage reduction in height as those with ATT associated with mixed/epithelioid melanomas. Pre-treatment height was only weakly correlated with regression. Additionally, significant spectral changes were observed following treatment. Ultrasonic spectrum analysis thus provides a noninvasive tool for classification, prediction and monitoring of tumor response to cobalt-60 plaque
Using Ridge Regression Models to Estimate Grain Yield from Field Spectral Data in Bread Wheat (Triticum Aestivum L. Grown under Three Water Regimes

Directory of Open Access Journals (Sweden)

Javier Hernandez

2015-02-01

Full Text Available Plant breeding based on grain yield (GY is an expensive and time-consuming method, so new indirect estimation techniques to evaluate the performance of crops represent an alternative method to improve grain yield. The present study evaluated the ability of canopy reflectance spectroscopy at the range from 350 to 2500 nm to predict GY in a large panel (368 genotypes of wheat (Triticum aestivum L. through multivariate ridge regression models. Plants were treated under three water regimes in the Mediterranean conditions of central Chile: severe water stress (SWS, rain fed, mild water stress (MWS; one irrigation event around booting and full irrigation (FI with mean GYs of 1655, 4739, and 7967 kg∙ha−1, respectively. Models developed from reflectance data during anthesis and grain filling under all water regimes explained between 77% and 91% of the GY variability, with the highest values in SWS condition. When individual models were used to predict yield in the rest of the trials assessed, models fitted during anthesis under MWS performed best. Combined models using data from different water regimes and each phenological stage were used to predict grain yield, and the coefficients of determination (R2 increased to 89.9% and 92.0% for anthesis and grain filling, respectively. The model generated during anthesis in MWS was the best at predicting yields when it was applied to other conditions. Comparisons against conventional reflectance indices were made, showing lower predictive abilities. It was concluded that a Ridge Regression Model using a data set based on spectral reflectance at anthesis or grain filling represents an effective method to predict grain yield in genotypes under different water regimes.
Genotype x environment interaction and stability analysis for yield ...

African Journals Online (AJOL)

etc

2015-05-06

. Combined analysis of variance (ANOVA) for yield and yield components revealed highly significant .... yield stability among varieties, multi-location trials with ... Mean grain yield (kg/ha) of 17 Kabuli-type chickpea genotypes ...
Detecting overdispersion in count data: A zero-inflated Poisson regression analysis

Science.gov (United States)

Afiqah Muhamad Jamil, Siti; Asrul Affendi Abdullah, M.; Kek, Sie Long; Nor, Maria Elena; Mohamed, Maryati; Ismail, Norradihah

2017-09-01

This study focusing on analysing count data of butterflies communities in Jasin, Melaka. In analysing count dependent variable, the Poisson regression model has been known as a benchmark model for regression analysis. Continuing from the previous literature that used Poisson regression analysis, this study comprising the used of zero-inflated Poisson (ZIP) regression analysis to gain acute precision on analysing the count data of butterfly communities in Jasin, Melaka. On the other hands, Poisson regression should be abandoned in the favour of count data models, which are capable of taking into account the extra zeros explicitly. By far, one of the most popular models include ZIP regression model. The data of butterfly communities which had been called as the number of subjects in this study had been taken in Jasin, Melaka and consisted of 131 number of subjects visits Jasin, Melaka. Since the researchers are considering the number of subjects, this data set consists of five families of butterfly and represent the five variables involve in the analysis which are the types of subjects. Besides, the analysis of ZIP used the SAS procedure of overdispersion in analysing zeros value and the main purpose of continuing the previous study is to compare which models would be better than when exists zero values for the observation of the count data. The analysis used AIC, BIC and Voung test of 5% level significance in order to achieve the objectives. The finding indicates that there is a presence of over-dispersion in analysing zero value. The ZIP regression model is better than Poisson regression model when zero values exist.
An Analysis of Bank Service Satisfaction Based on Quantile Regression and Grey Relational Analysis

Directory of Open Access Journals (Sweden)

Wen-Tsao Pan

2016-01-01

Full Text Available Bank service satisfaction is vital to the success of a bank. In this paper, we propose to use the grey relational analysis to gauge the levels of service satisfaction of the banks. With the grey relational analysis, we compared the effects of different variables on service satisfaction. We gave ranks to the banks according to their levels of service satisfaction. We further used the quantile regression model to find the variables that affected the satisfaction of a customer at a specific quantile of satisfaction level. The result of the quantile regression analysis provided a bank manager with information to formulate policies to further promote satisfaction of the customers at different quantiles of satisfaction level. We also compared the prediction accuracies of the regression models at different quantiles. The experiment result showed that, among the seven quantile regression models, the median regression model has the best performance in terms of RMSE, RTIC, and CE performance measures.
Research and analyze of physical health using multiple regression analysis

Directory of Open Access Journals (Sweden)

T. S. Kyi

2014-01-01

Full Text Available This paper represents the research which is trying to create a mathematical model of the "healthy people" using the method of regression analysis. The factors are the physical parameters of the person (such as heart rate, lung capacity, blood pressure, breath holding, weight height coefficient, flexibility of the spine, muscles of the shoulder belt, abdominal muscles, squatting, etc.., and the response variable is an indicator of physical working capacity. After performing multiple regression analysis, obtained useful multiple regression models that can predict the physical performance of boys the aged of fourteen to seventeen years. This paper represents the development of regression model for the sixteen year old boys and analyzed results.
An improved multiple linear regression and data analysis computer program package

Science.gov (United States)

Sidik, S. M.

1972-01-01

NEWRAP, an improved version of a previous multiple linear regression program called RAPIER, CREDUC, and CRSPLT, allows for a complete regression analysis including cross plots of the independent and dependent variables, correlation coefficients, regression coefficients, analysis of variance tables, t-statistics and their probability levels, rejection of independent variables, plots of residuals against the independent and dependent variables, and a canonical reduction of quadratic response functions useful in optimum seeking experimentation. A major improvement over RAPIER is that all regression calculations are done in double precision arithmetic.
Functional data analysis of generalized regression quantiles

KAUST Repository

Guo, Mengmeng; Zhou, Lan; Huang, Jianhua Z.; Hä rdle, Wolfgang Karl

2013-01-01

Generalized regression quantiles, including the conditional quantiles and expectiles as special cases, are useful alternatives to the conditional means for characterizing a conditional distribution, especially when the interest lies in the tails. We develop a functional data analysis approach to jointly estimate a family of generalized regression quantiles. Our approach assumes that the generalized regression quantiles share some common features that can be summarized by a small number of principal component functions. The principal component functions are modeled as splines and are estimated by minimizing a penalized asymmetric loss measure. An iterative least asymmetrically weighted squares algorithm is developed for computation. While separate estimation of individual generalized regression quantiles usually suffers from large variability due to lack of sufficient data, by borrowing strength across data sets, our joint estimation approach significantly improves the estimation efficiency, which is demonstrated in a simulation study. The proposed method is applied to data from 159 weather stations in China to obtain the generalized quantile curves of the volatility of the temperature at these stations. © 2013 Springer Science+Business Media New York.
Functional data analysis of generalized regression quantiles

KAUST Repository

Guo, Mengmeng

2013-11-05

Generalized regression quantiles, including the conditional quantiles and expectiles as special cases, are useful alternatives to the conditional means for characterizing a conditional distribution, especially when the interest lies in the tails. We develop a functional data analysis approach to jointly estimate a family of generalized regression quantiles. Our approach assumes that the generalized regression quantiles share some common features that can be summarized by a small number of principal component functions. The principal component functions are modeled as splines and are estimated by minimizing a penalized asymmetric loss measure. An iterative least asymmetrically weighted squares algorithm is developed for computation. While separate estimation of individual generalized regression quantiles usually suffers from large variability due to lack of sufficient data, by borrowing strength across data sets, our joint estimation approach significantly improves the estimation efficiency, which is demonstrated in a simulation study. The proposed method is applied to data from 159 weather stations in China to obtain the generalized quantile curves of the volatility of the temperature at these stations. © 2013 Springer Science+Business Media New York.
Analysis of Relationship Between Personality and Favorite Places with Poisson Regression Analysis

Directory of Open Access Journals (Sweden)

Yoon Song Ha

2018-01-01

Full Text Available A relationship between human personality and preferred locations have been a long conjecture for human mobility research. In this paper, we analyzed the relationship between personality and visiting place with Poisson Regression. Poisson Regression can analyze correlation between countable dependent variable and independent variable. For this analysis, 33 volunteers provided their personality data and 49 location categories data are used. Raw location data is preprocessed to be normalized into rates of visit and outlier data is prunned. For the regression analysis, independent variables are personality data and dependent variables are preprocessed location data. Several meaningful results are found. For example, persons with high tendency of frequent visiting to university laboratory has personality with high conscientiousness and low openness. As well, other meaningful location categories are presented in this paper.
application of multilinear regression analysis in modeling of soil

African Journals Online (AJOL)

Windows User

Accordingly [1, 3] in their work, they applied linear regression ... (MLRA) is a statistical technique that uses several explanatory ... order to check this, they adopted bivariate correlation analysis .... groups, namely A-1 through A-7, based on their relative expected ..... Multivariate Regression in Gorgan Province North of Iran” ...
Clinical and pathologic factors affecting lymph node yields in colorectal cancer.

Directory of Open Access Journals (Sweden)

Ta-Wen Hsu

Full Text Available OBJECTIVE: Lymph node yield is recommended as a benchmark of quality care in colorectal cancer. The objective of this study was to evaluate the impact of various factors upon lymph node yield and to identify independent factors associated with lymph node harvest. MATERIALS AND METHODS: The records of 162 patients with Stage I to Stage III colorectal cancers seen in one institution were reviewed. These patients underwent radical surgery as definitive therapy; high-risk patients then received adjuvant treatment. Pathologic and demographic data were recorded and analyzed. The subgroup analysis of lymph node yields was determined using a t-test and analysis of variants. Linear regression model and multivariable analysis were used to perform potential confounding and predicting variables. RESULTS: Five variables had significant association with lymph node yield after adjustment for other factors in a multiple linear regression model. These variables were: tumor size, surgical method, specimen length, and individual surgeon and pathologist. The model with these five significant variables interpreted 44.4% of the variation. CONCLUSIONS: Patients, tumor characteristics and surgical variables all influence the number of lymph nodes retrieved. Physicians are the main gatekeepers. Adequate training and optimized guidelines could greatly improve the quality of lymph node yields.
Multiple regression analysis of Jominy hardenability data for boron treated steels

International Nuclear Information System (INIS)

Komenda, J.; Sandstroem, R.; Tukiainen, M.

1997-01-01

The relations between chemical composition and their hardenability of boron treated steels have been investigated using a multiple regression analysis method. A linear model of regression was chosen. The free boron content that is effective for the hardenability was calculated using a model proposed by Jansson. The regression analysis for 1261 steel heats provided equations that were statistically significant at the 95% level. All heats met the specification according to the nordic countries producers classification. The variation in chemical composition explained typically 80 to 90% of the variation in the hardenability. In the regression analysis elements which did not significantly contribute to the calculated hardness according to the F test were eliminated. Carbon, silicon, manganese, phosphorus and chromium were of importance at all Jominy distances, nickel, vanadium, boron and nitrogen at distances above 6 mm. After the regression analysis it was demonstrated that very few outliers were present in the data set, i.e. data points outside four times the standard deviation. The model has successfully been used in industrial practice replacing some of the necessary Jominy tests. (orig.)
Partial Least Squares Regression for Determining the Control Factors for Runoff and Suspended Sediment Yield during Rainfall Events

Directory of Open Access Journals (Sweden)

Nufang Fang

2015-07-01

Full Text Available Multivariate statistics are commonly used to identify the factors that control the dynamics of runoff or sediment yields during hydrological processes. However, one issue with the use of conventional statistical methods to address relationships between variables and runoff or sediment yield is multicollinearity. The main objectives of this study were to apply a method for effectively identifying runoff and sediment control factors during hydrological processes and apply that method to a case study. The method combines the clustering approach and partial least squares regression (PLSR models. The case study was conducted in a mountainous watershed in the Three Gorges Area. A total of 29 flood events in three hydrological years in areas with different land uses were obtained. In total, fourteen related variables were separated from hydrographs using the classical hydrograph separation method. Twenty-nine rainfall events were classified into two rainfall regimes (heavy Rainfall Regime I and moderate Rainfall Regime II based on rainfall characteristics and K-means clustering. Four separate PLSR models were constructed to identify the main variables that control runoff and sediment yield for the two rainfall regimes. For Rainfall Regime I, the dominant first-order factors affecting the changes in sediment yield in our study were all of the four rainfall-related variables, flood peak discharge, maximum flood suspended sediment concentration, runoff, and the percentages of forest and farmland. For Rainfall Regime II, antecedent condition-related variables have more effects on both runoff and sediment yield than in Rainfall Regime I. The results suggest that the different control factors of the two rainfall regimes are determined by the rainfall characteristics and thus different runoff mechanisms.

Understanding logistic regression analysis

OpenAIRE

Sperandei, Sandro

2014-01-01

Logistic regression is used to obtain odds ratio in the presence of more than one explanatory variable. The procedure is quite similar to multiple linear regression, with the exception that the response variable is binomial. The result is the impact of each variable on the odds ratio of the observed event of interest. The main advantage is to avoid confounding effects by analyzing the association of all variables together. In this article, we explain the logistic regression procedure using ex...
correlation studies and path coefficient analysis for seed yield

African Journals Online (AJOL)

Prof. Adipala Ekwamu

African Crop Science Journal, Vol. 21, No. 1, pp. 51 - 59 ... Yield being a quantitative trait has complex inheritance, which is ... Analysis for seed yield and yield components in Ethiopian coriander. 53 ..... The financial assistance of Canadian.
Genotype x environment interaction for grain yield of wheat genotypes tested under water stress conditions

International Nuclear Information System (INIS)

Sail, M.A.; Dahot, M.U.; Mangrio, S.M.; Memon, S.

2007-01-01

Effect of water stress on grain yield in different wheat genotypes was studied under field conditions at various locations. Grain yield is a complex polygenic trait influenced by genotype, environment and genotype x environment (GxE) interaction. To understand the stability among genotypes for grain yield, twenty-one wheat genotypes developed Through hybridization and radiation-induced mutations at Nuclear Institute of Agriculture (NIA) TandoJam were evaluated with four local check varieties (Sarsabz, Thori, Margalla-99 and Chakwal-86) in multi-environmental trails (MET/sub s/). The experiments were conducted over 5 different water stress environments in Sindh. Data on grain yield were recorded from each site and statistically analyzed. Combined analysis of variance for all the environments indicated that the genotype, environment and genotype x environment (GxE) interaction were highly significant (P greater then 0.01) for grain yield. Genotypes differed in their response to various locations. The overall highest site mean yield (4031 kg/ha) recorded at Moro and the lowest (2326 kg/ha) at Thatta. Six genotypes produced significantly (P=0.01) the highest grain yield overall the environments. Stability analysis was applied to estimate stability parameters viz., regression coefficient (b), standard error of regression coefficient and variance due to deviation from regression (S/sub 2/d) genotypes 10/8, BWS-78 produced the highest mean yield over all the environments with low regression coefficient (b=0.68, 0.67 and 0.63 respectively and higher S/sup 2/ d value, showing specific adaptation to poor (un favorable) environments. Genotype 8/7 produced overall higher grain yield (3647 kg/ha) and ranked as third high yielding genotype had regression value close to unity (b=0.9) and low S/sup d/ value, indicating more stability and wide adaptation over the all environments. The knowledge of the presence and magnitude of genotype x environment (GE) interaction is important to
Background stratified Poisson regression analysis of cohort data.

Science.gov (United States)

Richardson, David B; Langholz, Bryan

2012-03-01

Background stratified Poisson regression is an approach that has been used in the analysis of data derived from a variety of epidemiologically important studies of radiation-exposed populations, including uranium miners, nuclear industry workers, and atomic bomb survivors. We describe a novel approach to fit Poisson regression models that adjust for a set of covariates through background stratification while directly estimating the radiation-disease association of primary interest. The approach makes use of an expression for the Poisson likelihood that treats the coefficients for stratum-specific indicator variables as 'nuisance' variables and avoids the need to explicitly estimate the coefficients for these stratum-specific parameters. Log-linear models, as well as other general relative rate models, are accommodated. This approach is illustrated using data from the Life Span Study of Japanese atomic bomb survivors and data from a study of underground uranium miners. The point estimate and confidence interval obtained from this 'conditional' regression approach are identical to the values obtained using unconditional Poisson regression with model terms for each background stratum. Moreover, it is shown that the proposed approach allows estimation of background stratified Poisson regression models of non-standard form, such as models that parameterize latency effects, as well as regression models in which the number of strata is large, thereby overcoming the limitations of previously available statistical software for fitting background stratified Poisson regression models.
Poisson Regression Analysis of Illness and Injury Surveillance Data

Energy Technology Data Exchange (ETDEWEB)

Frome E.L., Watkins J.P., Ellis E.D.

2012-12-12

The Department of Energy (DOE) uses illness and injury surveillance to monitor morbidity and assess the overall health of the work force. Data collected from each participating site include health events and a roster file with demographic information. The source data files are maintained in a relational data base, and are used to obtain stratified tables of health event counts and person time at risk that serve as the starting point for Poisson regression analysis. The explanatory variables that define these tables are age, gender, occupational group, and time. Typical response variables of interest are the number of absences due to illness or injury, i.e., the response variable is a count. Poisson regression methods are used to describe the effect of the explanatory variables on the health event rates using a log-linear main effects model. Results of fitting the main effects model are summarized in a tabular and graphical form and interpretation of model parameters is provided. An analysis of deviance table is used to evaluate the importance of each of the explanatory variables on the event rate of interest and to determine if interaction terms should be considered in the analysis. Although Poisson regression methods are widely used in the analysis of count data, there are situations in which over-dispersion occurs. This could be due to lack-of-fit of the regression model, extra-Poisson variation, or both. A score test statistic and regression diagnostics are used to identify over-dispersion. A quasi-likelihood method of moments procedure is used to evaluate and adjust for extra-Poisson variation when necessary. Two examples are presented using respiratory disease absence rates at two DOE sites to illustrate the methods and interpretation of the results. In the first example the Poisson main effects model is adequate. In the second example the score test indicates considerable over-dispersion and a more detailed analysis attributes the over-dispersion to extra
A Quality Assessment Tool for Non-Specialist Users of Regression Analysis

Science.gov (United States)

Argyrous, George

2015-01-01

This paper illustrates the use of a quality assessment tool for regression analysis. It is designed for non-specialist "consumers" of evidence, such as policy makers. The tool provides a series of questions such consumers of evidence can ask to interrogate regression analysis, and is illustrated with reference to a recent study published…
Background stratified Poisson regression analysis of cohort data

International Nuclear Information System (INIS)

Richardson, David B.; Langholz, Bryan

2012-01-01

Background stratified Poisson regression is an approach that has been used in the analysis of data derived from a variety of epidemiologically important studies of radiation-exposed populations, including uranium miners, nuclear industry workers, and atomic bomb survivors. We describe a novel approach to fit Poisson regression models that adjust for a set of covariates through background stratification while directly estimating the radiation-disease association of primary interest. The approach makes use of an expression for the Poisson likelihood that treats the coefficients for stratum-specific indicator variables as 'nuisance' variables and avoids the need to explicitly estimate the coefficients for these stratum-specific parameters. Log-linear models, as well as other general relative rate models, are accommodated. This approach is illustrated using data from the Life Span Study of Japanese atomic bomb survivors and data from a study of underground uranium miners. The point estimate and confidence interval obtained from this 'conditional' regression approach are identical to the values obtained using unconditional Poisson regression with model terms for each background stratum. Moreover, it is shown that the proposed approach allows estimation of background stratified Poisson regression models of non-standard form, such as models that parameterize latency effects, as well as regression models in which the number of strata is large, thereby overcoming the limitations of previously available statistical software for fitting background stratified Poisson regression models. (orig.)
Studies on the Effects of Climatic Factors on Dryland Wheat Grain Yield in Maragheh Region

Directory of Open Access Journals (Sweden)

V. Feiziasl

2011-01-01

Full Text Available Abstract In order to study the effects of climate variables on rainfed wheat grain yield, climate data and wheat yield for 10 years (1995-2005 collected from Dryland Agricultural Research Institute (DARI in Maragheh as the main station in cold and semi-cold areas. Collected data were analyzed by correlation coefficient, simple regression, stepwise regression and path analysis. The results showed that relationships between grain yield with average relative humidity and total rainfall of growing season was positive and significant at 5% and 1% probabilities, respectively. However, evaluation between grain yield with sunny hours and class A pan evaporation was negative and significant (p
Understanding logistic regression analysis.

Science.gov (United States)

Sperandei, Sandro

2014-01-01

Logistic regression is used to obtain odds ratio in the presence of more than one explanatory variable. The procedure is quite similar to multiple linear regression, with the exception that the response variable is binomial. The result is the impact of each variable on the odds ratio of the observed event of interest. The main advantage is to avoid confounding effects by analyzing the association of all variables together. In this article, we explain the logistic regression procedure using examples to make it as simple as possible. After definition of the technique, the basic interpretation of the results is highlighted and then some special issues are discussed.
Modelling fourier regression for time series data- a case study: modelling inflation in foods sector in Indonesia

Science.gov (United States)

Prahutama, Alan; Suparti; Wahyu Utami, Tiani

2018-03-01

Regression analysis is an analysis to model the relationship between response variables and predictor variables. The parametric approach to the regression model is very strict with the assumption, but nonparametric regression model isn’t need assumption of model. Time series data is the data of a variable that is observed based on a certain time, so if the time series data wanted to be modeled by regression, then we should determined the response and predictor variables first. Determination of the response variable in time series is variable in t-th (yt), while the predictor variable is a significant lag. In nonparametric regression modeling, one developing approach is to use the Fourier series approach. One of the advantages of nonparametric regression approach using Fourier series is able to overcome data having trigonometric distribution. In modeling using Fourier series needs parameter of K. To determine the number of K can be used Generalized Cross Validation method. In inflation modeling for the transportation sector, communication and financial services using Fourier series yields an optimal K of 120 parameters with R-square 99%. Whereas if it was modeled by multiple linear regression yield R-square 90%.
Retro-regression--another important multivariate regression improvement.

Science.gov (United States)

Randić, M

2001-01-01

We review the serious problem associated with instabilities of the coefficients of regression equations, referred to as the MRA (multivariate regression analysis) "nightmare of the first kind". This is manifested when in a stepwise regression a descriptor is included or excluded from a regression. The consequence is an unpredictable change of the coefficients of the descriptors that remain in the regression equation. We follow with consideration of an even more serious problem, referred to as the MRA "nightmare of the second kind", arising when optimal descriptors are selected from a large pool of descriptors. This process typically causes at different steps of the stepwise regression a replacement of several previously used descriptors by new ones. We describe a procedure that resolves these difficulties. The approach is illustrated on boiling points of nonanes which are considered (1) by using an ordered connectivity basis; (2) by using an ordering resulting from application of greedy algorithm; and (3) by using an ordering derived from an exhaustive search for optimal descriptors. A novel variant of multiple regression analysis, called retro-regression (RR), is outlined showing how it resolves the ambiguities associated with both "nightmares" of the first and the second kind of MRA.
Automated Detection of Connective Tissue by Tissue Counter Analysis and Classification and Regression Trees

Directory of Open Access Journals (Sweden)

Josef Smolle

2001-01-01

Full Text Available Objective: To evaluate the feasibility of the CART (Classification and Regression Tree procedure for the recognition of microscopic structures in tissue counter analysis. Methods: Digital microscopic images of H&E stained slides of normal human skin and of primary malignant melanoma were overlayed with regularly distributed square measuring masks (elements and grey value, texture and colour features within each mask were recorded. In the learning set, elements were interactively labeled as representing either connective tissue of the reticular dermis, other tissue components or background. Subsequently, CART models were based on these data sets. Results: Implementation of the CART classification rules into the image analysis program showed that in an independent test set 94.1% of elements classified as connective tissue of the reticular dermis were correctly labeled. Automated measurements of the total amount of tissue and of the amount of connective tissue within a slide showed high reproducibility (r=0.97 and r=0.94, respectively; p < 0.001. Conclusions: CART procedure in tissue counter analysis yields simple and reproducible classification rules for tissue elements.
Linear regression analysis: part 14 of a series on evaluation of scientific publications.

Science.gov (United States)

Schneider, Astrid; Hommel, Gerhard; Blettner, Maria

2010-11-01

Regression analysis is an important statistical method for the analysis of medical data. It enables the identification and characterization of relationships among multiple factors. It also enables the identification of prognostically relevant risk factors and the calculation of risk scores for individual prognostication. This article is based on selected textbooks of statistics, a selective review of the literature, and our own experience. After a brief introduction of the uni- and multivariable regression models, illustrative examples are given to explain what the important considerations are before a regression analysis is performed, and how the results should be interpreted. The reader should then be able to judge whether the method has been used correctly and interpret the results appropriately. The performance and interpretation of linear regression analysis are subject to a variety of pitfalls, which are discussed here in detail. The reader is made aware of common errors of interpretation through practical examples. Both the opportunities for applying linear regression analysis and its limitations are presented.
Management of Industrial Performance Indicators: Regression Analysis and Simulation

Directory of Open Access Journals (Sweden)

Walter Roberto Hernandez Vergara

2017-11-01

Full Text Available Stochastic methods can be used in problem solving and explanation of natural phenomena through the application of statistical procedures. The article aims to associate the regression analysis and systems simulation, in order to facilitate the practical understanding of data analysis. The algorithms were developed in Microsoft Office Excel software, using statistical techniques such as regression theory, ANOVA and Cholesky Factorization, which made it possible to create models of single and multiple systems with up to five independent variables. For the analysis of these models, the Monte Carlo simulation and analysis of industrial performance indicators were used, resulting in numerical indices that aim to improve the goals’ management for compliance indicators, by identifying systems’ instability, correlation and anomalies. The analytical models presented in the survey indicated satisfactory results with numerous possibilities for industrial and academic applications, as well as the potential for deployment in new analytical techniques.
Long Term Evaluation of Yield Stability Trend for Cereal Crops in Iran

Directory of Open Access Journals (Sweden)

mehdi nassiri mahalati

2016-05-01

Full Text Available During the last few decades cereals yield have increased drastically at the national level however, information about yield stability and its resistance to annual environmental variability are scare. In this study long term stability of grin yield of wheat, barley, rice, corn and overall cereals in Iran were evaluated during a 40-year period (1971-2011. Stability analysis was conducted using two different methods. In the first method the residuals of regression between crop yield and time (years were calculated as stability index. For this different segmented regression models including linear, bi-linear and tri-linear were fitted to yield trend data and the best model for each crop was selected based on statistical measures. Absolute residuals (the difference between actual and predicted yields for each year as well as relative residuals (absolute residuals as percent of predicted yield were estimated. In the second method yield stability was estimated from the slope of the regression line between average annual yield of all cereals (environmental index and the yield of each crop in the same year. Results indicted that in wheat and barley absolute and relative residuals were increased during the study period leading to reduction of stability despite considerable yield increment. However, for rice and corn residuals followed a decreasing trend and therefore yield stability of these crops was increased during the last 40 years. The same result was obtained with the environmental index but in this method reduction of yield stability in barley was lower than wheat. Based on the results, yield and yield stability of cereals crops in Iran increased during the last 40 years. However, the percentage increase in stability is lower than that of yield. Application of nitrogen fertilizers was led to reduction in stability. Yield stability of wheat, barley, rice, corn and overall cereals was improved with increasing their cultivated area.
Least-Squares Linear Regression and Schrodinger's Cat: Perspectives on the Analysis of Regression Residuals.

Science.gov (United States)

Hecht, Jeffrey B.

The analysis of regression residuals and detection of outliers are discussed, with emphasis on determining how deviant an individual data point must be to be considered an outlier and the impact that multiple suspected outlier data points have on the process of outlier determination and treatment. Only bivariate (one dependent and one independent)…
Risky decision making in Attention-Deficit/Hyperactivity Disorder: A meta-regression analysis.

Science.gov (United States)

Dekkers, Tycho J; Popma, Arne; Agelink van Rentergem, Joost A; Bexkens, Anika; Huizenga, Hilde M

2016-04-01

ADHD has been associated with various forms of risky real life decision making, for example risky driving, unsafe sex and substance abuse. However, results from laboratory studies on decision making deficits in ADHD have been inconsistent, probably because of between study differences. We therefore performed a meta-regression analysis in which 37 studies (n ADHD=1175; n Control=1222) were included, containing 52 effect sizes. The overall analysis yielded a small to medium effect size (standardized mean difference=.36, pdecision making than control groups. There was a trend for a moderating influence of co-morbid Disruptive Behavior Disorders (DBD): studies including more participants with co-morbid DBD had larger effect sizes. No moderating influence of co-morbid internalizing disorders, age or task explicitness was found. These results indicate that ADHD is related to increased risky decision making in laboratory settings, which tended to be more pronounced if ADHD is accompanied by DBD. We therefore argue that risky decision making should have a more prominent role in research on the neuropsychological and -biological mechanisms of ADHD, which can be useful in ADHD assessment and intervention. Copyright © 2016 Elsevier Ltd. All rights reserved.
Predicting Dropouts of University Freshmen: A Logit Regression Analysis.

Science.gov (United States)

Lam, Y. L. Jack

1984-01-01

Stepwise discriminant analysis coupled with logit regression analysis of freshmen data from Brandon University (Manitoba) indicated that six tested variables drawn from research on university dropouts were useful in predicting attrition: student status, residence, financial sources, distance from home town, goal fulfillment, and satisfaction with…
Simulation Experiments in Practice : Statistical Design and Regression Analysis

NARCIS (Netherlands)

Kleijnen, J.P.C.

2007-01-01

In practice, simulation analysts often change only one factor at a time, and use graphical analysis of the resulting Input/Output (I/O) data. Statistical theory proves that more information is obtained when applying Design Of Experiments (DOE) and linear regression analysis. Unfortunately, classic
Association between response rates and survival outcomes in patients with newly diagnosed multiple myeloma. A systematic review and meta-regression analysis.

Science.gov (United States)

Mainou, Maria; Madenidou, Anastasia-Vasiliki; Liakos, Aris; Paschos, Paschalis; Karagiannis, Thomas; Bekiari, Eleni; Vlachaki, Efthymia; Wang, Zhen; Murad, Mohammad Hassan; Kumar, Shaji; Tsapas, Apostolos

2017-06-01

We performed a systematic review and meta-regression analysis of randomized control trials to investigate the association between response to initial treatment and survival outcomes in patients with newly diagnosed multiple myeloma (MM). Response outcomes included complete response (CR) and the combined outcome of CR or very good partial response (VGPR), while survival outcomes were overall survival (OS) and progression-free survival (PFS). We used random-effect meta-regression models and conducted sensitivity analyses based on definition of CR and study quality. Seventy-two trials were included in the systematic review, 63 of which contributed data in meta-regression analyses. There was no association between OS and CR in patients without autologous stem cell transplant (ASCT) (regression coefficient: .02, 95% confidence interval [CI] -0.06, 0.10), in patients undergoing ASCT (-.11, 95% CI -0.44, 0.22) and in trials comparing ASCT with non-ASCT patients (.04, 95% CI -0.29, 0.38). Similarly, OS did not correlate with the combined metric of CR or VGPR, and no association was evident between response outcomes and PFS. Sensitivity analyses yielded similar results. This meta-regression analysis suggests that there is no association between conventional response outcomes and survival in patients with newly diagnosed MM. © 2017 John Wiley & Sons A/S. Published by John Wiley & Sons Ltd.

Production yield analysis in the poultry processing industry

NARCIS (Netherlands)

Somsen, D.J.; Capelle, A.; Tramper, J.

2004-01-01

The paper outlines a case study where the PYA-method (production yield analysis) was implemented at a poultry-slaughtering line, processing 9000 broiler chicks per hour. It was shown that the average live weight of a flock of broilers could be used to predict the maximum production yield of the
Quality of life in breast cancer patients--a quantile regression analysis.

Science.gov (United States)

Pourhoseingholi, Mohamad Amin; Safaee, Azadeh; Moghimi-Dehkordi, Bijan; Zeighami, Bahram; Faghihzadeh, Soghrat; Tabatabaee, Hamid Reza; Pourhoseingholi, Asma

2008-01-01

Quality of life study has an important role in health care especially in chronic diseases, in clinical judgment and in medical resources supplying. Statistical tools like linear regression are widely used to assess the predictors of quality of life. But when the response is not normal the results are misleading. The aim of this study is to determine the predictors of quality of life in breast cancer patients, using quantile regression model and compare to linear regression. A cross-sectional study conducted on 119 breast cancer patients that admitted and treated in chemotherapy ward of Namazi hospital in Shiraz. We used QLQ-C30 questionnaire to assessment quality of life in these patients. A quantile regression was employed to assess the assocciated factors and the results were compared to linear regression. All analysis carried out using SAS. The mean score for the global health status for breast cancer patients was 64.92+/-11.42. Linear regression showed that only grade of tumor, occupational status, menopausal status, financial difficulties and dyspnea were statistically significant. In spite of linear regression, financial difficulties were not significant in quantile regression analysis and dyspnea was only significant for first quartile. Also emotion functioning and duration of disease statistically predicted the QOL score in the third quartile. The results have demonstrated that using quantile regression leads to better interpretation and richer inference about predictors of the breast cancer patient quality of life.
Modelling of seed yield and its components in tall fescue (Festuca ...

African Journals Online (AJOL)

AJL

2011-10-03

Oct 3, 2011 ... Ridge regression analysis was used to derive a steady algorithmic .... included three replicates [3 × 6 = 18 plots (treatments), stochastic ..... The parameter estimates of the five seed yield components of a total of 327 samples.
Visual grading characteristics and ordinal regression analysis during optimisation of CT head examinations.

Science.gov (United States)

Zarb, Francis; McEntee, Mark F; Rainford, Louise

2015-06-01

To evaluate visual grading characteristics (VGC) and ordinal regression analysis during head CT optimisation as a potential alternative to visual grading assessment (VGA), traditionally employed to score anatomical visualisation. Patient images (n = 66) were obtained using current and optimised imaging protocols from two CT suites: a 16-slice scanner at the national Maltese centre for trauma and a 64-slice scanner in a private centre. Local resident radiologists (n = 6) performed VGA followed by VGC and ordinal regression analysis. VGC alone indicated that optimised protocols had similar image quality as current protocols. Ordinal logistic regression analysis provided an in-depth evaluation, criterion by criterion allowing the selective implementation of the protocols. The local radiology review panel supported the implementation of optimised protocols for brain CT examinations (including trauma) in one centre, achieving radiation dose reductions ranging from 24 % to 36 %. In the second centre a 29 % reduction in radiation dose was achieved for follow-up cases. The combined use of VGC and ordinal logistic regression analysis led to clinical decisions being taken on the implementation of the optimised protocols. This improved method of image quality analysis provided the evidence to support imaging protocol optimisation, resulting in significant radiation dose savings. • There is need for scientifically based image quality evaluation during CT optimisation. • VGC and ordinal regression analysis in combination led to better informed clinical decisions. • VGC and ordinal regression analysis led to dose reductions without compromising diagnostic efficacy.
Regression analysis of radiological parameters in nuclear power plants

International Nuclear Information System (INIS)

Bhargava, Pradeep; Verma, R.K.; Joshi, M.L.

2003-01-01

Indian Pressurized Heavy Water Reactors (PHWRs) have now attained maturity in their operations. Indian PHWR operation started in the year 1972. At present there are 12 operating PHWRs collectively producing nearly 2400 MWe. Sufficient radiological data are available for analysis to draw inferences which may be utilised for better understanding of radiological parameters influencing the collective internal dose. Tritium is the main contributor to the occupational internal dose originating in PHWRs. An attempt has been made to establish the relationship between radiological parameters, which may be useful to draw inferences about the internal dose. Regression analysis have been done to find out the relationship, if it exist, among the following variables: A. Specific tritium activity of heavy water (Moderator and PHT) and tritium concentration in air at various work locations. B. Internal collective occupational dose and tritium release to environment through air route. C. Specific tritium activity of heavy water (Moderator and PHT) and collective internal occupational dose. For this purpose multivariate regression analysis has been carried out. D. Tritium concentration in air at various work location and tritium release to environment through air route. For this purpose multivariate regression analysis has been carried out. This analysis reveals that collective internal dose has got very good correlation with the tritium activity release to the environment through air route. Whereas no correlation has been found between specific tritium activity in the heavy water systems and collective internal occupational dose. The good correlation has been found in case D and F test reveals that it is not by chance. (author)
Comparison of cranial sex determination by discriminant analysis and logistic regression.

Science.gov (United States)

Amores-Ampuero, Anabel; Alemán, Inmaculada

2016-04-05

Various methods have been proposed for estimating dimorphism. The objective of this study was to compare sex determination results from cranial measurements using discriminant analysis or logistic regression. The study sample comprised 130 individuals (70 males) of known sex, age, and cause of death from San José cemetery in Granada (Spain). Measurements of 19 neurocranial dimensions and 11 splanchnocranial dimensions were subjected to discriminant analysis and logistic regression, and the percentages of correct classification were compared between the sex functions obtained with each method. The discriminant capacity of the selected variables was evaluated with a cross-validation procedure. The percentage accuracy with discriminant analysis was 78.2% for the neurocranium (82.4% in females and 74.6% in males) and 73.7% for the splanchnocranium (79.6% in females and 68.8% in males). These percentages were higher with logistic regression analysis: 85.7% for the neurocranium (in both sexes) and 94.1% for the splanchnocranium (100% in females and 91.7% in males).
Regression Analysis: Instructional Resource for Cost/Managerial Accounting

Science.gov (United States)

Stout, David E.

2015-01-01

This paper describes a classroom-tested instructional resource, grounded in principles of active learning and a constructivism, that embraces two primary objectives: "demystify" for accounting students technical material from statistics regarding ordinary least-squares (OLS) regression analysis--material that students may find obscure or…
Regression analysis of case K interval-censored failure time data in the presence of informative censoring.

Science.gov (United States)

Wang, Peijie; Zhao, Hui; Sun, Jianguo

2016-12-01

Interval-censored failure time data occur in many fields such as demography, economics, medical research, and reliability and many inference procedures on them have been developed (Sun, 2006; Chen, Sun, and Peace, 2012). However, most of the existing approaches assume that the mechanism that yields interval censoring is independent of the failure time of interest and it is clear that this may not be true in practice (Zhang et al., 2007; Ma, Hu, and Sun, 2015). In this article, we consider regression analysis of case K interval-censored failure time data when the censoring mechanism may be related to the failure time of interest. For the problem, an estimated sieve maximum-likelihood approach is proposed for the data arising from the proportional hazards frailty model and for estimation, a two-step procedure is presented. In the addition, the asymptotic properties of the proposed estimators of regression parameters are established and an extensive simulation study suggests that the method works well. Finally, we apply the method to a set of real interval-censored data that motivated this study. © 2016, The International Biometric Society.
Robust Mediation Analysis Based on Median Regression

Science.gov (United States)

Yuan, Ying; MacKinnon, David P.

2014-01-01

Mediation analysis has many applications in psychology and the social sciences. The most prevalent methods typically assume that the error distribution is normal and homoscedastic. However, this assumption may rarely be met in practice, which can affect the validity of the mediation analysis. To address this problem, we propose robust mediation analysis based on median regression. Our approach is robust to various departures from the assumption of homoscedasticity and normality, including heavy-tailed, skewed, contaminated, and heteroscedastic distributions. Simulation studies show that under these circumstances, the proposed method is more efficient and powerful than standard mediation analysis. We further extend the proposed robust method to multilevel mediation analysis, and demonstrate through simulation studies that the new approach outperforms the standard multilevel mediation analysis. We illustrate the proposed method using data from a program designed to increase reemployment and enhance mental health of job seekers. PMID:24079925
Multiplication factor versus regression analysis in stature estimation from hand and foot dimensions.

Science.gov (United States)

Krishan, Kewal; Kanchan, Tanuj; Sharma, Abhilasha

2012-05-01

Estimation of stature is an important parameter in identification of human remains in forensic examinations. The present study is aimed to compare the reliability and accuracy of stature estimation and to demonstrate the variability in estimated stature and actual stature using multiplication factor and regression analysis methods. The study is based on a sample of 246 subjects (123 males and 123 females) from North India aged between 17 and 20 years. Four anthropometric measurements; hand length, hand breadth, foot length and foot breadth taken on the left side in each subject were included in the study. Stature was measured using standard anthropometric techniques. Multiplication factors were calculated and linear regression models were derived for estimation of stature from hand and foot dimensions. Derived multiplication factors and regression formula were applied to the hand and foot measurements in the study sample. The estimated stature from the multiplication factors and regression analysis was compared with the actual stature to find the error in estimated stature. The results indicate that the range of error in estimation of stature from regression analysis method is less than that of multiplication factor method thus, confirming that the regression analysis method is better than multiplication factor analysis in stature estimation. Copyright © 2012 Elsevier Ltd and Faculty of Forensic and Legal Medicine. All rights reserved.
Managment oriented analysis of sediment yield time compression

Science.gov (United States)

Smetanova, Anna; Le Bissonnais, Yves; Raclot, Damien; Nunes, João P.; Licciardello, Feliciana; Le Bouteiller, Caroline; Latron, Jérôme; Rodríguez Caballero, Emilio; Mathys, Nicolle; Klotz, Sébastien; Mekki, Insaf; Gallart, Francesc; Solé Benet, Albert; Pérez Gallego, Nuria; Andrieux, Patrick; Moussa, Roger; Planchon, Olivier; Marisa Santos, Juliana; Alshihabi, Omran; Chikhaoui, Mohamed

2016-04-01

The understanding of inter- and intra-annual variability of sediment yield is important for the land use planning and management decisions for sustainable landscapes. It is of particular importance in the regions where the annual sediment yield is often highly dependent on the occurrence of few large events which produce the majority of sediments, such as in the Mediterranean. This phenomenon is referred as time compression, and relevance of its consideration growths with the increase in magnitude and frequency of extreme events due to climate change in many other regions. So far, time compression has ben studied mainly on events datasets, providing high resolution, but (in terms of data amount, required data precision and methods), demanding analysis. In order to provide an alternative simplified approach, the monthly and yearly time compressions were evaluated in eight Mediterranean catchments (of the R-OSMed network), representing a wide range of Mediterranean landscapes. The annual sediment yield varied between 0 to ~27100 Mg•km-2•a-1, and the monthly sediment yield between 0 to ~11600 Mg•km-2•month-1. The catchment's sediment yield was un-equally distributed at inter- and intra-annual scale, and large differences were observed between the catchments. Two types of time compression were distinguished - (i) the inter-annual (based on annual values) and intra- annual (based on monthly values). Four different rainfall-runoff-sediment yield time compression patterns were observed: (i) no time-compression of rainfall, runoff, nor sediment yield, (ii) low time compression of rainfall and runoff, but high compression of sediment yield, (iii) low compression of rainfall and high of runoff and sediment yield, and (iv) low, medium and high compression of rainfall, runoff and sediment yield. All four patterns were present at inter-annual scale, while at intra-annual scale only the two latter were present. This implies that high sediment yields occurred in
A comparative study of multiple regression analysis and back ...

Indian Academy of Sciences (India)

Abhijit Sarkar

artificial neural network (ANN) models to predict weld bead geometry and HAZ width in submerged arc welding ... Keywords. Submerged arc welding (SAW); multi-regression analysis (MRA); artificial neural network ..... Degree of freedom.
Genetic analysis of yield in peanut ( Arachis hypogaea L.) using ...

African Journals Online (AJOL)

The yield had significant major gene effect and the results implied that not only should the two major genes' effects be considered but also the polygene's effect should be considered in breeding to increase peanut yield. Key words: Peanut, yield, major gene plus polygene inheritance model, genetic analysis.
Biases in Farm-Level Yield Risk Analysis due to Data Aggregation

NARCIS (Netherlands)

Finger, R.

2012-01-01

We investigate biases in farm-level yield risk analysis caused by data aggregation from the farm-level to regional and national levels using the example of Swiss wheat and barley yields. The estimated yield variability decreases significantly with increasing level of aggregation, with crop yield
Multilayer perceptron for robust nonlinear interval regression analysis using genetic algorithms.

Science.gov (United States)

Hu, Yi-Chung

2014-01-01

On the basis of fuzzy regression, computational models in intelligence such as neural networks have the capability to be applied to nonlinear interval regression analysis for dealing with uncertain and imprecise data. When training data are not contaminated by outliers, computational models perform well by including almost all given training data in the data interval. Nevertheless, since training data are often corrupted by outliers, robust learning algorithms employed to resist outliers for interval regression analysis have been an interesting area of research. Several approaches involving computational intelligence are effective for resisting outliers, but the required parameters for these approaches are related to whether the collected data contain outliers or not. Since it seems difficult to prespecify the degree of contamination beforehand, this paper uses multilayer perceptron to construct the robust nonlinear interval regression model using the genetic algorithm. Outliers beyond or beneath the data interval will impose slight effect on the determination of data interval. Simulation results demonstrate that the proposed method performs well for contaminated datasets.
Regression analysis for LED color detection of visual-MIMO system

Science.gov (United States)

Banik, Partha Pratim; Saha, Rappy; Kim, Ki-Doo

2018-04-01

Color detection from a light emitting diode (LED) array using a smartphone camera is very difficult in a visual multiple-input multiple-output (visual-MIMO) system. In this paper, we propose a method to determine the LED color using a smartphone camera by applying regression analysis. We employ a multivariate regression model to identify the LED color. After taking a picture of an LED array, we select the LED array region, and detect the LED using an image processing algorithm. We then apply the k-means clustering algorithm to determine the number of potential colors for feature extraction of each LED. Finally, we apply the multivariate regression model to predict the color of the transmitted LEDs. In this paper, we show our results for three types of environmental light condition: room environmental light, low environmental light (560 lux), and strong environmental light (2450 lux). We compare the results of our proposed algorithm from the analysis of training and test R-Square (%) values, percentage of closeness of transmitted and predicted colors, and we also mention about the number of distorted test data points from the analysis of distortion bar graph in CIE1931 color space.
Soybean yield modeling using bootstrap methods for small samples

Energy Technology Data Exchange (ETDEWEB)

Dalposso, G.A.; Uribe-Opazo, M.A.; Johann, J.A.

2016-11-01

One of the problems that occur when working with regression models is regarding the sample size; once the statistical methods used in inferential analyzes are asymptotic if the sample is small the analysis may be compromised because the estimates will be biased. An alternative is to use the bootstrap methodology, which in its non-parametric version does not need to guess or know the probability distribution that generated the original sample. In this work we used a set of soybean yield data and physical and chemical soil properties formed with fewer samples to determine a multiple linear regression model. Bootstrap methods were used for variable selection, identification of influential points and for determination of confidence intervals of the model parameters. The results showed that the bootstrap methods enabled us to select the physical and chemical soil properties, which were significant in the construction of the soybean yield regression model, construct the confidence intervals of the parameters and identify the points that had great influence on the estimated parameters. (Author)
Evaluation of syngas production unit cost of bio-gasification facility using regression analysis techniques

Energy Technology Data Exchange (ETDEWEB)

Deng, Yangyang; Parajuli, Prem B.

2011-08-10

Evaluation of economic feasibility of a bio-gasification facility needs understanding of its unit cost under different production capacities. The objective of this study was to evaluate the unit cost of syngas production at capacities from 60 through 1800Nm 3/h using an economic model with three regression analysis techniques (simple regression, reciprocal regression, and log-log regression). The preliminary result of this study showed that reciprocal regression analysis technique had the best fit curve between per unit cost and production capacity, with sum of error squares (SES) lower than 0.001 and coefficient of determination of (R 2) 0.996. The regression analysis techniques determined the minimum unit cost of syngas production for micro-scale bio-gasification facilities of $0.052/Nm 3, under the capacity of 2,880 Nm 3/h. The results of this study suggest that to reduce cost, facilities should run at a high production capacity. In addition, the contribution of this technique could be the new categorical criterion to evaluate micro-scale bio-gasification facility from the perspective of economic analysis.
Sequential Path Analysis for Determination of Relationship Between Yield and Yield Components in Bread Wheat (Triticum aestivum.L.

Directory of Open Access Journals (Sweden)

Mohtasham MOHAMMADI

2014-03-01

Full Text Available An experiment was conducted to evaluate 295 wheat genotypes in Alpha-Lattice design with two replications. The arithmetic mean and standard deviation of grain yield was 2706 and 950 (kg/ha,respectively. The results of correlation coefficients indicated that grain yield had significant and positive association with plant height, spike length, early growth vigor and agronomic score. Whereas there were negative correlation coefficients between grain yield and days to physiological maturity and canopy temperature before and during anthesis. Path analysis indicated agronomic score and plant height had high positive direct effects on grain yield, while canopy temperature before and during anthesis, and days to maturity, wes another trait having negative direct effect on grain yield. The results of sequential path analysis showed the traits that accounted as a criteria variable for high grain yield were agronomic score, plant height, canopy temperature, spike length, chlorophyll content and early growth vigor, which were determined as first, second and third order variables and had strong effects on grain yield via one or more paths. More important, as canopy temperature, agronomic score and early growth vigor can be evaluated quickly and easily, these traits may be used for evaluation of large populations.
REML/BLUP and sequential path analysis in estimating genotypic values and interrelationships among simple maize grain yield-related traits.

Science.gov (United States)

Olivoto, T; Nardino, M; Carvalho, I R; Follmann, D N; Ferrari, M; Szareski, V J; de Pelegrin, A J; de Souza, V Q

2017-03-22

Methodologies using restricted maximum likelihood/best linear unbiased prediction (REML/BLUP) in combination with sequential path analysis in maize are still limited in the literature. Therefore, the aims of this study were: i) to use REML/BLUP-based procedures in order to estimate variance components, genetic parameters, and genotypic values of simple maize hybrids, and ii) to fit stepwise regressions considering genotypic values to form a path diagram with multi-order predictors and minimum multicollinearity that explains the relationships of cause and effect among grain yield-related traits. Fifteen commercial simple maize hybrids were evaluated in multi-environment trials in a randomized complete block design with four replications. The environmental variance (78.80%) and genotype-vs-environment variance (20.83%) accounted for more than 99% of the phenotypic variance of grain yield, which difficult the direct selection of breeders for this trait. The sequential path analysis model allowed the selection of traits with high explanatory power and minimum multicollinearity, resulting in models with elevated fit (R 2 > 0.9 and ε analysis is effective in the evaluation of maize-breeding trials.

A primer for biomedical scientists on how to execute model II linear regression analysis.

Science.gov (United States)

Ludbrook, John

2012-04-01

1. There are two very different ways of executing linear regression analysis. One is Model I, when the x-values are fixed by the experimenter. The other is Model II, in which the x-values are free to vary and are subject to error. 2. I have received numerous complaints from biomedical scientists that they have great difficulty in executing Model II linear regression analysis. This may explain the results of a Google Scholar search, which showed that the authors of articles in journals of physiology, pharmacology and biochemistry rarely use Model II regression analysis. 3. I repeat my previous arguments in favour of using least products linear regression analysis for Model II regressions. I review three methods for executing ordinary least products (OLP) and weighted least products (WLP) regression analysis: (i) scientific calculator and/or computer spreadsheet; (ii) specific purpose computer programs; and (iii) general purpose computer programs. 4. Using a scientific calculator and/or computer spreadsheet, it is easy to obtain correct values for OLP slope and intercept, but the corresponding 95% confidence intervals (CI) are inaccurate. 5. Using specific purpose computer programs, the freeware computer program smatr gives the correct OLP regression coefficients and obtains 95% CI by bootstrapping. In addition, smatr can be used to compare the slopes of OLP lines. 6. When using general purpose computer programs, I recommend the commercial programs systat and Statistica for those who regularly undertake linear regression analysis and I give step-by-step instructions in the Supplementary Information as to how to use loss functions. © 2011 The Author. Clinical and Experimental Pharmacology and Physiology. © 2011 Blackwell Publishing Asia Pty Ltd.
External Tank Liquid Hydrogen (LH2) Prepress Regression Analysis Independent Review Technical Consultation Report

Science.gov (United States)

Parsons, Vickie s.

2009-01-01

The request to conduct an independent review of regression models, developed for determining the expected Launch Commit Criteria (LCC) External Tank (ET)-04 cycle count for the Space Shuttle ET tanking process, was submitted to the NASA Engineering and Safety Center NESC on September 20, 2005. The NESC team performed an independent review of regression models documented in Prepress Regression Analysis, Tom Clark and Angela Krenn, 10/27/05. This consultation consisted of a peer review by statistical experts of the proposed regression models provided in the Prepress Regression Analysis. This document is the consultation's final report.
Tutorial on Biostatistics: Linear Regression Analysis of Continuous Correlated Eye Data.

Science.gov (United States)

Ying, Gui-Shuang; Maguire, Maureen G; Glynn, Robert; Rosner, Bernard

2017-04-01

To describe and demonstrate appropriate linear regression methods for analyzing correlated continuous eye data. We describe several approaches to regression analysis involving both eyes, including mixed effects and marginal models under various covariance structures to account for inter-eye correlation. We demonstrate, with SAS statistical software, applications in a study comparing baseline refractive error between one eye with choroidal neovascularization (CNV) and the unaffected fellow eye, and in a study determining factors associated with visual field in the elderly. When refractive error from both eyes were analyzed with standard linear regression without accounting for inter-eye correlation (adjusting for demographic and ocular covariates), the difference between eyes with CNV and fellow eyes was 0.15 diopters (D; 95% confidence interval, CI -0.03 to 0.32D, p = 0.10). Using a mixed effects model or a marginal model, the estimated difference was the same but with narrower 95% CI (0.01 to 0.28D, p = 0.03). Standard regression for visual field data from both eyes provided biased estimates of standard error (generally underestimated) and smaller p-values, while analysis of the worse eye provided larger p-values than mixed effects models and marginal models. In research involving both eyes, ignoring inter-eye correlation can lead to invalid inferences. Analysis using only right or left eyes is valid, but decreases power. Worse-eye analysis can provide less power and biased estimates of effect. Mixed effects or marginal models using the eye as the unit of analysis should be used to appropriately account for inter-eye correlation and maximize power and precision.
Non-stationary hydrologic frequency analysis using B-spline quantile regression

Science.gov (United States)

Nasri, B.; Bouezmarni, T.; St-Hilaire, A.; Ouarda, T. B. M. J.

2017-11-01

Hydrologic frequency analysis is commonly used by engineers and hydrologists to provide the basic information on planning, design and management of hydraulic and water resources systems under the assumption of stationarity. However, with increasing evidence of climate change, it is possible that the assumption of stationarity, which is prerequisite for traditional frequency analysis and hence, the results of conventional analysis would become questionable. In this study, we consider a framework for frequency analysis of extremes based on B-Spline quantile regression which allows to model data in the presence of non-stationarity and/or dependence on covariates with linear and non-linear dependence. A Markov Chain Monte Carlo (MCMC) algorithm was used to estimate quantiles and their posterior distributions. A coefficient of determination and Bayesian information criterion (BIC) for quantile regression are used in order to select the best model, i.e. for each quantile, we choose the degree and number of knots of the adequate B-spline quantile regression model. The method is applied to annual maximum and minimum streamflow records in Ontario, Canada. Climate indices are considered to describe the non-stationarity in the variable of interest and to estimate the quantiles in this case. The results show large differences between the non-stationary quantiles and their stationary equivalents for an annual maximum and minimum discharge with high annual non-exceedance probabilities.
Alternative Methods of Regression

CERN Document Server

Birkes, David

2011-01-01

Of related interest. Nonlinear Regression Analysis and its Applications Douglas M. Bates and Donald G. Watts ".an extraordinary presentation of concepts and methods concerning the use and analysis of nonlinear regression models.highly recommend[ed].for anyone needing to use and/or understand issues concerning the analysis of nonlinear regression models." --Technometrics This book provides a balance between theory and practice supported by extensive displays of instructive geometrical constructs. Numerous in-depth case studies illustrate the use of nonlinear regression analysis--with all data s
modelling relationship between rainfall variability and yields

African Journals Online (AJOL)

, S. and ... factors to rice yield. Adebayo and Adebayo (1997) developed double log multiple regression model to predict rice yield in Adamawa State, Nigeria. The general form of .... the second are the crop yield/values for millet and sorghum ...
Path and correlation analysis of perennial ryegrass (Lolium perenne L.) seed yield components

DEFF Research Database (Denmark)

Abel, Simon; Gislum, René; Boelt, Birte

2017-01-01

Maximum perennial ryegrass seed production potential is substantially greater than harvested yields with harvested yields representing only 20% of calculated potential. Similar to wheat, maize and other agriculturally important crops, seed yield is highly dependent on a number of interacting seed...... yield components. This research was performed to apply and describe path analysis of perennial ryegrass seed yield components in relation to harvested seed yields. Utilising extensive yield components which included subdividing reproductive inflorescences into five size categories, path analysis...... was undertaken assuming a unidirectional causal-admissible relationship between seed yield components and harvested seed yield in six commercial seed production fields. Both spikelets per inflorescence and florets per spikelet had a significant (p seed yield; however, total...
On macroeconomic values investigation using fuzzy linear regression analysis

Directory of Open Access Journals (Sweden)

Richard Pospíšil

2017-06-01

Full Text Available The theoretical background for abstract formalization of the vague phenomenon of complex systems is the fuzzy set theory. In the paper, vague data is defined as specialized fuzzy sets - fuzzy numbers and there is described a fuzzy linear regression model as a fuzzy function with fuzzy numbers as vague parameters. To identify the fuzzy coefficients of the model, the genetic algorithm is used. The linear approximation of the vague function together with its possibility area is analytically and graphically expressed. A suitable application is performed in the tasks of the time series fuzzy regression analysis. The time-trend and seasonal cycles including their possibility areas are calculated and expressed. The examples are presented from the economy field, namely the time-development of unemployment, agricultural production and construction respectively between 2009 and 2011 in the Czech Republic. The results are shown in the form of the fuzzy regression models of variables of time series. For the period 2009-2011, the analysis assumptions about seasonal behaviour of variables and the relationship between them were confirmed; in 2010, the system behaved fuzzier and the relationships between the variables were vaguer, that has a lot of causes, from the different elasticity of demand, through state interventions to globalization and transnational impacts.
Remotely sensed rice yield prediction using multi-temporal NDVI data derived from NOAA's-AVHRR.

Science.gov (United States)

Huang, Jingfeng; Wang, Xiuzhen; Li, Xinxing; Tian, Hanqin; Pan, Zhuokun

2013-01-01

Grain-yield prediction using remotely sensed data have been intensively studied in wheat and maize, but such information is limited in rice, barley, oats and soybeans. The present study proposes a new framework for rice-yield prediction, which eliminates the influence of the technology development, fertilizer application, and management improvement and can be used for the development and implementation of provincial rice-yield predictions. The technique requires the collection of remotely sensed data over an adequate time frame and a corresponding record of the region's crop yields. Longer normalized-difference-vegetation-index (NDVI) time series are preferable to shorter ones for the purposes of rice-yield prediction because the well-contrasted seasons in a longer time series provide the opportunity to build regression models with a wide application range. A regression analysis of the yield versus the year indicated an annual gain in the rice yield of 50 to 128 kg ha(-1). Stepwise regression models for the remotely sensed rice-yield predictions have been developed for five typical rice-growing provinces in China. The prediction models for the remotely sensed rice yield indicated that the influences of the NDVIs on the rice yield were always positive. The association between the predicted and observed rice yields was highly significant without obvious outliers from 1982 to 2004. Independent validation found that the overall relative error is approximately 5.82%, and a majority of the relative errors were less than 5% in 2005 and 2006, depending on the study area. The proposed models can be used in an operational context to predict rice yields at the provincial level in China. The methodologies described in the present paper can be applied to any crop for which a sufficient time series of NDVI data and the corresponding historical yield information are available, as long as the historical yield increases significantly.
REGRESSION ANALYSIS OF SEA-SURFACE-TEMPERATURE PATTERNS FOR THE NORTH PACIFIC OCEAN.

Science.gov (United States)

SEA WATER, *SURFACE TEMPERATURE, *OCEANOGRAPHIC DATA, PACIFIC OCEAN, REGRESSION ANALYSIS , STATISTICAL ANALYSIS, UNDERWATER EQUIPMENT, DETECTION, UNDERWATER COMMUNICATIONS, DISTRIBUTION, THERMAL PROPERTIES, COMPUTERS.
Regression analysis understanding and building business and economic models using Excel

CERN Document Server

Wilson, J Holton

2012-01-01

The technique of regression analysis is used so often in business and economics today that an understanding of its use is necessary for almost everyone engaged in the field. This book will teach you the essential elements of building and understanding regression models in a business/economic context in an intuitive manner. The authors take a non-theoretical treatment that is accessible even if you have a limited statistical background. It is specifically designed to teach the correct use of regression, while advising you of its limitations and teaching about common pitfalls. This book describe
Yield gap analysis of Chickpea under semi-arid conditions: A simulation study

OpenAIRE

seyed Reza Amiri Deh ahmadi; mehdi parsa; mohammad bannayan aval; mahdi nassiri mahallati

2016-01-01

Yield gap analysis provides an essential framework to prioritize research and policy efforts aimed at reducing yield constraints. To identify options for increasing chickpea yield, the SSM-chickpea model was parameterized and evaluated to analyze yield potentials, water limited yields and yield gaps for nine regions representing major chickpea-growing areas of Razavi Khorasan province. The average potential yield of chickpea for the locations was 2251 kg ha-1, while the water limited yield wa...
Nonlinear regression analysis for evaluating tracer binding parameters using the programmable K1003 desk computer

International Nuclear Information System (INIS)

Sarrach, D.; Strohner, P.

1986-01-01

The Gauss-Newton algorithm has been used to evaluate tracer binding parameters of RIA by nonlinear regression analysis. The calculations were carried out on the K1003 desk computer. Equations for simple binding models and its derivatives are presented. The advantages of nonlinear regression analysis over linear regression are demonstrated
Regression analysis for the social sciences

CERN Document Server

Gordon, Rachel A

2010-01-01

The book provides graduate students in the social sciences with the basic skills that they need to estimate, interpret, present, and publish basic regression models using contemporary standards. Key features of the book include: interweaving the teaching of statistical concepts with examples developed for the course from publicly-available social science data or drawn from the literature. thorough integration of teaching statistical theory with teaching data processing and analysis. teaching of both SAS and Stata "side-by-side" and use of chapter exercises in which students practice programming and interpretation on the same data set and course exercises in which students can choose their own research questions and data set.
Multivariate Linear Regression and CART Regression Analysis of TBM Performance at Abu Hamour Phase-I Tunnel

Science.gov (United States)

Jakubowski, J.; Stypulkowski, J. B.; Bernardeau, F. G.

2017-12-01

The first phase of the Abu Hamour drainage and storm tunnel was completed in early 2017. The 9.5 km long, 3.7 m diameter tunnel was excavated with two Earth Pressure Balance (EPB) Tunnel Boring Machines from Herrenknecht. TBM operation processes were monitored and recorded by Data Acquisition and Evaluation System. The authors coupled collected TBM drive data with available information on rock mass properties, cleansed, completed with secondary variables and aggregated by weeks and shifts. Correlations and descriptive statistics charts were examined. Multivariate Linear Regression and CART regression tree models linking TBM penetration rate (PR), penetration per revolution (PPR) and field penetration index (FPI) with TBM operational and geotechnical characteristics were performed for the conditions of the weak/soft rock of Doha. Both regression methods are interpretable and the data were screened with different computational approaches allowing enriched insight. The primary goal of the analysis was to investigate empirical relations between multiple explanatory and responding variables, to search for best subsets of explanatory variables and to evaluate the strength of linear and non-linear relations. For each of the penetration indices, a predictive model coupling both regression methods was built and validated. The resultant models appeared to be stronger than constituent ones and indicated an opportunity for more accurate and robust TBM performance predictions.
Evolution of Grain Yield and its Components Relationships in Bread Wheat Genotypes under Full Irrigation and Terminal Water Stress Conditions Using Multivariate Statistical Analysis

Directory of Open Access Journals (Sweden)

S Mohammadi

2014-07-01

Full Text Available To study relationships between effective traits on wheat grain yield, the varieties Zarrin and Alvand, and some promising lines i.e. C-81-4, C-81-10, C-81-14 and C-82-12 were investigated at three sowing dates including 10 October, 1 November and 21 November. The experiment was carried out using strip plot in RCBD with three replications under two different water conditions including full-irrigation and terminal water stress at Miyandoab Agricultural Research Station in 2005-06 and 2006-07 cropping seasons. The results showed that under both full irrigation and terminal water stress conditions, grain yield had positive and significant correlation with days to heading, days to maturity, plant height, number of spikes/m2 and 1000 grain weight. Stepwise regression analysis revealed that 83 percent of yield variation under non-stressed conditions could be determined by days to maturity and number of spikes/m2 (R2 = 83% whereas these traits explained 87% of yield variation under stress conditions (R2= 87%. Path analysis indicated that number of spikes/m2 and days to maturity had the greatest positive direct and indirect effect on grain yield, under both conditions. The results of factor analysis under non-stressed condition showed that three factors explained 77% of total variation; these factors were called grain yield components, grain characteristics and plant phonology. Under non-stressed condition two factors (that were called grain yield and phenology, and plant morphology explained 88% of total variation. Cluster analysis through ward method, classified days to maturity and number of spikes/m2 in the same cluster where the grain yield was put under both conditions. It was concluded that under different sowing dates, selection based on days to maturity and number spikes/m2 could indirectly led to higher yield under both normal and water stress conditions.
Composite marginal quantile regression analysis for longitudinal adolescent body mass index data.

Science.gov (United States)

Yang, Chi-Chuan; Chen, Yi-Hau; Chang, Hsing-Yi

2017-09-20

Childhood and adolescenthood overweight or obesity, which may be quantified through the body mass index (BMI), is strongly associated with adult obesity and other health problems. Motivated by the child and adolescent behaviors in long-term evolution (CABLE) study, we are interested in individual, family, and school factors associated with marginal quantiles of longitudinal adolescent BMI values. We propose a new method for composite marginal quantile regression analysis for longitudinal outcome data, which performs marginal quantile regressions at multiple quantile levels simultaneously. The proposed method extends the quantile regression coefficient modeling method introduced by Frumento and Bottai (Biometrics 2016; 72:74-84) to longitudinal data accounting suitably for the correlation structure in longitudinal observations. A goodness-of-fit test for the proposed modeling is also developed. Simulation results show that the proposed method can be much more efficient than the analysis without taking correlation into account and the analysis performing separate quantile regressions at different quantile levels. The application to the longitudinal adolescent BMI data from the CABLE study demonstrates the practical utility of our proposal. Copyright © 2017 John Wiley & Sons, Ltd. Copyright © 2017 John Wiley & Sons, Ltd.
Treating experimental data of inverse kinetic method by unitary linear regression analysis

International Nuclear Information System (INIS)

Zhao Yusen; Chen Xiaoliang

2009-01-01

The theory of treating experimental data of inverse kinetic method by unitary linear regression analysis was described. Not only the reactivity, but also the effective neutron source intensity could be calculated by this method. Computer code was compiled base on the inverse kinetic method and unitary linear regression analysis. The data of zero power facility BFS-1 in Russia were processed and the results were compared. The results show that the reactivity and the effective neutron source intensity can be obtained correctly by treating experimental data of inverse kinetic method using unitary linear regression analysis and the precision of reactivity measurement is improved. The central element efficiency can be calculated by using the reactivity. The result also shows that the effect to reactivity measurement caused by external neutron source should be considered when the reactor power is low and the intensity of external neutron source is strong. (authors)
Regression analysis of informative current status data with the additive hazards model.

Science.gov (United States)

Zhao, Shishun; Hu, Tao; Ma, Ling; Wang, Peijie; Sun, Jianguo

2015-04-01

This paper discusses regression analysis of current status failure time data arising from the additive hazards model in the presence of informative censoring. Many methods have been developed for regression analysis of current status data under various regression models if the censoring is noninformative, and also there exists a large literature on parametric analysis of informative current status data in the context of tumorgenicity experiments. In this paper, a semiparametric maximum likelihood estimation procedure is presented and in the method, the copula model is employed to describe the relationship between the failure time of interest and the censoring time. Furthermore, I-splines are used to approximate the nonparametric functions involved and the asymptotic consistency and normality of the proposed estimators are established. A simulation study is conducted and indicates that the proposed approach works well for practical situations. An illustrative example is also provided.
YIELD STABILITY OF NEW HYBRID RICE ACROSS LOCATIONS

Directory of Open Access Journals (Sweden)

Satoto

2016-02-01

Full Text Available The adaptation of hybrid rice varieties mostly are in specific location and season, but there are some of the varieties have a wide adaptation then adopted by the farmer in the large area. Replicated yield trials were conducted to study the stability of hybrid rice yield and identify the best location to optimize their yield per ha. The trials were conducted in three location such as Sukamandi, Salatiga and Malang during two seasons in 2011. Data across location and season were analazed by using AMMI and Eberhart Russel methods. The AMMI analysis showed that the IR79156A/PK88 was adaptable to favorable environments but unstable. This hybrid is always performing well and produce the higher yield compare to check variety. Some of other hybrids were good only in specific location, i.e. IR62829A/BP2280-1E-12-22 and IR58029A/BP2280-1E-12-22. Those hybrids produced higher yield in Salatiga and Malang, respectively. Seem to AMMI analysis, the result of Eberhart and Russell method also showed that IR79156A/PK81 was the best hybrid with regression slope (b around 1 with the yield average higher than average of all hybrids. It indicated that this hybrid has a wide adaptation and probably can be cultivated in the wider ecosystem.

Credit Scoring Problem Based on Regression Analysis

OpenAIRE

Khassawneh, Bashar Suhil Jad Allah

2014-01-01

ABSTRACT: This thesis provides an explanatory introduction to the regression models of data mining and contains basic definitions of key terms in the linear, multiple and logistic regression models. Meanwhile, the aim of this study is to illustrate fitting models for the credit scoring problem using simple linear, multiple linear and logistic regression models and also to analyze the found model functions by statistical tools. Keywords: Data mining, linear regression, logistic regression....
MULGRES: a computer program for stepwise multiple regression analysis

Science.gov (United States)

A. Jeff Martin

1971-01-01

MULGRES is a computer program source deck that is designed for multiple regression analysis employing the technique of stepwise deletion in the search for most significant variables. The features of the program, along with inputs and outputs, are briefly described, with a note on machine compatibility.
Real-time regression analysis with deep convolutional neural networks

OpenAIRE

Huerta, E. A.; George, Daniel; Zhao, Zhizhen; Allen, Gabrielle

2018-01-01

We discuss the development of novel deep learning algorithms to enable real-time regression analysis for time series data. We showcase the application of this new method with a timely case study, and then discuss the applicability of this approach to tackle similar challenges across science domains.
[Comparison of application of Cochran-Armitage trend test and linear regression analysis for rate trend analysis in epidemiology study].

Science.gov (United States)

Wang, D Z; Wang, C; Shen, C F; Zhang, Y; Zhang, H; Song, G D; Xue, X D; Xu, Z L; Zhang, S; Jiang, G H

2017-05-10

We described the time trend of acute myocardial infarction (AMI) from 1999 to 2013 in Tianjin incidence rate with Cochran-Armitage trend (CAT) test and linear regression analysis, and the results were compared. Based on actual population, CAT test had much stronger statistical power than linear regression analysis for both overall incidence trend and age specific incidence trend (Cochran-Armitage trend P valuelinear regression P value). The statistical power of CAT test decreased, while the result of linear regression analysis remained the same when population size was reduced by 100 times and AMI incidence rate remained unchanged. The two statistical methods have their advantages and disadvantages. It is necessary to choose statistical method according the fitting degree of data, or comprehensively analyze the results of two methods.
Assessment of adaptability and stability of grain yield in bread wheat genotypes under different sowing times in Punjab

International Nuclear Information System (INIS)

Anwar, J.; Hussain, M.; Ali, M.A.; Subhani, G.M.; Munir, M.

2011-01-01

Twenty advanced lines/genotypes of wheat including two check varieties were sown under two different sowing times through out the Punjab province at 18 different locations with diverse environments to study their stability and adaptability. Normal sowing was done in second week of November 2007 while the delayed sowing was completed during second week of December 2007 during crop season 2007-08. The pooled analysis of variance showed significant differences among environments and genotypes for grain yield demonstrating the presence of considerable variations (p<0.01) among genotypes as well as diversity of growing environments at various locations for both normal and late sown wheat crops. The highest average grain yield was obtained at Jalandar Seed Farm, Arifwala and Pak. German Farm, Multan for normal and delayed sown crops, respectively. Most of the locations emerged as high yielding in normal sowing compared to late sown crop. Dendrograms of 18 locations based on the average yield of 20 wheat genotypes grown under normal and late sown crop revealed two main clusters. Under both normal and late sowing, none of the varieties exceeded the check Seher-2006, however, the check was followed by the advanced lines V-04022 and V-05066 for normal sown crop and Shafaq-2006, V-05066 and V-04022 under delayed sowing. All the genotypes revealed decline in grain yield for late sown wheat crop. The analysis of stability based on mean grain yield, regression coefficient and deviation from regression advocated that the cultivars V-05066 and V-03BT007 were most stable and adapted to diverse environmental conditions of Punjab. These cultivars revealed unit regression and non-significant deviations from regression. The check variety Seher-2006 produced maximum yield for both sowing times that suggested its consistent and stable performance across the environments. (author)
Yield trends and yield gap analysis of major crops in the world

OpenAIRE

Hengsdijk, H.; Langeveld, J.W.A.

2009-01-01

This study aims to quantify the gap between current and potential yields of major crops in the world, and the production constraints that contribute to this yield gap. Using an expert-based evaluation of yield gaps and the literature, global and regional yields and yield trends of major crops are quantified, yield gaps evaluated by crop experts, current yield progress by breeding estimated, and different yield projections compared. Results show decreasing yield growth for wheat and rice, but ...
Using yield gap analysis to give sustainable intensification local meaning

NARCIS (Netherlands)

Silva, João Vasco

2017-01-01

Yield gap analysis is useful to understand the relative contribution of growth-defining, -limiting and -reducing factors to actual yields. This is traditionally performed at the field level using mechanistic crop growth simulation models, and directly up-scaled to the regional and global levels
The evolution of GDP in USA using cyclic regression analysis

OpenAIRE

Catalin Angelo IOAN; Gina IOAN

2013-01-01

Based on the four major types of economic cycles (Kondratieff, Juglar, Kitchin, Kuznet), the paper aims to determine their actual length (for the U.S. economy) using cyclic regressions based on Fourier analysis.
THE YIELD ANALYSIS IN PRODUCTION COMPANIES BY AUDIT

Directory of Open Access Journals (Sweden)

Tuğçe Uzun Kocamış

2015-10-01

measure efficiency, trying to establish the relationship between cause and effect which should be. Yield and fire investigations are conducted on books and documents primarily. When the differences are occurred authorities can deduction to ex officio. it is obvious that there are some difficulties in practice and taxpayers does not have standards. In our article, inventory control were discussed within the concept of tax audits. Yield analysis examining all aspects and it is devoted to the application for ensuring that the subject more understandable.
Quantile regression for the statistical analysis of immunological data with many non-detects.

Science.gov (United States)

Eilers, Paul H C; Röder, Esther; Savelkoul, Huub F J; van Wijk, Roy Gerth

2012-07-07

Immunological parameters are hard to measure. A well-known problem is the occurrence of values below the detection limit, the non-detects. Non-detects are a nuisance, because classical statistical analyses, like ANOVA and regression, cannot be applied. The more advanced statistical techniques currently available for the analysis of datasets with non-detects can only be used if a small percentage of the data are non-detects. Quantile regression, a generalization of percentiles to regression models, models the median or higher percentiles and tolerates very high numbers of non-detects. We present a non-technical introduction and illustrate it with an implementation to real data from a clinical trial. We show that by using quantile regression, groups can be compared and that meaningful linear trends can be computed, even if more than half of the data consists of non-detects. Quantile regression is a valuable addition to the statistical methods that can be used for the analysis of immunological datasets with non-detects.
Optimal choice of basis functions in the linear regression analysis

International Nuclear Information System (INIS)

Khotinskij, A.M.

1988-01-01

Problem of optimal choice of basis functions in the linear regression analysis is investigated. Step algorithm with estimation of its efficiency, which holds true at finite number of measurements, is suggested. Conditions, providing the probability of correct choice close to 1 are formulated. Application of the step algorithm to analysis of decay curves is substantiated. 8 refs
Using historical wafermap data for automated yield analysis

International Nuclear Information System (INIS)

Tobin, K.W.; Karnowski, T.P.; Gleason, S.S.; Jensen, D.; Lakhani, F.

1999-01-01

To be productive and profitable in a modern semiconductor fabrication environment, large amounts of manufacturing data must be collected, analyzed, and maintained. This includes data collected from in- and off-line wafer inspection systems and from the process equipment itself. This data is increasingly being used to design new processes, control and maintain tools, and to provide the information needed for rapid yield learning and prediction. Because of increasing device complexity, the amount of data being generated is outstripping the yield engineer close-quote s ability to effectively monitor and correct unexpected trends and excursions. The 1997 SIA National Technology Roadmap for Semiconductors highlights a need to address these issues through open-quotes automated data reduction algorithms to source defects from multiple data sources and to reduce defect sourcing time.close quotes SEMATECH and the Oak Ridge National Laboratory have been developing new strategies and technologies for providing the yield engineer with higher levels of assisted data reduction for the purpose of automated yield analysis. In this article, we will discuss the current state of the art and trends in yield management automation. copyright 1999 American Vacuum Society
Genetic analysis and QTL mapping of maize yield and associate ...

African Journals Online (AJOL)

STORAGESEVER

2008-06-17

Jun 17, 2008 ... strongly influenced by both genotype and environment, and the interaction of ... associated with yield components as well as secondary ... QTLs that control grain yield under drought ... statistical analysis (ANOVA etc) of phenotypic traits was carried out .... which means that the loci had stable heredity.
Climatic and technological ceilings for Chinese rice stagnation based on yield gaps and yield trend pattern analysis.

Science.gov (United States)

Zhang, Tianyi; Yang, Xiaoguang; Wang, Hesong; Li, Yong; Ye, Qing

2014-04-01

Climatic or technological ceilings could cause yield stagnation. Thus, identifying the principal reasons for yield stagnation within the context of the local climate and socio-economic conditions are essential for informing regional agricultural policies. In this study, we identified the climatic and technological ceilings for seven rice-production regions in China based on yield gaps and on a yield trend pattern analysis for the period 1980-2010. The results indicate that 54.9% of the counties sampled experienced yield stagnation since the 1980. The potential yield ceilings in northern and eastern China decreased to a greater extent than in other regions due to the accompanying climate effects of increases in temperature and decreases in radiation. This may be associated with yield stagnation and halt occurring in approximately 49.8-57.0% of the sampled counties in these areas. South-western China exhibited a promising scope for yield improvement, showing the greatest yield gap (30.6%), whereas the yields were stagnant in 58.4% of the sampled counties. This finding suggests that efforts to overcome the technological ceiling must be given priority so that the available exploitable yield gap can be achieved. North-eastern China, however, represents a noteworthy exception. In the north-central area of this region, climate change has increased the yield potential ceiling, and this increase has been accompanied by the most rapid increase in actual yield: 1.02 ton ha(-1) per decade. Therefore, north-eastern China shows a great potential for rice production, which is favoured by the current climate conditions and available technology level. Additional environmentally friendly economic incentives might be considered in this region. © 2013 John Wiley & Sons Ltd.
Regression analysis for the social sciences

CERN Document Server

Gordon, Rachel A

2015-01-01

Provides graduate students in the social sciences with the basic skills they need to estimate, interpret, present, and publish basic regression models using contemporary standards. Key features of the book include: interweaving the teaching of statistical concepts with examples developed for the course from publicly-available social science data or drawn from the literature. thorough integration of teaching statistical theory with teaching data processing and analysis. teaching of Stata and use of chapter exercises in which students practice programming and interpretation on the same data set. A separate set of exercises allows students to select a data set to apply the concepts learned in each chapter to a research question of interest to them, all updated for this edition.
CHANGES IN CLIMATIC CHARACTERISTICS AND CROP YIELD IN KWARA STATE (NIGERIA

Directory of Open Access Journals (Sweden)

O. Oriola

2017-01-01

Full Text Available This paper assessed the vagaries of climatic elements on crop yield in Kwara State with a view to predicting the future climatic suitability level for selected crops in the state. Descriptive and infrential statistics analytical methods were used to examine the pattern of climatic elements for a period of 30 years. Analysis of variance was used to examine the variations in crop yield and also to determine whether or not significant differences in the harvests of the period under investigation. Correlation analysis was used to determine the relationship between climatic elements and crop yield while multiple regression analysis was used to determine the contribution of each climatic elements to crop yield. Time series analysis was used to project crop yield from 2014 to 2025. GAEZ model was adopted to determine the climatic suitability for the selected crops over time 1960 - 2050 and ArcGIS 10.3 software was used to produce the crop suitability maps. The result revealed that cassava, yam, maize and cowpea would be less suitable for production with the rate at which the climate is changing. The result also revealed that the climatic suitability level for cassava, yam, maize and cowpea would reduce drastically with time. The prediction shows severe impacts of changes in the selected climatic elements on both overall climatic suitability and crop the selected crops yield for by 2050.
A method for nonlinear exponential regression analysis

Science.gov (United States)

Junkin, B. G.

1971-01-01

A computer-oriented technique is presented for performing a nonlinear exponential regression analysis on decay-type experimental data. The technique involves the least squares procedure wherein the nonlinear problem is linearized by expansion in a Taylor series. A linear curve fitting procedure for determining the initial nominal estimates for the unknown exponential model parameters is included as an integral part of the technique. A correction matrix was derived and then applied to the nominal estimate to produce an improved set of model parameters. The solution cycle is repeated until some predetermined criterion is satisfied.
Analysis of Functional Data with Focus on Multinomial Regression and Multilevel Data

DEFF Research Database (Denmark)

Mousavi, Seyed Nourollah

Functional data analysis (FDA) is a fast growing area in statistical research with increasingly diverse range of application from economics, medicine, agriculture, chemometrics, etc. Functional regression is an area of FDA which has received the most attention both in aspects of application...... and methodological development. Our main Functional data analysis (FDA) is a fast growing area in statistical research with increasingly diverse range of application from economics, medicine, agriculture, chemometrics, etc. Functional regression is an area of FDA which has received the most attention both in aspects...
Quasi-experimental evidence on tobacco tax regressivity.

Science.gov (United States)

Koch, Steven F

2018-01-01

Tobacco taxes are known to reduce tobacco consumption and to be regressive, such that tobacco control policy may have the perverse effect of further harming the poor. However, if tobacco consumption falls faster amongst the poor than the rich, tobacco control policy can actually be progressive. We take advantage of persistent and committed tobacco control activities in South Africa to examine the household tobacco expenditure burden. For the analysis, we make use of two South African Income and Expenditure Surveys (2005/06 and 2010/11) that span a series of such tax increases and have been matched across the years, yielding 7806 matched pairs of tobacco consuming households and 4909 matched pairs of cigarette consuming households. By matching households across the surveys, we are able to examine both the regressivity of the household tobacco burden, and any change in that regressivity, and since tobacco taxes have been a consistent component of tobacco prices, our results also relate to the regressivity of tobacco taxes. Like previous research into cigarette and tobacco expenditures, we find that the tobacco burden is regressive; thus, so are tobacco taxes. However, we find that over the five-year period considered, the tobacco burden has decreased, and, most importantly, falls less heavily on the poor. Thus, the tobacco burden and the tobacco tax is less regressive in 2010/11 than in 2005/06. Thus, increased tobacco taxes can, in at least some circumstances, reduce the financial burden that tobacco places on households. Copyright © 2017 Elsevier Ltd. All rights reserved.
Regression analysis of a chemical reaction fouling model

International Nuclear Information System (INIS)

Vasak, F.; Epstein, N.

1996-01-01

A previously reported mathematical model for the initial chemical reaction fouling of a heated tube is critically examined in the light of the experimental data for which it was developed. A regression analysis of the model with respect to that data shows that the reference point upon which the two adjustable parameters of the model were originally based was well chosen, albeit fortuitously. (author). 3 refs., 2 tabs., 2 figs

Application of Artificial Neural Networks in Canola Crop Yield Prediction

Directory of Open Access Journals (Sweden)

S. J. Sajadi

2014-02-01

Full Text Available Crop yield prediction has an important role in agricultural policies such as specification of the crop price. Crop yield prediction researches have been based on regression analysis. In this research canola yield was predicted using Artificial Neural Networks (ANN using 11 crop year climate data (1998-2009 in Gonbad-e-Kavoos region of Golestan province. ANN inputs were mean weekly rainfall, mean weekly temperature, mean weekly relative humidity and mean weekly sun shine hours and ANN output was canola yield (kg/ha. Multi-Layer Perceptron networks (MLP with Levenberg-Marquardt backpropagation learning algorithm was used for crop yield prediction and Root Mean Square Error (RMSE and square of the Correlation Coefficient (R2 criterions were used to evaluate the performance of the ANN. The obtained results show that the 13-20-1 network has the lowest RMSE equal to 101.235 and maximum value of R2 equal to 0.997 and is suitable for predicting canola yield with climate factors.
[A SAS marco program for batch processing of univariate Cox regression analysis for great database].

Science.gov (United States)

Yang, Rendong; Xiong, Jie; Peng, Yangqin; Peng, Xiaoning; Zeng, Xiaomin

2015-02-01

To realize batch processing of univariate Cox regression analysis for great database by SAS marco program. We wrote a SAS macro program, which can filter, integrate, and export P values to Excel by SAS9.2. The program was used for screening survival correlated RNA molecules of ovarian cancer. A SAS marco program could finish the batch processing of univariate Cox regression analysis, the selection and export of the results. The SAS macro program has potential applications in reducing the workload of statistical analysis and providing a basis for batch processing of univariate Cox regression analysis.
Sensitivity analysis and optimization of system dynamics models : Regression analysis and statistical design of experiments

NARCIS (Netherlands)

Kleijnen, J.P.C.

1995-01-01

This tutorial discusses what-if analysis and optimization of System Dynamics models. These problems are solved, using the statistical techniques of regression analysis and design of experiments (DOE). These issues are illustrated by applying the statistical techniques to a System Dynamics model for
Application of multilinear regression analysis in modeling of soil ...

African Journals Online (AJOL)

The application of Multi-Linear Regression Analysis (MLRA) model for predicting soil properties in Calabar South offers a technical guide and solution in foundation designs problems in the area. Forty-five soil samples were collected from fifteen different boreholes at a different depth and 270 tests were carried out for CBR, ...
Genetic analysis of yield in peanut (Arachis hypogaea L.) using ...

African Journals Online (AJOL)

Jane

2011-07-20

Jul 20, 2011 ... only should the two major genes' effects be considered but also the polygene's effect should be considered in breeding to increase peanut yield. Key words: Peanut, yield, major gene plus polygene inheritance model, genetic analysis. INTRODUCTION. Peanut consists of diploid (2n = 2x = 20), tetraploid ...
Boosted beta regression.

Directory of Open Access Journals (Sweden)

Matthias Schmid

Full Text Available Regression analysis with a bounded outcome is a common problem in applied statistics. Typical examples include regression models for percentage outcomes and the analysis of ratings that are measured on a bounded scale. In this paper, we consider beta regression, which is a generalization of logit models to situations where the response is continuous on the interval (0,1. Consequently, beta regression is a convenient tool for analyzing percentage responses. The classical approach to fit a beta regression model is to use maximum likelihood estimation with subsequent AIC-based variable selection. As an alternative to this established - yet unstable - approach, we propose a new estimation technique called boosted beta regression. With boosted beta regression estimation and variable selection can be carried out simultaneously in a highly efficient way. Additionally, both the mean and the variance of a percentage response can be modeled using flexible nonlinear covariate effects. As a consequence, the new method accounts for common problems such as overdispersion and non-binomial variance structures.
Bias due to two-stage residual-outcome regression analysis in genetic association studies.

Science.gov (United States)

Demissie, Serkalem; Cupples, L Adrienne

2011-11-01

Association studies of risk factors and complex diseases require careful assessment of potential confounding factors. Two-stage regression analysis, sometimes referred to as residual- or adjusted-outcome analysis, has been increasingly used in association studies of single nucleotide polymorphisms (SNPs) and quantitative traits. In this analysis, first, a residual-outcome is calculated from a regression of the outcome variable on covariates and then the relationship between the adjusted-outcome and the SNP is evaluated by a simple linear regression of the adjusted-outcome on the SNP. In this article, we examine the performance of this two-stage analysis as compared with multiple linear regression (MLR) analysis. Our findings show that when a SNP and a covariate are correlated, the two-stage approach results in biased genotypic effect and loss of power. Bias is always toward the null and increases with the squared-correlation between the SNP and the covariate (). For example, for , 0.1, and 0.5, two-stage analysis results in, respectively, 0, 10, and 50% attenuation in the SNP effect. As expected, MLR was always unbiased. Since individual SNPs often show little or no correlation with covariates, a two-stage analysis is expected to perform as well as MLR in many genetic studies; however, it produces considerably different results from MLR and may lead to incorrect conclusions when independent variables are highly correlated. While a useful alternative to MLR under , the two -stage approach has serious limitations. Its use as a simple substitute for MLR should be avoided. © 2011 Wiley Periodicals, Inc.
Clonal stability of latex yield in eleven clones of Hevea brasiliensis Muell. Arg.

Directory of Open Access Journals (Sweden)

K.O. Omokhafe

2003-01-01

Full Text Available Eleven Hevea brasiliensis clones were evaluated for clonal stability of latex yield. A randomized complete block design was used with four replicates, two locations, seven years and three periods per year. Stability analysis was based on clone x year and clone x year x location interactions. Five stability parameters viz environmental variance, shukla's stability variance, regression of clonal latex yield on environmental index, variance due to regression and variance due to deviation from regression were applied. There was significant clone x environment effect at the two levels of interaction. Among the eleven clones, C 162 was outstanding for clonal stability and it can serve as donor parent for stability alleles. Three clones (C 76, C 150 and C 154 were also stable. The four stable clones (C 76, C 150, C 154 and C 162 are suitable for broad-spectrum recommendation for latex yield. Five clones (C 83, C 143, C 163, C 202 and RRIM 600 will require environment-specific recommendation because of their unstable phenotype. The stability feature of two clones (C 145 and C 159 was not clear and this will be investigated in subsequent studies.
An Original Stepwise Multilevel Logistic Regression Analysis of Discriminatory Accuracy

DEFF Research Database (Denmark)

Merlo, Juan; Wagner, Philippe; Ghith, Nermin

2016-01-01

BACKGROUND AND AIM: Many multilevel logistic regression analyses of "neighbourhood and health" focus on interpreting measures of associations (e.g., odds ratio, OR). In contrast, multilevel analysis of variance is rarely considered. We propose an original stepwise analytical approach that disting...
Yield trends and yield gap analysis of major crops in the world

NARCIS (Netherlands)

Hengsdijk, H.; Langeveld, J.W.A.

2009-01-01

This study aims to quantify the gap between current and potential yields of major crops in the world, and the production constraints that contribute to this yield gap. Using an expert-based evaluation of yield gaps and the literature, global and regional yields and yield trends of major crops are
Advanced statistics: linear regression, part I: simple linear regression.

Science.gov (United States)

Marill, Keith A

2004-01-01

Simple linear regression is a mathematical technique used to model the relationship between a single independent predictor variable and a single dependent outcome variable. In this, the first of a two-part series exploring concepts in linear regression analysis, the four fundamental assumptions and the mechanics of simple linear regression are reviewed. The most common technique used to derive the regression line, the method of least squares, is described. The reader will be acquainted with other important concepts in simple linear regression, including: variable transformations, dummy variables, relationship to inference testing, and leverage. Simplified clinical examples with small datasets and graphic models are used to illustrate the points. This will provide a foundation for the second article in this series: a discussion of multiple linear regression, in which there are multiple predictor variables.
Ordinary least square regression, orthogonal regression, geometric mean regression and their applications in aerosol science

International Nuclear Information System (INIS)

Leng Ling; Zhang Tianyi; Kleinman, Lawrence; Zhu Wei

2007-01-01

Regression analysis, especially the ordinary least squares method which assumes that errors are confined to the dependent variable, has seen a fair share of its applications in aerosol science. The ordinary least squares approach, however, could be problematic due to the fact that atmospheric data often does not lend itself to calling one variable independent and the other dependent. Errors often exist for both measurements. In this work, we examine two regression approaches available to accommodate this situation. They are orthogonal regression and geometric mean regression. Comparisons are made theoretically as well as numerically through an aerosol study examining whether the ratio of organic aerosol to CO would change with age
Yield gap analysis of Chickpea under semi-arid conditions: A simulation study

Directory of Open Access Journals (Sweden)

seyed Reza Amiri Deh ahmadi

2016-05-01

Full Text Available Yield gap analysis provides an essential framework to prioritize research and policy efforts aimed at reducing yield constraints. To identify options for increasing chickpea yield, the SSM-chickpea model was parameterized and evaluated to analyze yield potentials, water limited yields and yield gaps for nine regions representing major chickpea-growing areas of Razavi Khorasan province. The average potential yield of chickpea for the locations was 2251 kg ha-1, while the water limited yield was 1026 kg ha-1 indicating a 54% reduction in yield due to adverse soil moisture conditions. Also, the average irrigated and rainfed actual yields were respectively 64% and 79% less than simulated potential and water limited yields. Maximum and minimum yield gap between potential yield and actual yield were observed in Quchan and Torbat-jam respectively. Generally, yield gap showed an increasing trend from the north (including Nishabur, Mashhad, Quchan and Daregaz regions to the south of the province (Torbat- Jam and Gonabad. In addition, yield gap between simulated water limited potential yield and rainfed actual yield were very low because both simulated water limiting potential and average rainfed actual yields were low in these regions. Yield gap analysis provides an essential framework to prioritize research and policy efforts aimed at reducing yield constraints. To identify options for increasing chickpea yield, the SSM-chickpea model was parameterized and evaluated to analyze yield potentials, water limited yields and yield gaps for nine regions representing major chickpea-growing areas of Razavi Khorasan province. The average potential yield of chickpea for the locations was 2251 kg ha-1, while the water limited yield was 1026 kg ha-1 indicating a 54% reduction in yield due to adverse soil moisture conditions. Also, the average irrigated and rainfed actual yields were respectively 64% and 79% less than simulated potential and water limited yields
Genotype-environment interaction and phenotypic stability for girth growth and rubber yield of Hevea clones in São Paulo State, Brazil

Directory of Open Access Journals (Sweden)

Gonçalves Paulo de Souza

2003-01-01

Full Text Available The best-yielding, best vigour and most stable Hevea clones are identified by growing clones in different environments. However, research on the stability in Hevea brasiliensis (Willd. Adr. ex Juss. Muell.-Arg. is scarce. The objectives of this work were to assess genotype-environment interaction and determine stable genotypes. Stability analysis were performed on results for girth growth and rubber yield of seven clones from five comparative trials conducted over 10 years (girth growth and four years (rubber yield in São Paulo State, Brazil. Stability was estimated using the Eberhart and Russell (1966 method. Year by location and location variability were the dominant sources of interactions. The stability analysis identified GT 1 and IAN 873 as the most stable clones for girth growth and rubber yield respectively since their regression coefficients were almost the unity (b = 1 and they had one of the lowest deviations from regressions (S2di. Their coefficient of determination (R² was as high as 89.5% and 89.8% confirming their stability. In contrast, clones such as PB 235, PR 261, and RRIM 701 for girth growth and clones such as GT 1 for rubber yield with regression coefficients greater than one were regarded as sensitive to environment changes.
Yield gap analysis of feed-crop livestock systems

NARCIS (Netherlands)

Linden, van der Aart; Oosting, Simon J.; Ven, van de Gerrie W.J.; Veysset, Patrick; Boer, de Imke J.M.; Ittersum, van Martin K.

2018-01-01

Sustainable intensification is a strategy contributing to global food security. The scope for sustainable intensification in crop sciences can be assessed through yield gap analysis, using crop growth models based on concepts of production ecology. Recently, an analogous cattle production model
Forecasting urban water demand: A meta-regression analysis.

Science.gov (United States)

Sebri, Maamar

2016-12-01

Water managers and planners require accurate water demand forecasts over the short-, medium- and long-term for many purposes. These range from assessing water supply needs over spatial and temporal patterns to optimizing future investments and planning future allocations across competing sectors. This study surveys the empirical literature on the urban water demand forecasting using the meta-analytical approach. Specifically, using more than 600 estimates, a meta-regression analysis is conducted to identify explanations of cross-studies variation in accuracy of urban water demand forecasting. Our study finds that accuracy depends significantly on study characteristics, including demand periodicity, modeling method, forecasting horizon, model specification and sample size. The meta-regression results remain robust to different estimators employed as well as to a series of sensitivity checks performed. The importance of these findings lies in the conclusions and implications drawn out for regulators and policymakers and for academics alike. Copyright © 2016. Published by Elsevier Ltd.
A simple linear regression method for quantitative trait loci linkage analysis with censored observations.

Science.gov (United States)

Anderson, Carl A; McRae, Allan F; Visscher, Peter M

2006-07-01

Standard quantitative trait loci (QTL) mapping techniques commonly assume that the trait is both fully observed and normally distributed. When considering survival or age-at-onset traits these assumptions are often incorrect. Methods have been developed to map QTL for survival traits; however, they are both computationally intensive and not available in standard genome analysis software packages. We propose a grouped linear regression method for the analysis of continuous survival data. Using simulation we compare this method to both the Cox and Weibull proportional hazards models and a standard linear regression method that ignores censoring. The grouped linear regression method is of equivalent power to both the Cox and Weibull proportional hazards methods and is significantly better than the standard linear regression method when censored observations are present. The method is also robust to the proportion of censored individuals and the underlying distribution of the trait. On the basis of linear regression methodology, the grouped linear regression model is computationally simple and fast and can be implemented readily in freely available statistical software.
The use of cognitive ability measures as explanatory variables in regression analysis.

Science.gov (United States)

Junker, Brian; Schofield, Lynne Steuerle; Taylor, Lowell J

2012-12-01

Cognitive ability measures are often taken as explanatory variables in regression analysis, e.g., as a factor affecting a market outcome such as an individual's wage, or a decision such as an individual's education acquisition. Cognitive ability is a latent construct; its true value is unobserved. Nonetheless, researchers often assume that a test score , constructed via standard psychometric practice from individuals' responses to test items, can be safely used in regression analysis. We examine problems that can arise, and suggest that an alternative approach, a "mixed effects structural equations" (MESE) model, may be more appropriate in many circumstances.
Multiple Imputation of a Randomly Censored Covariate Improves Logistic Regression Analysis.

Science.gov (United States)

Atem, Folefac D; Qian, Jing; Maye, Jacqueline E; Johnson, Keith A; Betensky, Rebecca A

2016-01-01

Randomly censored covariates arise frequently in epidemiologic studies. The most commonly used methods, including complete case and single imputation or substitution, suffer from inefficiency and bias. They make strong parametric assumptions or they consider limit of detection censoring only. We employ multiple imputation, in conjunction with semi-parametric modeling of the censored covariate, to overcome these shortcomings and to facilitate robust estimation. We develop a multiple imputation approach for randomly censored covariates within the framework of a logistic regression model. We use the non-parametric estimate of the covariate distribution or the semiparametric Cox model estimate in the presence of additional covariates in the model. We evaluate this procedure in simulations, and compare its operating characteristics to those from the complete case analysis and a survival regression approach. We apply the procedures to an Alzheimer's study of the association between amyloid positivity and maternal age of onset of dementia. Multiple imputation achieves lower standard errors and higher power than the complete case approach under heavy and moderate censoring and is comparable under light censoring. The survival regression approach achieves the highest power among all procedures, but does not produce interpretable estimates of association. Multiple imputation offers a favorable alternative to complete case analysis and ad hoc substitution methods in the presence of randomly censored covariates within the framework of logistic regression.
Temporal trends in sperm count: a systematic review and meta-regression analysis.

Science.gov (United States)

Levine, Hagai; Jørgensen, Niels; Martino-Andrade, Anderson; Mendiola, Jaime; Weksler-Derri, Dan; Mindlis, Irina; Pinotti, Rachel; Swan, Shanna H

2017-11-01

Reported declines in sperm counts remain controversial today and recent trends are unknown. A definitive meta-analysis is critical given the predictive value of sperm count for fertility, morbidity and mortality. To provide a systematic review and meta-regression analysis of recent trends in sperm counts as measured by sperm concentration (SC) and total sperm count (TSC), and their modification by fertility and geographic group. PubMed/MEDLINE and EMBASE were searched for English language studies of human SC published in 1981-2013. Following a predefined protocol 7518 abstracts were screened and 2510 full articles reporting primary data on SC were reviewed. A total of 244 estimates of SC and TSC from 185 studies of 42 935 men who provided semen samples in 1973-2011 were extracted for meta-regression analysis, as well as information on years of sample collection and covariates [fertility group ('Unselected by fertility' versus 'Fertile'), geographic group ('Western', including North America, Europe Australia and New Zealand versus 'Other', including South America, Asia and Africa), age, ejaculation abstinence time, semen collection method, method of measuring SC and semen volume, exclusion criteria and indicators of completeness of covariate data]. The slopes of SC and TSC were estimated as functions of sample collection year using both simple linear regression and weighted meta-regression models and the latter were adjusted for pre-determined covariates and modification by fertility and geographic group. Assumptions were examined using multiple sensitivity analyses and nonlinear models. SC declined significantly between 1973 and 2011 (slope in unadjusted simple regression models -0.70 million/ml/year; 95% CI: -0.72 to -0.69; P regression analysis reports a significant decline in sperm counts (as measured by SC and TSC) between 1973 and 2011, driven by a 50-60% decline among men unselected by fertility from North America, Europe, Australia and New Zealand. Because

Neighborhood social capital and crime victimization: comparison of spatial regression analysis and hierarchical regression analysis.

Science.gov (United States)

Takagi, Daisuke; Ikeda, Ken'ichi; Kawachi, Ichiro

2012-11-01

Crime is an important determinant of public health outcomes, including quality of life, mental well-being, and health behavior. A body of research has documented the association between community social capital and crime victimization. The association between social capital and crime victimization has been examined at multiple levels of spatial aggregation, ranging from entire countries, to states, metropolitan areas, counties, and neighborhoods. In multilevel analysis, the spatial boundaries at level 2 are most often drawn from administrative boundaries (e.g., Census tracts in the U.S.). One problem with adopting administrative definitions of neighborhoods is that it ignores spatial spillover. We conducted a study of social capital and crime victimization in one ward of Tokyo city, using a spatial Durbin model with an inverse-distance weighting matrix that assigned each respondent a unique level of "exposure" to social capital based on all other residents' perceptions. The study is based on a postal questionnaire sent to 20-69 years old residents of Arakawa Ward, Tokyo. The response rate was 43.7%. We examined the contextual influence of generalized trust, perceptions of reciprocity, two types of social network variables, as well as two principal components of social capital (constructed from the above four variables). Our outcome measure was self-reported crime victimization in the last five years. In the spatial Durbin model, we found that neighborhood generalized trust, reciprocity, supportive networks and two principal components of social capital were each inversely associated with crime victimization. By contrast, a multilevel regression performed with the same data (using administrative neighborhood boundaries) found generally null associations between neighborhood social capital and crime. Spatial regression methods may be more appropriate for investigating the contextual influence of social capital in homogeneous cultural settings such as Japan. Copyright
Evaluation of Logistic Regression and Multivariate Adaptive Regression Spline Models for Groundwater Potential Mapping Using R and GIS

Directory of Open Access Journals (Sweden)

Soyoung Park

2017-07-01

Full Text Available This study mapped and analyzed groundwater potential using two different models, logistic regression (LR and multivariate adaptive regression splines (MARS, and compared the results. A spatial database was constructed for groundwater well data and groundwater influence factors. Groundwater well data with a high potential yield of ≥70 m3/d were extracted, and 859 locations (70% were used for model training, whereas the other 365 locations (30% were used for model validation. We analyzed 16 groundwater influence factors including altitude, slope degree, slope aspect, plan curvature, profile curvature, topographic wetness index, stream power index, sediment transport index, distance from drainage, drainage density, lithology, distance from fault, fault density, distance from lineament, lineament density, and land cover. Groundwater potential maps (GPMs were constructed using LR and MARS models and tested using a receiver operating characteristics curve. Based on this analysis, the area under the curve (AUC for the success rate curve of GPMs created using the MARS and LR models was 0.867 and 0.838, and the AUC for the prediction rate curve was 0.836 and 0.801, respectively. This implies that the MARS model is useful and effective for groundwater potential analysis in the study area.
Introduction to regression graphics

CERN Document Server

Cook, R Dennis

2009-01-01

Covers the use of dynamic and interactive computer graphics in linear regression analysis, focusing on analytical graphics. Features new techniques like plot rotation. The authors have composed their own regression code, using Xlisp-Stat language called R-code, which is a nearly complete system for linear regression analysis and can be utilized as the main computer program in a linear regression course. The accompanying disks, for both Macintosh and Windows computers, contain the R-code and Xlisp-Stat. An Instructor's Manual presenting detailed solutions to all the problems in the book is ava
Diallel analysis for seed yield and its component traits in Cuphea procumbens

Directory of Open Access Journals (Sweden)

Singh S.P.

2006-01-01

Full Text Available The Cuphea procumbens Orteg. is an important annual plant source of medium chain fatty acids. The present study was conducted to estimate different gene systems involved in the inheritance of important quantitative traits viz. plant height, branches/plant, fruits/plant, seeds/fruit and seed yield/plant in F1 and F2 generations following 6 parents half diallel. Diallel assumptions were fulfilled for all the characters. Wr - Vr graph and component analysis revealed the major influence of over dominance for all the traits except branches/plant in F1. The arrays scattered all along the regression line below limiting parabola in two groups, Dominance and recessive and was confirmed by standardized deviation graph. The ranking on the basis of breeding value (Yr of the parents and per se performance was closely associated (r=0.83**. On the basis of ranking, parents 'NBC-01', 'NBC-25' and 'NBC-30' were found most promising and possessed more dominant alleles for most of the characters. Considering the gene action involved, the breeding plan was discussed.
Analysis of γ spectra in airborne radioactivity measurements using multiple linear regressions

International Nuclear Information System (INIS)

Bao Min; Shi Quanlin; Zhang Jiamei

2004-01-01

This paper describes the net peak counts calculating of nuclide 137 Cs at 662 keV of γ spectra in airborne radioactivity measurements using multiple linear regressions. Mathematic model is founded by analyzing every factor that has contribution to Cs peak counts in spectra, and multiple linear regression function is established. Calculating process adopts stepwise regression, and the indistinctive factors are eliminated by F check. The regression results and its uncertainty are calculated using Least Square Estimation, then the Cs peak net counts and its uncertainty can be gotten. The analysis results for experimental spectrum are displayed. The influence of energy shift and energy resolution on the analyzing result is discussed. In comparison with the stripping spectra method, multiple linear regression method needn't stripping radios, and the calculating result has relation with the counts in Cs peak only, and the calculating uncertainty is reduced. (authors)
Regression Analysis and Calibration Recommendations for the Characterization of Balance Temperature Effects

Science.gov (United States)

Ulbrich, N.; Volden, T.

2018-01-01

Analysis and use of temperature-dependent wind tunnel strain-gage balance calibration data are discussed in the paper. First, three different methods are presented and compared that may be used to process temperature-dependent strain-gage balance data. The first method uses an extended set of independent variables in order to process the data and predict balance loads. The second method applies an extended load iteration equation during the analysis of balance calibration data. The third method uses temperature-dependent sensitivities for the data analysis. Physical interpretations of the most important temperature-dependent regression model terms are provided that relate temperature compensation imperfections and the temperature-dependent nature of the gage factor to sets of regression model terms. Finally, balance calibration recommendations are listed so that temperature-dependent calibration data can be obtained and successfully processed using the reviewed analysis methods.
Clinical evaluation of a novel population-based regression analysis for detecting glaucomatous visual field progression.

Science.gov (United States)

Kovalska, M P; Bürki, E; Schoetzau, A; Orguel, S F; Orguel, S; Grieshaber, M C

2011-04-01

The distinction of real progression from test variability in visual field (VF) series may be based on clinical judgment, on trend analysis based on follow-up of test parameters over time, or on identification of a significant change related to the mean of baseline exams (event analysis). The aim of this study was to compare a new population-based method (Octopus field analysis, OFA) with classic regression analyses and clinical judgment for detecting glaucomatous VF changes. 240 VF series of 240 patients with at least 9 consecutive examinations available were included into this study. They were independently classified by two experienced investigators. The results of such a classification served as a reference for comparison for the following statistical tests: (a) t-test global, (b) r-test global, (c) regression analysis of 10 VF clusters and (d) point-wise linear regression analysis. 32.5 % of the VF series were classified as progressive by the investigators. The sensitivity and specificity were 89.7 % and 92.0 % for r-test, and 73.1 % and 93.8 % for the t-test, respectively. In the point-wise linear regression analysis, the specificity was comparable (89.5 % versus 92 %), but the sensitivity was clearly lower than in the r-test (22.4 % versus 89.7 %) at a significance level of p = 0.01. A regression analysis for the 10 VF clusters showed a markedly higher sensitivity for the r-test (37.7 %) than the t-test (14.1 %) at a similar specificity (88.3 % versus 93.8 %) for a significant trend (p = 0.005). In regard to the cluster distribution, the paracentral clusters and the superior nasal hemifield progressed most frequently. The population-based regression analysis seems to be superior to the trend analysis in detecting VF progression in glaucoma, and may eliminate the drawbacks of the event analysis. Further, it may assist the clinician in the evaluation of VF series and may allow better visualization of the correlation between function and structure owing to VF
Selective principal component regression analysis of fluorescence hyperspectral image to assess aflatoxin contamination in corn

Science.gov (United States)

Selective principal component regression analysis (SPCR) uses a subset of the original image bands for principal component transformation and regression. For optimal band selection before the transformation, this paper used genetic algorithms (GA). In this case, the GA process used the regression co...
Covariate Imbalance and Adjustment for Logistic Regression Analysis of Clinical Trial Data

Science.gov (United States)

Ciolino, Jody D.; Martin, Reneé H.; Zhao, Wenle; Jauch, Edward C.; Hill, Michael D.; Palesch, Yuko Y.

2014-01-01

In logistic regression analysis for binary clinical trial data, adjusted treatment effect estimates are often not equivalent to unadjusted estimates in the presence of influential covariates. This paper uses simulation to quantify the benefit of covariate adjustment in logistic regression. However, International Conference on Harmonization guidelines suggest that covariate adjustment be pre-specified. Unplanned adjusted analyses should be considered secondary. Results suggest that that if adjustment is not possible or unplanned in a logistic setting, balance in continuous covariates can alleviate some (but never all) of the shortcomings of unadjusted analyses. The case of log binomial regression is also explored. PMID:24138438
Declining Bias and Gender Wage Discrimination? A Meta-Regression Analysis

Science.gov (United States)

Jarrell, Stephen B.; Stanley, T. D.

2004-01-01

The meta-regression analysis reveals that there is a strong tendency for discrimination estimates to fall and wage discrimination exist against the woman. The biasing effect of researchers' gender of not correcting for selection bias has weakened and changes in labor market have made it less important.
Yield stability analysis of pearl millet hybrids in Nigeria

African Journals Online (AJOL)

hope&shola

2006-02-02

.] was ... Genotype x environment interaction was observed, a large component of which was accounted ... The importance of evaluating many potential genotypes .... Pooled analysis of variance for stability of grain yield (t/ha).
Length bias correction in gene ontology enrichment analysis using logistic regression.

Science.gov (United States)

Mi, Gu; Di, Yanming; Emerson, Sarah; Cumbie, Jason S; Chang, Jeff H

2012-01-01

When assessing differential gene expression from RNA sequencing data, commonly used statistical tests tend to have greater power to detect differential expression of genes encoding longer transcripts. This phenomenon, called "length bias", will influence subsequent analyses such as Gene Ontology enrichment analysis. In the presence of length bias, Gene Ontology categories that include longer genes are more likely to be identified as enriched. These categories, however, are not necessarily biologically more relevant. We show that one can effectively adjust for length bias in Gene Ontology analysis by including transcript length as a covariate in a logistic regression model. The logistic regression model makes the statistical issue underlying length bias more transparent: transcript length becomes a confounding factor when it correlates with both the Gene Ontology membership and the significance of the differential expression test. The inclusion of the transcript length as a covariate allows one to investigate the direct correlation between the Gene Ontology membership and the significance of testing differential expression, conditional on the transcript length. We present both real and simulated data examples to show that the logistic regression approach is simple, effective, and flexible.
Correlation and path-cofficient analysis of seed yield and yield ...

African Journals Online (AJOL)

This study was undertaken in order to determine the association among yield components and their direct and indirect effects on the seed yield of confectionery sunflower. 36 confectionery sunflower populations originated from different regions of Northwest Iran were characterized using 11 agromorphological traits ...
Regression models of ultimate methane yields of fruits and vegetable solid wastes, sorghum and napiergrass on chemical composition

Energy Technology Data Exchange (ETDEWEB)

Gunaseelan, V.N. [PSG College of Arts and Science, Coimbatore (India). Department of Zoology

2007-04-15

Several fractions of fruits and vegetable solid wastes (FVSW), sorghum and napiergrass were analyzed for total solids (TS), volatile solids (VS), total organic carbon, total kjeldahl nitrogen, total soluble carbohydrate, extractable protein, acid-detergent fiber (ADF), lignin, cellulose and ash contents. Their ultimate methane yields (B{sub o}) were determined using the biochemical methane potential (BMP) assay. A series of simple and multiple regression models relating the B{sub o} to the various substrate constituents were generated and evaluated using computer statistical software, Statistical Package for Social Sciences (SPSS). The results of simple regression analyses revealed that, only weak relationship existed between the individual components such as carbohydrate, protein, ADF, lignin and cellulose versus B{sub o}. A regression of B{sub o} versus combination of two variables as a single independent variable such as carbohydrate/ADF and carbohydrate + protein/ADF also showed that the relationship is not strong. Thus it does not appear possible to relate the B{sub o} of FVSW, sorghum and napiergrass with single compositional characteristics. The results of multiple regression analyses showed promise and the relationship appeared to be good. When ADF and lignin/ADF were used as independent variables, the percentage of variation accounted for by the model is low for FVSW (r{sup 2}=0.665) and sorghum and napiergrass (r{sup 2}=0.746). Addition of nitrogen, ash and total soluble carbohydrate data to the model had a significantly higher effect on prediction of B{sub o} of these wastes with the r{sup 2} values ranging from 0.9 to 0.99. More than 90% of variation in B{sub o} of FVSW could be accounted for by the models when the variables carbohydrate, lignin, lignin/ADF, nitrogen and ash (r{sup 2}=0.904), carbohydrate, ADF, lignin/ADF, nitrogen and ash (r{sup 2}=0.90) and carbohydrate/ADF, lignin/ADF, lignin and ash (r{sup 2}=0.901) were used. All the models have
Analysis of designed experiments by stabilised PLS Regression and jack-knifing

DEFF Research Database (Denmark)

Martens, Harald; Høy, M.; Westad, F.

2001-01-01

Pragmatical, visually oriented methods for assessing and optimising bi-linear regression models are described, and applied to PLS Regression (PLSR) analysis of multi-response data from controlled experiments. The paper outlines some ways to stabilise the PLSR method to extend its range...... the reliability of the linear and bi-linear model parameter estimates. The paper illustrates how the obtained PLSR "significance" probabilities are similar to those from conventional factorial ANOVA, but the PLSR is shown to give important additional overview plots of the main relevant structures in the multi....... An Introduction, Wiley, Chichester, UK, 2001]....
Blackleg (Leptosphaeria maculans) Severity and Yield Loss in Canola in Alberta, Canada

Science.gov (United States)

Hwang, Sheau-Fang; Strelkov, Stephen E.; Peng, Gary; Ahmed, Hafiz; Zhou, Qixing; Turnbull, George

2016-01-01

Blackleg, caused by Leptosphaeria maculans, is an important disease of oilseed rape (Brassica napus L.) in Canada and throughout the world. Severe epidemics of blackleg can result in significant yield losses. Understanding disease-yield relationships is a prerequisite for measuring the agronomic efficacy and economic benefits of control methods. Field experiments were conducted in 2013, 2014, and 2015 to determine the relationship between blackleg disease severity and yield in a susceptible cultivar and in moderately resistant to resistant canola hybrids. Disease severity was lower, and seed yield was 120%–128% greater, in the moderately resistant to resistant hybrids compared with the susceptible cultivar. Regression analysis showed that pod number and seed yield declined linearly as blackleg severity increased. Seed yield per plant decreased by 1.8 g for each unit increase in disease severity, corresponding to a decline in yield of 17.2% for each unit increase in disease severity. Pyraclostrobin fungicide reduced disease severity in all site-years and increased yield. These results show that the reduction of blackleg in canola crops substantially improves yields. PMID:27447676
Blackleg (Leptosphaeria maculans Severity and Yield Loss in Canola in Alberta, Canada

Directory of Open Access Journals (Sweden)

Sheau-Fang Hwang

2016-07-01

Full Text Available Blackleg, caused by Leptosphaeria maculans, is an important disease of oilseed rape (Brassica napus L. in Canada and throughout the world. Severe epidemics of blackleg can result in significant yield losses. Understanding disease-yield relationships is a prerequisite for measuring the agronomic efficacy and economic benefits of control methods. Field experiments were conducted in 2013, 2014, and 2015 to determine the relationship between blackleg disease severity and yield in a susceptible cultivar and in moderately resistant to resistant canola hybrids. Disease severity was lower, and seed yield was 120%–128% greater, in the moderately resistant to resistant hybrids compared with the susceptible cultivar. Regression analysis showed that pod number and seed yield declined linearly as blackleg severity increased. Seed yield per plant decreased by 1.8 g for each unit increase in disease severity, corresponding to a decline in yield of 17.2% for each unit increase in disease severity. Pyraclostrobin fungicide reduced disease severity in all site-years and increased yield. These results show that the reduction of blackleg in canola crops substantially improves yields.
Replica analysis of overfitting in regression models for time-to-event data

Science.gov (United States)

Coolen, A. C. C.; Barrett, J. E.; Paga, P.; Perez-Vicente, C. J.

2017-09-01

Overfitting, which happens when the number of parameters in a model is too large compared to the number of data points available for determining these parameters, is a serious and growing problem in survival analysis. While modern medicine presents us with data of unprecedented dimensionality, these data cannot yet be used effectively for clinical outcome prediction. Standard error measures in maximum likelihood regression, such as p-values and z-scores, are blind to overfitting, and even for Cox’s proportional hazards model (the main tool of medical statisticians), one finds in literature only rules of thumb on the number of samples required to avoid overfitting. In this paper we present a mathematical theory of overfitting in regression models for time-to-event data, which aims to increase our quantitative understanding of the problem and provide practical tools with which to correct regression outcomes for the impact of overfitting. It is based on the replica method, a statistical mechanical technique for the analysis of heterogeneous many-variable systems that has been used successfully for several decades in physics, biology, and computer science, but not yet in medical statistics. We develop the theory initially for arbitrary regression models for time-to-event data, and verify its predictions in detail for the popular Cox model.
Predictors of postoperative outcomes of cubital tunnel syndrome treatments using multiple logistic regression analysis.

Science.gov (United States)

Suzuki, Taku; Iwamoto, Takuji; Shizu, Kanae; Suzuki, Katsuji; Yamada, Harumoto; Sato, Kazuki

2017-05-01

This retrospective study was designed to investigate prognostic factors for postoperative outcomes for cubital tunnel syndrome (CubTS) using multiple logistic regression analysis with a large number of patients. Eighty-three patients with CubTS who underwent surgeries were enrolled. The following potential prognostic factors for disease severity were selected according to previous reports: sex, age, type of surgery, disease duration, body mass index, cervical lesion, presence of diabetes mellitus, Workers' Compensation status, preoperative severity, and preoperative electrodiagnostic testing. Postoperative severity of disease was assessed 2 years after surgery by Messina's criteria which is an outcome measure specifically for CubTS. Bivariate analysis was performed to select candidate prognostic factors for multiple linear regression analyses. Multiple logistic regression analysis was conducted to identify the association between postoperative severity and selected prognostic factors. Both bivariate and multiple linear regression analysis revealed only preoperative severity as an independent risk factor for poor prognosis, while other factors did not show any significant association. Although conflicting results exist regarding prognosis of CubTS, this study supports evidence from previous studies and concludes early surgical intervention portends the most favorable prognosis. Copyright © 2017 The Japanese Orthopaedic Association. Published by Elsevier B.V. All rights reserved.
Multiple linear regression analysis

Science.gov (United States)

Edwards, T. R.

1980-01-01

Program rapidly selects best-suited set of coefficients. User supplies only vectors of independent and dependent data and specifies confidence level required. Program uses stepwise statistical procedure for relating minimal set of variables to set of observations; final regression contains only most statistically significant coefficients. Program is written in FORTRAN IV for batch execution and has been implemented on NOVA 1200.

Post-processing through linear regression

Science.gov (United States)

van Schaeybroeck, B.; Vannitsem, S.

2011-03-01

Various post-processing techniques are compared for both deterministic and ensemble forecasts, all based on linear regression between forecast data and observations. In order to evaluate the quality of the regression methods, three criteria are proposed, related to the effective correction of forecast error, the optimal variability of the corrected forecast and multicollinearity. The regression schemes under consideration include the ordinary least-square (OLS) method, a new time-dependent Tikhonov regularization (TDTR) method, the total least-square method, a new geometric-mean regression (GM), a recently introduced error-in-variables (EVMOS) method and, finally, a "best member" OLS method. The advantages and drawbacks of each method are clarified. These techniques are applied in the context of the 63 Lorenz system, whose model version is affected by both initial condition and model errors. For short forecast lead times, the number and choice of predictors plays an important role. Contrarily to the other techniques, GM degrades when the number of predictors increases. At intermediate lead times, linear regression is unable to provide corrections to the forecast and can sometimes degrade the performance (GM and the best member OLS with noise). At long lead times the regression schemes (EVMOS, TDTR) which yield the correct variability and the largest correlation between ensemble error and spread, should be preferred.
Diagnostic accuracy of atypical p-ANCA in autoimmune hepatitis using ROC- and multivariate regression analysis.

Science.gov (United States)

Terjung, B; Bogsch, F; Klein, R; Söhne, J; Reichel, C; Wasmuth, J-C; Beuers, U; Sauerbruch, T; Spengler, U

2004-09-29

Antineutrophil cytoplasmic antibodies (atypical p-ANCA) are detected at high prevalence in sera from patients with autoimmune hepatitis (AIH), but their diagnostic relevance for AIH has not been systematically evaluated so far. Here, we studied sera from 357 patients with autoimmune (autoimmune hepatitis n=175, primary sclerosing cholangitis (PSC) n=35, primary biliary cirrhosis n=45), non-autoimmune chronic liver disease (alcoholic liver cirrhosis n=62; chronic hepatitis C virus infection (HCV) n=21) or healthy controls (n=19) for the presence of various non-organ specific autoantibodies. Atypical p-ANCA, antinuclear antibodies (ANA), antibodies against smooth muscles (SMA), antibodies against liver/kidney microsomes (anti-Lkm1) and antimitochondrial antibodies (AMA) were detected by indirect immunofluorescence microscopy, antibodies against the M2 antigen (anti-M2), antibodies against soluble liver antigen (anti-SLA/LP) and anti-Lkm1 by using enzyme linked immunosorbent assays. To define the diagnostic precision of the autoantibodies, results of autoantibody testing were analyzed by receiver operating characteristics (ROC) and forward conditional logistic regression analysis. Atypical p-ANCA were detected at high prevalence in sera from patients with AIH (81%) and PSC (94%). ROC- and logistic regression analysis revealed atypical p-ANCA and SMA, but not ANA as significant diagnostic seromarkers for AIH (atypical p-ANCA: AUC 0.754+/-0.026, odds ratio [OR] 3.4; SMA: 0.652+/-0.028, OR 4.1). Atypical p-ANCA also emerged as the only diagnostically relevant seromarker for PSC (AUC 0.690+/-0.04, OR 3.4). None of the tested antibodies yielded a significant diagnostic accuracy for patients with alcoholic liver cirrhosis, HCV or healthy controls. Atypical p-ANCA along with SMA represent a seromarker with high diagnostic accuracy for AIH and should be explicitly considered in a revised version of the diagnostic score for AIH.
Exploring factors associated with traumatic dental injuries in preschool children: a Poisson regression analysis.

Science.gov (United States)

Feldens, Carlos Alberto; Kramer, Paulo Floriani; Ferreira, Simone Helena; Spiguel, Mônica Hermann; Marquezan, Marcela

2010-04-01

This cross-sectional study aimed to investigate the factors associated with dental trauma in preschool children using Poisson regression analysis with robust variance. The study population comprised 888 children aged 3- to 5-year-old attending public nurseries in Canoas, southern Brazil. Questionnaires assessing information related to the independent variables (age, gender, race, mother's educational level and family income) were completed by the parents. Clinical examinations were carried out by five trained examiners in order to assess traumatic dental injuries (TDI) according to Andreasen's classification. One of the five examiners was calibrated to assess orthodontic characteristics (open bite and overjet). Multivariable Poisson regression analysis with robust variance was used to determine the factors associated with dental trauma as well as the strengths of association. Traditional logistic regression was also performed in order to compare the estimates obtained by both methods of statistical analysis. 36.4% (323/888) of the children suffered dental trauma and there was no difference in prevalence rates from 3 to 5 years of age. Poisson regression analysis showed that the probability of the outcome was almost 30% higher for children whose mothers had more than 8 years of education (Prevalence Ratio = 1.28; 95% CI = 1.03-1.60) and 63% higher for children with an overjet greater than 2 mm (Prevalence Ratio = 1.63; 95% CI = 1.31-2.03). Odds ratios clearly overestimated the size of the effect when compared with prevalence ratios. These findings indicate the need for preventive orientation regarding TDI, in order to educate parents and caregivers about supervising infants, particularly those with increased overjet and whose mothers have a higher level of education. Poisson regression with robust variance represents a better alternative than logistic regression to estimate the risk of dental trauma in preschool children.
Application of range-test in multiple linear regression analysis in ...

African Journals Online (AJOL)

Application of range-test in multiple linear regression analysis in the presence of outliers is studied in this paper. First, the plot of the explanatory variables (i.e. Administration, Social/Commercial, Economic services and Transfer) on the dependent variable (i.e. GDP) was done to identify the statistical trend over the years.
Statistical evaluation of fuel yield and morphological variates for some promising energy plantation tree species in western Rajasthan

Energy Technology Data Exchange (ETDEWEB)

Kalla, J.C.

1977-01-01

Stepwise regression analysis suggested that tree height and collar diameter were, in general, the morphological parameters that most reliably predicted fuel yield in Acacia nilotica, A. tortilis, Albizzia lebbek, Azadirachta indica and Prosopis juliflora.
Multiple regression analysis of anthropometric measurements influencing the cephalic index of male Japanese university students.

Science.gov (United States)

Hossain, Md Golam; Saw, Aik; Alam, Rashidul; Ohtsuki, Fumio; Kamarul, Tunku

2013-09-01

Cephalic index (CI), the ratio of head breadth to head length, is widely used to categorise human populations. The aim of this study was to access the impact of anthropometric measurements on the CI of male Japanese university students. This study included 1,215 male university students from Tokyo and Kyoto, selected using convenient sampling. Multiple regression analysis was used to determine the effect of anthropometric measurements on CI. The variance inflation factor (VIF) showed no evidence of a multicollinearity problem among independent variables. The coefficients of the regression line demonstrated a significant positive relationship between CI and minimum frontal breadth (p regression analysis showed a greater likelihood for minimum frontal breadth (p regression analysis revealed bizygomatic breadth, head circumference, minimum frontal breadth, head height and morphological facial height to be the best predictor craniofacial measurements with respect to CI. The results suggest that most of the variables considered in this study appear to influence the CI of adult male Japanese students.
PATH COEFFICIENT ANALYSIS OF SEVERAL COMPONENTS OIL YIELD IN SUNFLOWER (HELIANTHUS ANNUUS L.

Directory of Open Access Journals (Sweden)

A. MIjić

2006-06-01

Full Text Available The objective of investigation was to analyse oil yield components and their relations by simple coefficient correlations as well as direct and indirect effects to oil yield by path analysis. Twenty-four sunflower hybrids were included in the investigation and their seven traits (plant height, head diameter, 1000 seed weight, hec- tolitar mass, grain yield, oil content and oil yield. Very strong positive correlation was estimated between grain yield and oil yield, strong positive correlation between hectolitar mass and oil yield, and middle corre- lation among oil yield and: 1000 seed weight, plaint height and oil content. There was no correlation between grain yields and oil content. Grain yield showed the strongest effect to oil yield. Oil content had lower effect to oil yield. Other traits showed no significant effect to oil yield, and their effect to oil yield was covered by indirect effect of grain yield.
Remote sensing and GIS-based landslide hazard analysis and cross-validation using multivariate logistic regression model on three test areas in Malaysia

Science.gov (United States)

Pradhan, Biswajeet

2010-05-01

This paper presents the results of the cross-validation of a multivariate logistic regression model using remote sensing data and GIS for landslide hazard analysis on the Penang, Cameron, and Selangor areas in Malaysia. Landslide locations in the study areas were identified by interpreting aerial photographs and satellite images, supported by field surveys. SPOT 5 and Landsat TM satellite imagery were used to map landcover and vegetation index, respectively. Maps of topography, soil type, lineaments and land cover were constructed from the spatial datasets. Ten factors which influence landslide occurrence, i.e., slope, aspect, curvature, distance from drainage, lithology, distance from lineaments, soil type, landcover, rainfall precipitation, and normalized difference vegetation index (ndvi), were extracted from the spatial database and the logistic regression coefficient of each factor was computed. Then the landslide hazard was analysed using the multivariate logistic regression coefficients derived not only from the data for the respective area but also using the logistic regression coefficients calculated from each of the other two areas (nine hazard maps in all) as a cross-validation of the model. For verification of the model, the results of the analyses were then compared with the field-verified landslide locations. Among the three cases of the application of logistic regression coefficient in the same study area, the case of Selangor based on the Selangor logistic regression coefficients showed the highest accuracy (94%), where as Penang based on the Penang coefficients showed the lowest accuracy (86%). Similarly, among the six cases from the cross application of logistic regression coefficient in other two areas, the case of Selangor based on logistic coefficient of Cameron showed highest (90%) prediction accuracy where as the case of Penang based on the Selangor logistic regression coefficients showed the lowest accuracy (79%). Qualitatively, the cross
Understanding poisson regression.

Science.gov (United States)

Hayat, Matthew J; Higgins, Melinda

2014-04-01

Nurse investigators often collect study data in the form of counts. Traditional methods of data analysis have historically approached analysis of count data either as if the count data were continuous and normally distributed or with dichotomization of the counts into the categories of occurred or did not occur. These outdated methods for analyzing count data have been replaced with more appropriate statistical methods that make use of the Poisson probability distribution, which is useful for analyzing count data. The purpose of this article is to provide an overview of the Poisson distribution and its use in Poisson regression. Assumption violations for the standard Poisson regression model are addressed with alternative approaches, including addition of an overdispersion parameter or negative binomial regression. An illustrative example is presented with an application from the ENSPIRE study, and regression modeling of comorbidity data is included for illustrative purposes. Copyright 2014, SLACK Incorporated.
Yield stress fluids slowly yield to analysis

NARCIS (Netherlands)

Bonn, D.; Denn, M.M.

2009-01-01

We are surrounded in everyday life by yield stress fluids: materials that behave as solids under small stresses but flow like liquids beyond a critical stress. For example, paint must flow under the brush, but remain fixed in a vertical film despite the force of gravity. Food products (such as
Line tester analysis of yield and yield related attributed in different sunflower genotypes

International Nuclear Information System (INIS)

Din, S.U.; Khan, M.A.; Usman, K.; Sayal, O.U.

2014-01-01

This paper encompasses the study of line * tester analysis to chalk out genetic implications regarding yield and yield relating components in different genotypes of sunflower. Eight parents (four CMS lines and four restorers) along with their sixteen F1 hybrids were considered and planted in Randomized Complete Block Design (RCBD) replicated thrice at experimental area of Oilseed Research Program, National Agriculture Research Centre (NARC), Islamabad, Pakistan in 2011. Combining ability for some important morphological traits included days to flower initiation, days to flower completion, days to maturity, plant height, head diameter and seed yield plant-1. In this concern general combining ability (GCA), reciprocals combining ability (RCA) and specific combining ability (SCA) for all traits were studied. The GCA and SCA variances due to lines and testers interaction were significant for all the characters. However, the magnitude of GCAs from CMS lines (females) and restorers (pollinators) were higher than the SCA indicating preponderance of additive genes in the expression of all the traits. Among the lines, CMS-HA-54 whereas in testers, RHP-71, by manifesting maximum GCA effects were considered as the best general combiners for almost all the traits indicating the presence of more additive gene effects in these parents, therefore may serve as potential parents for hybridization and to improve the characters studied. Among the F1 hybrids, CMS HA-99 * RHP-76 (1.54, 212.65) and CMS HA-101 * RHP-73 (0.91, 432.73) were found as the best specific combiners for head or capitulum and seed yield. Hence, if farming community and researchers include these hybrids in their selection and hybridization program for the trait under study optimum result may be obtained. (author)
Analysis of the influence of quantile regression model on mainland tourists' service satisfaction performance.

Science.gov (United States)

Wang, Wen-Cheng; Cho, Wen-Chien; Chen, Yin-Jen

2014-01-01

It is estimated that mainland Chinese tourists travelling to Taiwan can bring annual revenues of 400 billion NTD to the Taiwan economy. Thus, how the Taiwanese Government formulates relevant measures to satisfy both sides is the focus of most concern. Taiwan must improve the facilities and service quality of its tourism industry so as to attract more mainland tourists. This paper conducted a questionnaire survey of mainland tourists and used grey relational analysis in grey mathematics to analyze the satisfaction performance of all satisfaction question items. The first eight satisfaction items were used as independent variables, and the overall satisfaction performance was used as a dependent variable for quantile regression model analysis to discuss the relationship between the dependent variable under different quantiles and independent variables. Finally, this study further discussed the predictive accuracy of the least mean regression model and each quantile regression model, as a reference for research personnel. The analysis results showed that other variables could also affect the overall satisfaction performance of mainland tourists, in addition to occupation and age. The overall predictive accuracy of quantile regression model Q0.25 was higher than that of the other three models.
Analysis of the Influence of Quantile Regression Model on Mainland Tourists' Service Satisfaction Performance

Science.gov (United States)

Wang, Wen-Cheng; Cho, Wen-Chien; Chen, Yin-Jen

2014-01-01

It is estimated that mainland Chinese tourists travelling to Taiwan can bring annual revenues of 400 billion NTD to the Taiwan economy. Thus, how the Taiwanese Government formulates relevant measures to satisfy both sides is the focus of most concern. Taiwan must improve the facilities and service quality of its tourism industry so as to attract more mainland tourists. This paper conducted a questionnaire survey of mainland tourists and used grey relational analysis in grey mathematics to analyze the satisfaction performance of all satisfaction question items. The first eight satisfaction items were used as independent variables, and the overall satisfaction performance was used as a dependent variable for quantile regression model analysis to discuss the relationship between the dependent variable under different quantiles and independent variables. Finally, this study further discussed the predictive accuracy of the least mean regression model and each quantile regression model, as a reference for research personnel. The analysis results showed that other variables could also affect the overall satisfaction performance of mainland tourists, in addition to occupation and age. The overall predictive accuracy of quantile regression model Q0.25 was higher than that of the other three models. PMID:24574916
Analysis of the Influence of Quantile Regression Model on Mainland Tourists’ Service Satisfaction Performance

Directory of Open Access Journals (Sweden)

Wen-Cheng Wang

2014-01-01

Full Text Available It is estimated that mainland Chinese tourists travelling to Taiwan can bring annual revenues of 400 billion NTD to the Taiwan economy. Thus, how the Taiwanese Government formulates relevant measures to satisfy both sides is the focus of most concern. Taiwan must improve the facilities and service quality of its tourism industry so as to attract more mainland tourists. This paper conducted a questionnaire survey of mainland tourists and used grey relational analysis in grey mathematics to analyze the satisfaction performance of all satisfaction question items. The first eight satisfaction items were used as independent variables, and the overall satisfaction performance was used as a dependent variable for quantile regression model analysis to discuss the relationship between the dependent variable under different quantiles and independent variables. Finally, this study further discussed the predictive accuracy of the least mean regression model and each quantile regression model, as a reference for research personnel. The analysis results showed that other variables could also affect the overall satisfaction performance of mainland tourists, in addition to occupation and age. The overall predictive accuracy of quantile regression model Q0.25 was higher than that of the other three models.
Econometric analysis of realised covariation: high frequency covariance, regression and correlation in financial economics

OpenAIRE

Ole E. Barndorff-Nielsen; Neil Shephard

2002-01-01

This paper analyses multivariate high frequency financial data using realised covariation. We provide a new asymptotic distribution theory for standard methods such as regression, correlation analysis and covariance. It will be based on a fixed interval of time (e.g. a day or week), allowing the number of high frequency returns during this period to go to infinity. Our analysis allows us to study how high frequency correlations, regressions and covariances change through time. In particular w...
Prediction of radiation levels in residences: A methodological comparison of CART [Classification and Regression Tree Analysis] and conventional regression

International Nuclear Information System (INIS)

Janssen, I.; Stebbings, J.H.

1990-01-01

In environmental epidemiology, trace and toxic substance concentrations frequently have very highly skewed distributions ranging over one or more orders of magnitude, and prediction by conventional regression is often poor. Classification and Regression Tree Analysis (CART) is an alternative in such contexts. To compare the techniques, two Pennsylvania data sets and three independent variables are used: house radon progeny (RnD) and gamma levels as predicted by construction characteristics in 1330 houses; and ∼200 house radon (Rn) measurements as predicted by topographic parameters. CART may identify structural variables of interest not identified by conventional regression, and vice versa, but in general the regression models are similar. CART has major advantages in dealing with other common characteristics of environmental data sets, such as missing values, continuous variables requiring transformations, and large sets of potential independent variables. CART is most useful in the identification and screening of independent variables, greatly reducing the need for cross-tabulations and nested breakdown analyses. There is no need to discard cases with missing values for the independent variables because surrogate variables are intrinsic to CART. The tree-structured approach is also independent of the scale on which the independent variables are measured, so that transformations are unnecessary. CART identifies important interactions as well as main effects. The major advantages of CART appear to be in exploring data. Once the important variables are identified, conventional regressions seem to lead to results similar but more interpretable by most audiences. 12 refs., 8 figs., 10 tabs
Study on the Method of Grass Yield Model in the Source Region of Three Rivers with Multivariate Data

International Nuclear Information System (INIS)

You, Haoyan; Luo, Chengfeng; Liu, Zhengjun; Wang, Jiao

2014-01-01

This paper uses remote sensing and GIS technology to analyse the Source Region of Three Rivers (SRTR) to establish a grass yield estimation model during 2010 with remote sensing data, meteorological data, grassland type data and ground measured data. Analysis of the correlation between ground measured data, vegetation index based HJ-1A/B satellite data, meteorological data and grassland type data were used to establish the grass yield model. The grass yield model was studied by several statistical methods, such as multiple linear regression and Geographically Weighted Regression (GWR). The model's precision was validated. Finally, the best model to estimate the grass yield of Maduo County in SRTR was contrasted with the TM degraded grassland interpretation image of Maduo County from 2009. The result shows that: (1) Comparing with the multiple linear regression model, the GWR model gave a much better fitting result with the quality of fit increasing significantly from less than 0.3 to more than 0.8; (2) The most sensitive factors affecting the grass yield in SRTR were precipitation from May to August and drought index from May to August. From calculation of the five vegetation indices, MSAVI fitted the best; (3) The Maduo County grass yield estimated by the optimal model was consistent with the TM degraded grassland interpretation image, the spatial distribution of grass yield in Maduo County for 2010 showed a ''high south and low north'' pattern
A rotor optimization using regression analysis

Science.gov (United States)

Giansante, N.

1984-01-01

The design and development of helicopter rotors is subject to the many design variables and their interactions that effect rotor operation. Until recently, selection of rotor design variables to achieve specified rotor operational qualities has been a costly, time consuming, repetitive task. For the past several years, Kaman Aerospace Corporation has successfully applied multiple linear regression analysis, coupled with optimization and sensitivity procedures, in the analytical design of rotor systems. It is concluded that approximating equations can be developed rapidly for a multiplicity of objective and constraint functions and optimizations can be performed in a rapid and cost effective manner; the number and/or range of design variables can be increased by expanding the data base and developing approximating functions to reflect the expanded design space; the order of the approximating equations can be expanded easily to improve correlation between analyzer results and the approximating equations; gradients of the approximating equations can be calculated easily and these gradients are smooth functions reducing the risk of numerical problems in the optimization; the use of approximating functions allows the problem to be started easily and rapidly from various initial designs to enhance the probability of finding a global optimum; and the approximating equations are independent of the analysis or optimization codes used.
Regression and local control rates after radiotherapy for jugulotympanic paragangliomas: Systematic review and meta-analysis

International Nuclear Information System (INIS)

Hulsteijn, Leonie T. van; Corssmit, Eleonora P.M.; Coremans, Ida E.M.; Smit, Johannes W.A.; Jansen, Jeroen C.; Dekkers, Olaf M.

2013-01-01

The primary treatment goal of radiotherapy for paragangliomas of the head and neck region (HNPGLs) is local control of the tumor, i.e. stabilization of tumor volume. Interestingly, regression of tumor volume has also been reported. Up to the present, no meta-analysis has been performed giving an overview of regression rates after radiotherapy in HNPGLs. The main objective was to perform a systematic review and meta-analysis to assess regression of tumor volume in HNPGL-patients after radiotherapy. A second outcome was local tumor control. Design of the study is systematic review and meta-analysis. PubMed, EMBASE, Web of Science, COCHRANE and Academic Search Premier and references of key articles were searched in March 2012 to identify potentially relevant studies. Considering the indolent course of HNPGLs, only studies with ⩾12 months follow-up were eligible. Main outcomes were the pooled proportions of regression and local control after radiotherapy as initial, combined (i.e. directly post-operatively or post-embolization) or salvage treatment (i.e. after initial treatment has failed) for HNPGLs. A meta-analysis was performed with an exact likelihood approach using a logistic regression with a random effect at the study level. Pooled proportions with 95% confidence intervals (CI) were reported. Fifteen studies were included, concerning a total of 283 jugulotympanic HNPGLs in 276 patients. Pooled regression proportions for initial, combined and salvage treatment were respectively 21%, 33% and 52% in radiosurgery studies and 4%, 0% and 64% in external beam radiotherapy studies. Pooled local control proportions for radiotherapy as initial, combined and salvage treatment ranged from 79% to 100%. Radiotherapy for jugulotympanic paragangliomas results in excellent local tumor control and therefore is a valuable treatment for these types of tumors. The effects of radiotherapy on regression of tumor volume remain ambiguous, although the data suggest that regression can
Prediction of Pectin Yield and Quality by FTIR and Carbohydrate Microarray Analysis

DEFF Research Database (Denmark)

Baum, Andreas; Dominiak, Malgorzata Maria; Vidal-Melgosa, Silvia

2017-01-01

and carbohydrate microarray analysis were performed directly on the crude lime peel extracts during the time course of the extractions. Multivariate analysis of the data was carried out to predict final pectin yields. Fourier transform infrared spectroscopy (FTIR) was found applicable for determining the optimal...... extraction time for the enzymatic and acidic extraction processes, respectively. The combined results of FTIR and carbohydrate microarray analysis suggested major differences in the crude pectin extracts obtained by enzymatic and acid extraction, respectively. Enzymatically extracted pectin, thus, showed......, and that FTIR and carbohydrate microarray analysis have potential to be developed into online process analysis tools for prediction of pectin extraction yields and pectin features from measurements on crude pectin extracts....

Modelling crop yield in Iberia under drought conditions

Science.gov (United States)

Ribeiro, Andreia; Páscoa, Patrícia; Russo, Ana; Gouveia, Célia

2017-04-01

The improved assessment of the cereal yield and crop loss under drought conditions are essential to meet the increasing economy demands. The growing frequency and severity of the extreme drought conditions in the Iberian Peninsula (IP) has been likely responsible for negative impacts on agriculture, namely on crop yield losses. Therefore, a continuous monitoring of vegetation activity and a reliable estimation of drought impacts is crucial to contribute for the agricultural drought management and development of suitable information tools. This works aims to assess the influence of drought conditions in agricultural yields over the IP, considering cereal yields from mainly rainfed agriculture for the provinces with higher productivity. The main target is to develop a strategy to model drought risk on agriculture for wheat yield at a province level. In order to achieve this goal a combined assessment was made using a drought indicator (Standardized Precipitation Evapotranspiration Index, SPEI) to evaluate drought conditions together with a widely used vegetation index (Normalized Difference Vegetation Index, NDVI) to monitor vegetation activity. A correlation analysis between detrended wheat yield and SPEI was performed in order to assess the vegetation response to each time scale of drought occurrence and also identify the moment of the vegetative cycle when the crop yields are more vulnerable to drought conditions. The time scales and months of SPEI, together with the months of NDVI, better related with wheat yield were chosen to perform a multivariate regression analysis to simulate crop yield. Model results are satisfactory and highlighted the usefulness of such analysis in the framework of developing a drought risk model for crop yields. In terms of an operational point of view, the results aim to contribute to an improved understanding of crop yield management under dry conditions, particularly adding substantial information on the advantages of combining
Robust analysis of trends in noisy tokamak confinement data using geodesic least squares regression

Energy Technology Data Exchange (ETDEWEB)

Verdoolaege, G., E-mail: geert.verdoolaege@ugent.be [Department of Applied Physics, Ghent University, B-9000 Ghent (Belgium); Laboratory for Plasma Physics, Royal Military Academy, B-1000 Brussels (Belgium); Shabbir, A. [Department of Applied Physics, Ghent University, B-9000 Ghent (Belgium); Max Planck Institute for Plasma Physics, Boltzmannstr. 2, 85748 Garching (Germany); Hornung, G. [Department of Applied Physics, Ghent University, B-9000 Ghent (Belgium)

2016-11-15

Regression analysis is a very common activity in fusion science for unveiling trends and parametric dependencies, but it can be a difficult matter. We have recently developed the method of geodesic least squares (GLS) regression that is able to handle errors in all variables, is robust against data outliers and uncertainty in the regression model, and can be used with arbitrary distribution models and regression functions. We here report on first results of application of GLS to estimation of the multi-machine scaling law for the energy confinement time in tokamaks, demonstrating improved consistency of the GLS results compared to standard least squares.
Meta-regression analysis of commensal and pathogenic Escherichia coli survival in soil and water.

Science.gov (United States)

Franz, Eelco; Schijven, Jack; de Roda Husman, Ana Maria; Blaak, Hetty

2014-06-17

The extent to which pathogenic and commensal E. coli (respectively PEC and CEC) can survive, and which factors predominantly determine the rate of decline, are crucial issues from a public health point of view. The goal of this study was to provide a quantitative summary of the variability in E. coli survival in soil and water over a broad range of individual studies and to identify the most important sources of variability. To that end, a meta-regression analysis on available literature data was conducted. The considerable variation in reported decline rates indicated that the persistence of E. coli is not easily predictable. The meta-analysis demonstrated that for soil and water, the type of experiment (laboratory or field), the matrix subtype (type of water and soil), and temperature were the main factors included in the regression analysis. A higher average decline rate in soil of PEC compared with CEC was observed. The regression models explained at best 57% of the variation in decline rate in soil and 41% of the variation in decline rate in water. This indicates that additional factors, not included in the current meta-regression analysis, are of importance but rarely reported. More complete reporting of experimental conditions may allow future inference on the global effects of these variables on the decline rate of E. coli.
Automatic yield-line analysis of slabs using discontinuity layout optimization.

Science.gov (United States)

Gilbert, Matthew; He, Linwei; Smith, Colin C; Le, Canh V

2014-08-08

The yield-line method of analysis is a long established and extremely effective means of estimating the maximum load sustainable by a slab or plate. However, although numerous attempts to automate the process of directly identifying the critical pattern of yield-lines have been made over the past few decades, to date none has proved capable of reliably analysing slabs of arbitrary geometry. Here, it is demonstrated that the discontinuity layout optimization (DLO) procedure can successfully be applied to such problems. The procedure involves discretization of the problem using nodes inter-connected by potential yield-line discontinuities, with the critical layout of these then identified using linear programming. The procedure is applied to various benchmark problems, demonstrating that highly accurate solutions can be obtained, and showing that DLO provides a truly systematic means of directly and reliably automatically identifying yield-line patterns. Finally, since the critical yield-line patterns for many problems are found to be quite complex in form, a means of automatically simplifying these is presented.
Vector regression introduced

Directory of Open Access Journals (Sweden)

Mok Tik

2014-06-01

Full Text Available This study formulates regression of vector data that will enable statistical analysis of various geodetic phenomena such as, polar motion, ocean currents, typhoon/hurricane tracking, crustal deformations, and precursory earthquake signals. The observed vector variable of an event (dependent vector variable is expressed as a function of a number of hypothesized phenomena realized also as vector variables (independent vector variables and/or scalar variables that are likely to impact the dependent vector variable. The proposed representation has the unique property of solving the coefficients of independent vector variables (explanatory variables also as vectors, hence it supersedes multivariate multiple regression models, in which the unknown coefficients are scalar quantities. For the solution, complex numbers are used to rep- resent vector information, and the method of least squares is deployed to estimate the vector model parameters after transforming the complex vector regression model into a real vector regression model through isomorphism. Various operational statistics for testing the predictive significance of the estimated vector parameter coefficients are also derived. A simple numerical example demonstrates the use of the proposed vector regression analysis in modeling typhoon paths.
Climate Change Impact on Rainfall: How will Threaten Wheat Yield?

Science.gov (United States)

Tafoughalti, K.; El Faleh, E. M.; Moujahid, Y.; Ouargaga, F.

2018-05-01

Climate change has a significant impact on the environmental condition of the agricultural region. Meknes has an agrarian economy and wheat production is of paramount importance. As most arable area are under rainfed system, Meknes is one of the sensitive regions to rainfall variability and consequently to climate change. Therefore, the use of changes in rainfall is vital for detecting the influence of climate system on agricultural productivity. This article identifies rainfall temporal variability and its impact on wheat yields. We used monthly rainfall records for three decades and wheat yields records of fifteen years. Rainfall variability is assessed utilizing the precipitation concentration index and the variation coefficient. The association between wheat yields and cumulative rainfall amounts of different scales was calculated based on a regression model. The analysis shown moderate seasonal and irregular annual rainfall distribution. Yields fluctuated from 210 to 4500 Kg/ha with 52% of coefficient of variation. The correlation results shows that wheat yields are strongly correlated with rainfall of the period January to March. This investigation concluded that climate change is altering wheat yield and it is crucial to adept the necessary adaptation to challenge the risk.
Post-processing through linear regression

Directory of Open Access Journals (Sweden)

B. Van Schaeybroeck

2011-03-01

Full Text Available Various post-processing techniques are compared for both deterministic and ensemble forecasts, all based on linear regression between forecast data and observations. In order to evaluate the quality of the regression methods, three criteria are proposed, related to the effective correction of forecast error, the optimal variability of the corrected forecast and multicollinearity. The regression schemes under consideration include the ordinary least-square (OLS method, a new time-dependent Tikhonov regularization (TDTR method, the total least-square method, a new geometric-mean regression (GM, a recently introduced error-in-variables (EVMOS method and, finally, a "best member" OLS method. The advantages and drawbacks of each method are clarified.

These techniques are applied in the context of the 63 Lorenz system, whose model version is affected by both initial condition and model errors. For short forecast lead times, the number and choice of predictors plays an important role. Contrarily to the other techniques, GM degrades when the number of predictors increases. At intermediate lead times, linear regression is unable to provide corrections to the forecast and can sometimes degrade the performance (GM and the best member OLS with noise. At long lead times the regression schemes (EVMOS, TDTR which yield the correct variability and the largest correlation between ensemble error and spread, should be preferred.
Estimating the impact of mineral aerosols on crop yields in food insecure regions using statistical crop models

Science.gov (United States)

Hoffman, A.; Forest, C. E.; Kemanian, A.

2016-12-01

A significant number of food-insecure nations exist in regions of the world where dust plays a large role in the climate system. While the impacts of common climate variables (e.g. temperature, precipitation, ozone, and carbon dioxide) on crop yields are relatively well understood, the impact of mineral aerosols on yields have not yet been thoroughly investigated. This research aims to develop the data and tools to progress our understanding of mineral aerosol impacts on crop yields. Suspended dust affects crop yields by altering the amount and type of radiation reaching the plant, modifying local temperature and precipitation. While dust events (i.e. dust storms) affect crop yields by depleting the soil of nutrients or by defoliation via particle abrasion. The impact of dust on yields is modeled statistically because we are uncertain which impacts will dominate the response on national and regional scales considered in this study. Multiple linear regression is used in a number of large-scale statistical crop modeling studies to estimate yield responses to various climate variables. In alignment with previous work, we develop linear crop models, but build upon this simple method of regression with machine-learning techniques (e.g. random forests) to identify important statistical predictors and isolate how dust affects yields on the scales of interest. To perform this analysis, we develop a crop-climate dataset for maize, soybean, groundnut, sorghum, rice, and wheat for the regions of West Africa, East Africa, South Africa, and the Sahel. Random forest regression models consistently model historic crop yields better than the linear models. In several instances, the random forest models accurately capture the temperature and precipitation threshold behavior in crops. Additionally, improving agricultural technology has caused a well-documented positive trend that dominates time series of global and regional yields. This trend is often removed before regression with
Repeated Results Analysis for Middleware Regression Benchmarking

Czech Academy of Sciences Publication Activity Database

Bulej, Lubomír; Kalibera, T.; Tůma, P.

2005-01-01

Roč. 60, - (2005), s. 345-358 ISSN 0166-5316 R&D Projects: GA ČR GA102/03/0672 Institutional research plan: CEZ:AV0Z10300504 Keywords : middleware benchmarking * regression benchmarking * regression testing Subject RIV: JD - Computer Applications, Robotics Impact factor: 0.756, year: 2005
Systematic review of treatment modalities for gingival depigmentation: a random-effects poisson regression analysis.

Science.gov (United States)

Lin, Yi Hung; Tu, Yu Kang; Lu, Chun Tai; Chung, Wen Chen; Huang, Chiung Fang; Huang, Mao Suan; Lu, Hsein Kun

2014-01-01

Repigmentation variably occurs with different treatment methods in patients with gingival pigmentation. A systemic review was conducted of various treatment modalities for eliminating melanin pigmentation of the gingiva, comprising bur abrasion, scalpel surgery, cryosurgery, electrosurgery, gingival grafts, and laser techniques, to compare the recurrence rates (Rrs) of these treatment procedures. Electronic databases, including PubMed, Web of Science, Google, and Medline were comprehensively searched, and manual searches were conducted for studies published from January 1951 to June 2013. After applying inclusion and exclusion criteria, the final list of articles was reviewed in depth to achieve the objectives of this review. A Poisson regression was used to analyze the outcome of depigmentation using the various treatment methods. The systematic review was based on case reports mainly. In total, 61 eligible publications met the defined criteria. The various therapeutic procedures showed variable clinical results with a wide range of Rrs. A random-effects Poisson regression showed that cryosurgery (Rr = 0.32%), electrosurgery (Rr = 0.74%), and laser depigmentation (Rr = 1.16%) yielded superior result, whereas bur abrasion yielded the highest Rr (8.89%). Within the limit of the sampling level, the present evidence-based results show that cryosurgery exhibits the optimal predictability for depigmentation of the gingiva among all procedures examined, followed by electrosurgery and laser techniques. It is possible to treat melanin pigmentation of the gingiva with various methods and prevent repigmentation. Among those treatment modalities, cryosurgery, electrosurgery, and laser surgery appear to be the best choices for treating gingival pigmentation. © 2014 Wiley Periodicals, Inc.
Bayesian Analysis for Penalized Spline Regression Using WinBUGS

Directory of Open Access Journals (Sweden)

Ciprian M. Crainiceanu

2005-09-01

Full Text Available Penalized splines can be viewed as BLUPs in a mixed model framework, which allows the use of mixed model software for smoothing. Thus, software originally developed for Bayesian analysis of mixed models can be used for penalized spline regression. Bayesian inference for nonparametric models enjoys the flexibility of nonparametric models and the exact inference provided by the Bayesian inferential machinery. This paper provides a simple, yet comprehensive, set of programs for the implementation of nonparametric Bayesian analysis in WinBUGS. Good mixing properties of the MCMC chains are obtained by using low-rank thin-plate splines, while simulation times per iteration are reduced employing WinBUGS specific computational tricks.
Forecasting municipal solid waste generation using prognostic tools and regression analysis.

Science.gov (United States)

Ghinea, Cristina; Drăgoi, Elena Niculina; Comăniţă, Elena-Diana; Gavrilescu, Marius; Câmpean, Teofil; Curteanu, Silvia; Gavrilescu, Maria

2016-11-01

For an adequate planning of waste management systems the accurate forecast of waste generation is an essential step, since various factors can affect waste trends. The application of predictive and prognosis models are useful tools, as reliable support for decision making processes. In this paper some indicators such as: number of residents, population age, urban life expectancy, total municipal solid waste were used as input variables in prognostic models in order to predict the amount of solid waste fractions. We applied Waste Prognostic Tool, regression analysis and time series analysis to forecast municipal solid waste generation and composition by considering the Iasi Romania case study. Regression equations were determined for six solid waste fractions (paper, plastic, metal, glass, biodegradable and other waste). Accuracy Measures were calculated and the results showed that S-curve trend model is the most suitable for municipal solid waste (MSW) prediction. Copyright © 2016 Elsevier Ltd. All rights reserved.
Controls and forecasts of nitrate yields in forested watersheds: A view over mainland Portugal.

Science.gov (United States)

Pacheco, F A L; Santos, R M B; Sanches Fernandes, L F; Pereira, M G; Cortes, R M V

2015-12-15

A study on nitrate yields was conducted in forested watersheds of mainland Portugal. The prime goal was to rank parameters in descending order of their contribution to the export of nitrate towards streams and lakes. To attain the goal, variables like soil loss, rainfall intensity, topography, soil type, forest composition and environmental disturbances such as hardwood harvesting or wildfires were organized in a conceptual yield model. Because some parameters were potentially collinear, a robust multivariate statistical technique was selected to execute the conceptual model and perform the aforementioned ranking, namely Partial Least Squares (PLS) regression. This technique was tested with a sample of 60 forested watersheds (>70% of forest occupation), being subject to a double-validation process to ensure prediction capability. According to final regression coefficients, soil erosion seems to regulate nitrate distribution across the basins, because soil loss and type, rainfall intensity and topography explained around 60% of nitrate yield variance. The major importance of erosion is followed by a moderate role of biochemical processes such as nitrification or nutrient uptake, which accounted for approximately 15% of nitrate yield variance. In this case, deciduous forests and scrubland seem to behave as net sinks of nitrate while coniferous and mixed forests seem to act dually, as net sources or sinks. The least important parameters are the environmental disturbances, explaining no more than 5% of nitrate yield variance. The results of PLS regression were coupled in a scenario analysis with measures designed to protect soil from erosion and surface water from eutrophication. These interventions are to be implemented until 2045, according to regional plans of forest management. Considering the key role of erosion in explaining nitrate dynamics across the catchments, it was not surprising to verify that soil protection measures may reduce nitrate yields by some 35
Development of Compressive Failure Strength for Composite Laminate Using Regression Analysis Method

Energy Technology Data Exchange (ETDEWEB)

Lee, Myoung Keon [Agency for Defense Development, Daejeon (Korea, Republic of); Lee, Jeong Won; Yoon, Dong Hyun; Kim, Jae Hoon [Chungnam Nat’l Univ., Daejeon (Korea, Republic of)

2016-10-15

This paper provides the compressive failure strength value of composite laminate developed by using regression analysis method. Composite material in this document is a Carbon/Epoxy unidirection(UD) tape prepreg(Cycom G40-800/5276-1) cured at 350°F(177°C). The operating temperature is –60°F~+200°F(-55°C - +95°C). A total of 56 compression tests were conducted on specimens from eight (8) distinct laminates that were laid up by standard angle layers (0°, +45°, –45° and 90°). The ASTM-D-6484 standard was used for test method. The regression analysis was performed with the response variable being the laminate ultimate fracture strength and the regressor variables being two ply orientations (0° and ±45°)
Development of Compressive Failure Strength for Composite Laminate Using Regression Analysis Method

International Nuclear Information System (INIS)

Lee, Myoung Keon; Lee, Jeong Won; Yoon, Dong Hyun; Kim, Jae Hoon

2016-01-01

This paper provides the compressive failure strength value of composite laminate developed by using regression analysis method. Composite material in this document is a Carbon/Epoxy unidirection(UD) tape prepreg(Cycom G40-800/5276-1) cured at 350°F(177°C). The operating temperature is –60°F~+200°F(-55°C - +95°C). A total of 56 compression tests were conducted on specimens from eight (8) distinct laminates that were laid up by standard angle layers (0°, +45°, –45° and 90°). The ASTM-D-6484 standard was used for test method. The regression analysis was performed with the response variable being the laminate ultimate fracture strength and the regressor variables being two ply orientations (0° and ±45°)
Standards for Standardized Logistic Regression Coefficients

Science.gov (United States)

Menard, Scott

2011-01-01

Standardized coefficients in logistic regression analysis have the same utility as standardized coefficients in linear regression analysis. Although there has been no consensus on the best way to construct standardized logistic regression coefficients, there is now sufficient evidence to suggest a single best approach to the construction of a…
CADDIS Volume 4. Data Analysis: PECBO Appendix - R Scripts for Non-Parametric Regressions

Science.gov (United States)

Script for computing nonparametric regression analysis. Overview of using scripts to infer environmental conditions from biological observations, statistically estimating species-environment relationships, statistical scripts.
Aneurysmal subarachnoid hemorrhage prognostic decision-making algorithm using classification and regression tree analysis.

Science.gov (United States)

Lo, Benjamin W Y; Fukuda, Hitoshi; Angle, Mark; Teitelbaum, Jeanne; Macdonald, R Loch; Farrokhyar, Forough; Thabane, Lehana; Levine, Mitchell A H

2016-01-01

Classification and regression tree analysis involves the creation of a decision tree by recursive partitioning of a dataset into more homogeneous subgroups. Thus far, there is scarce literature on using this technique to create clinical prediction tools for aneurysmal subarachnoid hemorrhage (SAH). The classification and regression tree analysis technique was applied to the multicenter Tirilazad database (3551 patients) in order to create the decision-making algorithm. In order to elucidate prognostic subgroups in aneurysmal SAH, neurologic, systemic, and demographic factors were taken into account. The dependent variable used for analysis was the dichotomized Glasgow Outcome Score at 3 months. Classification and regression tree analysis revealed seven prognostic subgroups. Neurological grade, occurrence of post-admission stroke, occurrence of post-admission fever, and age represented the explanatory nodes of this decision tree. Split sample validation revealed classification accuracy of 79% for the training dataset and 77% for the testing dataset. In addition, the occurrence of fever at 1-week post-aneurysmal SAH is associated with increased odds of post-admission stroke (odds ratio: 1.83, 95% confidence interval: 1.56-2.45, P tree was generated, which serves as a prediction tool to guide bedside prognostication and clinical treatment decision making. This prognostic decision-making algorithm also shed light on the complex interactions between a number of risk factors in determining outcome after aneurysmal SAH.
Inferring gene expression dynamics via functional regression analysis

Directory of Open Access Journals (Sweden)

Leng Xiaoyan

2008-01-01

Full Text Available Abstract Background Temporal gene expression profiles characterize the time-dynamics of expression of specific genes and are increasingly collected in current gene expression experiments. In the analysis of experiments where gene expression is obtained over the life cycle, it is of interest to relate temporal patterns of gene expression associated with different developmental stages to each other to study patterns of long-term developmental gene regulation. We use tools from functional data analysis to study dynamic changes by relating temporal gene expression profiles of different developmental stages to each other. Results We demonstrate that functional regression methodology can pinpoint relationships that exist between temporary gene expression profiles for different life cycle phases and incorporates dimension reduction as needed for these high-dimensional data. By applying these tools, gene expression profiles for pupa and adult phases are found to be strongly related to the profiles of the same genes obtained during the embryo phase. Moreover, one can distinguish between gene groups that exhibit relationships with positive and others with negative associations between later life and embryonal expression profiles. Specifically, we find a positive relationship in expression for muscle development related genes, and a negative relationship for strictly maternal genes for Drosophila, using temporal gene expression profiles. Conclusion Our findings point to specific reactivation patterns of gene expression during the Drosophila life cycle which differ in characteristic ways between various gene groups. Functional regression emerges as a useful tool for relating gene expression patterns from different developmental stages, and avoids the problems with large numbers of parameters and multiple testing that affect alternative approaches.
In-Season Yield Prediction of Cabbage with a Hand-Held Active Canopy Sensor.

Science.gov (United States)

Ji, Rongting; Min, Ju; Wang, Yuan; Cheng, Hu; Zhang, Hailin; Shi, Weiming

2017-10-08

Efficient and precise yield prediction is critical to optimize cabbage yields and guide fertilizer application. A two-year field experiment was conducted to establish a yield prediction model for cabbage by using the Greenseeker hand-held optical sensor. Two cabbage cultivars (Jianbao and Pingbao) were used and Jianbao cultivar was grown for 2 consecutive seasons but Pingbao was only grown in the second season. Four chemical nitrogen application rates were implemented: 0, 80, 140, and 200 kg·N·ha -1 . Normalized difference vegetation index (NDVI) was collected 20, 50, 70, 80, 90, 100, 110, 120, 130, and 140 days after transplanting (DAT). Pearson correlation analysis and regression analysis were performed to identify the relationship between the NDVI measurements and harvested yields of cabbage. NDVI measurements obtained at 110 DAT were significantly correlated to yield and explained 87-89% and 75-82% of the cabbage yield variation of Jianbao cultivar over the two-year experiment and 77-81% of the yield variability of Pingbao cultivar. Adjusting the yield prediction models with CGDD (cumulative growing degree days) could make remarkable improvement to the accuracy of the prediction model and increase the determination coefficient to 0.82, while the modification with DFP (days from transplanting when GDD > 0) values did not. The integrated exponential yield prediction equation was better than linear or quadratic functions and could accurately make in-season estimation of cabbage yields with different cultivars between years.

Survival analysis II: Cox regression

NARCIS (Netherlands)

Stel, Vianda S.; Dekker, Friedo W.; Tripepi, Giovanni; Zoccali, Carmine; Jager, Kitty J.

2011-01-01

In contrast to the Kaplan-Meier method, Cox proportional hazards regression can provide an effect estimate by quantifying the difference in survival between patient groups and can adjust for confounding effects of other variables. The purpose of this article is to explain the basic concepts of the
Use of generalized ordered logistic regression for the analysis of multidrug resistance data.

Science.gov (United States)

Agga, Getahun E; Scott, H Morgan

2015-10-01

Statistical analysis of antimicrobial resistance data largely focuses on individual antimicrobial's binary outcome (susceptible or resistant). However, bacteria are becoming increasingly multidrug resistant (MDR). Statistical analysis of MDR data is mostly descriptive often with tabular or graphical presentations. Here we report the applicability of generalized ordinal logistic regression model for the analysis of MDR data. A total of 1,152 Escherichia coli, isolated from the feces of weaned pigs experimentally supplemented with chlortetracycline (CTC) and copper, were tested for susceptibilities against 15 antimicrobials and were binary classified into resistant or susceptible. The 15 antimicrobial agents tested were grouped into eight different antimicrobial classes. We defined MDR as the number of antimicrobial classes to which E. coli isolates were resistant ranging from 0 to 8. Proportionality of the odds assumption of the ordinal logistic regression model was violated only for the effect of treatment period (pre-treatment, during-treatment and post-treatment); but not for the effect of CTC or copper supplementation. Subsequently, a partially constrained generalized ordinal logistic model was built that allows for the effect of treatment period to vary while constraining the effects of treatment (CTC and copper supplementation) to be constant across the levels of MDR classes. Copper (Proportional Odds Ratio [Prop OR]=1.03; 95% CI=0.73-1.47) and CTC (Prop OR=1.1; 95% CI=0.78-1.56) supplementation were not significantly associated with the level of MDR adjusted for the effect of treatment period. MDR generally declined over the trial period. In conclusion, generalized ordered logistic regression can be used for the analysis of ordinal data such as MDR data when the proportionality assumptions for ordered logistic regression are violated. Published by Elsevier B.V.
Regression analysis of growth responses to water depth in three wetland plant species

DEFF Research Database (Denmark)

Sorrell, Brian K; Tanner, Chris C; Brix, Hans

2012-01-01

depths from 0 – 0.5 m. Morphological and growth responses to depth were followed for 54 days before harvest, and then analysed by repeated measures analysis of covariance, and non-linear and quantile regression analysis (QRA), to compare flooding tolerances. Principal results Growth responses to depth...
A SOCIOLOGICAL ANALYSIS OF THE CHILDBEARING COEFFICIENT IN THE ALTAI REGION BASED ON METHOD OF FUZZY LINEAR REGRESSION

Directory of Open Access Journals (Sweden)

Sergei Vladimirovich Varaksin

2017-06-01

Full Text Available Purpose. Construction of a mathematical model of the dynamics of childbearing change in the Altai region in 2000–2016, analysis of the dynamics of changes in birth rates for multiple age categories of women of childbearing age. Methodology. A auxiliary analysis element is the construction of linear mathematical models of the dynamics of childbearing by using fuzzy linear regression method based on fuzzy numbers. Fuzzy linear regression is considered as an alternative to standard statistical linear regression for short time series and unknown distribution law. The parameters of fuzzy linear and standard statistical regressions for childbearing time series were defined with using the built in language MatLab algorithm. Method of fuzzy linear regression is not used in sociological researches yet. Results. There are made the conclusions about the socio-demographic changes in society, the high efficiency of the demographic policy of the leadership of the region and the country, and the applicability of the method of fuzzy linear regression for sociological analysis.
Correlation, path analysis and heritability estimation for agronomic traits contribute to yield on soybean

Science.gov (United States)

Sulistyo, A.; Purwantoro; Sari, K. P.

2018-01-01

Selection is a routine activity in plant breeding programs that must be done by plant breeders in obtaining superior plant genotypes. The use of appropriate selection criteria will determine the effectiveness of selection activities. The purpose of this study was to analysis the inheritable agronomic traits that contribute to soybean yield. A total of 91 soybean lines were planted in Muneng Experimental Station, Probolinggo District, East Java Province, Indonesia in 2016. All soybean lines were arranged in randomized complete block design with two replicates. Correlation analysis, path analysis and heritability estimation were performed on days to flowering, days to maturing, plant height, number of branches, number of fertile nodes, number of filled pods, weight of 100 seeds, and yield to determine selection criteria on soybean breeding program. The results showed that the heritability value of almost all agronomic traits observed is high except for the number of fertile nodes with low heritability. The result of correlation analysis shows that days to flowering, plant height and number of fertile nodes have positive correlation with seed yield per plot (0.056, 0.444, and 0.100, respectively). In addition, path analysis showed that plant height and number of fertile nodes have highest positive direct effect on soybean yield. Based on this result, plant height can be selected as one of selection criteria in soybean breeding program to obtain high yielding soybean variety.
[Hyperspectral Estimation of Apple Tree Canopy LAI Based on SVM and RF Regression].

Science.gov (United States)

Han, Zhao-ying; Zhu, Xi-cun; Fang, Xian-yi; Wang, Zhuo-yuan; Wang, Ling; Zhao, Geng-Xing; Jiang, Yuan-mao

2016-03-01

Leaf area index (LAI) is the dynamic index of crop population size. Hyperspectral technology can be used to estimate apple canopy LAI rapidly and nondestructively. It can be provide a reference for monitoring the tree growing and yield estimation. The Red Fuji apple trees of full bearing fruit are the researching objects. Ninety apple trees canopies spectral reflectance and LAI values were measured by the ASD Fieldspec3 spectrometer and LAI-2200 in thirty orchards in constant two years in Qixia research area of Shandong Province. The optimal vegetation indices were selected by the method of correlation analysis of the original spectral reflectance and vegetation indices. The models of predicting the LAI were built with the multivariate regression analysis method of support vector machine (SVM) and random forest (RF). The new vegetation indices, GNDVI527, ND-VI676, RVI682, FD-NVI656 and GRVI517 and the previous two main vegetation indices, NDVI670 and NDVI705, are in accordance with LAI. In the RF regression model, the calibration set decision coefficient C-R2 of 0.920 and validation set decision coefficient V-R2 of 0.889 are higher than the SVM regression model by 0.045 and 0.033 respectively. The root mean square error of calibration set C-RMSE of 0.249, the root mean square error validation set V-RMSE of 0.236 are lower than that of the SVM regression model by 0.054 and 0.058 respectively. Relative analysis of calibrating error C-RPD and relative analysis of validation set V-RPD reached 3.363 and 2.520, 0.598 and 0.262, respectively, which were higher than the SVM regression model. The measured and predicted the scatterplot trend line slope of the calibration set and validation set C-S and V-S are close to 1. The estimation result of RF regression model is better than that of the SVM. RF regression model can be used to estimate the LAI of red Fuji apple trees in full fruit period.
Multiple Logistic Regression Analysis of Cigarette Use among High School Students

Science.gov (United States)

Adwere-Boamah, Joseph

2011-01-01

A binary logistic regression analysis was performed to predict high school students' cigarette smoking behavior from selected predictors from 2009 CDC Youth Risk Behavior Surveillance Survey. The specific target student behavior of interest was frequent cigarette use. Five predictor variables included in the model were: a) race, b) frequency of…
Stability Parameters for Grain Yield and its Component Traits in Maize Hybrids of Different FAO Maturity Groups

Directory of Open Access Journals (Sweden)

Dragan Djurovic

2014-12-01

Full Text Available An objective evaluation of maize hybrids in intensive cropping systems requires identification not only of yield components and other agronomically important traits but also of stability parameters. Grain yield and its components were assessed in 11 maize hybrids with different lengths of growing season (FAO 300-700 maturity groups using analysis of variance and regression analysis at three different locations in Western Serbia. The test hybrids and locations showed significant differences in grain yield, grain moisture content at maturity, 1,000-kernel weight and ear length. A significant interaction was observed between all traits and the environment. The hybrids with higher mean values of the traits, regardless of maturity group, generally exhibited sensitivity i.e. adaptation to more favourable environmental conditions as compared to those having lower mean values. Regression coefficient (bi values for grain yield mostly suggested no significant differences relative to the mean. The medium-season hybrid gave high yields and less favourable values of stability parameters at most locations and in most years, as compared to mediumlate hybrids. As compared to medium-early hybrids, medium-late hybrids (FAO 600 and 700 mostly exhibited unfavourable values of stability parameters i.e. a specific response and better adaptation to favourable environmental conditions, and gave higher average yields. Apart from producing lower average yields, FAO 300 and 400 hybrids showed higher yield stability as compared to the other hybrids tested. Medium-late hybrids had higher yields and showed a better response to favourable environmental conditions compared to early-maturing hybrids. Therefore, they can be recommended for intensive cultural practices and low-stress environments. Due to their more favourable stability parameter values, medium-early hybrids can be recommended for low-intensity cultural practices and stressful environments.
THE PROGNOSIS OF RUSSIAN DEFENSE INDUSTRY DEVELOPMENT IMPLEMENTED THROUGH REGRESSION ANALYSIS

Directory of Open Access Journals (Sweden)

L.M. Kapustina

2007-03-01

Full Text Available The article illustrates the results of investigation the major internal and external factors which influence the development of the defense industry, as well as the results of regression analysis which quantitatively displays the factorial contribution in the growth rate of Russian defense industry. On the basis of calculated regression dependences the authors fulfilled the medium-term prognosis of defense industry. Optimistic and inertial versions of defense product growth rate for the period up to 2009 are based on scenario conditions in Russian economy worked out by the Ministry of economy and development. In conclusion authors point out which factors and conditions have the largest impact on successful and stable operation of Russian defense industry.
Lower bound limit analysis of slabs with nonlinear yield criteria

DEFF Research Database (Denmark)

Krabbenhøft, Kristian; Damkilde, Lars

2002-01-01

A finite element formulation of the limit analysis of perfectly plastic slabs is given. An element with linear moment fields for which equilibrium is satisfied exactly is used in connection with an optimization algorithm taking into account the full nonlinearity of the yield criteria. Both load...... and material optimization problems are formulated and by means of the duality theory of linear programming the displacements are extracted from the dual variables. Numerical examples demonstrating the capabilities of the method and the effects of using a more refined representation of the yield criteria...
Applied linear regression

CERN Document Server

Weisberg, Sanford

2013-01-01

Praise for the Third Edition ""...this is an excellent book which could easily be used as a course text...""-International Statistical Institute The Fourth Edition of Applied Linear Regression provides a thorough update of the basic theory and methodology of linear regression modeling. Demonstrating the practical applications of linear regression analysis techniques, the Fourth Edition uses interesting, real-world exercises and examples. Stressing central concepts such as model building, understanding parameters, assessing fit and reliability, and drawing conclusions, the new edition illus
A Bayesian goodness of fit test and semiparametric generalization of logistic regression with measurement data.

Science.gov (United States)

Schörgendorfer, Angela; Branscum, Adam J; Hanson, Timothy E

2013-06-01

Logistic regression is a popular tool for risk analysis in medical and population health science. With continuous response data, it is common to create a dichotomous outcome for logistic regression analysis by specifying a threshold for positivity. Fitting a linear regression to the nondichotomized response variable assuming a logistic sampling model for the data has been empirically shown to yield more efficient estimates of odds ratios than ordinary logistic regression of the dichotomized endpoint. We illustrate that risk inference is not robust to departures from the parametric logistic distribution. Moreover, the model assumption of proportional odds is generally not satisfied when the condition of a logistic distribution for the data is violated, leading to biased inference from a parametric logistic analysis. We develop novel Bayesian semiparametric methodology for testing goodness of fit of parametric logistic regression with continuous measurement data. The testing procedures hold for any cutoff threshold and our approach simultaneously provides the ability to perform semiparametric risk estimation. Bayes factors are calculated using the Savage-Dickey ratio for testing the null hypothesis of logistic regression versus a semiparametric generalization. We propose a fully Bayesian and a computationally efficient empirical Bayesian approach to testing, and we present methods for semiparametric estimation of risks, relative risks, and odds ratios when parametric logistic regression fails. Theoretical results establish the consistency of the empirical Bayes test. Results from simulated data show that the proposed approach provides accurate inference irrespective of whether parametric assumptions hold or not. Evaluation of risk factors for obesity shows that different inferences are derived from an analysis of a real data set when deviations from a logistic distribution are permissible in a flexible semiparametric framework. © 2013, The International Biometric
Use of empirical likelihood to calibrate auxiliary information in partly linear monotone regression models.

Science.gov (United States)

Chen, Baojiang; Qin, Jing

2014-05-10

In statistical analysis, a regression model is needed if one is interested in finding the relationship between a response variable and covariates. When the response depends on the covariate, then it may also depend on the function of this covariate. If one has no knowledge of this functional form but expect for monotonic increasing or decreasing, then the isotonic regression model is preferable. Estimation of parameters for isotonic regression models is based on the pool-adjacent-violators algorithm (PAVA), where the monotonicity constraints are built in. With missing data, people often employ the augmented estimating method to improve estimation efficiency by incorporating auxiliary information through a working regression model. However, under the framework of the isotonic regression model, the PAVA does not work as the monotonicity constraints are violated. In this paper, we develop an empirical likelihood-based method for isotonic regression model to incorporate the auxiliary information. Because the monotonicity constraints still hold, the PAVA can be used for parameter estimation. Simulation studies demonstrate that the proposed method can yield more efficient estimates, and in some situations, the efficiency improvement is substantial. We apply this method to a dementia study. Copyright © 2013 John Wiley & Sons, Ltd.
Variational formulation based analysis on growth of yield front in ...

African Journals Online (AJOL)

user

The analysis of rotating disk behavior has been of great interest to many ... strain hardening using Tresca's yield condition and its associated flow rule ...... Determination of Stresses in Gas-Turbine Disks Subjected to Plastic Flow and Creep.
Noninvasive spectral imaging of skin chromophores based on multiple regression analysis aided by Monte Carlo simulation

Science.gov (United States)

Nishidate, Izumi; Wiswadarma, Aditya; Hase, Yota; Tanaka, Noriyuki; Maeda, Takaaki; Niizeki, Kyuichi; Aizu, Yoshihisa

2011-08-01

In order to visualize melanin and blood concentrations and oxygen saturation in human skin tissue, a simple imaging technique based on multispectral diffuse reflectance images acquired at six wavelengths (500, 520, 540, 560, 580 and 600nm) was developed. The technique utilizes multiple regression analysis aided by Monte Carlo simulation for diffuse reflectance spectra. Using the absorbance spectrum as a response variable and the extinction coefficients of melanin, oxygenated hemoglobin, and deoxygenated hemoglobin as predictor variables, multiple regression analysis provides regression coefficients. Concentrations of melanin and total blood are then determined from the regression coefficients using conversion vectors that are deduced numerically in advance, while oxygen saturation is obtained directly from the regression coefficients. Experiments with a tissue-like agar gel phantom validated the method. In vivo experiments with human skin of the human hand during upper limb occlusion and of the inner forearm exposed to UV irradiation demonstrated the ability of the method to evaluate physiological reactions of human skin tissue.
Nitrogen utilization and biomass yield in trickle bed air biofilters.

Science.gov (United States)

Kim, Daekeun; Sorial, George A

2010-10-15

Nitrogen utilization and subsequent biomass yield were investigated in four independent lab-scale trickle bed air biofilters (TBABs) fed with different VOCs substrate. The VOCs considered were two aromatic (toluene, styrene) and two oxygenated (methyl ethyl ketone (MEK), methyl isobutyl ketone (MIBK)). Long-term observations of TBABs performances show that more nitrogen was required to sustain high VOC removal, but the one fed with a high loading of VOC utilized much more nitrogen for sustaining biomass yield. The ratio N(consumption)/N(growth) was an effective indicator in evaluating nitrogen utilization in the system. Substrate VOC availability in the system was significant in determining nitrogen utilization and biomass yield. VOC substrate availability in the TBAB system was effectively identified by using maximum practical concentrations in the biofilm. Biomass yield coefficient, which was driven from the regression analysis between CO(2) production rate and substrate consumption rate, was effective in evaluating the TBAB performance with respect to nitrogen utilization and VOC removal. Biomass yield coefficients (g biomass/g substrate, dry weight basis) were observed to be 0.668, 0.642, 0.737, and 0.939 for toluene, styrene, MEK, and MIBK, respectively. 2010 Elsevier B.V. All rights reserved.
Multiple Regression Analysis of Unconfined Compression Strength of Mine Tailings Matrices

Directory of Open Access Journals (Sweden)

Mahmood Ali A.

2017-01-01

Full Text Available As part of a novel approach of sustainable development of mine tailings, experimental and numerical analysis is carried out on newly formulated tailings matrices. Several physical characteristic tests are carried out including the unconfined compression strength test to ascertain the integrity of these matrices when subjected to loading. The current paper attempts a multiple regression analysis of the unconfined compressive strength test results of these matrices to investigate the most pertinent factors affecting their strength. Results of this analysis showed that the suggested equation is reasonably applicable to the range of binder combinations used.
The Regression Analysis of Individual Financial Performance: Evidence from Croatia

OpenAIRE

Bahovec, Vlasta; Barbić, Dajana; Palić, Irena

2017-01-01

Background: A large body of empirical literature indicates that gender and financial literacy are significant determinants of individual financial performance. Objectives: The purpose of this paper is to recognize the impact of the variable financial literacy and the variable gender on the variation of the financial performance using the regression analysis. Methods/Approach: The survey was conducted using the systematically chosen random sample of Croatian financial consumers. The cross sect...
A systematic review and meta-regression analysis of mivacurium for tracheal intubation

NARCIS (Netherlands)

Vanlinthout, L.E.H.; Mesfin, S.H.; Hens, N.; Vanacker, B.F.; Robertson, E.N.; Booij, L.H.D.J.

2014-01-01

We systematically reviewed factors associated with intubation conditions in randomised controlled trials of mivacurium, using random-effects meta-regression analysis. We included 29 studies of 1050 healthy participants. Four factors explained 72.9% of the variation in the probability of excellent
No rationale for 1 variable per 10 events criterion for binary logistic regression analysis.

Science.gov (United States)

van Smeden, Maarten; de Groot, Joris A H; Moons, Karel G M; Collins, Gary S; Altman, Douglas G; Eijkemans, Marinus J C; Reitsma, Johannes B

2016-11-24

Ten events per variable (EPV) is a widely advocated minimal criterion for sample size considerations in logistic regression analysis. Of three previous simulation studies that examined this minimal EPV criterion only one supports the use of a minimum of 10 EPV. In this paper, we examine the reasons for substantial differences between these extensive simulation studies. The current study uses Monte Carlo simulations to evaluate small sample bias, coverage of confidence intervals and mean square error of logit coefficients. Logistic regression models fitted by maximum likelihood and a modified estimation procedure, known as Firth's correction, are compared. The results show that besides EPV, the problems associated with low EPV depend on other factors such as the total sample size. It is also demonstrated that simulation results can be dominated by even a few simulated data sets for which the prediction of the outcome by the covariates is perfect ('separation'). We reveal that different approaches for identifying and handling separation leads to substantially different simulation results. We further show that Firth's correction can be used to improve the accuracy of regression coefficients and alleviate the problems associated with separation. The current evidence supporting EPV rules for binary logistic regression is weak. Given our findings, there is an urgent need for new research to provide guidance for supporting sample size considerations for binary logistic regression analysis.

Yield QTL analysis of Oryza sativa x O. glumaepatula introgression lines

Directory of Open Access Journals (Sweden)

Priscila Nascimento Rangel

2013-03-01

Full Text Available The objective of this work was to evaluate the yield performance of two generations (BC2F2 and BC2F9 of introgression lines developed from the interspecific cross between Oryza sativa and O. glumaepatula, and to identify the SSR markers associated to yield. The wild accession RS‑16 (O. glumaepatula was used as donor parent in the backcross with the high yielding cultivar Cica‑8 (O. sativa. A set of 114 BC2F1 introgression lines was genotyped with 141 polymorphic SSR loci distributed across the whole rice genome. Molecular analysis showed that in average 22% of the O. glumaepatula genome was introgressed into BC2F1 generation. Nine BC2F9 introgression lines had a significantly higher yield than the genitor Cica‑8, thus showing a positive genome interaction among cultivated rice and the wild O. glumaepatula. Seven QTL were identified in the overall BC2F2, with one marker interval (4879‑EST20 of great effect on yield. The alleles with positive effect on yield came from the cultivated parent Cica‑8.
Estimate the contribution of incubation parameters influence egg hatchability using multiple linear regression analysis.

Science.gov (United States)

Khalil, Mohamed H; Shebl, Mostafa K; Kosba, Mohamed A; El-Sabrout, Karim; Zaki, Nesma

2016-08-01

This research was conducted to determine the most affecting parameters on hatchability of indigenous and improved local chickens' eggs. Five parameters were studied (fertility, early and late embryonic mortalities, shape index, egg weight, and egg weight loss) on four strains, namely Fayoumi, Alexandria, Matrouh, and Montazah. Multiple linear regression was performed on the studied parameters to determine the most influencing one on hatchability. The results showed significant differences in commercial and scientific hatchability among strains. Alexandria strain has the highest significant commercial hatchability (80.70%). Regarding the studied strains, highly significant differences in hatching chick weight among strains were observed. Using multiple linear regression analysis, fertility made the greatest percent contribution (71.31%) to hatchability, and the lowest percent contributions were made by shape index and egg weight loss. A prediction of hatchability using multiple regression analysis could be a good tool to improve hatchability percentage in chickens.
Logistic Regression: Concept and Application

Science.gov (United States)

Cokluk, Omay

2010-01-01

The main focus of logistic regression analysis is classification of individuals in different groups. The aim of the present study is to explain basic concepts and processes of binary logistic regression analysis intended to determine the combination of independent variables which best explain the membership in certain groups called dichotomous…
Lowland rice yield estimates based on air temperature and solar radiation

International Nuclear Information System (INIS)

Pedro Júnior, M.J.; Sentelhas, P.C.; Moraes, A.V.C.; Villela, O.V.

1995-01-01

Two regression equations were developed to estimate lowland rice yield as a function of air temperature and incoming solar radiation, during the crop yield production period in Pindamonhangaba, SP, Brazil. The following rice cultivars were used: IAC-242, IAC-100, IAC-101 and IAC-102. The value of optimum air temperature obtained was 25.0°C and of optimum global solar radiation was 475 cal.cm -2 , day -1 . The best agrometeorological model was the one that related least deviation of air temperature and solar radiation in relation to the optimum value obtained through a multiple linear regression. The yield values estimated by the model showed good fit to actual yields of lowland rice (less than 10%). (author) [pt
Estimation Methods for Non-Homogeneous Regression - Minimum CRPS vs Maximum Likelihood

Science.gov (United States)

Gebetsberger, Manuel; Messner, Jakob W.; Mayr, Georg J.; Zeileis, Achim

2017-04-01

Non-homogeneous regression models are widely used to statistically post-process numerical weather prediction models. Such regression models correct for errors in mean and variance and are capable to forecast a full probability distribution. In order to estimate the corresponding regression coefficients, CRPS minimization is performed in many meteorological post-processing studies since the last decade. In contrast to maximum likelihood estimation, CRPS minimization is claimed to yield more calibrated forecasts. Theoretically, both scoring rules used as an optimization score should be able to locate a similar and unknown optimum. Discrepancies might result from a wrong distributional assumption of the observed quantity. To address this theoretical concept, this study compares maximum likelihood and minimum CRPS estimation for different distributional assumptions. First, a synthetic case study shows that, for an appropriate distributional assumption, both estimation methods yield to similar regression coefficients. The log-likelihood estimator is slightly more efficient. A real world case study for surface temperature forecasts at different sites in Europe confirms these results but shows that surface temperature does not always follow the classical assumption of a Gaussian distribution. KEYWORDS: ensemble post-processing, maximum likelihood estimation, CRPS minimization, probabilistic temperature forecasting, distributional regression models
A multiple regression analysis for accurate background subtraction in 99Tcm-DTPA renography

International Nuclear Information System (INIS)

Middleton, G.W.; Thomson, W.H.; Davies, I.H.; Morgan, A.

1989-01-01

A technique for accurate background subtraction in 99 Tc m -DTPA renography is described. The technique is based on a multiple regression analysis of the renal curves and separate heart and soft tissue curves which together represent background activity. It is compared, in over 100 renograms, with a previously described linear regression technique. Results show that the method provides accurate background subtraction, even in very poorly functioning kidneys, thus enabling relative renal filtration and excretion to be accurately estimated. (author)
Random regression models to account for the effect of genotype by environment interaction due to heat stress on the milk yield of Holstein cows under tropical conditions.

Science.gov (United States)

Santana, Mário L; Bignardi, Annaiza Braga; Pereira, Rodrigo Junqueira; Menéndez-Buxadera, Alberto; El Faro, Lenira

2016-02-01

The present study had the following objectives: to compare random regression models (RRM) considering the time-dependent (days in milk, DIM) and/or temperature × humidity-dependent (THI) covariate for genetic evaluation; to identify the effect of genotype by environment interaction (G×E) due to heat stress on milk yield; and to quantify the loss of milk yield due to heat stress across lactation of cows under tropical conditions. A total of 937,771 test-day records from 3603 first lactations of Brazilian Holstein cows obtained between 2007 and 2013 were analyzed. An important reduction in milk yield due to heat stress was observed for THI values above 66 (-0.23 kg/day/THI). Three phases of milk yield loss were identified during lactation, the most damaging one at the end of lactation (-0.27 kg/day/THI). Using the most complex RRM, the additive genetic variance could be altered simultaneously as a function of both DIM and THI values. This model could be recommended for the genetic evaluation taking into account the effect of G×E. The response to selection in the comfort zone (THI ≤ 66) is expected to be higher than that obtained in the heat stress zone (THI > 66) of the animals. The genetic correlations between milk yield in the comfort and heat stress zones were less than unity at opposite extremes of the environmental gradient. Thus, the best animals for milk yield in the comfort zone are not necessarily the best in the zone of heat stress and, therefore, G×E due to heat stress should not be neglected in the genetic evaluation.
Development of an empirical model of turbine efficiency using the Taylor expansion and regression analysis

International Nuclear Information System (INIS)

Fang, Xiande; Xu, Yu

2011-01-01

The empirical model of turbine efficiency is necessary for the control- and/or diagnosis-oriented simulation and useful for the simulation and analysis of dynamic performances of the turbine equipment and systems, such as air cycle refrigeration systems, power plants, turbine engines, and turbochargers. Existing empirical models of turbine efficiency are insufficient because there is no suitable form available for air cycle refrigeration turbines. This work performs a critical review of empirical models (called mean value models in some literature) of turbine efficiency and develops an empirical model in the desired form for air cycle refrigeration, the dominant cooling approach in aircraft environmental control systems. The Taylor series and regression analysis are used to build the model, with the Taylor series being used to expand functions with the polytropic exponent and the regression analysis to finalize the model. The measured data of a turbocharger turbine and two air cycle refrigeration turbines are used for the regression analysis. The proposed model is compact and able to present the turbine efficiency map. Its predictions agree with the measured data very well, with the corrected coefficient of determination R c 2 ≥ 0.96 and the mean absolute percentage deviation = 1.19% for the three turbines. -- Highlights: → Performed a critical review of empirical models of turbine efficiency. → Developed an empirical model in the desired form for air cycle refrigeration, using the Taylor expansion and regression analysis. → Verified the method for developing the empirical model. → Verified the model.
Econometric analysis of realized covariation: high frequency based covariance, regression, and correlation in financial economics

DEFF Research Database (Denmark)

Barndorff-Nielsen, Ole Eiler; Shephard, N.

2004-01-01

This paper analyses multivariate high frequency financial data using realized covariation. We provide a new asymptotic distribution theory for standard methods such as regression, correlation analysis, and covariance. It will be based on a fixed interval of time (e.g., a day or week), allowing...... the number of high frequency returns during this period to go to infinity. Our analysis allows us to study how high frequency correlations, regressions, and covariances change through time. In particular we provide confidence intervals for each of these quantities....
Health care: necessity or luxury good? A meta-regression analysis

OpenAIRE

Iordache, Ioana Raluca

2014-01-01

When estimating the influence income per capita exerts on health care expenditure, the research in the field offers mixed results. Studies employ different data, estimation techniques and models, which brings about the question whether these differences in research design play any part in explaining the heterogeneity of reported outcomes. By employing meta-regression analysis, the present paper analyzes 220 estimates of health spending income elasticity collected from 54 studies and finds tha...
Inheritance of grain yield and its correlation with yield components in ...

African Journals Online (AJOL)

SAM

2014-03-19

Mar 19, 2014 ... average yield of wheat in China is 4.75 t ha-1, which is low compared to other .... Analysis of variance for combining ability for grain yield plant-1. Source of variation ..... Hayman BI (1954). The theory and analysis of diallel crosses. .... Analysis and prospect of China wheat market in 2011. Food and Oil.
Distance Based Root Cause Analysis and Change Impact Analysis of Performance Regressions

Directory of Open Access Journals (Sweden)

Junzan Zhou

2015-01-01

Full Text Available Performance regression testing is applied to uncover both performance and functional problems of software releases. A performance problem revealed by performance testing can be high response time, low throughput, or even being out of service. Mature performance testing process helps systematically detect software performance problems. However, it is difficult to identify the root cause and evaluate the potential change impact. In this paper, we present an approach leveraging server side logs for identifying root causes of performance problems. Firstly, server side logs are used to recover call tree of each business transaction. We define a novel distance based metric computed from call trees for root cause analysis and apply inverted index from methods to business transactions for change impact analysis. Empirical studies show that our approach can effectively and efficiently help developers diagnose root cause of performance problems.
On yield gaps and yield gains in intercropping

NARCIS (Netherlands)

Gou, Fang; Yin, Wen; Hong, Yu; Werf, van der Wopke; Chai, Qiang; Heerink, Nico; Ittersum, van Martin K.

2017-01-01

Wheat-maize relay intercropping has been widely used by farmers in northwest China, and based on field experiments agronomists report it has a higher productivity than sole crops. However, the yields from farmers’ fields have not been investigated yet. Yield gap analysis provides a framework to
Sparse multivariate factor analysis regression models and its applications to integrative genomics analysis.

Science.gov (United States)

Zhou, Yan; Wang, Pei; Wang, Xianlong; Zhu, Ji; Song, Peter X-K

2017-01-01

The multivariate regression model is a useful tool to explore complex associations between two kinds of molecular markers, which enables the understanding of the biological pathways underlying disease etiology. For a set of correlated response variables, accounting for such dependency can increase statistical power. Motivated by integrative genomic data analyses, we propose a new methodology-sparse multivariate factor analysis regression model (smFARM), in which correlations of response variables are assumed to follow a factor analysis model with latent factors. This proposed method not only allows us to address the challenge that the number of association parameters is larger than the sample size, but also to adjust for unobserved genetic and/or nongenetic factors that potentially conceal the underlying response-predictor associations. The proposed smFARM is implemented by the EM algorithm and the blockwise coordinate descent algorithm. The proposed methodology is evaluated and compared to the existing methods through extensive simulation studies. Our results show that accounting for latent factors through the proposed smFARM can improve sensitivity of signal detection and accuracy of sparse association map estimation. We illustrate smFARM by two integrative genomics analysis examples, a breast cancer dataset, and an ovarian cancer dataset, to assess the relationship between DNA copy numbers and gene expression arrays to understand genetic regulatory patterns relevant to the disease. We identify two trans-hub regions: one in cytoband 17q12 whose amplification influences the RNA expression levels of important breast cancer genes, and the other in cytoband 9q21.32-33, which is associated with chemoresistance in ovarian cancer. © 2016 WILEY PERIODICALS, INC.
Evaluation of quantitative relationships between saffron yield and nutrition (on farm trial

Directory of Open Access Journals (Sweden)

mohamad ali behdani

2009-06-01

Full Text Available In order to relate production of saffron and utilization of nutrients, a study was conducted in 2001 and 2002. Four selected locations for this study were Birjand, Gonabad, Qaen and Torbat-Haydariah, which are the main saffron production centers in Iran. This study was performed in 160 saffron farms, aged between 1 and 5 years. Manure, nitrogen and phosphorous fertilizers showed a positive linear relation with yield and length of flowering, while nitrogen and phosphorous showed a negative linear relation with start of flowering period. Yield of saffron showed a significant and positive correlation with the amount of applied manure and the saffron farms with age 4-5 year had highest yield. Our results showed that manure was the most effective factor in production of saffron. The beneficial effects of manure could be due to slow release of nutrients and enhancing soil physical properties. Stepwise regression analysis of yield and fertilizer application showed that 67 percent of yield variations was attributed to manure and phosphorous application.
Time series regression-based pairs trading in the Korean equities market

Science.gov (United States)

Kim, Saejoon; Heo, Jun

2017-07-01

Pairs trading is an instance of statistical arbitrage that relies on heavy quantitative data analysis to profit by capitalising low-risk trading opportunities provided by anomalies of related assets. A key element in pairs trading is the rule by which open and close trading triggers are defined. This paper investigates the use of time series regression to define the rule which has previously been identified with fixed threshold-based approaches. Empirical results indicate that our approach may yield significantly increased excess returns compared to ones obtained by previous approaches on large capitalisation stocks in the Korean equities market.
No rationale for 1 variable per 10 events criterion for binary logistic regression analysis

Directory of Open Access Journals (Sweden)

Maarten van Smeden

2016-11-01

Full Text Available Abstract Background Ten events per variable (EPV is a widely advocated minimal criterion for sample size considerations in logistic regression analysis. Of three previous simulation studies that examined this minimal EPV criterion only one supports the use of a minimum of 10 EPV. In this paper, we examine the reasons for substantial differences between these extensive simulation studies. Methods The current study uses Monte Carlo simulations to evaluate small sample bias, coverage of confidence intervals and mean square error of logit coefficients. Logistic regression models fitted by maximum likelihood and a modified estimation procedure, known as Firth’s correction, are compared. Results The results show that besides EPV, the problems associated with low EPV depend on other factors such as the total sample size. It is also demonstrated that simulation results can be dominated by even a few simulated data sets for which the prediction of the outcome by the covariates is perfect (‘separation’. We reveal that different approaches for identifying and handling separation leads to substantially different simulation results. We further show that Firth’s correction can be used to improve the accuracy of regression coefficients and alleviate the problems associated with separation. Conclusions The current evidence supporting EPV rules for binary logistic regression is weak. Given our findings, there is an urgent need for new research to provide guidance for supporting sample size considerations for binary logistic regression analysis.
Sparse Regression by Projection and Sparse Discriminant Analysis

KAUST Repository

Qi, Xin; Luo, Ruiyan; Carroll, Raymond J.; Zhao, Hongyu

2015-01-01

predictions. We introduce a new framework, regression by projection, and its sparse version to analyze high-dimensional data. The unique nature of this framework is that the directions of the regression coefficients are inferred first, and the lengths
Relationship Between Seed Yield And Some of Fruit Traits in Iranian Squash (Cucurbita pepo L. Accissions

Directory of Open Access Journals (Sweden)

R. Barzegar

2016-02-01

Full Text Available In order to evaluation of squash (Cucurbita pepo seed yield per fruit and its relations with other characteristics of fruit include: length, diameter, length: diameter ratio (fruit shape, flesh thickness, thousand seed weight and fruit weight, an experiment was conducted using 24 accessions of squash as a randomized complete-block design with three replications. Morphological traits were evaluated according to UPOV descriptor and UPGMA clustering algorithm clustered the accessions in 4 groups (predominantly on the basis of fruit shape. Correlation, regression and path analysis were done for mentioned characteristics in 4 type-fruit groups. There was negative correlation between seed yield of individual fruit and its length and fruit length: diameter ratio. But fruit weight, fruit diameter, and thousand seeds weight had positive correlation with seed yield. Seed weight: fruit weight ratio had negative relationship with fruit weight. Therefore small size fruit is more suitable for seed yield per area. Path analysis was showed fruit weight had the most positive direct effect on seed yield per fruit in all groups.
GENETIC ANALYSIS OF YIELD AND YIELD COMPONENTS IN ...

African Journals Online (AJOL)

ACSS

2017-11-16

Nov 16, 2017 ... used different genotypes and the environmental conditions under which their ... and Jinks (1971):. Y = m + aa + βd + a2aa + 2aβad +β2dd … .... /plant, 100-grain weight per plant and Grain yield per plant (g) of six generations in IET6279 X IR70445-146-3-. 3 cross. Traits. Generation. Mean. Standard. Range.

Statistical methods and regression analysis of stratospheric ozone and meteorological variables in Isfahan

Science.gov (United States)

Hassanzadeh, S.; Hosseinibalam, F.; Omidvari, M.

2008-04-01

Data of seven meteorological variables (relative humidity, wet temperature, dry temperature, maximum temperature, minimum temperature, ground temperature and sun radiation time) and ozone values have been used for statistical analysis. Meteorological variables and ozone values were analyzed using both multiple linear regression and principal component methods. Data for the period 1999-2004 are analyzed jointly using both methods. For all periods, temperature dependent variables were highly correlated, but were all negatively correlated with relative humidity. Multiple regression analysis was used to fit the meteorological variables using the meteorological variables as predictors. A variable selection method based on high loading of varimax rotated principal components was used to obtain subsets of the predictor variables to be included in the linear regression model of the meteorological variables. In 1999, 2001 and 2002 one of the meteorological variables was weakly influenced predominantly by the ozone concentrations. However, the model did not predict that the meteorological variables for the year 2000 were not influenced predominantly by the ozone concentrations that point to variation in sun radiation. This could be due to other factors that were not explicitly considered in this study.
Analysis of yield advantage in mixed cropping

NARCIS (Netherlands)

Ranganathan, R.

1993-01-01

It has long been recognized that mixed cropping can give yield advantages over sole cropping, but methods that can identify such yield benefits are still being developed. This thesis presents a method that combines physiological and economic principles in the evaluation of yield advantage.
Determining the Most Important Soil Properties Affecting the Yield of Saffron in the Ghayenat Area

Directory of Open Access Journals (Sweden)

amir ranjbar

2016-02-01

Full Text Available Introduction: Saffron is one of the most important economic plants in the Khorasan province. Awareness of soil quality in agricultural lands is essential for the best management of lands and for obtaining maximum economic benefit. In general, plant growth is a function of environmental factors especially chemical and physical properties of soil (20. It has been demonstrated that there was a positive and high correlation between soil organic matter and saffron yield. Increasing the yield of saffron due to organic matter is probably due to soil nutrient, especially phosphorous and nitrogen and also improvement of soil physical quality (6, 28, 29. The yield of saffron in soils with high nitrogen as a result of vegetative growth is high (8. Shahandeh (6 found that most of the variation of saffron yield depends on soil properties. Due to the economic importance of saffron and the role of soil properties on saffron yield, this research was conducted to find the relationship between saffron yield and some soil physical and chemical properties, and to determine the contribution of soil properties that have the greatest impact on saffron yield in the Ghayenat area. Materials and Methods: This research was performed in 30 saffron fields (30 soil samples of the Ghayenat area (longitude 59° 10΄ 10.37˝ - 59° 11΄ 38.41˝ and latitude 33° 43΄ 35.08˝ - 33΄ 44΄ 02.78˝, which is located in the Khrasan province of Iran. In this research, 21 soil properties were regarded as the total data set (TDS. Then the principal component analysis (PCA was used to determine the most important soil properties affecting saffron yield as a minimum data set (MDS and the stepwise regression to estimate saffron yield. To estimate the yield of saffron in stepwise regression method, saffron yield was considered as a dependent variable and soil physical and chemical properties were considered to be independent variables. Results and Discussion: According to the PCA method
Multivariate regression analysis for determining short-term values of radon and its decay products from filter measurements

International Nuclear Information System (INIS)

Kraut, W.; Schwarz, W.; Wilhelm, A.

1994-01-01

A multivariate regression analysis is applied to decay measurements of α-resp. β-filter activcity. Activity concentrations for Po-218, Pb-214 and Bi-214, resp. for the Rn-222 equilibrium equivalent concentration are obtained explicitly. The regression analysis takes into account properly the variances of the measured count rates and their influence on the resulting activity concentrations. (orig.) [de
Predictive ability of machine learning methods for massive crop yield prediction

Directory of Open Access Journals (Sweden)

Alberto Gonzalez-Sanchez

2014-04-01

Full Text Available An important issue for agricultural planning purposes is the accurate yield estimation for the numerous crops involved in the planning. Machine learning (ML is an essential approach for achieving practical and effective solutions for this problem. Many comparisons of ML methods for yield prediction have been made, seeking for the most accurate technique. Generally, the number of evaluated crops and techniques is too low and does not provide enough information for agricultural planning purposes. This paper compares the predictive accuracy of ML and linear regression techniques for crop yield prediction in ten crop datasets. Multiple linear regression, M5-Prime regression trees, perceptron multilayer neural networks, support vector regression and k-nearest neighbor methods were ranked. Four accuracy metrics were used to validate the models: the root mean square error (RMS, root relative square error (RRSE, normalized mean absolute error (MAE, and correlation factor (R. Real data of an irrigation zone of Mexico were used for building the models. Models were tested with samples of two consecutive years. The results show that M5-Prime and k-nearest neighbor techniques obtain the lowest average RMSE errors (5.14 and 4.91, the lowest RRSE errors (79.46% and 79.78%, the lowest average MAE errors (18.12% and 19.42%, and the highest average correlation factors (0.41 and 0.42. Since M5-Prime achieves the largest number of crop yield models with the lowest errors, it is a very suitable tool for massive crop yield prediction in agricultural planning.
A comparison of three methods of assessing differential item functioning (DIF) in the Hospital Anxiety Depression Scale: ordinal logistic regression, Rasch analysis and the Mantel chi-square procedure.

Science.gov (United States)

Cameron, Isobel M; Scott, Neil W; Adler, Mats; Reid, Ian C

2014-12-01

It is important for clinical practice and research that measurement scales of well-being and quality of life exhibit only minimal differential item functioning (DIF). DIF occurs where different groups of people endorse items in a scale to different extents after being matched by the intended scale attribute. We investigate the equivalence or otherwise of common methods of assessing DIF. Three methods of measuring age- and sex-related DIF (ordinal logistic regression, Rasch analysis and Mantel χ(2) procedure) were applied to Hospital Anxiety Depression Scale (HADS) data pertaining to a sample of 1,068 patients consulting primary care practitioners. Three items were flagged by all three approaches as having either age- or sex-related DIF with a consistent direction of effect; a further three items identified did not meet stricter criteria for important DIF using at least one method. When applying strict criteria for significant DIF, ordinal logistic regression was slightly less sensitive. Ordinal logistic regression, Rasch analysis and contingency table methods yielded consistent results when identifying DIF in the HADS depression and HADS anxiety scales. Regardless of methods applied, investigators should use a combination of statistical significance, magnitude of the DIF effect and investigator judgement when interpreting the results.
A Preliminary Study on Rainfall Interception Loss and Water Yield Analysis on Arabica Coffee Plants in Central Aceh Regency, Indonesia

Directory of Open Access Journals (Sweden)

Reza Benara

2012-12-01

Full Text Available Rainfall interception loss from plants or trees can reduce a net rainfall as source of water yield. The amount of rainfall interception loss depends on kinds of plants and hydro-meteorological characteristics. Therefore, it is important to study rainfall interception loss such as from Arabica Coffee plantation which is as main agricultural commodity for Central Aceh Regency. In this study, rainfall interception loss from Arabica Coffee plants was studied in Kebet Village of Central Aceh Regency, Indonesia from January 20 to March 9, 2011. Arabica coffee plants used in this study was 15 years old, height of 1.5 m and canopy of 4.567 m2. Rainfall interception loss was determined based on water balance approach of daily rainfall, throughfall, and stemflow data. Empirical regression equation between rainfall interception loss and rainfall were adopted as a model to estimate rainfall interception loss from Arabica Coffee plantation, which the coefficient of correlation, r is 0.98. In water yield analysis, this formula was applied and founded that Arabica Coffee plants intercept 76% of annual rainfall or it leaved over annual net rainfall 24% of annual rainfall. Using this net rainfall, water yield produced from Paya Bener River which is the catchment area covered by Arabica Coffee plantation was analyzed in a planning of water supply project for water needs domestic of 3 sub-districts in Central Aceh Regency. Based on increasing population until year of 2025, the results showed that the water yield will be not enough from year of 2015. However, if the catchment area is covered by forest, the water yield is still enough until year of 2025
An Econometric Analysis of Modulated Realised Covariance, Regression and Correlation in Noisy Diffusion Models

DEFF Research Database (Denmark)

Kinnebrock, Silja; Podolskij, Mark

This paper introduces a new estimator to measure the ex-post covariation between high-frequency financial time series under market microstructure noise. We provide an asymptotic limit theory (including feasible central limit theorems) for standard methods such as regression, correlation analysis...... process can be relaxed and how our method can be applied to non-synchronous observations. We also present an empirical study of how high-frequency correlations, regressions and covariances change through time....
Identifying the most promising genotypes in lentil for cultivation in a wide range of environments of Pakistan using various yield stability measures

International Nuclear Information System (INIS)

Ali, A.; Zahid, M.A.

2012-01-01

The present study was aimed to identify the most promising high yielding lentil genotype for a wide range of environments of Pakistan using 8 stability measures. The experiment consisted of 12 lentil genotypes grown at 11 locations falling in different agro-ecological zones of Pakistan for 2 years during 2006/07 and 2007/08 under national uniform yield testing. The General Linear Model (GLM) of MINITAB (version 15) was used for two-way analysis of variance for lentil yield data to examine the total variation into genotypes, environments and genotype x environment interaction. The percent variation of 2 major contributors, environment and GxE interaction, was permissible to perform stability analysis to evaluate stable genotypes across the environments. The genotype x environment interaction means were used for eight stability measures (genotype mean, genotype variance, coefficient of variation, ecovalence, interaction variance, regression slope, deviation mean square, coefficient of determination). The stability measures depicted that the genotype NARC-06-1 with high mean yield (1140 kg/ha -1/), regression slope (1.09) close to unity and less statistics of remaining stability measures except high value of R/sup 2/ for yield proved to be the best within the pool of studied genotypes. The results clearly suggest that the genotype NARC-06-1 may prove to be a widely adapted high yielding stable variety for a broad spectrum of environments of Pakistan. (author)
A menu-driven software package of Bayesian nonparametric (and parametric) mixed models for regression analysis and density estimation.

Science.gov (United States)

Karabatsos, George

2017-02-01

Most of applied statistics involves regression analysis of data. In practice, it is important to specify a regression model that has minimal assumptions which are not violated by data, to ensure that statistical inferences from the model are informative and not misleading. This paper presents a stand-alone and menu-driven software package, Bayesian Regression: Nonparametric and Parametric Models, constructed from MATLAB Compiler. Currently, this package gives the user a choice from 83 Bayesian models for data analysis. They include 47 Bayesian nonparametric (BNP) infinite-mixture regression models; 5 BNP infinite-mixture models for density estimation; and 31 normal random effects models (HLMs), including normal linear models. Each of the 78 regression models handles either a continuous, binary, or ordinal dependent variable, and can handle multi-level (grouped) data. All 83 Bayesian models can handle the analysis of weighted observations (e.g., for meta-analysis), and the analysis of left-censored, right-censored, and/or interval-censored data. Each BNP infinite-mixture model has a mixture distribution assigned one of various BNP prior distributions, including priors defined by either the Dirichlet process, Pitman-Yor process (including the normalized stable process), beta (two-parameter) process, normalized inverse-Gaussian process, geometric weights prior, dependent Dirichlet process, or the dependent infinite-probits prior. The software user can mouse-click to select a Bayesian model and perform data analysis via Markov chain Monte Carlo (MCMC) sampling. After the sampling completes, the software automatically opens text output that reports MCMC-based estimates of the model's posterior distribution and model predictive fit to the data. Additional text and/or graphical output can be generated by mouse-clicking other menu options. This includes output of MCMC convergence analyses, and estimates of the model's posterior predictive distribution, for selected
Regression analysis of mixed recurrent-event and panel-count data.

Science.gov (United States)

Zhu, Liang; Tong, Xinwei; Sun, Jianguo; Chen, Manhua; Srivastava, Deo Kumar; Leisenring, Wendy; Robison, Leslie L

2014-07-01

In event history studies concerning recurrent events, two types of data have been extensively discussed. One is recurrent-event data (Cook and Lawless, 2007. The Analysis of Recurrent Event Data. New York: Springer), and the other is panel-count data (Zhao and others, 2010. Nonparametric inference based on panel-count data. Test 20: , 1-42). In the former case, all study subjects are monitored continuously; thus, complete information is available for the underlying recurrent-event processes of interest. In the latter case, study subjects are monitored periodically; thus, only incomplete information is available for the processes of interest. In reality, however, a third type of data could occur in which some study subjects are monitored continuously, but others are monitored periodically. When this occurs, we have mixed recurrent-event and panel-count data. This paper discusses regression analysis of such mixed data and presents two estimation procedures for the problem. One is a maximum likelihood estimation procedure, and the other is an estimating equation procedure. The asymptotic properties of both resulting estimators of regression parameters are established. Also, the methods are applied to a set of mixed recurrent-event and panel-count data that arose from a Childhood Cancer Survivor Study and motivated this investigation. © The Author 2014. Published by Oxford University Press. All rights reserved. For permissions, please e-mail: journals.permissions@oup.com.
Search for the Higgs Boson in the H{yields} ZZ{sup (*)}{yields}4{mu} Channel in CMS Using a Multivariate Analysis; Busqueda del Boson de Higgs en el Canal H{yields} ZZ{sup (*)}{yields}4{mu} en CMS Empleando un Metodo de Analisis Multivariado

Energy Technology Data Exchange (ETDEWEB)

Alonso Diaz, A.

2007-12-28

This note presents a Higgs boson search analysis in the CMS detector of the LHC accelerator (CERN, Geneva, Switzerland) in the H{yields} ZZ{sup (*)}{yields}4{mu} channel, using a multivariate method. This analysis, based in a Higgs boson mass dependent likelihood, constructed from discriminant variables, provides a significant improvement of the Higgs boson discovery potential in a wide mass range with respect to the official analysis published by CMS, based in orthogonal cuts independent of the Higgs boson mass. (Author) 8 refs.
[Application of negative binomial regression and modified Poisson regression in the research of risk factors for injury frequency].

Science.gov (United States)

Cao, Qingqing; Wu, Zhenqiang; Sun, Ying; Wang, Tiezhu; Han, Tengwei; Gu, Chaomei; Sun, Yehuan

2011-11-01

To Eexplore the application of negative binomial regression and modified Poisson regression analysis in analyzing the influential factors for injury frequency and the risk factors leading to the increase of injury frequency. 2917 primary and secondary school students were selected from Hefei by cluster random sampling method and surveyed by questionnaire. The data on the count event-based injuries used to fitted modified Poisson regression and negative binomial regression model. The risk factors incurring the increase of unintentional injury frequency for juvenile students was explored, so as to probe the efficiency of these two models in studying the influential factors for injury frequency. The Poisson model existed over-dispersion (P Poisson regression and negative binomial regression model, was fitted better. respectively. Both showed that male gender, younger age, father working outside of the hometown, the level of the guardian being above junior high school and smoking might be the results of higher injury frequencies. On a tendency of clustered frequency data on injury event, both the modified Poisson regression analysis and negative binomial regression analysis can be used. However, based on our data, the modified Poisson regression fitted better and this model could give a more accurate interpretation of relevant factors affecting the frequency of injury.
Tomato Yield and Water Use Efficiency - Coupling Effects between Growth Stage Specific Soil Water Deficits

DEFF Research Database (Denmark)

Chen, Si; Zhenjiang, Zhou; Andersen, Mathias Neumann

2015-01-01

To investigate the sensitivity of tomato yield and water use efficiency (WUE) to soil water content at different growth stages, the central composite rotatable design (CCRD) was employed in a five-factor-five-level pot experiment under regulated deficit irrigation. Two regression models concerning...... the effects of stage-specific soil water content on tomato yield and WUE were established. The results showed that the lowest available soil water (ASW) content (around 28%) during vegetative growth stage (here denoted θ1) resulted in high yield and WUE. Moderate (around 69% ASW) during blooming and fruit...... effects of ASW in two growth stages were between θ2 and θ5, θ3. In both cases a moderate θ2 was a precondition for maximum yield response to increasing θ5 and θ3. Sensitivity analysis revealed that yield was most sensitive to soil water content at fruit maturity (θ5). Numerical inspection...
Forecasting Model for IPTV Service in Korea Using Bootstrap Ridge Regression Analysis

Science.gov (United States)

Lee, Byoung Chul; Kee, Seho; Kim, Jae Bum; Kim, Yun Bae

The telecom firms in Korea are taking new step to prepare for the next generation of convergence services, IPTV. In this paper we described our analysis on the effective method for demand forecasting about IPTV broadcasting. We have tried according to 3 types of scenarios based on some aspects of IPTV potential market and made a comparison among the results. The forecasting method used in this paper is the multi generation substitution model with bootstrap ridge regression analysis.
Assessment of genotype x environment interaction on yield and ...

African Journals Online (AJOL)

Days to heading, plant height, number of spikes per square meter, number of kernels per spike, spike weight, 1000 kernel weight and grain yield of the genotypes were evaluated in each location. The regression coefficient (bi) of Finlay and Wilkinson (1963) and mean square of deviation from regression (S2d) of Eberhart ...
An Additive-Multiplicative Cox-Aalen Regression Model

DEFF Research Database (Denmark)

Scheike, Thomas H.; Zhang, Mei-Jie

2002-01-01

Aalen model; additive risk model; counting processes; Cox regression; survival analysis; time-varying effects......Aalen model; additive risk model; counting processes; Cox regression; survival analysis; time-varying effects...
Marital status integration and suicide: A meta-analysis and meta-regression.

Science.gov (United States)

Kyung-Sook, Woo; SangSoo, Shin; Sangjin, Shin; Young-Jeon, Shin

2018-01-01

Marital status is an index of the phenomenon of social integration within social structures and has long been identified as an important predictor suicide. However, previous meta-analyses have focused only on a particular marital status, or not sufficiently explored moderators. A meta-analysis of observational studies was conducted to explore the relationships between marital status and suicide and to understand the important moderating factors in this association. Electronic databases were searched to identify studies conducted between January 1, 2000 and June 30, 2016. We performed a meta-analysis, subgroup analysis, and meta-regression of 170 suicide risk estimates from 36 publications. Using random effects model with adjustment for covariates, the study found that the suicide risk for non-married versus married was OR = 1.92 (95% CI: 1.75-2.12). The suicide risk was higher for non-married individuals aged analysis by gender, non-married men exhibited a greater risk of suicide than their married counterparts in all sub-analyses, but women aged 65 years or older showed no significant association between marital status and suicide. The suicide risk in divorced individuals was higher than for non-married individuals in both men and women. The meta-regression showed that gender, age, and sample size affected between-study variation. The results of the study indicated that non-married individuals have an aggregate higher suicide risk than married ones. In addition, gender and age were confirmed as important moderating factors in the relationship between marital status and suicide. Copyright © 2017 Elsevier Ltd. All rights reserved.
Statistical modelling for precision agriculture: A case study in optimal environmental schedules for Agaricus Bisporus production via variable domain functional regression

Science.gov (United States)

Panayi, Efstathios; Kyriakides, George

2017-01-01

Quantifying the effects of environmental factors over the duration of the growing process on Agaricus Bisporus (button mushroom) yields has been difficult, as common functional data analysis approaches require fixed length functional data. The data available from commercial growers, however, is of variable duration, due to commercial considerations. We employ a recently proposed regression technique termed Variable-Domain Functional Regression in order to be able to accommodate these irregular-length datasets. In this way, we are able to quantify the contribution of covariates such as temperature, humidity and water spraying volumes across the growing process, and for different lengths of growing processes. Our results indicate that optimal oxygen and temperature levels vary across the growing cycle and we propose environmental schedules for these covariates to optimise overall yields. PMID:28961254
Grain yield increase in cereal variety mixtures: A meta-analysis of field trials

DEFF Research Database (Denmark)

Kiær, Lars Pødenphant; Skovgaard, Ib; Østergård, Hanne

2009-01-01

on grain yield. To investigate the prevalence and preconditions for positive mixing effects, reported grain yields of variety mixtures and pure variety stands were obtained from previously published variety trials, converted into relative mixing effects and combined using meta-analysis. Furthermore...... as meeting the criteria for inclusion in the meta-analysis; on the other hand, nearly 200 studies were discarded. The accepted studies reported results on both winter and spring types of each crop species. Relative mixing effects ranged from 30% to 100% with an overall meta-estimate of at least 2.7% (p

Wheat yield vulnerability: relation to rainfall and suggestions for adaptation

Directory of Open Access Journals (Sweden)

Khalid Tafoughalti

2018-04-01

Full Text Available Wheat production is of paramount importance in the region of Meknes, which is mainly produced under rainfed conditions. It is the dominant cereal, the greater proportion being the soft type. During the past few decades, rainfall flaws have caused a number of cases of droughts. These flaws have seriously affecting wheat production. The main objective of this study is the assessment of rainfall variability at monthly, seasonal and annual scales and to determine their impact on wheat yields. To reduce this impact we suggested some mechanisms of adaptation. We used monthly rainfall records for three decades and wheat yields records of fifteen years. Rainfall variability is assessed utilizing the precipitation concentration index and the variation coefficient. The association between wheat yields and cumulative rainfall amounts of different scales was calculated based on a regression model to evaluate the impact of rainfall on wheat yields. Data analysis shown moderate seasonal and irregular annual rainfall distribution. Yields fluctuated from 210 to 4500 Kg/ha with 52% of coefficient of variation. The correlation results shows that soft wheat and hard wheat are strongly correlated with the period of January to March than with the whole growing-season. While they are adversely correlated with the mid-spring. This investigation concluded that synchronizing appropriate adaptation with the period of January to March was crucial to achieving success yield of wheat.
Driven Factors Analysis of China’s Irrigation Water Use Efficiency by Stepwise Regression and Principal Component Analysis

Directory of Open Access Journals (Sweden)

Renfu Jia

2016-01-01

Full Text Available This paper introduces an integrated approach to find out the major factors influencing efficiency of irrigation water use in China. It combines multiple stepwise regression (MSR and principal component analysis (PCA to obtain more realistic results. In real world case studies, classical linear regression model often involves too many explanatory variables and the linear correlation issue among variables cannot be eliminated. Linearly correlated variables will cause the invalidity of the factor analysis results. To overcome this issue and reduce the number of the variables, PCA technique has been used combining with MSR. As such, the irrigation water use status in China was analyzed to find out the five major factors that have significant impacts on irrigation water use efficiency. To illustrate the performance of the proposed approach, the calculation based on real data was conducted and the results were shown in this paper.
Effects of Hydrological Parameters on Palm Oil Fresh Fruit Bunch Yield)

Science.gov (United States)

Nda, M.; Adnan, M. S.; Suhadak, M. A.; Zakaria, M. S.; Lopa, R. T.

2018-04-01

Climate change effects and variability have been studied by many researchers in diverse geophysical fields. Malaysia produces large volume of palm oil, the effects of climate change on hydrological parameters (rainfall and precipitation) could have adverse effects on palm oil fresh fruit bunch (FFB) production with implications at both local and international market. It is important to understand the effects of climate change on crop yield to adopt new cultivation techniques and guaranteeing food security globally. Based on this background, the paper’s objective is to investigate the effects of rainfall and temperature pattern on crop yield (FFB) within five years period (2013 - 2017) at Batu Pahat District. The Man - Kendall rank technique (trend test) and statistical analyses (correlation and regression) were applied to the dataset used for the study. The results reveal that there are variabilities in rainfall and temperature from one month to the other and the statistical analysis reveals that the hydrological parameters have an insignificant effect on crop yield.
A regression analysis of the effect of energy use in agriculture

International Nuclear Information System (INIS)

Karkacier, Osman; Gokalp Goktolga, Z.; Cicek, Adnan

2006-01-01

This study investigates the impacts of energy use on productivity of Turkey's agriculture. It reports the results of a regression analysis of the relationship between energy use and agricultural productivity. The study is based on the analysis of the yearbook data for the period 1971-2003. Agricultural productivity was specified as a function of its energy consumption (TOE) and gross additions of fixed assets during the year. Least square (LS) was employed to estimate equation parameters. The data of this study comes from the State Institute of Statistics (SIS) and The Ministry of Energy of Turkey
Genetic Variation and Association Analysis of the SSR Markers Linked to the Major Drought-Yield QTLs of Rice.

Science.gov (United States)

Tabkhkar, Narjes; Rabiei, Babak; Samizadeh Lahiji, Habibollah; Hosseini Chaleshtori, Maryam

2018-02-24

Drought is one of the major abiotic stresses, which hampers the production of rice worldwide. Informative molecular markers are valuable tools for improving the drought tolerance in various varieties of rice. The present study was conducted to evaluate the informative simple sequence repeat (SSR) markers in a diverse set of rice genotypes. The genetic diversity analyses of the 83 studied rice genotypes were performed using 34 SSR markers closely linked to the major quantitative trait loci (QTLs) of grain yield under drought stress (qDTYs). In general, our results indicated high levels of polymorphism. In addition, we screened these rice genotypes at the reproductive stage under both drought stress and nonstressful conditions. The results of the regression analysis demonstrated a significant relationship between 11 SSR marker alleles and the plant paddy weight under stressful conditions. Under the nonstressful conditions, 16 SSR marker alleles showed a significant correlation with the plant paddy weight. Finally, four markers (RM279, RM231, RM166, and RM231) demonstrated a significant association with the plant paddy weight under both stressful and nonstressful conditions. These informative-associated alleles may be useful for improving the crop yield under both drought stress and nonstressful conditions in breeding programs.
Use of Linear Discriminant Function Analysis in Five Yield Sub ...

African Journals Online (AJOL)

K-means cluster analysis grouped the 134 accessions into four distinct groups. Pairwise Mahalanobis 2 distance (D) among some of the groups was highly significant. From the study the yield sub-characters pod length, pod width, peduncle length and 100-seed weight contributed most to group separation in the cowpea ...
Simulation-based production planning for engineer-to-order systems with random yield

NARCIS (Netherlands)

Akcay, Alp; Martagan, Tugce

2018-01-01

We consider an engineer-to-order production system with unknown yield. We model the yield as a random variable which represents the percentage output obtained from one unit of production quantity. We develop a beta-regression model in which the mean value of the yield depends on the unique
Regional intensity-duration-frequency analysis in the Eastern Black Sea Basin, Turkey, by using L-moments and regression analysis

Science.gov (United States)

Ghiaei, Farhad; Kankal, Murat; Anilan, Tugce; Yuksek, Omer

2018-01-01

The analysis of rainfall frequency is an important step in hydrology and water resources engineering. However, a lack of measuring stations, short duration of statistical periods, and unreliable outliers are among the most important problems when designing hydrology projects. In this study, regional rainfall analysis based on L-moments was used to overcome these problems in the Eastern Black Sea Basin (EBSB) of Turkey. The L-moments technique was applied at all stages of the regional analysis, including determining homogeneous regions, in addition to fitting and estimating parameters from appropriate distribution functions in each homogeneous region. We studied annual maximum rainfall height values of various durations (5 min to 24 h) from seven rain gauge stations located in the EBSB in Turkey, which have gauging periods of 39 to 70 years. Homogeneity of the region was evaluated by using L-moments. The goodness-of-fit criterion for each distribution was defined as the ZDIST statistics, depending on various distributions, including generalized logistic (GLO), generalized extreme value (GEV), generalized normal (GNO), Pearson type 3 (PE3), and generalized Pareto (GPA). GLO and GEV determined the best distributions for short (5 to 30 min) and long (1 to 24 h) period data, respectively. Based on the distribution functions, the governing equations were extracted for calculation of intensities of 2, 5, 25, 50, 100, 250, and 500 years return periods (T). Subsequently, the T values for different rainfall intensities were estimated using data quantifying maximum amount of rainfall at different times. Using these T values, duration, altitude, latitude, and longitude values were used as independent variables in a regression model of the data. The determination coefficient ( R 2) value indicated that the model yields suitable results for the regional relationship of intensity-duration-frequency (IDF), which is necessary for the design of hydraulic structures in small and
Prediction of hearing outcomes by multiple regression analysis in patients with idiopathic sudden sensorineural hearing loss.

Science.gov (United States)

Suzuki, Hideaki; Tabata, Takahisa; Koizumi, Hiroki; Hohchi, Nobusuke; Takeuchi, Shoko; Kitamura, Takuro; Fujino, Yoshihisa; Ohbuchi, Toyoaki

2014-12-01

This study aimed to create a multiple regression model for predicting hearing outcomes of idiopathic sudden sensorineural hearing loss (ISSNHL). The participants were 205 consecutive patients (205 ears) with ISSNHL (hearing level ≥ 40 dB, interval between onset and treatment ≤ 30 days). They received systemic steroid administration combined with intratympanic steroid injection. Data were examined by simple and multiple regression analyses. Three hearing indices (percentage hearing improvement, hearing gain, and posttreatment hearing level [HLpost]) and 7 prognostic factors (age, days from onset to treatment, initial hearing level, initial hearing level at low frequencies, initial hearing level at high frequencies, presence of vertigo, and contralateral hearing level) were included in the multiple regression analysis as dependent and explanatory variables, respectively. In the simple regression analysis, the percentage hearing improvement, hearing gain, and HLpost showed significant correlation with 2, 5, and 6 of the 7 prognostic factors, respectively. The multiple correlation coefficients were 0.396, 0.503, and 0.714 for the percentage hearing improvement, hearing gain, and HLpost, respectively. Predicted values of HLpost calculated by the multiple regression equation were reliable with 70% probability with a 40-dB-width prediction interval. Prediction of HLpost by the multiple regression model may be useful to estimate the hearing prognosis of ISSNHL. © The Author(s) 2014.
Uso de modelos de regressão aleatória para descrever a variação genética da produção de leite na raça Holandesa Random regressions models to describe the genetic variation of milk yield in Holstein breed

Directory of Open Access Journals (Sweden)

Cláudio Vieira de Araújo

2006-06-01

Full Text Available Registros de produção de leite de 68.523 controles leiteiros de 8.536 vacas da raça Holandesa, com parições nos anos de 1996 a 2001, foram utilizados na comparação entre modelos de regressão aleatória para estimação de componentes de variância. Os registros de controle leiteiro foram analisados como características múltiplas, considerando cada controle uma característica distinta. Os mesmos registros de controle leiteiro foram analisados como dados longitudinais, por meio de modelos de regressão aleatória, que diferiram entre si pela função utilizada para descrever a trajetória da curva de lactação dos animais. As funções utilizadas foram a exponencial de Wilmink, a função de Ali e Schaeffer e os polinômios de Legendre de segundo e quarto graus. A comparação entre modelos foi realizada com base nos seguintes critérios: estimativas de componentes de variância, obtidas no modelo multicaractístico e por regressão aleatória; valores da variância residual; e valores do logaritmo da função de verossimilhança. As estimativas de herdabilidade obtidas por meio dos modelos de características múltiplas variaram de 0,110 a 0,244. Para os modelos de regressão aleatória, esses valores oscilaram de 0,127 a 0,301, observando-se as maiores estimativas nos modelos com maior número de parâmetros. Verificou-se que os modelos de regressão aleatória que utilizaram os polinômios de Legendre descreveram melhor a variação genética da produção de leite.Data comprising 68,523 test day milk yield of 8,536 cows of the Holstein breed, calving from 1996 to 2001, were used to compare random regression models, for estimating variance components. Test day records (TD were analyzed as multiple traits, considering each TD as a different trait. The test day records were analyzed as longitudinal traits by different random regression models regarding the function used to describe the trajectory of the lactation curve of the animals
Regression of environmental noise in LIGO data

International Nuclear Information System (INIS)

Tiwari, V; Klimenko, S; Mitselmakher, G; Necula, V; Drago, M; Prodi, G; Frolov, V; Yakushin, I; Re, V; Salemi, F; Vedovato, G

2015-01-01

We address the problem of noise regression in the output of gravitational-wave (GW) interferometers, using data from the physical environmental monitors (PEM). The objective of the regression analysis is to predict environmental noise in the GW channel from the PEM measurements. One of the most promising regression methods is based on the construction of Wiener–Kolmogorov (WK) filters. Using this method, the seismic noise cancellation from the LIGO GW channel has already been performed. In the presented approach the WK method has been extended, incorporating banks of Wiener filters in the time–frequency domain, multi-channel analysis and regulation schemes, which greatly enhance the versatility of the regression analysis. Also we present the first results on regression of the bi-coherent noise in the LIGO data. (paper)
[Multiple linear regression analysis of X-ray measurement and WOMAC scores of knee osteoarthritis].

Science.gov (United States)

Ma, Yu-Feng; Wang, Qing-Fu; Chen, Zhao-Jun; Du, Chun-Lin; Li, Jun-Hai; Huang, Hu; Shi, Zong-Ting; Yin, Yue-Shan; Zhang, Lei; A-Di, Li-Jiang; Dong, Shi-Yu; Wu, Ji

2012-05-01

To perform Multiple Linear Regression analysis of X-ray measurement and WOMAC scores of knee osteoarthritis, and to analyze their relationship with clinical and biomechanical concepts. From March 2011 to July 2011, 140 patients (250 knees) were reviewed, including 132 knees in the left and 118 knees in the right; ranging in age from 40 to 71 years, with an average of 54.68 years. The MB-RULER measurement software was applied to measure femoral angle, tibial angle, femorotibial angle, joint gap angle from antero-posterir and lateral position of X-rays. The WOMAC scores were also collected. Then multiple regression equations was applied for the linear regression analysis of correlation between the X-ray measurement and WOMAC scores. There was statistical significance in the regression equation of AP X-rays value and WOMAC scores (Pregression equation of lateral X-ray value and WOMAC scores (P>0.05). 1) X-ray measurement of knee joint can reflect the WOMAC scores to a certain extent. 2) It is necessary to measure the X-ray mechanical axis of knee, which is important for diagnosis and treatment of osteoarthritis. 3) The correlation between tibial angle,joint gap angle on antero-posterior X-ray and WOMAC scores is significant, which can be used to assess the functional recovery of patients before and after treatment.
Neck-focused panic attacks among Cambodian refugees; a logistic and linear regression analysis.

Science.gov (United States)

Hinton, Devon E; Chhean, Dara; Pich, Vuth; Um, Khin; Fama, Jeanne M; Pollack, Mark H

2006-01-01

Consecutive Cambodian refugees attending a psychiatric clinic were assessed for the presence and severity of current--i.e., at least one episode in the last month--neck-focused panic. Among the whole sample (N=130), in a logistic regression analysis, the Anxiety Sensitivity Index (ASI; odds ratio=3.70) and the Clinician-Administered PTSD Scale (CAPS; odds ratio=2.61) significantly predicted the presence of current neck panic (NP). Among the neck panic patients (N=60), in the linear regression analysis, NP severity was significantly predicted by NP-associated flashbacks (beta=.42), NP-associated catastrophic cognitions (beta=.22), and CAPS score (beta=.28). Further analysis revealed the effect of the CAPS score to be significantly mediated (Sobel test [Baron, R. M., & Kenny, D. A. (1986). The moderator-mediator variable distinction in social psychological research: conceptual, strategic, and statistical considerations. Journal of Personality and Social Psychology, 51, 1173-1182]) by both NP-associated flashbacks and catastrophic cognitions. In the care of traumatized Cambodian refugees, NP severity, as well as NP-associated flashbacks and catastrophic cognitions, should be specifically assessed and treated.
Finding determinants of audit delay by pooled OLS regression analysis

OpenAIRE

Vuko, Tina; Čular, Marko

2014-01-01

The aim of this paper is to investigate determinants of audit delay. Audit delay is measured as the length of time (i.e. the number of calendar days) from the fiscal year-end to the audit report date. It is important to understand factors that influence audit delay since it directly affects the timeliness of financial reporting. The research is conducted on a sample of Croatian listed companies, covering the period of four years (from 2008 to 2011). We use pooled OLS regression analysis, mode...
A Seemingly Unrelated Poisson Regression Model

OpenAIRE

King, Gary

1989-01-01

This article introduces a new estimator for the analysis of two contemporaneously correlated endogenous event count variables. This seemingly unrelated Poisson regression model (SUPREME) estimator combines the efficiencies created by single equation Poisson regression model estimators and insights from "seemingly unrelated" linear regression models.
Quantile regression theory and applications

CERN Document Server

Davino, Cristina; Vistocco, Domenico

2013-01-01

A guide to the implementation and interpretation of Quantile Regression models This book explores the theory and numerous applications of quantile regression, offering empirical data analysis as well as the software tools to implement the methods. The main focus of this book is to provide the reader with a comprehensivedescription of the main issues concerning quantile regression; these include basic modeling, geometrical interpretation, estimation and inference for quantile regression, as well as issues on validity of the model, diagnostic tools. Each methodological aspect is explored and
Spatial analysis of the dairy yield using a conditional autoregressive model / Análise espacial da produção leiteira usando um modelo autoregressivo condicional

Directory of Open Access Journals (Sweden)

João Domingos Scalon

2010-07-01

Full Text Available The dairy yield is one of the most important activities for the Brazilian economy and the use of statistical models may improve the decision making in this productive sector. The aim of this paper was to compare the performance of both the traditional linear regression model and the spatial regression model called conditional autoregressive (CAR to explain how some covariates may contribute for the dairy yield. This work used a database on dairy yield supplied by the Brazilian Institute of Geography and Statistics (IBGE and another database on geographical information of the state of Minas Gerais provided by the Integrated Program of Technological Use of Geographical Information (GEOMINAS. The results showed the superiority of the CAR model over the traditional linear regression model to explain the dairy yield. The CAR model allowed the identification of two different spatial clusters of counties related to the dairy yield in the state of Minas Gerais. The first cluster represents the region where one observes the biggest levels of dairy yield. It is formed by the counties of the Triângulo Mineiro. The second cluster is formed by the northern counties of the state that present the lesser levels of dairy yield. A produção de leite é uma das atividades mais importantes para a economia brasileira e o uso de modelos estatísticos pode auxiliar a tomada de decisão neste setor produtivo. O objetivo deste artigo foi comparar o desempenho do modelo de regressão linear tradicional e do modelo de regressão espacial, denominado de autoregressivo condicional (CAR, para explicar como algumas variáveis preditoras contribuem para a quantidade de leite produzido. Este trabalho usou uma base de dados sobre a produção de leite fornecida pelo Instituto Brasileiro de Geografia e Estatística (IBGE e outra base de dados sobre informações geográficas do estado de Minas Gerais, fornecida pelo Programa Integrado de Uso da Tecnologia de Geoprocessamento
Logistic Regression and Path Analysis Method to Analyze Factors influencing Students’ Achievement

Science.gov (United States)

Noeryanti, N.; Suryowati, K.; Setyawan, Y.; Aulia, R. R.

2018-04-01

Students' academic achievement cannot be separated from the influence of two factors namely internal and external factors. The first factors of the student (internal factors) consist of intelligence (X1), health (X2), interest (X3), and motivation of students (X4). The external factors consist of family environment (X5), school environment (X6), and society environment (X7). The objects of this research are eighth grade students of the school year 2016/2017 at SMPN 1 Jiwan Madiun sampled by using simple random sampling. Primary data are obtained by distributing questionnaires. The method used in this study is binary logistic regression analysis that aims to identify internal and external factors that affect student’s achievement and how the trends of them. Path Analysis was used to determine the factors that influence directly, indirectly or totally on student’s achievement. Based on the results of binary logistic regression, variables that affect student’s achievement are interest and motivation. And based on the results obtained by path analysis, factors that have a direct impact on student’s achievement are students’ interest (59%) and students’ motivation (27%). While the factors that have indirect influences on students’ achievement, are family environment (97%) and school environment (37).
Inheritance of grain yield and its correlation with yield components in ...

African Journals Online (AJOL)

SAM

2014-03-19

Mar 19, 2014 ... 7 × 7 incomplete diallel cross of seven wheat parents during the crop season of 2009 to 2010. Mean square of general ... Genetic background and yield traits of the seven parents. Parent. Pedigree. Released year ..... Correlation and path analysis for yield and yield contributing characters in wheat (Triticum ...
Beyond the mean estimate: a quantile regression analysis of inequalities in educational outcomes using INVALSI survey data

Directory of Open Access Journals (Sweden)

Antonella Costanzo

2017-09-01

Full Text Available Abstract The number of studies addressing issues of inequality in educational outcomes using cognitive achievement tests and variables from large-scale assessment data has increased. Here the value of using a quantile regression approach is compared with a classical regression analysis approach to study the relationships between educational outcomes and likely predictor variables. Italian primary school data from INVALSI large-scale assessments were analyzed using both quantile and standard regression approaches. Mathematics and reading scores were regressed on students' characteristics and geographical variables selected for their theoretical and policy relevance. The results demonstrated that, in Italy, the role of gender and immigrant status varied across the entire conditional distribution of students’ performance. Analogous results emerged pertaining to the difference in students’ performance across Italian geographic areas. These findings suggest that quantile regression analysis is a useful tool to explore the determinants and mechanisms of inequality in educational outcomes. A proper interpretation of quantile estimates may enable teachers to identify effective learning activities and help policymakers to develop tailored programs that increase equity in education.

Weighted functional linear regression models for gene-based association analysis.

Science.gov (United States)

Belonogova, Nadezhda M; Svishcheva, Gulnara R; Wilson, James F; Campbell, Harry; Axenovich, Tatiana I

2018-01-01

Functional linear regression models are effectively used in gene-based association analysis of complex traits. These models combine information about individual genetic variants, taking into account their positions and reducing the influence of noise and/or observation errors. To increase the power of methods, where several differently informative components are combined, weights are introduced to give the advantage to more informative components. Allele-specific weights have been introduced to collapsing and kernel-based approaches to gene-based association analysis. Here we have for the first time introduced weights to functional linear regression models adapted for both independent and family samples. Using data simulated on the basis of GAW17 genotypes and weights defined by allele frequencies via the beta distribution, we demonstrated that type I errors correspond to declared values and that increasing the weights of causal variants allows the power of functional linear models to be increased. We applied the new method to real data on blood pressure from the ORCADES sample. Five of the six known genes with P models. Moreover, we found an association between diastolic blood pressure and the VMP1 gene (P = 8.18×10-6), when we used a weighted functional model. For this gene, the unweighted functional and weighted kernel-based models had P = 0.004 and 0.006, respectively. The new method has been implemented in the program package FREGAT, which is freely available at https://cran.r-project.org/web/packages/FREGAT/index.html.
Data analysis and approximate models model choice, location-scale, analysis of variance, nonparametric regression and image analysis

CERN Document Server

Davies, Patrick Laurie

2014-01-01

Introduction IntroductionApproximate Models Notation Two Modes of Statistical AnalysisTowards One Mode of Analysis Approximation, Randomness, Chaos, Determinism ApproximationA Concept of Approximation Approximation Approximating a Data Set by a Model Approximation Regions Functionals and EquivarianceRegularization and Optimality Metrics and DiscrepanciesStrong and Weak Topologies On Being (almost) Honest Simulations and Tables Degree of Approximation and p-values ScalesStability of Analysis The Choice of En(α, P) Independence Procedures, Approximation and VaguenessDiscrete Models The Empirical Density Metrics and Discrepancies The Total Variation Metric The Kullback-Leibler and Chi-Squared Discrepancies The Po(λ) ModelThe b(k, p) and nb(k, p) Models The Flying Bomb Data The Student Study Times Data OutliersOutliers, Data Analysis and Models Breakdown Points and Equivariance Identifying Outliers and Breakdown Outliers in Multivariate Data Outliers in Linear Regression Outliers in Structured Data The Location...
7755 EFFECT OF NPK FERTILIZER ON FRUIT YIELD AND YIELD ...

African Journals Online (AJOL)

Win7Ent

2013-06-03

Jun 3, 2013 ... peasant farmers in Nigeria. With the increased ... did not significantly (p=0.05) increase the fruit yield nor the seed yield. Key words: NPK fertilizer, Fruit ..... SAS (Statistical Analysis System) Version 9.1. SAS Institute Inc., Cary, ...
Climatically driven yield variability of major crops in Khakassia (South Siberia)

Science.gov (United States)

Babushkina, Elena A.; Belokopytova, Liliana V.; Zhirnova, Dina F.; Shah, Santosh K.; Kostyakova, Tatiana V.

2017-12-01

We investigated the variability of yield of the three main crop cultures in the Khakassia Republic: spring wheat, spring barley, and oats. In terms of yield values, variability characteristics, and climatic response, the agricultural territory of Khakassia can be divided into three zones: (1) the Northern Zone, where crops yield has a high positive response to the amount of precipitation, May-July, and a moderately negative one to the temperatures of the same period; (2) the Central Zone, where crops yield depends mainly on temperatures; and (3) the Southern Zone, where climate has the least expressed impact on yield. The dominant pattern in the crops yield is caused by water stress during periods of high temperatures and low moisture supply with heat stress as additional reason. Differences between zones are due to combinations of temperature latitudinal gradient, precipitation altitudinal gradient, and the presence of a well-developed hydrological network and the irrigational system as moisture sources in the Central Zone. More detailed analysis shows differences in the climatic sensitivity of crops during phases of their vegetative growth and grain development and, to a lesser extent, during harvesting period. Multifactor linear regression models were constructed to estimate climate- and autocorrelation-induced variability of the crops yield. These models allowed prediction of the possibility of yield decreasing by at least 2-11% in the next decade due to increasing of the regional summer temperatures.
Regression analysis of mixed panel count data with dependent terminal events.

Science.gov (United States)

Yu, Guanglei; Zhu, Liang; Li, Yang; Sun, Jianguo; Robison, Leslie L

2017-05-10

Event history studies are commonly conducted in many fields, and a great deal of literature has been established for the analysis of the two types of data commonly arising from these studies: recurrent event data and panel count data. The former arises if all study subjects are followed continuously, while the latter means that each study subject is observed only at discrete time points. In reality, a third type of data, a mixture of the two types of the data earlier, may occur and furthermore, as with the first two types of the data, there may exist a dependent terminal event, which may preclude the occurrences of recurrent events of interest. This paper discusses regression analysis of mixed recurrent event and panel count data in the presence of a terminal event and an estimating equation-based approach is proposed for estimation of regression parameters of interest. In addition, the asymptotic properties of the proposed estimator are established, and a simulation study conducted to assess the finite-sample performance of the proposed method suggests that it works well in practical situations. Finally, the methodology is applied to a childhood cancer study that motivated this study. Copyright © 2017 John Wiley & Sons, Ltd. Copyright © 2017 John Wiley & Sons, Ltd.
Comparative Transcriptome Analysis Reveals Different Silk Yields of Two Silkworm Strains.

Directory of Open Access Journals (Sweden)

Juan Li

Full Text Available Cocoon and silk yields are the most important characteristics of sericulture. However, few studies have examined the genes that modulate these features. Further studies of these genes will be useful for improving the products of sericulture. JingSong (JS and Lan10 (L10 are two strains having significantly different cocoon and silk yields. In the current study, RNA-Seq and quantitative polymerase chain reaction (qPCR were performed on both strains in order to determine divergence of the silk gland, which controls silk biosynthesis in silkworms. Compared with L10, JS had 1375 differentially expressed genes (DEGs; 738 up-regulated genes and 673 down-regulated genes. Nine enriched gene ontology (GO terms were identified by GO enrichment analysis based on these DEGs. KEGG enrichment analysis results showed that the DEGs were enriched in three pathways, which were mainly associated with the processing and biosynthesis of proteins. The representative genes in the enrichment pathways and ten significant DEGs were further verified by qPCR, the results of which were consistent with the RNA-Seq data. Our study has revealed differences in silk glands between the two silkworm strains and provides a perspective for understanding the molecular mechanisms determining silk yield.
Reconstruction of missing daily streamflow data using dynamic regression models

Science.gov (United States)

Tencaliec, Patricia; Favre, Anne-Catherine; Prieur, Clémentine; Mathevet, Thibault

2015-12-01

River discharge is one of the most important quantities in hydrology. It provides fundamental records for water resources management and climate change monitoring. Even very short data-gaps in this information can cause extremely different analysis outputs. Therefore, reconstructing missing data of incomplete data sets is an important step regarding the performance of the environmental models, engineering, and research applications, thus it presents a great challenge. The objective of this paper is to introduce an effective technique for reconstructing missing daily discharge data when one has access to only daily streamflow data. The proposed procedure uses a combination of regression and autoregressive integrated moving average models (ARIMA) called dynamic regression model. This model uses the linear relationship between neighbor and correlated stations and then adjusts the residual term by fitting an ARIMA structure. Application of the model to eight daily streamflow data for the Durance river watershed showed that the model yields reliable estimates for the missing data in the time series. Simulation studies were also conducted to evaluate the performance of the procedure.
CUSUM-Logistic Regression analysis for the rapid detection of errors in clinical laboratory test results.

Science.gov (United States)

Sampson, Maureen L; Gounden, Verena; van Deventer, Hendrik E; Remaley, Alan T

2016-02-01

The main drawback of the periodic analysis of quality control (QC) material is that test performance is not monitored in time periods between QC analyses, potentially leading to the reporting of faulty test results. The objective of this study was to develop a patient based QC procedure for the more timely detection of test errors. Results from a Chem-14 panel measured on the Beckman LX20 analyzer were used to develop the model. Each test result was predicted from the other 13 members of the panel by multiple regression, which resulted in correlation coefficients between the predicted and measured result of >0.7 for 8 of the 14 tests. A logistic regression model, which utilized the measured test result, the predicted test result, the day of the week and time of day, was then developed for predicting test errors. The output of the logistic regression was tallied by a daily CUSUM approach and used to predict test errors, with a fixed specificity of 90%. The mean average run length (ARL) before error detection by CUSUM-Logistic Regression (CSLR) was 20 with a mean sensitivity of 97%, which was considerably shorter than the mean ARL of 53 (sensitivity 87.5%) for a simple prediction model that only used the measured result for error detection. A CUSUM-Logistic Regression analysis of patient laboratory data can be an effective approach for the rapid and sensitive detection of clinical laboratory errors. Published by Elsevier Inc.
Interrelationships between morphometric variables and rounded fish body yields evaluated by path analysis

Directory of Open Access Journals (Sweden)

Rafael Vilhena Reis Neto

2012-07-01

Full Text Available The objective of this study was to verify which morphometric measures and ratios are more directly associated with the weight and body yields of rounded fish. A total of 225 specimens of rounded fish (59 pacus, 61 tambaquis, 52 tambacus and 53 paquis with average weight of 972.43 g (±115.52 g were sampled, stunned, slaughtered, weighed, measured, and processed for morphometric and processing yield analysis. The morphometric measures taken were: standard length (CP; head length (CC; head height (AC; body height (A1; and body width (L1. For completeness, the following morphometric ratios were calculated: CC/CP, AC/CP, A1/CP, L1/CP, CC/A1, AC/A1, L1/A1, CC/AC and L1/CC. The yields of carcass, filet, rib and filet with rib were estimated after processing. Initially, a "stepwise" procedure was performed in order to eliminate multicollinearity problems among the morphometric variables, and the phenotypic correlations were then calculated for the dependent variables (weight and body yields and independent variables (morphometric measurements and ratios. These correlations were later deployed in direct and indirect effects through path analysis, and the direct and indirect contributions of each variable were measured in percentage terms. The CC and A1 measures were important for determining the weight of rounded fish. The CC/A1 ratio was the variable most directly associated with carcass yield. For filet, filet with rib and rib yields, the L1/CC ratio was found to be more appropriate and can be used directly.
Response of Yield and Yield Components of Tef [Eragrostis Tef ...

African Journals Online (AJOL)

The partial budget analysis also indicates that applications of 46 kg. N ha-1 and 10 kg P ha-1 are ..... (1994) indicated that where the grain yield response is negative, yield reduction is primarily caused by a .... An Economic Training. Manual.
Analysis of sparse data in logistic regression in medical research: A newer approach

Directory of Open Access Journals (Sweden)

S Devika

2016-01-01

Full Text Available Background and Objective: In the analysis of dichotomous type response variable, logistic regression is usually used. However, the performance of logistic regression in the presence of sparse data is questionable. In such a situation, a common problem is the presence of high odds ratios (ORs with very wide 95% confidence interval (CI (OR: >999.999, 95% CI: 999.999. In this paper, we addressed this issue by using penalized logistic regression (PLR method. Materials and Methods: Data from case-control study on hyponatremia and hiccups conducted in Christian Medical College, Vellore, Tamil Nadu, India was used. The outcome variable was the presence/absence of hiccups and the main exposure variable was the status of hyponatremia. Simulation dataset was created with different sample sizes and with a different number of covariates. Results: A total of 23 cases and 50 controls were used for the analysis of ordinary and PLR methods. The main exposure variable hyponatremia was present in nine (39.13% of the cases and in four (8.0% of the controls. Of the 23 hiccup cases, all were males and among the controls, 46 (92.0% were males. Thus, the complete separation between gender and the disease group led into an infinite OR with 95% CI (OR: >999.999, 95% CI: 999.999 whereas there was a finite and consistent regression coefficient for gender (OR: 5.35; 95% CI: 0.42, 816.48 using PLR. After adjusting for all the confounding variables, hyponatremia entailed 7.9 (95% CI: 2.06, 38.86 times higher risk for the development of hiccups as was found using PLR whereas there was an overestimation of risk OR: 10.76 (95% CI: 2.17, 53.41 using the conventional method. Simulation experiment shows that the estimated coverage probability of this method is near the nominal level of 95% even for small sample sizes and for a large number of covariates. Conclusions: PLR is almost equal to the ordinary logistic regression when the sample size is large and is superior in small cell
Multiple Linear Regression Analysis of Factors Affecting Real Property Price Index From Case Study Research In Istanbul/Turkey

Science.gov (United States)

Denli, H. H.; Koc, Z.

2015-12-01

Estimation of real properties depending on standards is difficult to apply in time and location. Regression analysis construct mathematical models which describe or explain relationships that may exist between variables. The problem of identifying price differences of properties to obtain a price index can be converted into a regression problem, and standard techniques of regression analysis can be used to estimate the index. Considering regression analysis for real estate valuation, which are presented in real marketing process with its current characteristics and quantifiers, the method will help us to find the effective factors or variables in the formation of the value. In this study, prices of housing for sale in Zeytinburnu, a district in Istanbul, are associated with its characteristics to find a price index, based on information received from a real estate web page. The associated variables used for the analysis are age, size in m2, number of floors having the house, floor number of the estate and number of rooms. The price of the estate represents the dependent variable, whereas the rest are independent variables. Prices from 60 real estates have been used for the analysis. Same price valued locations have been found and plotted on the map and equivalence curves have been drawn identifying the same valued zones as lines.
Correlation Analysis of some Growth, Yield, Yield Components and ...

African Journals Online (AJOL)

three critical growth stages which was imposed by withholding water (at ... November, 5th December, 19th December and 2nd January) laid out in a split ... Simple correlation coefficient ® of different crop parameters and grain yield ... The husk bran and germ are rich sources of ..... heat in 2009/2010 dry season at Fadam a ...
Prevalence of treponema species detected in endodontic infections: systematic review and meta-regression analysis.

Science.gov (United States)

Leite, Fábio R M; Nascimento, Gustavo G; Demarco, Flávio F; Gomes, Brenda P F A; Pucci, Cesar R; Martinho, Frederico C

2015-05-01

This systematic review and meta-regression analysis aimed to calculate a combined prevalence estimate and evaluate the prevalence of different Treponema species in primary and secondary endodontic infections, including symptomatic and asymptomatic cases. The MEDLINE/PubMed, Embase, Scielo, Web of Knowledge, and Scopus databases were searched without starting date restriction up to and including March 2014. Only reports in English were included. The selected literature was reviewed by 2 authors and classified as suitable or not to be included in this review. Lists were compared, and, in case of disagreements, decisions were made after a discussion based on inclusion and exclusion criteria. A pooled prevalence of Treponema species in endodontic infections was estimated. Additionally, a meta-regression analysis was performed. Among the 265 articles identified in the initial search, only 51 were included in the final analysis. The studies were classified into 2 different groups according to the type of endodontic infection and whether it was an exclusively primary/secondary study (n = 36) or a primary/secondary comparison (n = 15). The pooled prevalence of Treponema species was 41.5% (95% confidence interval, 35.9-47.0). In the multivariate model of meta-regression analysis, primary endodontic infections (P apical abscess, symptomatic apical periodontitis (P < .001), and concomitant presence of 2 or more species (P = .028) explained the heterogeneity regarding the prevalence rates of Treponema species. Our findings suggest that Treponema species are important pathogens involved in endodontic infections, particularly in cases of primary and acute infections. Copyright © 2015 American Association of Endodontists. Published by Elsevier Inc. All rights reserved.
Yield evaluation and stability analysis in newly selected `KSA' cotton ...

African Journals Online (AJOL)

Yield evaluation and stability analysis in newly selected `KSA' cotton cultivars in Western Kenya. R M Opondo, G A Ombakho. Abstract. (African Crop Science Journal, 1997 5(2): 119-126). http://dx.doi.org/10.4314/acsj.v5i2.27854 · AJOL African Journals Online. HOW TO USE AJOL... for Researchers · for Librarians ...
Image superresolution using support vector regression.

Science.gov (United States)

Ni, Karl S; Nguyen, Truong Q

2007-06-01

A thorough investigation of the application of support vector regression (SVR) to the superresolution problem is conducted through various frameworks. Prior to the study, the SVR problem is enhanced by finding the optimal kernel. This is done by formulating the kernel learning problem in SVR form as a convex optimization problem, specifically a semi-definite programming (SDP) problem. An additional constraint is added to reduce the SDP to a quadratically constrained quadratic programming (QCQP) problem. After this optimization, investigation of the relevancy of SVR to superresolution proceeds with the possibility of using a single and general support vector regression for all image content, and the results are impressive for small training sets. This idea is improved upon by observing structural properties in the discrete cosine transform (DCT) domain to aid in learning the regression. Further improvement involves a combination of classification and SVR-based techniques, extending works in resolution synthesis. This method, termed kernel resolution synthesis, uses specific regressors for isolated image content to describe the domain through a partitioned look of the vector space, thereby yielding good results.
What Satisfies Students?: Mining Student-Opinion Data with Regression and Decision Tree Analysis

Science.gov (United States)

Thomas, Emily H.; Galambos, Nora

2004-01-01

To investigate how students' characteristics and experiences affect satisfaction, this study uses regression and decision tree analysis with the CHAID algorithm to analyze student-opinion data. A data mining approach identifies the specific aspects of students' university experience that most influence three measures of general satisfaction. The…
Ca analysis: an Excel based program for the analysis of intracellular calcium transients including multiple, simultaneous regression analysis.

Science.gov (United States)

Greensmith, David J

2014-01-01

Here I present an Excel based program for the analysis of intracellular Ca transients recorded using fluorescent indicators. The program can perform all the necessary steps which convert recorded raw voltage changes into meaningful physiological information. The program performs two fundamental processes. (1) It can prepare the raw signal by several methods. (2) It can then be used to analyze the prepared data to provide information such as absolute intracellular Ca levels. Also, the rates of change of Ca can be measured using multiple, simultaneous regression analysis. I demonstrate that this program performs equally well as commercially available software, but has numerous advantages, namely creating a simplified, self-contained analysis workflow. Copyright © 2013 The Author. Published by Elsevier Ireland Ltd.. All rights reserved.
The spatial prediction of landslide susceptibility applying artificial neural network and logistic regression models: A case study of Inje, Korea

Science.gov (United States)

Saro, Lee; Woo, Jeon Seong; Kwan-Young, Oh; Moung-Jin, Lee

2016-02-01

The aim of this study is to predict landslide susceptibility caused using the spatial analysis by the application of a statistical methodology based on the GIS. Logistic regression models along with artificial neutral network were applied and validated to analyze landslide susceptibility in Inje, Korea. Landslide occurrence area in the study were identified based on interpretations of optical remote sensing data (Aerial photographs) followed by field surveys. A spatial database considering forest, geophysical, soil and topographic data, was built on the study area using the Geographical Information System (GIS). These factors were analysed using artificial neural network (ANN) and logistic regression models to generate a landslide susceptibility map. The study validates the landslide susceptibility map by comparing them with landslide occurrence areas. The locations of landslide occurrence were divided randomly into a training set (50%) and a test set (50%). A training set analyse the landslide susceptibility map using the artificial network along with logistic regression models, and a test set was retained to validate the prediction map. The validation results revealed that the artificial neural network model (with an accuracy of 80.10%) was better at predicting landslides than the logistic regression model (with an accuracy of 77.05%). Of the weights used in the artificial neural network model, `slope' yielded the highest weight value (1.330), and `aspect' yielded the lowest value (1.000). This research applied two statistical analysis methods in a GIS and compared their results. Based on the findings, we were able to derive a more effective method for analyzing landslide susceptibility.
The spatial prediction of landslide susceptibility applying artificial neural network and logistic regression models: A case study of Inje, Korea

Directory of Open Access Journals (Sweden)

Saro Lee

2016-02-01

Full Text Available The aim of this study is to predict landslide susceptibility caused using the spatial analysis by the application of a statistical methodology based on the GIS. Logistic regression models along with artificial neutral network were applied and validated to analyze landslide susceptibility in Inje, Korea. Landslide occurrence area in the study were identified based on interpretations of optical remote sensing data (Aerial photographs followed by field surveys. A spatial database considering forest, geophysical, soil and topographic data, was built on the study area using the Geographical Information System (GIS. These factors were analysed using artificial neural network (ANN and logistic regression models to generate a landslide susceptibility map. The study validates the landslide susceptibility map by comparing them with landslide occurrence areas. The locations of landslide occurrence were divided randomly into a training set (50% and a test set (50%. A training set analyse the landslide susceptibility map using the artificial network along with logistic regression models, and a test set was retained to validate the prediction map. The validation results revealed that the artificial neural network model (with an accuracy of 80.10% was better at predicting landslides than the logistic regression model (with an accuracy of 77.05%. Of the weights used in the artificial neural network model, ‘slope’ yielded the highest weight value (1.330, and ‘aspect’ yielded the lowest value (1.000. This research applied two statistical analysis methods in a GIS and compared their results. Based on the findings, we were able to derive a more effective method for analyzing landslide susceptibility.

N-terminal pro-B-type natriuretic peptide measurement is useful in predicting left ventricular hypertrophy regression after aortic valve replacement in patients with severe aortic stenosis.

Science.gov (United States)

Lee, Mirae; Choi, Jin-Oh; Park, Sung-Ji; Kim, Eun Young; Park, PyoWon; Oh, Jae K; Jeon, Eun-Seok

2015-01-01

The predictive factors for early left ventricular hypertrophy (LVH) regression after aortic valve replacement (AVR) have not been fully elucidated. This study was conducted to investigate which preoperative parameters predict early LVH regression after AVR. 87 consecutive patients who underwent AVR due to isolated severe aortic stenosis (AS) were analysed. Patients with ejection fraction regression of LVH at the midterm follow-up was determined. In multivariate analysis, including preoperative echocardiographic parameters, only E/e' ratio was associated with midterm LVH regression (OR 1.11, 95% CI 1.01 to 1.22; p=0.035). When preoperative NT-proBNP was added to the analysis, logNT-proBNP was found to be the single significant predictor of midterm LVH regression (OR 2.00, 95% CI 1.08 to 3.71; p=0.028). By receiver operating characteristic curve analysis, a cut-off value of 440 pg/mL for NT-proBNP yielded a sensitivity of 72% and a specificity of 77% for the prediction of LVH regression after AVR. Preoperative NT-proBNP was an independent predictor for early LVH regression after AVR in patients with isolated severe AS.
Robust Regression and its Application in Financial Data Analysis

OpenAIRE

Mansoor Momeni; Mahmoud Dehghan Nayeri; Ali Faal Ghayoumi; Hoda Ghorbani

2010-01-01

This research is aimed to describe the application of robust regression and its advantages over the least square regression method in analyzing financial data. To do this, relationship between earning per share, book value of equity per share and share price as price model and earning per share, annual change of earning per share and return of stock as return model is discussed using both robust and least square regressions, and finally the outcomes are compared. Comparing the results from th...
6 Grain Yield

African Journals Online (AJOL)

create a favourable environment for rice ... developing lines adaptable to many ... have stable, not too short crop duration with ..... Analysis of variance of the effect of site and season on maturity, grain yield and plant ..... and yield components.
Grain filling parameters and yield components in wheat

OpenAIRE

Brdar Milka; Kobiljski Borislav; Balalić-Kraljević Marija

2006-01-01

Grain yield of wheat (Triticum aestivum L.) is influenced by number of grains per unit area and grain weight, which is result of grain filling duration and rate. The aim of the study was to investigate the relationships between grain filling parameters in 4 wheat genotypes of different earliness and yield components. Nonlinear regression estimated and observed parameters were analyzed. Rang of estimated parameters corresponds to rang of observed parameters. Stepwise MANOVA indicated that the ...
Support vector methods for survival analysis: a comparison between ranking and regression approaches.

Science.gov (United States)

Van Belle, Vanya; Pelckmans, Kristiaan; Van Huffel, Sabine; Suykens, Johan A K

2011-10-01

To compare and evaluate ranking, regression and combined machine learning approaches for the analysis of survival data. The literature describes two approaches based on support vector machines to deal with censored observations. In the first approach the key idea is to rephrase the task as a ranking problem via the concordance index, a problem which can be solved efficiently in a context of structural risk minimization and convex optimization techniques. In a second approach, one uses a regression approach, dealing with censoring by means of inequality constraints. The goal of this paper is then twofold: (i) introducing a new model combining the ranking and regression strategy, which retains the link with existing survival models such as the proportional hazards model via transformation models; and (ii) comparison of the three techniques on 6 clinical and 3 high-dimensional datasets and discussing the relevance of these techniques over classical approaches fur survival data. We compare svm-based survival models based on ranking constraints, based on regression constraints and models based on both ranking and regression constraints. The performance of the models is compared by means of three different measures: (i) the concordance index, measuring the model's discriminating ability; (ii) the logrank test statistic, indicating whether patients with a prognostic index lower than the median prognostic index have a significant different survival than patients with a prognostic index higher than the median; and (iii) the hazard ratio after normalization to restrict the prognostic index between 0 and 1. Our results indicate a significantly better performance for models including regression constraints above models only based on ranking constraints. This work gives empirical evidence that svm-based models using regression constraints perform significantly better than svm-based models based on ranking constraints. Our experiments show a comparable performance for methods
Regression Analysis for Multivariate Dependent Count Data Using Convolved Gaussian Processes

OpenAIRE

Sofro, A'yunin; Shi, Jian Qing; Cao, Chunzheng

2017-01-01

Research on Poisson regression analysis for dependent data has been developed rapidly in the last decade. One of difficult problems in a multivariate case is how to construct a cross-correlation structure and at the meantime make sure that the covariance matrix is positive definite. To address the issue, we propose to use convolved Gaussian process (CGP) in this paper. The approach provides a semi-parametric model and offers a natural framework for modeling common mean structure and covarianc...
Robust best linear estimation for regression analysis using surrogate and instrumental variables.

Science.gov (United States)

Wang, C Y

2012-04-01

We investigate methods for regression analysis when covariates are measured with errors. In a subset of the whole cohort, a surrogate variable is available for the true unobserved exposure variable. The surrogate variable satisfies the classical measurement error model, but it may not have repeated measurements. In addition to the surrogate variables that are available among the subjects in the calibration sample, we assume that there is an instrumental variable (IV) that is available for all study subjects. An IV is correlated with the unobserved true exposure variable and hence can be useful in the estimation of the regression coefficients. We propose a robust best linear estimator that uses all the available data, which is the most efficient among a class of consistent estimators. The proposed estimator is shown to be consistent and asymptotically normal under very weak distributional assumptions. For Poisson or linear regression, the proposed estimator is consistent even if the measurement error from the surrogate or IV is heteroscedastic. Finite-sample performance of the proposed estimator is examined and compared with other estimators via intensive simulation studies. The proposed method and other methods are applied to a bladder cancer case-control study.
Sensitivity of Microstructural Factors Influencing the Impact Toughness of Hypoeutectoid Steels with Ferrite-Pearlite Structure using Multiple Regression Analysis

International Nuclear Information System (INIS)

Lee, Seung-Yong; Lee, Sang-In; Hwang, Byoung-chul

2016-01-01

In this study, the effect of microstructural factors on the impact toughness of hypoeutectoid steels with ferrite-pearlite structure was quantitatively investigated using multiple regression analysis. Microstructural analysis results showed that the pearlite fraction increased with increasing austenitizing temperature and decreasing transformation temperature which substantially decreased the pearlite interlamellar spacing and cementite thickness depending on carbon content. The impact toughness of hypoeutectoid steels usually increased as interlamellar spacing or cementite thickness decreased, although the impact toughness was largely associated with pearlite fraction. Based on these results, multiple regression analysis was performed to understand the individual effect of pearlite fraction, interlamellar spacing, and cementite thickness on the impact toughness. The regression analysis results revealed that pearlite fraction significantly affected impact toughness at room temperature, while cementite thickness did at low temperature.
The Norwegian Healthier Goats program--modeling lactation curves using a multilevel cubic spline regression model.

Science.gov (United States)

Nagel-Alne, G E; Krontveit, R; Bohlin, J; Valle, P S; Skjerve, E; Sølverød, L S

2014-07-01

In 2001, the Norwegian Goat Health Service initiated the Healthier Goats program (HG), with the aim of eradicating caprine arthritis encephalitis, caseous lymphadenitis, and Johne's disease (caprine paratuberculosis) in Norwegian goat herds. The aim of the present study was to explore how control and eradication of the above-mentioned diseases by enrolling in HG affected milk yield by comparison with herds not enrolled in HG. Lactation curves were modeled using a multilevel cubic spline regression model where farm, goat, and lactation were included as random effect parameters. The data material contained 135,446 registrations of daily milk yield from 28,829 lactations in 43 herds. The multilevel cubic spline regression model was applied to 4 categories of data: enrolled early, control early, enrolled late, and control late. For enrolled herds, the early and late notations refer to the situation before and after enrolling in HG; for nonenrolled herds (controls), they refer to development over time, independent of HG. Total milk yield increased in the enrolled herds after eradication: the total milk yields in the fourth lactation were 634.2 and 873.3 kg in enrolled early and enrolled late herds, respectively, and 613.2 and 701.4 kg in the control early and control late herds, respectively. Day of peak yield differed between enrolled and control herds. The day of peak yield came on d 6 of lactation for the control early category for parities 2, 3, and 4, indicating an inability of the goats to further increase their milk yield from the initial level. For enrolled herds, on the other hand, peak yield came between d 49 and 56, indicating a gradual increase in milk yield after kidding. Our results indicate that enrollment in the HG disease eradication program improved the milk yield of dairy goats considerably, and that the multilevel cubic spline regression was a suitable model for exploring effects of disease control and eradication on milk yield. Copyright © 2014
Influence of Previous Crop on Durum Wheat Yield and Yield Stability in a Long-term Experiment

Directory of Open Access Journals (Sweden)

Anna Maria Stellacci

2011-02-01

Full Text Available Long-term experiments are leading indicators of sustainability and serve as an early warning system to detect problems that may compromise future productivity. So the stability of yield is an important parameter to be considered when judging the value of a cropping system relative to others. In a long-term rotation experiment set up in 1972 the influence of different crop sequences on the yields and on yield stability of durum wheat (Triticum durum Desf. was studied. The complete field experiment is a split-split plot in a randomized complete block design with two replications; the whole experiment considers three crop sequences: 1 three-year crop rotation: sugar-beet, wheat + catch crop, wheat; 2 one-year crop rotation: wheat + catch crop; 3 wheat continuous crop; the split treatments are two different crop residue managements; the split-split plot treatments are 18 different fertilization formulas. Each phase of every crop rotation occurred every year. In this paper only one crop residue management and only one fertilization treatment have been analized. Wheat crops in different rotations are coded as follows: F1: wheat after sugar-beet in three-year crop rotation; F2: wheat after wheat in three-year crop rotation; Fc+i: wheat in wheat + catch crop rotation; Fc: continuous wheat. The following two variables were analysed: grain yield and hectolitre weight. Repeated measures analyses of variance and stability analyses have been perfomed for the two variables. The stability analysis was conducted using: three variance methods, namely the coefficient of variability of Francis and Kannenberg, the ecovalence index of Wricke and the stability variance index of Shukla; the regression method of Eberhart and Russell; a method, proposed by Piepho, that computes the probability of one system outperforming another system. It has turned out that each of the stability methods used has enriched of information the simple variance analysis. The Piepho�
Predicting Insolvency : A comparison between discriminant analysis and logistic regression using principal components

OpenAIRE

Geroukis, Asterios; Brorson, Erik

2014-01-01

In this study, we compare the two statistical techniques logistic regression and discriminant analysis to see how well they classify companies based on clusters – made from the solvency ratio – using principal components as independent variables. The principal components are made with different financial ratios. We use cluster analysis to find groups with low, medium and high solvency ratio of 1200 different companies found on the NASDAQ stock market and use this as an apriori definition of ...
A REVIEW ON THE USE OF REGRESSION ANALYSIS IN STUDIES OF AUDIT QUALITY

Directory of Open Access Journals (Sweden)

Agung Dodit Muliawan

2015-07-01

Full Text Available This study aimed to review how regression analysis has been used in studies of abstract phenomenon, such as audit quality, an importance concept in the auditing practice (Schroeder et al., 1986, yet is not well defined. The articles reviewed were the research articles that include audit quality as research variable, either as dependent or independent variables. The articles were purposefully selected to represent balance combination between audit specific and more general accounting journals and between Anglo Saxon and Anglo American journals. The articles were published between 1983-2011 and from the A/A class journal based on ERA 2010’s classifications. The study found that most of the articles reviewed used multiple regression analysis and treated audit quality as dependent variable and measured it by using a proxy. This study also highlights the size of data sample used and the lack of discussions about the assumptions of the statistical analysis used in most of the articles reviewed. This study concluded that the effectiveness and validity of multiple regressions do not only depends on its application by the researchers but also on how the researchers communicate their findings to the audience. KEYWORDS Audit quality, regression analysis ABSTRAK Kajian ini bertujuan untuk mereviu bagaimana analisa regresi digunakan dalam suatu fenomena abstrak seperti kualitas audit, suatu konsep yang penting dalam praktik audit (Schroeder et al., 1986 namun belum terdefinisi dengan jelas. Artikel yang direviu dalam kajian ini adalah artikel penelitian yang memasukkan kualitas audit sebagai variabel penelitian, baik sebagai variabel independen maupun dependen. Artikel-artikel tersebut dipilih dengan cara purposif sampling untuk mendapatkan keterwakilan yang seimbang antara artikel jurnal khusus audit dan akuntansi secara umum, serta mewakili jurnal Anglo Saxon dan Anglo American. Artikel yang direviu diterbitkan pada periode 1983-2011 oleh jurnal yang
Accelerated failure time regression for backward recurrence times and current durations

DEFF Research Database (Denmark)

Keiding, N; Fine, J P; Hansen, O H

2011-01-01

Backward recurrence times in stationary renewal processes and current durations in dynamic populations observed at a cross-section may yield estimates of underlying interarrival times or survival distributions under suitable stationarity assumptions. Regression models have been proposed for these......Backward recurrence times in stationary renewal processes and current durations in dynamic populations observed at a cross-section may yield estimates of underlying interarrival times or survival distributions under suitable stationarity assumptions. Regression models have been proposed...... for these situations, but accelerated failure time models have the particularly attractive feature that they are preserved when going from the backward recurrence times to the underlying survival distribution of interest. This simple fact has recently been noticed in a sociological context and is here illustrated...... by a study of current duration of time to pregnancy...
Logistic regression applied to natural hazards: rare event logistic regression with replications

OpenAIRE

Guns, M.; Vanacker, Veerle

2012-01-01

Statistical analysis of natural hazards needs particular attention, as most of these phenomena are rare events. This study shows that the ordinary rare event logistic regression, as it is now commonly used in geomorphologic studies, does not always lead to a robust detection of controlling factors, as the results can be strongly sample-dependent. In this paper, we introduce some concepts of Monte Carlo simulations in rare event logistic regression. This technique, so-called rare event logisti...
Differentiating regressed melanoma from regressed lichenoid keratosis.

Science.gov (United States)

Chan, Aegean H; Shulman, Kenneth J; Lee, Bonnie A

2017-04-01

Distinguishing regressed lichen planus-like keratosis (LPLK) from regressed melanoma can be difficult on histopathologic examination, potentially resulting in mismanagement of patients. We aimed to identify histopathologic features by which regressed melanoma can be differentiated from regressed LPLK. Twenty actively inflamed LPLK, 12 LPLK with regression and 15 melanomas with regression were compared and evaluated by hematoxylin and eosin staining as well as Melan-A, microphthalmia transcription factor (MiTF) and cytokeratin (AE1/AE3) immunostaining. (1) A total of 40% of regressed melanomas showed complete or near complete loss of melanocytes within the epidermis with Melan-A and MiTF immunostaining, while 8% of regressed LPLK exhibited this finding. (2) Necrotic keratinocytes were seen in the epidermis in 33% regressed melanomas as opposed to all of the regressed LPLK. (3) A dense infiltrate of melanophages in the papillary dermis was seen in 40% of regressed melanomas, a feature not seen in regressed LPLK. In summary, our findings suggest that a complete or near complete loss of melanocytes within the epidermis strongly favors a regressed melanoma over a regressed LPLK. In addition, necrotic epidermal keratinocytes and the presence of a dense band-like distribution of dermal melanophages can be helpful in differentiating these lesions. © 2016 John Wiley & Sons A/S. Published by John Wiley & Sons Ltd.
Economical analysis and relation between energy inputs and yield of greenhouse cucumber production in Iran

Energy Technology Data Exchange (ETDEWEB)

Mohammadi, Ali; Omid, Mahmoud [Department of Agricultural Machinery Engineering, Faculty of Agricultural Engineering and Technology, University of Tehran, Karaj (Iran)

2010-01-15

This paper studies the energy balance between the input and the output per unit area for greenhouse cucumber production. For this purpose, the data on 43 cucumber production greenhouses in the Tehran province, Iran, were collected and analyzed. The results indicated that a total energy input of 148836.76 MJ ha{sup -1} was consumed for cucumber production. Diesel fuel (with 41.94%) and chemical fertilizers (with 19.69%) were amongst the highest energy inputs for cucumber production. The energy productivity was estimated as 0.80 kg MJ{sup -1}. The ratio of energy output to energy input was approximately 0.64. Results indicate 10.93% and 89.07% of total energy input was in renewable and non-renewable forms, respectively. The regression results revealed that the contribution of energy inputs on crop yield (except for fertilizers and seeds energies) was significant. The human labour energy had the highest impact (0.35) among the other inputs in greenhouse cucumber production. Econometric analysis indicated that the total cost of production for one hectare of cucumber production was around 33425.70$. Accordingly, the benefit-cost ratio was estimated as 2.58. (author)
High-throughput quantitative biochemical characterization of algal biomass by NIR spectroscopy; multiple linear regression and multivariate linear regression analysis.

Science.gov (United States)

Laurens, L M L; Wolfrum, E J

2013-12-18

One of the challenges associated with microalgal biomass characterization and the comparison of microalgal strains and conversion processes is the rapid determination of the composition of algae. We have developed and applied a high-throughput screening technology based on near-infrared (NIR) spectroscopy for the rapid and accurate determination of algal biomass composition. We show that NIR spectroscopy can accurately predict the full composition using multivariate linear regression analysis of varying lipid, protein, and carbohydrate content of algal biomass samples from three strains. We also demonstrate a high quality of predictions of an independent validation set. A high-throughput 96-well configuration for spectroscopy gives equally good prediction relative to a ring-cup configuration, and thus, spectra can be obtained from as little as 10-20 mg of material. We found that lipids exhibit a dominant, distinct, and unique fingerprint in the NIR spectrum that allows for the use of single and multiple linear regression of respective wavelengths for the prediction of the biomass lipid content. This is not the case for carbohydrate and protein content, and thus, the use of multivariate statistical modeling approaches remains necessary.
Identification of cotton properties to improve yarn count quality by using regression analysis

International Nuclear Information System (INIS)

Amin, M.; Ullah, M.; Akbar, A.

2014-01-01

Identification of raw material characteristics towards yarn count variation was studied by using statistical techniques. Regression analysis is used to meet the objective. Stepwise regression is used for mode) selection, and coefficient of determination and mean squared error (MSE) criteria are used to identify the contributing factors of cotton properties for yam count. Statistical assumptions of normality, autocorrelation and multicollinearity are evaluated by using probability plot, Durbin Watson test, variance inflation factor (VIF), and then model fitting is carried out. It is found that, invisible (INV), nepness (Nep), grayness (RD), cotton trash (TR) and uniformity index (VI) are the main contributing cotton properties for yarn count variation. The results are also verified by Pareto chart. (author)
A tandem regression-outlier analysis of a ligand cellular system for key structural modifications around ligand binding.

Science.gov (United States)

Lin, Ying-Ting

2013-04-30

A tandem technique of hard equipment is often used for the chemical analysis of a single cell to first isolate and then detect the wanted identities. The first part is the separation of wanted chemicals from the bulk of a cell; the second part is the actual detection of the important identities. To identify the key structural modifications around ligand binding, the present study aims to develop a counterpart of tandem technique for cheminformatics. A statistical regression and its outliers act as a computational technique for separation. A PPARγ (peroxisome proliferator-activated receptor gamma) agonist cellular system was subjected to such an investigation. Results show that this tandem regression-outlier analysis, or the prioritization of the context equations tagged with features of the outliers, is an effective regression technique of cheminformatics to detect key structural modifications, as well as their tendency of impact to ligand binding. The key structural modifications around ligand binding are effectively extracted or characterized out of cellular reactions. This is because molecular binding is the paramount factor in such ligand cellular system and key structural modifications around ligand binding are expected to create outliers. Therefore, such outliers can be captured by this tandem regression-outlier analysis.
The characteristics of high-yield genotype of early-mature mutant lines in barley

International Nuclear Information System (INIS)

Chen Xiulan; Han Yuepeng; He Zhentian; Yang Hefeng

2000-01-01

The correlation and genetic parameters of eight agronomic traits of 36 early mature mutant lines induced from barley Sunong 9052 were studied by stepwise regression and path analysis. The results showed that: (1) the growing period of early mutants was shortened 2-13 days from that of their parent and the trait of yield had a great mutation range; (2) the number of grain per panicle significantly correlated with the days from sowing to heading; (3) according to direct path coefficients, the main characters related with individual plant-yield were in order of productive panicle per plant > 1000-grain-weight > number of grain per panicle > fertility, the high-yield genotype had more productive panicle and higher 10000-grain-weight, and to increase the yield in the breeding of early mature mutation was to select the lines with more tillers and productive panicles, higher 1000-grain-weight and lower number of grain per panicle; (4) the higher broad-sense heritability and genetic variation coefficient were found in 1000-grain-weight and the days from sowing to heading

ESTIMATION OF PEA GRAIN YIELD STABILITY (Pisum sativum L.

Directory of Open Access Journals (Sweden)

Tihomir Čupić

2003-06-01

Full Text Available The paper aimed to determine yield and estimate pea grain yield stability of newly-created lines JSG-1 (cultivar in recognition process as well as compare with foreign origin cultivars in agroecological area of east Slavonia. The trial was set up by a randomized block design on the experimental field of Agricultural Institute Osijek in four replicates in the five-year period (1998 – 2002. Six (five foreign and one inland cultivars were included by the trial: Eiffil, Erbi, JP-5, JSG-1 (in a recognition process, Torsz and Baccara. Stability parameters were calculated by the grouping method after Francis and Kannenberg (1978 and by the model of individual stability estimation after Eberhart and Russel method (1966. According to Francis and Kannenberg, cultivars Eiffil, Erbi, JSG-1 and Baccara belonged to group I known for high yield and low trait varying coefficient, thus, represent stabile yield cultivars. According to regression coefficient and regression deviation variance the most stabile cultivar appeared to be cultivar JSG-1 (bi =1.06 and S2 di=0.010 and the lowest one was Torsz (bi =0.67 and S2 di =0.160. Cultivar Baccara (bi = 1.22 and S2 di =0.034 was comprised by the group of unstabile and adaptible for high-yielding environments.
Multiple predictor smoothing methods for sensitivity analysis.

Energy Technology Data Exchange (ETDEWEB)

Helton, Jon Craig; Storlie, Curtis B.

2006-08-01

The use of multiple predictor smoothing methods in sampling-based sensitivity analyses of complex models is investigated. Specifically, sensitivity analysis procedures based on smoothing methods employing the stepwise application of the following nonparametric regression techniques are described: (1) locally weighted regression (LOESS), (2) additive models, (3) projection pursuit regression, and (4) recursive partitioning regression. The indicated procedures are illustrated with both simple test problems and results from a performance assessment for a radioactive waste disposal facility (i.e., the Waste Isolation Pilot Plant). As shown by the example illustrations, the use of smoothing procedures based on nonparametric regression techniques can yield more informative sensitivity analysis results than can be obtained with more traditional sensitivity analysis procedures based on linear regression, rank regression or quadratic regression when nonlinear relationships between model inputs and model predictions are present.
Multiple predictor smoothing methods for sensitivity analysis

International Nuclear Information System (INIS)

Helton, Jon Craig; Storlie, Curtis B.

2006-01-01

The use of multiple predictor smoothing methods in sampling-based sensitivity analyses of complex models is investigated. Specifically, sensitivity analysis procedures based on smoothing methods employing the stepwise application of the following nonparametric regression techniques are described: (1) locally weighted regression (LOESS), (2) additive models, (3) projection pursuit regression, and (4) recursive partitioning regression. The indicated procedures are illustrated with both simple test problems and results from a performance assessment for a radioactive waste disposal facility (i.e., the Waste Isolation Pilot Plant). As shown by the example illustrations, the use of smoothing procedures based on nonparametric regression techniques can yield more informative sensitivity analysis results than can be obtained with more traditional sensitivity analysis procedures based on linear regression, rank regression or quadratic regression when nonlinear relationships between model inputs and model predictions are present
Estimating milk yield and value losses from increased somatic cell count on US dairy farms.

Science.gov (United States)

Hadrich, J C; Wolf, C A; Lombard, J; Dolak, T M

2018-04-01

Milk loss due to increased somatic cell counts (SCC) results in economic losses for dairy producers. This research uses 10 mo of consecutive dairy herd improvement data from 2013 and 2014 to estimate milk yield loss using SCC as a proxy for clinical and subclinical mastitis. A fixed effects regression was used to examine factors that affected milk yield while controlling for herd-level management. Breed, milking frequency, days in milk, seasonality, SCC, cumulative months with SCC greater than 100,000 cells/mL, lactation, and herd size were variables included in the regression analysis. The cumulative months with SCC above a threshold was included as a proxy for chronic mastitis. Milk yield loss increased as the number of test days with SCC ≥100,000 cells/mL increased. Results from the regression were used to estimate a monetary value of milk loss related to SCC as a function of cow and operation related explanatory variables for a representative dairy cow. The largest losses occurred from increased cumulative test days with a SCC ≥100,000 cells/mL, with daily losses of $1.20/cow per day in the first month to $2.06/cow per day in mo 10. Results demonstrate the importance of including the duration of months above a threshold SCC when estimating milk yield losses. Cows with chronic mastitis, measured by increased consecutive test days with SCC ≥100,000 cells/mL, resulted in higher milk losses than cows with a new infection. This provides farm managers with a method to evaluate the trade-off between treatment and culling decisions as it relates to mastitis control and early detection. Copyright © 2018 American Dairy Science Association. Published by Elsevier Inc. All rights reserved.
Regression analysis: An evaluation of the inuences behindthe pricing of beer

OpenAIRE

Eriksson, Sara; Häggmark, Jonas

2017-01-01

This bachelor thesis in applied mathematics is an analysis of which factors affect the pricing of beer at the Swedish market. A multiple linear regression model is created with the statistical programming language R through a study of the influences for several explanatory variables. For example these variables include country of origin, beer style, volume sold and a Bayesian weighted mean rating from RateBeer, a popular website for beer enthusiasts. The main goal of the project is to find si...
Few crystal balls are crystal clear : eyeballing regression

International Nuclear Information System (INIS)

Wittebrood, R.T.

1998-01-01

The theory of regression and statistical analysis as it applies to reservoir analysis was discussed. It was argued that regression lines are not always the final truth. It was suggested that regression lines and eyeballed lines are often equally accurate. The many conditions that must be fulfilled to calculate a proper regression were discussed. Mentioned among these conditions were the distribution of the data, hidden variables, knowledge of how the data was obtained, the need for causal correlation of the variables, and knowledge of the manner in which the regression results are going to be used. 1 tab., 13 figs
Statistical Analysis of Large Simulated Yield Datasets for Studying Climate Effects

Science.gov (United States)

Makowski, David; Asseng, Senthold; Ewert, Frank; Bassu, Simona; Durand, Jean-Louis; Martre, Pierre; Adam, Myriam; Aggarwal, Pramod K.; Angulo, Carlos; Baron, Chritian;

2015-01-01

process-based crop models is a rather new idea. We demonstrate herewith that statistical methods can play an important role in analyzing simulated yield data sets obtained from the ensembles of process-based crop models. Formal statistical analysis is helpful to estimate the effects of different climatic variables on yield, and to describe the between-model variability of these effects.

An adapted yield criterion for the evolution of subsequent yield surfaces

Science.gov (United States)

Küsters, N.; Brosius, A.

2017-09-01

In numerical analysis of sheet metal forming processes, the anisotropic material behaviour is often modelled with isotropic work hardening and an average Lankford coefficient. In contrast, experimental observations show an evolution of the Lankford coefficients, which can be associated with a yield surface change due to kinematic and distortional hardening. Commonly, extensive efforts are carried out to describe these phenomena. In this paper an isotropic material model based on the Yld2000-2d criterion is adapted with an evolving yield exponent in order to change the yield surface shape. The yield exponent is linked to the accumulative plastic strain. This change has the effect of a rotating yield surface normal. As the normal is directly related to the Lankford coefficient, the change can be used to model the evolution of the Lankford coefficient during yielding. The paper will focus on the numerical implementation of the adapted material model for the FE-code LS-Dyna, mpi-version R7.1.2-d. A recently introduced identification scheme [1] is used to obtain the parameters for the evolving yield surface and will be briefly described for the proposed model. The suitability for numerical analysis will be discussed for deep drawing processes in general. Efforts for material characterization and modelling will be compared to other common yield surface descriptions. Besides experimental efforts and achieved accuracy, the potential of flexibility in material models and the risk of ambiguity during identification are of major interest in this paper.
Role of regression analysis and variation of rheological data in calculation of pressure drop for sludge pipelines.

Science.gov (United States)

Farno, E; Coventry, K; Slatter, P; Eshtiaghi, N

2018-06-15

Sludge pumps in wastewater treatment plants are often oversized due to uncertainty in calculation of pressure drop. This issue costs millions of dollars for industry to purchase and operate the oversized pumps. Besides costs, higher electricity consumption is associated with extra CO 2 emission which creates huge environmental impacts. Calculation of pressure drop via current pipe flow theory requires model estimation of flow curve data which depends on regression analysis and also varies with natural variation of rheological data. This study investigates impact of variation of rheological data and regression analysis on variation of pressure drop calculated via current pipe flow theories. Results compare the variation of calculated pressure drop between different models and regression methods and suggest on the suitability of each method. Copyright © 2018 Elsevier Ltd. All rights reserved.
A Comparison of Regression Techniques for Estimation of Above-Ground Winter Wheat Biomass Using Near-Surface Spectroscopy

Directory of Open Access Journals (Sweden)

Jibo Yue

2018-01-01

Full Text Available Above-ground biomass (AGB provides a vital link between solar energy consumption and yield, so its correct estimation is crucial to accurately monitor crop growth and predict yield. In this work, we estimate AGB by using 54 vegetation indexes (e.g., Normalized Difference Vegetation Index, Soil-Adjusted Vegetation Index and eight statistical regression techniques: artificial neural network (ANN, multivariable linear regression (MLR, decision-tree regression (DT, boosted binary regression tree (BBRT, partial least squares regression (PLSR, random forest regression (RF, support vector machine regression (SVM, and principal component regression (PCR, which are used to analyze hyperspectral data acquired by using a field spectrophotometer. The vegetation indexes (VIs determined from the spectra were first used to train regression techniques for modeling and validation to select the best VI input, and then summed with white Gaussian noise to study how remote sensing errors affect the regression techniques. Next, the VIs were divided into groups of different sizes by using various sampling methods for modeling and validation to test the stability of the techniques. Finally, the AGB was estimated by using a leave-one-out cross validation with these powerful techniques. The results of the study demonstrate that, of the eight techniques investigated, PLSR and MLR perform best in terms of stability and are most suitable when high-accuracy and stable estimates are required from relatively few samples. In addition, RF is extremely robust against noise and is best suited to deal with repeated observations involving remote-sensing data (i.e., data affected by atmosphere, clouds, observation times, and/or sensor noise. Finally, the leave-one-out cross-validation method indicates that PLSR provides the highest accuracy (R2 = 0.89, RMSE = 1.20 t/ha, MAE = 0.90 t/ha, NRMSE = 0.07, CV (RMSE = 0.18; thus, PLSR is best suited for works requiring high
COLOR IMAGE RETRIEVAL BASED ON FEATURE FUSION THROUGH MULTIPLE LINEAR REGRESSION ANALYSIS

Directory of Open Access Journals (Sweden)

K. Seetharaman

2015-08-01

Full Text Available This paper proposes a novel technique based on feature fusion using multiple linear regression analysis, and the least-square estimation method is employed to estimate the parameters. The given input query image is segmented into various regions according to the structure of the image. The color and texture features are extracted on each region of the query image, and the features are fused together using the multiple linear regression model. The estimated parameters of the model, which is modeled based on the features, are formed as a vector called a feature vector. The Canberra distance measure is adopted to compare the feature vectors of the query and target images. The F-measure is applied to evaluate the performance of the proposed technique. The obtained results expose that the proposed technique is comparable to the other existing techniques.
Linear regression in astronomy. II

Science.gov (United States)

Feigelson, Eric D.; Babu, Gutti J.

1992-01-01

A wide variety of least-squares linear regression procedures used in observational astronomy, particularly investigations of the cosmic distance scale, are presented and discussed. The classes of linear models considered are (1) unweighted regression lines, with bootstrap and jackknife resampling; (2) regression solutions when measurement error, in one or both variables, dominates the scatter; (3) methods to apply a calibration line to new data; (4) truncated regression models, which apply to flux-limited data sets; and (5) censored regression models, which apply when nondetections are present. For the calibration problem we develop two new procedures: a formula for the intercept offset between two parallel data sets, which propagates slope errors from one regression to the other; and a generalization of the Working-Hotelling confidence bands to nonstandard least-squares lines. They can provide improved error analysis for Faber-Jackson, Tully-Fisher, and similar cosmic distance scale relations.
Phenotypic Correlation Between Yield and Yield components of Read wheat (Triticum Aestivum L) in Drought Simulated Conditions in Kenya

International Nuclear Information System (INIS)

Kimurto, P.K.

2002-01-01

Establishing the presence and magnitude of x watering regimes interaction and stability of yield under drought simulated conditions would allow plant breeders select the drought tolerant wheat genotypes based on their performance at different rainfall patterns in different locations, not on overall mean yield. Development of drought tolerant wheat varieties in Kenya in an easier, cheaper and more efficient way is required most of it's land area is marginal. Four moisture stress regimes which simulated terminal, early, mid and late drought were created under rain shelter by supplying 70, 82, 94, 106 mm of moisture up to seedling stage, tillering, anthesis and grain filling, respectively. control had 118 mm of moisture applied at all stages. Four test genotypes R748, R830, R831 and R833 were tested together with one check variety, Duma. Yields for each genotype in two seasons were analysed using ANOVA and genotype x watering regimes assessed. Yield stability was also analysed using regression analysis. The result showed that genotype x watering regimes interaction was highly significant, suggesting that genotypes responded differently to increases water levels in each season. This indicated that selecting of drought tolerant genotypes for marginal areas under rain shelter should be based on those rainfall regimes. Yield stability across watering regimes varied among genotypes with Duma and R830 being the most stable cultivars, indicating that they only do well in low water levels. Genotypes R748 and R831 were the most unstable among all the test cultivars. R748 was the most responsive to increasing levels, indicating that it can be grown in low and high rainfall areas. The study showed that selection of stable drought tolerant cultivars using mobile rain shelters is possible
PATH ANALYSIS WITH LOGISTIC REGRESSION MODELS : EFFECT ANALYSIS OF FULLY RECURSIVE CAUSAL SYSTEMS OF CATEGORICAL VARIABLES

OpenAIRE

Nobuoki, Eshima; Minoru, Tabata; Geng, Zhi; Department of Medical Information Analysis, Faculty of Medicine, Oita Medical University; Department of Applied Mathematics, Faculty of Engineering, Kobe University; Department of Probability and Statistics, Peking University

2001-01-01

This paper discusses path analysis of categorical variables with logistic regression models. The total, direct and indirect effects in fully recursive causal systems are considered by using model parameters. These effects can be explained in terms of log odds ratios, uncertainty differences, and an inner product of explanatory variables and a response variable. A study on food choice of alligators as a numerical exampleis reanalysed to illustrate the present approach.
Análise de fatores e regressão bissegmentada em estudos de estratificação ambiental e adaptabilidade em milho Factor analysis and bissegmented regression for studies about environmental stratification and maize adaptability

Directory of Open Access Journals (Sweden)

Deoclécio Domingos Garbuglio

2007-02-01

Full Text Available O objetivo deste trabalho foi verificar possíveis divergências entre os resultados obtidos nas avaliações da adaptabilidade de 27 genótipos de milho (Zea mays L., e na estratificação de 22 ambientes no Estado do Paraná, por meio de técnicas baseadas na análise de fatores e regressão bissegmentada. As estratificações ambientais foram feitas por meio do método tradicional e por análise de fatores, aliada ao porcentual da porção simples da interação GxA (PS%. As análises de adaptabilidade foram realizadas por meio de regressão bissegmentada e análise de fatores. Pela análise de regressão bissegmentada, os genótipos estudados apresentaram alta performance produtiva; no entanto, não foi constatado o genótipo considerado como ideal. A adaptabilidade dos genótipos, analisada por meio de plotagens gráficas, apresentou respostas diferenciadas quando comparada à regressão bissegmentada. A análise de fatores mostrou-se eficiente nos processos de estratificação ambiental e adaptabilidade dos genótipos de milho.The objective of this work was to verify possible divergences among results obtained on adaptability evaluations of 27 maize genotypes (Zea mays L., and on stratification of 22 environments on Paraná State, Brazil, through techniques of factor analysis and bissegmented regression. The environmental stratifications were made through the traditional methodology and by factor analysis, allied to the percentage of the simple portion of GxE interaction (PS%. Adaptability analyses were carried out through bissegmented regression and factor analysis. By the analysis of bissegmented regression, studied genotypes had presented high productive performance; however, it was not evidenced the genotype considered as ideal. The adaptability of the genotypes, analyzed through graphs, presented different answers when compared to bissegmented regression. Factor analysis was efficient in the processes of environment stratification and
Lawrence Livermore National Laboratory seismic yield determination for the NPE

Energy Technology Data Exchange (ETDEWEB)

Rohrer, R. [Lawrence Livermore National Lab., CA (United States)

1994-12-31

The Lawrence Livermore National Laboratory recorded seismic signals from the Non-Proliferation experiment at the Nevada Test Site on September 22, 1993, at seismic stations near Mina, Nevada; Kanab Utah; Landers, California; and Elko, Nevada. Yields were calculated from these recorded seismic amplitudes at the stations using statistical amplitude- yield regression curves from earlier nuclear experiments performed near the Non-Proliferation experiment. The weighted seismic yield average using these amplitudes is 1.9 kt with a standard deviation of 19%. The calibrating experiments were nuclear, so this yield is equivalent to a 1.9-kt nuclear experiment.
Determinants of orphan drugs prices in France: a regression analysis.

Science.gov (United States)

Korchagina, Daria; Millier, Aurelie; Vataire, Anne-Lise; Aballea, Samuel; Falissard, Bruno; Toumi, Mondher

2017-04-21

The introduction of the orphan drug legislation led to the increase in the number of available orphan drugs, but the access to them is often limited due to the high price. Social preferences regarding funding orphan drugs as well as the criteria taken into consideration while setting the price remain unclear. The study aimed at identifying the determinant of orphan drug prices in France using a regression analysis. All drugs with a valid orphan designation at the moment of launch for which the price was available in France were included in the analysis. The selection of covariates was based on a literature review and included drug characteristics (Anatomical Therapeutic Chemical (ATC) class, treatment line, age of target population), diseases characteristics (severity, prevalence, availability of alternative therapeutic options), health technology assessment (HTA) details (actual benefit (AB) and improvement in actual benefit (IAB) scores, delay between the HTA and commercialisation), and study characteristics (type of study, comparator, type of endpoint). The main data sources were European public assessment reports, HTA reports, summaries of opinion on orphan designation of the European Medicines Agency, and the French insurance database of drugs and tariffs. A generalized regression model was developed to test the association between the annual treatment cost and selected covariates. A total of 68 drugs were included. The mean annual treatment cost was €96,518. In the univariate analysis, the ATC class (p = 0.01), availability of alternative treatment options (p = 0.02) and the prevalence (p = 0.02) showed a significant correlation with the annual cost. The multivariate analysis demonstrated significant association between the annual cost and availability of alternative treatment options, ATC class, IAB score, type of comparator in the pivotal clinical trial, as well as commercialisation date and delay between the HTA and commercialisation. The
Logistic regression applied to natural hazards: rare event logistic regression with replications

Science.gov (United States)

Guns, M.; Vanacker, V.

2012-06-01

Statistical analysis of natural hazards needs particular attention, as most of these phenomena are rare events. This study shows that the ordinary rare event logistic regression, as it is now commonly used in geomorphologic studies, does not always lead to a robust detection of controlling factors, as the results can be strongly sample-dependent. In this paper, we introduce some concepts of Monte Carlo simulations in rare event logistic regression. This technique, so-called rare event logistic regression with replications, combines the strength of probabilistic and statistical methods, and allows overcoming some of the limitations of previous developments through robust variable selection. This technique was here developed for the analyses of landslide controlling factors, but the concept is widely applicable for statistical analyses of natural hazards.
Stepwise versus Hierarchical Regression: Pros and Cons

Science.gov (United States)

Lewis, Mitzi

2007-01-01

Multiple regression is commonly used in social and behavioral data analysis. In multiple regression contexts, researchers are very often interested in determining the "best" predictors in the analysis. This focus may stem from a need to identify those predictors that are supportive of theory. Alternatively, the researcher may simply be interested…
Bayesian Nonparametric Regression Analysis of Data with Random Effects Covariates from Longitudinal Measurements

KAUST Repository

Ryu, Duchwan

2010-09-28

We consider nonparametric regression analysis in a generalized linear model (GLM) framework for data with covariates that are the subject-specific random effects of longitudinal measurements. The usual assumption that the effects of the longitudinal covariate processes are linear in the GLM may be unrealistic and if this happens it can cast doubt on the inference of observed covariate effects. Allowing the regression functions to be unknown, we propose to apply Bayesian nonparametric methods including cubic smoothing splines or P-splines for the possible nonlinearity and use an additive model in this complex setting. To improve computational efficiency, we propose the use of data-augmentation schemes. The approach allows flexible covariance structures for the random effects and within-subject measurement errors of the longitudinal processes. The posterior model space is explored through a Markov chain Monte Carlo (MCMC) sampler. The proposed methods are illustrated and compared to other approaches, the "naive" approach and the regression calibration, via simulations and by an application that investigates the relationship between obesity in adulthood and childhood growth curves. © 2010, The International Biometric Society.

Detrended fluctuation analysis as a regression framework: Estimating dependence at different scales

Czech Academy of Sciences Publication Activity Database

Krištoufek, Ladislav

2015-01-01

Roč. 91, č. 1 (2015), 022802-1-022802-5 ISSN 1539-3755 R&D Projects: GA ČR(CZ) GP14-11402P Grant - others:GA ČR(CZ) GAP402/11/0948 Program:GA Institutional support: RVO:67985556 Keywords : Detrended cross-correlation analysis * Regression * Scales Subject RIV: AH - Economics Impact factor: 2.288, year: 2014 http://library.utia.cas.cz/separaty/2015/E/kristoufek-0452315.pdf
MULTIPLE LINEAR REGRESSION ANALYSIS FOR PREDICTION OF BOILER LOSSES AND BOILER EFFICIENCY

OpenAIRE

Chayalakshmi C.L

2018-01-01

MULTIPLE LINEAR REGRESSION ANALYSIS FOR PREDICTION OF BOILER LOSSES AND BOILER EFFICIENCY ABSTRACT Calculation of boiler efficiency is essential if its parameters need to be controlled for either maintaining or enhancing its efficiency. But determination of boiler efficiency using conventional method is time consuming and very expensive. Hence, it is not recommended to find boiler efficiency frequently. The work presented in this paper deals with establishing the statistical mo...
Statistical learning method in regression analysis of simulated positron spectral data

International Nuclear Information System (INIS)

Avdic, S. Dz.

2005-01-01

Positron lifetime spectroscopy is a non-destructive tool for detection of radiation induced defects in nuclear reactor materials. This work concerns the applicability of the support vector machines method for the input data compression in the neural network analysis of positron lifetime spectra. It has been demonstrated that the SVM technique can be successfully applied to regression analysis of positron spectra. A substantial data compression of about 50 % and 8 % of the whole training set with two and three spectral components respectively has been achieved including a high accuracy of the spectra approximation. However, some parameters in the SVM approach such as the insensitivity zone e and the penalty parameter C have to be chosen carefully to obtain a good performance. (author)
Use of generalized regression models for the analysis of stress-rupture data

International Nuclear Information System (INIS)

Booker, M.K.

1978-01-01

The design of components for operation in an elevated-temperature environment often requires a detailed consideration of the creep and creep-rupture properties of the construction materials involved. Techniques for the analysis and extrapolation of creep data have been widely discussed. The paper presents a generalized regression approach to the analysis of such data. This approach has been applied to multiple heat data sets for types 304 and 316 austenitic stainless steel, ferritic 2 1 / 4 Cr-1 Mo steel, and the high-nickel austenitic alloy 800H. Analyses of data for single heats of several materials are also presented. All results appear good. The techniques presented represent a simple yet flexible and powerful means for the analysis and extrapolation of creep and creep-rupture data
Regression: The Apple Does Not Fall Far From the Tree.

Science.gov (United States)

Vetter, Thomas R; Schober, Patrick

2018-05-15

Researchers and clinicians are frequently interested in either: (1) assessing whether there is a relationship or association between 2 or more variables and quantifying this association; or (2) determining whether 1 or more variables can predict another variable. The strength of such an association is mainly described by the correlation. However, regression analysis and regression models can be used not only to identify whether there is a significant relationship or association between variables but also to generate estimations of such a predictive relationship between variables. This basic statistical tutorial discusses the fundamental concepts and techniques related to the most common types of regression analysis and modeling, including simple linear regression, multiple regression, logistic regression, ordinal regression, and Poisson regression, as well as the common yet often underrecognized phenomenon of regression toward the mean. The various types of regression analysis are powerful statistical techniques, which when appropriately applied, can allow for the valid interpretation of complex, multifactorial data. Regression analysis and models can assess whether there is a relationship or association between 2 or more observed variables and estimate the strength of this association, as well as determine whether 1 or more variables can predict another variable. Regression is thus being applied more commonly in anesthesia, perioperative, critical care, and pain research. However, it is crucial to note that regression can identify plausible risk factors; it does not prove causation (a definitive cause and effect relationship). The results of a regression analysis instead identify independent (predictor) variable(s) associated with the dependent (outcome) variable. As with other statistical methods, applying regression requires that certain assumptions be met, which can be tested with specific diagnostics.
Regression Analysis

CERN Document Server

Freund, Rudolf J; Sa, Ping

2006-01-01

The book provides complete coverage of the classical methods of statistical analysis. It is designed to give students an understanding of the purpose of statistical analyses, to allow the student to determine, at least to some degree, the correct type of statistical analyses to be performed in a given situation, and have some appreciation of what constitutes good experimental design
Systematic review, meta-analysis, and meta-regression: Successful second-line treatment for Helicobacter pylori.

Science.gov (United States)

Muñoz, Neus; Sánchez-Delgado, Jordi; Baylina, Mireia; Puig, Ignasi; López-Góngora, Sheila; Suarez, David; Calvet, Xavier

2018-06-01

Multiple Helicobacter pylori second-line schedules have been described as potentially useful. It remains unclear, however, which are the best combinations, and which features of second-line treatments are related to better cure rates. The aim of this study was to determine that second-line treatments achieved excellent (>90%) cure rates by performing a systematic review and when possible a meta-analysis. A meta-regression was planned to determine the characteristics of treatments achieving excellent cure rates. A systematic review for studies evaluating second-line Helicobacter pylori treatment was carried out in multiple databases. A formal meta-analysis was performed when an adequate number of comparative studies was found, using RevMan5.3. A meta-regression for evaluating factors predicting cure rates >90% was performed using Stata Statistical Software. The systematic review identified 115 eligible studies, including 203 evaluable treatment arms. The results were extremely heterogeneous, with 61 treatment arms (30%) achieving optimal (>90%) cure rates. The meta-analysis favored quadruple therapies over triple (83.2% vs 76.1%, OR: 0.59:0.38-0.93; P = .02) and 14-day quadruple treatments over 7-day treatments (91.2% vs 81.5%, OR; 95% CI: 0.42:0.24-0.73; P = .002), although the differences were significant only in the per-protocol analysis. The meta-regression did not find any particular characteristics of the studies to be associated with excellent cure rates. Second-line Helicobacter pylori treatments achieving>90% cure rates are extremely heterogeneous. Quadruple therapy and 14-day treatments seem better than triple therapies and 7-day ones. No single characteristic of the treatments was related to excellent cure rates. Future approaches suitable for infectious diseases-thus considering antibiotic resistances-are needed to design rescue treatments that consistently achieve excellent cure rates. © 2018 John Wiley & Sons Ltd.
Multiple regression and beyond an introduction to multiple regression and structural equation modeling

CERN Document Server

Keith, Timothy Z

2014-01-01

Multiple Regression and Beyond offers a conceptually oriented introduction to multiple regression (MR) analysis and structural equation modeling (SEM), along with analyses that flow naturally from those methods. By focusing on the concepts and purposes of MR and related methods, rather than the derivation and calculation of formulae, this book introduces material to students more clearly, and in a less threatening way. In addition to illuminating content necessary for coursework, the accessibility of this approach means students are more likely to be able to conduct research using MR or SEM--and more likely to use the methods wisely. Covers both MR and SEM, while explaining their relevance to one another Also includes path analysis, confirmatory factor analysis, and latent growth modeling Figures and tables throughout provide examples and illustrate key concepts and techniques For additional resources, please visit: http://tzkeith.com/.
Analysis of chlorophyll content and its correlation with yield attributing traits on early varieties of maize (Zea mays L.

Directory of Open Access Journals (Sweden)

Bikal Ghimire

2015-12-01

Full Text Available Chlorophyll has direct roles on photosynthesis and hence closely relates to capacity for photosynthesis, development and yield of crops. With object to explore the roles of chlorophyll content and its relation with other yield attributing traits a field research was conducted using fourteen early genotypes of maize in RCBD design with three replications. Observations were made for Soil Plant Analysis Development (SPAD reading, ear weight, number of kernel row/ear, number of kernel/row, five hundred kernel weight and grain yield/hectare and these traits were analyzed using Analysis of Variance (ANOVA and correlation coefficient analysis. SPAD reading showed a non-significant variation among the genotypes while it revealed significant correlation with no. of kernel/row, grain yield/hectare and highly significant correlation with no. of kernel row/ear and ear weight which are the most yield determinative traits. For the trait grain yield/ha followed by number of kernel row/ear genotype ARUN-1EV has been found comparatively superior to ARUN-2 (standard check. Grain Yield/hectare was highly heritable (>0.6 while no. of kernel / row, SPAD reading, ear weight, number of kernel row/ear were moderately heritable (0.3-0.6. Correlation analysis and ANOVA revealed ARUN-1EV, comparatively superior to ARUN-2 (standard check, had higher SPAD reading than mean SPAD reading with significant correlation with no. of kernel/row, no. of kernel row/ear, ear weight and grain yield/ha which are all yield determinative traits . This showed positive and significant effect of chlorophyll content in grain yield of the maize.
Satellite-based studies of maize yield spatial variations and their causes in China

Science.gov (United States)

Zhao, Y.

2013-12-01

Maize production in China has been expanding significantly in the past two decades, but yield has become relatively stagnant in the past few years, and needs to be improved to meet increasing demand. Multiple studies found that the gap between potential and actual yield of maize is as large as 40% to 60% of yield potential. Although a few major causes of yield gap have been qualitatively identified with surveys, there has not been spatial analysis aimed at quantifying relative importance of specific biophysical and socio-economic causes, information which would be useful for targeting interventions. This study analyzes the causes of yield variation at field and village level in Quzhou county of North China Plain (NCP). We combine remote sensing and crop modeling to estimate yields in 2009-2012, and identify fields that are consistently high or low yielding. To establish the relationship between yield and potential factors, we gather data on those factors through a household survey. We select targeted survey fields such that not only both extremes of yield distribution but also all soil texture categories in the county is covered. Our survey assesses management and biophysical factors as well as social factors such as farmers' access to agronomic knowledge, which is approximated by distance to the closest demonstration plot or 'Science and technology backyard'. Our survey covers 10 townships, 53 villages and 180 fields. Three to ten farmers are surveyed depending on the amount of variation present among sub pixels of each field. According to survey results, we extract the amount of variation within as well as between villages and or soil type. The higher within village or within field variation, the higher importance of management factors. Factors such as soil type and access to knowledge are more represented by between village variation. Through regression and analysis of variance, we gain more quantitative and thorough understanding of causes to yield variation at
Comparison of logistic regression and neural models in predicting the outcome of biopsy in breast cancer from MRI findings

International Nuclear Information System (INIS)

Abdolmaleki, P.; Yarmohammadi, M.; Gity, M.

2004-01-01

Background: We designed an algorithmic model based on regression analysis and a non-algorithmic model based on the Artificial Neural Network. Materials and methods: The ability of these models was compared together in clinical application to differentiate malignant from benign breast tumors in a study group of 161 patient's records. Each patient's record consisted of 6 subjective features extracted from MRI appearance. These findings were enclosed as features extracted for an Artificial Neural Network as well as a logistic regression model to predict biopsy outcome. After both models had been trained perfectly on samples (n=100), the validation samples (n=61) were presented to the trained network as well as the established logistic regression models. Finally, the diagnostic performance of models were compared to the that of the radiologist in terms of sensitivity, specificity and accuracy, using receiver operating characteristic curve analysis. Results: The average out put of the Artificial Neural Network yielded a perfect sensitivity (98%) and high accuracy (90%) similar to that one of an expert radiologist (96% and 92%) while specificity was smaller than that (67%) verses 80%). The output of the logistic regression model using significant features showed improvement in specificity from 60% for the logistic regression model using all features to 93% for the reduced logistic regression model, keeping the accuracy around 90%. Conclusion: Results show that Artificial Neural Network and logistic regression model prove the relationship between extracted morphological features and biopsy results. Using statistically significant variables reduced logistic regression model outperformed of Artificial Neural Network with remarkable specificity while keeping high sensitivity is achieved
Two levels ARIMAX and regression models for forecasting time series data with calendar variation effects

Science.gov (United States)

Suhartono, Lee, Muhammad Hisyam; Prastyo, Dedy Dwi

2015-12-01

The aim of this research is to develop a calendar variation model for forecasting retail sales data with the Eid ul-Fitr effect. The proposed model is based on two methods, namely two levels ARIMAX and regression methods. Two levels ARIMAX and regression models are built by using ARIMAX for the first level and regression for the second level. Monthly men's jeans and women's trousers sales in a retail company for the period January 2002 to September 2009 are used as case study. In general, two levels of calendar variation model yields two models, namely the first model to reconstruct the sales pattern that already occurred, and the second model to forecast the effect of increasing sales due to Eid ul-Fitr that affected sales at the same and the previous months. The results show that the proposed two level calendar variation model based on ARIMAX and regression methods yields better forecast compared to the seasonal ARIMA model and Neural Networks.
Intermediate and advanced topics in multilevel logistic regression analysis.

Science.gov (United States)

Austin, Peter C; Merlo, Juan

2017-09-10

Multilevel data occur frequently in health services, population and public health, and epidemiologic research. In such research, binary outcomes are common. Multilevel logistic regression models allow one to account for the clustering of subjects within clusters of higher-level units when estimating the effect of subject and cluster characteristics on subject outcomes. A search of the PubMed database demonstrated that the use of multilevel or hierarchical regression models is increasing rapidly. However, our impression is that many analysts simply use multilevel regression models to account for the nuisance of within-cluster homogeneity that is induced by clustering. In this article, we describe a suite of analyses that can complement the fitting of multilevel logistic regression models. These ancillary analyses permit analysts to estimate the marginal or population-average effect of covariates measured at the subject and cluster level, in contrast to the within-cluster or cluster-specific effects arising from the original multilevel logistic regression model. We describe the interval odds ratio and the proportion of opposed odds ratios, which are summary measures of effect for cluster-level covariates. We describe the variance partition coefficient and the median odds ratio which are measures of components of variance and heterogeneity in outcomes. These measures allow one to quantify the magnitude of the general contextual effect. We describe an R 2 measure that allows analysts to quantify the proportion of variation explained by different multilevel logistic regression models. We illustrate the application and interpretation of these measures by analyzing mortality in patients hospitalized with a diagnosis of acute myocardial infarction. © 2017 The Authors. Statistics in Medicine published by John Wiley & Sons Ltd. © 2017 The Authors. Statistics in Medicine published by John Wiley & Sons Ltd.
Regression models of reactor diagnostic signals

International Nuclear Information System (INIS)

Vavrin, J.

1989-01-01

The application is described of an autoregression model as the simplest regression model of diagnostic signals in experimental analysis of diagnostic systems, in in-service monitoring of normal and anomalous conditions and their diagnostics. The method of diagnostics is described using a regression type diagnostic data base and regression spectral diagnostics. The diagnostics is described of neutron noise signals from anomalous modes in the experimental fuel assembly of a reactor. (author)
Analysis of Yield and Yield Related Traits Variability of Winter Wheat (Triticum aestivum L. Cv. Izolda and Double Haploid Lines

Directory of Open Access Journals (Sweden)

Kozdój Janusz

2015-12-01

Full Text Available The yield-forming potential of winter wheat is determined by several factors, namely total number of shoots per plant and total number of spikelets per spike. The field experiments were conducted during three vegetation seasons at the Plant Breeding and Acclimatization Institute – National Research Institute (PBAI–NRI, located in Radzików, Poland. The objective of this study was a comparative analysis of the structural yield-forming factor levels, which determine grain yield per spike and per plant of the DH lines and standard Izolda cultivar. Results indicate that several DH lines showed some differences in tested morphological structures of plant, yield factor levels and in grain yield per spike and per plant in comparison to standard Izolda, regardless of the year. Mean grain yield per plant of DH lines was 26.5% lower in comparison to standard Izolda only in the second year of study. It was caused by a reduction of productive tillers number. Structural yield-forming potential of DH lines was used in 38% and 59% and in case of Izolda in 47% and 61% (the second and the third year of experiment, respectively. The mean grain yield per spike of DH lines was 14.8% lower than Izolda cultivar only in third year of experiment and it was caused by about 12% lower number of grains per spike. Structural yield-forming potential of DH spikes was used in 82.4%, 85.4% and 84.9% and in case of Izolda in 83.8%, 87% and 89.5% (the first, the second and the third year of experiment, respectively. The grain yield per winter wheat plant (both DH lines and standard Izolda was significantly correlated with the number of productive tillers per plant (r = 0.80. The grain yield per winter wheat spike (both DH lines and Izolda cultivar was significantly and highly correlated with the number of grains per spike (r = 0.96, number of fertile spikelets per spike (r = 0.87 and the spike length (r = 0.80. Variation of spike and plant structural yield-forming factors
Comparison of Classical Linear Regression and Orthogonal Regression According to the Sum of Squares Perpendicular Distances

OpenAIRE

KELEŞ, Taliha; ALTUN, Murat

2016-01-01

Regression analysis is a statistical technique for investigating and modeling the relationship between variables. The purpose of this study was the trivial presentation of the equation for orthogonal regression (OR) and the comparison of classical linear regression (CLR) and OR techniques with respect to the sum of squared perpendicular distances. For that purpose, the analyses were shown by an example. It was found that the sum of squared perpendicular distances of OR is smaller. Thus, it wa...
Genetic Analysis of Seed Yield Components and its Association with Forage Production in Wild and Cultivated Species of Sainfoin

Directory of Open Access Journals (Sweden)

A. Najafipoor

2017-02-01

Full Text Available Little is known about genetic variation of seed related traits and their association with forage characters in sainfoin. In order to investigate the variation and relationship among seed yield and its components, 93 genotypes from 21 wild and cultivated species of genus Onobrychis were evaluated using a randomized complete block design with four replications at Isfahan University of Technology Research Farm, Isfahan, Iran. Analysis of variance showed that there was significant difference among genotypes, indicating existence of considerable genetic variation in this germplasm. Panicle fertility and panicle length had the most variation in cultivated and the wild genotypes, respectively. Results of correlation analysis showed that seed yield was positively correlated with number of stems per plant and number of seeds per panicle and negatively correlated with panicle length and days to 50% flowering. Seed yield had positive correlation with forage yield in wild species while this correlation was not significant in cultivated one. Cluster analysis classified the genotypes into three groups which separate wild and cultivated species. Based on principal component analysis the first component was related to seed yield and the second one was related to components of forage yield which can be used for selection of high forage and seed yielding genotypes.
Online Statistical Modeling (Regression Analysis) for Independent Responses

Science.gov (United States)

Made Tirta, I.; Anggraeni, Dian; Pandutama, Martinus

2017-06-01

Regression analysis (statistical analmodelling) are among statistical methods which are frequently needed in analyzing quantitative data, especially to model relationship between response and explanatory variables. Nowadays, statistical models have been developed into various directions to model various type and complex relationship of data. Rich varieties of advanced and recent statistical modelling are mostly available on open source software (one of them is R). However, these advanced statistical modelling, are not very friendly to novice R users, since they are based on programming script or command line interface. Our research aims to developed web interface (based on R and shiny), so that most recent and advanced statistical modelling are readily available, accessible and applicable on web. We have previously made interface in the form of e-tutorial for several modern and advanced statistical modelling on R especially for independent responses (including linear models/LM, generalized linier models/GLM, generalized additive model/GAM and generalized additive model for location scale and shape/GAMLSS). In this research we unified them in the form of data analysis, including model using Computer Intensive Statistics (Bootstrap and Markov Chain Monte Carlo/ MCMC). All are readily accessible on our online Virtual Statistics Laboratory. The web (interface) make the statistical modeling becomes easier to apply and easier to compare them in order to find the most appropriate model for the data.
A Simple Linear Regression Method for Quantitative Trait Loci Linkage Analysis With Censored Observations

OpenAIRE

Anderson, Carl A.; McRae, Allan F.; Visscher, Peter M.

2006-01-01

Standard quantitative trait loci (QTL) mapping techniques commonly assume that the trait is both fully observed and normally distributed. When considering survival or age-at-onset traits these assumptions are often incorrect. Methods have been developed to map QTL for survival traits; however, they are both computationally intensive and not available in standard genome analysis software packages. We propose a grouped linear regression method for the analysis of continuous survival data. Using...
Logistic regression applied to natural hazards: rare event logistic regression with replications

Directory of Open Access Journals (Sweden)

M. Guns

2012-06-01

Full Text Available Statistical analysis of natural hazards needs particular attention, as most of these phenomena are rare events. This study shows that the ordinary rare event logistic regression, as it is now commonly used in geomorphologic studies, does not always lead to a robust detection of controlling factors, as the results can be strongly sample-dependent. In this paper, we introduce some concepts of Monte Carlo simulations in rare event logistic regression. This technique, so-called rare event logistic regression with replications, combines the strength of probabilistic and statistical methods, and allows overcoming some of the limitations of previous developments through robust variable selection. This technique was here developed for the analyses of landslide controlling factors, but the concept is widely applicable for statistical analyses of natural hazards.

Dicentric yields induced in rabbit blood lymphocytes after exposure in vitro to X-rays

International Nuclear Information System (INIS)

Inoue, Yoshinori

1995-06-01

For the purpose of biological dosimetry, it is essential to establish the relationship between dicentric yields and absorbed doses. The present experiment was carried out to obtain data for rabbit lymphocytes as a reference for this relationship. As data at low dose level are scanty, rabbit lymphocytes were exposed to various doses, especially below 0.5 Gy, of 150 kVp X-rays and analysed at their first mitotic division for dicentric yields. The yields at high dose level were compared with data reported by other authors. The linear-quadratic equation, which is generally accepted, for the dose-response relationship was obtained by the iteratively reweighted least squares method. However, as the present experiment result showed that the dose-response relationship at low dose-levels was likely to be linear, a dose-response line was calculated by the linear regression analysis. As the result of the chi-square tests, it was found that the dicentric yield was better fitted to the linear model at low doses below 0.5 Gy than the linear quadratic model. (author)
Robust Methods for Moderation Analysis with a Two-Level Regression Model.

Science.gov (United States)

Yang, Miao; Yuan, Ke-Hai

2016-01-01

Moderation analysis has many applications in social sciences. Most widely used estimation methods for moderation analysis assume that errors are normally distributed and homoscedastic. When these assumptions are not met, the results from a classical moderation analysis can be misleading. For more reliable moderation analysis, this article proposes two robust methods with a two-level regression model when the predictors do not contain measurement error. One method is based on maximum likelihood with Student's t distribution and the other is based on M-estimators with Huber-type weights. An algorithm for obtaining the robust estimators is developed. Consistent estimates of standard errors of the robust estimators are provided. The robust approaches are compared against normal-distribution-based maximum likelihood (NML) with respect to power and accuracy of parameter estimates through a simulation study. Results show that the robust approaches outperform NML under various distributional conditions. Application of the robust methods is illustrated through a real data example. An R program is developed and documented to facilitate the application of the robust methods.
Oil and gas pipeline construction cost analysis and developing regression models for cost estimation

Science.gov (United States)

Thaduri, Ravi Kiran

In this study, cost data for 180 pipelines and 136 compressor stations have been analyzed. On the basis of the distribution analysis, regression models have been developed. Material, Labor, ROW and miscellaneous costs make up the total cost of a pipeline construction. The pipelines are analyzed based on different pipeline lengths, diameter, location, pipeline volume and year of completion. In a pipeline construction, labor costs dominate the total costs with a share of about 40%. Multiple non-linear regression models are developed to estimate the component costs of pipelines for various cross-sectional areas, lengths and locations. The Compressor stations are analyzed based on the capacity, year of completion and location. Unlike the pipeline costs, material costs dominate the total costs in the construction of compressor station, with an average share of about 50.6%. Land costs have very little influence on the total costs. Similar regression models are developed to estimate the component costs of compressor station for various capacities and locations.
Salience Assignment for Multiple-Instance Data and Its Application to Crop Yield Prediction

Science.gov (United States)

Wagstaff, Kiri L.; Lane, Terran

2010-01-01

An algorithm was developed to generate crop yield predictions from orbital remote sensing observations, by analyzing thousands of pixels per county and the associated historical crop yield data for those counties. The algorithm determines which pixels contain which crop. Since each known yield value is associated with thousands of individual pixels, this is a multiple instance learning problem. Because individual crop growth is related to the resulting yield, this relationship has been leveraged to identify pixels that are individually related to corn, wheat, cotton, and soybean yield. Those that have the strongest relationship to a given crop s yield values are most likely to contain fields with that crop. Remote sensing time series data (a new observation every 8 days) was examined for each pixel, which contains information for that pixel s growth curve, peak greenness, and other relevant features. An alternating-projection (AP) technique was used to first estimate the "salience" of each pixel, with respect to the given target (crop yield), and then those estimates were used to build a regression model that relates input data (remote sensing observations) to the target. This is achieved by constructing an exemplar for each crop in each county that is a weighted average of all the pixels within the county; the pixels are weighted according to the salience values. The new regression model estimate then informs the next estimate of the salience values. By iterating between these two steps, the algorithm converges to a stable estimate of both the salience of each pixel and the regression model. The salience values indicate which pixels are most relevant to each crop under consideration.
Total Suspended Load and Sediment Yield of Kayan River, Bulungan District, East Kalimantan

Directory of Open Access Journals (Sweden)

Suprapto Dibyosaputro

2016-12-01

Full Text Available This research was carried out the the drainage system of Kayan river, Bulungan District, East Kalimantan. The purpose of the research were to study the physical conditions of the Kayan catchment area, calculate the suspended sediment load, and to define the total sediment yield of Kayan River. Observation method were used in this research both of direct field observation as well as laboratory observation. Data acquired in this study were include of climatic data, geology, geomorphology, soil and land cover data. Besides also rain-fall data, temperature, river discharge and suspended sediment load. The total sediment yield were calculated by mean of mathematical and statistical analysis especially of linier regression analysis. The result of the research show that total the sediment yield of Kayan River with drainage area of 6,329.452 km² is about 236,921.25 m³/km²/year. The interesting result of the statistical analysis was that the existing negative correlation between river discharge and suspended sediment load. It is the effect of the location of discharge and suspended measurement. This condition caused by sea tide effect on river discharge at the apex delta. During high tide water river trend rising up on discharge but not on suspended sediment load. Instead, also existing setting down processes takes places of the suspended sediment load into the river bottom upper stream and the apex.
Yield Gap analysis - Rationale, methods and applications - Introduction to the Special Issue

NARCIS (Netherlands)

Ittersum, van M.K.; Cassman, K.G.

2013-01-01

Yield gap analysis is an increasingly popular concept. It is a powerful method to reveal and understand the biophysical opportunities to meet the projected increase in demand for agricultural products towards 2050, and to support decision making on research, policies, development and investment that
Weibull and lognormal Taguchi analysis using multiple linear regression

International Nuclear Information System (INIS)

Piña-Monarrez, Manuel R.; Ortiz-Yañez, Jesús F.

2015-01-01

The paper provides to reliability practitioners with a method (1) to estimate the robust Weibull family when the Taguchi method (TM) is applied, (2) to estimate the normal operational Weibull family in an accelerated life testing (ALT) analysis to give confidence to the extrapolation and (3) to perform the ANOVA analysis to both the robust and the normal operational Weibull family. On the other hand, because the Weibull distribution neither has the normal additive property nor has a direct relationship with the normal parameters (µ, σ), in this paper, the issues of estimating a Weibull family by using a design of experiment (DOE) are first addressed by using an L_9 (3"4) orthogonal array (OA) in both the TM and in the Weibull proportional hazard model approach (WPHM). Then, by using the Weibull/Gumbel and the lognormal/normal relationships and multiple linear regression, the direct relationships between the Weibull and the lifetime parameters are derived and used to formulate the proposed method. Moreover, since the derived direct relationships always hold, the method is generalized to the lognormal and ALT analysis. Finally, the method’s efficiency is shown through its application to the used OA and to a set of ALT data. - Highlights: • It gives the statistical relations and steps to use the Taguchi Method (TM) to analyze Weibull data. • It gives the steps to determine the unknown Weibull family to both the robust TM setting and the normal ALT level. • It gives a method to determine the expected lifetimes and to perform its ANOVA analysis in TM and ALT analysis. • It gives a method to give confidence to the extrapolation in an ALT analysis by using the Weibull family of the normal level.
The impact of mineral fertilization and atmospheric precipitation on yield of field crops on family farms

Directory of Open Access Journals (Sweden)

Munćan Mihajlo

2016-01-01

Full Text Available The field crop production, as the most important branch of plant production of the Republic of Serbia, in the period 2002-2011, was carried out on an average of over 2.7 million hectares, 82.7% of which took place on the individual farms/family holdings. Hence, the subject of research in this paper covers yields of major field crops realized on family farms in the region of Vojvodina in the period 1972-2011. The main objective of the research is to study the interdependence of utilization of mineral fertilizers and atmospheric precipitation during the vegetation period and realized yields of major field crops on family farms in the observed period. The regression analysis was applied in order to verify dependencies and determine the form of dependence of achieved yields from examined variables. The results showed that the main limiting factors for obtaining high and stable yields of field crops is inadequate use of fertilizers and the lack of precipitation during the vegetation period.
Framing an Nuclear Emergency Plan using Qualitative Regression Analysis

International Nuclear Information System (INIS)

Amy Hamijah Abdul Hamid; Ibrahim, M.Z.A.; Deris, S.R.

2014-01-01

Since the arising on safety maintenance issues due to post-Fukushima disaster, as well as, lack of literatures on disaster scenario investigation and theory development. This study is dealing with the initiation difficulty on the research purpose which is related to content and problem setting of the phenomenon. Therefore, the research design of this study refers to inductive approach which is interpreted and codified qualitatively according to primary findings and written reports. These data need to be classified inductively into thematic analysis as to develop conceptual framework related to several theoretical lenses. Moreover, the framing of the expected framework of the respective emergency plan as the improvised business process models are abundant of unstructured data abstraction and simplification. The structural methods of Qualitative Regression Analysis (QRA) and Work System snapshot applied to form the data into the proposed model conceptualization using rigorous analyses. These methods were helpful in organising and summarizing the snapshot into an ' as-is ' work system that being recommended as ' to-be' w ork system towards business process modelling. We conclude that these methods are useful to develop comprehensive and structured research framework for future enhancement in business process simulation. (author)
Estimation of the yield of poplars in plantations of fast-growing species within current results

Directory of Open Access Journals (Sweden)

Martin Fajman

2009-01-01

Full Text Available Current results are presented of allometric yield estimates of the poplar short rotation coppice. According to a literature review it is obvious that yield estimates, based on measurable quantities of a growing stand, depend not only on the selected tree specie or its clone, but also on the site location. The Jap-105 poplar clone (P. nigra x P. maximowiczii allometric relations were analyzed by regression methods aimed at the creation of the yield estimation methodology at a testing site in Domanínek. Altogether, the twelve polynomial dependences of particular measured quantities approved the high empirical data conformity with the tested regression model (correlation index from 0.9033 to 0.9967. Within the forward stepwise regression, factors were selected, which explain best examined estimates of the total biomass DM; i.e. d.b.h. and stem height. Furthermore, the KESTEMONT’s (1971 model was verified with a satisfying conformity as well. Approving presented yield estimation methods, the presented models will be checked in a large-scale field trial.
Effect of sulfur and iron fertilizers on yield, yield components and ...

African Journals Online (AJOL)

Jane

2011-06-13

Jun 13, 2011 ... per plant. Interaction between water stress and combination of iron and sulfur fertilizers had significant .... Results of analysis of variance (ANOVA) of water stress (W), sulfur (B) and iron (C), and their interaction with gain yield, yield components and ... the soil structure and it increased the usefulness of other.
A new approach to nuclear reactor design optimization using genetic algorithms and regression analysis

International Nuclear Information System (INIS)

Kumar, Akansha; Tsvetkov, Pavel V.

2015-01-01

Highlights: • This paper presents a new method useful for the optimization of complex dynamic systems. • The method uses the strengths of; genetic algorithms (GA), and regression splines. • The method is applied to the design of a gas cooled fast breeder reactor design. • Tools like Java, R, and codes like MCNP, Matlab are used in this research. - Abstract: A module based optimization method using genetic algorithms (GA), and multivariate regression analysis has been developed to optimize a set of parameters in the design of a nuclear reactor. GA simulates natural evolution to perform optimization, and is widely used in recent times by the scientific community. The GA fits a population of random solutions to the optimized solution of a specific problem. In this work, we have developed a genetic algorithm to determine the values for a set of nuclear reactor parameters to design a gas cooled fast breeder reactor core including a basis thermal–hydraulics analysis, and energy transfer. Multivariate regression is implemented using regression splines (RS). Reactor designs are usually complex and a simulation needs a significantly large amount of time to execute, hence the implementation of GA or any other global optimization techniques is not feasible, therefore we present a new method of using RS in conjunction with GA. Due to using RS, we do not necessarily need to run the neutronics simulation for all the inputs generated from the GA module rather, run the simulations for a predefined set of inputs, build a multivariate regression fit to the input and the output parameters, and then use this fit to predict the output parameters for the inputs generated by GA. The reactor parameters are given by the, radius of a fuel pin cell, isotopic enrichment of the fissile material in the fuel, mass flow rate of the coolant, and temperature of the coolant at the core inlet. And, the optimization objectives for the reactor core are, high breeding of U-233 and Pu-239 in
CORRELATION ANALYSIS OF AGRONOMIC CHARACTERS AND GRAIN YIELD OF RICE FOR TIDAL SWAMP AREAS

Directory of Open Access Journals (Sweden)

Aris Hairmansis

2013-05-01

Full Text Available Development of rice varieties for tidal swamp areas is emphasized on the improvement of rice yield potential in specific environment. However, grain yield is a complex trait and highly dependent on the other agronomic characters; while information related to the relationship between agronomic characters and grain yield in the breeding program particularly for tidal swamp areas is very limited. The objective of this study was to investigate relationship between agronomic characters and grain yield of rice as a basis for selection of high yielding rice varieties for tidal swamp areas. Agronomic characters and grain yield of nine advanced rice breeding lines and two rice varieties were evaluated in a series of experiments in tidal swamp areas, Karang Agung Ulu Village, Banyuasin, South Sumatra, for four cropping seasons in dry season (DS 2005, wet season (WS 2005/2006, DS 2006, and DS 2007. Result from path analysis revealed that the following characters had positive direct effect on grain yield, i.e. number of productive tillers per hill (p = 0.356, number of filled grains per panicle (p = 0.544, and spikelet fertility (p = 0.215. Plant height had negative direct effect (p = -0.332 on grain yield, while maturity, number of spikelets per panicle, and 1000-grain weight showed negligible effect on rice grain yield. Present study suggests that indirect selection of high yielding tidal swamp rice can be done by selecting breeding lines which have many product tive tillers, dense filled grains, and high spikelet fertility.
Robust estimation for homoscedastic regression in the secondary analysis of case-control data

KAUST Repository

Wei, Jiawei; Carroll, Raymond J.; Mü ller, Ursula U.; Keilegom, Ingrid Van; Chatterjee, Nilanjan

2012-01-01

Primary analysis of case-control studies focuses on the relationship between disease D and a set of covariates of interest (Y, X). A secondary application of the case-control study, which is often invoked in modern genetic epidemiologic association studies, is to investigate the interrelationship between the covariates themselves. The task is complicated owing to the case-control sampling, where the regression of Y on X is different from what it is in the population. Previous work has assumed a parametric distribution for Y given X and derived semiparametric efficient estimation and inference without any distributional assumptions about X. We take up the issue of estimation of a regression function when Y given X follows a homoscedastic regression model, but otherwise the distribution of Y is unspecified. The semiparametric efficient approaches can be used to construct semiparametric efficient estimates, but they suffer from a lack of robustness to the assumed model for Y given X. We take an entirely different approach. We show how to estimate the regression parameters consistently even if the assumed model for Y given X is incorrect, and thus the estimates are model robust. For this we make the assumption that the disease rate is known or well estimated. The assumption can be dropped when the disease is rare, which is typically so for most case-control studies, and the estimation algorithm simplifies. Simulations and empirical examples are used to illustrate the approach.
Robust estimation for homoscedastic regression in the secondary analysis of case-control data

KAUST Repository

Wei, Jiawei

2012-12-04

Primary analysis of case-control studies focuses on the relationship between disease D and a set of covariates of interest (Y, X). A secondary application of the case-control study, which is often invoked in modern genetic epidemiologic association studies, is to investigate the interrelationship between the covariates themselves. The task is complicated owing to the case-control sampling, where the regression of Y on X is different from what it is in the population. Previous work has assumed a parametric distribution for Y given X and derived semiparametric efficient estimation and inference without any distributional assumptions about X. We take up the issue of estimation of a regression function when Y given X follows a homoscedastic regression model, but otherwise the distribution of Y is unspecified. The semiparametric efficient approaches can be used to construct semiparametric efficient estimates, but they suffer from a lack of robustness to the assumed model for Y given X. We take an entirely different approach. We show how to estimate the regression parameters consistently even if the assumed model for Y given X is incorrect, and thus the estimates are model robust. For this we make the assumption that the disease rate is known or well estimated. The assumption can be dropped when the disease is rare, which is typically so for most case-control studies, and the estimation algorithm simplifies. Simulations and empirical examples are used to illustrate the approach.
Variable and subset selection in PLS regression

DEFF Research Database (Denmark)

Høskuldsson, Agnar

2001-01-01

The purpose of this paper is to present some useful methods for introductory analysis of variables and subsets in relation to PLS regression. We present here methods that are efficient in finding the appropriate variables or subset to use in the PLS regression. The general conclusion...... is that variable selection is important for successful analysis of chemometric data. An important aspect of the results presented is that lack of variable selection can spoil the PLS regression, and that cross-validation measures using a test set can show larger variation, when we use different subsets of X, than...
A study of the irradiation of sawdust culture medium and the yield of fresh lentinus

International Nuclear Information System (INIS)

Xu Meiyu; Wu Jinshui

1989-01-01

60 Co γ-rays at the doses of 1.0-1.8 Mrad were used to irradiate sawdust material for sterilization, It was found that 60 Co γ-rays could promote the breakdown of cellulose into soluble sugar, and thus promoted the growth of hyphae and increased the yield of L.edode. Determination of regressive analysis and correlation coefficients indicated that there was a high positive correlation between soluble sugar and the growth of hyphae(r = 0.9546**) , which can be expressed as Y = 3.12 + 1.30X; A high positive correlation between soluble sugar and the yield of L.edode(r = 0.9935**) can be expressed as Y = 31.95 + 86.69X; and a high positive correlation between the hyphae growing rate and the yield of L.edode(r = 0.9531**) can be expressed as Y = -145.34 + 60.92X
Logistic regression analysis of conventional ultrasonography, strain elastosonography, and contrast-enhanced ultrasound characteristics for the differentiation of benign and malignant thyroid nodules.

Science.gov (United States)

Pang, Tiantian; Huang, Leidan; Deng, Yingyuan; Wang, Tianfu; Chen, Siping; Gong, Xuehao; Liu, Weixiang

2017-01-01

The aim of the study is to screen the significant sonographic features by logistic regression analysis and fit a model to diagnose thyroid nodules. A total of 525 pathological thyroid nodules were retrospectively analyzed. All the nodules underwent conventional ultrasonography (US), strain elastosonography (SE), and contrast -enhanced ultrasound (CEUS). Those nodules' 12 suspicious sonographic features were used to assess thyroid nodules. The significant features of diagnosing thyroid nodules were picked out by logistic regression analysis. All variables that were statistically related to diagnosis of thyroid nodules, at a level of p regression analysis model. The significant features in the logistic regression model of diagnosing thyroid nodules were calcification, suspected cervical lymph node metastasis, hypoenhancement pattern, margin, shape, vascularity, posterior acoustic, echogenicity, and elastography score. According to the results of logistic regression analysis, the formula that could predict whether or not thyroid nodules are malignant was established. The area under the receiver operating curve (ROC) was 0.930 and the sensitivity, specificity, accuracy, positive predictive value, and negative predictive value were 83.77%, 89.56%, 87.05%, 86.04%, and 87.79% respectively.
Interactions between cadmium and decabrominated diphenyl ether on blood cells count in rats—Multiple factorial regression analysis

International Nuclear Information System (INIS)

Curcic, Marijana; Buha, Aleksandra; Stankovic, Sanja; Milovanovic, Vesna; Bulat, Zorica; Đukić-Ćosić, Danijela; Antonijević, Evica; Vučinić, Slavica; Matović, Vesna; Antonijevic, Biljana

2017-01-01

The objective of this study was to assess toxicity of Cd and BDE-209 mixture on haematological parameters in subacutely exposed rats and to determine the presence and type of interactions between these two chemicals using multiple factorial regression analysis. Furthermore, for the assessment of interaction type, an isobologram based methodology was applied and compared with multiple factorial regression analysis. Chemicals were given by oral gavage to the male Wistar rats weighing 200–240 g for 28 days. Animals were divided in 16 groups (8/group): control vehiculum group, three groups of rats were treated with 2.5, 7.5 or 15 mg Cd/kg/day. These doses were chosen on the bases of literature data and reflect relatively high Cd environmental exposure, three groups of rats were treated with 1000, 2000 or 4000 mg BDE-209/kg/bw/day, doses proved to induce toxic effects in rats. Furthermore, nine groups of animals were treated with different mixtures of Cd and BDE-209 containing doses of Cd and BDE-209 stated above. Blood samples were taken at the end of experiment and red blood cells, white blood cells and platelets counts were determined. For interaction assessment multiple factorial regression analysis and fitted isobologram approach were used. In this study, we focused on multiple factorial regression analysis as a method for interaction assessment. We also investigated the interactions between Cd and BDE-209 by the derived model for the description of the obtained fitted isobologram curves. Current study indicated that co-exposure to Cd and BDE-209 can result in significant decrease in RBC count, increase in WBC count and decrease in PLT count, when compared with controls. Multiple factorial regression analysis used for the assessment of interactions type between Cd and BDE-209 indicated synergism for the effect on RBC count and no interactions i.e. additivity for the effects on WBC and PLT counts. On the other hand, isobologram based approach showed slight
Multicollinearity in Regression Analyses Conducted in Epidemiologic Studies.

Science.gov (United States)

Vatcheva, Kristina P; Lee, MinJae; McCormick, Joseph B; Rahbar, Mohammad H

2016-04-01

The adverse impact of ignoring multicollinearity on findings and data interpretation in regression analysis is very well documented in the statistical literature. The failure to identify and report multicollinearity could result in misleading interpretations of the results. A review of epidemiological literature in PubMed from January 2004 to December 2013, illustrated the need for a greater attention to identifying and minimizing the effect of multicollinearity in analysis of data from epidemiologic studies. We used simulated datasets and real life data from the Cameron County Hispanic Cohort to demonstrate the adverse effects of multicollinearity in the regression analysis and encourage researchers to consider the diagnostic for multicollinearity as one of the steps in regression analysis.

The relative importance of imaging markers for the prediction of Alzheimer's disease dementia in mild cognitive impairment — Beyond classical regression

Directory of Open Access Journals (Sweden)

Stefan J. Teipel

2015-01-01

Penalized regression yielded more parsimonious models than unpenalized stepwise regression for the integration of multiregional and multimodal imaging information. The advantage of penalized regression was particularly strong with a high number of collinear predictors.
Study of Yield and Effective Traits in Bread Wheat Recombinant Inbred Lines (Triticum aestivum L. under Water Deficit Condition

Directory of Open Access Journals (Sweden)

S. Mohammad zadeh

2013-11-01

Full Text Available The effects some traits on seed yield of recombinant inbred lines of wheat under water deficit stress was studied. This research was done at the Agricultural Research Stations, Islamic Azad University, Tabriz Branch in 2010- 2011. 28 recombinant inbred lines of wheat bread with two parents (Norstar and Zagros in split plot experiment based on a randomized complete block design with three replications at two irrigation levels (70 and 140 mm evaporation from pan class A were studied. Analysis of variance indicated a significant genetic differences in all traits under study among the lines. Lines No. 32, 163 and 182 produced highest yield under both irrigation levels. Number of spikes, grains per spike and harvest index had the highest positive correlation with grain yield. Path analysis based on stepwise regression showed that under the normal irrigation conditions, number spike (0.556, number of grains per spike (0.278, weight of 1000 grain (0.259 and the drought stress number spike (0.430, straw yield (0.276 and peduncle length (0.323 had the most direct and positive effect on yield respectively.
Testing contingency hypotheses in budgetary research: An evaluation of the use of moderated regression analysis

NARCIS (Netherlands)

Hartmann, Frank G.H.; Moers, Frank

1999-01-01

In the contingency literature on the behavioral and organizational effects of budgeting, use of the Moderated Regression Analysis (MRA) technique is prevalent. This technique is used to test contingency hypotheses that predict interaction effects between budgetary and contextual variables. This
Study on Yield and Yield Components of Wheat Genotypes under Different Moisture Regimes

Directory of Open Access Journals (Sweden)

E. Mogtader

2012-10-01

Full Text Available In order to study grain yield and yield components of 16 advanced wheat lines under rainfed and supplementary irrigation conditions, this research was conducted in randomized block design with 3 replications at Maragheh Research Station during 2008-09 seasons. Analysis of variance revealed significant differences for date to heading, plant height, 1000 kernel weight, tiller number, spike length, seed number per spike, spikelet number per spike, peduncle length, harvest index, leaf, sheath length and grain yield. Results also showed that the lines No. 4 (91-142 a 61/3/F35.70/MO73//1D13.1/MLT and 16 (Azar2 with 1895 and 1878 Kg/ha, lines No. 4 and 7 (YUMAI13/5/NAI60/3/14.53/ODIN//CI13441 with 2132 and 2285 Kg/ha had highest grain yield under rainfed and supplementary irrigated conditions respectively. Based on results these 16 lines and cultivars were grouped in 4 and 3 distinct classes using Ward’s Method of cluster analysis under rainfed and irrigated conditions. Path analysis indicated that vigor at shooting stage, seed number per spike and HI were positive important traits to select lines for high yielding potential in this study. HI and TKW had also positive effects on grain under supplementary irrigation.
Modified Regression Correlation Coefficient for Poisson Regression Model

Science.gov (United States)

Kaengthong, Nattacha; Domthong, Uthumporn

2017-09-01

This study gives attention to indicators in predictive power of the Generalized Linear Model (GLM) which are widely used; however, often having some restrictions. We are interested in regression correlation coefficient for a Poisson regression model. This is a measure of predictive power, and defined by the relationship between the dependent variable (Y) and the expected value of the dependent variable given the independent variables [E(Y|X)] for the Poisson regression model. The dependent variable is distributed as Poisson. The purpose of this research was modifying regression correlation coefficient for Poisson regression model. We also compare the proposed modified regression correlation coefficient with the traditional regression correlation coefficient in the case of two or more independent variables, and having multicollinearity in independent variables. The result shows that the proposed regression correlation coefficient is better than the traditional regression correlation coefficient based on Bias and the Root Mean Square Error (RMSE).
Historical effects of temperature and precipitation on California crop yields

Energy Technology Data Exchange (ETDEWEB)

Lobell, D.B. [Energy and Environment Directorate, Lawrence Livermore National Laboratory, Livermore, CA 94550 (United States); Cahill, K.N. [Interdisciplinary Graduate Program in Environment and Resources, Stanford University, Stanford, CA 94305 (United States); Field, C.B. [Department of Global Ecology, Carnegie Institution, Stanford, CA 94305 (United States)

2007-03-15

For the 1980-2003 period, we analyzed the relationship between crop yield and three climatic variables (minimum temperature, maximum temperature, and precipitation) for 12 major Californian crops: wine grapes, lettuce, almonds, strawberries, table grapes, hay, oranges, cotton, tomatoes, walnuts, avocados, and pistachios. The months and climatic variables of greatest importance to each crop were used to develop regressions relating yield to climatic conditions. For most crops, fairly simple equations using only 2-3 variables explained more than two-thirds of observed yield variance. The types of variables and months identified suggest that relatively poorly understood processes such as crop infection, pollination, and dormancy may be important mechanisms by which climate influences crop yield. Recent climatic trends have had mixed effects on crop yields, with orange and walnut yields aided, avocado yields hurt, and most crops little affected by recent climatic trends. Yield-climate relationships can provide a foundation for forecasting crop production within a year and for projecting the impact of future climate changes.
Clinical benefit from pharmacological elevation of high-density lipoprotein cholesterol: meta-regression analysis.

Science.gov (United States)

Hourcade-Potelleret, F; Laporte, S; Lehnert, V; Delmar, P; Benghozi, Renée; Torriani, U; Koch, R; Mismetti, P

2015-06-01

Epidemiological evidence that the risk of coronary heart disease is inversely associated with the level of high-density lipoprotein cholesterol (HDL-C) has motivated several phase III programmes with cholesteryl ester transfer protein (CETP) inhibitors. To assess alternative methods to predict clinical response of CETP inhibitors. Meta-regression analysis on raising HDL-C drugs (statins, fibrates, niacin) in randomised controlled trials. 51 trials in secondary prevention with a total of 167,311 patients for a follow-up >1 year where HDL-C was measured at baseline and during treatment. The meta-regression analysis showed no significant association between change in HDL-C (treatment vs comparator) and log risk ratio (RR) of clinical endpoint (non-fatal myocardial infarction or cardiac death). CETP inhibitors data are consistent with this finding (RR: 1.03; P5-P95: 0.99-1.21). A prespecified sensitivity analysis by drug class suggested that the strength of relationship might differ between pharmacological groups. A significant association for both statins (p<0.02, log RR=-0.169-0.0499*HDL-C change, R(2)=0.21) and niacin (p=0.02, log RR=1.07-0.185*HDL-C change, R(2)=0.61) but not fibrates (p=0.18, log RR=-0.367+0.077*HDL-C change, R(2)=0.40) was shown. However, the association was no longer detectable after adjustment for low-density lipoprotein cholesterol for statins or exclusion of open trials for niacin. Meta-regression suggested that CETP inhibitors might not influence coronary risk. The relation between change in HDL-C level and clinical endpoint may be drug dependent, which limits the use of HDL-C as a surrogate marker of coronary events. Other markers of HDL function may be more relevant. Published by the BMJ Publishing Group Limited. For permission to use (where not already granted under a licence) please go to http://group.bmj.com/group/rights-licensing/permissions.
Trend Analysis of Cancer Mortality and Incidence in Panama, Using Joinpoint Regression Analysis.

Science.gov (United States)

Politis, Michael; Higuera, Gladys; Chang, Lissette Raquel; Gomez, Beatriz; Bares, Juan; Motta, Jorge

2015-06-01

Cancer is one of the leading causes of death worldwide and its incidence is expected to increase in the future. In Panama, cancer is also one of the leading causes of death. In 1964, a nationwide cancer registry was started and it was restructured and improved in 2012. The aim of this study is to utilize Joinpoint regression analysis to study the trends of the incidence and mortality of cancer in Panama in the last decade. Cancer mortality was estimated from the Panamanian National Institute of Census and Statistics Registry for the period 2001 to 2011. Cancer incidence was estimated from the Panamanian National Cancer Registry for the period 2000 to 2009. The Joinpoint Regression Analysis program, version 4.0.4, was used to calculate trends by age-adjusted incidence and mortality rates for selected cancers. Overall, the trend of age-adjusted cancer mortality in Panama has declined over the last 10 years (-1.12% per year). The cancers for which there was a significant increase in the trend of mortality were female breast cancer and ovarian cancer; while the highest increases in incidence were shown for breast cancer, liver cancer, and prostate cancer. Significant decrease in the trend of mortality was evidenced for the following: prostate cancer, lung and bronchus cancer, and cervical cancer; with respect to incidence, only oral and pharynx cancer in both sexes had a significant decrease. Some cancers showed no significant trends in incidence or mortality. This study reveals contrasting trends in cancer incidence and mortality in Panama in the last decade. Although Panama is considered an upper middle income nation, this study demonstrates that some cancer mortality trends, like the ones seen in cervical and lung cancer, behave similarly to the ones seen in high income countries. In contrast, other types, like breast cancer, follow a pattern seen in countries undergoing a transition to a developed economy with its associated lifestyle, nutrition, and body weight
Selenium Exposure and Cancer Risk: an Updated Meta-analysis and Meta-regression

Science.gov (United States)

Cai, Xianlei; Wang, Chen; Yu, Wanqi; Fan, Wenjie; Wang, Shan; Shen, Ning; Wu, Pengcheng; Li, Xiuyang; Wang, Fudi

2016-01-01

The objective of this study was to investigate the associations between selenium exposure and cancer risk. We identified 69 studies and applied meta-analysis, meta-regression and dose-response analysis to obtain available evidence. The results indicated that high selenium exposure had a protective effect on cancer risk (pooled OR = 0.78; 95%CI: 0.73–0.83). The results of linear and nonlinear dose-response analysis indicated that high serum/plasma selenium and toenail selenium had the efficacy on cancer prevention. However, we did not find a protective efficacy of selenium supplement. High selenium exposure may have different effects on specific types of cancer. It decreased the risk of breast cancer, lung cancer, esophageal cancer, gastric cancer, and prostate cancer, but it was not associated with colorectal cancer, bladder cancer, and skin cancer. PMID:26786590
Similar estimates of temperature impacts on global wheat yield by three independent methods

DEFF Research Database (Denmark)

Liu, Bing; Asseng, Senthold; Müller, Christoph

2016-01-01

The potential impact of global temperature change on global crop yield has recently been assessed with different methods. Here we show that grid-based and point-based simulations and statistical regressions (from historic records), without deliberate adaptation or CO2 fertilization effects, produ......-method ensemble, it was possible to quantify ‘method uncertainty’ in addition to model uncertainty. This significantly improves confidence in estimates of climate impacts on global food security.......The potential impact of global temperature change on global crop yield has recently been assessed with different methods. Here we show that grid-based and point-based simulations and statistical regressions (from historic records), without deliberate adaptation or CO2 fertilization effects, produce...... similar estimates of temperature impact on wheat yields at global and national scales. With a 1 °C global temperature increase, global wheat yield is projected to decline between 4.1% and 6.4%. Projected relative temperature impacts from different methods were similar for major wheat-producing countries...
Similar Estimates of Temperature Impacts on Global Wheat Yield by Three Independent Methods

Science.gov (United States)

Liu, Bing; Asseng, Senthold; Muller, Christoph; Ewart, Frank; Elliott, Joshua; Lobell, David B.; Martre, Pierre; Ruane, Alex C.; Wallach, Daniel; Jones, James W.;

2016-01-01

The potential impact of global temperature change on global crop yield has recently been assessed with different methods. Here we show that grid-based and point-based simulations and statistical regressions (from historic records), without deliberate adaptation or CO2 fertilization effects, produce similar estimates of temperature impact on wheat yields at global and national scales. With a 1 C global temperature increase, global wheat yield is projected to decline between 4.1% and 6.4%. Projected relative temperature impacts from different methods were similar for major wheat-producing countries China, India, USA and France, but less so for Russia. Point-based and grid-based simulations, and to some extent the statistical regressions, were consistent in projecting that warmer regions are likely to suffer more yield loss with increasing temperature than cooler regions. By forming a multi-method ensemble, it was possible to quantify 'method uncertainty' in addition to model uncertainty. This significantly improves confidence in estimates of climate impacts on global food security.

Similar estimates of temperature impacts on global wheat yield by three independent methods

Science.gov (United States)

Liu, Bing; Asseng, Senthold; Müller, Christoph; Ewert, Frank; Elliott, Joshua; Lobell, David B.; Martre, Pierre; Ruane, Alex C.; Wallach, Daniel; Jones, James W.; Rosenzweig, Cynthia; Aggarwal, Pramod K.; Alderman, Phillip D.; Anothai, Jakarat; Basso, Bruno; Biernath, Christian; Cammarano, Davide; Challinor, Andy; Deryng, Delphine; Sanctis, Giacomo De; Doltra, Jordi; Fereres, Elias; Folberth, Christian; Garcia-Vila, Margarita; Gayler, Sebastian; Hoogenboom, Gerrit; Hunt, Leslie A.; Izaurralde, Roberto C.; Jabloun, Mohamed; Jones, Curtis D.; Kersebaum, Kurt C.; Kimball, Bruce A.; Koehler, Ann-Kristin; Kumar, Soora Naresh; Nendel, Claas; O'Leary, Garry J.; Olesen, Jørgen E.; Ottman, Michael J.; Palosuo, Taru; Prasad, P. V. Vara; Priesack, Eckart; Pugh, Thomas A. M.; Reynolds, Matthew; Rezaei, Ehsan E.; Rötter, Reimund P.; Schmid, Erwin; Semenov, Mikhail A.; Shcherbak, Iurii; Stehfest, Elke; Stöckle, Claudio O.; Stratonovitch, Pierre; Streck, Thilo; Supit, Iwan; Tao, Fulu; Thorburn, Peter; Waha, Katharina; Wall, Gerard W.; Wang, Enli; White, Jeffrey W.; Wolf, Joost; Zhao, Zhigan; Zhu, Yan

2016-12-01

The potential impact of global temperature change on global crop yield has recently been assessed with different methods. Here we show that grid-based and point-based simulations and statistical regressions (from historic records), without deliberate adaptation or CO2 fertilization effects, produce similar estimates of temperature impact on wheat yields at global and national scales. With a 1 °C global temperature increase, global wheat yield is projected to decline between 4.1% and 6.4%. Projected relative temperature impacts from different methods were similar for major wheat-producing countries China, India, USA and France, but less so for Russia. Point-based and grid-based simulations, and to some extent the statistical regressions, were consistent in projecting that warmer regions are likely to suffer more yield loss with increasing temperature than cooler regions. By forming a multi-method ensemble, it was possible to quantify `method uncertainty’ in addition to model uncertainty. This significantly improves confidence in estimates of climate impacts on global food security.
Predictive model of Amorphophallus muelleri growth in some agroforestry in East Java by multiple regression analysis

Directory of Open Access Journals (Sweden)

BUDIMAN

2012-01-01

Full Text Available Budiman, Arisoesilaningsih E. 2012. Predictive model of Amorphophallus muelleri growth in some agroforestry in East Java by multiple regression analysis. Biodiversitas 13: 18-22. The aims of this research was to determine the multiple regression models of vegetative and corm growth of Amorphophallus muelleri Blume in some age variations and habitat conditions of agroforestry in East Java. Descriptive exploratory research method was conducted by systematic random sampling at five agroforestries on four plantations in East Java: Saradan, Bojonegoro, Nganjuk and Blitar. In each agroforestry, we observed A. muelleri vegetative and corm growth on four growing age (1, 2, 3 and 4 years old respectively as well as environmental variables such as altitude, vegetation, climate and soil conditions. Data were analyzed using descriptive statistics to compare A. muelleri habitat in five agroforestries. Meanwhile, the influence and contribution of each environmental variable to the growth of A. muelleri vegetative and corm were determined using multiple regression analysis of SPSS 17.0. The multiple regression models of A. muelleri vegetative and corm growth were generated based on some characteristics of agroforestries and age showed high validity with R2 = 88-99%. Regression model showed that age, monthly temperatures, percentage of radiation and soil calcium (Ca content either simultaneously or partially determined the growth of A. muelleri vegetative and corm. Based on these models, the A. muelleri corm reached the optimal growth after four years of cultivation and they will be ready to be harvested. Additionally, the soil Ca content should reach 25.3 me.hg-1 as Sugihwaras agroforestry, with the maximal radiation of 60%.
Response of Barley Double Haploid Lines to the Grain Yield and Morphological Traits under Water Deficit Stress Conditions

Directory of Open Access Journals (Sweden)

Maroof Khalily

2017-04-01

Full Text Available To study the relationships of grain yield and some of agro-morphological traits in 40 doubled haploid (DH lines along with parental and three check genotypes in a randomized complete block design with two replications under two water regimes (normal and stress were evaluated during 2011-2012 and 2012-2013 growing seasons. Combined analysis of variance showed significant difference for all the traits in terms of the year, water regimes, lines, and and line × year. Comparison of group means, between non-stress and stress conditions, showed that DH lines had the lowest reduction percentage for the number of grains per spike, thousand grain weight, grain yield and biological yield as opposed to check genotypes. The correlation between grain yield with biological yield, harvest index, thousand grain weight, and hectoliter of kernel weight in both conditions, were highly significant and positive. Based on stepwise regression the peduncle length, number of seeds per spike, thousand seed weight, and hectoliter of kernel weight had important effect on increasing seed yield. The result of path analysis showed that these traits had the highest direct effect on grain yield. Based on mean comparisons of morphological characters as well as STI and GMP indices it can be concluded that lines No.11, 13, 14, 24, 29, 30, 35 and 39 were distinguished to be desirable lines for grain yield and their related traits and also tolerant lines in terms of response to drought stress conditions.
Determination of Selection Index of Cocoa (Theobroma Cacao L.) Yield Traits Using Regression Methods

OpenAIRE

Setyawan, Bayu; Taryono; Mitrowihardjo, Suyadi

2016-01-01

The increasing chocolate consumption has not been followed by growing production of dry cocoa beans. In order to support the increase in cocoa production, planting materials with high yield are needed. The objective of this research was to determine the components of cocoa traits affecting weight of dry cocoa beans, and set a selection index for superior cocoa trees. The experiment material were four cocoa hybrid populations of which their family ancestry were unknown, and were planted on Sam...
VCS-SSA Mainz Experiment. Measurement of the beam spin asymmetry in (e polarized p {yields} ep{gamma}) and (e polarized p {yields} ep{pi}{sup 0}). Final analysis - MEMO I

Energy Technology Data Exchange (ETDEWEB)

Fonvieille, H.; Bensafa, I. [LPC-Clermont-Fd, Universite Blaise Pascal, F-63170 Aubiere Cedex (France)

2006-02-15

This note gives details on the final analysis of the VCS-SSA experiment in terms of Beam Spin Asymmetry. It summarizes the changes between the first and second pass analysis. Then the measured asymmetry is presented for both channels e polarized p {yields} ep{gamma} and e polarized p {yields} ep{pi}{sup 0} including systematic studies. The final experimental result is briefly compared to some model predictions. (authors)
Interpreting Bivariate Regression Coefficients: Going beyond the Average

Science.gov (United States)

Halcoussis, Dennis; Phillips, G. Michael

2010-01-01

Statistics, econometrics, investment analysis, and data analysis classes often review the calculation of several types of averages, including the arithmetic mean, geometric mean, harmonic mean, and various weighted averages. This note shows how each of these can be computed using a basic regression framework. By recognizing when a regression model…
Statistical analysis of sediment toxicity by additive monotone regression splines

NARCIS (Netherlands)

Boer, de W.J.; Besten, den P.J.; Braak, ter C.J.F.

2002-01-01

Modeling nonlinearity and thresholds in dose-effect relations is a major challenge, particularly in noisy data sets. Here we show the utility of nonlinear regression with additive monotone regression splines. These splines lead almost automatically to the estimation of thresholds. We applied this
Modeling maximum daily temperature using a varying coefficient regression model

Science.gov (United States)

Han Li; Xinwei Deng; Dong-Yum Kim; Eric P. Smith

2014-01-01

Relationships between stream water and air temperatures are often modeled using linear or nonlinear regression methods. Despite a strong relationship between water and air temperatures and a variety of models that are effective for data summarized on a weekly basis, such models did not yield consistently good predictions for summaries such as daily maximum temperature...
Targeting: Logistic Regression, Special Cases and Extensions

Directory of Open Access Journals (Sweden)

Helmut Schaeben

2014-12-01

Full Text Available Logistic regression is a classical linear model for logit-transformed conditional probabilities of a binary target variable. It recovers the true conditional probabilities if the joint distribution of predictors and the target is of log-linear form. Weights-of-evidence is an ordinary logistic regression with parameters equal to the differences of the weights of evidence if all predictor variables are discrete and conditionally independent given the target variable. The hypothesis of conditional independence can be tested in terms of log-linear models. If the assumption of conditional independence is violated, the application of weights-of-evidence does not only corrupt the predicted conditional probabilities, but also their rank transform. Logistic regression models, including the interaction terms, can account for the lack of conditional independence, appropriate interaction terms compensate exactly for violations of conditional independence. Multilayer artificial neural nets may be seen as nested regression-like models, with some sigmoidal activation function. Most often, the logistic function is used as the activation function. If the net topology, i.e., its control, is sufficiently versatile to mimic interaction terms, artificial neural nets are able to account for violations of conditional independence and yield very similar results. Weights-of-evidence cannot reasonably include interaction terms; subsequent modifications of the weights, as often suggested, cannot emulate the effect of interaction terms.

A Visual Analytics Approach for Correlation, Classification, and Regression Analysis

Energy Technology Data Exchange (ETDEWEB)

Steed, Chad A [ORNL; SwanII, J. Edward [Mississippi State University (MSU); Fitzpatrick, Patrick J. [Mississippi State University (MSU); Jankun-Kelly, T.J. [Mississippi State University (MSU)

2012-02-01

New approaches that combine the strengths of humans and machines are necessary to equip analysts with the proper tools for exploring today's increasing complex, multivariate data sets. In this paper, a novel visual data mining framework, called the Multidimensional Data eXplorer (MDX), is described that addresses the challenges of today's data by combining automated statistical analytics with a highly interactive parallel coordinates based canvas. In addition to several intuitive interaction capabilities, this framework offers a rich set of graphical statistical indicators, interactive regression analysis, visual correlation mining, automated axis arrangements and filtering, and data classification techniques. The current work provides a detailed description of the system as well as a discussion of key design aspects and critical feedback from domain experts.
MODELING POLLINATION FACTORS THAT INFLUENCE ALFALFA SEED YIELD IN NORTH-CENTRAL NEVADA

OpenAIRE

BREAZEALE, Don; FERNANDEZ, George; NARAYANAN, Rangesan

2008-01-01

The relative importance of both environmental and management factors on alfalfa seed yield was investigated on North–Central Nevada farms. Multiple linear regression models using 2002-2003 data revealed that cumulative tripped fl owers increased seed yield in both years. Field location does not appear to make a difference in the observed variation in tripped fl ower production. The results suggest that seed yield can be increased by (a) by placing bee shelters closer and (b) cultural practice...
Spatial-Temporal Variations of Turbidity and Ocean Current Velocity of the Ariake Sea Area, Kyushu, Japan Through Regression Analysis with Remote Sensing Satellite Data

OpenAIRE

Yuichi Sarusawa; Kohei Arai

2013-01-01

Regression analysis based method for turbidity and ocean current velocity estimation with remote sensing satellite data is proposed. Through regressive analysis with MODIS data and measured data of turbidity and ocean current velocity, regressive equation which allows estimation of turbidity and ocean current velocity is obtained. With the regressive equation as well as long term MODIS data, turbidity and ocean current velocity trends in Ariake Sea area are clarified. It is also confirmed tha...
Characterization of sonographically indeterminate ovarian tumors with MR imaging. A logistic regression analysis

International Nuclear Information System (INIS)

Yamashita, Y.; Hatanaka, Y.; Torashima, M.; Takahashi, M.; Miyazaki, K.; Okamura, H.

1997-01-01

Purpose: The goal of this study was to maximize the discrimination between benign and malignant masses in patients with sonographically indeterminate ovarian lesions by means of unenhanced and contrast-enhanced MR imaging, and to develop a computer-assisted diagnosis system. Material and Methods: Findings in precontrast and Gd-DTPA contrast-enhanced MR images of 104 patients with 115 sonographically indeterminate ovarian masses were analyzed, and the results were correlated with histopathological findings. Of 115 lesions, 65 were benign (23 cystadenomas, 13 complex cysts, 11 teratomas, 6 fibrothecomas, 12 others) and 50 were malignant (32 ovarian carcinomas, 7 metastatic tumors of the ovary, 4 carcinomas of the fallopian tubes, 7 others). A logistic regression analysis was performed to discriminate between benign and malignant lesions, and a model of a computer-assisted diagnosis was developed. This model was prospectively tested in 75 cases of ovarian tumors found at other institutions. Results: From the univariate analysis, the following parameters were selected as significant for predicting malignancy (p≤0.05): A solid or cystic mass with a large solid component or wall thickness greater than 3 mm; complex internal architecture; ascites; and bilaterality. Based on these parameters, a model of a computer-assisted diagnosis system was developed with the logistic regression analysis. To distinguish benign from malignant lesions, the maximum cut-off point was obtained between 0.47 and 0.51. In a prospective application of this model, 87% of the lesions were accurately identified as benign or malignant. (orig.)
Mixed kernel function support vector regression for global sensitivity analysis

Science.gov (United States)

Cheng, Kai; Lu, Zhenzhou; Wei, Yuhao; Shi, Yan; Zhou, Yicheng

2017-11-01

Global sensitivity analysis (GSA) plays an important role in exploring the respective effects of input variables on an assigned output response. Amongst the wide sensitivity analyses in literature, the Sobol indices have attracted much attention since they can provide accurate information for most models. In this paper, a mixed kernel function (MKF) based support vector regression (SVR) model is employed to evaluate the Sobol indices at low computational cost. By the proposed derivation, the estimation of the Sobol indices can be obtained by post-processing the coefficients of the SVR meta-model. The MKF is constituted by the orthogonal polynomials kernel function and Gaussian radial basis kernel function, thus the MKF possesses both the global characteristic advantage of the polynomials kernel function and the local characteristic advantage of the Gaussian radial basis kernel function. The proposed approach is suitable for high-dimensional and non-linear problems. Performance of the proposed approach is validated by various analytical functions and compared with the popular polynomial chaos expansion (PCE). Results demonstrate that the proposed approach is an efficient method for global sensitivity analysis.
Video image analysis in the Australian meat industry - precision and accuracy of predicting lean meat yield in lamb carcasses.

Science.gov (United States)

Hopkins, D L; Safari, E; Thompson, J M; Smith, C R

2004-06-01

A wide selection of lamb types of mixed sex (ewes and wethers) were slaughtered at a commercial abattoir and during this process images of 360 carcasses were obtained online using the VIAScan® system developed by Meat and Livestock Australia. Soft tissue depth at the GR site (thickness of tissue over the 12th rib 110 mm from the midline) was measured by an abattoir employee using the AUS-MEAT sheep probe (PGR). Another measure of this thickness was taken in the chiller using a GR knife (NGR). Each carcass was subsequently broken down to a range of trimmed boneless retail cuts and the lean meat yield determined. The current industry model for predicting meat yield uses hot carcass weight (HCW) and tissue depth at the GR site. A low level of accuracy and precision was found when HCW and PGR were used to predict lean meat yield (R(2)=0.19, r.s.d.=2.80%), which could be improved markedly when PGR was replaced by NGR (R(2)=0.41, r.s.d.=2.39%). If the GR measures were replaced by 8 VIAScan® measures then greater prediction accuracy could be achieved (R(2)=0.52, r.s.d.=2.17%). A similar result was achieved when the model was based on principal components (PCs) computed from the 8 VIAScan® measures (R(2)=0.52, r.s.d.=2.17%). The use of PCs also improved the stability of the model compared to a regression model based on HCW and NGR. The transportability of the models was tested by randomly dividing the data set and comparing coefficients and the level of accuracy and precision. Those models based on PCs were superior to those based on regression. It is demonstrated that with the appropriate modeling the VIAScan® system offers a workable method for predicting lean meat yield automatically.
Analysis of the spatial variability of crop yield and soil properties in small agricultural plots

Directory of Open Access Journals (Sweden)

Vieira Sidney Rosa

2003-01-01

Full Text Available The objective of this study was to assess spatial variability of soil properties and crop yield under no tillage as a function of time, in two soil/climate conditions in São Paulo State, Brazil. The two sites measured approximately one hectare each and were cultivated with crop sequences which included corn, soybean, cotton, oats, black oats, wheat, rye, rice and green manure. Soil fertility, soil physical properties and crop yield were measured in a 10-m grid. The soils were a Dusky Red Latossol (Oxisol and a Red Yellow Latossol (Ultisol. Soil sampling was performed in each field every two years after harvesting of the summer crop. Crop yield was measured at the end of each crop cycle, in 2 x 2.5 m sub plots. Data were analysed using semivariogram analysis and kriging interpolation for contour map generation. Yield maps were constructed in order to visually compare the variability of yields, the variability of the yield components and related soil properties. The results show that the factors affecting the variability of crop yield varies from one crop to another. The changes in yield from one year to another suggest that the causes of variability may change with time. The changes with time for the cross semivariogram between phosphorus in leaves and soybean yield is another evidence of this result.
Examination of influential observations in penalized spline regression

Science.gov (United States)

Türkan, Semra

2013-10-01

In parametric or nonparametric regression models, the results of regression analysis are affected by some anomalous observations in the data set. Thus, detection of these observations is one of the major steps in regression analysis. These observations are precisely detected by well-known influence measures. Pena's statistic is one of them. In this study, Pena's approach is formulated for penalized spline regression in terms of ordinary residuals and leverages. The real data and artificial data are used to see illustrate the effectiveness of Pena's statistic as to Cook's distance on detecting influential observations. The results of the study clearly reveal that the proposed measure is superior to Cook's Distance to detect these observations in large data set.
Ca analysis: An Excel based program for the analysis of intracellular calcium transients including multiple, simultaneous regression analysis☆

Science.gov (United States)

Greensmith, David J.

2014-01-01

Here I present an Excel based program for the analysis of intracellular Ca transients recorded using fluorescent indicators. The program can perform all the necessary steps which convert recorded raw voltage changes into meaningful physiological information. The program performs two fundamental processes. (1) It can prepare the raw signal by several methods. (2) It can then be used to analyze the prepared data to provide information such as absolute intracellular Ca levels. Also, the rates of change of Ca can be measured using multiple, simultaneous regression analysis. I demonstrate that this program performs equally well as commercially available software, but has numerous advantages, namely creating a simplified, self-contained analysis workflow. PMID:24125908
Nonparametric Mixture of Regression Models.

Science.gov (United States)

Huang, Mian; Li, Runze; Wang, Shaoli

2013-07-01

Motivated by an analysis of US house price index data, we propose nonparametric finite mixture of regression models. We study the identifiability issue of the proposed models, and develop an estimation procedure by employing kernel regression. We further systematically study the sampling properties of the proposed estimators, and establish their asymptotic normality. A modified EM algorithm is proposed to carry out the estimation procedure. We show that our algorithm preserves the ascent property of the EM algorithm in an asymptotic sense. Monte Carlo simulations are conducted to examine the finite sample performance of the proposed estimation procedure. An empirical analysis of the US house price index data is illustrated for the proposed methodology.
Importance of growth characteristics for yield of barley in different growing systems: will growth characteristics describe yield diffently in different growing systems?

DEFF Research Database (Denmark)

Kristensen, Kristian; Ericson, Lars

2008-01-01

The interest in organic grown cereals has increased the need for variety tests under organic growing systems and/or the knowledge on whether growth characteristics describe yield differently under conventional and organic conditions. This paper is a contribution to that question by examining...... the relationships between some important growth characteristics in barley trials in both systems in Northern Sweden and in Denmark. Mixed model analyses were used for regressions of growth characteristics (or transformations of those) on yield (and log-transformed yield), allowing the slope to depend on the growing...... system. The analyses showed that diseases seemed to have a less negative effect on yield in the organic growing system than in the conventional growing system if pesticides were not applied. For other characteristics the effect depended on the country. This was the case for grain characteristics where...
Evaluation of Visual Field Progression in Glaucoma: Quasar Regression Program and Event Analysis.

Science.gov (United States)

Díaz-Alemán, Valentín T; González-Hernández, Marta; Perera-Sanz, Daniel; Armas-Domínguez, Karintia

2016-01-01

To determine the sensitivity, specificity and agreement between the Quasar program, glaucoma progression analysis (GPA II) event analysis and expert opinion in the detection of glaucomatous progression. The Quasar program is based on linear regression analysis of both mean defect (MD) and pattern standard deviation (PSD). Each series of visual fields was evaluated by three methods; Quasar, GPA II and four experts. The sensitivity, specificity and agreement (kappa) for each method was calculated, using expert opinion as the reference standard. The study included 439 SITA Standard visual fields of 56 eyes of 42 patients, with a mean of 7.8 ± 0.8 visual fields per eye. When suspected cases of progression were considered stable, sensitivity and specificity of Quasar, GPA II and the experts were 86.6% and 70.7%, 26.6% and 95.1%, and 86.6% and 92.6% respectively. When suspected cases of progression were considered as progressing, sensitivity and specificity of Quasar, GPA II and the experts were 79.1% and 81.2%, 45.8% and 90.6%, and 85.4% and 90.6% respectively. The agreement between Quasar and GPA II when suspected cases were considered stable or progressing was 0.03 and 0.28 respectively. The degree of agreement between Quasar and the experts when suspected cases were considered stable or progressing was 0.472 and 0.507. The degree of agreement between GPA II and the experts when suspected cases were considered stable or progressing was 0.262 and 0.342. The combination of MD and PSD regression analysis in the Quasar program showed better agreement with the experts and higher sensitivity than GPA II.
Regression analysis of mixed recurrent-event and panel-count data with additive rate models.

Science.gov (United States)

Zhu, Liang; Zhao, Hui; Sun, Jianguo; Leisenring, Wendy; Robison, Leslie L

2015-03-01

Event-history studies of recurrent events are often conducted in fields such as demography, epidemiology, medicine, and social sciences (Cook and Lawless, 2007, The Statistical Analysis of Recurrent Events. New York: Springer-Verlag; Zhao et al., 2011, Test 20, 1-42). For such analysis, two types of data have been extensively investigated: recurrent-event data and panel-count data. However, in practice, one may face a third type of data, mixed recurrent-event and panel-count data or mixed event-history data. Such data occur if some study subjects are monitored or observed continuously and thus provide recurrent-event data, while the others are observed only at discrete times and hence give only panel-count data. A more general situation is that each subject is observed continuously over certain time periods but only at discrete times over other time periods. There exists little literature on the analysis of such mixed data except that published by Zhu et al. (2013, Statistics in Medicine 32, 1954-1963). In this article, we consider the regression analysis of mixed data using the additive rate model and develop some estimating equation-based approaches to estimate the regression parameters of interest. Both finite sample and asymptotic properties of the resulting estimators are established, and the numerical studies suggest that the proposed methodology works well for practical situations. The approach is applied to a Childhood Cancer Survivor Study that motivated this study. © 2014, The International Biometric Society.
Yield gap determinants for wheat production in major irrigated cropping zones of punjab, pakistan

International Nuclear Information System (INIS)

Hussain, A.; Aujla, K.M.; Badar, N.

2014-01-01

Yield gap is useful measurement for crop productivity and the extent to which crop productivity falls below some potential level. The study was carried out to analyze the yield gap and determinants of wheat production in the Punjab province of Pakistan. It is based on cross sectional data from 210 farmers for the crop year 2009-10. Results suggest that farm level wheat yields are less than the potential yield level by 33.0%, 43.0% and 50.6% in the mixed-cropping, cotton-wheat and rice-wheat zones of the province, respectively. Ordinary least square regression analysis of wheat production by assuming Cobb-Douglas specification reveals that the number of irrigations, usage of farm yard manure and fertilizers contribute positively and significantly to wheat crop production. Coefficients of dummy variables for cropping zones indicate that farmers in the mixed cropping zone are obtaining better yield of the wheat crop as compared to their counterparts in other selected cropping zones. These results suggested that farmers can increase wheat productivity by increasing the use of factor inputs; however, poverty may be a constraint on realizing these gains. Thus, wheat production can be increased in the country by helping resource poor farmers through suitable support mechanisms. (author)
Laser-induced Breakdown spectroscopy quantitative analysis method via adaptive analytical line selection and relevance vector machine regression model

International Nuclear Information System (INIS)

Yang, Jianhong; Yi, Cancan; Xu, Jinwu; Ma, Xianghong

2015-01-01

A new LIBS quantitative analysis method based on analytical line adaptive selection and Relevance Vector Machine (RVM) regression model is proposed. First, a scheme of adaptively selecting analytical line is put forward in order to overcome the drawback of high dependency on a priori knowledge. The candidate analytical lines are automatically selected based on the built-in characteristics of spectral lines, such as spectral intensity, wavelength and width at half height. The analytical lines which will be used as input variables of regression model are determined adaptively according to the samples for both training and testing. Second, an LIBS quantitative analysis method based on RVM is presented. The intensities of analytical lines and the elemental concentrations of certified standard samples are used to train the RVM regression model. The predicted elemental concentration analysis results will be given with a form of confidence interval of probabilistic distribution, which is helpful for evaluating the uncertainness contained in the measured spectra. Chromium concentration analysis experiments of 23 certified standard high-alloy steel samples have been carried out. The multiple correlation coefficient of the prediction was up to 98.85%, and the average relative error of the prediction was 4.01%. The experiment results showed that the proposed LIBS quantitative analysis method achieved better prediction accuracy and better modeling robustness compared with the methods based on partial least squares regression, artificial neural network and standard support vector machine. - Highlights: • Both training and testing samples are considered for analytical lines selection. • The analytical lines are auto-selected based on the built-in characteristics of spectral lines. • The new method can achieve better prediction accuracy and modeling robustness. • Model predictions are given with confidence interval of probabilistic distribution
Regression modeling of ground-water flow

Science.gov (United States)

Cooley, R.L.; Naff, R.L.

1985-01-01

Nonlinear multiple regression methods are developed to model and analyze groundwater flow systems. Complete descriptions of regression methodology as applied to groundwater flow models allow scientists and engineers engaged in flow modeling to apply the methods to a wide range of problems. Organization of the text proceeds from an introduction that discusses the general topic of groundwater flow modeling, to a review of basic statistics necessary to properly apply regression techniques, and then to the main topic: exposition and use of linear and nonlinear regression to model groundwater flow. Statistical procedures are given to analyze and use the regression models. A number of exercises and answers are included to exercise the student on nearly all the methods that are presented for modeling and statistical analysis. Three computer programs implement the more complex methods. These three are a general two-dimensional, steady-state regression model for flow in an anisotropic, heterogeneous porous medium, a program to calculate a measure of model nonlinearity with respect to the regression parameters, and a program to analyze model errors in computed dependent variables such as hydraulic head. (USGS)
Principal coordinate analysis of genotype × environment interaction for grain yield of bread wheat in the semi-arid regions

Directory of Open Access Journals (Sweden)

Sabaghnia Naser

2013-01-01

Full Text Available Multi-environmental trials have significant main effects and significant multiplicative genotype × environment (GE interaction effect. Principal coordinate analysis (PCOA offers a more appropriate statistical analysis to deal with such situations, compared to traditional statistical methods. Eighteen bread wheat genotypes were grown in four semi-arid regions over three year seasons to study the GE interaction and yield stability and obtained data on grain yield were analyzed using PCOA. Combined analysis of variance indicated that all of the studied effects including the main effects of genotype and environments as well as the GE interaction were highly significant. According to grand means and total mean yield, test environments were grouped to two main groups as high mean yield (H and low mean yield (L. There were five H test environments and six L test environments which analyzed in the sequential cycles. For each cycle, both scatter point diagram and minimum spanning tree plot were drawn. The identified most stable genotypes with dynamic stability concept and based on the minimum spanning tree plots and centroid distances were G1 (3310.2 kg ha-1 and G5 (3065.6 kg ha-1, and therefore could be recommended for unfavorable or poor conditions. Also, genotypes G7 (3047.2 kg ha-1 and G16 (3132.3 kg ha-1 were located several times in the vertex positions of high cycles according to the principal coordinates analysis. The principal coordinates analysis provided useful and interesting ways of investigating GE interaction of barley genotypes. Finally, the results of principal coordinates analysis in general confirmed the breeding value of the genotypes, obtained on the basis of the yield stability evaluation.
Influence diagnostics in meta-regression model.

Science.gov (United States)

Shi, Lei; Zuo, ShanShan; Yu, Dalei; Zhou, Xiaohua

2017-09-01

This paper studies the influence diagnostics in meta-regression model including case deletion diagnostic and local influence analysis. We derive the subset deletion formulae for the estimation of regression coefficient and heterogeneity variance and obtain the corresponding influence measures. The DerSimonian and Laird estimation and maximum likelihood estimation methods in meta-regression are considered, respectively, to derive the results. Internal and external residual and leverage measure are defined. The local influence analysis based on case-weights perturbation scheme, responses perturbation scheme, covariate perturbation scheme, and within-variance perturbation scheme are explored. We introduce a method by simultaneous perturbing responses, covariate, and within-variance to obtain the local influence measure, which has an advantage of capable to compare the influence magnitude of influential studies from different perturbations. An example is used to illustrate the proposed methodology. Copyright © 2017 John Wiley & Sons, Ltd.
Interactions between cadmium and decabrominated diphenyl ether on blood cells count in rats-Multiple factorial regression analysis.

Science.gov (United States)

Curcic, Marijana; Buha, Aleksandra; Stankovic, Sanja; Milovanovic, Vesna; Bulat, Zorica; Đukić-Ćosić, Danijela; Antonijević, Evica; Vučinić, Slavica; Matović, Vesna; Antonijevic, Biljana

2017-02-01

The objective of this study was to assess toxicity of Cd and BDE-209 mixture on haematological parameters in subacutely exposed rats and to determine the presence and type of interactions between these two chemicals using multiple factorial regression analysis. Furthermore, for the assessment of interaction type, an isobologram based methodology was applied and compared with multiple factorial regression analysis. Chemicals were given by oral gavage to the male Wistar rats weighing 200-240g for 28days. Animals were divided in 16 groups (8/group): control vehiculum group, three groups of rats were treated with 2.5, 7.5 or 15mg Cd/kg/day. These doses were chosen on the bases of literature data and reflect relatively high Cd environmental exposure, three groups of rats were treated with 1000, 2000 or 4000mg BDE-209/kg/bw/day, doses proved to induce toxic effects in rats. Furthermore, nine groups of animals were treated with different mixtures of Cd and BDE-209 containing doses of Cd and BDE-209 stated above. Blood samples were taken at the end of experiment and red blood cells, white blood cells and platelets counts were determined. For interaction assessment multiple factorial regression analysis and fitted isobologram approach were used. In this study, we focused on multiple factorial regression analysis as a method for interaction assessment. We also investigated the interactions between Cd and BDE-209 by the derived model for the description of the obtained fitted isobologram curves. Current study indicated that co-exposure to Cd and BDE-209 can result in significant decrease in RBC count, increase in WBC count and decrease in PLT count, when compared with controls. Multiple factorial regression analysis used for the assessment of interactions type between Cd and BDE-209 indicated synergism for the effect on RBC count and no interactions i.e. additivity for the effects on WBC and PLT counts. On the other hand, isobologram based approach showed slight antagonism
Skeletal height estimation from regression analysis of sternal lengths in a Northwest Indian population of Chandigarh region: a postmortem study.

Science.gov (United States)

Singh, Jagmahender; Pathak, R K; Chavali, Krishnadutt H

2011-03-20

Skeletal height estimation from regression analysis of eight sternal lengths in the subjects of Chandigarh zone of Northwest India is the topic of discussion in this study. Analysis of eight sternal lengths (length of manubrium, length of mesosternum, combined length of manubrium and mesosternum, total sternal length and first four intercostals lengths of mesosternum) measured from 252 male and 91 female sternums obtained at postmortems revealed that mean cadaver stature and sternal lengths were more in North Indians and males than the South Indians and females. Except intercostal lengths, all the sternal lengths were positively correlated with stature of the deceased in both sexes (P regression analysis of sternal lengths was found more useful than the linear regression for stature estimation. Using multivariate regression analysis, the combined length of manubrium and mesosternum in both sexes and the length of manubrium along with 2nd and 3rd intercostal lengths of mesosternum in males were selected as best estimators of stature. Nonetheless, the stature of males can be predicted with SEE of 6.66 (R(2) = 0.16, r = 0.318) from combination of MBL+BL_3+LM+BL_2, and in females from MBL only, it can be estimated with SEE of 6.65 (R(2) = 0.10, r = 0.318), whereas from the multiple regression analysis of pooled data, stature can be known with SEE of 6.97 (R(2) = 0.387, r = 575) from the combination of MBL+LM+BL_2+TSL+BL_3. The R(2) and F-ratio were found to be statistically significant for almost all the variables in both the sexes, except 4th intercostal length in males and 2nd to 4th intercostal lengths in females. The 'major' sternal lengths were more useful than the 'minor' ones for stature estimation The universal regression analysis used by Kanchan et al. [39] when applied to sternal lengths, gave satisfactory estimates of stature for males only but female stature was comparatively better estimated from simple linear regressions. But they are not proposed for the

Applied Regression Modeling A Business Approach

CERN Document Server

Pardoe, Iain

2012-01-01

An applied and concise treatment of statistical regression techniques for business students and professionals who have little or no background in calculusRegression analysis is an invaluable statistical methodology in business settings and is vital to model the relationship between a response variable and one or more predictor variables, as well as the prediction of a response value given values of the predictors. In view of the inherent uncertainty of business processes, such as the volatility of consumer spending and the presence of market uncertainty, business professionals use regression a
A novel simple QSAR model for the prediction of anti-HIV activity using multiple linear regression analysis.

Science.gov (United States)

Afantitis, Antreas; Melagraki, Georgia; Sarimveis, Haralambos; Koutentis, Panayiotis A; Markopoulos, John; Igglessi-Markopoulou, Olga

2006-08-01

A quantitative-structure activity relationship was obtained by applying Multiple Linear Regression Analysis to a series of 80 1-[2-hydroxyethoxy-methyl]-6-(phenylthio) thymine (HEPT) derivatives with significant anti-HIV activity. For the selection of the best among 37 different descriptors, the Elimination Selection Stepwise Regression Method (ES-SWR) was utilized. The resulting QSAR model (R (2) (CV) = 0.8160; S (PRESS) = 0.5680) proved to be very accurate both in training and predictive stages.
Comparison of beta-binomial regression model approaches to analyze health-related quality of life data.

Science.gov (United States)

Najera-Zuloaga, Josu; Lee, Dae-Jin; Arostegui, Inmaculada

2017-01-01

Health-related quality of life has become an increasingly important indicator of health status in clinical trials and epidemiological research. Moreover, the study of the relationship of health-related quality of life with patients and disease characteristics has become one of the primary aims of many health-related quality of life studies. Health-related quality of life scores are usually assumed to be distributed as binomial random variables and often highly skewed. The use of the beta-binomial distribution in the regression context has been proposed to model such data; however, the beta-binomial regression has been performed by means of two different approaches in the literature: (i) beta-binomial distribution with a logistic link; and (ii) hierarchical generalized linear models. None of the existing literature in the analysis of health-related quality of life survey data has performed a comparison of both approaches in terms of adequacy and regression parameter interpretation context. This paper is motivated by the analysis of a real data application of health-related quality of life outcomes in patients with Chronic Obstructive Pulmonary Disease, where the use of both approaches yields to contradictory results in terms of covariate effects significance and consequently the interpretation of the most relevant factors in health-related quality of life. We present an explanation of the results in both methodologies through a simulation study and address the need to apply the proper approach in the analysis of health-related quality of life survey data for practitioners, providing an R package.
Mediation analysis for logistic regression with interactions: Application of a surrogate marker in ophthalmology

DEFF Research Database (Denmark)

Jensen, Signe Marie; Hauger, Hanne; Ritz, Christian

2018-01-01

Mediation analysis is often based on fitting two models, one including and another excluding a potential mediator, and subsequently quantify the mediated effects by combining parameter estimates from these two models. Standard errors of such derived parameters may be approximated using the delta...... method. For a study evaluating a treatment effect on visual acuity, a binary outcome, we demonstrate how mediation analysis may conveniently be carried out by means of marginally fitted logistic regression models in combination with the delta method. Several metrics of mediation are estimated and results...
Evaluation agriculture morphological and analysis of yield components in twelve cañahua agreements (Chenopodium pallidicaule Aellen

Directory of Open Access Journals (Sweden)

Mayta-Mamani Adelio

2015-11-01

Full Text Available In the present investigation, they were carried out the analysis descriptive statistic, variance analysis, comparison of stockings of Duncan, multiple correlation and analysis of path coefficient, this I finish to determine the direct and indirect effects, in cañahua (Chenopodium pallidicaule Aellen. The study behaved at random under the design of blocks with four repetitions and twelve treatments (cañahua agreements, during the campaign agricultural 2009 and 2010. The results show, the grain production on the average general of 7.67 grain/plant, with a coefficient of variability of 23.1%, and with better yields they were the agreements 455, 222, ILLPA-INIA and CUPI, achieved to produce 12.28 grain/plant on the average, 10.58 grain/plant, 10.29 grain/plant and 10.16 grain/plant, respectively. Through the path analysis, selection approaches the components of yields were identified. In the agreement 455, the coefficient of determination of the system plants dear it reached 36.5%, and main yield components were constituted, number of branches, vegetable covering and height of the plant, with an association degree and direct effect of r=0.505 P=0.327, r=0.446 P=0.168 and r=0.417 P=0.196, respectively. In the agreement 222, the characters morphological that influence in the yield this explained in 40.5%, and as better yield component the number of branches has been constituted (r=0.462 (P=0.261, vegetable covering (r=0.514 (P=0.271 and height of the plant (r=0.383 (P=0.047. In the agreement ILLPA-INIA, it has been determined as better yield component, the number of branches (r=0.514 (P=0.318 and diameter of the shaft (r=0.479 (P=0.524, these characters in its group justify that the system plants defined total it was 44.7% of the direct influence. Agreement CUPI, in this cultivation the characters that influenced in the grain yield are implied in 36.4% and with better characters morphological that have been able to influence in a direct way they are
PREDICTION MODELS OF GRAIN YIELD AND CHARACTERIZATION

Directory of Open Access Journals (Sweden)

Narciso Ysac Avila Serrano

2009-06-01

Full Text Available With the objective to characterize the grain yield of five cowpea cultivars and to find linear regression models to predict it, a study was developed in La Paz, Baja California Sur, Mexico. A complete randomized blocks design was used. Simple and multivariate analyses of variance were carried out using the canonical variables to characterize the cultivars. The variables cluster per plant, pods per plant, pods per cluster, seeds weight per plant, seeds hectoliter weight, 100-seed weight, seeds length, seeds wide, seeds thickness, pods length, pods wide, pods weight, seeds per pods, and seeds weight per pods, showed significant differences (Pâ‰¤ 0.05 among cultivars. PaceÃ±o and IT90K-277-2 cultivars showed the higher seeds weight per plant. The linear regression models showed correlation coefficients â‰¥0.92. In these models, the seeds weight per plant, pods per cluster, pods per plant, cluster per plant and pods length showed significant correlations (Pâ‰¤ 0.05. In conclusion, the results showed that grain yield differ among cultivars and for its estimation, the prediction models showed determination coefficients highly dependable.
Improved Regression Analysis of Temperature-Dependent Strain-Gage Balance Calibration Data

Science.gov (United States)

Ulbrich, N.

2015-01-01

An improved approach is discussed that may be used to directly include first and second order temperature effects in the load prediction algorithm of a wind tunnel strain-gage balance. The improved approach was designed for the Iterative Method that fits strain-gage outputs as a function of calibration loads and uses a load iteration scheme during the wind tunnel test to predict loads from measured gage outputs. The improved approach assumes that the strain-gage balance is at a constant uniform temperature when it is calibrated and used. First, the method introduces a new independent variable for the regression analysis of the balance calibration data. The new variable is designed as the difference between the uniform temperature of the balance and a global reference temperature. This reference temperature should be the primary calibration temperature of the balance so that, if needed, a tare load iteration can be performed. Then, two temperature{dependent terms are included in the regression models of the gage outputs. They are the temperature difference itself and the square of the temperature difference. Simulated temperature{dependent data obtained from Triumph Aerospace's 2013 calibration of NASA's ARC-30K five component semi{span balance is used to illustrate the application of the improved approach.
Classification of Effective Soil Depth by Using Multinomial Logistic Regression Analysis

Science.gov (United States)

Chang, C. H.; Chan, H. C.; Chen, B. A.

2016-12-01

Classification of effective soil depth is a task of determining the slopeland utilizable limitation in Taiwan. The "Slopeland Conservation and Utilization Act" categorizes the slopeland into agriculture and husbandry land, land suitable for forestry and land for enhanced conservation according to the factors including average slope, effective soil depth, soil erosion and parental rock. However, sit investigation of the effective soil depth requires a cost-effective field work. This research aimed to classify the effective soil depth by using multinomial logistic regression with the environmental factors. The Wen-Shui Watershed located at the central Taiwan was selected as the study areas. The analysis of multinomial logistic regression is performed by the assistance of a Geographic Information Systems (GIS). The effective soil depth was categorized into four levels including deeper, deep, shallow and shallower. The environmental factors of slope, aspect, digital elevation model (DEM), curvature and normalized difference vegetation index (NDVI) were selected for classifying the soil depth. An Error Matrix was then used to assess the model accuracy. The results showed an overall accuracy of 75%. At the end, a map of effective soil depth was produced to help planners and decision makers in determining the slopeland utilizable limitation in the study areas.
Regression and kriging analysis for grid power factor estimation

Directory of Open Access Journals (Sweden)

Rajesh Guntaka

2014-12-01

Full Text Available The measurement of power factor (PF in electrical utility grids is a mainstay of load balancing and is also a critical element of transmission and distribution efficiency. The measurement of PF dates back to the earliest periods of electrical power distribution to public grids. In the wide-area distribution grid, measurement of current waveforms is trivial and may be accomplished at any point in the grid using a current tap transformer. However, voltage measurement requires reference to ground and so is more problematic and measurements are normally constrained to points that have ready and easy access to a ground source. We present two mathematical analysis methods based on kriging and linear least square estimation (LLSE (regression to derive PF at nodes with unknown voltages that are within a perimeter of sample nodes with ground reference across a selected power grid. Our results indicate an error average of 1.884% that is within acceptable tolerances for PF measurements that are used in load balancing tasks.
Computational Tools for Probing Interactions in Multiple Linear Regression, Multilevel Modeling, and Latent Curve Analysis

Science.gov (United States)

Preacher, Kristopher J.; Curran, Patrick J.; Bauer, Daniel J.

2006-01-01

Simple slopes, regions of significance, and confidence bands are commonly used to evaluate interactions in multiple linear regression (MLR) models, and the use of these techniques has recently been extended to multilevel or hierarchical linear modeling (HLM) and latent curve analysis (LCA). However, conducting these tests and plotting the…
Drug treatment rates with beta-blockers and ACE-inhibitors/angiotensin receptor blockers and recurrences in takotsubo cardiomyopathy: A meta-regression analysis.

Science.gov (United States)

Brunetti, Natale Daniele; Santoro, Francesco; De Gennaro, Luisa; Correale, Michele; Gaglione, Antonio; Di Biase, Matteo

2016-07-01

In a recent paper Singh et al. analyzed the effect of drug treatment on recurrence of takotsubo cardiomyopathy (TTC) in a comprehensive meta-analysis. The study found that recurrence rates were independent of clinic utilization of BB prescription, but inversely correlated with ACEi/ARB prescription: authors therefore conclude that ACEi/ARB rather than BB may reduce risk of recurrence. We aimed to re-analyze data reported in the study, now weighted for populations' size, in a meta-regression analysis. After multiple meta-regression analysis, we found a significant regression between rates of prescription of ACEi and rates of recurrence of TTC; regression was not statistically significant for BBs. On the bases of our re-analysis, we confirm that rates of recurrence of TTC are lower in populations of patients with higher rates of treatment with ACEi/ARB. That could not necessarily imply that ACEi may prevent recurrence of TTC, but barely that, for example, rates of recurrence are lower in cohorts more compliant with therapy or more prescribed with ACEi because more carefully followed. Randomized prospective studies are surely warranted. Copyright © 2016 Elsevier Ireland Ltd. All rights reserved.
Cubic-spline interpolation to estimate effects of inbreeding on milk yield in first lactation Holstein cows

Directory of Open Access Journals (Sweden)

Makram J. Geha

2011-01-01

Full Text Available Milk yield records (305d, 2X, actual milk yield of 123,639 registered first lactation Holstein cows were used to compare linear regression (y = β0 + β1X + e ,quadratic regression, (y = β0 + β1X + β2X2 + e cubic regression (y = β0 + β1X + β2X2 + β3X3 + e and fixed factor models, with cubic-spline interpolation models, for estimating the effects of inbreeding on milk yield. Ten animal models, all with herd-year-season of calving as fixed effect, were compared using the Akaike corrected-Information Criterion (AICc. The cubic-spline interpolation model with seven knots had the lowest AICc, whereas for all those labeled as "traditional", AICc was higher than the best model. Results from fitting inbreeding using a cubic-spline with seven knots were compared to results from fitting inbreeding as a linear covariate or as a fixed factor with seven levels. Estimates of inbreeding effects were not significantly different between the cubic-spline model and the fixed factor model, but were significantly different from the linear regression model. Milk yield decreased significantly at inbreeding levels greater than 9%. Variance component estimates were similar for the three models. Ranking of the top 100 sires with daughter records remained unaffected by the model used.
Evaluation of Hail Simulated Damage on Marketable Tuber Yield of Potato Agria Cultivar in Ardabil Region

Directory of Open Access Journals (Sweden)

D. Hassanpanah

2012-07-01

Full Text Available This study was conducted at Ardabil Agriculture and Natural Resources Research Station during the year of 2010. A factorial experiment based on randomized complete block design with four replications and two factors were used to evaluate the effect of simulated hail damage to foliage at different growth stages of potato Agria cultivar on marketable tuber yield. The first factor consisted of six levels of foliar damage (0, 20, 40, 60, 80 and 100 percent and the second factor of five levels of plant growth stages (2, 5, 8, 11 and 15 weeks after the growing. Analysis of variance showed that there were significant differences among plants for levels and times of hail damage and their interactions in terms of marketable tuber yield. Percentage of marketable yield reduction at early stages of vegetative growth (2 weeks after growing was minimal. Occurrence of hail damage at the tuberization and bulking stages (5, 8 and 11 weeks after growing severely reduced marketable tuber yield. While, its damage at late growing stages of (14 weeks after growing on tuber yield was not appreciable. Times of hail damage on marketable tuber yield reduction was calculated through the regression. Relative reduction of marketable tuber yield at the early stages of vegetative growth, due to hail damage, against non-marketable tuber yield was higher than of bulking stage.
Assessing risk factors for periodontitis using regression

Science.gov (United States)

Lobo Pereira, J. A.; Ferreira, Maria Cristina; Oliveira, Teresa

2013-10-01

Multivariate statistical analysis is indispensable to assess the associations and interactions between different factors and the risk of periodontitis. Among others, regression analysis is a statistical technique widely used in healthcare to investigate and model the relationship between variables. In our work we study the impact of socio-demographic, medical and behavioral factors on periodontal health. Using regression, linear and logistic models, we can assess the relevance, as risk factors for periodontitis disease, of the following independent variables (IVs): Age, Gender, Diabetic Status, Education, Smoking status and Plaque Index. The multiple linear regression analysis model was built to evaluate the influence of IVs on mean Attachment Loss (AL). Thus, the regression coefficients along with respective p-values will be obtained as well as the respective p-values from the significance tests. The classification of a case (individual) adopted in the logistic model was the extent of the destruction of periodontal tissues defined by an Attachment Loss greater than or equal to 4 mm in 25% (AL≥4mm/≥25%) of sites surveyed. The association measures include the Odds Ratios together with the correspondent 95% confidence intervals.
Sub-pixel estimation of tree cover and bare surface densities using regression tree analysis

Directory of Open Access Journals (Sweden)

Carlos Augusto Zangrando Toneli

2011-09-01

Full Text Available Sub-pixel analysis is capable of generating continuous fields, which represent the spatial variability of certain thematic classes. The aim of this work was to develop numerical models to represent the variability of tree cover and bare surfaces within the study area. This research was conducted in the riparian buffer within a watershed of the São Francisco River in the North of Minas Gerais, Brazil. IKONOS and Landsat TM imagery were used with the GUIDE algorithm to construct the models. The results were two index images derived with regression trees for the entire study area, one representing tree cover and the other representing bare surface. The use of non-parametric and non-linear regression tree models presented satisfactory results to characterize wetland, deciduous and savanna patterns of forest formation.
Evaluation of yield and yield components and some agronomic traits of white bean genotypes under Karaj climate

Directory of Open Access Journals (Sweden)

M. Ebrahimi

2016-04-01

Full Text Available In order to study compatibility of 30 white bean genotypes under Karaj climate, an experiment was conducted based on randomized Complete Block Design with four replications. Evaluation and statistical analysis was performed for 18 important traits. Analysis of variance results showed that there are significant differences between varieties for all traits. Results of genotypes means comparison with Duncan’s multiple range test showed that genotype No. 29 was better than others in plant height, yield, and biological yield, seed no. per plant and pod weight traits. Simple correlation coefficients were significant between yield and weight of pod, biological yield, number of seed per plant, number of pod per plant, plant height, width of pod and number of seed per pod. Only pod length was negative correlated with yield between all investigated traits. Cluster analysis with UPGMA method arrangement genotypes into three groups. According to this experiment results we can recommend 29 and 30 genotypes for Karaj condition
Multiple predictor smoothing methods for sensitivity analysis: Description of techniques

International Nuclear Information System (INIS)

Storlie, Curtis B.; Helton, Jon C.

2008-01-01

The use of multiple predictor smoothing methods in sampling-based sensitivity analyses of complex models is investigated. Specifically, sensitivity analysis procedures based on smoothing methods employing the stepwise application of the following nonparametric regression techniques are described: (i) locally weighted regression (LOESS), (ii) additive models, (iii) projection pursuit regression, and (iv) recursive partitioning regression. Then, in the second and concluding part of this presentation, the indicated procedures are illustrated with both simple test problems and results from a performance assessment for a radioactive waste disposal facility (i.e., the Waste Isolation Pilot Plant). As shown by the example illustrations, the use of smoothing procedures based on nonparametric regression techniques can yield more informative sensitivity analysis results than can be obtained with more traditional sensitivity analysis procedures based on linear regression, rank regression or quadratic regression when nonlinear relationships between model inputs and model predictions are present
Dual Regression

OpenAIRE

Spady, Richard; Stouli, Sami

2012-01-01

We propose dual regression as an alternative to the quantile regression process for the global estimation of conditional distribution functions under minimal assumptions. Dual regression provides all the interpretational power of the quantile regression process while avoiding the need for repairing the intersecting conditional quantile surfaces that quantile regression often produces in practice. Our approach introduces a mathematical programming characterization of conditional distribution f...
Multiple predictor smoothing methods for sensitivity analysis: Example results

International Nuclear Information System (INIS)

Storlie, Curtis B.; Helton, Jon C.

2008-01-01

The use of multiple predictor smoothing methods in sampling-based sensitivity analyses of complex models is investigated. Specifically, sensitivity analysis procedures based on smoothing methods employing the stepwise application of the following nonparametric regression techniques are described in the first part of this presentation: (i) locally weighted regression (LOESS), (ii) additive models, (iii) projection pursuit regression, and (iv) recursive partitioning regression. In this, the second and concluding part of the presentation, the indicated procedures are illustrated with both simple test problems and results from a performance assessment for a radioactive waste disposal facility (i.e., the Waste Isolation Pilot Plant). As shown by the example illustrations, the use of smoothing procedures based on nonparametric regression techniques can yield more informative sensitivity analysis results than can be obtained with more traditional sensitivity analysis procedures based on linear regression, rank regression or quadratic regression when nonlinear relationships between model inputs and model predictions are present
Dispersive analysis of {omega}{yields}3{pi} and {phi}{yields}3{pi} decays

Energy Technology Data Exchange (ETDEWEB)

Niecknig, Franz; Kubis, Bastian; Schneider, Sebastian P. [Universitaet Bonn, Helmholtz-Institut fuer Strahlen- und Kernphysik (Theorie) and Bethe Center for Theoretical Physics, Bonn (Germany)

2012-05-15

We study the three-pion decays of the lightest isoscalar vector mesons, {omega} and {phi}, in a dispersive framework that allows for a consistent description of final-state interactions between all three pions. Our results are solely dependent on the phenomenological input for the pion-pion P-wave scattering phase shift. We predict the Dalitz plot distributions for both decays and compare our findings to recent measurements of the {phi}{yields}3{pi} Dalitz plot by the KLOE and CMD-2 collaborations. Dalitz plot parameters for future precision measurements of {omega}{yields}3{pi} are predicted. We also calculate the {pi}{pi} P-wave inelasticity contribution from {omega}{pi} intermediate states. (orig.)

Some links on this page may take you to non-federal websites. Their policies may differ from this site.