Linear regression of sampling distributions of the mean
Torres, D. J.; Vasilic, A.; Pacheco, J.
Show abstract
We show that the simple and multiple linear regression coefficients and the coefficient of determination R2 computed from sampling distributions of the mean (with or without replacement) are equal to the regression coefficients and coefficient of determination computed with individual data. Moreover, the standard error of estimate is reduced by the square root of the group size for sampling distributions of the mean. The result has applications when formulating a distance measure between two genes in a hierarchical clustering algorithm. We show that the Pearson R coefficient can measure how differential expression in one gene correlates with differential expression in a second gene.
Matching journals
The top 6 journals account for 50% of the predicted probability mass.
Similar papers in this journal
- Theoretical properties of nearest-neighbor distance distributions and novel metrics for high dimensional bioinformatics data 97%
- Time Series Experimental Design Under One-Shot Sampling: The Importance of Condition Diversity 96%
- Theory on the rate equations of Michaelis-Menten type enzyme kinetics with competitive inhibition 95%
Similar papers in this journal
- Statistical inference of the rates of cell proliferation and phenotypic switching in cancer 96%
- Reliable and efficient parameter estimation using approximate continuum limit descriptions of stochastic models 96%
- Age and Generation-Based Model of Metastatic Cancer: From Micrometastases to Macrometastases 95%
Similar papers in this journal
Similar papers in this journal
Similar papers in this journal
- GIFT: New method for the genetic analysis of small gene effects involving small sample sizes. 95%
- Evolutionary stability of social interaction rules in collective decision-making 94%
- Urn models for regulated gene expression yield physically intuitive solutions for probability distributions of single-cell counts 93%
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.