%svy_freqs: A generic SAS macro for cross-tabulation between a factor and a by-group variable given a third variable and creating publication-quality tables using data from complex surveys
Muthusi, J.; Mwalili, S.; Young, P.
Show abstract
IntroductionIn epidemiological studies, cross-tabulations are a simple but important tool for understanding the distribution of socio-demographic characteristics among study participants. They become more useful when comparisons are presented using a by-group variable such as key demographic characteristic or an outcome status; for instance, sex or the presence or absence of a disease status. Most available statistical analysis software can easily perform cross-tabulations, however, output from these must be processed further to make it readily available for review and use in a publication. In addition, performing three-way cross-tabulations of complex survey data such as those required to show the distribution of disease prevalence across multiple factors and a by-group variable is not easily implemented directly using available standard procedures of commonly used statistical software.\n\nMethodsWe developed a generic SAS macro, %svy_freqs, to create quality publication-ready tables from cross-tabulations between a factor and a by-group variable given a third variable using survey or non-survey data. The SAS macro also performs classical two-way cross-tabulations and refines output into publication-quality tables. It provides extra features not available in existing procedures such as ability to incorporate parameters for survey design and replication-based variance estimation methods, performing validation checks for input parameters, transparently formatting character variable values into numeric ones and allowing for generalizability.\n\nResultsWe demonstrate the application of the SAS macro in the analysis of data from the 2013-2014 National Health and Nutrition Examination Survey (NHANES), a complex survey designed to assess the health and nutritional status of adults and children in the United States (U.S.).\n\nConclusionThe SAS code use to develop the macro is simple yet comprehensive, easy to follow, straightforward for the end user and simple for a SAS programmer to extend. The SAS macro has shown to shorten turn-around time for statistical analysis, eliminate errors when preparing output, and support reproducible research.
Matching journals
The top 1 journal accounts for 50% of the predicted probability mass.
Similar papers in this journal
- %svy_logistic_regression: A generic SAS(R) macro for simple and multiple logistic regression and creating quality publication-ready tables using survey or non-survey data 97%
- Common misconceptions held by health researchers when interpreting linear regression assumptions, a cross-sectional study 96%
- Introducing riskCommunicator: an R package to obtain interpretable effect estimates for public health 94%
Similar papers in this journal
Similar papers in this journal
- Epidemiological Correlates of Overweight and Obesity in the Northern Cape Province, South Africa 93%
- A machine learning approach for identification of gastrointestinal predictors for the risk of COVID-19 related hospitalization 91%
- covid19.Explorer : A web application and R package to explore United States COVID-19 data 90%
Similar papers in this journal
- dsSurvival: Privacy preserving survival models for federated individual patient meta-analysis in DataSHIELD 90%
- The Aliment to Bodily Condition knowledgebase (ABCkb): A database connecting plants and human health 90%
- Understanding bias when estimating life expectancy from age at death: A simulation approach applied to Morquio Syndrome A 90%
Similar papers in this journal
"Similar papers" are the closest papers from that journal in the model's embedding space. They show what the match is built on, but the ranking comes mostly from a classifier over the whole training set, not from these examples alone.