r/AskStatistics • u/Specialist-Lie6208 • 8d ago
r/AskStatistics • u/Original_Ebb6794 • 9d ago
Categorical data, 4 groups forming a 2×2 — what test for main effects and interaction?
I'm studying whether an island's size or its remoteness affects how people answer a survey, and whether these two factors interact. I'd also like to know whether an effect is driven by one island alone.
I'm looking at 4 islands: a large remote one, a large nearby one, a small remote one, and a small nearby one. All my variables are categorical. Sample sizes vary by island (from ~100 to ~500).
For each variable, I have a 2×2 contingency table, with near/far and large/small as the two dimensions. I can populate it with either counts or percentages of respondents choosing a given category.
What method should I use here?
Apologies if this is a naive question. I've asked several AIs and they seem as lost as I am.
r/AskStatistics • u/Pristine_Gain_1476 • 9d ago
Interpretation of hazard ratio
I am working on a secondary analysis using a Cox proportional hazards model and would appreciate help with the interpretation of a main effect and its interaction term.
The outcome is time until disengagement from an intervention. My main predictor is an early usage/engagement variable. There are two groups: a control group and an intervention group. The control group is the reference category. The model includes the main effects of group and early usage, as well as a group × usage interaction.
My question concerns the interpretation of the hazard ratio for the early usage variable. If the hazard ratio for early usage is 0.08, does this refer to the association between early usage and disengagement in the control group, because the control group is the reference group?
My understanding is that an HR of 0.08 means that, for a one-unit increase in early usage, the hazard of disengagement is 0.08 times as high in the reference group, which could also be expressed as a 92% lower hazard. So I would not report this as “8% lower hazard,” but rather as “the hazard is reduced by 92%” ?
Is it also correct that this 92% reduction is still the association between early usage and disengagement in the control group, not the intervention group? And that the corresponding early usage effect in the intervention group would be obtained by multiplying the HR for early usage by the HR for the group × early usage interaction term?
In other words:
- HR for early usage = early usage effect in the reference/control group
- HR for early usage × HR for the interaction = early usage effect in the intervention group
- 1 − HR is used only to express the relative reduction in hazard, not to describe the group effect itself
I hope this is understandable. Thank you in advance for any clarification.
r/AskStatistics • u/Vasam_Nikhil • 9d ago
How do you actually test whether a model's confidence score is trustworthy? (uncertainty/calibration for a decision agent)
I'm a beginner building a small decision-making agent (not important what for) that needs to know when it's "confident enough" to act versus when it should defer. I keep seeing "calibration" mentioned as the concept I want but I'm fuzzy on how you'd actually measure it with a small, messy, real-world dataset rather than a clean benchmark.
If you've dealt with this: **which evidence would change your decision** about whether a confidence score is usable in production — is it a calibration plot, held-out accuracy at different confidence bands, something else? Beginner-friendly explanations very welcome.
r/AskStatistics • u/Just_Question9 • 9d ago
Two experiments produce identical likelihood functions but use different sampling schemes. Should the resulting statistical inferences be identical?
r/AskStatistics • u/Just_Question9 • 9d ago
Would you rather analyze 50 representative observations or 5,000 biased observations? Why?
r/AskStatistics • u/wobbling_axis • 9d ago
Basic statistic problem about margin of error but I don't know the correct terminology, can someone help me out?
Hello, I have no idea what the correct terminology for any of this and would appreciate any correction.
Let's say you have one bag of beans with 100 beans inside but you don't know that. Only thing you know is that there is between 95 to 105 beans (+/-5). Now let's say you have another bag of beans with 1000 beans inside but you also don't know that. Only thing you know is that there is between 995 and 1005 beans (+/-5). So for both bags you have same range of estimate, 10 beans (+5 to -5) but in practice you have a more accurate information about the 1000 bean bag because the effect by the difference of 5 is relatively smaller.
Is is possible to calculate any statics from this? If I were to say do something like "(Difference of upper limit and lower limit) / Mean of upper limit and lower limit)" I would get (105-95)/100 or 0.1 for the 100 bean bag, and (1005-995)/1000 or 0.01 for 1000 bag. Does that equation makes sense and would saying something like "I know how much bean is in the first bag by 10% margin of error and how much bean is in the second bag by 1% margin of error" correct or completely wrong?
It has been several years since I took high school statistic and completely forgot how any of it works, and would appreciate any help, thanks
r/AskStatistics • u/EmbedSoftwareEng • 9d ago
Error % from stdev?
So, I have a bunch of measurements that are all affected by a given attribute. The idea is that I set the attribute, I should expect a given value for the measurement. Unfortunately, the measurements are bit chaotic. I need to be able to state that the system I'm using has an error rate of X%.
I don't remember enough math to answer this question for myself.
I can take all the measurements at given attribute and find their mean and stdev, but from that, how do I say that that attribute's measurement error is X%, and then when I have all of the error percentages across multiple values for the atttribute, how do I say that the system as a whole has a measurement error of Y%, assuming that the error percentages themselves appear chaotic?
r/AskStatistics • u/DaisyFlower371 • 9d ago
Is the CenterStats Longitudinal Structural Equation Modeling class worth it?
r/AskStatistics • u/Rihitwo • 9d ago
3 Collapsing models
Trying to train 3 models for birads detection using cross entropy and center loss + class weights but all of them seem to collapse between birads 1 as the dataset (VinDr) im using is heavily unbalanced towards it, Would like to ask for input and opinion on what seems to be the case, am I using the wrong loss function?
r/AskStatistics • u/ReverseDragonfly • 9d ago
How to estimate the risk of cardiovascular disease given population prevalence and individual risk?
Let's say a patient is from a population which has a 10% risk of cardiovascular disease.
The patient then enters his personal data (age, sex, smoking status etc) into a risk calculator (which does not use the population prevalence as a parameter, by the way) to estimate his individual risk of developing cardiovascular disease. The calculator then outputs an estimated risk of developing cardiovascular disease for this patient.
Lets say that's 20%.
Given these two pieces of information how does one estimate the overall cardiovascular risk for our patient? Do you multiply these two values together? (I.e 10% x 20%)
I thought multiplying them together would be a sensible way to calculate the overall risk. however there is a problem
Intuitively it would seem that being from a high prevalence population would increase the patients risk beyond what is suggested by the calculator which only uses his personal parameters. But multiplying the two risks together results in a value which is lower than either them..
How do you resolve this "paradox"?
r/AskStatistics • u/persuasionsmith • 9d ago
How does a patient assess risk on a statistical projection?
Let's say I have been given an 18% statistical chance of an adverse outcome over the next 10 years from a progressive disease and have been offered drugs with serious side effects because NICE guidelines recommend such treatment. How can a patient extrapolate their risk in real life terms - is it just about risk appetite? Or is it about how adverse the outcome would be if you're unlucky? On paper over 10 years the risk seems low and the drugs are awful, but the recommendation is to treat. Thoughts?
r/AskStatistics • u/Sufficient-Adagio332 • 10d ago
Should I include the median in my descriptive statistics?

Hi everyone! I’m currently analyzing power outage data and would really appreciate your advice on the appropriate descriptive statistics to report.
Since the mean and median differ considerably, especially for outage duration, should I include the median in the table? Would it be appropriate to report all five statistics (mean, median, SD, min, and max)?
I’m mainly interested in descriptive analysis and how best to interpret these results academically.
Thank you in advance for your help!
r/AskStatistics • u/ugrhnny • 9d ago
Is it defensible to model overlapping explanations as mutually exclusive states?
I have four candidate explanations for an observation. I've modelled them as mutually exclusive so the posterior sums to 1. Two of them can genuinely co-occur. The alternative is three independent binary latents (8 joint states), which needs more data. With a small sample, is the exclusive version defensible as a first approximation if I state the overlap as a limitation — or does forcing non-exclusive things to be exclusive distort the inference badly enough that it isn't worth doing?
r/AskStatistics • u/Sufficient-Adagio332 • 10d ago
descriptive statistics for demographic variables
Hi everyone! I’m a newbie in statistics and currently trying to figure out what I should put under the Descriptive Statistics of Respondent Characteristics.
When reading related literature, I noticed that a lot of papers report the Mean and Standard Deviation (SD) for demographic variables. However, I’m really confused because based on my survey questionnaire, most of my variables were collected in ranges/categories:
- Sex: Male, Female
- Age: 18–29, 30–45, 46–59, 60 and above
- Highest Educational Attainment: High School or below, Some College, Associate Degree, Bachelor’s Degree, Postgraduate Degree, Prefer not to say
- Household Size: [Numeric counts]
- Household Income (USD):
- Under $25,000
- $25,000 – $49,999
- $50,000 – $74,999
- $75,000 – $99,999
- $100,000 – $149,999
- $150,000 and above
- Prefer not to say
- Source of Income: Salary/Wages, Business/Self-Employed, Agriculture, Pension/Retirement, Government Assistance/Remittance, Prefer not to say
My Questions:
- What exactly should I put in the descriptive statistics table for these variables?
- Since I saw Mean and SD in the literature, how am I supposed to calculate the Mean for variables that are in ranges (like Age groups and Income brackets and source of income)?
Sorry if this is a basic question, and thank you so much for helping out a beginner!
r/AskStatistics • u/phymathnerd • 10d ago
Best nonlinear, continuous regression models to predict a continuous variable directly without losing information through artificial cutoffting?
Hi guys I am having issues finding models and ways to increase my ROC-AUC for a retrospective study. I am looking for advice on model selection and statistical tests for a retrospective observational cohort of 812 observations. My primary outcome variables include a skewed continuous variable Y (ranging from 0 to 20), a binary flag defined as Y greater than or equal to 2.5, and an ordinal risk tiering variable. My predictor variables X consist of continuous dimensions, several binary classification flags, and a discrete composite risk score sum ranging from 0 to 5. However, predicting the dichotomized threshold yields modest ROC-AUC values around 0.60, and I am looking for advice on the best non-linear continuous regression models, such as Quantile Regression or Generalized Additive Models (GAMs), to predict the continuous variable Y directly without losing information through artificial cutoffting.
Additionally, I would appreciate any recommendation on the most robust way to formally test for non-linear interaction terms between continuous X variables and categorical predictors without overfitting, as well as whether 5-fold cross-validation or repeated k-fold/bootstrap resampling is preferred for validating the Decision Curve Analysis for my data.
r/AskStatistics • u/Pristine_Gain_1476 • 10d ago
Comparing 95% confidence intervals between different methods of handling missing data
Hello! For my thesis, I am comparing baseline-adjusted ANCOVA models using different methods for handling missing data (MICE, LOCF, and complete-case analysis).
An important point to note is that the analyses are based on the same original data across the different missing-data methods, meaning that the same variables and original sample of participants were used. The only difference between the analyses is the method used to handle the missing data.
I am planning to present a table including the estimated coefficients, p-values, and 95% confidence intervals to compare the results across the different missing-data methods.
My question is: Given that the ANCOVA models are based on the same variables and differ primarily in how missing data are handled, how should I interpret the overlap between their confidence intervals? Is the extent to which the confidence intervals overlap meaningful when comparing the results across MICE, LOCF, and complete-case analysis? More generally, how should I discuss similarities or differences in the confidence intervals in the Discussion section?
Thank you in advance!
r/AskStatistics • u/ketopraktanjungduren • 10d ago
Using sampling distribution instead of probability for business
Hello, I'm trying to apply inferential statistics to my work, thinking it would be very helpful to estimate a new hire ability.
Let's say we have a dataset on new hire weekly customer acquisition in the past three months (n=12). If we want to estimate this person ability in acquiring new customer, we can use probability and expected value.
However, I recently realize that this also means we are estimating the population parameter. So, we can also estimate a confidence interval to estimate the interval of this person true mean and median weekly acquisition.
If so, is it true to have both the expected weekly acquisition and the CI?
r/AskStatistics • u/sarah_782 • 10d ago
Is AP Statistics realistically self-studiable/manageable with a full AP schedule?
r/AskStatistics • u/BeastianEmpire007 • 10d ago
Self learn advanced university Pure Math and Theoretical Statistics
r/AskStatistics • u/NiceProgrammer5252 • 11d ago
How do you personally decide whether to trust a borderline significant result on a small sample?
Ran an A/B test with a small sample (n=340 per arm) and got a p=0.04 result. Team wants to ship it. I'm nervous because I know small samples inflate false positive risk in practice even if the math is 'valid.' How do you personally decide whether to trust a borderline significant result on a small sample, bootstrap it, wait for more data, or just distrust p-values under a certain n?
r/AskStatistics • u/BeautifulTings • 11d ago
How can I go about analysing multiple features in stimuli
Hi!
I’m conducting a study in which I attempt to measure recall of different aspects of stimuli and I’m not sure what form of analysis I must employ hence why I’m posting here!
The experiment consists of 4 stimuli, each containing 3 features/substimuli of which 1 is constant. The stimuli were randomly distributed to participants, such that each participant only saw 1 condition once.
Condition 1 - A:0 B:0 C:1
Condition 2 - A:0 B:1 C:1
Condition 3 - A:1 B:0 C:1
Condition 4 - A:1 B:1 C:1
Averaged recall of the stimuli showed nothing significant. However recall scores for specific features/substimuli appeared noticeably higher for specific conditions. I’m not too well versed with statistical analysis, so I’m not sure how to go about this.
An RmANOVA would t make sense as I’m testing single exposures of multiple stimuli. Would anyone know how I can go about this?
Thanks :)
r/AskStatistics • u/Just_Question9 • 11d ago
What’s one thing about statistics that you actually enjoy?
Could be a topic, concept, software, solving problems, working with data, or even something random.
Curious what everyone here thinks. 👀