Confidence intervals: rather than a single point estimate, a range of plausible values. A narrow interval means precise measurement; a wide interval means measurement is imprecise. Wider intervals (more uncertainty) are honest about limitation.
Measurement error in home scales, with a worked standard deviation — the long version
Absence of evidence and evidence of absence: if a study is small and finds no effect, that is absence of evidence, not evidence of absence. A larger study might find an effect that a small study missed.
Worth separating two things that post #33 runs together.
Multiplicity and multiple comparisons: if you test many hypotheses, the chance of finding a false positive by random chance increases. That is why pre-specifying the primary hypothesis matters.
For anyone arriving from a search: the marked solution above is the direct answer, and the replies underneath it add the caveats that make it safe to use.
post #55 answers the question as asked. The question underneath it is different.
Confounding: a third variable explains an apparent association. In randomised data, randomisation balances confounders. In observational data, confounders can be adjusted for but unknown ones cannot.
post #73 is right about the mechanism and I think understates the practical bit.
Power and sample size: a study might be too small to detect a real effect (low power). Sample size calculations help determine how many participants are needed to detect an effect of a given magnitude.
Worth separating two things that post #84 runs together.
Two things before anyone answers the substance.
First, the context in the first post is clear and specific. Second, the question is framed so that an answer can actually address it. Both are the norm here and both matter more than they sound.
Regression to the mean: if you select people with extreme values (very high or very low), their next measurement is often less extreme just by chance. This can look like a treatment effect when it is just statistics.
Read the full topic (106 posts)
This topic was referenced in
- [2026 update] Correlation in a self-tracked dataset: what it can supportResearch Methods › Statistics · 2 replies
Suggested topics
| Topic | Participants | Replies | Views | Activity |
|---|---|---|---|---|
|
[2026 update] Correlation in a self-tracked dataset: what it can support
Posting this under the heading it deserves: Correlation in a self-tracked dataset: what it can support Everything below is what sits behind that. I have seen SURMOUNT-4 ( JAMA , 2024) cited in support of a…
|
2 | 31k | 6h | |
|
What a confidence interval means, from scratch — the long version
The question in the title: What a confidence interval means, from scratch — the long version I will give what I have already checked below so nobody repeats it. I have seen SURMOUNT-2 ( Lancet , 2023) cited…
|
+16 | 20 | 2.3k | 8mo |
|
About the Statistics category
Effect sizes, intervals, multiplicity, and the difference between absent and undetected. This post is a community wiki: any member at trust level 3 or above can edit it, and every edit is recorded with its…
|
+4 | 8 | 2.3k | 7mo |
|
Multiplicity when you track fifteen variables
Multiplicity when you track fifteen variables — setting out what I have, and where I think it stops being reliable. Comparing LEADER ( N Engl J Med , 2016) with SURMOUNT-OSA ( N Engl J Med , 2024) and finding…
|
2 | 60k | 11mo | |
|
Measurement error in home scales, with a worked standard deviation — a second dataset
Measurement error in home scales, with a worked standard deviation — a second dataset — setting out what I have, and where I think it stops being reliable. Comparing FLOW ( N Engl J Med , 2024) with SURPASS-2…
|
2 | 13k | 7mo |
Related topics — sharing the tags number needed to treat, worked example, heterogeneity
| Topic | Participants | Replies | Views | Activity |
|---|---|---|---|---|
|
Revisiting: Journal club: SURMOUNT-OSA and a hard endpoint in a soft field
Revisiting: Journal club: SURMOUNT-OSA and a hard endpoint in a soft field — setting out what I have, and where I think it stops being reliable. Comparing LEADER ( N Engl J Med , 2016) with SUSTAIN 6 ( N Engl…
|
+108 | 122 | 52k | 7d |
|
Second pass at: Absence of evidence and evidence of absence
On the subject in the title: Second pass at: Absence of evidence and evidence of absence Working notes rather than a conclusion. Session topic: SELECT ( N Engl J Med , 2023). Please read it before posting;…
|
+50 | 55 | 18k | 16h |
|
What to include when your question involves a chromatogram — one year on
Asking directly, because I could not find a straight answer: What to include when your question involves a chromatogram — one year on Question in the title. Context below, and I have tried to include the…
|
+69 | 75 | 862 | 14d |
|
Semaglutide in people without diabetes: what the evidence base looks like
Semaglutide in people without diabetes: what the evidence base looks like Writing it up because I had to work it out twice and would rather nobody else did. Session topic: SUSTAIN 6 ( N Engl J Med , 2016).…
|
+2 | 6 | 18k | 17mo |
|
Coming back to: Prediction intervals and why they are more honest than confidence intervals
Prediction intervals and why they are more honest than confidence intervals Writing it up because I had to work it out twice and would rather nobody else did. Comparing SURMOUNT-OSA ( N Engl J Med , 2024)…
|
2 | 65 | 22mo |