Pre-registering a personal experiment, seriously posts 91–120
This is a continuation of a long topic, addressed by post number rather than by page. Start at post 1.
Sample size in n-of-1: you are the sample. Repeated measurements (weekly weighings, daily mood scores) increase the power to detect a real effect even though n=1.
Collapsed as off-topic by two members at trust level 3 or above
post #92 is right about the mechanism and I think understates the practical bit.
Generalisability: a robust n-of-1 result applies to you. It does not tell you much about whether the effect generalises to others similar to you, much less to people different from you.
Worth separating two things that post #90 runs together.
Two things before anyone answers the substance.
First, the context in the first post is clear and specific. Second, the question is framed so that an answer can actually address it. Both are the norm here and both matter more than they sound.
Stopping rules: decide in advance when you will stop measuring (after a defined duration, after a defined number of measurements, or after a defined condition is met). Not deciding in advance means stopping when the result satisfies you, which is bias.
For anyone arriving from a search: the marked solution above is the direct answer, and the replies underneath it add the caveats that make it safe to use.
On post #94 — agreed on the reasoning, with one qualification.
Designing a personal experiment that could actually change your mind: that is the standard for an n-of-1 design. An experiment designed so that any result confirms what you already believed has not changed anything.
This follows post #96 rather than contradicting it.
Blinding: blinding yourself (not knowing which condition you are in) removes expectation bias. This is hard to do with these compounds (the appetite suppression is hard to miss) but partial blinding is possible (measuring something objective without knowing whether you took it today).
I read post #98 twice before replying, because I had assumed the opposite.
Two things before anyone answers the substance.
First, the context in the first post is clear and specific. Second, the question is framed so that an answer can actually address it. Both are the norm here and both matter more than they sound.
Washout periods: after stopping a medication, how long does it take for the effect to wash out? For compounds with a week-long half-life, roughly a month is needed to reach baseline. Using that washout period in a before-after design strengthens the inference.
Blinding: blinding yourself (not knowing which condition you are in) removes expectation bias. This is hard to do with these compounds (the appetite suppression is hard to miss) but partial blinding is possible (measuring something objective without knowing whether you took it today).
post #102 is right about the mechanism and I think understates the practical bit.
Stopping rules: decide in advance when you will stop measuring (after a defined duration, after a defined number of measurements, or after a defined condition is met). Not deciding in advance means stopping when the result satisfies you, which is bias.
Worth separating two things that post #100 runs together.
Having read the exchange above, I think I was wrong earlier in this topic and I want to say so plainly rather than quietly editing.
The correction was fair and I had been repeating something I had not checked carefully enough.
Picking up post #102: that is the part I would want checked first.
Statistical analysis of n-of-1 data: comparing before versus after with a t-test or similar is one approach. Plotting the data visually is another. Both are valid.
Confounding in personal experiments: other things change when you start a medication (season, exercise, diet, stress). Documenting those confounders helps you understand their contribution to the result.
On post #104 — agreed on the reasoning, with one qualification.
Sample size in n-of-1: you are the sample. Repeated measurements (weekly weighings, daily mood scores) increase the power to detect a real effect even though n=1.
This follows post #106 rather than contradicting it.
Objective versus subjective measures: subjective measures (how you feel) are vulnerable to bias. Objective measures (weight, strength on a specific exercise) are less vulnerable but not immune.
I read post #108 twice before replying, because I had assumed the opposite.
Generalisability: a robust n-of-1 result applies to you. It does not tell you much about whether the effect generalises to others similar to you, much less to people different from you.
When to run an n-of-1: this design works when you want to know whether a treatment works for you, not whether it works in general. For that purpose, it is efficient.
Designing a personal experiment that could actually change your mind: that is the standard for an n-of-1 design. An experiment designed so that any result confirms what you already believed has not changed anything.
Thank you for the correction. I have edited my earlier post with a note rather than silently, so the thread still makes sense to read. The error was mine and it was the kind that comes from remembering a figure instead of looking it up.
When to run an n-of-1: this design works when you want to know whether a treatment works for you, not whether it works in general. For that purpose, it is efficient.
Generalisability: a robust n-of-1 result applies to you. It does not tell you much about whether the effect generalises to others similar to you, much less to people different from you.
Coming back to post #115, because the follow-up matters more than the original answer.
Having read the exchange above, I think I was wrong earlier in this topic and I want to say so plainly rather than quietly editing.
The correction was fair and I had been repeating something I had not checked carefully enough.
Picking up post #115: that is the part I would want checked first.
Stopping rules: decide in advance when you will stop measuring (after a defined duration, after a defined number of measurements, or after a defined condition is met). Not deciding in advance means stopping when the result satisfies you, which is bias.
post #119 is right about the mechanism and I think understates the practical bit.
Objective versus subjective measures: subjective measures (how you feel) are vulnerable to bias. Objective measures (weight, strength on a specific exercise) are less vulnerable but not immune.
This topic was referenced in
- An ABAB design with a data table and honest limitations — a second datasetResearch Methods › N-of-1 designs · 118 replies
- Second pass at: Designing a personal experiment that could change your mindResearch Methods › N-of-1 designs · 44 replies
Suggested topics
| Topic | Participants | Replies | Views | Activity |
|---|---|---|---|---|
|
Designing a personal experiment that could change your mind
Posting this under the heading it deserves: Designing a personal experiment that could change your mind Everything below is what sits behind that. Session topic: STEP 8 ( JAMA , 2022). Please read it before…
|
+2 | 6 | 695 | 6mo |
|
Washout with a one-week half-life: the arithmetic — does this still hold?
Washout with a one-week half-life: the arithmetic — does this still hold? — that is the question, and I have not found it answered plainly anywhere I have looked. Comparing SURPASS-2 ( N Engl J Med , 2021)…
|
+41 | 45 | 12k | 1d |
|
Blinding yourself: practical methods and their limits — the long version
Blinding yourself: practical methods and their limits — the long version Writing it up because I had to work it out twice and would rather nobody else did. I have seen SURMOUNT-OSA ( N Engl J Med , 2024)…
|
+7 | 11 | 1.8k | 22h |
|
An ABAB design with a data table and honest limitations
An ABAB design with a data table and honest limitations Writing it up because I had to work it out twice and would rather nobody else did. Session topic: SURPASS-4 ( Lancet , 2021). Please read it before…
|
2 | 45k | 5mo | |
|
An ABAB design with a data table and honest limitations — a second dataset
On the subject in the title: An ABAB design with a data table and honest limitations — a second dataset Working notes rather than a conclusion. Session topic: STEP 2 ( Lancet , 2021). Please read it before…
|
+106 | 118 | 25k | 4mo |
Related topics — sharing the tags worked example, data table, effect size
| Topic | Participants | Replies | Views | Activity |
|---|---|---|---|---|
|
Carryover and the ghost peak from last week's standard — what changed since
Carryover and the ghost peak from last week's standard — what changed since Writing it up because I had to work it out twice and would rather nobody else did. Posting the method first, because I know what the…
|
+98 | 114 | 37k | 22d |
|
Journal club: SURMOUNT-4 and continuation versus withdrawal — a second dataset
Posting this under the heading it deserves: Journal club: SURMOUNT-4 and continuation versus withdrawal — a second dataset Everything below is what sits behind that. Comparing STEP 8 ( JAMA , 2022) with FLOW…
|
+58 | 64 | 17k | 13mo |
|
List price versus negotiated price versus what you pay
On the subject in the title: List price versus negotiated price versus what you pay Working notes rather than a conclusion. Documenting an access outcome, dated, because everything in this category expires.…
|
+99 | 115 | 11k | 3mo |
|
Preprint servers and what screening they do — the long version
On the subject in the title: Preprint servers and what screening they do — the long version Working notes rather than a conclusion. Session topic: PIONEER 6 ( N Engl J Med , 2019). Please read it before…
|
+7 | 11 | 26k | 4mo |
|
Coming back to: Column chemistry choices for a 40-residue peptide
On the subject in the title: Column chemistry choices for a 40-residue peptide Working notes rather than a conclusion. Posting the method first, because I know what the first three replies will otherwise be.…
|
2 | 43k | 16mo |