Limitations of datasets: all community-collected data has limitations. The population is self-selected (people in this community are not representative of all people using these compounds). Reporting bias is real (remarkable outcomes get reported; mundane outcomes do not).
Topic summary
Whether an aggregate is worth publishing at all: a disputed topic — does this still hold?
This is a generated summary. It shows the 5 most-liked posts from a topic of 19, in their original order, with the accepted answer included where one exists. It is a reading aid and it will miss nuance — the full topic is the record.
16 likes 7mo
h.falk, post #2: the opening post is right about the mechanism and I think understates the practical bit. I disagree with the reply above, and I think the disagreement is substantive rather than terminological. The distinction being drawn does not survive when you look at the published data for this specific question. I would be glad to be shown wrong… Go to post
Using data in discussions: datasets are useful as reference points when someone claims something unusual. "I have not seen that reported in the data" is different from "that is impossible", but data gives you something to say.
On post #8 — agreed on the reasoning, with one qualification.
How to contribute: if you have longitudinal data you want to add, the format is simple: date, measurement, context. Contact the maintainer of the specific dataset.
20 likes 4mo
Worth separating two things that post #12 runs together.
How to contribute: if you have longitudinal data you want to add, the format is simple: date, measurement, context. Contact the maintainer of the specific dataset.
14 likes 3mo
Read the full topic (19 posts)
This topic was referenced in
- Contributing data without breaching anyone's privacy — what changed sinceData & Tools › Datasets · 2 replies
- Cleaning a self-reported dataset and what you throw away — what changed sinceData & Tools › Datasets · 2 replies
- Revisiting: Sample size in a voluntary survey: the selection problemData & Tools › Datasets · 2 replies
Suggested topics
| Topic | Participants | Replies | Views | Activity |
|---|---|---|---|---|
|
Sample size in a voluntary survey: the selection problem
Posting this under the heading it deserves: Sample size in a voluntary survey: the selection problem Everything below is what sits behind that. Posting the method first, because I know what the first three…
|
+1 | 5 | 12k | 15mo |
|
Contributing data without breaching anyone's privacy — what changed since
On the subject in the title: Contributing data without breaching anyone's privacy — what changed since Working notes rather than a conclusion. A documentation question rather than an analytical one. I have a…
|
2 | 58k | 17mo | |
|
Coming back to: A community side-effect dataset, with its response rate and biases
Posting this under the heading it deserves: A community side-effect dataset, with its response rate and biases Everything below is what sits behind that. Working through the identity arithmetic and I would…
|
+96 | 101 | 33k | 2d |
|
About the Datasets category
Community-collected datasets, their collection methods, and their limitations. This post is a community wiki: any member at trust level 3 or above can edit it, and every edit is recorded with its author and a…
|
+6 | 11 | 42k | 12mo |
|
A community side-effect dataset, with its response rate and biases
A community side-effect dataset, with its response rate and biases — setting out what I have, and where I think it stops being reliable. Posting the method first, because I know what the first three replies…
|
2 | 5.1k | 19h |
Related topics — sharing the tags data table, confounding, observational data
| Topic | Participants | Replies | Views | Activity |
|---|---|---|---|---|
|
Deamidation and the close-eluting pair it produces
Deamidation and the close-eluting pair it produces Writing it up because I had to work it out twice and would rather nobody else did. I would like to understand what this number means before I repeat it…
|
+33 | 37 | 46k | 15mo |
|
What a confidence interval means, from scratch — the long version
The question in the title: What a confidence interval means, from scratch — the long version I will give what I have already checked below so nobody repeats it. I have seen SURMOUNT-2 ( Lancet , 2023) cited…
|
+16 | 20 | 2.3k | 8mo |
|
A claim built entirely on a subgroup analysis — what changed since
On the subject in the title: A claim built entirely on a subgroup analysis — what changed since Working notes rather than a conclusion. I have seen PIONEER 6 ( N Engl J Med , 2019) cited in support of a claim…
|
2 | 2k | 18mo | |
|
Coming back to: Why the cheapest option is often not the cheapest
Why the cheapest option is often not the cheapest — that is the question, and I have not found it answered plainly anywhere I have looked. Documenting an access outcome, dated, because everything in this…
|
+21 | 25 | 55k | 13mo |
|
About the Datasets category
Community-collected datasets, their collection methods, and their limitations. This post is a community wiki: any member at trust level 3 or above can edit it, and every edit is recorded with its author and a…
|
+6 | 11 | 42k | 12mo |