The Peptide CommonsEst. May 2024
Independent. We sell nothing and are affiliated with no manufacturer or pharmacy. Every moderation action is logged in public
Data & Tools · Datasets · continued

Coming back to: A community side-effect dataset, with its response rate and biases posts 91–102

This is a continuation of a long topic, addressed by post number rather than by page. Start at post 1 · go to the accepted answer.

RV
r.villalobosTL25 Jul 2026#91

Limitations of datasets: all community-collected data has limitations. The population is self-selected (people in this community are not representative of all people using these compounds). Reporting bias is real (remarkable outcomes get reported; mundane outcomes do not).

The short version is the first sentence; the rest is why.

20 likes 23d
HF
h.ferrariTL27 Jul 2026#92

The most useful dataset this community could hold is boring: lot, supplier, service, method, date, result. That is enough to answer most of the questions people ask badly.

I checked the source rather than the summary, and they differ.

8 likes 21d
RF
r.friskTL29 Jul 2026#93

I had written a reply contradicting post #89 and deleted it. Here is what survived.

The practical version of community side-effect dataset is three sentences long. The rigorous version is three pages and reaches the same conclusion with the conditions attached.

2 likes 19d
LA
l.aguirreTL211 Jul 2026#94
m.mwangi, post #29: What I would want before treating community side-effect dataset as settled: the method, the sample, and whether anyone tried to find the opposite result. Two of the three are usually missing. Go to post

On community side-effect dataset the community has more anecdote than the confidence in this thread implies, and I include my own contribution in that.

0 likes in reply to #29 17d
FF
f.fenwickTL3Regular13 Jul 2026#95

Agreed, and I will stop repeating the version of this I had been repeating.

27 likes 15d
AP
ar.petrovTL215 Jul 2026#96

The arithmetic in post #93 is right; the assumption feeding it is the part to check.

Units in the column header, always, and the same units down the whole column. Mixed units in one field is the commonest defect in shared spreadsheets here.

13 likes 13d
K
KForsbergTL2Member17 Jul 2026#97
n.bridgewater, post #62: Bias toward positive outcomes: datasets collected by members are biased toward people who found the compounds useful. People who did not respond do not return. People who had bad outcomes might have left the community. Go to post

Community side-effect dataset: I have looked for the primary source twice and failed twice. Either it does not exist or it is somewhere I do not know to look, and I would like to know which.

4 likes in reply to #62 11d
SH
s.hartmannTL219 Jul 2026 · edited#98
sleep_log, post #41: I would rather this thread reach "we do not know" about community side-effect dataset than reach a confident answer that nobody can support when asked. Go to post

Posting my community side-effect dataset numbers with the method attached so they can be discounted properly. Uncontrolled, unblinded, and collected by someone who wanted a particular answer.

0 likes in reply to #41 9d
DI
diluent_indexTL1Member21 Jul 2026#99
h.frisk, post #19: Where I part company with post #17, and it is a narrow parting. Aggregating first-hand accounts does not produce evidence of the kind a trial produces. It produces a description of who chose to post, which is a real thing and a different thing. Not disagreeing with anyone above, just adding the bit I keep having to look up. Go to post

Worth separating two things that post #97 runs together.

A codebook describing each field takes twenty minutes and is what makes the file usable by anyone but you. Most shared datasets here do not have one.

9 likes in reply to #19 7d
ON
o.nybergTL223 Jul 2026#100

This follows post #97 rather than contradicting it.

Community side-effect dataset is worth one more sentence than it usually gets, and the sentence is the one about how the number was arrived at.

2 likes 5d
MM
maintenance_modeTL3Regular24 Jul 2026#101

I have three months of notes on community side-effect dataset and the honest summary is that the trend is real and the week-to-week numbers are noise. I nearly drew the opposite conclusion from the first fortnight.

0 likes 3d
AP
au.pereiraTL226 Jul 2026#102

Post #101 answers the question as asked. The question underneath it is different.

Aggregating first-hand accounts does not produce evidence of the kind a trial produces. It produces a description of who chose to post, which is a real thing and a different thing.

I would want the raw data before agreeing with my own summary of it.

23 likes 2d

Suggested topics

TopicParticipantsRepliesViewsActivity
About the Datasets category
Community-collected datasets, their collection methods, and their limitations. This post is a community wiki: any member at trust level 3 or above can edit it, and every edit is recorded with its author and a…
BNOFSVPSAA+6 11 42k 12mo
Whether an aggregate is worth publishing at all: a disputed topic — does this still hold?
Whether an aggregate is worth publishing at all: a disputed topic — does this still hold? I have a specific reason for asking rather than idle curiosity, and the context is below. An honest uncertainty about…
BDHFBWIBOF+14 18 3.4k 2mo
Cleaning a self-reported dataset and what you throw away — what changed since
Posting this under the heading it deserves: Cleaning a self-reported dataset and what you throw away — what changed since Everything below is what sits behind that. What changes if the standard account of…
HMSCPW 2 2k 1h
Whether an aggregate is worth publishing at all: a disputed topic
On the subject in the title: Whether an aggregate is worth publishing at all: a disputed topic Working notes rather than a conclusion. An honest uncertainty about whether an aggregate rather than a disguised…
PEGLIGHNMD+20 24 18k 13mo
Contributing data without breaching anyone's privacy — what changed since
On the subject in the title: Contributing data without breaching anyone's privacy — what changed since Working notes rather than a conclusion. What changes if the standard account of Contributing data without…
RIMSBJ 2 58k 17mo

Related topics — sharing the tags observational data, site feedback, data table

TopicParticipantsRepliesViewsActivity
Impurity thresholds: where the common numbers come from
Impurity thresholds: where the common numbers come from — setting out what I have, and where I think it stops being reliable. I have spent a fortnight trying to pin impurity thresholds down and I want to set…
MRDNLAPOBP+11 15 825 2d
Area percent versus weight percent: the confusion that causes most arguments
Posting this under the heading it deserves: Area percent versus weight percent: the confusion that causes most arguments Everything below is what sits behind that. Proposing that area percent versus weight…
ADSVCDFNAD+10 14 2.1k 2mo
Tag proliferation and whether we should prune
On the subject in the title: Tag proliferation and whether we should prune Working notes rather than a conclusion. I have spent a fortnight trying to pin tag proliferation down and I want to set out where I…
ALADAW 2 11k 11mo
[2026 update] Why 99.2% and 97.8% on the same lot can both be correct
Why 99.2% and 97.8% on the same lot can both be correct — that is the question, and I have not found it answered plainly anywhere I have looked. I would like to understand what this number means before I…
BBWSZDSTD+1 5 719 2d
A friction point in the promotion workflow
A friction point in the promotion workflow — setting out what I have, and where I think it stops being reliable. Asking about friction point directly, because I have read four threads on it and each answered…
FPWPSAFNCA+4 8 491 11mo