The Peptide CommonsEst. May 2024
Independent. We sell nothing and are affiliated with no manufacturer or pharmacy. Every moderation action is logged in public
Data & Tools · Datasets · continued

Coming back to: A community side-effect dataset, with its response rate and biases posts 31–60

This is a continuation of a long topic, addressed by post number rather than by page. Start at post 1 · go to the accepted answer.

BN
b.nwosuTL227 Feb 2026#31

Everything in post #27 holds. The case it does not cover is the one I have.

Aggregating first-hand accounts does not produce evidence of the kind a trial produces. It produces a description of who chose to post, which is a real thing and a different thing.

I would be glad to be shown a cleaner way of putting this.

0 likes 5mo
BP
baseline_peakTL2Member1 Mar 2026#32

Bookmarking this. I will come back when I have something worth adding.

32 likes 5mo
ZI
z.iyerTL24 Mar 2026#33

On community side-effect dataset I would separate what is worth knowing from what is worth acting on. The first list is long and the second is short, and conflating them is how threads get heated.

11 likes 5mo
MS
m.stephanopoulosTL3Regular6 Mar 2026 · edited#34
st.diallo, post #23: Where a dataset is used to support a claim in a maintained document, the version used should be cited. Otherwise the document and the data drift apart silently. Posted with less confidence than the sentence structure implies. Go to post

Reporting rather than recommending, on community side-effect dataset. What happened is above. Whether it should have is a different question and not one I am qualified to answer.

3 likes in reply to #23 5mo
WM
w.moreauTL28 Mar 2026#35

Units in the column header, always, and the same units down the whole column. Mixed units in one field is the commonest defect in shared spreadsheets here.

0 likes 5mo
W
WickramasingheTL2Member11 Mar 2026#36

Post #35 answers the question as asked. The question underneath it is different.

A codebook describing each field takes twenty minutes and is what makes the file usable by anyone but you. Most shared datasets here do not have one.

The mechanism is plausible, which is not the same as established.

24 likes 5mo
BJ
b.jansenTL213 Mar 2026#37

Worth separating two things that post #35 runs together.

Sample size is not the only thing that determines what a dataset can support. Selection is usually the larger problem here and it does not improve with volume.

Worth reading the earlier posts in this thread before acting on mine.

7 likes 5mo
Z
ZieglerTL3Regular15 Mar 2026#38
j.vandermolen, post #22: Coming back to post #20, because the follow-up matters more than the original answer. Where a value is derived rather than measured, mark it. Derived columns get treated as observations the moment the file leaves your hands. Not the answer, but possibly the question that gets there. Go to post

Summarising the community side-effect dataset thread so far, since it is long and the answer is buried: the first reply has the method, the fourth has the correction to it, and the rest is people agreeing at length.

1 like in reply to #22 4mo
PO
p.onwukaTL218 Mar 2026#39
e.lehtinen, post #10: Independent test results are the most valuable data this community collects, and they are only comparable when the method is captured alongside the number. Go to post

How to contribute: if you have longitudinal data you want to add, the format is simple: date, measurement, context. Contact the maintainer of the specific dataset.

3 likes in reply to #10 4mo
HN
h.nicolaidesTL3Regular20 Mar 2026#40

The arithmetic on community side-effect dataset is the easy part and it is where the errors are, which is an uncomfortable combination. Show your working and someone will catch it.

0 likes 4mo
SL
sleep_logTL2Regular22 Mar 2026#41

I would rather this thread reach "we do not know" about community side-effect dataset than reach a confident answer that nobody can support when asked.

8 likes 4mo
II
i.ilungaTL225 Mar 2026#42

Clear enough that I do not think I have a follow-up, which is unusual.

20 likes 4mo
IA
i.aranda_esTL2Translator · ES27 Mar 2026#43
c.grimaldi, post #6: Where a value is derived rather than measured, mark it. Derived columns get treated as observations the moment the file leaves your hands. It is the sort of thing that seems obvious in retrospect and was not at the time. Go to post

Units in the column header, always, and the same units down the whole column. Mixed units in one field is the commonest defect in shared spreadsheets here.

0 likes in reply to #6 4mo
SR
sa.rasmussenTL229 Mar 2026#44
SG
s.grigorescuTL2Member1 Apr 2026 · edited#45

Where the community side-effect dataset reasoning breaks down for me is the step from the group result to the individual case. That step is almost never argued for.

5 likes 4mo
WV
w.verhoevenTL23 Apr 2026#46

Publishing the raw records alongside the summary is what makes a dataset checkable. A summary alone asks for trust that nobody has earned.

The answer changed when I changed how I was measuring, which was informative.

14 likes 4mo
N
NicolaidesTL3Regular5 Apr 2026#47

Adding the measurement that post #45 says would settle it.

Version the file rather than editing in place. A dataset that changes silently under an analysis makes the analysis unreproducible.

28 likes 4mo
GT
g.tammTL27 Apr 2026#48
impurity_table, post #12: Small methodological point on community side-effect dataset: repeating a measurement is cheap and resolves most of what is being argued about here at no cost to anyone. Go to post

Where a dataset is used to support a claim in a maintained document, the version used should be cited. Otherwise the document and the data drift apart silently.

0 likes in reply to #12 4mo
EF
endo_fellow_rkTL3Endocrinology fellow10 Apr 2026#49

Sensible. I would want the same detail before I acted on it either.

2 likes 4mo
RE
r.ekstromTL212 Apr 2026#50

The most useful dataset this community could hold is boring: lot, supplier, service, method, date, result. That is enough to answer most of the questions people ask badly.

9 likes 4mo
VS
v.stanescuTL214 Apr 2026#51
t.nardone, post #28: Date every record. A dataset assembled over two years without dates cannot distinguish a change over time from a change in who was contributing. Go to post

Answering the question post #50 raises rather than the one it answers.

The honest answer on community side-effect dataset is that it depends, and the useful part is the list of what it depends on. Four items, in rough order of how much they matter.

Most people get the first two right and then argue about the fourth.

4 likes in reply to #28 3mo
AL
aliquot_lineTL3Regular16 Apr 2026#52

Checked the community side-effect dataset claim against the primary source this morning. It survives, with a narrower scope than the version quoted here. Posting the narrower scope.

0 likes 3mo
EN
e.nilsenTL218 Apr 2026#53

Self-reported, unblinded, self-selected data has known biases and is still worth collecting, provided every one of those words appears in the description.

0 likes 3mo
FT
fr.translation_moTL2Translator · FR21 Apr 2026#54
Nicolaides, post #47: Adding the measurement that post #45 says would settle it. Version the file rather than editing in place. A dataset that changes silently under an analysis makes the analysis unreproducible. Go to post

Post #53 is the version of this I will quote in future. One addition.

Bias toward positive outcomes: datasets collected by members are biased toward people who found the compounds useful. People who did not respond do not return. People who had bad outcomes might have left the community.

Reading it back, the second half matters more than the first.

18 likes in reply to #47 3mo
LK
l.krastevTL223 Apr 2026#55
GD
glossary_deskTL3Regular25 Apr 2026#56

Reading rather than contributing, but this is the most useful thread I have found on it.

0 likes 3mo
FP
f.piresTL227 Apr 2026#57

Where the community side-effect dataset discussion usually stalls is that nobody wants to say "I do not know" and everyone is willing to say "it varies". Those are the same sentence with different clothes on.

26 likes 3mo
N
NicolaidesTL3Regular29 Apr 2026#58

Temporal bias: older data in a dataset might reflect conditions (supplier, formulation, context) that have changed. Newer data is more current.

12 likes 3mo
AK
a.kirchnerTL21 May 2026#59
s.grigorescu, post #45: Where the community side-effect dataset reasoning breaks down for me is the step from the group result to the individual case. That step is almost never argued for. Go to post

Good question, well framed, and I would like to see it answered properly.

0 likes in reply to #45 3mo
AD
appeals_deskTL3Regular4 May 2026#60
baseline_peak, post #32: Bookmarking this. I will come back when I have something worth adding. Go to post

Picking up post #57: that is the part I would want checked first.

Where a dataset is used to support a claim in a maintained document, the version used should be cited. Otherwise the document and the data drift apart silently.

0 likes in reply to #32 3mo