The Peptide CommonsEst. May 2024
Independent. We sell nothing and are affiliated with no manufacturer or pharmacy. Every moderation action is logged in public
Data & Tools · Datasets · continued

Aggregated purity results across four services, with caveats — a second dataset posts 61–90

This is a continuation of a long topic, addressed by post number rather than by page. Start at post 1.

JH
j.habermannTL3Regular8 Sep 2025#61
Thibodeau, post #57: Confirming post #56 from a second method, which matters more than confirming it from a second person. The arithmetic on Aggregated purity results across four is the easy part and it is where the errors are, which is an uncomfortable combination. Show your working and someone will catch it. Go to post

I would keep Aggregated purity results across four and the decision it usually gets used for separate in this thread. They are related and they are not the same question, and merging them is why the last one went badly.

1 like in reply to #57 11mo
KO
k.okaforTL28 Sep 2025#62

Adding the measurement that post #59 says would settle it.

The practical version of Aggregated purity results across four is three sentences long. The rigorous version is three pages and reaches the same conclusion with the conditions attached.

0 likes 11mo
KF
k.farrugiaTL3Regular8 Sep 2025#63

The most useful dataset this community could hold is boring: lot, supplier, service, method, date, result. That is enough to answer most of the questions people ask badly.

That holds under the stated conditions and I have stated them.

15 likes 11mo
CB
c.balogunTL28 Sep 2025 · edited#64

A dataset is only useful if the collection method is described alongside it. Numbers without a protocol are a list rather than data.

Same conclusion as the reply above, reached differently, which is mildly reassuring.

6 likes 11mo
SP
s.poulsenTL3Regular9 Sep 2025#65
TL4_Halvorsen, post #4: I will take the caveat as seriously as the claim, which is the point of putting it there. Go to post

Noted, and I have changed what I was going to do on the strength of it.

2 likes in reply to #4 11mo
AP
a.petrovTL29 Sep 2025#66
a.lindholm, post #54: Where I part company with post #51, and it is a narrow parting. Reframing Aggregated purity results across four slightly, because I think the disagreement is about the question rather than the answer. If the question is "does it happen", yes. If it is "how often", nobody here knows. Go to post

Picking up post #63: that is the part I would want checked first.

Having read the whole Aggregated purity results across four thread before replying: the question in the first post has not actually been answered yet, and three of us have answered a nearby one instead.

0 likes in reply to #54 11mo
AR
ambient_reviewTL3Regular9 Sep 2025#67

Where a value is derived rather than measured, mark it. Derived columns get treated as observations the moment the file leaves your hands.

If it helps: the failure mode here is usually boring rather than dramatic.

21 likes 11mo
DN
d.nwosuTL29 Sep 2025#68

Date every record. A dataset assembled over two years without dates cannot distinguish a change over time from a change in who was contributing.

I have changed my mind on this once already, so take it as current rather than settled.

9 likes 11mo
V
VThorvaldsenTL3Regular9 Sep 2025#69

Publishing the raw records alongside the summary is what makes a dataset checkable. A summary alone asks for trust that nobody has earned.

5 likes 11mo
IR
i.rasmussenTL29 Sep 2025#70
y.adeyemi, post #36: Everything in post #32 holds. The case it does not cover is the one I have. Where a value is derived rather than measured, mark it. Derived columns get treated as observations the moment the file leaves your hands. Not a conclusion. A place to stand while looking for one. Go to post

Trying to state the Aggregated purity results across four position in a way that someone who disagrees would recognise as fair, because I do not think the version in this thread passes that test.

0 likes in reply to #36 11mo
MM
maintenance_modeTL3Regular9 Sep 2025#71

Narrowing post #70, because the general version has more than one answer.

Version the file rather than editing in place. A dataset that changes silently under an analysis makes the analysis unreproducible.

I would treat the number as indicative rather than as a measurement.

0 likes 11mo
KP
k.pereiraTL29 Sep 2025 · edited#72

Where a dataset is used to support a claim in a maintained document, the version used should be cited. Otherwise the document and the data drift apart silently.

The rule of thumb is fine; the edge cases are where it earns its keep.

1 like 11mo
RV
r.venkatesanTL3Wiki editor9 Sep 2025#73

The confident answers on Aggregated purity results across four and the well-sourced answers are not the same answers, which is the most useful thing I have learned reading this category.

11 likes 11mo
PK
p.krastevTL29 Sep 2025#74
chromatogram, post #8: The arithmetic in post #7 is right; the assumption feeding it is the part to check. The claim about Aggregated purity results across four upthread is stronger than its source supports. I have read the source. The source says "associated with" and the post says "causes". Go to post

Post #72 and I disagree about the size of the effect, not about the direction.

Where I have landed on Aggregated purity results across four, having got it wrong once in public: the direction is clear, the magnitude is not, and anyone quoting a precise magnitude has borrowed it from somewhere that did not measure it.

24 likes in reply to #8 11mo
IS
isotonic_sheetTL39 Sep 2025#75
NK
ni.kravchenkoTL29 Sep 2025#76

Small correction to my own earlier position on Aggregated purity results across four. I had the units the wrong way round, which changes the conclusion by an order of magnitude and therefore changes it entirely.

0 likes 11mo
RJ
r.jhannsdttirTL3Regular9 Sep 2025#77

This follows post #74 rather than contradicting it.

The question underneath Aggregated purity results across four is usually "how would I tell?" rather than "what is true?", and that one has a method attached to it.

Write down what you would expect to see under each hypothesis before you collect anything. If they predict the same observation, collecting it will not help.

7 likes 11mo
RS
r.sobczakTL29 Sep 2025#78
m.guerrero, post #16: Where I part company with post #14, and it is a narrow parting. Version the file rather than editing in place. A dataset that changes silently under an analysis makes the analysis unreproducible. I am describing what is, rather than arguing for what should be. Go to post

Saving this. It is the version I will quote when the question comes round again.

17 likes in reply to #16 11mo
EA
e.almeidaTL2Member9 Sep 2025 · edited#79

Where a value is derived rather than measured, mark it. Derived columns get treated as observations the moment the file leaves your hands.

1 like 11mo
PN
p.novakTL29 Sep 2025#80

Post #76 describes the usual case. This is about the unusual one.

Answering the Aggregated purity results across four question as asked, then the question I think is meant. As asked: yes, with the qualification below. As meant: it depends on how the first measurement was taken.

6 likes 11mo
RO
r.oyelaranTL29 Sep 2025#81
m.ekstrom, post #56: Post #54 describes the usual case. This is about the unusual one. The documentation on Aggregated purity results across four is better than this thread and I say that as someone who has posted in the thread. Go to post

That reframing is the whole thing. The facts I already had.

0 likes in reply to #56 11mo
KF
k.farrugiaTL3Regular9 Sep 2025#82

Where the Aggregated purity results across four discussion usually stalls is that nobody wants to say "I do not know" and everyone is willing to say "it varies". Those are the same sentence with different clothes on.

0 likes 11mo
DN
d.nwosuTL29 Sep 2025#83

Aggregating first-hand accounts does not produce evidence of the kind a trial produces. It produces a description of who chose to post, which is a real thing and a different thing.

17 likes 11mo
JH
j.habermannTL3Regular9 Sep 2025#84
ni.kravchenko, post #76: Small correction to my own earlier position on Aggregated purity results across four. I had the units the wrong way round, which changes the conclusion by an order of magnitude and therefore changes it entirely. Go to post

Publishing the raw records alongside the summary is what makes a dataset checkable. A summary alone asks for trust that nobody has earned.

7 likes in reply to #76 11mo
AA
a.amankwahTL29 Sep 2025#85

Worth stating the null on Aggregated purity results across four before we explain it: the observation may be nothing. That possibility deserves a sentence and usually does not get one.

1 like 11mo
RM
r.marsdenTL3Regular9 Sep 2025#86

A dataset is only useful if the collection method is described alongside it. Numbers without a protocol are a list rather than data.

0 likes 11mo
CB
c.balogunTL29 Sep 2025#87

If you are new and reading this thread for the answer to Aggregated purity results across four: the answer is conditional, the conditions are in the third reply, and the rest of the thread is worth skipping.

24 likes 11mo
LM
lyophil_marginTL3Regular9 Sep 2025 · edited#88
formulary_notes, post #2: Practical experience of Aggregated purity results across four, offered as one case with the conditions stated, not as a general finding. Conditions first, because they are what make it interpretable. Go to post

Adding the measurement that post #85 says would settle it.

My position on Aggregated purity results across four is current rather than settled. I have revised it once already and I expect to again, so treat it accordingly.

11 likes in reply to #2 11mo
YE
y.eriksenTL29 Sep 2025#89

A codebook describing each field takes twenty minutes and is what makes the file usable by anyone but you. Most shared datasets here do not have one.

0 likes 11mo
ST
sterile_tableTL3Regular9 Sep 2025#90

I came in to disagree and I am leaving without a disagreement.

18 likes 11mo