The Peptide CommonsEst. May 2024
Independent. We sell nothing and are affiliated with no manufacturer or pharmacy. Every moderation action is logged in public
Research Methods · Statistics

Regression to the mean in progress reports

HA
h.almeidaTL2Member10 Apr 2025#1

On the subject in the title: Regression to the mean in progress reports Working notes rather than a conclusion.

Something about regression does not reconcile and I would like a second pair of eyes before I decide which half is wrong.

Two sources, both reputable, giving figures that cannot both be right unless they are measuring different quantities. My suspicion is that they are, and I cannot see how.

58 likes 16mo
AM
a.molnarTL217 Apr 2025#2

The thing about regression that took me longest to accept is that a plausible mechanism is not evidence of an effect. It is a reason to look, not a result.

0 likes 15mo
TF
taper_fileTL3Regular21 Apr 2025#3

Bayesian and frequentist analyses answer different questions and both are legitimate. What matters is that the reader knows which is on offer.

I have deliberately not rounded that, because the rounding is where the argument starts.

3 likes 15mo
AV
a.villalobosTL225 Apr 2025#4
h.almeida, post #1: On the subject in the title: Regression to the mean in progress reports Working notes rather than a conclusion. Something about regression does not reconcile and I would like a second pair of eyes before I decide which half is wrong. Two sources, both reputable, giving figures that cannot both be right unless they are measuring… Go to post

The opening post describes the usual case. This is about the unusual one.

Regression: I have looked for the primary source twice and failed twice. Either it does not exist or it is somewhere I do not know to look, and I would like to know which.

11 likes in reply to #1 15mo
DM
d.magalhesTL2Member29 Apr 2025#5

Building on post #2 rather than restating it.

P-values and significance: p<0.05 means the data would be surprising if the null hypothesis were true, not that the null hypothesis is false. A non-significant p-value does not mean "no effect".

I have said this before in a thread nobody could find, so it is worth repeating.

0 likes 15mo
AC
a.coelhoTL23 May 2025#6

Noted, and I have changed what I was going to do on the strength of it.

1 like 15mo
BT
baseline_tableTL2Member6 May 2025#7

Regression is a good example of a question where the honest answer is boring and the interesting answers are unsupported. I would go with boring.

6 likes 15mo
TB
t.brandtTL210 May 2025#8
h.almeida, post #1: On the subject in the title: Regression to the mean in progress reports Working notes rather than a conclusion. Something about regression does not reconcile and I would like a second pair of eyes before I decide which half is wrong. Two sources, both reputable, giving figures that cannot both be right unless they are measuring… Go to post

A standard deviation describes the spread of individuals and a standard error describes the precision of the mean. Quoting one where the other belongs changes the apparent result substantially.

Adding this to the thread rather than to the wiki, because I am not confident enough for the wiki.

16 likes in reply to #1 15mo
N
NicolaidesTL3Regular13 May 2025#9

Reading this regression thread as someone who came in with a fixed view: the third and seventh replies moved me and the confident ones did not.

0 likes 15mo
FP
f.piresTL216 May 2025#10

Post #8 and I disagree about the size of the effect, not about the direction.

Baseline imbalance in a randomised trial is expected by chance and adjusting for it post hoc is a choice that should have been pre-specified.

Posted with less confidence than the sentence structure implies.

3 likes 14mo
AF
a.finnegan_rdTL219 May 2025#11
VK
v.kirchnerTL222 May 2025 · edited#12
Nicolaides, post #9: Reading this regression thread as someone who came in with a fixed view: the third and seventh replies moved me and the confident ones did not. Go to post

Adding the measurement that post #10 says would settle it.

The most common statistical error in this community is not technical: it is treating a self-selected collection of reports as a sample from a population.

2 likes in reply to #9 14mo
CL
coldchain_liuTL3Regular25 May 2025#13
a.coelho, post #6: Noted, and I have changed what I was going to do on the strength of it. Go to post

This is the first time the answer has come with its own limits attached. Appreciated.

0 likes in reply to #6 14mo
MN
m.nascimentoTL228 May 2025#14

What would change my mind on regression is a second dataset collected by someone with no stake in the first. Until then I hold it loosely and I would rather say so than pretend to more.

22 likes 14mo
LE
logbook_erinTL3Regular31 May 2025#15

Answering the question post #12 raises rather than the one it answers.

I changed my mind about regression after someone here asked me for the source and I could not produce one. That is worth saying out loud because it is the ordinary way it happens.

6 likes 14mo
AI
a.ilungaTL23 Jun 2025#16

The arithmetic in post #15 is right; the assumption feeding it is the part to check.

Number needed to treat: how many people need to be treated to prevent one bad outcome or achieve one good outcome. More intuitive than relative risk reduction.

Written quickly, so the reasoning may be tighter than the wording.

1 like 14mo
CO
c.okaforTL3Regular5 Jun 2025#17
Nicolaides, post #9: Reading this regression thread as someone who came in with a fixed view: the third and seventh replies moved me and the confident ones did not. Go to post

Confounding: a third variable explains an apparent association. In randomised data, randomisation balances confounders. In observational data, confounders can be adjusted for but unknown ones cannot.

If anyone has run this properly I would rather read that than my own guess.

30 likes in reply to #9 14mo
SG
s.girardTL28 Jun 2025#18

Regression to the mean: if you select people with extreme values (very high or very low), their next measurement is often less extreme just by chance. This can look like a treatment effect when it is just statistics.

Not a conclusion. A place to stand while looking for one.

15 likes 14mo
CC
crossref_checkTL3Wiki editor11 Jun 2025#19
t.brandt, post #8: A standard deviation describes the spread of individuals and a standard error describes the precision of the mean. Quoting one where the other belongs changes the apparent result substantially. Adding this to the thread rather than to the wiki, because I am not confident enough for the wiki. Go to post

Before the thread moves on from regression — what is the sample size behind the claim? I am not being difficult; I have seen the same figure quoted from an n of four and from an n of four hundred.

21 likes in reply to #8 14mo
FD
f.danquahTL213 Jun 2025#20
a.finnegan_rd, post #11: Medians and means diverge for skewed distributions, and most of the quantities discussed here are skewed. Which one a paper reports is a choice worth noticing. Go to post

Multiple testing inflates the false-positive rate in a way that is entirely predictable and entirely correctable. The correction should be declared in advance.

Genuinely open to being wrong about this one.

9 likes in reply to #11 13mo
GH
g.haalandTL3Regular16 Jun 2025 · edited#21

Post #20 answers the question as asked. The question underneath it is different.

What I want from this regression thread is the list of things that would need to be true for the claim to hold. If we can write that list, we can check it.

5 likes 13mo
SB
s.beaulieuTL219 Jun 2025#22

I read post #18 twice before replying, because I had assumed the opposite.

Posting my regression numbers with the method attached so they can be discounted properly. Uncontrolled, unblinded, and collected by someone who wanted a particular answer.

14 likes 13mo
FA
f.abrahamsenTL2Member21 Jun 2025#23

Seconded. It reads as careful rather than confident, which is the right register.

28 likes 13mo
KH
ka.haddadTL224 Jun 2025#24
f.danquah, post #20: Multiple testing inflates the false-positive rate in a way that is entirely predictable and entirely correctable. The correction should be declared in advance. Genuinely open to being wrong about this one. Go to post

A p-value is the probability of data at least this extreme given the null hypothesis. It is not the probability the hypothesis is false, and almost every plain-language gloss gets that backwards.

The short answer was in the first line; everything after is the working.

0 likes in reply to #20 13mo
AS
a.salcedoTL326 Jun 2025#25
MY
m.yilmazTL229 Jun 2025#26

Everything in post #22 holds. The case it does not cover is the one I have.

Two questions I would want answered before drawing anything from the regression data above: how were the cases selected, and what happened to the ones that dropped out.

20 likes 13mo
SD
s.duarteTL21 Jul 2025#27

Percentages of small denominators should be reported with the denominator. Two out of three is not sixty-seven per cent in any useful sense.

0 likes 13mo
VB
v.bhattacharyaTL24 Jul 2025#28
f.danquah, post #20: Multiple testing inflates the false-positive rate in a way that is entirely predictable and entirely correctable. The correction should be declared in advance. Genuinely open to being wrong about this one. Go to post

Regression to the mean: if you select people with extreme values (very high or very low), their next measurement is often less extreme just by chance. This can look like a treatment effect when it is just statistics.

0 likes in reply to #20 13mo
EM
endpoint_marginTL2Member6 Jul 2025#29

Rounding and significant figures carry information about precision. A figure quoted to four significant figures from a method with two per cent variability is overstating what is known.

13 likes 13mo
NZ
n.zielinskiTL28 Jul 2025#30

On regression, I would rather understate and be corrected upward than overstate and be quoted. That is a house style here and it is a good one.

27 likes 13mo