The Ledger
a living recordYou meet advice like this every day: “it takes 21 days to build a habit”, “power poses change your hormones”. The Ledger is our public record of such claims — each one checked against the published research and stamped with one of the five verdicts below. So far, only 2 have fully survived. Every claim is below; open a card and its record shows the verdict, the strength of the evidence, and every source it was judged against.
Confirmed holds up · Nuanced true, with conditions · Overstated stretched past the evidence · Refuted does not hold · Declined not graded — the evidence was too thin to judge responsibly
What the five verdicts mean
Confirmed holds up · Nuanced true, with conditions · Overstated stretched past the evidence · Refuted does not hold · Declined not graded — the evidence was too thin to judge responsibly
Select a band to filter the wall · select again to clear
“It takes 21 days to build a habit.”
The evidence contradicts the popular '21 days' rule. In the original habit-formation study, participants took a median of around 66 days for a new behaviour to become automatic, with individual times ranging from 18 to 254 days depending on the person and the habit. A later pooled review of the evidence found closely similar results — medians of roughly 59 to 66 days — and its authors state plainly that their findings refute the popular notion that habits form in approximately 21 days.
Worth knowingMissing a single day of practice did not meaningfully disrupt habit formation in the original study, which undercuts the urgency implied by a fixed deadline.
Graded against
- Lally et al. 2010cohort · Euro J Social Psych ·
10.1002/ejsp.674 counters - Gardner et al. 2012commentary · Br J Gen Pract ·
10.3399/bjgp12x659466 context - Singh et al. 2024meta-analysis · Healthcare ·
10.3390/healthcare12232488 counters
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“Standing in a power pose changes your hormones.”
The idea that striking a brief high-power posture shifts your hormones traces to a single small laboratory study, which reported that holding an expansive pose for about a minute raised testosterone and lowered cortisol. That hormonal finding has not held up: several independent, larger, and more rigorously controlled attempts to reproduce it — including a close conceptual replication with blinded experimenters and a field study nested inside a real competition — found no significant change in testosterone or cortisol from power posing. The researchers behind the original study later co-authored a joint statement, alongside many other investigators who ran a coordinated set of preregistered replications, concluding that pose type has essentially no effect on any hormonal or behavioural measure. Standing in a power pose does not reliably change your hormones.
Worth knowingThis verdict addresses hormones only. The hormonal claim and the felt-power claim are different outcomes and must not be conflated: while testosterone and cortisol changes did not replicate, the effect of posing on self-reported feelings of power has been repeatedly confirmed, including in a dedicated Bayesian meta-analysis restricted specifically to that outcome. Power posing is not 'debunked' wholesale — only its hormonal mechanism is.
Graded against
- Carney et al. 2010RCT · Psychol Sci ·
10.1177/0956797610383437 supports - Ranehill et al. 2015RCT · Psychol Sci ·
10.1177/0956797614553946 counters - Smith & Apicella 2017RCT · Hormones and Behavior ·
10.1016/j.yhbeh.2016.11.003 counters - Metzler & Grèzes 2019RCT · PeerJ ·
10.7717/peerj.6726 counters - Jonas et al. 2017experiment · Comprehensive Results in Social Psychology ·
10.1080/23743603.2017.1342447 context - Gronau et al. 2017meta-analysis · Comprehensive Results in Social Psychology ·
10.1080/23743603.2017.1326760 context
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“Magnesium supplements improve sleep.”
Controlled trials show a small, real reduction in time to fall asleep in older adults and magnesium-deficient individuals, but the RCT record is explicitly contradictory and evidence quality is rated low to very low across all meta-analyses. The popular claim implies broad, reliable benefit that the evidence does not support.
Worth knowingA large systematic review found that while observational data links magnesium status to better sleep quality, the RCT findings are contradictory — making a confident population-wide recommendation impossible on current evidence.
Graded against
- Mah & Pitre 2021meta-analysis · BMC Complement Med Ther ·
10.1186/s12906-021-03297-z supports - Arab et al. 2023review · Biol Trace Elem Res ·
10.1007/s12011-022-03162-1 counters - Rawji et al. 2024review · Cureus ·
10.7759/cureus.59317 supports - Schuster et al. 2025RCT · NSS ·
10.2147/nss.s524348 supports - He et al. 2025review · NSS ·
10.2147/nss.s552646 context
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“Grip strength predicts longevity.”
Grip strength is among the most replicated predictors of all-cause mortality in epidemiology, outperforming systolic blood pressure in large multi-country cohorts. Prospective studies spanning millions of participants consistently find that lower grip strength is associated with higher mortality from cardiovascular disease, respiratory disease, and cancer. The relationship holds as a predictive signal; grip strength reflects overall physiological reserve rather than acting as a direct cause of longer life.
Worth knowingPrediction is not causation. Grip strength functions as a proxy for overall muscular and physiological reserve. No randomised controlled trial has demonstrated that specifically training to improve grip strength extends lifespan.
Graded against
- Leong et al. 2015cohort · The Lancet ·
10.1016/s0140-6736(14)62000-6 supports - Celis-Morales et al. 2018cohort · BMJ ·
10.1136/bmj.k1651 context - Wu et al. 2017meta-analysis · Journal of the American Medical Directors Association ·
10.1016/j.jamda.2017.03.011 supports - Zhuo et al. 2022RCT · Front. Cardiovasc. Med. ·
10.3389/fcvm.2022.930077 context
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“It takes 10,000 hours of practice to master a skill.”
Deliberate, structured practice is a genuine predictor of skill, but its power varies hugely by domain — explaining 26% of the variance in performance for games, 21% for music, 18% for sports, 4% for education, and less than 1% for professions. The specific '10 000 hour rule' popularised by Malcolm Gladwell is not, however, what the source science shows. It traces to a small study of 30 violin students and 12 pianists at one Berlin academy, which reported that skill tier corresponded to average accumulated practice time across the group — not a fixed per-person threshold for mastery. A pre-registered direct replication did not reproduce that core finding: by age 20, both the best and good violinists in the replication sample had already passed 10 000 hours of solo practice, with no reliable gap between the top two tiers. Ericsson himself later said there was no evidence for a 'magical number' of hours, and separately estimated that reaching elite international-level piano performance would take around 25,000 hours — roughly two and a half times the popularised figure.
Worth knowingThe original Ericsson study itself was small and correlational — 30 violinists and 12 pianists from a single Berlin conservatoire, sorted into skill tiers after the fact — and reported a group-average pattern, not evidence that any individual guaranteed mastery at 10 000 hours.
Graded against
- Ericsson et al. 1993cross-sectional · Psychological Review ·
10.1037/0033-295x.100.3.363 context - Macnamara & Maitra 2019replication · R. Soc. open sci. ·
10.1098/rsos.190327 context - Macnamara et al. 2014meta-analysis · Psychol Sci ·
10.1177/0956797614535810 supports - Ericsson & Harwell 2019commentary · Front. Psychol. ·
10.3389/fpsyg.2019.02396 counters
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“Melatonin is an effective treatment for insomnia.”
Melatonin produces a real but modest reduction in time to fall asleep (sleep onset latency) and acts on the circadian system, which suits sleep problems with a circadian component — jet lag, shift work, delayed sleep phase. But the American Academy of Sleep Medicine explicitly recommends against it for chronic primary insomnia, and its mechanism is chronobiotic (circadian timing) rather than sedative: it shifts sleep timing rather than inducing sleep. Calling it an effective treatment for insomnia without qualification overstates what the evidence supports.
Worth knowingThe sleep-onset benefit is statistically significant across multiple large meta-analyses but modest in magnitude — substantially smaller than the effects seen with approved hypnotic medications.
Graded against
- Sateia et al. 2017guideline · J Clin Sleep Med ·
10.5664/jcsm.6470 counters - Cruz‐Sanabria et al. 2024meta-analysis · Journal of Pineal Research ·
10.1111/jpi.12985 supports - Ferracioli-Oda et al. 2013meta-analysis · PLoS ONE ·
10.1371/journal.pone.0063773 context - Zisapel 2018review · British J Pharmacology ·
10.1111/bph.14116 context - Cohen et al. 2023experiment · JAMA ·
10.1001/jama.2023.2296 context
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
- GlossaryMelatonin: Definition, Mechanism, and Dosage for Sleep Timing
- GlossarySleep Onset Latency: Definition and Why 15 Minutes Is the Healthy Target
- GlossaryCircadian Rhythm: Definition and the 24-Hour Biological Clock
- Deep diveHow Your Circadian Clock Actually Controls Your Brain: Not Just Your Sleep
“Spacing your study out beats cramming.”
Spacing study sessions out over time produces better long-term retention than cramming them into one sitting — this is one of the best-replicated findings in cognitive psychology, drawn from a meta-analysis of 317 experiments and confirmed in a real-world field study on a large employee training dataset. But the advantage is conditional, not universal. First, how far apart sessions should be depends on how long you need to remember the material: a large factorial study found the optimal gap between sessions declined from about 20 to 40% of a 1-week test delay down to about 5 to 10% of a 1-year test delay — meaning spacing's advantage shrinks toward zero, and can favour massed practice instead, when the test is imminent. Second, the benefit is well-established for verbal and factual material but inconsistent for procedural skills: one study found spacing nearly doubled four-week retention of a maths procedure using 10 practice problems, while another equally well-powered study found no spacing benefit at all for a different maths procedure.
Worth knowingHow large the spacing effect is depends on the retention interval: the optimal gap between study sessions is not a fixed number of days but shrinks as a proportion of how long you need to remember the material — from about 20 to 40% of a 1-week delay down to 5 to 10% of a 1-year delay. This is also the mechanistic reason cramming can outperform spacing on an immediate test: massed repetition creates a close match between the study context and the test context that spaced study cannot offer.
Graded against
- Cepeda et al. 2006meta-analysis · Psychological Bulletin ·
10.1037/0033-2909.132.3.354 supports - Cepeda et al. 2008experiment · Psychol Sci ·
10.1111/j.1467-9280.2008.02209.x supports - Dunlosky et al. 2013review · Psychol Sci Public Interest ·
10.1177/1529100612453266 supports - Kang 2016review · Policy Insights from the Behavioral and Brain Sciences ·
10.1177/2372732215624708 supports - Rohrer & Taylor 2006RCT · Appl. Cognit. Psychol. ·
10.1002/acp.1266 supports - Ebersbach & Barzagar Nazari 2020RCT · Front. Psychol. ·
10.3389/fpsyg.2020.00811 counters - Kim et al. 2019field study · Behav Res ·
10.3758/s13428-018-1184-7 supports - Smith & Scarf 2017commentary · Front. Psychol. ·
10.3389/fpsyg.2017.00962 counters
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“Loneliness is as harmful to health as smoking 15 cigarettes a day.”
The underlying finding is real: people with poor social connection face a meaningfully higher risk of earlier death, a result replicated across three separate meta-analyses. But the specific 'as harmful as smoking 15 cigarettes a day' figure does not come from the meta-analysis usually cited for it — that paper describes the effect only in general terms, as comparable with quitting smoking, and reports a 50% relative-survival advantage for people with adequate social relationships, without ever naming a cigarette count. Researchers examining the claim's origins describe the '15 cigarettes/day' figure only as an oft-repeated claim whose derivation is unclear. The claim also collapses three distinct exposures into one: loneliness alone carries the smallest mortality risk (OR 1.14 to 1.26 across the two most recent meta-analyses), smaller than social isolation (OR 1.29 to 1.32) or living alone (OR 1.32) — yet the popular claim treats 'loneliness' as though it carried the combined weight of all three. When smoking is benchmarked on a comparable relative-risk scale, its own mortality risk in the moderate range (RR 2.02 for 10 to 20 cigarettes a day) is substantially larger than any of these odds ratios, undercutting the claimed equivalence even on its own quantitative terms.
Worth knowingThe mortality association between poor social connection and earlier death is genuine and consistently replicated — this is not a case of loneliness being harmless. The issue is specifically with the popular quantification, not the existence of a real effect.
Graded against
- Holt-Lunstad et al. 2010meta-analysis · PLoS Med ·
10.1371/journal.pmed.1000316 context - Holt-Lunstad et al. 2015meta-analysis · Perspect Psychol Sci ·
10.1177/1745691614568352 context - Wang et al. 2023meta-analysis · Nat Hum Behav ·
10.1038/s41562-023-01617-6 context - Smith et al. 2023commentary · American Journal of Epidemiology ·
10.1093/aje/kwad121 counters - VanderWeele & Kim 2024commentary · American Journal of Epidemiology ·
10.1093/aje/kwae059 counters
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“Teaching people in their preferred learning style improves learning.”
People do have stable preferences for how they like information presented — nobody disputes that. What the popular claim actually asserts, though, is something stronger: that giving someone lessons tailored to their assessed 'learning style' produces better learning than a mismatched format. Tested properly, that specific mechanism does not show up. The foundational review of the field found no adequate evidence base for using learning-styles assessments in teaching, a direct test in a real course found no relationship between students' assessed style and their exam performance, and a classroom trial built specifically to detect a matching benefit found none. Even the most sympathetic meta-analysis designed to rehabilitate the idea could only confirm the required pattern in 26% of the outcome measures it examined, and its own authors concluded that was too small and inconsistent to justify style-matched teaching.
Worth knowingThe foundational critical review of the learning-styles literature concluded there is no adequate evidence base to justify incorporating learning-styles assessments into general educational practice, despite the idea's popularity in classrooms.
Graded against
- Pashler et al. 2008review · Psychol Sci Public Interest ·
10.1111/j.1539-6053.2009.01038.x counters - Rohrer & Pashler 2012commentary · Medical Education ·
10.1111/j.1365-2923.2012.04273.x counters - Husmann & O'Loughlin 2019cross-sectional · Anatomical Sciences Ed ·
10.1002/ase.1777 counters - Rogowsky et al. 2020RCT · Front. Psychol. ·
10.3389/fpsyg.2020.00164 counters - Newton 2015study · Front. Psychol. ·
10.3389/fpsyg.2015.01908 context - Clinton-Lisell & Litzinger 2024meta-analysis · Front. Psychol. ·
10.3389/fpsyg.2024.1428732 context
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“Your personality is fixed by the time you reach adulthood.”
False as stated. A study of 132,515 adults aged 21-60 directly tested the idea that personality traits stop changing by age 30 and found that Conscientiousness and Agreeableness kept increasing throughout early and middle adulthood, well past 30. A separate meta-analysis found 4 of the 6 broad trait categories it studied showed significant mean-level change in middle and old age — change does not stop at any fixed adult cutoff. Personality can also be shifted deliberately: a meta-analysis of intervention studies found meaningful trait change (d = .37) after an average of 24 weeks, persisting beyond the intervention itself. Personality is not fixed by adulthood, and it keeps changing well beyond it.
Worth knowingThe claim survives because it mistakes a different, real phenomenon for trait fixity: people's personality relative to their peers becomes fairly consistent with age. Test-retest stability of trait rankings rises from .31 in childhood to .54 during the college years, to .64 at age 30, and plateaus around .74 between ages 50 and 70 — high, but not perfect. That is about relative ranking staying similar, not about trait levels staying the same.
Graded against
- Srivastava et al. 2003cohort · Journal of Personality and Social Psychology ·
10.1037/0022-3514.84.5.1041 counters - Roberts et al. 2006meta-analysis · Psychological Bulletin ·
10.1037/0033-2909.132.1.1 counters - Roberts et al. 2017meta-analysis · Psychological Bulletin ·
10.1037/bul0000088 counters - Roberts & DelVecchio 2000meta-analysis · Psychological Bulletin ·
10.1037/0033-2909.126.1.3 context
Every source Crossref-verified — tap to open
“Mentally rehearsing a skill improves how well you perform it.”
Four decades of meta-analyses agree: mentally rehearsing a motor skill improves how well you later perform it, compared with no practice at all. The first major synthesis, pooling 60 studies, found mental practice improved performance with an average effect size of .48. A second, highly cited synthesis confirmed the effect was positive and significant, and found it was moderated by the type of task, the gap between practice and performance, and how long the mental practice lasted. A follow-up meta-analysis found an even larger average effect of .68, and identified that imagining the movement from the inside, as if you are the one performing it, works better than picturing yourself from the outside. A more recent, bias-corrected replication of the whole field, run in the shadow of psychology's reproducibility crisis, still found a small but significant positive effect (r = 0.131), smaller than the earlier estimates but the same direction and still real. And in an applied test with real stakes, novice surgeons who mentally rehearsed a laparoscopic procedure scored significantly higher on technical-skill ratings than those who did not, across every one of their practice sessions. So the effect holds up under scrutiny: it just shrinks once the field is corrected for publication bias and re-tested rigorously.
Worth knowingThis verdict covers rehearsing a specific motor skill, a sport movement, a surgical procedure, a musical piece, in your mind before performing it. It does not cover imagining yourself having already achieved a goal or outcome, which is a separate claim with different and largely opposite evidence, and the two should not be treated as the same thing.
Graded against
- Feltz & Landers 1983meta-analysis · Journal of Sport Psychology ·
10.1123/jsp.5.1.25 supports - Driskell et al. 1994meta-analysis · Journal of Applied Psychology ·
10.1037/0021-9010.79.4.481 supports - Hinshaw 1991meta-analysis · Imagination, Cognition and Personality ·
10.2190/x9ba-kj68-07an-qmj8 supports - Toth et al. 2020replication · Psychology of Sport and Exercise ·
10.1016/j.psychsport.2020.101672 supports - Arora et al. 2011RCT · Annals of Surgery ·
10.1097/sla.0b013e318207a789 supports
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“A 20-minute power nap restores focus and performance.”
The specific “20-minute nap” figure in this claim traces to a canonical dose-response trial that directly compared 10-, 20-, and 30-minute naps: the 20-minute nap produced measurable improvements in alertness and cognitive performance, though the benefits took 35 minutes to emerge and then lasted up to 125 minutes, while the 10-minute nap acted faster, with some benefits maintained for up to 155 minutes. Two independent meta-analyses pooling many controlled napping studies corroborate a real, replicated small-to-medium benefit of afternoon napping on memory, vigilance, and processing speed, and found this benefit held across a range of nap durations rather than being unique to any single length. So the core mechanism — a brief nap improving subsequent focus and performance — is well supported. But the popular claim strips out real conditions: the trial establishing the 20-minute figure was run in participants under shortened nocturnal sleep, not fully rested people, and the safe window behaves like a cliff rather than a plateau — a nap running just 10 minutes longer, to 30 minutes, produced measurable sleep inertia (impaired alertness and performance right after waking) in two separate studies before any benefit appeared.
Worth knowingThe direct evidence for the 20-minute duration specifically comes from a single lab trial in participants under shortened nocturnal sleep, not fully rested people. Despite an explicit search for a null result in well-rested, non-sleep-deprived adults, none was found in the literature — generalizing this to a fully rested person taking a 20-minute nap is inferred, not directly tested.
Graded against
- Brooks & Lack 2006experiment · Sleep ·
10.1093/sleep/29.6.831 supports - Leong et al. 2022meta-analysis · Sleep Medicine Reviews ·
10.1016/j.smrv.2022.101666 supports - Dutheil et al. 2021meta-analysis · IJERPH ·
10.3390/ijerph181910212 context - Hilditch et al. 2017review · Sleep Medicine ·
10.1016/j.sleep.2016.12.016 counters - Hilditch et al. 2016RCT · Sleep ·
10.5665/sleep.5550 counters
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“Ashwagandha lowers cortisol.”
Standardised ashwagandha extracts reliably lower serum cortisol in stressed adults, with reductions ranging from 11% to 32.63% across multiple independent RCTs and systematic reviews. The effect is specific to proprietary standardised formulations — not raw root powder or arbitrary products labelled ashwagandha. Critically, lower serum cortisol does not reliably translate to less felt stress: one meta-analysis confirmed cortisol fell significantly across trials but perceived stress scores did not improve.
Worth knowingThe cortisol-lowering effect is extract-specific: all consistent trial evidence uses standardised, withanolide-calibrated root extracts of the kind used in clinical trials. The evidence cannot be generalised to uncharacterised products sold as ashwagandha.
Graded against
- Chandrasekhar et al. 2012RCT · Indian Journal of Psychological Medicine ·
10.4103/0253-7176.106022 supports - Lopresti et al. 2019RCT · Medicine ·
10.1097/md.0000000000017186 supports - Della Porta et al. 2023systematic review · Nutrients ·
10.3390/nu15245015 supports - Bachour et al. 2025meta-analysis · BJPsych open ·
10.1192/bjo.2025.10136 supports - Albalawi 2025meta-analysis · Nutr Health ·
10.1177/02601060251363647 counters - Philips et al. 2023case series · Hepatology Communications ·
10.1097/hc9.0000000000000270 context
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
- Trend breakdownDoes Ashwagandha Lower Cortisol? The Evidence Reviewed
- GlossaryAshwagandha: Definition, Function and Evidence on the Adaptogenic Stress Response
- GlossaryCortisol: Definition, Function & What High Cortisol Means for Performance
- Deep diveCortisol: The Complete Science of Your Body’s Primary Stress Hormone
“Beta-alanine improves high-intensity performance.”
Beta-alanine raises muscle carnosine, which buffers hydrogen ions during intense exercise and delays fatigue-driving acidosis. The performance benefit is real but modest, and it is concentrated in sustained high-intensity efforts lasting roughly one to four minutes — efforts where muscle acid build-up is the limiter. Very short maximal sprints under a minute and extended aerobic efforts beyond roughly twenty-five minutes show no consistent benefit.
Worth knowingThe mechanism is well-established: beta-alanine supplementation elevates muscle carnosine, an intracellular proton buffer that slows the pH drop limiting high-intensity output — making it genuinely effective within its duration window.
Graded against
- Saunders et al. 2017meta-analysis · Br J Sports Med ·
10.1136/bjsports-2016-096396 supports - Trexler et al. 2015study · Journal of the International Society of Sports Nutrition ·
10.1186/s12970-015-0090-y supports - Hobson et al. 2012meta-analysis · Amino Acids ·
10.1007/s00726-011-1200-z counters
Every source Crossref-verified — tap to open
“Breathwork lowers stress as effectively as meditation.”
Breathwork is broadly comparable to mindfulness meditation for reducing stress and anxiety, based on the limited direct comparative evidence available. The one rigorous head-to-head trial found that cyclic sighing matched mindfulness meditation for anxiety reduction and exceeded it for positive mood. However, this single trial covered one specific breathwork technique over a short duration, and the broader meta-analytic literature could not address the comparison for lack of qualifying head-to-head studies. Whether the equivalence extends to other breathwork types, longer durations, or clinical populations remains unknown.
Worth knowingThe comparative claim rests on a single randomised controlled trial; at the meta-analytic level, no qualifying head-to-head trials between breathwork and meditation existed at the time of the most recent breathwork meta-analysis, which could only establish that breathwork outperforms non-breathwork control conditions.
Graded against
- Balban et al. 2023RCT · Cell Reports Medicine ·
10.1016/j.xcrm.2022.100895 supports - Fincham et al. 2023meta-analysis · Sci Rep ·
10.1038/s41598-022-27247-y context - Zaccaro et al. 2018review · Front. Hum. Neurosci. ·
10.3389/fnhum.2018.00353 context - Blades et al. 2024RCT · Comprehensive Psychoneuroendocrinology ·
10.1016/j.cpnec.2024.100272 counters - Fox et al. 2025RCT · Sci Rep ·
10.1038/s41598-025-29187-9 context
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“Cold plunges speed up muscle recovery.”
Cold water immersion reliably reduces delayed-onset muscle soreness and biochemical markers of muscle damage, making you feel recovered faster after hard training. But when applied habitually after resistance training it suppresses the molecular signals that drive muscle growth, trading long-term adaptation for short-term comfort.
Worth knowingThe soreness-relief benefit is well-supported across a large body of RCTs: cold water immersion was most effective for biochemical markers and neuromuscular recovery, and best for alleviating muscle soreness.
Graded against
- Wang et al. 2025meta-analysis · Front. Physiol. ·
10.3389/fphys.2025.1525726 supports - Piñero et al. 2024meta-analysis · European Journal of Sport Science ·
10.1002/ejsc.12074 counters - Roberts et al. 2015RCT · The Journal of Physiology ·
10.1113/jp270570 counters
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“Collagen supplements improve skin and joint health.”
Collagen supplementation produces a small-to-moderate, well-supported reduction in joint pain and functional impairment in osteoarthritis, backed by moderate-to-high certainty evidence from a large trial sequential meta-analysis. The skin half of the claim does not hold up under scrutiny: pooled positive effects in major meta-analyses disappear entirely when restricted to non-industry-funded or high-quality trials, leaving the popular skin claim without independent support.
Worth knowingFor joint health in osteoarthritis, the evidence is relatively robust: a large trial sequential meta-analysis confirmed consistent pain relief and function improvement at moderate-to-high certainty by GRADE — making joints the stronger, more defensible half of the compound claim.
Graded against
- Liang et al. 2024meta-analysis · Osteoarthritis and Cartilage ·
10.1016/j.joca.2023.12.010 supports - Myung & Park 2025meta-analysis · The American Journal of Medicine ·
10.1016/j.amjmed.2025.04.034 counters - Pu et al. 2023meta-analysis · Nutrients ·
10.3390/nu15092080 context - Dewi et al. 2023meta-analysis · Cureus ·
10.7759/cureus.50231 context - Lin et al. 2023meta-analysis · J Orthop Surg Res ·
10.1186/s13018-023-04182-w supports - Shaw et al. 2017RCT · The American Journal of Clinical Nutrition ·
10.3945/ajcn.116.138594 context
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“Extroverts make better leaders.”
The honest answer turns on a distinction the popular claim erases: who emerges as a leader versus who leads effectively once in the role. A large meta-analysis across many studies confirms extraversion is the single most consistent Big Five personality correlate of leadership, but that same analysis found extraversion tracks who gets seen and chosen as a leader more strongly than it tracks who actually performs well as one -- the emergence effect is not the same as the effectiveness effect. A field study of store managers, paired with a laboratory experiment, found the relationship between extraverted leadership and objective group performance actually reverses depending on the people being led: extraverted leadership lifts measurable output such as profit only when followers are passive, but when followers are proactive -- voicing ideas, taking initiative -- extraverted leadership is linked to lower objective performance, because extraverted leaders are perceived as less receptive to that input. A large modern update of the original meta-analysis, spanning many countries, corroborates that the extraversion-effectiveness link is not a fixed law: it is stronger in collectivist cultures and appears to work through leader behaviour rather than personality alone. So 'extroverts make better leaders' holds for who tends to be recognised and selected as a leader, but the case for extroverts leading more effectively once in charge is conditional -- on follower proactivity and on cultural context -- not a blanket truth.
Worth knowingThe meta-analysis pools results from many independent samples across business, government or military, and student settings, and separates leader emergence from leader effectiveness as distinct outcomes -- extraversion's link to effectiveness is real but consistently weaker than its link to emergence.
Graded against
- Judge et al. 2002meta-analysis · Journal of Applied Psychology ·
10.1037/0021-9010.87.4.765 supports - Grant et al. 2011survey · AMJ ·
10.5465/amj.2011.61968043 counters - Javalagi et al. 2024meta-analysis · Journal of Applied Psychology ·
10.1037/apl0001182 context
Every source Crossref-verified — tap to open
“Having a growth mindset improves achievement.”
The evidence is genuinely split, and the honest picture is small-to-null on average, with any real signal concentrated in specific groups rather than in students generally. The largest, best-designed randomised trial — a nationally representative U.S. study (Yeager and colleagues, National Study of Learning Mindsets) — found a real intervention effect, but it improved grades specifically among lower-achieving students, and it only held in schools where peer norms aligned with the intervention's message; it was not a general effect across all students. An earlier longitudinal study of 373 7th-graders found a similar pattern: believing intelligence is malleable predicted rising grades, and a brief classroom intervention reversed a grade decline seen in the control group. Set against this, the field's most rigorous systematic review (Macnamara and Burgoyne, 63 intervention studies, N=97,672) found only a small overall effect (d=0.05, 95% CI 0.02-0.09) that became statistically nonsignificant once publication bias was corrected for, and the authors traced the apparent benefit substantially to weak study design and to researchers with a financial incentive in the outcome publishing larger effects. A separate pair of meta-analyses (Sisk and colleagues) likewise found weak overall effects, with any real benefit confined to students who are low-income or academically at risk. A construct-validation study testing the theory's core premises directly found little support for them, concluding that mind-set claims 'appear to be overstated.'
Worth knowingThis is not evidence that mindset beliefs are irrelevant to learning — the effect is real but narrow: it shows up in lower-achieving students and depends on classroom or school peer-norm context, not as a general lever that raises achievement for students at large.
Graded against
- Yeager et al. 2019RCT · Nature ·
10.1038/s41586-019-1466-y supports - Blackwell et al. 2007RCT · Child Development ·
10.1111/j.1467-8624.2007.00995.x supports - Sisk et al. 2018meta-analysis · Psychol Sci ·
10.1177/0956797617739704 counters - Macnamara & Burgoyne 2023meta-analysis · Psychological Bulletin ·
10.1037/bul0000352 counters - Burgoyne et al. 2020study · Psychol Sci ·
10.1177/0956797619897588 counters
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“Having more options makes people less likely to choose.”
The evidence is genuinely split, not simply for or against. The original jam-display experiment found a dramatic effect in one setting: only 3% of shoppers bought when shown an extensive display, versus 30% who bought from a limited display. But the largest attempt to average across the whole choice-overload literature — 63 experimental conditions from 50 studies, pooling 5,036 participants — found the effect shrinks to virtually zero once pooled, with considerable unexplained variance between individual studies. A separate, larger re-analysis of that same body of research argues the pooled-null result is misleading because it averages over very different conditions, and that assortment size does reliably matter once factors such as how complex the choice set is and how settled the decision-maker's preferences already are get modelled — a conclusion its authors present as directly contradicting the pooled-null finding. The two research teams have not reconciled: whether 'more options' reliably backfires depends on unresolved, actively disputed questions about which conditions matter, not on one settled number.
Worth knowingNo peer-reviewed, Crossref-verifiable direct replication of the original jam-display study could be located. Reports that replication attempts failed trace only to secondary sources (blog posts, a magazine feature) rather than a citable primary study, so this verdict does not assert a failed replication.
Graded against
- Iyengar & Lepper 2000experiment · Journal of Personality and Social Psychology ·
10.1037/0022-3514.79.6.995 supports - Scheibehenne et al. 2010meta-analysis · J Consum Res ·
10.1086/651235 counters - Chernev et al. 2015meta-analysis · J Consum Psychol ·
10.1016/j.jcps.2014.08.002 context - Chernev et al. 2010commentary · J Consum Res ·
10.1086/655200 context
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“Making a specific if-then plan makes you more likely to follow through.”
Making a specific if-then plan measurably increases follow-through, but how much depends heavily on who is measuring it and what kind of behaviour is being changed. The two largest meta-analyses supporting the effect (94 independent tests, d = .65; a later synthesis of 642 tests) come from the theory's own originators. Independent researchers replicate a similarly strong effect for one-off, discrete actions — a clinical/mental-health meta-analysis outside that research group found an even larger effect (d + = 0.99) after excluding one outlier. But independent researchers studying repeated, habitual behaviours like exercise find the opposite: a pooled analysis of randomised physical-activity trials found no significant effect (0.15) without added reinforcement, and a large field experiment on gym attendance found a tightly estimated null result. The claim holds well for one-off, novel actions; it does not reliably hold for sustained behaviour change on its own.
Worth knowingThe two headline meta-analyses defining this effect (94 independent tests, d = .65, and a later update spanning 642 independent tests) were conducted by the theory's own originators (Gollwitzer and Sheeran), so on their own they cannot settle whether the effect replicates independently of the lab that built it.
Graded against
- Gollwitzer 1999commentary · American Psychologist ·
10.1037/0003-066x.54.7.493 supports - Gollwitzer & Sheeran 2006meta-analysis · Advances in Experimental Social Psychology ·
10.1016/s0065-2601(06)38002-1 supports - Sheeran et al. 2025meta-analysis · European Review of Social Psychology ·
10.1080/10463283.2024.2334563 context - Adriaanse et al. 2011meta-analysis · Appetite ·
10.1016/j.appet.2010.10.012 context - Toli et al. 2016meta-analysis · British J Clinic Psychol ·
10.1111/bjc.12086 supports - Silva et al. 2018meta-analysis · PLoS ONE ·
10.1371/journal.pone.0206294 counters - Carrera et al. 2018RCT · Journal of Health Economics ·
10.1016/j.jhealeco.2018.09.002 counters
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“Multitasking makes you less productive.”
The core mechanism behind this claim is well established: the brain does not run two demanding tasks in true parallel, it switches between them, and each switch carries a measurable cost in speed and accuracy. That cost, however, is not universal — it shows up specifically when the tasks compete for the same underlying cognitive resource. When two tasks draw on separate resources instead (for example one perceptual-motor, one cognitive), modelling of that exact task pairing shows they can eventually be interleaved with little or no measurable interference — but only after many trials of practice; the same pairing still showed real interference before that training took hold. So the escape hatch from switch costs is narrow: it requires both resource-disjoint tasks and a well-practised pairing, which does not describe the everyday case of switching between demanding, overlapping tasks, and 'multitasking always costs you' still overstates what the switch-cost research shows for that narrow, trained exception. The stronger popular version of this claim — that habitual heavy media multitasking causes a lasting decline in general cognitive ability — rests on much shakier ground: pre-registered replication attempts of the founding study mostly failed, and a meta-analysis correcting for small-study bias found the association was no longer significant. Finally, every citation behind this verdict measures laboratory reaction time, accuracy, or distractor filtering — none directly measures real-world workplace output, so extending these findings to a claim about 'productivity' specifically involves a genuine gap beyond the direct evidence.
Worth knowingSwitch costs are a robust, replicated laboratory finding: alternating between two tasks produces slower, more error-prone performance than repeating the same task, and this cost is only partly reduced by advance preparation or cueing.
Graded against
- Rubinstein et al. 2001experiment · Journal of Experimental Psychology: Human Perception and Performance ·
10.1037/0096-1523.27.4.763 supports - Monsell 2003review · Trends in Cognitive Sciences ·
10.1016/s1364-6613(03)00028-7 supports - Salvucci & Taatgen 2008commentary · Psychological Review ·
10.1037/0033-295x.115.1.101 counters - Wiradhany & Nieuwenstein 2017meta-analysis · Atten Percept Psychophys ·
10.3758/s13414-017-1408-4 counters - Ophir et al. 2009cross-sectional · Proc. Natl. Acad. Sci. U.S.A. ·
10.1073/pnas.0903620106 context - Uncapher & Wagner 2018review · Proc. Natl. Acad. Sci. U.S.A. ·
10.1073/pnas.1611612115 context
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“People choke under pressure.”
The picture is genuinely two-sided, and both sides are well evidenced. In controlled experiments, expert performers on a proceduralised sensorimotor skill -- golf putting -- reliably choked when a cash incentive raised the stakes, while performers on a task that stays in working memory rather than becoming automatic did not choke under the same pressure; choking was eliminated, and performance even improved under pressure, when performers had trained under conditions that raised self-consciousness during practice. An earlier, foundational experiment found the same pattern from a different angle: golfers who learned putting implicitly, without building up a large store of explicit step-by-step rules, were less likely to break down under evaluative and financial pressure than golfers who learned the same skill explicitly. A review of the wider behavioural and neuroimaging literature situates this within a broader pattern -- that high incentives can produce performance decrements rather than the performance boost simple economic models predict. But a field study of actual PGA, Senior PGA and LPGA Tour final rounds, where the pressure and stakes are as real as competitive golf gets, found no support for the choking hypothesis at all: players leading going into the final round won more often than not rather than folding. So choking is a real, mechanistically demonstrated phenomenon for skills that have become automatic, and it is trainable away -- but it is not a universal law that applies uniformly whenever stakes rise, and the clearest direct test among elite professionals under real competitive pressure did not find it.
Worth knowingThe core choking-under-pressure experiments used undergraduate and collegiate golfers in controlled laboratory settings with cash incentives standing in for real stakes -- strong for isolating the explicit-monitoring mechanism, but a different context from professional, career-stakes competition.
Graded against
- Beilock & Carr 2001experiment · Journal of Experimental Psychology: General ·
10.1037/0096-3445.130.4.701 supports - Masters 1992experiment · British J of Psychology ·
10.1111/j.2044-8295.1992.tb02446.x supports - Yu 2015review · Front. Behav. Neurosci. ·
10.3389/fnbeh.2015.00019 context - Clark 2002field study · Percept Mot Skills ·
10.2466/pms.2002.94.3c.1124 counters
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“Positive affirmations improve self-esteem.”
The popular claim usually means one specific thing: repeating a fixed positive self-statement about yourself, such as ‘I am a lovable person’. Tested directly, that practice backfired for people with low self-esteem -- the exact population it is marketed to help -- leaving them feeling worse, not better, while people who already had high self-esteem gained only a small benefit (the same study's own meta-analytic comparison found the low-self-esteem effect, d = 0.72, was even larger than the high-self-esteem benefit, d = 0.66). That backfire finding is not settled: a later two-study replication attempt found no self-esteem-moderated effect at all, and in its second study found no self-esteem benefit from repeating positive self-statements or from writing about personal values either -- so neither harm nor benefit for this specific practice is currently well established. A different, often-confused intervention that researchers actually call 'self-affirmation' -- writing about your own core personal values, not reciting a scripted positive trait -- has a real applied evidence base in education and health, and one large adolescent trial found it improved self-esteem specifically. But that is a different exercise from reciting affirmations, and defenders of the popular practice regularly borrow its evidence without acknowledging the swap.
Worth knowingThe strongest single experimental test of the popular practice (repeating a fixed positive self-statement) found it made low-self-esteem participants feel worse, not better -- the reverse of the marketed benefit -- with an effect size (d = 0.72) larger than the modest gain seen in high-self-esteem participants (d = 0.66).
Graded against
- Wood et al. 2009experiment · Psychol Sci ·
10.1111/j.1467-9280.2009.02370.x counters - Flynn & Bordieri 2020replication · Journal of Contextual Behavioral Science ·
10.1016/j.jcbs.2020.03.003 context - Cohen & Sherman 2014review · Annual Review of Psychology ·
10.1146/annurev-psych-010213-115137 context - Yan et al. 2024RCT · Applied Psych Health & Well ·
10.1111/aphw.12516 supports
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“Regular sauna use lowers heart-disease risk.”
Frequent sauna use — several sessions per week sustained over years — is consistently linked to lower risk of fatal cardiovascular events and incident hypertension in long-running prospective cohort data (large groups of people tracked over many years, with researchers recording who develops disease), with a clear dose-response pattern that persists after adjustment for conventional risk factors. Causation is not established: the supporting studies all draw from the same overlapping Finnish cohort, healthy-user confounding (the possibility that sauna users are already healthier in ways that statistics cannot fully account for) cannot be fully excluded, and short-term randomized trials of passive heating have not confirmed cardiometabolic benefit.
Worth knowingAll three major supporting studies draw from the same Finnish research cohort (the Kuopio Ischaemic Heart Disease Risk Factor Study), so the dose-response association reflects one population studied repeatedly — not independent replication across diverse groups.
Graded against
- Laukkanen et al. 2015cohort · JAMA Intern Med ·
10.1001/jamainternmed.2014.8187 supports - Laukkanen et al. 2018cohort · BMC Med ·
10.1186/s12916-018-1198-0 supports - Zaccardi et al. 2017cohort · American Journal of Hypertension ·
10.1093/ajh/hpx102 supports - Hamaya et al. 2025meta-analysis · American Journal of Preventive Cardiology ·
10.1016/j.ajpc.2025.101082 counters - Laukkanen et al. 2018review · Mayo Clinic Proceedings ·
10.1016/j.mayocp.2018.04.008 context
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“Sleeping on a decision leads to a better choice.”
This claim bundles two different things, and the evidence treats them very differently. Actual sleep — falling asleep overnight and waking up — has solid, replicated support: in a controlled sleep-lab study, people who slept between learning a hidden rule and being retested were more than twice as likely to consciously discover that rule as people who stayed awake for the same stretch, and a separate study found sleep deprivation measurably pushed people toward riskier, worse choices on a gambling task relative to their own rested baseline. The other half of the claim — that a brief spell of quiet, unconscious 'mulling it over' without actually sleeping produces demonstrably better complex decisions than consciously thinking it through — traces to an influential study on consumer choice, but has since failed its largest, best-powered, pre-registered direct replication and an earlier independent meta-analysis, neither of which found a reliable unconscious-thought advantage. In short: if 'sleep on it' means literal sleep, the advice holds up; if it means merely stepping away to let your unconscious work on it, the evidence does not support a reliable benefit.
Worth knowingThe two literatures are easy to conflate in popular advice, and 'sleep on it' is often defended by citing unconscious-thought research — the weaker of the two bodies of evidence — rather than the sleep research, which is the stronger one.
Graded against
- Dijksterhuis et al. 2006experiment · Science ·
10.1126/science.1121629 context - Nieuwenstein et al. 2015meta-analysis · Judgm. decis. mak. ·
10.1017/s1930297500003144 counters - Acker 2008meta-analysis · Judgm. decis. mak. ·
10.1017/s1930297500000863 counters - Wagner et al. 2004RCT · Nature ·
10.1038/nature02223 supports - KILLGORE et al. 2006experiment · Journal of Sleep Research ·
10.1111/j.1365-2869.2006.00487.x supports - Walker & Stickgold 2006review · Annu. Rev. Psychol. ·
10.1146/annurev.psych.56.091103.070307 context
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“The first number named anchors the outcome of a negotiation.”
This is one of the better-supported classic decision biases. In controlled negotiation experiments, whichever side named the first number obtained a better final outcome, and that first offer strongly predicted the eventual settlement price, in both face-to-face and email negotiations. Anchoring more generally is also one of the few classic psychology effects that held up in a large, preregistered replication project spanning many independent samples worldwide, with very large effects — unlike several other classic findings tested in that same project, which failed to replicate. But the negotiation advantage is not unconditional: the same research that established it also found the first-offer advantage was eliminated, not just reduced, when the other side deliberately focused on their own walk-away alternatives, the other side's likely reservation price, or their own target figure instead of anchoring on the number just named. And it is not a novice-only effect — professional real-estate agents were anchored by a manipulated listing price just as much as inexperienced participants, so expertise alone does not confer immunity.
Worth knowingThe first-offer advantage was eliminated, not merely reduced, when the other party used one of three specific counter-strategies: focusing on their own alternatives, the other side's likely reservation price, or their own target outcome — so a first offer is a strong default, not a guaranteed win.
Graded against
- Tversky & Kahneman 1974commentary · Science ·
10.1126/science.185.4157.1124 context - Galinsky & Mussweiler 2001experiment · Journal of Personality and Social Psychology ·
10.1037/0022-3514.81.4.657 supports - Klein et al. 2014replication · Social Psychology ·
10.1027/1864-9335/a000178 supports - Furnham & Boo 2011review · The Journal of Socio-Economics ·
10.1016/j.socec.2010.10.008 context - Northcraft & Neale 1987experiment · Organizational Behavior and Human Decision Processes ·
10.1016/0749-5978(87)90046-x context
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“Zone 2 training is the best way to build an endurance base.”
Zone two training reliably builds an aerobic base — low-intensity endurance exercise increases skeletal muscle mitochondrial content and fat-oxidation capacity (the cellular engines that burn fat and sustain aerobic effort), and elite endurance athletes structure the bulk of their training at low intensity. But the superlative 'best way' is unsupported: the only direct comparative trial found that a polarized model pairing low-intensity volume with high-intensity sessions outperformed the high-volume low-intensity approach on every key performance measure, and a recent review found no controlled evidence that zone two beats higher intensities for the mitochondrial adaptations the claim implies it excels at.
Worth knowingThe mechanistic and observational case for zone two as a core base-building method is solid: regular endurance exercise drives increases in skeletal muscle mitochondrial content and respiratory capacity, and elite endurance athletes consistently structure the bulk of their training at low intensity.
Graded against
- Holloszy & Coyle 1984review · Journal of Applied Physiology ·
10.1152/jappl.1984.56.4.831 context - San-Millán & Brooks 2018cross-sectional · Sports Med ·
10.1007/s40279-017-0751-x supports - Seiler 2010review · International Journal of Sports Physiology and Performance ·
10.1123/ijspp.5.3.276 context - Stöggl & Sperlich 2014RCT · Front. Physiol. ·
10.3389/fphys.2014.00033 counters - Storoschuk et al. 2025review · Sports Med ·
10.1007/s40279-025-02261-y counters
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
- Trend breakdownIs training slow really the key to fitness and longevity?
- GlossaryMitochondria: Definition, Function and How Training Builds Cellular Energy Capacity
- GlossaryVO2 Max: Definition, Function & Why It Predicts All-Cause Mortality
- Deep diveMitochondrial Health: The Science of Sustainable Energy & Peak Cognitive Output
“Blue-light glasses improve sleep.”
High-blocking amber lenses worn before bed may benefit people with existing insomnia symptoms, but the broad consumer claim is not supported by the overall evidence base. Meta-analytic syntheses find no statistically significant effect on objective sleep outcomes across healthy and mixed adult populations.
Worth knowingThe mechanistic basis is real: evening blue-light exposure suppresses melatonin and delays sleep onset, giving glasses a plausible theoretical rationale.
Graded against
- Shechter et al. 2018RCT · Journal of Psychiatric Research ·
10.1016/j.jpsychires.2017.10.015 supports - Singh et al. 2023systematic review · Cochrane Database of Systematic Reviews ·
10.1002/14651858.cd013244.pub2 counters - Luna-Rangel et al. 2025meta-analysis · Front. Neurol. ·
10.3389/fneur.2025.1699303 counters - Bigalke et al. 2021RCT · Sleep Health ·
10.1016/j.sleh.2021.02.004 counters - Silvani et al. 2022systematic review · Front. Physiol. ·
10.3389/fphys.2022.943108 context
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“Caffeine after 2pm ruins your sleep.”
Afternoon caffeine can reduce sleep time and quality for typical consumers — a standard coffee (107 mg) is best consumed at least 8.8 h before bedtime. But 'ruins' overstates: at 100 mg, evidence shows no significant sleep disruption at the pre-bed timing tested, and individual metabolism varies 5-6 fold, making a flat time rule unreliable.
Worth knowingThe 8.8 h cutoff applies specifically to a 107 mg coffee; higher-dose products require an even earlier cutoff.
Graded against
- Gardiner et al. 2023meta-analysis · Sleep Medicine Reviews ·
10.1016/j.smrv.2023.101764 supports - Gardiner et al. 2025RCT · SLEEP ·
10.1093/sleep/zsae230 counters - Grzegorzewski et al. 2022systematic review · Front. Pharmacol. ·
10.3389/fphar.2021.752826 context
Every source Crossref-verified — tap to open
“Intermittent fasting burns more fat than ordinary calorie restriction.”
Intermittent fasting is a genuine fat-loss tool — it outperforms no diet — but when calories are equated, head-to-head trials and pooled meta-analyses consistently find it produces no more fat loss than continuous calorie restriction. The popular claim mistakes 'IF works' for 'IF works better.'
Worth knowingMultiple randomised trials — including alternate-day fasting and sixteen-eight time-restricted eating protocols — find no significant between-group difference in weight or fat loss compared with calorie-matched continuous restriction.
Graded against
- Trepanowski et al. 2017RCT · JAMA Intern Med ·
10.1001/jamainternmed.2017.0936 counters - Lowe et al. 2020RCT · JAMA Intern Med ·
10.1001/jamainternmed.2020.4153 counters - Cioffi et al. 2018meta-analysis · J Transl Med ·
10.1186/s12967-018-1748-4 counters - Hamsho et al. 2025meta-analysis · Nutrition, Metabolism and Cardiovascular Diseases ·
10.1016/j.numecd.2024.103805 counters - Liu et al. 2022meta-analysis · The Journal of Clinical Endocrinology & Metabolism ·
10.1210/clinem/dgac570 context - Sutton et al. 2018RCT · Cell Metabolism ·
10.1016/j.cmet.2018.04.010 context
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“It takes 23 minutes to refocus after an interruption.”
The specific figure of twenty-three minutes does not appear in either of the two peer-reviewed papers it is almost universally credited to. A field study of information workers measured an average of 11 minutes 4 seconds spent in a working sphere before switching to another task or being interrupted, and a separate average of 25 minutes 26 seconds for same-day resumption of an interrupted task -- though that resumption figure was reached only after roughly two other working spheres intervened, so it is not a single interruption's recovery time. A separate lab study found the opposite of the popular framing: people who were interrupted completed the same task in less time than an uninterrupted baseline of 22.77 minutes, at the cost of more self-reported stress. An independently logged field study at a technology company found roughly 10 minutes to handle an alert plus a further 10 to 15 minutes to return to focused work; for an immediate response to an email alert specifically, the resumption phase averaged 16 minutes 33 seconds. Real interruption costs are well documented across all three studies and roughly cluster in a ten-to-twenty-five-minute range; the specific number popularly quoted is not.
Worth knowingThe nearest a source comes to the popular number -- the field study's 25 minutes 26 seconds same-day resumption figure -- is a same-day-resumption average reached only after roughly two intervening working spheres, not a clean single-interruption recovery time, so even the closest real number differs in what it actually measures.
Graded against
- Mark et al. 2005field study · Proceedings of the SIGCHI Conference on Human Factors in Computing Systems ·
10.1145/1054972.1055017 counters - Mark et al. 2008experiment · Proceedings of the SIGCHI Conference on Human Factors in Computing Systems ·
10.1145/1357054.1357072 counters - Iqbal & Horvitz 2007field study · Proceedings of the SIGCHI Conference on Human Factors in Computing Systems ·
10.1145/1240624.1240730 context
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“Judges grant fewer paroles as they get hungrier before a break.”
The original Israeli parole-board study found favourable rulings fell from around 65% at the start of a session to nearly zero by the end, then jumped back to around 65% after each food break — the 'hungry judge' pattern behind this claim. But this is a genuinely contested single-dataset finding, not a settled effect. A reanalysis of the same case records found that scheduling is not random: unrepresented prisoners are systematically heard last within each session, right before a break, and are less likely to be granted parole regardless of any hunger effect. The original authors dispute this, reporting that the meal-break pattern survives once legal representation is added as a control. Separately, a later simulation study argues that even a purely rational, non-fatigued judge — one who simply takes longer to write up a grant than a denial, with only limited foresight about session length — could produce an order effect of a similar size through statistical artifact alone. A swing this large, from a majority of favourable rulings to almost none, is an unusually big effect for any single psychological mechanism to carry on its own.
Worth knowingThe dispute is not resolved: the original authors' reply reports the meal-break pattern survives after controlling for legal representation, while the reanalysis authors maintain the pattern is better explained by non-random case scheduling than by hunger or fatigue.
Graded against
- Danziger et al. 2011field study · Proc. Natl. Acad. Sci. U.S.A. ·
10.1073/pnas.1018033108 supports - Weinshall-Margel & Shapard 2011commentary · Proc. Natl. Acad. Sci. U.S.A. ·
10.1073/pnas.1110910108 counters - Danziger et al. 2011study · Proc. Natl. Acad. Sci. U.S.A. ·
10.1073/pnas.1112190108 supports - Glöckner 2016commentary · Judgm. decis. mak. ·
10.1017/s1930297500004812 counters
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“Low HRV means you should skip your workout.”
Heart rate variability (HRV)-guided training — adjusting daily session intensity based on readiness — has genuine support across multiple RCTs and a meta-analysis: it produces similar or superior fitness adaptations compared to fixed periodization plans. The kernel of truth ends there. Every studied protocol responds to a depressed HRV reading by prescribing a low-intensity session, not complete rest; furthermore, single-day readings are unreliable noise — protocols require multi-day rolling averages to detect genuine readiness shifts. The specific consumer rule 'low HRV = skip your workout' maps onto no tested protocol and exceeds what the evidence actually endorses.
Worth knowingNo published RCT protocol instructs athletes to skip their session on low-HRV days — every studied protocol prescribes a low-intensity session as the readiness-informed response to a depressed reading, not complete rest.
Graded against
- VESTERINEN et al. 2016RCT · Medicine & Science in Sports & Exercise ·
10.1249/mss.0000000000000910 supports - Javaloyes et al. 2019RCT · International Journal of Sports Physiology and Performance ·
10.1123/ijspp.2018-0122 supports - Granero-Gallegos et al. 2020meta-analysis · IJERPH ·
10.3390/ijerph17217999 supports - Manresa-Rocamora et al. 2021meta-analysis · IJERPH ·
10.3390/ijerph181910299 counters - Buchheit 2014review · Front. Physiol. ·
10.3389/fphys.2014.00073 counters
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“Moderate alcohol is good for your heart.”
Decades of observational studies produced a consistent J-shaped association between moderate drinking and cardiovascular risk — the real foundation of the popular belief. Two more rigorous methodological approaches undermine that apparent benefit: correcting for the misclassification of former drinkers as abstainers eliminates the mortality advantage, and Mendelian randomization studies find that genetic variants predicting lower alcohol intake are associated with better cardiovascular outcomes, not worse. The best available causal evidence indicates that lower alcohol consumption benefits cardiovascular health: it does not merely fail to support a net heart benefit — it points the opposite way, toward harm.
Worth knowingThe observational J-curve literature is large and internally consistent: dozens of prospective cohort studies found that light-to-moderate drinkers had lower cardiovascular mortality and coronary heart disease incidence than abstainers. This is the genuine observational kernel the popular claim rests on.
Graded against
- Ronksley et al. 2011meta-analysis · BMJ ·
10.1136/bmj.d671 supports - Holmes et al. 2014meta-analysis · BMJ ·
10.1136/bmj.g4164 counters - Stockwell et al. 2016systematic review · J. Stud. Alcohol Drugs ·
10.15288/jsad.2016.77.185 counters - Millwood et al. 2019RCT · The Lancet ·
10.1016/s0140-6736(18)31772-0 counters - Wood et al. 2018meta-analysis · The Lancet ·
10.1016/s0140-6736(18)30134-x context
Every source Crossref-verified — tap to open
“Money stops buying happiness after about $75,000 a year.”
The '$75,000 happiness plateau' comes from the original Kahneman & Deaton study, which found day-to-day emotional well-being stopped improving with income above roughly $75,000 a year, even though people's overall life satisfaction kept climbing with income regardless of level. But the popular, sweeping version of this claim has not held up as later, larger research refined the picture. A later study using over one million real-time smartphone reports of in-the-moment feelings found happiness kept rising with income on an equally steep slope above $80,000 as below it -- no plateau at all for most people. A subsequent adversarial collaboration, in which the original researcher reanalysed the newer data himself, resolved the disagreement: the flattening pattern is real, but it applies only to the least happy roughly 15% of people, whose happiness levels off abruptly around $100,000; for everyone else, happiness keeps rising steadily with income well beyond $75,000. A separate global study estimated the threshold differently again -- around $95,000 for overall life evaluation and $60,000 to $75,000 for emotional well-being specifically, with the exact figure shifting by which well-being measure is used and by world region. So 'money stops buying happiness after about $75,000' overstates a narrower, more conditional finding: for most people happiness keeps climbing with income well past that point, and a hard ceiling near $75,000 applies only to an unhappy minority, and even then closer to $100,000.
Worth knowingThe original finding was specific to day-to-day emotional well-being, not overall life satisfaction: the same study reported that high income continued to raise life evaluation with no plateau, even as emotional well-being stalled beyond ~$75,000.
Graded against
- Kahneman & Deaton 2010cross-sectional · Proc. Natl. Acad. Sci. U.S.A. ·
10.1073/pnas.1011492107 supports - Killingsworth 2021study · Proc. Natl. Acad. Sci. U.S.A. ·
10.1073/pnas.2016976118 counters - Killingsworth et al. 2023experiment · Proc. Natl. Acad. Sci. U.S.A. ·
10.1073/pnas.2208661120 counters - Jebb et al. 2018cross-sectional · Nat Hum Behav ·
10.1038/s41562-017-0277-0 context
Every source Crossref-verified — tap to open
“Psychological safety drives team performance.”
There is a genuine kernel of truth here, but 'drives' claims more direct causal force than the evidence supports. In the foundational field study, psychological safety predicted team learning behaviour, and it was learning behaviour — not psychological safety on its own — that predicted team performance: once both were entered into the same statistical model, psychological safety's own direct effect on performance became non-significant (B = .25, p = .42), while learning behaviour remained a significant predictor (B = .60, p < .05). An independent replication two decades later, in South Korean sales teams rather than a US manufacturer, found the same pattern: no significant direct effect of psychological safety on team effectiveness (β = 0.037), with the relationship holding only through a full double-mediation pathway via learning behaviour and team efficacy. A further meta-analysis shows the link is also task-contingent, stronger in complex, creative, sensemaking-heavy work and possibly absent where tasks do not require learning. So the popular claim's real mechanism — psychological safety enabling the learning behaviours that in turn improve performance, mainly in complex or creative work — is well supported; the popular claim's implied direct, universal 'drive' is not.
Worth knowingThe performance link is mediated, not direct: psychological safety predicts team learning behaviour, and it is learning behaviour that predicts performance, not psychological safety acting on its own. Once both are entered into the same model, psychological safety's direct effect on performance is non-significant while learning behaviour's effect remains significant.
Graded against
- Edmondson 1999survey · Administrative Science Quarterly ·
10.2307/2666999 supports - Kim et al. 2020survey · Front. Psychol. ·
10.3389/fpsyg.2020.01581 counters - Sanner & Bunderson 2015meta-analysis · Organizational Psychology Review ·
10.1177/2041386614565145 counters - Frazier et al. 2017meta-analysis · Personnel Psychology ·
10.1111/peps.12183 context - Newman et al. 2017review · Human Resource Management Review ·
10.1016/j.hrmr.2017.01.001 context
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“Rewarding people for something they enjoy destroys their motivation for it.”
There is a real 'undermining effect', but it is narrower than the popular claim suggests. A large meta-analysis found that engagement-contingent, completion-contingent, and performance-contingent rewards significantly reduced people's later free-choice engagement and interest in an activity (d = -0.40, -0.36, and -0.28, respectively). But the same meta-analysis found the opposite for a different kind of reward: positive verbal feedback increased both later free-choice engagement (d = 0.33) and self-reported interest (d = 0.31). A rival meta-analysis concluded that, overall, reward does not decrease intrinsic motivation, and traced the only reliable negative effect to expected, tangible rewards given to people simply for doing a task. So 'rewarding people destroys their motivation' holds for a specific, narrow kind of reward — tangible, expected, and tied to doing or finishing the task — not for reward in general, and not for praise.
Worth knowingOther reviews in this literature converge on the same narrower conclusion from the opposite rhetorical direction: detrimental reward effects are real, but occur only under highly restricted, easily avoidable conditions, and reward can even be used to boost generalised creativity.
Graded against
- Deci et al. 1999meta-analysis · Psychological Bulletin ·
10.1037/0033-2909.125.6.627 supports - Cameron & Pierce 1994meta-analysis · Review of Educational Research ·
10.3102/00346543064003363 counters - Eisenberger & Cameron 1996review · American Psychologist ·
10.1037/0003-066X.51.11.1153 counters - Cerasoli et al. 2014meta-analysis · Psychological Bulletin ·
10.1037/a0035661 context
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“Simply having your phone nearby drains your attention.”
The original study, published in the Journal of the Association for Consumer Research, found that people scored worse on tests of working memory and reasoning when their own smartphone sat nearby on the desk rather than in another room, even though it was silenced and never touched. But the claim has not held up well since it was published. A pre-registered study using the exact same tasks and the exact same phone-location conditions found no difference in performance at all. A separate study testing short-term and prospective memory instead found no overall effect of phone presence either. A large meta-analysis pooling many studies since then did find a negative effect overall, but one that varies substantially depending on which cognitive skill is being tested, rather than a single clean 'brain drain'. So the mere-presence effect is real in the sense that it was published with a significant result, but the most direct replication attempts have failed to reproduce it, and the wider evidence since is mixed rather than settled.
Worth knowingPhone notifications are a different, better-established distraction mechanism from mere presence: pings measurably disrupt attention even without the phone being touched, but that is a separate claim from the passive 'sitting nearby' effect this verdict addresses.
Graded against
- Ward et al. 2017experiment · Journal of the Association for Consumer Research ·
10.1086/691462 supports - Hartmann et al. 2020replication · Consciousness and Cognition ·
10.1016/j.concog.2020.103033 counters - Ruiz Pardo & Minda 2022replication · Acta Psychologica ·
10.1016/j.actpsy.2022.103717 counters - Böttger et al. 2023meta-analysis · Behavioral Sciences ·
10.3390/bs13090751 context - Stothart et al. 2015experiment · Journal of Experimental Psychology: Human Perception and Performance ·
10.1037/xhp0000100 context
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“Sitting is the new smoking.”
Sitting for long periods is a real, independent mortality risk, but it is far smaller than smoking's, and unlike smoking's risk it is not fixed — daily physical activity substantially reduces or removes it. The best available meta-analysis puts sedentary behavior's hazard ratio for all-cause mortality at 1.22, and a separate, independent meta-analysis found a closely matching hazard ratio of 1.240 for sedentary time and all-cause mortality. By contrast, current smokers carry a relative risk of death of 2.80 (men) and 2.76 (women) compared with never-smokers — more than double sitting's excess risk — and a second, independent meta-analysis of older adults found smokers' relative mortality was 1.83 compared with never-smokers, corroborating that smoking's own risk sits well above sitting's from a different data source. The equivalence breaks down further because sitting's risk, unlike smoking's, is not fixed: a pooled analysis of over a million adults found that about 60-75 minutes of moderate-intensity physical activity per day appears to eliminate the increased mortality risk associated with high sitting time. 'Sitting is the new smoking' borrows real evidence of harm but claims a magnitude and an irreversibility that the evidence does not support.
Worth knowingSedentary behavior's association with mortality is genuine and independently significant — this is not a case of sitting being harmless. The issue is specifically with the popular magnitude/equivalence claim, not the existence of a real effect.
Graded against
- Vallance et al. 2018commentary · Am J Public Health ·
10.2105/AJPH.2018.304649 counters - Gellert et al. 2012meta-analysis · Arch Intern Med ·
10.1001/archinternmed.2012.1397 context - Biswas et al. 2015meta-analysis · Ann Intern Med ·
10.7326/M14-1651 supports - Ekelund et al. 2016meta-analysis · The Lancet ·
10.1016/S0140-6736(16)30370-1 counters - Chen et al. 2019meta-analysis · European Journal of Public Health ·
10.1093/eurpub/cky121 context
Every source Crossref-verified — tap to open
“Standing desks burn significantly more calories than sitting.”
Standing at a desk burns a fraction of a calorie more per minute than sitting — a real but trivially small difference that falls far below what 'significantly' implies to a lay reader. No randomised trial has shown meaningful caloric benefit or body-composition change from standing desk use alone.
Worth knowingThe best available meta-analysis confirms a real energy expenditure advantage for standing, but the effect is so small it is unlikely to produce detectable weight change in practice, especially given evidence that workers may compensate with increased sedentary time outside work.
Graded against
- Saeidifard et al. 2018meta-analysis · Eur J Prev Cardiolog ·
10.1177/2047487317752186 context - Burns et al. 2017experiment · Hum Factors ·
10.1177/0018720817719167 counters - Mantzari et al. 2019RCT · Preventive Medicine Reports ·
10.1016/j.pmedr.2018.11.012 counters - MacEwen et al. 2015systematic review · Preventive Medicine ·
10.1016/j.ypmed.2014.11.011 context
Every source Crossref-verified — tap to open
“Taking notes by hand beats typing them.”
The original study found that students who took notes on laptops performed worse on conceptual questions than students who took notes longhand, and pinned this on laptop note-takers transcribing lectures word for word instead of processing and reframing the material in their own words. But the specific performance advantage has not held up well since. A large preregistered replication reproduced the behavioural pattern behind the claim, laptop users writing more and copying more verbatim, but did not find that longhand users actually scored better on the later test, and an accompanying meta-analysis of several similar studies echoed the same null result. A separate preregistered replication, which added an e-writer condition and even a no-notes condition, found no consistent difference between any of the note-taking methods. So the mechanism behind the claim, that verbatim transcription is a shallower way to process material, does appear to be real and reproducible, but the specific benefit, that handwriting beats typing on a later test, has not held up under direct replication.
Worth knowingThe clearest replicated finding is behavioural rather than a performance benefit: laptop note-takers reliably write more and copy more verbatim, but this has not been shown to reliably translate into worse quiz performance.
Graded against
- Mueller & Oppenheimer 2014experiment · Psychol Sci ·
10.1177/0956797614524581 supports - Morehead et al. 2019replication · Educ Psychol Rev ·
10.1007/s10648-019-09468-2 counters - Urry et al. 2021meta-analysis · Psychol Sci ·
10.1177/0956797620965541 counters
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“You can only maintain about 150 relationships.”
Dunbar's own foundational analysis, extrapolating from a primate neocortex-to-group-size regression, predicts a human group size of 147.8 -- the figure popularly rounded to 150 -- and Dunbar cross-checked it against hunter-gatherer groupings, farming-community splits and army units that cluster near the same range. But even Dunbar's own confidence interval around that estimate was wide, running from roughly a hundred to well over two hundred, and a modern re-analysis using updated primate datasets and different statistical methods found the underlying regression cannot reliably produce one number at all, concluding that a cognitive limit on human group size cannot be derived this way. Independent evidence does support the idea of tiered relationship layers: one large analysis of mobile-phone call patterns found strong evidence for a layered social structure broadly consistent with Dunbar's tiers -- inner five, middle fifteen, outer 150 -- though with large variability in the middle layers, and a large online survey found real person-to-person variation in how people allocate relational energy across those layers, with extraversion unrelated to the pattern. Taken together, 150 is a genuine, historically corroborated central estimate from Dunbar's own methodology, not a fabricated number, but treating it as a precise, hard ceiling overstates the evidence: the original analysis carried a very wide margin of error, and a modern statistical re-analysis could not derive a reliable single number at all.
Worth knowingThe commonly cited number traces to Dunbar's original Journal of Human Evolution paper, which is paywalled and could not be independently verified from any available source; the 147.8 figure and its confidence interval used here instead come from Dunbar's own companion paper, which reports the identical regression and result with fully verified text.
Graded against
- Dunbar 1993commentary · Behav Brain Sci ·
10.1017/s0140525x00032325 supports - Lindenfors et al. 2021study · Biology Letters ·
10.1098/rsbl.2021.0158 counters - Mac Carron et al. 2016observational · Social Networks ·
10.1016/j.socnet.2016.06.003 context - Li et al. 2025survey · PLoS ONE ·
10.1371/journal.pone.0319604 context
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“You need 10,000 steps a day for health.”
Walking more steps each day cuts mortality risk — a consistent finding across large cohorts and meta-analyses. But the ten-thousand-step target has no physiological basis; it traces to a decades-old Japanese pedometer marketing campaign, not to physiology. The mortality-benefit curve levels off for adults over sixty at roughly 6,000–8,000 steps per day, and at approximately 7,500 steps per day in older women specifically. The claim that you 'need' ten thousand steps overstates the evidence: the majority of the survival benefit accrues well below that figure.
Worth knowingThe mortality-benefit plateau for older adults (aged ≥60 years) falls at approximately 6,000–8,000 steps per day according to a meta-analysis of fifteen international cohorts; the ten-thousand-step target lies above this range without delivering proportionally greater survival benefit in that age group.
Graded against
- Lee et al. 2019cohort · JAMA Intern Med ·
10.1001/jamainternmed.2019.0899 counters - Paluch et al. 2022meta-analysis · The Lancet Public Health ·
10.1016/s2468-2667(21)00302-9 counters - Banach et al. 2023meta-analysis · European Journal of Preventive Cardiology ·
10.1093/eurjpc/zwad229 context - Saint-Maurice et al. 2020cohort · JAMA ·
10.1001/jama.2020.1382 context
Every source Crossref-verified — tap to open
“93% of communication is non-verbal.”
There is no peer-reviewed evidence for a general claim that 93% of communication is non-verbal. The figure traces to two 1967 experiments in which 37 female psychology majors judged a speaker's feelings from a single ambiguous spoken word ('maybe') paired with varying vocal tones and facial photographs — a design built to make the verbal channel almost irrelevant, not a test of communication as a whole, and no single study measured all three channels together. Mehrabian himself has called extending his 7% figure to all verbal communication absurd, and a widely used nonverbal-communication handbook states plainly that the popular 93% estimate rests on faulty analysis. Yet a content analysis of 79 public websites citing the figure found 63 of them (80%) using it as a general statement about communication overall — the opposite of what the original narrow research actually measured.
Worth knowingThe two foundational 1967 studies (Mehrabian & Ferris; Mehrabian & Wiener) could not be quoted directly for this verdict — Crossref carries no abstract for either record and no accessible full text was located — so the scope facts here are sourced via a later peer-reviewed paper that itself quotes both originals, and Mehrabian's own correspondence, verbatim; the 1967 papers were not read directly.
Graded against
- Lapakko 2015review · Communication and Theater Association of Minnesota Journal ·
10.56816/2471-0032.1000 context - Trimboli & Walker 1987RCT · J Nonverbal Behav ·
10.1007/bf00990236 counters - Lapakko 1997review · Communication Education ·
10.1080/03634529709379073 counters - Hall et al. 2019review · Annu. Rev. Psychol. ·
10.1146/annurev-psych-010418-103145 context
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“Creatine harms your kidneys.”
Serum creatinine rises with creatine supplementation — because creatine is metabolised to creatinine — but this is a metabolic artefact, not organ damage. GFR, the true measure of kidney filtration, is unchanged across meta-analyses of randomised trials and a gold-standard radioisotope clearance RCT. The popular claim conflates a benign biomarker shift with kidney harm.
Worth knowingThese studies predominantly recruited healthy adults; individuals with pre-existing renal disease were largely excluded. The refutation applies to healthy populations; caution in that sub-group cannot be assumed away from this evidence base.
Graded against
- Naeini et al. 2025meta-analysis · BMC Nephrol ·
10.1186/s12882-025-04558-6 counters - Lugaresi et al. 2013RCT · Journal of the International Society of Sports Nutrition ·
10.1186/1550-2783-10-26 counters - Kreider et al. 2017review · Journal of the International Society of Sports Nutrition ·
10.1186/s12970-017-0173-z counters
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“Open-plan offices increase collaboration.”
The strongest direct evidence runs against this claim. Two intervention field studies used wearable sociometric badges plus email and messaging logs to measure face-to-face interaction before and after a switch to open-plan offices at large corporate headquarters: contact fell sharply post-redesign -- a drop of roughly 70% -- with electronic messaging rising to compensate, the opposite of the collaboration boost the redesign was meant to produce. A systematic review pooling a large body of the existing comparative literature on open-plan versus enclosed offices corroborates this at scale, finding open-plan associated with more negative outcomes across health, satisfaction, productivity and social-relationship measures. One narrower single-firm survey found a different pattern for one specific slice of the picture -- high office density and low privacy were positively linked to expressive personal relations among coworkers -- but the same sample showed worse satisfaction, engagement and well-being overall, so it does not amount to independent, generalisable support for the claim as stated. No sound peer-reviewed evidence surfaced showing open-plan conversions increase collaboration; the industry claims that circulate to that effect are anecdotal and were excluded from this evidence base.
Worth knowingThe two headline sociometric-badge field studies were natural experiments drawn from a single broad corporate-redesign context, not randomised trials, and used small samples -- at most about a hundred employees per study -- so they demonstrate strong within-company change but limited generalisability across industries or office types.
Graded against
- Bernstein & Turban 2018experiment · Phil. Trans. R. Soc. B ·
10.1098/rstb.2017.0239 counters - James et al. 2021systematic review · Sage Open ·
10.1177/2158244020988869 counters - Węziak-Białowolska et al. 2018cross-sectional · Front. Psychol. ·
10.3389/fpsyg.2018.02178 context
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“Opposites attract.”
The evidence overwhelmingly contradicts 'opposites attract' as a general law of interpersonal attraction. The largest synthesis to date -- a meta-analysis of many personal traits across a large body of studies of real couples, cross-checked against a national biobank sample -- finds that partners are overwhelmingly similar to each other, not opposite, across nearly every trait examined, from attitudes and values to habits and education; genuinely opposite pairings are the rare exception, not the rule. A separate meta-analysis of laboratory and field attraction research reaches the same conclusion from a different angle: both real and perceived similarity between two people are large, reliable predictors of how attracted they feel to each other, echoing decades of foundational research establishing that attitude similarity, not dissimilarity, drives interpersonal evaluation. The one genuine exception is narrow: in controlled interaction experiments, people paired with a partner who took the opposite role on the specific dimension of dominance versus submission reported more satisfaction with that single interaction than people paired with a similarly dominant or submissive partner -- real evidence that complementarity can work, but scoped to one personality dimension and to satisfaction with an interaction, not to romantic attraction, mate choice, or lasting relationships broadly.
Worth knowingThe complementarity finding that offers real support for 'opposites attract' is confined to a single interpersonal dimension -- dominance versus submission in a specific interaction -- not personality or values broadly, and the same study found that satisfied participants in complementary pairings still perceived their partner as similar to themselves, undercutting a clean opposites-versus-similar story.
Graded against
- Horwitz et al. 2023meta-analysis · Nat Hum Behav ·
10.1038/s41562-023-01672-z counters - Montoya et al. 2008meta-analysis · Journal of Social and Personal Relationships ·
10.1177/0265407508096700 counters - Dryer & Horowitz 1997experiment · Journal of Personality and Social Psychology ·
10.1037/0022-3514.72.3.592 supports - Byrne 1997commentary · Journal of Social and Personal Relationships ·
10.1177/0265407597143008 context
Every source Crossref-verified — tap to open
“The hot hand in basketball is a fallacy.”
The claim that the hot hand is a fallacy rests on an archival analysis of shot records that found essentially no advantage after a made shot: hit rate after a hit was, if anything, lower than after a miss (weighted mean: 51% versus weighted mean: 54%), and this was read for decades as proof that fans' and players' belief in shooting streaks is a cognitive illusion. Two independent, peer-reviewed statistical papers later showed that the method behind that finding is itself biased: conditioning on a streak of hits systematically undercounts hits in finite sequences, mechanically pulling the measured hit rate down. Reapplying a bias-corrected method to that same original archival data reverses the result, finding real streak shooting with large effect sizes, and the paper states directly that the hot hand is not a myth and the belief in it is not a cognitive illusion. Separately, shot-tracking data that models shot difficulty (defender distance, shot distance, game situation) finds a modest real hot-hand effect, in the range of 1.2 to 2.4 percentage points, once difficulty is held constant — though this comes from a working paper, not a peer-reviewed journal article. The picture is not fully settled: the newest peer-reviewed reanalysis, using a modern shot-difficulty model, finds a reverse hot-hand pattern (worse performance after streaks of difficult makes) rather than a clean confirmation of the folk hot-hand belief. Taken together, the specific statistical case for calling the hot hand a fallacy has been undercut by rigorous, peer-reviewed correction of the original method, even though the underlying phenomenon is still being actively remeasured.
Worth knowingTwo independent, peer-reviewed papers — one in Econometrica, one predating it in The American Statistician — prove, as a mathematical result rather than a matter of interpretation, that the classic method for detecting hot-hand streaks in finite shot sequences is a biased estimator that mechanically understates real streakiness.
Graded against
- Gilovich et al. 1985archival · Cognitive Psychology ·
10.1016/0010-0285(85)90010-6 supports - Miller & Sanjurjo 2018commentary · Econometrica ·
10.3982/ecta14943 counters - Stone 2012commentary · The American Statistician ·
10.1080/00031305.2012.676467 counters - Bocskocsky et al. 2014review · SSRN Journal ·
10.2139/ssrn.2481494 counters - Kondur & Shen 2026archival · International Journal of Sports Science & Coaching ·
10.1177/17479541251374788 context
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“Venting your anger gets it out of your system.”
The evidence does not support the popular idea that venting anger -- hitting something, yelling, or otherwise physically releasing it -- gets it out of your system. In the largest meta-analysis to date of anger-management activities, arousal-increasing activities such as venting were ineffective overall (g = -0.02, [-0.13, 0.09]), while arousal-decreasing activities such as breathing and mindfulness reliably reduced anger and aggression (g = -0.63, [-0.82, -0.43]); the researchers concluded these findings do not support the idea that venting anger is an effective anger-management activity. Controlled experiments point the same way: participants who hit a punching bag while ruminating on the person who angered them ended up angrier and more aggressive afterward than participants who were distracted or who did nothing at all -- doing nothing outperformed venting. Even participants who were first persuaded to believe in catharsis theory became more aggressive after venting than participants given an anti-catharsis message. A separate written-venting experiment found unstructured catharsis produced more aggressive behaviour afterward than a distraction task, and was no better than plain distraction at relieving anger. The one genuine counter-finding comes from a real-world study of offenders: venting did reduce aggression in forensic psychiatric offenders, but the identical venting opportunity did not help -- and by one measure worsened -- aggression in penitentiary offenders, marking this as a narrow, population-specific effect rather than support for the general 'get it out of your system' claim.
Worth knowingThis verdict draws on the largest meta-analysis to date of anger-management activities, two independent randomised experiments manipulating venting directly, a written-catharsis experiment, and a real-world offender study. The arousal-increasing/arousal-decreasing distinction from the meta-analysis (g = -0.02, [-0.13, 0.09] for arousal-increasing activities like venting versus g = -0.63, [-0.82, -0.43] for arousal-decreasing activities like breathing and mindfulness) explains the core pattern: it is not that anger-regulation strategies never work, it is that raising physiological arousal through venting does not work, while lowering it does.
Graded against
- Kjærvik & Bushman 2024meta-analysis · Clinical Psychology Review ·
10.1016/j.cpr.2024.102414 counters - Bushman 2002RCT · Pers Soc Psychol Bull ·
10.1177/0146167202289002 counters - Bushman et al. 1999experiment · Journal of Personality and Social Psychology ·
10.1037/0022-3514.76.3.367 counters - Zhan et al. 2021experiment · PsyCh Journal ·
10.1002/pchj.490 counters - Tonnaer et al. 2020experiment · Journal of Aggression, Maltreatment & Trauma ·
10.1080/10926771.2019.1575303 supports
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“Visualising success makes you more likely to achieve it.”
The evidence points the opposite way from the popular advice. Longitudinal studies found that people who spent more time positively fantasising about a desired future put in less effort and did worse at reaching it, weeks to years later — while people who judged the outcome as realistically likely (a different mental act from fantasising about it) did better. Follow-up experiments identified a likely reason: imagining the successful outcome measurably lowers the physiological and behavioural energy needed to pursue it. A separate research programme that distinguished simulating the process of reaching a goal from simulating its successful completion found the same pattern: process simulation produced progress toward goals, but envisioning successful completion of the goal did not. The one future-oriented technique with meta-analytic support for goal attainment is structurally different from pure success-visualisation — it works only when the positive fantasy is deliberately paired with a real obstacle and a plan for overcoming it.
Worth knowingThis verdict rests on two independent research programmes: Oettingen's lab (a four-study longitudinal cohort finding that positive fantasies predicted lower effort and worse attainment, and a four-experiment causal follow-up identifying reduced energy as the mechanism) and Taylor and colleagues' programme distinguishing 'process' from 'outcome' simulation, which found that mental simulation of the process for reaching a goal produced progress toward it, while envisioning successful completion of the goal did not.
Graded against
- Oettingen & Mayer 2002cohort · Journal of Personality and Social Psychology ·
10.1037/0022-3514.83.5.1198 counters - Kappes & Oettingen 2011RCT · Journal of Experimental Social Psychology ·
10.1016/j.jesp.2011.02.003 counters - Taylor et al. 1998review · American Psychologist ·
10.1037/0003-066X.53.4.429 counters - Wang et al. 2021meta-analysis · Front. Psychol. ·
10.3389/fpsyg.2021.565202 context
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“Willpower is a finite resource that gets used up.”
The idea that willpower is a single, finite tank that empties with use rests on small early laboratory experiments and a first meta-analysis of that literature — a later bias-correction re-analysis of that same meta-analysis described it as having concluded the depletion effect was 'robust and medium in magnitude (d = 0.62)'. That headline result has not held up under scrutiny. When the very same dataset was re-analysed using methods built to detect and correct for publication bias, the depletion effect became statistically indistinguishable from zero. Two large, independent, pre-registered multi-laboratory replication projects then tested the effect directly, each using a different protocol and far larger combined samples than any single study in the original literature. The first found a small pooled effect with a 95% confidence interval spanning zero (d = 0.04, 95% CI [-0.07, 0.15]). The second used a Bayesian analysis and found the data favoured no effect over even a modest true effect. A popular fallback explanation — that depletion only appears in people who believe willpower is limited — has also failed a pre-registered direct replication of that specific claim, and a proposed biological mechanism, that exerting self-control burns through blood glucose, has been tested directly and not supported. On the strongest currently available evidence, a literal, finite, depletable willpower resource is not what the data show.
Worth knowingThe claim is usually defended by pointing to the original small lab studies and the first meta-analysis of that literature, which a later bias-correction re-analysis described as having concluded willpower depletion was 'robust and medium in magnitude (d = 0.62)' — but that meta-analysis has since been shown to be an artefact of publication bias once bias-correction methods were applied to its own dataset, and it has not survived two much larger, pre-registered, multi-laboratory replication attempts that used different methodologies and both converged on a null result.
Graded against
- Baumeister et al. 1998experiment · Journal of Personality and Social Psychology ·
10.1037/0022-3514.74.5.1252 supports - Hagger et al. 2010meta-analysis · Psychological Bulletin ·
10.1037/a0019486 supports - Carter & McCullough 2014meta-analysis · Front. Psychol. ·
10.3389/fpsyg.2014.00823 counters - Hagger et al. 2016replication · Perspect Psychol Sci ·
10.1177/1745691616652873 counters - Vohs et al. 2021replication · Psychol Sci ·
10.1177/0956797621989733 counters - Job et al. 2010cohort · Psychol Sci ·
10.1177/0956797610384745 context - Carruth et al. 2023replication · PLoS ONE ·
10.1371/journal.pone.0287911 counters - Molden et al. 2012experiment · Psychol Sci ·
10.1177/0956797612439069 counters
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“You must eat protein within 30 minutes of training or the workout is wasted.”
The strict version of this claim — that missing a fixed, narrow post-workout window makes a workout's protein 'wasted' — is directly contradicted by the strongest available evidence. A meta-regression pooling many resistance-training studies found that timing protein around a workout had no significant effect once total daily protein intake was accounted for, and identified total protein intake, not timing, as the key driver of muscle gains. A direct head-to-head trial giving trained men an identical protein dose either immediately before or immediately after lifting found no meaningful difference in strength, muscle size, or body composition, and concluded the useful window for protein intake may be several hours wide rather than a matter of minutes. A separate narrative review found the very existence of a fixed 'window of opportunity' is questionable and depends heavily on other factors, such as whether a pre-workout meal was already eaten. The one study in this evidence set that does show a timing effect looked at untrained elderly men: those given protein immediately after training built significantly more quadriceps muscle than a matched group given the same protein roughly two hours later — a population and a gap far removed from the specific claim of a thirty-minute cutoff in trained gym-goers. Skipping a shake right after the gym does not 'waste' a workout; what matters most is getting enough total protein across the day.
Worth knowingNone of the cited evidence tests a literal thirty-minute cutoff; the closest population-specific timing study compared protein taken immediately after training versus roughly two hours later in previously untrained elderly men, not a narrow window in trained adults.
Graded against
- Schoenfeld et al. 2013meta-analysis · Journal of the International Society of Sports Nutrition ·
10.1186/1550-2783-10-53 counters - Aragon & Schoenfeld 2013review · Journal of the International Society of Sports Nutrition ·
10.1186/1550-2783-10-5 counters - Schoenfeld et al. 2017RCT · PeerJ ·
10.7717/peerj.2825 counters - Esmarck et al. 2001RCT · The Journal of Physiology ·
10.1111/j.1469-7793.2001.00301.x supports - Morton et al. 2018meta-analysis · Br J Sports Med ·
10.1136/bjsports-2017-097608 context
Every source Crossref-verified — tap to open
“You need eight glasses of water a day.”
No scientific evidence supports the rule. A formal literature search found no proof that every person must 'drink at least eight glasses of water a day'; individual water needs vary substantially based on body size, activity, age, and climate.
Worth knowingHydration itself matters — the refuted element is the specific universal figure, not the importance of adequate fluid intake.
Graded against
- Valtin 2002systematic review · American Journal of Physiology-Regulatory, Integrative and Comparative Physiology ·
10.1152/ajpregu.00365.2002 counters - Yamada et al. 2022cross-sectional · Science ·
10.1126/science.abm8668 counters - Popkin et al. 2010review · Nutrition Reviews ·
10.1111/j.1753-4887.2010.00304.x counters
Every source Crossref-verified — tap to open
“Cold showers build discipline and willpower.”
This claim has not actually been tested. No peer-reviewed study measures discipline, willpower, or self-control as an outcome of taking cold showers. The two most-cited cold-shower studies measure something else entirely: one is a randomised trial that looked at sickness absence and quality of life, not discipline, and the other is an untested hypothesis paper about mood and depression, not willpower. The closest indirect evidence, a meta-analysis on training self-control through repeated effortful acts in general, found only a small-to-medium effect that shrank further once publication bias was corrected, and its own authors said the mechanism driving any such effect is poorly understood. So the honest answer is not that cold showers fail to build discipline, but that nobody has actually measured whether they do, which means the claim cannot be confirmed, refuted, or even graded as nuanced or overstated.
Worth knowingThe best-known cold-shower trial did find a real health benefit, a reduction in self-reported sickness absence, but health outcomes are not the same as discipline or willpower, and the trial did not test either.
Graded against
- Buijze et al. 2016RCT · PLoS ONE ·
10.1371/journal.pone.0161749 context - Shevchuk 2008commentary · Medical Hypotheses ·
10.1016/j.mehy.2007.04.052 context - Friese et al. 2017meta-analysis · Perspect Psychol Sci ·
10.1177/1745691617697076 context
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
“Mouth taping improves sleep quality.”
We decline to rule on this claim as stated. The only controlled evidence measures breathing proxies — apnoea events and snoring — in diagnosed mild-OSA habitual mouth-breathers, not sleep quality in general healthy sleepers, which is the claim's actual subject. Sleep quality was not a primary endpoint in any trial and no controlled study exists in healthy sleepers, so on current evidence the claim can be neither supported nor refuted.
Worth knowingAll positive evidence is confined to mild obstructive sleep apnoea patients who are habitual mouth-breathers; no controlled study has demonstrated benefit in healthy sleepers without a sleep-breathing disorder.
Graded against
- Lee et al. 2022study · Healthcare ·
10.3390/healthcare10091755 supports - Huang & Young 2015RCT · Otolaryngol.--head neck surg. ·
10.1177/0194599814559383 supports - Yang et al. 2024RCT · JAMA Otolaryngol Head Neck Surg ·
10.1001/jamaoto.2024.3319 counters - Rhee et al. 2025systematic review · PLoS One ·
10.1371/journal.pone.0323643 counters - Fangmeyer et al. 2025review · American Journal of Otolaryngology ·
10.1016/j.amjoto.2024.104545 context
Every source Crossref-verified — tap to open
How we grade — the methodology
Further on the site
No claims match.
Grades A–C reflect the strength of the underlying evidence, not the popularity of the claim — A means consistent high-quality evidence; C means limited or mixed — independent of whether the claim itself holds up. Declined means the evidence base was too thin or conflicted to grade responsibly — we say so rather than force a verdict.
Changelog
- 2026-09-04 — 1 claim added
- 2026-09-03 — 1 claim added
- 2026-09-02 — 1 claim added
- 2026-09-01 — 1 claim added
- 2026-08-31 — 1 claim added
- 2026-08-30 — 1 claim added
- 2026-07-17 — 30 claims added
- 2026-07-07 — 12 claims added
- 2026-07-05 — 8 claims added