Bearings

Aug 31, 2026
Aug 31, 2026
SOCIAL PSYCHOLOGY · a field experiment, contested by a 2021 replication attempt · 2021

A Photo Of Eyes Raised Museum Donations

Setup. Eyes are among the strongest social signals humans read, and a body of psychology research holds that even a picture of eyes can trigger the sense of being watched. The claim matters beyond the lab: it implies a poster, not a camera, could nudge people toward honesty and generosity.

The finding. In one field test at a University of Virginia children's museum, a donation box sign alternated weekly between images of eyes and images of neutral objects like chairs, over 28 weeks and more than 34,100 visitors. Museum patrons donated more in weeks the sign showed eyes than in weeks it showed an inanimate object. Similar designs have cut littering and bicycle theft and reduced dishonesty in economic games, and the effect holds even when people are told their choices are anonymous.

Where it stands. The donation and littering studies are real field data, not lab artifacts, which is the strongest evidence tier here. The effect is not settled. A 2021 attempted replication with a larger, more diverse sample found no effect and specifically tested for individual differences the original studies missed.

Caroline Kelsey · Wikipedia · 2021
MEMORY RESEARCH · a 2025 meta-analysis overturning a century-old finding · 2025

The Memory Effect Behind Cliffhangers Did Not Replicate

Setup. In the 1920s the Gestalt psychologist Kurt Lewin noticed a waiter who remembered unpaid orders in detail but forgot them the moment the bill was settled. His student Bluma Zeigarnik ran experiments in 1927 and reported that people remember interrupted tasks better than finished ones. The finding entered popular use to explain cliffhangers, onboarding progress bars, and study breaks.

What happened. Attempts to reproduce the original result have struggled for decades. A 2025 systematic review and meta-analysis of the accumulated research found no memory advantage for unfinished tasks, though it confirmed a separate, real tendency to want to go back and finish them. That second tendency, called the Ovsiankina effect after Zeigarnik's colleague, is the part that survived. The memory claim did not.

Where it stands. This is a meta-analysis, the strongest evidence tier short of a fresh mega-study, and it directly overturns a finding still taught as settled. The urge-to-resume effect is now the better-supported half of the pair. Products and writers built on "people remember what's unfinished" are leaning on the half that broke.

Bluma Zeigarnik · Wikipedia · 2025
DEVELOPMENTAL PSYCHOLOGY · a psychologist's synthesis of twin and adoption studies · 1998

Peers, Not Parents, Shape A Child's Character

Setup. Psychology has long assumed two forces shape adult personality: genes and the home environment parents provide. Judith Rich Harris, a textbook writer with no university post, questioned the second half. She read the twin and adoption literature that psychologists cite for parental influence and asked whether it actually shows what it is used to show.

The argument. Identical twins raised apart differ from each other about as much as identical twins raised together, and adopted siblings resemble each other no more than strangers, which undercuts a home-environment effect independent of genes. Harris placed the missing non-genetic influence outside the family, in the peer group: children of immigrants learn their parents' language but speak it in their playmates' accent. She named this group socialization theory and argued birth-order effects mostly vanish under close reanalysis.

Where it stands. This is a synthesis of published studies, not new data, and it split the field. Steven Pinker called it a turning point, the developmental psychologist Jerome Kagan called it selective. Later behavioral genetics work keeps finding a real but small home-environment effect that Harris's model does not fully explain.

Judith Rich Harris · Wikipedia · 1998
COGNITIVE SCIENCE · a research program built on replicated judgment-error experiments · 2009

Quantum Math Predicts Judgments Classical Probability Forbids

Setup. Classical probability assumes beliefs sit in one fixed space, so learning more can never make an either/or judgment more likely. People routinely break that rule. Since the 1990s, researchers led by Diederik Aerts, Jerome Busemeyer and Peter Bruza have modeled the errors with quantum probability math, not because brains are quantum, but because the same equations that describe interference in physics also predict how people actually judge.

The finding. In the Linda problem, told a woman looks like a feminist, subjects rank "feminist and bank teller" as more probable than "bank teller" alone, which classical probability forbids. Quantum probability predicts this directly, because judging "bank teller" after "feminist" is a different, context-dependent measurement, not one fixed prior. The same framework explains why people who do not yet know a coin flip's result decline a second flip they would accept after either winning or losing it.

Where it stands. This is a competing model, not a settled theory of cognition. It fits documented judgment errors and predicts new ones, including a tested violation of Bell's inequality in how people combine concepts. Its founders make no claim that neurons compute quantum mechanically.

Diederik Aerts and Jerome Busemeyer · Wikipedia · 2009
RESEARCH METHODOLOGY · a documented pattern traced through citation-history cases · 2017

A Misquoted Letter Got Cited 608 Times

Setup. A citation is supposed to mean a claim was checked. The woozle effect describes what happens when it is not: a source gets cited for a claim it never supported, other writers copy the citation without reading the original, and repetition alone builds false authority. Beverly Houghton named the pattern in 1979 after watching an unsupported claim harden into consensus.

The finding. The clearest case is a five-sentence 1980 letter to the New England Journal of Medicine, which reported low addiction rates among hospitalized patients given narcotics, based on hospital records, not home use. A 2017 NEJM review found the letter cited 608 times, with a sharp rise after OxyContin launched in 1995, most citations stretching its narrow finding into a general claim about opioid safety. Purdue Pharma used the letter in marketing and later pleaded guilty to misleading regulators about addiction risk.

Where it stands. This is documented citation history, not a controlled study, and the same pattern is attested across unrelated fields: human trafficking statistics, a wartime poster's disputed model. The open question is scale. Nobody has measured what share of citations in a given field are woozles versus sound.

Beverly Houghton · Wikipedia · 1979
GENETICS AND BREEDING · magazine essay built on a genome study of 135,000 horses · Aug 2026

Thoroughbred Race Times Stopped Improving Around 1910

Setup. In the 1660s Charles II built a racing and breeding center at Newmarket, and English breeders began selecting for speed at scale. From 1791 the General Stud Book recorded every thoroughbred's ancestry.

The finding. Reliable timing began in the mid nineteenth century, after which speed improved for about fifty years and stopped. The total gain was 1 to 2 percent, or 2 to 4 seconds over a 1.5-mile race. Prominent race times have not significantly fallen since around 1910. Secretariat's 1973 Belmont record still stands. Celebrity stallions sire perhaps a hundred foals a year, so paternal lines sweep the population. Across 135,000 Australian thoroughbreds, inbreeding tracks slower and less lucrative careers, and ten eighteenth-century ancestors account for over 80 percent of it.

Where it stands. The ceiling is a measured record, not a projection, and the inbreeding analysis covers a full population, not a sample. The mechanism is standard quantitative genetics: hard selection on one trait exhausts usable variation and concentrates harmful recessives. That ceiling is for elite horses. Between 1997 and 2012 the average British thoroughbred still gained 0.011 yards per second per year.

Sophie Fessl · Works in Progress · Aug 2026
TIME-USE ECONOMICS · thesis and data from a 2008 six-country study

Spare Time Cannot Tell Necessity From Choice

Setup. Time-use research compares people by spare time, the hours left after paid work, housework and personal care. Robert Goodin, James Mahmud Rice, Antti Parpo and Lina Eriksson argued in 2008 that this measure hides what it claims to show. A person can lack spare time from necessity or from choice.

The argument. They built a second measure. Discretionary time counts the hours left after the time a person needs in those three activities, where need is set against a relative poverty line rather than against actual behavior. The ranking then changes. Dual-earner couples without children and lone mothers look alike on spare time and differ dramatically on discretionary time. Averages ran from 76 hours per week in France to 85 in Sweden. Taxes, transfers and childcare subsidies raised discretionary time for parents in Sweden and Finland and lowered it in the United States and Australia.

Where it stands. This is a measured comparison of six countries, not a projection, and it won the 2009 Stein Rokkan Prize. The construction of necessary time carries all the weight. Michael Bittman accepted the central idea and questioned exactly that, above all for unpaid household labor and personal care.

Robert E. Goodin, James Mahmud Rice, Antti Parpo and Lina Eriksson · Cambridge University Press · 2008
URBAN POLICY · an architecture researcher's argument built on national demolition statistics · Jul 2026

Zoning Turned Strict Before Modernism Arrived

Setup. Western countries once let landowners build almost anything. Now strict rules freeze most neighborhoods against change. A popular theory blames modernist architecture: new buildings turned ugly, so residents voted to exclude them. Samuel Hughes tested that theory against dates.

The argument. The order is wrong. Restrictive zoning arrived in Germany and Austria-Hungary in the 1890s, Britain in 1909, and France in 1919. Modernism emerged only in the 1920s and won globally in the 1950s. The great downzoning happened while nearly every new building still carried cornices and pedimented windows. Hughes moves ugliness to a second group, people who live nowhere near a site and object anyway. He calls them NITBYs, and credits them with the victory of heritage conservation after 1960.

Where it stands. The timing evidence is public and hard to dispute. The demolition numbers come from government statistics: Berlin demolished nine pre-1919 buildings out of 50,337 in 2022. The NITBY mechanism is an inference from that timing gap, not a measured result, and Hughes runs no survey of who objects. He names the awkward case himself. I'On in Charleston is beautiful and still drew thousands of petitioners.

Samuel Hughes · Works in Progress · Jul 2026
SOCIOLOGY OF RISK · a researcher's argument built on national survey series · Jul 2026

The Internet Arrived Too Late To Kill Deviance

Setup. Adam Mastroianni, a psychologist, argued in October 2025 that American risk-taking and rule-breaking fell sharply after the 1990s. In 1995 half of high school students drank, 35 percent smoked, 40 percent had tried marijuana, and about 6 percent of girls aged 15 to 19 were pregnant. All of those fell over the next thirty years.

The argument. The common explanation blames the internet, through surveillance or algorithmic flattening. Mastroianni rules it out on timing. A majority of Americans lacked broadband until 2007, and most people did not carry a smartphone until about 2012, long after the trends began. His own explanation is prosperity. As life gets safer and longer, the price of taking a risk rises. Marijuana is legal in 24 states and teenagers now rate it as less dangerous, yet they smoke less than 1990s teenagers did.

Where it stands. The trends come from national survey series, not one study, and the timing argument against the internet is hard to dispute. The prosperity mechanism is an interpretation, not a test, and he offers no direct measure of it. He grants the awkward part: people reliably believe culture peaked when they were young.

Adam Mastroianni · Experimental History · Jul 2026
RADIATION EPIDEMIOLOGY · magazine reanalysis of published cohort studies · Jun 2026

Seventy-Seven Cancer Subtypes Manufactured A Radiation Risk

Setup. Between 1982 and 1984, Taiwanese builders unknowingly used recycled steel contaminated with cobalt-60 in more than 180 buildings. Over two decades, more than 10,000 residents absorbed an average total dose of 400 millisieverts, roughly seven times normal background. The buildings became an accidental test of whether slow, low-dose radiation causes cancer.

The finding. Two studies, Hwang and colleagues in 2006 and Hsieh and colleagues in 2017, reported higher rates of thyroid cancer, leukemia and breast cancer. They found them by splitting cases into 77 subtypes and testing each one. Chance alone produces about four significant results across 77 tests, and every significant site-specific result rested on seven or fewer observed cases. Total cancer in the exposed group ran about 35 percent below the national rate.

Where it stands. This is a reanalysis by Works in Progress, not new fieldwork, and it does not show that radiation is safe. The authors accept INWORKS, a study of 300,000 nuclear workers, as the best evidence for a real effect. Even there the effect is small. 100 millisieverts raises cancer mortality by five percent, and the median worker received four.

Ben Southwood and Alex Chalmers · Works in Progress · Jun 2026
DECISION THEORY · a theorem published by Matthew Rabin in 2000

Turning Down A Small Bet Breaks Utility Theory

Setup. Economics explains caution about risk through the diminishing marginal utility of wealth. Each extra dollar is worth slightly less than the last, so a fair coin flip loses value. One curve is supposed to explain a $10 bet and a $10 million bet at once. Matthew Rabin asked whether it can.

The finding. In 2000 Rabin proved that it cannot. Take a person who turns down a coin flip that wins $125 and loses $100, at every wealth level up to $300,000. The curvature needed to reject that small bet compounds as wealth rises. Receiving $1,000 must cut marginal utility to about 37 percent of its value, and receiving $10,000 must cut it to about 0.005 percent. At $290,000 of wealth the same person must also reject a coin flip that wins $160 billion and loses $1,000.

Where it stands. This is a proof, not a survey, and the arithmetic is not disputed. Its force rests on a premise it assumes rather than measures, that people do turn down modest favorable bets. The escape route is narrow. Zvi Safra and Uzi Segal extended the result to any model with a differentiable utility over lotteries, including rank-dependent expected utility and disappointment aversion.

Matthew Rabin · Wikipedia · 2000
EDUCATIONAL SIGNALING · an essay built on placement data and national grade statistics · Jul 2026

Cheap Signals Collapse Into Expensive Ones

Setup. A grade point average is a cheap signal. It costs a student little to produce and an employer almost nothing to read. Matt Duffy argues that American education destroyed this signal, and that the destruction did not remove screening. It moved screening somewhere more expensive.

The argument. Duffy builds on Campbell's Law from 1976, which holds that a measure used for decisions gets gamed until it stops tracking what it measured. Average high school GPA rose from 3.17 in 2010 to 3.36 in 2021 while average ACT scores fell to their lowest of the decade. A 2025 report at UC San Diego found that students placed in the lowest remedial math track carried an average high school math GPA of A minus. Employers and admissions offices then moved to internships, code portfolios, essays and extracurriculars, which cost far more and track family income.

Where it stands. The numbers are public and specific, and Campbell's Law is well established. Two limits are real. UC San Diego is one campus, and Duffy calls it representative by assertion. His remedy, AI-run assessment, rests on Alpha School, whose students score well on the national tests the school teaches toward and have no college or job record yet.

Matt Duffy · Palladium · Jul 2026
THERMAL BIOLOGY · an 1877 ecogeographic rule, with a developmental mechanism tested in mice

Cold Shortens Limbs Within One Lifetime

Setup. Joel Asaph Allen proposed in 1877 that animals adapted to cold climates carry shorter, thicker limbs and appendages than animals adapted to warm ones. Polar bears have stocky legs and short ears. The standard reading is evolutionary. A low ratio of surface area to volume conserves heat, so selection favors it over many generations.

The finding. Part of the pattern needs no generations. Experimenters raised mice at 7, 21 and 27 degrees Celsius. The cold-raised mice grew significantly shorter tails and ears at the same body weight, and showed less blood flow in their extremities. Bone samples grown warm produced significantly more cartilage. Cartilage growth responds to temperature directly, so the animal's own environment shapes its proportions during development. Human populations fit too. In Peru, people living at altitude have shorter limbs than people from the same population on the coast.

Where it stands. The mouse experiment is controlled and supplies a clear proximate mechanism, which sits underneath selection rather than replacing it. The rule itself is weaker than its fame. Nudds and Oswald argued in 2007 that empirical support is poor, because tests across many species are confounded by Bergmann's rule on body mass.

Joel Asaph Allen · Wikipedia · 1877
NEUROSCIENCE · a review paper by two researchers, Nature Reviews Neuroscience · Aug 2026

The Body's Energy Budget Picks The Category

Setup. Categorization is how a brain compresses a flood of sensory signal into objects, people and concepts. The textbook account runs one way. The senses deliver features, the brain matches them against stored templates, and a label comes out at the end. Lisa Feldman Barrett of Northeastern University and Earl Miller of MIT published an alternative in Nature Reviews Neuroscience.

The argument. They invert the flow. The brain projects categories outward, driven by the body's energy needs, before the senses finish reporting. Their anatomical evidence is a count. Inside the visual cortex, 90 percent of synaptic connections carry feedback rather than feedforward signals. They place the source of prediction in the limbic core, beside the hypothalamus, which tracks temperature, heart rate and hunger. Sensory signals compress inward, bodily signals compress upward, and the two meet there. The same scratch on a leg is nothing at home and a snake in tall grass.

Where it stands. The connectivity figure and the compression gradient are established anatomy. The framework built on them is a proposal, and it is the authors' own synthesis of their two research programs. Luiz Pessoa of the University of Maryland calls the energy-constraint idea important to pursue, which is endorsement of a direction, not of a result.

Conor Feehly · Quanta Magazine · Aug 2026
COLONIAL HISTORY · a magazine essay built on the nineteenth-century diplomatic record · Aug 2026

Siam Performed Each Empire's Own Legitimacy Test

Setup. European powers colonized nearly every territory on earth over five centuries. Siam, now Thailand, is one of the few exceptions. Conquest needed a justification as well as an army. Europeans held that technologically advanced Christian nations had a duty to bring government, writing and refinement to backward peoples.

The argument. Derek Hopper argues that Siam attacked the justification rather than the army. In 1833 the monk Mongkut found the Ramkhamhaeng Stele, 124 lines of archaic Thai purportedly from 1292, and the court promoted it as proof of an old literary civilization. As King Rama IV he received the British envoy John Bowring in 1855 over cigars and wine and spoke English without interpreters. Each power received the credential it valued: science for France, law and free trade for Britain. Burma reformed as well, but only after Britain invaded in 1824, and Britain annexed it anyway in 1885.

Where it stands. The contrast with Burma is the strongest part of the case, because it isolates timing rather than reform itself. The central claim stays unmeasurable, since nobody can run the counterfactual where Siam stayed aloof. The sovereignty Siam kept was also partly formal. The 1855 treaty capped Siamese duties at 3 percent and put British subjects beyond Siamese courts.

Derek Hopper · Works in Progress · Aug 2026
Aug 31, 2026
Aug 30, 2026
Aug 31, 2026
Aug 30, 2026