Here is the link:
https://www.ssoar.info/ssoar/bitstream/handle/document/42104...
Years after pondering this article, I felt I had figured it out: psychology is secretly the study of particular subset of 18 to 21 year old American women. The ones who study psychology because it's a female dominated study and all those psych students need their credit and 'participating' in research is part of it. Most of them are American because psychology is a bigger thing in the US than in Europe, or at least it seemed to be regarding well-known theories, so I presume most research happens there.
There is another big group. A lot of dead mice (neuroscience).
"After the tragedies caused by the use of thalidomide in pregnant women, the FDA issued “General Considerations for the Clinical Evaluation of Drugs” in 1977. This guidance document stated that women of child-bearing potential should be excluded from Phase 1 and early Phase 2 research, except if these studies were being conducted to test a drug for a life-threatening illness. If a drug appeared to have a favorable risk-benefit assessment, women could then be included in later Phase 2 and Phase 3 trials if animal teratogenicity and fertility studies were finished...In 1993, FDA reversed the 1977 guidance with another guidance document entitled Guidelines for the Study and Evaluation of Gender Differences in the Clinical Evaluation of Drugs."
Also:
"In a study that evaluated the inclusion and analysis of sex in the results of federally-funded randomized clinical trials in nine major medical journals in 2009, researchers found most studies that were not sex-specific had an average enrollment of 37% women."
From "Women’s involvement in clinical trials: historical perspective and future implications" - https://www.ncbi.nlm.nih.gov/pmc/articles/PMC4800017/
There are many analogs of that in Computer science too. For example in computer vision, CIFAR-10 is a de facto standard for measuring performance. A set of 60k 32x32 images. Good results on that data set doesn't necessarily translate well into real-world performance. But what can you do? Gathering huge data sets and having humans annotate them is incredibly expensive.
Another ml example is music recognition. There are several state-of-the-art methods for detecting notes in music. So one Chinese researcher tried to apply them for Jingju music https://www.youtube.com/watch?v=NJsTl342RhI. The results were... less than stellar.
> subset of 18 to 21 year old American women.
This is perfect!
I had a similar epiphany about linguistics, much of which appears to be the study of example sentences that linguists come up with.
Since you could invariably tell when something was an example sentence, this was obviously a different language from what people actually used. (Of course there are linguists who go out and study real language-use, but some actively dismiss this as mostly irrelevant "performance")
I'm not sure about the amount of research but the amount of psychology students here in the Netherlands is huge. And it seems to function in much the same way. Often first-years get some form of credit for mandatory participation in experiments for higher-year students and sometimes regular faculty members. I'm also aware of some researchers going out of their way to get a different sample from the population but in the end it's a lot easier to get participants when you make them.
I don't know any easy way out here. Just because "folk psychology" can say formal psychology is too simple doesn't mean folk psychology and anecdote are more useful.
Wait, does participating in a study as a subject count for credit, or does working on a study count?
Aren't studies explicit about the demographics of their participants? And don't studies that make generalized claims usually control for things like gender and age?
I'd add in Beagles, Capuchins, Zebrafish, and Mongolian Gerbils there too.
The entire field is murky because it's very hard to scientifically measure human psychology and behavior, and the tools of the trade seem almost laughably simplistic (like the aforementioned 5 point rating questionnaires). So our body of knowledge doesn't even reliably describe WEIRD people. Rather its a crude proxy, that just might contain some elements of truth warranting further investigation.
But better crude tools than no tools, and better locally available subjects than no subjects (because most studies don't have the budget to go to Zambia). Similarly in other fields, research is done on rats pigs and monkeys with conclusions drawn to humans. Obviously not perfect, but again it's at best a starting point for later studies.
I think the real problem is the over-zealous interpretation of study results as "truth".
If you're trying to measure a target 1cm across from 1 mile away with a ruler and a squint anything you say is not only likely wrong, but woefully deceptive.
My view is that this shouldn't even be attempted because it just generates superstitious theories. Psychology is the skinner box pidegon.
There's a fallacy here that's hard to pin down clearly but roughly: to measure inaccurately isnt to measure approximately. Its to measure totally in error.
The errors in social psychology are not just "second decimal place", they're angels pushing stars.
Surely, the solution isn't to throw up our hands in the West and be like "we can't find any help for the mentally ill because some subject might not apply to our discoveries from around here!"--but of course, if they are extrapolating those results to other cultures, that's worrisome and inaccurate in many studies.
Human behavior has some universals, but a lot we figure are universal aren't.
There will always be a conflict between what we can know with confidence and the decisions that we need to make to handle situations that arise. But it has been shown that having someone in the position of authority as a 'scientist' giving their approval to something based upon insufficient evidence or reason tends to often lead to the most severe and large scale suffering. So I'd have to recommend against giving any credence to any crude tool.
The point of this is that the average study in psychology is much more likely to be wrong than right. And so by indulging psychology you are not giving yourself a crude tool for understanding, you are actively misinforming yourself! Imagine I wrote a newspaper where 64% of the articles were fake or misleading. If you'd like to be as well informed as possible, you'd be better off never reading that paper, even if there are some true things in it.
Science is not a 0 or positive game. Bad science can and does send societies and progress backwards.
https://en.wikipedia.org/wiki/Replication_crisis#Psychology_...
> A report by the Open Science Collaboration in August 2015 that was coordinated by Brian Nosek estimated the reproducibility of 100 studies in psychological science from three high-ranking psychology journals.[38] Overall, 36% of the replications yielded significant findings (p value below 0.05) compared to 97% of the original studies that had significant effects. The mean effect size in the replications was approximately half the magnitude of the effects reported in the original studies.
Journals are filters to cherry-pick 'study space'. By the way they're constituted, they publish new studies that have overstated significance.
Oh god, having filled out a bunch of these for diagnosis and such I hate these with a passion. I always wondered how well these actually work.
I've seen grammatical nonsense like, "Do you often do X? -- always, often, sometimes, rarely, never". What, I often rarely do X? And what does often mean, anyway? Like once a week? Every day?
Then, there are the abstract or vague questions that you then have to interpret what concrete situation it could apply to. Hard to think up an example off the top of my head, but how people reply to these surely depends on what exactly they think it might mean.
Then you start losing patience after about 3 minutes of this shit, not to mention 15 or 30 minutes, and just go through them barely reading the questions, but for the first couple of questions you were pondering whether you "agree" or "somewhat agree" for ages.
I grant that with a questions like "Do you often do X?", examples are necessary to specify what "often" means.
From the article:
> Some people may refuse to answer. Others prefer to answer simply yes or no. Sometimes they respond with no difficulty.
That just sounds like some people boycott the Likert questions, but we don't know why.
Not saying it's right, but not saying it's a singular question "do you believe X agree? slightly agree? etc." it's a bit more deep and nuanced than that. And the statistics tend to back it up.
Especially since a study with 300 participants is worth immensely more than 10 studies with 30 participants.
The worst is it’s contextual loyalties against critiquing the relationsips people have with one another. Power struggles, to be particular.
I don't have any issue with the general message. This surely is a legitimate problem. But the intro reads a bit like "psychology is such a great endeavour, if there wasn't this little issue."
Everyone following science news should be aware by now that psychology suffers from a whole range of systemic methodological problems, notably publication bias, widespread p-hacking and failed replications.
Think children growing in warzone or poor and violent environment. Their behaviour as adults is often sexually more promiscuous, aggressive and their impulse control seems to be less than 'the baseline'. They show trust issues.
How much of that is just damage and disorder as psychology seems to assume, and how much is adaption to survive and procreate in an environment where lifespans are short and life is uncertain. Maybe childhood stress and stress hormones trigger survival strategies that work well in hard environment. They are maladaptive only in the culture and safety of the developed world.
It would mean we are effectively blind(er) to unfamiliar shapes, even though they are extremely simple, like a triangle or square are.
So just by simple fractions we know something is off. A more careful study by subject and region could be required.
They should be considered temporary interpretations of statistical data or in short meta statistics cause that's really what they are.
Much damage is being done by treating these fields as science and the article is only mentioning a few of those problems.
This is what a lot of experimental sciences are, even physical ones, when the systems being studied are complex.
There are very few areas of scientific study anymore that offer convenient, deterministic results. That fruit was picked a long time ago.
Even at the cutting end of physics, researchers have to infer from statistical results.
The difference is only that some of these fields have more reproducible results than others, often because they are studying less complex phenomena, whose causal factors and mechanisms can be more directly observed.
Psychology is at one end of that spectrum, because it is studying the output of the mind, a biological information system whose mechanisms are among the most complex and obscure that people have ever studied.