Data quality in online human-subjects research: Comparisons between MTurk, Prolific, CloudResearch, Qualtrics, and SONA
2023/03/14 by Benjamin D Douglas, Patrick J. Ewell, Markus Bräuer · 1 voice · 70 citations
Computer Science · Social Sciences · #Privacy-Preserving Technologies in Data #Privacy, Security, and Data Protection #Mobile Crowdsensing and Crowdsourcing
paper · pdf · doi:10.1371/journal.pone.0279720
Abstract
With the proliferation of online data collection in human-subjects research, concerns have been raised over the presence of inattentive survey participants and non-human respondents (bots). We compared the quality of the data collected through five commonly used platforms. Data quality was indicated by the percentage of participants who meaningfully respond to the researcher's question (high quality) versus those who only contribute noise (low quality). We found that compared to MTurk, Qualtrics, or an undergraduate student sample (i.e., SONA), participants on Prolific and CloudResearch were more likely to pass various attention checks, provide meaningful answers, follow instructions, remember previously presented information, have a unique IP address and geolocation, and work slowly enough to be able to read all the items. We divided the samples into high- and low-quality respondents and computed the cost we paid per high-quality respondent. Prolific (1.90) and CloudResearch (2.00) were cheaper than MTurk (4.36) and Qualtrics (8.17). SONA cost 0.00, yet took the longest to collect the data.
Citations
Cited by
- Decision making under extinction risk
- AI and metaverse as transformative tools for knowledge management: behavioral reasoning and its implications for organizational performance
- Why do people buy NFTs: a configurational approach to understanding the adoption of non-fungible tokens
- Is Crowdsourcing a Puppet Show? Detecting a New Type of Fraud in Online Platforms
- Prevalence of Security and Privacy Risk-Inducing Usage of AI-based Conversational Agents
- Beyond the Uncanny Valley: A Mixed-Method Investigation of Anthropomorphism in Protective Responses to Robot Abuse
- I don't Want You to Die: A Shared Responsibility Framework for Safeguarding Child-Robot Companionship
- Effect of a Short, Animated Storytelling Video on Transphobia Among US Parents: Randomized Controlled Trial
- Delegating Destruction: Coercive Threats and Automated Nuclear Systems
- “Men should watch football games instead of soap operas”: Masculine role norm endorsement as a population-level construct associated with harmful gambling
- Public Opinion and Prosecutor-Initiated Resentencing in California: Can a Racial Equity Norms Message Increase Support?
- Warrior or guardian? Public perceptions of law enforcement during traffic encounters and domestic disputes
- The Burden for High-Quality Online Data Collection Lies With Researchers, Not Recruitment Platforms
- Semantic memory network resilience in aging: The role of abstract and semantically diverse words
- The Myth of Misleading Labels: Examining Consumer Understanding of Plant-based Meat and Milk Analogues
- From hashtags to ballots: Conceptualizing political influencers and evaluating their impact on election outcomes
- Talk about shared money: Account pooling is associated with financial communication
- The role of defendant race, expert testimony and interrogation coerciveness on Canadian mock jurors' perceptions of recanted confessions
- Hazardous Drinking Amplifies the Association Between Emotion-Based Impulsivity and Negative Thoughts Related to Suicide Ideation Among Adults
- Potential for Harm and Privacy Are Top AI Ethics Concerns: Evidence From a Representative UK Survey
- Men, masculinity, and veganism: Experimental insights and branding solutions to foster sustainable food choices
- Higher cognitive ability linked to weaker moral foundations in UK adults
- Online In-Context Distillation for Low-Resource Vision Language Models
- Developing and testing an AI-based risk information-seeking model (ARISM): A structural equation modeling approach
- The combined effect of patient classification systems and availability of resources can bias the judgments of treatment effectiveness
- Defensive reactions to a meat reduction intervention
- Externalities and norms in the context of a novel virus
- Framing Allais: Is the paradox robust to the pictorial framing of probabilities?
- In generative artificial intelligence we trust: unpacking determinants and outcomes for cognitive trust
- Extending the Gamer’s Dilemma: empirically investigating the paradox of fictionally going too far across media
- Occupational prestige and occupational social value in the United Kingdom: New indices for the modern British economy
- Examining the replicability of online experiments selected by a decision market
- Responses to belief-conflicting information: Justification of support for Donald Trump
- Assessing Police Recruit Disqualifiers: Public Attitudes Toward Drug Use, Mental Illness, Education, and Criminal Offending
- The Proust effect and hoarding symptoms: relationships among memory vividness, object type, and urge to save
- Attention checks and how to use them: Review and practical recommendations
- Future Shock or Future Shrug? Public Responses to Varied Artificial Intelligence Development Timelines
- How accurate are witnesses of first suspected seizures in recalling semiology at clinically relevant timepoints? A UK experimental study with a pilot intervention
- Sounds of Silence: Using the Stereotype Content Model to Understand Perceptions of Introverts and Extraverts at Work
- An Expert Guide to Planning Experimental Tasks For Evidence-Accumulation Modeling
- Cross-country differences in willingness to use conditionally automated driving systems: Impact of technology affinity, driving skills, and perceived traffic climate
- Is Deliberate Control of Behavior Rare? A Test of the Automaticity Dominance Perspective
- Social Imagery and Subjective Ideological Proximity to the Supreme Court: Evidence From Evangelical Christians
- Open minds, tied hands: Awareness, behavior, and reasoning on open science and irresponsible research behavior
- Just a diversity hire?: the effects of competency microaggressions in collaborative work
- A crowdsourced megastudy of 12 digital single-session interventions for depression in US adults
- Chatbots Are Undermining Crowdsourced Research in the Behavioral Sciences: Detecting Artificial Intelligence–Assisted Cheating With a Keystroke-Based Tool
- More Rights, More Danger for Police: An Experimental Look at How Messaging Shapes Public Views about Constitutional Carry
- Representativeness and Response Validity Across Nine Opt-In Online Samples
- ConViS-Bench: Estimating Video Similarity Through Semantic Concepts
- Communicating diagnostic uncertainty reduces expectations of receiving antibiotics: Two online experiments with hypothetical patients
- Individual utilities of life satisfaction reveal inequality aversion unrelated to political alignment
- Articulating scholarship in human resource management: Guidance for researchers
- Moral Foundations, Welcomeness and Belonging, and Visitation to Manassas National Battlefield Park: The Necessity of Nuance for Understanding Interpretive Preferences
- Bridging Cultures in the Era of Big Data: A Cross-Language Equivalence Framework in Machine-Learning Research With Social Media Texts
- Conversational Inoculation to Enhance Resistance to Misinformation
- Capturing and Sharing Know-How through Visual Process Representations: A Human-Centred Approach to Teacher Workflows
- A diffusion of innovations measurement scale for reinvention, relative advantage, compatibility, complexity, trialability and observability
- When trust turns digital: why relational cues matter in online crime-reporting portals
- The Path to Driving Aggression and Crash Risk: The Role of Metacognition and Anger Rumination in Anger Expression Among Chinese Drivers
- SCOOTER: A Human Evaluation Framework for Unrestricted Adversarial Examples
- Interaction Techniques that Encourage Longer Prompts Can Improve Psychological Ownership when Writing with AI
- Requirements Elicitation Follow-Up Question Generation
- Who's Sorry Now: User Preferences Among Rote, Empathic, and Explanatory Apologies from LLM Chatbots
- The Impact of Co-Opetition on Cognitive Flexibility
- If You Had to Pitch Your Ideal Software -- Evaluating Large Language Models to Support User Scenario Writing for User Experience Experts and Laypersons
- Connecting Through Fear of Missing Out (FoMO): Social Media Involvement and Team Identification Among Sports Fans
- Reputation-based reciprocity in human–bot and human–human networks
- Fantasy Worlds, Real-Life Impact: The Benefits of RPGs for Transgender Identity Exploration
- Stress and episodic future thinking: Effects on temporal window and alcohol demand
- Benchmarking Music Generation Models and Metrics via Human Preference Studies
Discussions
Related