The Measurement of Observer Agreement for Categorical Data
1977/03/01 by J. Richard Landis, Gary G. Koch · 1 voice · 778 citations
Decision Sciences · Mathematics · #Meta-analysis and systematic reviews #Reliability and Agreement in Measurement #Statistical Methods in Epidemiology
paper · doi:10.2307/2529310
openalex publication_date 1977/03/01 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/31
Abstract
This paper presents a general statistical methodology for the analysis of multivariate categorical data arising from observer reliability studies. The procedure essentially involves the construction of functions of the observed proportions which are directed at the extent to which the observers agree among themselves and the construction of test statistics for hypotheses involving these functions. Tests for interobserver bias are presented in terms of first-order marginal homogeneity and measures of interobserver agreement are developed as generalized kappa-type statistics. These procedures are illustrated with a clinical diagnosis example from the epidemiological literature.
Citations
Cited by
- Exploring the Design Space of LLM-Based Programming Support in CS Education: A Scoping Review through the Lens of Assistance Governance
- Inter‐ and intrarater reliability of Hurley staging for hidradenitis suppurativa
- The Boston bowel preparation scale: a valid and reliable instrument for colonoscopy-oriented research
- Visual Function Classification System for children with cerebral palsy: development and validation
- The Relation Between Perceived Mental Effort, Monitoring Judgments, and Learning Outcomes: A Meta-Analysis
- Behavioural and psychological symptoms in dementia and the challenges for family carers: Systematic review
- Test-retest Reliability and Construct Validity of the Satisfaction with Treatment Result Questionnaire in Patients with Hand and Wrist Conditions: A Prospective Study
- Children's Early Awareness of Comprehension as Evident in Their Spontaneous Corrections of Speech Errors
- Relational energy at work: Implications for job engagement and job performance.
- Overinterpretation of Research Findings: Evidence of “Spin” in Systematic Reviews of Diagnostic Accuracy Studies
- Situational factors and police use of force across micro‐time intervals: A video systematic social observation and panel regression analysis
- Trauma and Hallucinatory Experience in Psychosis
- Is It Good to Cooperate? Testing the Theory of Morality-as-Cooperation in 60 Societies
- Exploration as a mediator of the relation between the attainment of motor milestones and the development of spatial cognition and spatial language.
- Assessing Validity of ICD‐9‐CM and ICD‐10 Administrative Data in Recording Clinical Conditions in a Unique Dually Coded Database
- Suggested Novel Risk Factors for Swimming-Induced Pulmonary Edema
- Anxiety and depressive disorders in people with epilepsy: A meta‐analysis
- A Systematic Review of Agreement Between Perceived and Objective Neighborhood Environment Measures and Associations With Physical Activity Outcomes
- Factors That Can Undermine the Psychological Benefits of Coastal Environments
- The Accuracy of Citizen Science Data: A Quantitative Review
- A practical approach to measuring the visual field component of fitness to drive
- Malignancy grading of the deep invasive margins of oral squamous cell carcinomas has high prognostic value
- Back-Channel Representation: A Study of the Strategic Communication of Senators with the US Department of Labor
- Predicting Under- and Overperforming SKUs within the Distribution–Market Share Relationship
- The Prevalence of Dental Caries and Fluorosis in Japanese Communities with Up to 1.4 ppm of Naturally Occurring Fluoride
- The staging of gastritis with the OLGA system by using intestinal metaplasia as an accurate alternative for atrophic gastritis
- Diagnostic performance of Japan NBI Expert Team classification for differentiation among noninvasive, superficially invasive, and deeply invasive colorectal neoplasia
- How Commercial Video Games Engage with Biodiversity and Conservation: A Systematic Map of Literature
- Tell Me a Good Story and I May Lend you Money: The Role of Narratives in Peer-to-Peer Lending Decisions
- 15% of Talar Osteochondral Lesions Are Present Bilaterally While Only 1 in 3 Bilateral Lesions Are Bilaterally Symptomatic
- Newspaper Coverage of Intimate Partner Violence: Skewing Representations of Risk
- A meta-analysis of confidence and judgment accuracy in clinical decision making.
- Alcohol interventions for mandated college students: A meta-analytic review.
- Uncovering biomarkers for chronic toxoplasmosis detection highlights alternative pathways shaping parasite dormancy
- A Revised Diagnostic Classification of Canine Glioma: Towards Validation of the Canine Glioma Patient as a Naturally Occurring Preclinical Model for Human Glioma
- Multisite Assessment of Aging-Related Tau Astrogliopathy (ARTAG)
- Why Network Segmentation Projects Fail
- Proficiency Reporting Practices in Research on Second Language Acquisition: Have We Made any Progress?
- Tracking the Real‐Time Evolution of a Writing Event: Second Language Writers at Different Proficiency Levels
- Early detection of nasopharyngeal carcinoma: performance of a short contrast-free screening magnetic resonance imaging
- Tumor growth rate of invasive breast cancers during wait times for surgery assessed by ultrasonography
- Risk Factors for Failure of Nonoperative Treatment of Posterior Shoulder Labral Tears on Magnetic Resonance Imaging
- MedKGent: A Large Language Model Agent Framework for Constructing Temporally Evolving Medical Knowledge Graph
- Concordance Between Self-Report and Electronic Medical Record Diagnoses of Insomnia and Sleep Apnea
- COMPASS 31: A Refined and Abbreviated Composite Autonomic Symptom Score
- Precision of Health-Related Quality-of-Life Data Compared With Other Clinical Measures
- Dialogic Instruction in a Chinese EFL Classroom: A Practitioner Perspective
- Reliability of the PEDro Scale for Rating Quality of Randomized Controlled Trials
- Reliability, Internal Consistency, and Validity of Data Obtained With the Functional Gait Assessment
- Board Characteristics and Corporate Social Responsibility: A Meta-Analytic Investigation
- Accuracy of real-time polymerase chain reaction to detect <i>Schistosoma mansoni –</i> infected individuals from an endemic area with low parasite loads
- Green Governance: Boards of Directors’ Composition and Environmental Corporate Social Responsibility
- Reducing time in lethality assay (LD50) for Bothrops jararaca and Crotalus durissus terrificus venoms and lethality neutralizing assay (ED50) for their respective antivenoms: A 3Rs-based retrospective data validation
- The Kappa Statistic in Reliability Studies: Use, Interpretation, and Sample Size Requirements
- Development and Validation of a Scale for Rating Motor Compensations Used for Reaching in Patients With Hemiparesis: The Reaching Performance Scale
- IMPREGNATION OF SUPERCONDUCTING COILS TO REDUCE TRAINING.
- An eating pattern characterised by skipped or delayed breakfast is associated with mood disorders among an Australian adult cohort
- Wheat Crown Rot Pathogens <i>Fusarium graminearum</i> and <i>F. pseudograminearum</i> Lack Specialization
- When philosophy (of science) meets formal methods: a citation analysis of early approaches between research fields
- Unfit for stranding assessment: a panel-scale multimodal-LLM audit of building-decarbonisation disclosure (BeDA)
- Reasoning Denoiser: Denoising Reasoning Traces for Hallucination Detection in Large Reasoning Models
- A Latent Class Extension of Signal Detection Theory, with Applications
- The Populist Style in American Politics: Presidential Campaign Discourse, 1952–1996
- Race and Ethnic Representations of Lawbreakers and Victims in Crime News: A National Study of Television Coverage
- Discordance between pain specialists and patients on the perception of dependence on pain medication: A multi-centre cross-sectional study
- Evaluating shared decision‐making between companion animal veterinarians and their clients using the Observer OPTION <sup>5</sup> instrument
- Reporting community involvement in autism research: Findings from the journal Autism
- Carvone and its pharmacological activities: A systematic review
- A global systematic review of the cultural ecosystem services provided by wetlands
- Scripting police escalation of use of force through conjunctive analysis of body-worn camera footage: A systematic social observational pilot study
- Dynamic Capability Scoping for Enterprise AI Agents: A Synthetic Dataset and Three-Source Permission Architecture
- Identifying and Supporting Academically Low-Performing Schools in a Developing Country: An Application of a Specialized Multilevel IRT Model to PISA-D Assessment Data
- Relationships of Teachers’ Language and Explicit Vocabulary Instruction to Students’ Vocabulary Growth in Kindergarten
- Mother–Child Relationships of Children with ADHD: The Role of Maternal Depressive Symptoms and Depression-Related Distortions
- A systematic review of information practices research
- Perceived Control as a Potential Protective Factor for Suicidal Thoughts and Behaviors in Cancer Patients and Survivors: A Systematic Review With Meta‐Analysis
- Gender bias in video game dialogue
- How to Tell More is More: Quantity Discrimination in Eastern Box Turtles (Emydidae: Terrapene carolina)
- A Systematic Review and Meta-Analysis of the Risk of Disruptive Behavioral Disorders in the Offspring of Parents with Severe Psychiatric Disorders
- Longitudinal Assessment of Self-Harm Statements of Youth in Foster Care: Rates, Reporters, and Related Factors
- The perils of averaging crime data: Consequences of a common practice
- Systematic Review of Gender-Specific Child and Adolescent Mental Health Care
- Interrater Reliability in Systematic Review Methodology: Exploring Variation in Coder Decision-Making
- What Do Books in the Home Proxy For? A Cautionary Tale
- Investigating the Association Between Obsessive-Compulsive Disorder Symptom Subtypes and Health Anxiety as Impacted by the COVID-19 Pandemic: A Cross-Sectional Study
- Inter-Rater Reliability Methods in Qualitative Case Study Research
- Motivations for violent extremism: Evidence from lone offenders’ manifestos
- Sources of Variance in the Accuracy of Interviewer Observations
- Male suicide risk and recovery factors: A systematic review and qualitative metasynthesis of two decades of research.
- A Unified Moral-Value Dataset for Instruction Tuning
- Endometrial cytology and ultrasonography for the detection of subclinical endometritis in postpartum dairy cows
- Active Listening in Integrative Negotiation
- Utilizing a domain-specific large language model for LI-RADS v2018 categorization of free-text MRI reports: a feasibility study
- SciHazard: A Benchmark for Measuring Scientific Safety Risks with Decomposed Harm Scoring
- Real-World Evaluation of an AI Agent Drafting Translational Impact Summaries
- ChannelGuard: Safe Models Do Not Compose into Safe Multi-Agent Systems
- SpEmoC: A Balanced Speaker-Segment Multimodal Emotion Benchmark
- Exposure is not manifestation: measurement target and output resolution jointly determine which behavioural-faithfulness evaluator wins
- Safety That Does Not Transfer: Cross-Lingual Clinical Correctness Drift in Deployable Medical Language Models
- When LLMs Over-Answer: Measuring and Mitigating Quality Issues in LLM-Based Hardware Description Language Question Answering
- Driver Behavior Under Traffic Complexity: Variance Decomposition and Cross-Driver Generalization for Driver Monitoring
- Orientation Reading by Production Vision-Language Models on Optotype Charts: A Controlled Multi-Model Evaluation Across Reasoning Modes, Prompts, and Access Modalities
- What Does It Take to Research with AI? A Rapid Review of Competencies to Train LLM-Literate Researchers
- Precise but Uncoupled: Reviewer Precision Does Not Guarantee Critique Uptake in Multi-Agent Math Reasoning
- WrAFT: a Modularized Automated Writing Evaluation System for Argumentative Essays
- Grokipedia vs Wikipedia: An LLM-Based Audit of Political Neutrality along Ideologies
- RobustMAD: Evaluating Real-World Robustness of Multimodal Small Language Models for Deployable Anomaly Detection Assistants
- The Riddle Riddle: Testing Flexible Reasoning in Large Language Models and Humans
- Students' Perceptions of Peer Grading
- Large Language Models Hack Rewards, and Society
- Can AR Embedded Visualizations Foster Appropriate Reliance on AI in Spatial Decision-Making? A Comparative Study of AR X-Ray vs. 2D Minimap
- ARGUS: Seeing the Influence of Narrative Features on Persuasion in Argumentative Texts
- Measuring the State of Open Science in Transportation Using Large Language Models
- Understanding academic burnout: A systematic literature review of its risk and protective factors
- Losses that Cook: Topological Optimal Transport for Structured Recipe Generation
- Characterizing Language Use in a Collaborative Situated Game
- Echoing: Identity Failures when LLM Agents Talk to Each Other
- Victimhood claims in German political manifestos
- Partnership through Play: Investigating How Long-Distance Couples Use Digital Games to Facilitate Intimacy
- Information transfer within and between autistic and non-autistic people
- Types of bullying behaviour and their correlates
- Empirical Evaluation of the Impact of Object-Oriented Code Refactoring on Quality Attributes: A Systematic Literature Review
- Exploring color metaphor with Behavioral Profiles: A usage-based analysis on the metaphorical meanings of the Chinese color term bái “white”
- Activity Demands During Multi-Directional Team Sports: A Systematic Review
- Consonant cluster production in children with cochlear implants: A comparison with normally hearing peers
- A multi‐dimensional measure of vocational identity status
- Identity statuses and psychosocial functioning in Turkish youth: A person‐centered approach
- Factors Relating to Sprint Swimming Performance: A Systematic Review
- Students with Disabilities in Life Science Undergraduate Research Experiences: Challenges and Opportunities
- Measuring complex constructs in large-scale text with computational social mixed methods
- Comparison of metrics for identifying Key Biodiversity Areas
- The Effect of Prequestions on Learning: A Multilevel Meta-Analysis
- Can we really reduce ethnic prejudice outside the lab? A meta‐analysis of direct and indirect contact interventions
- Understanding barriers and facilitators to non-pharmaceutical chronic pain research engagement among people living with chronic pain in the UK: a two-phase mixed-methods approach
- Systematic reviews of empirical bioethics
- Validity and reliability of a low-cost digital dynamometer for measuring isometric strength of lower limb
- Straight sprinting is the most frequent action in goal situations in professional football
- Neuro4Neuro: A neural network approach for neural tract segmentation using large-scale population-based diffusion imaging
- Equity How and Equity for Whom? Incorporating Equity Into Local Government Budgeting Processes
- Random forest predictive modeling of mineral prospectivity with small number of prospects and data with missing values in Abra (Philippines)
- Reclaiming narrative sovereignty: education as counter-narrative in Occupied Palestine
- Lilobot: A Cognitive Conversational Agent to Train Counsellors at Children’s Helplines
- Exploring the determinants of the 2023 Quebec, Canada, mega-wildfire burn severity
- Global tracking of marine megafauna space use reveals how to achieve conservation targets
- Dyadic differences in empathy scores are associated with kinematic similarity during conversational question–answer pairs
- Algorithmic bias: review, synthesis, and future research directions
- WNUT-2020 Task 2: Identification of Informative COVID-19 English Tweets
- The Security of Smart Buildings: a Systematic Literature Review
- Harbor Porpoise and Beluga Whale Habitat Use in the Saguenay‐St. Lawrence Marine Park (Canada) Revealed by a Combination of Visual and Acoustic Survey
- A Practical MRI Grading System for Lumbar Foraminal Stenosis
- Plans Work in Mysterious Ways: Evaluating a Plan Mode for Spreadsheet Agents
- TARGET: Automated Scenario Generation from Traffic Rules for Testing Autonomous Vehicles via Validated LLM-Guided Knowledge Extraction
- The relationship between teacher talk and students’ academic achievement: A meta-analysis
- A Structured Cyber Threat Intelligence Dataset Using STIX 2.1 Entities and MITRE ATT&CK Mappings
- DravidianMultiModality: A Dataset for Multi-modal Sentiment Analysis in\n Tamil and Malayalam
- Decoding of speech acoustics from EEG: going beyond the amplitude envelope
- Beyond Exact Match: How Evaluation Methodology Dominates Model Choice in LLM-Based Product Attribute Extraction
- AssumptionMiner: Extracting, Tracing, and Revising Implicit Assumptions in LLM Code Generation
- Evaluating and Mitigating the Misguidance Effect of Buggy Code in LLM-Generated Unit Tests
- Reinforcement Learning Based Argument Component Detection
- TrajAudit: Automated Failure Diagnosis for Agentic Coding Systems
- Predicting residential building age from map data
- Translation Quality Assessment: A Brief Survey on Manual and Automatic Methods
- TalkDown: A Corpus for Condescension Detection in Context
- Tracing the evolution of personality cognition in early human civilisations: A computational analysis of the Gilgamesh epic
- Ara-HOPE: Human-Centric Post-Editing Evaluation for Dialectal Arabic to Modern Standard Arabic Translation
- From Pilots to Practices: A Scoping Review of GenAI-Enabled Personalization in Computer Science Education
- Improving ML Training Data with Gold-Standard Quality Metrics
- Generative AI-powered social robots in education: opportunities and challenges from a Delphi study
- How well do Large Language Models Recognize Instructional Moves? Establishing Baselines for Foundation Models in Educational Discourse
- CienaLLM: Generative Climate-Impact Extraction from News Articles with Autoregressive LLMs
- Beyond the Prompt: An Empirical Study of Cursor Rules
- Measuring fidelity of implementation of named active learning methods in physics
- A Benchmark and Agentic Framework for Omni-Modal Reasoning and Tool Use in Long Videos
- Thematic Dispersion in Arabic Applied Linguistics: A Bibliometric Analysis using Brookes' Measure
- Integrating Large Language Models and Knowledge Graphs to Capture Political Viewpoints in News Media
- Agreement Between Large Language Models and Human Raters in Essay Scoring: A Research Synthesis
- Historical and experimental evidence that inherent properties are overweighted in early scientific explanation
- How participation in Covid‐19 mutual aid groups affects subjective well‐being and how political identity moderates these effects
- A Multifaceted Analysis of Social Biases in Large Language Models
- JointAVBench: A Benchmark for Joint Audio-Visual Reasoning Evaluation
- LexRel: Benchmarking Legal Relation Extraction for Chinese Civil Cases
- Unveiling Malicious Logic: Towards a Statement-Level Taxonomy and Dataset for Securing Python Packages
- Mining Legal Arguments to Study Judicial Formalism
- Analyzing developer discussions on EU and US privacy legislation compliance in GitHub repositories
- Generate-Then-Validate: A Novel Question Generation Approach Using Small Language Models
- Chasing Shadows: Pitfalls in LLM Security Research
- Source Coverage and Citation Bias in LLM-based vs. Traditional Search Engines
- Privacy in the Digital Age: a Review of Information Privacy Research in Information Systems1
- Configuration Defects in Kubernetes
- Can gestures speak louder than words? The effect of gestural discourse markers on discourse expectations
- "Dragon Slayer Becomes the Dragon": How Players Perceive and Respond to Inequality in the Game World of Whiteout Survival
- Overcoming State Inertia: Minimally Invasive Temporal Alignment for Evolving Contexts
- An Empirical Study of Agent Developer Practices in AI Agent Frameworks
- Evaluating the Robustness of Large Language Model Safety Guardrails Against Adversarial Attacks
- Code Comments for Quantum Software Development Kits: An Empirical Study on Qiskit
- Work strain predictors in construction work
- Sparse Overcomplete Word Vector Representations
- CentaurEval: Benchmarking Human-in-the-Loop Value in Agentic Coding
- A Customer Journey in the Land of Oz: Leveraging the Wizard of Oz Technique to Model Emotions in Customer Service Interactions
- Mitigating Semantic Drift: Evaluating LLMs' Efficacy in Psychotherapy through MI Dialogue Summarization
- Eyebrow movements as signals of communicative problems in human face-to-face interaction
- Learning Quick Fixes from Code Repositories
- Reliability and validity of using the global school-based student health survey to assess 24 hour movement behaviours in adolescents from Saudi Arabia
- How strongly are trait positive and negative affectivity associated with anxiety symptoms? A multilevel meta-analysis of cross-sectional studies in anxiety disorders
- From affordance offers to uptake depth: examining hybrid intelligence in human–AI collaborative academic writing
- A Reproducible Framework for Neural Topic Modeling in Focus Group Analysis
- Cross-replication Reliability -- An Empirical Approach to Interpreting Inter-rater Reliability
- StealthCup: Realistic, Multi-Stage, Evasion-Focused CTF for Benchmarking IDS
- Estonian WinoGrande Dataset: Comparative Analysis of LLM Performance on Human and Machine Translation
- MUCH: A Multilingual Claim Hallucination Benchmark
- Do Vision-Language Models Understand Visual Persuasiveness?
- The Linguistic Architecture of Reflective Thought: Evaluation of a Large Language Model as a Tool to Isolate the Formal Structure of Mentalization
- The Oracle and The Prism: A Decoupled and Efficient Framework for Generative Recommendation Explanation
- Exploring Learner Prompting Behavior and Its Effect on ChatGPT-Assisted English Writing Revision
- Acoustic cardiography helps to identify heart failure and its phenotypes
- Training higher education teachers’ critical thinking and attitudes towards teaching it
- From Code Smells to Best Practices: Tackling Resource Leaks in PyTorch, TensorFlow, and Keras
- Recovery Experiences for Work and Health Outcomes: A Meta-Analysis and Recovery-Engagement-Exhaustion Model
- EngChain: A Symbolic Benchmark for Verifiable Multi-Step Reasoning in Engineering
- ConInstruct: Evaluating Large Language Models on Conflict Detection and Resolution in Instructions
- RAG-Driven Data Quality Governance for Enterprise ERP Systems
- Who grades best? Comparing ChatGPT, peer, and instructor evaluations across varying levels of student project quality
- Aspect-Level Obfuscated Sentiment in Thai Financial Disclosures and Its Impact on Abnormal Returns
- Enhancement of conservation knowledge through increased access to botanical information
- Lay Beliefs About Attention to and Awareness of the Present: Implicit Mindfulness Theory (IMT) and Its Workplace Implications
- Augmenting expert detection of early coronary artery occlusion from 12\n lead electrocardiograms using deep learning
- Why Do People (Not) Take Breaks? An Investigation of Individuals’ Reasons for Taking and for Not Taking Breaks at Work
- Testing and Validating Two Morphological Flare Predictors by Logistic Regression Machine Learning
- Fostering netizens to engage in rumour-refuting messages of government social media: a view of persuasion theory
- Evaluating consistency of deterministic streamline tractography in\n non-linearly warped DTI data
- De-identification of Privacy-related Entities in Job Postings
- Spatial prioritisation for conserving ecosystem services: comparing hotspots with heuristic optimisation
- Classifying discourse in a CSCL platform to evaluate correlations with Teacher Participation and Progress
- Playing Video Games During the COVID-19 Pandemic and Effects on Players’ Well-Being
- Analyzing Dataset Annotation Quality Management in the Wild
- AI Annotation Orchestration: Evaluating LLM verifiers to Improve the Quality of LLM Annotations in Learning Analytics
- On User Interfaces for Large-Scale Document-Level Human Evaluation of Machine Translation Outputs
- A Discursive Turn in Journalism Studies? A Systematic Review of Discourse Analysis in Leading Journalism Journals
- Recognition, Workload and Sustainability: Perspectives of <scp>A</scp> ustralian Journal Editors
- Test–Retest Reliability of Choice Experiments in Environmental Valuation
- CC30k: A Citation Contexts Dataset for Reproducibility-Oriented Sentiment Analysis
- Digital déjà vus: an investigation into how AI image generators s(t)imulate historical thinking
- Concrete images, diverse ideas: The role of pictures in learning from multimedia texts of varying abstractness
- SPOT: An Annotated French Corpus and Benchmark for Detecting Critical Interventions in Online Conversations
- FinRpt: Dataset, Evaluation System and LLM-based Multi-agent Framework for Equity Research Report Generation
- Place Matters: Comparing LLM Hallucination Rates for Place-Based Legal Queries
- Outpatient management of cancer patients with febrile neutropenia: a systematic review and meta-analysis
- Indicating Robot Vision Capabilities with Augmented Reality
- Validation of the ?SUV <sub>max</sub> for Interim PET Interpretation in Diffuse Large B-Cell Lymphoma on the Basis of the GAINED Clinical Trial
- Fear, Pain, Denial, and Spiritual Experiences in Dying Processes
- Inter-Rater Reliability of the Phase of Illness Tool in Pediatric Palliative Care
- Consistently Simulating Human Personas with Multi-Turn Reinforcement Learning
- Dataset Creation and Baseline Models for Sexism Detection in Hausa
- Semantically-Aware LLM Agent to Enhance Privacy in Conversational AI Services
- Envisioning Future Interactive Web Development: Editing Webpage with Natural Language
- SynBullying: A Multi LLM Synthetic Conversational Dataset for Cyberbullying Detection
- ConceptScope: Characterizing Dataset Bias via Disentangled Visual Concepts
- User Misconceptions of LLM-Based Conversational Programming Assistants
- Publication activity in medical education research: A descriptive analysis of submissions to the GMS Zeitschrift für Medizinische Ausbildung in 2007-2015
- Luxury value perceptions and consumer outcomes: A meta‐analysis
- Multimodal information density is highest in question beginnings, and early entropy is associated with fewer but longer visual signals
- Towards accurate differential diagnosis with large language models
- Systematic review and meta-analysis of randomised controlled trials on the effects of yoga in people with Parkinson’s disease
- Getting Comfortable on Audits: Understanding Firms’ Usage of Forensic Specialists
- Caregiver encouragement to act on objects is related with crawling infants' receptive language
- What are they thinking? Exploring college students’ mental processing and decision making about COVID-19 (mis)information on social media.
- A meta-analysis of self-determination theory-informed intervention studies in the health domain: effects on motivation, health behavior, physical, and psychological health
- Audit and feedback using the brief Decision Support Analysis Tool (DSAT-10) to evaluate nurse–standardized patient encounters
- Die Kooperation zwischen Regel- und sonderpädagogischen Lehrpersonen in der Praxis verstehen – Ein multimethodischer Blick auf Kooperation im Kontext schulischer Inklusion
- Comparison between methods of vascular calibre characterization and measurement protocols: Influence of vessels number considered
- Digital Prompting in Education: A Design Framework and Bibliometric Analysis
- Rethinking linguistic feedback: A modality-agnostic and holistic approach to multimodal addressee signals in spoken and signed dyadic interaction
- Birthday memories: an experimental think-aloud study on autobiographical remembering in the digital age
- Stone tool shaping without direct cultural transmission
- Peek a boo! Information seeking about food and functionality in capuchin monkeys
- Responding to ostracism in children: The role of parental recollections of the playground
- AI, representation, and critical digital literacy: Navigating visual bias in the digital age
- Interaction coding in leadership research: A critical review and best-practice recommendations to measure behavior
- Diagnostic value of a coronal STIR sequence in conjoined lumbar nerve root detection: an MRI accuracy study
- Developing behavioural indicators for intellectual functioning and adaptive behaviour for <scp>ICD</scp>‐11 disorders of intellectual development
- The impact of acquired communication impairments on sexuality and intimacy: A scoping review
- Motor unit discharge behavior in human muscles throughout force gradation: a systematic review and meta-analysis with meta-regression
- Population Assessment of Tobacco and Health (PATH) reliability and validity study: selected reliability and validity estimates
- Interpretable Image-Level Acne Severity Grading via EfficientNet-B0 Transfer Learning and Grad-CAM
- Handwriting in primary school: comparing standardized tests and evaluating impact of grapho-motor parameters
- StealthBench: Measuring Operational Stealth in Autonomous Offensive-Security Agents
- Classifying Developers into Core and Peripheral: An Empirical Study on\n Count and Network Metrics
- Learning to Paraphrase: An Unsupervised Approach Using Multiple-Sequence Alignment
- Identifying Relevant Messages in a Twitter-based Citizen Channel for\n Natural Disaster Situations
- Role of triggers and dysphoria in mind-wandering about past, present and future: A laboratory study
- The circumventricular organs of the brain: conspicuity on clinical 3T MRI and a review of functional anatomy
- Do risk assessment tools help manage and reduce risk of violence and reoffending? A systematic review.
- Mapping the extent and exploring the drivers of cocoa agroforestry in Nigeria, insights into trends for climate change adaptation
- Determining the Validity, Reliability, and Utility of the Forgotten Joint Score: A Systematic Review
- Does Media Coverage of Partisan Polarization Affect Political Attitudes?
- A Qualitative Exploration of Black and White Women’s Perceptions of Procedural Justice
- Outcomes of intuitive eating interventions: a systematic review and meta-analysis
- Beliefs about the relationship between work and personal life: development and validation of three work-life belief scales
- The reliability of the mankin score for osteoarthritis
- Codificación de egresos hospitalarios con CIE-11 en Chile: principales resultados y perspectivas para su implementación
- Inter-Rater Reliability in Assessing the Methodological Quality of Research Papers in Psychology
- Social regulation of epistemic thinking during collaborative inquiry with multiple documents
- Bone Metastasis Detection at CT with Deep Learning Models Trained Using Multicenter, Multimodal Reference Standards: Development and Evaluation
- Generating units of cultural analysis with large language models: methods and validation for scalable cross-cultural research
- Public Support for Pro-environment and Environment-Critical Movements
- 2023 global heatwave causes mass mortality of a keystone coral on shallow Western Atlantic reefs
- Crowdfunding Success Factors: A Meta-Analytic Investigation
- Combating COVID-19 with charisma: Evidence on governor speeches in the United States
- Individual contributions in student-led collaborative learning: Insights from two analytical approaches to explain the quality of group outcome
- Artificial Intelligence Versus Human Intelligence in Presurgical Implant Planning: A Preclinical Validation
- Physical Activity Preferences Among Older Adults: A Systematic Review
- Development of a research tool leveraging theoretical frameworks to better understand One Health systems thinking among livestock farmers
- Short text classification with machine learning in the social sciences: The case of climate change on Twitter
- Microplastics in food and drink: perceptions of the risks, challenges, and solutions among individuals in the ‘farm-to-fork’ food chain
- Does technique matter? A multilevel meta-analysis on the association between psychotherapeutic techniques and outcome
- PSMA PET Evaluation with a Deep Learning Platform Compared with a Standard Image Viewer and Histopathology
- Exploring students’ (mis)conceptions about ChatGPT-generated text: A qualitative study
- Co-designing a research agenda for UK agroforestry using a multi-actor approach
- 40 Years of Rural Research in Public Administration: Conceptualization, Evidence, and Future Avenues for Research
- Planning promotes studying – The higher the plan quality the better: A micro-randomized trial
- Online Exogamy Reconsidered: Estimating the Internet’s Effects on Racial, Educational, Religious, Political and Age Assortative Mating
- The impact of gamification in educational settings on student learning outcomes: a meta-analysis
- Constructing social media links to formal learning: A knowledge Graph Approach
- Identifying psychological distress data available in nationally representative surveys: A scoping review and case study of Australian surveys
- <i>Histoplasma</i> Antigenuria Prevalence in Patients With Advanced HIV Disease in Côte d’Ivoire: A Prospective Trial Ancillary Study
- Mediated autobiographical remembering in the digital age: insights from an experimental think-aloud study
- Evidence-inclusive communication: steering crisis leadership outcomes in Portugal, Brazil, New Zealand, and the US
- Role of Hip Internal Rotation Range and Foot Progression Angle for Preventing Jones Fracture During Crossover Cutting
- Inter-rater variability for the American Society of Anesthesiologists classification in patients undergoing hepato-pancreato-biliary surgery (MILESTONE-2): international survey among surgeons and anaesthesiologists
- Towards reproducible systematic reviews in Open, Distance, and Digital Education—An umbrella mapping review
- Revealing Rubric Relations: Investigating the Interdependence of a Research-Informed and a Machine Learning-Based Rubric in Assessing Student Reasoning in Chemistry
- MoDL-QSM: Model-based Deep Learning for Quantitative Susceptibility Mapping
- Diagnostic performance of deep learning–based reconstruction algorithm in 3D MR neurography
- Engagement in language learning: A systematic review of 20 years of research methods and definitions
- How do incumbent firms innovate their business models for the circular economy? Identifying micro‐foundations of dynamic capabilities
- Thresholds in forest bird occurrence as a function of the amount of early‐seral broadleaf forest at landscape scales
- Subluxation of the extensor carpi ulnaris on magnetic resonance imaging on neutral wrist position: correlation with tenosynovitis of the extensor carpi ulnaris and translation of the distal radioulnar joint
- Intercoder Reliability in Qualitative Research: Debates and Practical Guidelines
- A genetic, demographic and habitat evaluation of an endangered ephemeral species Xerothamnella herbacea from Australia’s Brigalow belt
- Digital Games-Based Learning Pedagogy Enhances the Quality of Medical Education: A Systematic Review and Meta-Analysis
- Radiographic Analysis of Bone Graft Changes and Anatomical Factors Influencing Maxillary Sinus Floor Augmentation
- Implementation and Randomized Controlled Trial Evaluation of Universal Postnatal Nurse Home Visiting
- How Second Order Are Local Elections? Voting Motives and Party Preferences in Belgian Municipal Elections
- A reliable and valid questionnaire was developed to measure computer vision syndrome at the workplace
- Does Google Scholar contain all highly cited documents (1950-2013)?
- Land-cover change in the Kruger to Canyons Biosphere Reserve (1993– 2006): A first step towards creating a conservation plan for the subregion
- Knee imaging: Rapid three‐dimensional fast spin‐echo using compressed sensing
- Cancer patients’ trust as a motivator to seek a second opinion and its effects on trust
- School-level mental health protective and promotive factors for sexual and gender minority youth: A systematic review
- Patterns of Communication Breakdowns Resulting in Injury to Surgical Patients
- An Empirical Study on Deployment Faults of Deep Learning Based Mobile Applications
- ChatGPT and me: First-time and experienced users’ perceptions of ChatGPT’s communicative ability as a dialogue partner
- Nursing Staff Factors Contributing to Seclusion in Acute Mental Health Care – An Explorative Cohort Study
- Prevalence, phenomenology, aetiology and predictors of challenging behaviour in Smith‐Magenis syndrome
- The efficacy of conjoint behavioral consultation on parents and children in the home setting: Results of a randomized controlled trial
- Min-Mid-Max Scaling, Limits of Agreement, and Agreement Score
- A Science Model Driven Retrieval Prototype
- What Is the Specificity of the Aortic Dissection Detection Risk Score in a Low‐prevalence Population?
- An Industrial Case Study on Measuring the Quality of the Requirements\n Scoping Process
- Contribution of full-thickness supraspinatus tendon tears to acquired subcoracoid impingement
- Análisis del artículo 19 de la Ley del Impuesto sobre Sucesiones y Donaciones y su posible inconstitucionalidad
- Nomogram for sample size calculation on a straightforward basis for the kappa statistic
- Consultorio Médico Local Tipo I. Vícar, Almería
- High‐resolution riparian vegetation mapping to prioritize conservation and restoration in an impaired desert river
- A valid and reliable belief elicitation method for Bayesian priors
- ¿Está presente la inclusión educativa en la formación inicial del personal orientador? Un análisis de su incorporación en los planes de estudio
- Cognitive Behavioral Therapy-Based Short-Term Abstinence Intervention for Problematic Social Media Use: Improved Well-Being and Underlying Mechanisms
- How Online Fraud Victims are Targeted in China: A Crime Script Analysis of Baidu Tieba C2C Fraud
- The positive effects of aquarium visits on children's behaviour: A behavioural observation
- Encouraging innovation in a modern foreign language initial teacher education programme: What do beginning teachers make of task-based language teaching?
- DECEN: A deep learning model enhanced by depressive emotions for depression detection from social media content
- Personal and Social Facets of Job Identity: A Person-Centered Approach
- Broadcasting Messages via Telegram: Pro-Government Social Media Control During the 2020 Protests in Belarus and 2022 Anti-War Protests in Russia
- Biases in the Blind Spot: Detecting What LLMs Fail to Mention
- Performance of the Pediatric Glasgow Coma Scale Score in the Evaluation of Children With Blunt Head Trauma
- The Role of Triage Nurse Ordering on Mitigating Overcrowding in Emergency Departments: A Systematic Review
- What Questions Should Robots Be Able to Answer? A Dataset of User Questions for Explainable Robotics
- The Localized Scleroderma Skin Severity Index and Physician Global Assessment of Disease Activity: A Work in Progress Toward Development of Localized Scleroderma Outcome Measures
- The Fast and Spurious: Developer Productivity with GenAI
- The influence of social relationship on food tolerance in wolves and dogs
- MuSaG: A Multimodal German Sarcasm Dataset with Full-Modal Annotations
- Teacher education for artificial intelligence literacy through a self-determination theory perspective
- A Scoping Review on Conversational Memory and Characteristics of Conversations in Alzheimer's Disease
- “Why Santa but not witches?”: Parents' reasoning behind encouraging and discouraging fantasy beliefs in children
- Can We Stand Together? Measuring Racial Avoidance in Shared Spaces
- The Paris 1976 Wine Tastings Revisited Once More: Comparing Ratings of Consistent and Inconsistent Tasters
- From simple to sophisticated: characterization of new signals in the expanding vocal repertoire of the East Indian Ocean pygmy blue whale
- A framework for collaborative identification of geographical information for map-based dashboards to support pandemic response policy-making
- Evolution of semi-quantitative whole joint assessment of knee OA: MOAKS (MRI Osteoarthritis Knee Score)
- Analysis of hotspot areas in China's satellite internet innovation policies and research on policy evolution
- Participants’ Reported Discomfort with Live Video as a Mode for Answering a Sensitive Survey Question
- Mutual Wanting in Human--AI Interaction: Empirical Evidence from Large-Scale Analysis of GPT Model Transitions
- Understanding collective flight responses to (mis)perceived hostile threats in Britain 2010-2019: a systematic review of ten years of false alarms in crowded spaces
- Moral Minds in Gaming
- Restricting outdoor advertising of unhealthy food: can Australia’s food category-based classification system be applied consistently?
- A systematic review of weight stigma and disordered eating cognitions and behaviors
- Explaining Support for Russian Narratives about the Events in Ukraine among Japanese Scholars and Intellectuals in 2014–19
- Characterizing social and ecological values expressed in <scp>US</scp> Forest Service public comments using a computational approach
- Not in My Schoolyard: Disability Discrimination in Educational Access
- Grey Literature in Software Engineering: A Critical Review
- Loop mediated isothermal amplification assay for detection of <i>Trichomonas vaginalis</i> in vaginal swabs among symptomatic women from North India
- Understanding Self-Admitted Technical Debt in Test Code: An Empirical Study
- Synthetic-to-Real Transfer Learning for Chromatin-Sensitive PWS Microscopy
- “ <i>YOU</i> Were Adopted?!”
- On making causal claims: A review and recommendations
- Relationship of immunodiagnostic assays for tuberculosis and numbers of circulating CD4+ T-cells in HIV infection
- Redefining Retrieval Evaluation in the Era of LLMs
- Explainable machine learning improves interpretability in the predictive modeling of biological stream conditions in the Chesapeake Bay Watershed, USA
- BDiff: Block-aware and Accurate Text-based Code Differencing
- A Grounded Conceptual Model for Ownership Types in Rust
- Do 8- to 18-year-old children/adolescents with chronic physical health conditions have worse health-related quality of life than their healthy peers? a meta-analysis of studies using the KIDSCREEN questionnaires
- Update of the BEVQ‐15, a beverage intake questionnaire for habitual beverage intake for adults: determining comparative validity and reproducibility
- Implicit and explicit processes in phonological concept learning
- BioCAP: Exploiting Synthetic Captions Beyond Labels in Biological Foundation Models
- A performance‐based measure of emotion response control: A preliminary MRI study
- Calibration of two objective measures of physical activity for children
- Reasons People Want Explanations After Unrecoverable Pre-Handover\n Failures
- An Experimental Study of Real-Life LLM-Proposed Performance Improvements
- Shared Medical Decision Making Reconsidered: Challenging an Overly Cognitivist Perspective with a Linguistic Approach
- Amplifying Quiet Voices
- The advantages of the Matthews correlation coefficient (MCC) over F1 score and accuracy in binary classification evaluation
- Elbow radiographic anatomy: measurement techniques and normative data
- PoSh: Using Scene Graphs To Guide LLMs-as-a-Judge For Detailed Image Descriptions
- Zygapophyseal Joint Adhesions After Induced Hypomobility
- Reliability of Manual Muscle Testing in Applied Kinesiology: A Systematic Review
- Climate Sceptics or Climate Nationalists? Understanding and Explaining Populist Radical Right Parties’ Positions towards Climate Change (1990–2022)
- LexChain: Modeling Legal Reasoning Chains for Chinese Tort Case Analysis
- Suicidal Comment Tree Dataset: Enhancing Risk Assessment and Prediction Through Contextual Analysis
- A Review and Taxonomy of Choice Architecture Techniques
- Image quality and lesion detectability of deep learning-accelerated T2-weighted Dixon imaging of the cervical spine
- Reliability and validity of self-reported physical activity in the Nord-Trøndelag Health Study (HUNT 2)
- IVEBench: Modern Benchmark Suite for Instruction-Guided Video Editing Assessment
- Exposure, access, and inequities: Central themes, emerging trends, and key gaps in Canadian environmental justice literature from 2006 to 2017
- Defects4C: Benchmarking Large Language Model Repair Capability with C/C++ Bugs
- Detecting Gender Stereotypes in Scratch Programming Tutorials
- Benchmarking Deep Learning Models for Laryngeal Cancer Staging Using the LaryngealCT Dataset
- Toward coherence in curriculum, instruction, and assessment: A review of learning progression literature
- Detecting Hallucinations in Authentic LLM-Human Interactions
- LLM-Based Multi-Task Bangla Hate Speech Detection: Type, Severity, and Target
- Inter-rater Agreement on Sentence Formality
- Agentic Troubleshooting Guide Automation for Incident Management
- Questionnaires for evaluating virtual reality: A systematic scoping review
- The Royal Zoological Society of Scotland’s Approach to Assessing and Promoting Animal Welfare in Collaboration with Universities
- Computational Drama Analysis
- Intellectual humility and religion/spirituality: a scoping review of research
- Association of Stress and Neighborhood Social Context With Actigraphy-Measured and Self-Reported Adolescent Sleep Outcomes
- Recent Diachronic Change in Affiliative Vocatives in British English
- Masculine Republicans and Feminine Democrats: Gender and Americans’ Explicit and Implicit Images of the Political Parties
- Relative benefits of different active learning methods to conceptual physics learning
- Testing the Efficacy of a Tier 2 Mathematics Intervention
- Ordonnancement d'entités pour la rencontre du web des documents et du web des données
- From Birdwatch to Community Notes, from Twitter to X: four years of community-based content moderation
- Judge's Verdict: A Comprehensive Analysis of LLM Judge Capability Through Human Agreement
- ReTraceQA: Evaluating Reasoning Traces of Small Language Models in Commonsense Question Answering
- Inflated Excellence or True Performance? Rethinking Medical Diagnostic Benchmarks with Dynamic Evaluation
- Identifying Experts in Software Libraries and Frameworks among GitHub Users
- PyMigTool: a tool for end-to-end Python library migration
- Ironically enjoyed music: An investigation of the unique self-regulatory value of irony as part of the enjoyment of music
- Sources of support for learning words in conversation: evidence from mealtimes
- PEER INTERACTION AND CORRECTIVE FEEDBACK FOR ACCURACY AND FLUENCY DEVELOPMENT
- Learning How to Say What One Means: A Longitudinal Study of Children's Speech Act Use*
- Modeling Developer Burnout with GenAI Adoption
- When Do Generics Feel Justifiable? A Registered Report Bridging Key Theories
- Artificial intelligence tools expand scientists’ impact but contract science’s focus
- Applying machine learning to classify table olives using bacterial metataxonomic data
- Balancing Acts: The Communicative Roles of Cabinet Ministers on Social Media
- Self‐Supervised App‐Based Speech Training for Children With Speech Sound Disorder—A Single‐Case Experimental Design Study
- Aligning EU policies to address biological invasions: assessing invasion impacts across sectors
- Automated Program Repair of Uncompilable Student Code
- Refactoring with LLMs: Bridging Human Expertise and Machine Understanding
- Conservative treatment for stable osteochondritis dissecans of the elbow before epiphyseal closure: effectiveness of elbow immobilization for healing
- Commonsense for Generative Multi-Hop Question Answering Tasks
- Person-Centric Annotations of LAION-400M: Auditing Bias and Its Transfer to Models
- Referenceless Quality Estimation for Natural Language Generation
- Reproducibility and methodological issues of skin post-occlusive and thermal hyperemia assessed by single-point laser Doppler flowmetry
- Beyond named methods: A typology of active learning based on classroom observation networks
- Humanly Certifying Superhuman Classifiers
- Boosting bug localization in software models of video games with simulations and component-specific genetic operations
- Generative Value Conflicts Reveal LLM Priorities
- Case reports unlocked: Harnessing large language models to advance research on child maltreatment
- Metamorphic Testing for Audio Content Moderation Software
- Artificial Intelligence in Multimodal Learning Process Analytics
- Pre‐Imaging Clinical Factors Associated With Cardiac <scp>MR</scp> Image Quality Using Large Language Model‐Enabled Data Extraction
- Beyond kappa: A review of interrater agreement measures
- LGBT+ Inclusion and Human Rights in Taiwan: A Scoping Review of the Literature
- Higher Pitch, Slower Tempo, and Greater Stability in Singing than in Conversation among Mandarin speakers in Auckland: A Registered Report Replicating Ozaki et al. (2024)
- Moving towards Europe-wide freshwater restoration through model-based integration of policy objectives
- Revisiting Maturity Data: Using Oocyte Diameter and Gonadosomatic Index to Retroactively Apply a New Maturity Scale to Greenland Halibut ( <i>Reinhardtius hippoglossoides</i> )
- Understanding trust development in negotiations: An interdependent approach
- Evaluation of the CAN SPAM Act: Testing Deterrence and Other Influences of E-mail Spammer Legal Compliance Over Time
- Provision of knee bracing for knee osteoarthritis (PROP OA): multicentre, parallel group, superiority, statistician blinded, randomised controlled trial
- Distribution and Timing of Verbal Backchannels in Conversational Speech: A Quantitative Study
- Meta-analysis of Interventions for Monitoring Accuracy in Problem Solving
- A Systematic Realist Review of School-Based Working Memory Training
- Do teachers have the knowledge and skills to facilitate effective parental engagement? Findings from a national survey in England
- Comparing Morphometric and Mitochondrial DNA Data from Honeybees and Honey Samples for Identifying Apis mellifera ligustica Subspecies at the Colony Level
- Can growth mindset interventions improve academic achievement? A structured review of the existing evidence
- Predicting Melt Pond Coverage on Arctic Sea Ice From Pre‐Melt Surface Topography
- Becoming visible with limited resources: Non-profit journalists’ perspectives on search engine optimization
- Teacher collaborative knowledge building in Reciprocal Peer Observation
- Safety-netting advice documentation in out-of-hours primary care: a retrospective cohort from 2013 to 2020
- How Parents Play: Play Style as a Function of Gender of Parent, Gender of Child, and Play Context
- A systematic review of AI literacy conceptualization, constructs, and implementation and assessment efforts (2019–2023)
- Student Engagement with GenAI's Tutoring Feedback: A Mixed Methods Study
- LLMs Behind the Scenes: Enabling Narrative Scene Illustration
- How to conduct more systematic reviews of agent-based models and foster theory development - Taking stock and looking ahead
- How News Coverage of Misinformation Shapes Perceptions and Trust
- Empirically evaluating modeling language ontologies: the Peira framework
- The impact of bionic prostheses on users' self-perceptions: A qualitative study
- How Many Interviews Are Enough to Identify Metathemes in Multisited and Cross-cultural Research? Another Perspective on Guest, Bunce, and Johnson’s (2006) Landmark Study
- Personality and prosocial behavior: A theoretical framework and meta-analysis.
- The Climate-Health Communication Gap: Connecting the Dots to Drive Climate Health Action
- A Systematic Review of Teachers’ Knowledge of Self-Regulated Learning and its Assessment
- Building a Pilot Software Quality-in-Use Benchmark Dataset
- Evaluation of drug susceptibility profile of Mycobacterium tuberculosis Lineage 1 from Brazil based on whole genome sequencing and phenotypic methods
- Can We Stop Malicious AI? KILLBENCH: A Benchmark for External AI Kill Switch Feasibility
- The GDN-CC Dataset: Automatic Corpus Clarification for AI-enhanced Democratic Citizen Consultations
- PseudoBridge: Pseudo Code as the Bridge for Better Semantic and Logic Alignment in Code Retrieval
- Technology Acceptance and User Experience
- Age differences in credibility judgments of online health information
- Teachers’ assessment competence in self-regulated learning – presence of non-diagnostic cues diminishes judgment accuracy
- Open-ended Structured Question Assessment with Human-LLM Collaboration
- Exploring Communication and Collaboration in Distributed AR Escape Rooms: Design Opportunities to Support Social Play
- Cognitive Control and Creativity in Children: The Role of Task‐Specific Demands in Divergent Thinking
- Developmental Trajectories of Multicompetent Writers: An Ecological-Historical Approach to L1/L2 Writing Abilities and L2 Proficiency
- ‘AI is like fire’: a discourse analysis of language teachers’ attitudes towards artificial intelligence
- Gender Discrimination in the Visual Representation of Athletes on the Official Instagram Accounts of Sports Federations in Indonesia
- Automated Quantification of Affect Synchrony: Links to Autism Risk and Social Communication in Infancy
- AHA-Memes: A Fine-Grained Multimodal Benchmark for Understanding Hate in Arabic Memes
- Expiratory Muscle Strength Training to Improve Voice and Respiratory Outcomes After Laryngectomy: A Feasibility Study
- Attractive synthetic voices
- Consonant articulation accuracy in paediatric cochlear implant recipients
- Evaluating the environmental sustainability of AI in radiology: a systematic review of current practice
- Digital twins in fertility, assisted reproductive technology and pregnancy: a systematic review
- Pubescence color classification in soybean breeding using aerial images and the Random Forest machine learning algorithm
- Large language models as first-pass filters for corpus annotation: semantic disambiguation of Galician <i>pobo</i>
- Evaluating <scp>UAV</scp> Multispectral Imagery, Machine Learning and Image Analysis Techniques for Mapping Taro and Sweet Potato in a Smallholder Cropland in Swayimane, South Africa
- The revised Cochrane risk of bias tool for randomized trials (RoB 2) showed low interrater reliability and challenges in its application
- Interindividual differences in the dynamics of the homeostatic process are trait‐like and distinct for sleep versus wakefulness
- Interobserver Agreement of Ankle Fracture Classification Among Emergency Physicians
- Blame games, problem denial, and relational distance
- The T-Word Phenomenon: Analyzing Pre-Service Teachers’ Technology Metaphors Through Feenberg’s Critical Theory
- Examining the consistency of instructor versus large language model ratings on summary content: Toward checklist-based feedback provision with second language writers
- “It Became My Buddy, But I’m Not Afraid to Disagree”: A Multi-Session Study of UX Evaluators Collaborating with Conversational AI Assistants
- The Development of an Issue Public: Evidence from The Eras Tour
- The six C’s of successful higher education-industry collaboration in engineering education: a systematic literature review
- Inter‐rater reliability in the assessment of consciousness in patients receiving palliative care in intensive care: A prospective cross sectional observational study
- How Do Players Perceive Gender Discrimination? On the Differences of Harassment in Online Games
- Clinical development and performance of the First to Know Syphilis Self-Test for over-the-counter usage: a <i>de novo</i> rapid test for treponemal antibody
- Ictal emotional features in pediatric and young adult patients with frontal lobe epilepsy
- Tipo de inferencias y comprensión de textos narrativos en básica primaria: Un análisis desde la teoría de sistemas dinámicos
- Revealing local adaptation of Quercus suber L. populations under climate change through Genome Scans and Environmental Association Analysis
- Secondary and 2-Year Outcomes of a Sexual Assault Resistance Program for University Women
- Adverb placement in L1 and L2 spoken production
- Predicting symptom worsening in remitted depression on maintenance pharmacotherapy using digital biomarkers: A prognostic modeling study using machine learning
- School partnered approaches to emotionally based school avoidance in UK primary and secondary school‐age learners: A systematic review
- What Do Case-Control Studies Estimate? Survey of Methods and Assumptions in Published Case-Control Research
- The influence of seductive details in learning environments with low and high extrinsic motivation
- Automated Algorithm for Accurate Waking Sitting and Physical Activity Estimates Without Diaries Using Thigh-Worn Fibion Accelerometers in 10- to 12-Year-Old Children
- The Effectiveness of Physical Literacy Interventions: A Systematic Review with Meta-Analysis
- Turning the Tide: A 2°C Increase in Heat Tolerance Can Halve Climate Change‐Induced Losses in Four Cold‐Adapted Kelp Species
- Artificial Light at Night Reduces Flashing in Photinus and Photuris Fireflies During Courtship and Predation
- Does reciprocal peer observation promote the transfer of learning to teaching practice?
- Relationship between the time course of Burden of Amplitudes and Epileptiform Discharges scores and relapse in children with infantile epileptic spasms syndrome
- Psychopathology in children before and after epilepsy surgery: a prospective controlled study
- Exploring artificial intelligence-powered virtual assistants to understand their potential to support older adults’ search needs
- Blame-validation: Beyond rationality? Effect of causal link on the relationship between evaluation and causal judgment
- Examining item content across nine psychological (in)flexibility scales: What do they measure?
- QRATER: a collaborative and centralized imaging quality control web-based application
- The short-lived hope for contagion: Brexit in social media communication of the populist right
- Primary school students’ awareness of their monitoring and regulation judgment accuracy
- Modeling competences in enterprise architecture: from knowledge, skills, and attitudes to organizational capabilities
- NarrativeTime: Dense Temporal Annotation on a Timeline
- Fake news as a rhetorical weapon: Strategic delegitimization and selective amplification in Italian newspapers
- How do mathematics teachers grade tests? Different scoring and content priorities in micro-decisions, yet similar final grades
- Current Practices and a Novel Operational Framework for Planning Research on Digital Health Promotion Interventions From Development to Implementation: Scoping Review
- Identifying <i>Connectional Silence</i> in Palliative Care Consultations: A Tandem Machine-Learning and Human Coding Method
- Politics and Religion in Secular Societies: The Prominence and Framing of Religion in Nordic Election Campaigns
- Protocol for a Scoping Review of Participatory Action Research on Student Voice in Language Education
- Co-Design and User Evaluation of a Robotic Mental Well-Being Coach to Support University Students’ Public Speaking Anxiety
- Evaluating the Use of Google Street View to Visually Verify the Locations of Cannabis Retailers in the United States Extracted from Websites, 2015–2018
- Turkish Reliability and Validity of the Intensive Care Unit Specific Pressure Injury Risk Assessment Scale ( <scp>RAPS</scp> ‐ <scp>ICU</scp> ): A Tool Validation
- "Can You Tell Me?": Designing Copilots to Support Human Judgement in Online Information Seeking
- MR imaging and CT in osteoarthritis of the lumbar facet joints
- Constructing Policy Narratives in 140 Characters or Less: The Case of Gun Policy Organizations
- Physical Impairments in Adults With Ankle Osteoarthritis: A Systematic Review and Meta-analysis
- Comparison of the Accuracy of WhipPredict to That of a Modified Version of the Short-Form Örebro Musculoskeletal Pain Screening Questionnaire to Predict Poor Recovery After Whiplash Injury
- A Comparison of 3 Methodological Approaches to Defining Major Clinically Important Improvement of 4 Performance Measures in Patients With Hip Osteoarthritis
- Understanding and Detecting Dangerous Speech in Social Media
- Quote RTs on Twitter: Usage of the New Feature for Political Discourse
- A Gesture Recognition System for Detecting Behavioral Patterns of ADHD
- Query-Focused Opinion Summarization for User-Generated Content
- A closed-loop AI framework for hypothesis-driven and interpretable materials design
- AccessEval: Benchmarking Disability Bias in Large Language Models
- Evaluating Generative AI as an Educational Tool for Radiology Resident Report Drafting
- Robustness and Reliability of Gender Bias Assessment in Word Embeddings: The Role of Base Pairs
- DIWALI: Diversity and Inclusivity aWare cuLture specific Items for India: Dataset and Assessment of LLMs for Cultural Text Adaptation in Indian Context
- Prompt-with-Me: in-IDE Structured Prompt Management for LLM-Driven Software Engineering
- Controlled Yet Natural: A Hybrid BDI-LLM Conversational Agent for Child Helpline Training
- Computational-Assisted Systematic Review and Meta-Analysis (CASMA): Effect of a Subclass of GnRH-a on Endometriosis Recurrence
- Mental Multi-class Classification on Social Media: Benchmarking Transformer Architectures against LSTM Models
- Screening for the loss of protective sensation in people without a history of diabetic foot ulceration: Validation of two simple tests in India
- Building Data-Driven Occupation Taxonomies: A Bottom-Up Multi-Stage Approach via Semantic Clustering and Multi-Agent Collaboration
- The importance of being external. methodological insights for the external validation of machine learning models in medicine
- It Depends: Resolving Referential Ambiguity in Minimal Contexts with Commonsense Knowledge
- Codebook LLMs: Evaluating LLMs as Measurement Tools for Political Science Concepts
- On the Use of Agentic Coding: An Empirical Study of Pull Requests on GitHub
- MEENA (PersianMMMU): Multimodal-Multilingual Educational Exams for N-level Assessment
- When Large Language Models Meet UAV Projects: An Empirical Study from Developers' Perspective
- ClaimGen-CN: A Large-scale Chinese Dataset for Legal Claim Generation
- Classification systems for assessing acute muscle injuries: a retrospective comparison of inter-reader agreements
- Poor sleep in organ transplant recipients: self‐reports and actigraphy
- Interobserver and Intraobserver Reliability of the Goutallier Classification Using Magnetic Resonance Imaging
- Derivation of the Screening of Nutritional Risk in Intensive Care risk prediction score: A secondary analysis of a prospective cohort study
- User eXperience Perception Insights Dataset (UXPID): Synthetic User Feedback from Public Industrial Forums
- Policy Change: An Advocacy Coalition Framework Perspective
- SCENIC: A Location-based System to Foster Cognitive Development in Children During Car Rides
- Auto-Slides: An Interactive Multi-Agent System for Creating and Customizing Research Presentations
- AutoOEP -- A Multi-modal Framework for Online Exam Proctoring
- My Favorite Streamer is an LLM: Discovering, Bonding, and Co-Creating in AI VTuber Fandom
- Issue Framing and Engagement: Rhetorical Strategy in Public Policy Debates
- Factors Associated With Background Parenchymal Enhancement on Contrast-Enhanced Mammography
- Sporadic Primary Pheochromocytoma: A Prospective Intraindividual Comparison of Six Imaging Tests (CT, MRI, and PET/CT Using <sup>68</sup>Ga-DOTATATE, FDG, <sup>18</sup>F-FDOPA, and <sup>18</sup>F-FDA)
- Reduced Acquisition Time per Bed Position for PET/MRI Using <sup>68</sup>Ga-RM2 or <sup>68</sup>Ga-PSMA-11 in Patients With Prostate Cancer: A Retrospective Analysis
- Examining teaching assistant pedagogies in traditional laboratories and recitations
- Food insecurity and mental health: a systematic review and meta-analysis
- AOSpine Thoracolumbar Spine Injury Classification System
- Assessment of Hip Abductor Power in Patients With Foot Drop
- What Were You Thinking? An LLM-Driven Large-Scale Study of Refactoring Motivations in Open-Source Projects
- The Knowledge Graph Track at OAEI -- Gold Standards, Baselines, and the Golden Hammer Bias
- MedFactEval and MedAgentBrief: A Framework and Workflow for Generating and Evaluating Factual Clinical Summaries
- Towards an Automated Framework to Audit Youth Safety on TikTok
- From Vision to Validation: A Theory- and Data-Driven Construction of a GCC-Specific AI Adoption Index
- Ultrasound features of traumatic digital nerve injuries of the hand with surgical confirmation
- Methods and uncertainties in bioclimatic envelope modelling under climate change
- A Conceptual Framework for Implicit Evaluation of Conversational Search Interfaces
- Interaction Quality: Assessing the quality of ongoing spoken dialog interaction by experts—And how it relates to user satisfaction
- Social Cognition Psychometric Evaluation: Results of the Initial Psychometric Study
- A review of methods for the assessment of prediction errors in conservation presence/absence models
- Collective Leadership Development: Emerging Themes From Urban, Suburban, and Rural High Schools
- Comparative effectiveness of instructional design features in simulation-based education: Systematic review and meta-analysis
- Generative Goal Modeling
- Waste-Bench: A Comprehensive Benchmark for Evaluating VLLMs in Cluttered Environments
- Inter-rater concordance of basal cell carcinoma subtypes: influences on reporting format and opportunities for further classification modifications
- A Benchmark for Modeling Violation-of-Expectation in Physical Reasoning Across Event Categories
- Guideline Concordance of Large Language Model Responses to Parent-oriented Guideline Prompts About Pediatric Acute Bacterial Arthritis
- People in Tight Cultures and Tight Situations Wear Masks More: Evidence From Three Large-Scale Studies in China
- JADES: A Universal Framework for Jailbreak Assessment via Decompositional Scoring
- Automated Quality Assessment for LLM-Based Complex Qualitative Coding: A Confidence-Diversity Framework
- Examining pre-service teacher competence in lesson planning pertaining to collaborative learning
- Clinical and radiographic outcomes of zirconia dental implants—A systematic review and meta‐analysis
- The effect of different abutment materials on peri‐implant tissues—A systematic review and meta‐analysis
- Multimodal Emotion-Cause Pair Extraction in Conversations
- Preschoolers’ inquisitiveness and science-relevant problem solving
- Generative Interfaces for Language Models
- Assessing the reliability of retrospective reports of adverse childhood experiences among adult HMO members attending a primary care clinic
- M3HG: Multimodal, Multi-scale, and Multi-type Node Heterogeneous Graph for Emotion Cause Triplet Extraction in Conversations
- The Social Position of Pupils with Special Needs in Regular Schools
- LiveMCP-101: Stress Testing and Diagnosing MCP-enabled Agents on Challenging Queries
- DBpedia NIF: Open, Large-Scale and Multilingual Knowledge Extraction\n Corpus
- Systematic Review Of Collaborative Learning Activities For Promoting AI Literacy
- Interrater variation in scoring radiological discrepancies
- Development, validation and field evaluation of a quantitative real-time PCR able to differentiate between field <i>Mycoplasma synoviae</i> and the MS-H-live vaccine strain
- A reappraisal of adult thoracic surface anatomy
- Redefining the projectional and clinical anatomy of the duodenojejunal flexure in children
- Real-Time Classification of Twitter Trends
- Head Pain Referral During Examination of the Neck in Migraine and Tension‐Type Headache
- Parkinson's Disease Diagnosis Using Deep Learning
- Social inclusion through mixed-income development: Design and practice in the Choice Neighborhoods Initiative
- Self-presentation and gender on MySpace
- Avaliação de eficiência na leitura: uma abordagem baseada em PLN
- Is GPT-OSS Good? A Comprehensive Evaluation of OpenAI's Latest Open Source Models
- "My productivity is boosted, but ..." Demystifying Users' Perception on AI Coding Assistants
- A matter of perspective – the role of interpersonal relationships in supply chain risk management
- The self‐administered comorbidity questionnaire: A new method to assess comorbidity for clinical and health services research
- How Do Practitioners Interpret Conditionals in Requirements?
- ToxiFrench: Benchmarking and Enhancing Language Models via CoT Fine-Tuning for French Toxicity Detection
- Credible leadership signals
- The impact of metaphorical language on social media engagement: evidence from the presidential run-off and major parliamentary parties during the 2023 Turkish general election
- Safety of a large language model-based clinical decision support system in African primary healthcare
- Alcohol‐cancer risk communication on social media: A content analysis of alcohol‐related Instagram and TikTok posts
- Evaluating AI-Assisted Deductive Coding in MAXQDA: A Methodological Analysis of Inputs and Outputs
- Identifying Potentially Irregular Electoral Ads in Facebook during the Brazilian Elections
- Small language models applied in text summarization task of health-related news to improve public health audit: an experimental case study
- Pass on the grass? The unexpected “last supper” of hypselodont Pachyrukhinae (Notoungulata, Mammalia) from the late Neogene of northwestern Argentina
- Hybrid Projections Improve Prediction of Distributional Shifts of Invasive and Native Seaweeds Under Climate Change
- Insights into Sexism: Male Status and Performance Moderates Female-Directed Hostile and Amicable Behaviour
- A comparison of prevalence and risk factor profiles of prolonged grief disorder among French and Togolese bereaved adults
- Workplace and non-workplace loneliness: a cross-sectional comparative study on risk factors and impacts on absenteeism and mental health among employees in Spain
- Optimizing Peer Grading: A Systematic Literature Review of Reviewer Assignment Strategies and Quantity of Reviewers
- Scaling Success: A Systematic Review of Peer Grading Strategies for Accuracy, Efficiency, and Learning in Contemporary Education
- Streamlining Admission with LOR Insights: AI-Based Leadership Assessment in Online Master's Program
- Bench-2-CoP: Can We Trust Benchmarking for EU AI Compliance?
- Posterior-GRPO: Rewarding Reasoning Processes in Code Generation
- Beneath the Surface of the Sexual Harassment Label: A Mixed Methods Study of Young Working Women
- Modelling and Classifying the Components of a Literature Review
- From App Features to Explanation Needs: Analyzing Correlations and Predictive Potential
- Optimized imaging prefiltering for enhanced image segmentation
- FilBench: Can LLMs Understand and Generate Filipino?
- On doing relevant and rigorous experiments: Review and recommendations
- Cross-lingual Opinions and Emotions Mining in Comparable Documents
- DeepLung: Deep 3D Dual Path Nets for Automated Pulmonary Nodule Detection and Classification
- A Confidence-Diversity Framework for Calibrating AI Judgement in Accessible Qualitative Coding Tasks
- The eating disorder assessment for DSM‐5 (EDA‐5): Development and validation of a structured interview for feeding and eating disorders
- Knots in the Discourse of Innovation: Investigating Multiple Tensions in a Reacquired Spin-off
- Comparison of Personal Versus Fictional Narratives of Children With Language Impairment
- Exploring Direct Instruction and Summary-Mediated Prompting in LLM-Assisted Code Modification
- UrBLiMP: A Benchmark for Evaluating the Linguistic Competence of Large Language Models in Urdu
- HateBuffer: Safeguarding Content Moderators' Mental Well-Being through Hate Speech Content Modification
- Simultaneous cemented and cementless total knee replacement in the same patients
- Global Study on the Accuracy of Human Epidermal Growth Factor Receptor 2-Low Diagnosis in Breast Cancer
- Automating Thematic Review of Prevention of Future Deaths Reports: Replicating the ONS Child Suicide Study using Large Language Models
- Feedback 2.0 in online writing instruction: Combining audio-visual and text-based commentary to enhance student revision and writing competency
- A Formal Ontology-Based Classification of Lexemes and its Applications
- Knowledge Injection into Dialogue Generation via Language Models
- Product Innovation Processes in Small Firms: Combining Entrepreneurial Effectuation and Managerial Causation
- Leveraging Fine-Tuned Large Language Models for Interpretable Pancreatic Cystic Lesion Feature Extraction and Risk Categorization
- CANDLE: A Cross-Modal Agentic Knowledge Distillation Framework for Interpretable Sarcopenia Diagnosis
- ShEMO -- A Large-Scale Validated Database for Persian Speech Emotion\n Detection
- Tip of the Tongue Known-Item Retrieval: A Case Study in Movie Identification
- "I Would Not Be This Version of Myself Today": Elaborating on the Effects of Eudaimonic Gaming Experiences
- Educating Boundary Crossing Planners: Evidence for Student Learning in the Multistakeholder Regional Learning Environment
- Multi-Year Evaluations of an FTA Card–Based Detection Protocol for Four Vector-Borne Viruses Affecting Potato
- Enhancing digital literacy in children and adolescents: a meta-analysis of school-based interventions
- The effectiveness of ChatGPT as a lexical tool for English, compared with a bilingual dictionary and a monolingual learner’s dictionary
- Gender and Digital Rights: An Empirical Study Among Young Entrepreneurs
- Compliance with the national and WHO antibiotic treatment guidelines for respiratory tract infections and their association with clinical and economic outcomes in Vietnam: an observational study
- The Harmful Dysfunction Analysis applied to the concept of behavioral addiction: toward a new theoretical framework
- The Manitoba Joint Replacement Registry: validation of a provincial hip and knee arthroplasty registry
- Land use transformation and carbon sequestration in the Chittagong Hill Tracts, Bangladesh: a spatiotemporal and predictive analysis with economic implications
- The AIR framework for research transparency: a critical analysis of stage-specific AI disclosure in the context of accessibility and research integrity
- Urban greening to cool towns and cities: A systematic review of the empirical evidence
- COVIDRead: A Large-scale Question Answering Dataset on COVID-19
- STUDY QUALITY IN SLA
- Mediators of change in cognitive behavior therapy and interpersonal psychotherapy for eating disorders: A secondary analysis of a transdiagnostic randomized controlled trial
- <scp>AI</scp> support in self‐regulated learning: A decade of technological evolution and meta‐analysis
- Quantifying Controversy in Social Media
- How Do Dieticians on Instagram Teach? The Potential of the Kirkpatrick Model in the Evaluation of the Effectiveness of Nutritional Education in Social Media
- Learning grammar the explorative way: Integrating interactive grammar animations into CALL
- The role of oral dysbiosis in pregnancy complications: a systematic review and meta-analysis of preterm birth
- Global warming promotes biological invasion of a honey bee pest
- Couples’ interpersonal dynamics and relationship quality in donor-conceived families: a systematic review
- Screening for REM Sleep Behaviour Disorder with Minimal Sensors
- Is "moby dick" a Whale or a Bird? Named Entities and Terminology in\n Speech Translation
- Individual differences in computational psychiatry: A review of current challenges
- TreCap: A wearable device to measure and assess tremor data of visually guided hand movements in real time
- Surgical technique and effectiveness of microendoscopic discectomy for large uncontained lumbar disc herniations: a prospective, randomized, controlled study with 8 years of follow-up
- Atom Responding Machine for Dialog Generation
- Osseointegration in osteoporotic‐like condition: A systematic review of preclinical studies
- What Should I Learn First: Introducing LectureBank for NLP Education and Prerequisite Chain Learning
- A Natural Language Query Interface for Searching Personal Information on\n Smartwatches
- Reliability of CT attenuation value for adrenal masses
- Review on Requirements Modeling and Analysis for Self-Adaptive Systems: A Ten-Year Perspective
- Parental perceptions of children's oral health: the Early Childhood Oral Health Impact Scale (ECOHIS). [europepmc]
- Mean 20-year followup of Bernese periacetabular osteotomy. [europepmc]
- Validity/reliability of PHQ-9 and PHQ-2 depression scales among adults living with HIV/AIDS in western Kenya. [europepmc]
- Reliability of the interRAI suite of assessment instruments: a 12-country study of an integrated health information system. [europepmc]
- Feasibility, reliability, and validity of the EQ-5D-Y: results from a multinational study. [europepmc]
- The sustainability of new programs and innovations: a review of the empirical literature and recommendations for future research. [europepmc]
- Validation of the theoretical domains framework for use in behaviour change and implementation research. [europepmc]
- A systematic review of publications assessing reliability and validity of the Behavioral Risk Factor Surveillance System (BRFSS), 2004-2011. [europepmc]
- Development of a framework and coding system for modifications and adaptations of evidence-based interventions. [europepmc]
- Macrophage polarisation: an immunohistochemical approach for identifying M1 and M2 macrophages. [europepmc]
- Newcastle-Ottawa Scale: comparing reviewers' to authors' assessments. [europepmc]
- Validity of the global physical activity questionnaire (GPAQ) in assessing levels and change in moderate-vigorous physical activity and sedentary behaviour. [europepmc]
- Low health literacy and evaluation of online health information: a systematic review of the literature. [europepmc]
- How to reduce sitting time? A review of behaviour change strategies used in sedentary behaviour reduction interventions among adults. [europepmc]
- Systematic review of the validity and reliability of consumer-wearable activity trackers. [europepmc]
- A guide to using the Theoretical Domains Framework of behaviour change to investigate implementation problems. [europepmc]
- Trauma and PTSD in the WHO World Mental Health Surveys. [europepmc]
- The 2016 WHO classification and diagnostic criteria for myeloproliferative neoplasms: document summary and in-depth discussion. [europepmc]
- Opioids for Chronic Noncancer Pain: A Systematic Review and Meta-analysis. [europepmc]
- Assessment of 68Ga-PSMA-11 PET Accuracy in Localizing Recurrent Prostate Cancer: A Prospective Single-Arm Clinical Trial. [europepmc]
- Understanding the care and support needs of older people: a scoping review and categorisation using the WHO international classification of functioning, disability and health framework (ICF). [europepmc]
- The advantages of the Matthews correlation coefficient (MCC) over F1 score and accuracy in binary classification evaluation. [europepmc]
- Distribution patterns of tau pathology in progressive supranuclear palsy. [europepmc]
- Clinical, pathological, and PAM50 gene expression features of HER2-low breast cancer. [europepmc]
- European Resuscitation Council and European Society of Intensive Care Medicine guidelines 2021: post-resuscitation care. [europepmc]
Discussions
Related