A NEW MEASURE OF RANK CORRELATION
1938/06/01 by M. G. Kendall, M. G. KENDALL · 5,865 citations
Mathematics · #Combinatorics #Computer science #Correlation #Data mining #Geometry #Mathematics #Measure (data warehouse) #Rank (graph theory) #Rank correlation #Statistical Methods and Inference #Statistics #Volume (thermodynamics)
paper · doi:10.1093/biomet/30.1-2.81
published in Biometrika 30(1-2), 81-93 (Oxford University Press)
openalex publication_date 1938/06/01 · openalex created_date 2025/10/10 · openalex updated_date 2026/08/05
Abstract
A NEW MEASURE OF RANK CORRELATION M. G. KENDALL M. G. KENDALL Search for other works by this author on: Oxford Academic Google Scholar Biometrika, Volume 30, Issue 1-2, June 1938, Pages 81–93, 10.1093/biomet/30.1-2.81 Published: 01 June 1938
Cited by
- Hamiltonian expressibility for ansatz selection in variational quantum algorithms
- A Non-Parametric Test of Independence
- Wave climate on the southwestern coast of Lake Michigan: Perspectives from wave directionality
- Gradient forests: calculating importance gradients on physical predictors
- Age‐related trends in niche position and specialization in Neotropical vertebrates
- Determining the Minimum Reliability Standard Based on a Decision Criterion
- Explaining individual predictions when features are dependent: More accurate approximations to Shapley values
- Axiomatic characterization of committee scoring rules
- Observations of a Magellanic Corona
- A Long Way to the Top
- Copula Models for Aggregating Expert Opinions
- Global subnational estimates of migration of scientists reveal large disparities in internal and international flows
- Third Party Tracking in the Mobile Ecosystem
- Khondo: A Multimodal Benchmark for Document Packet Splitting of Bangla Forms
- smartcor: Intelligent Correlation Method Selection for Mixed Variable Types
- Limiting spectral distribution of large dimensional Spearman's rank correlation matrices
- Data Annotations as Pedagogical Hints: From Subjective Labels to Critical Thinking
- GATE-3D: Geometry-Aware Test-time Adaptive Reranking for Open-Set 3D Shape Retrieval
- WorldCupArena: Fine-Grained Evaluation of Language Models and Deep-Research Agents on Football Forecasting
- Leveraging Dissimilarity Invariance as a Robust Anchor for Learning with Noisy Labels
- K-IPO: Kendall-constrained Importance Preserving Oversampling for Imbalanced Tabular Data
- Debiasing Text-to-Image Evaluation via Implicit Cultural Alignment Reward Modeling
- Data Balancing Strategies: A Systematic Survey of Resampling and Augmentation Methods
- Are Performance-Optimization Benchmarks Reliably Measuring Coding Agents?
- Can We Hide Machines in the Crowd? Quantifying Equivalence in LLM-in-the-loop Annotation Tasks
- Epistemic Diversity and Knowledge Collapse in Large Language Models
- ProxAnn: Use-Oriented Evaluations of Topic Models and Document Clustering
- RAVENEA: A Benchmark for Multimodal Retrieval-Augmented Visual Culture Understanding
- Distribution‐Based Global Sensitivity Analysis in Hydrology
- Investigating the Characteristics of One-Sided Matching Mechanisms Under Various Preferences and Risk Attitudes
- The Lehmer factorial norm on Sn
- GAIA: A Transfer Learning System of Object Detection that Fits Your Needs
- Discovering Association with Copula Entropy
- Exploring Memorization in Adversarial Training
- Domain matters: Towards domain-informed evaluation for link prediction
- An empirical appraisal of eLife’s assessment vocabulary
- The Lovasz-Bregman Divergence and connections to rank aggregation, clustering, and web ranking
- FasterPy: An LLM-based Code Execution Efficiency Optimization Framework
- proxymate: Diagnosis and Adjustment of Proxy Estimates for Reliable Inference
- Monotone Retargeting for Unsupervised Rank Aggregation with Object Features
- Node importance estimation in complex networks based on entropic Modularity-Perturbed gravity model
- Collective Schedules: Scheduling Meets Computational Social Choice
- Learning Optimal Representations with the Decodable Information Bottleneck
- Ties, Tails and Spectra: On Rank-Based Dependency Measures in High Dimensions
- Structure-aware Relative Policy Optimization for Ranking
- Spectra of high-dimensional Spearman correlation matrices under scale-mixture dependence
- Translation Quality Assessment: A Brief Survey on Manual and Automatic Methods
- Knowledge without Wisdom: Measuring Misalignment between LLMs and Intended Impact
- A Robust Similarity Estimator
- Narrative Consolidation: Formulating a New Task for Unifying Multi-Perspective Accounts
- Symmetric Rank Covariances: a Generalised Framework for Nonparametric Measures of Dependence
- Perturb Your Data: Paraphrase-Guided Training Data Watermarking
- Asymptotic Inference for Rank Correlations
- Two CFG Nahuatl for automatic corpora expansion
- Embedding-Based Rankings of Educational Resources based on Learning Outcome Alignment: Benchmarking, Expert Validation, and Learner Performance
- What Matters in Evaluating Book-Length Stories? A Systematic Study of Long Story Evaluation
- The Effect of Document Summarization on LLM-Based Relevance Judgments
- NAS-Bench-x11 and the Power of Learning Curves
- A fine-grained look at causal effects in causal spaces
- Assessing Test Case Prioritization on Real Faults and Mutants
- Probabilistic Multi-Agent Aircraft Landing Time Prediction
- VLD: Visual Language Goal Distance for Reinforcement Learning Navigation
- SimGNN: A Neural Network Approach to Fast Graph Similarity Computation
- GreedyNAS: Towards Fast One-Shot NAS with Greedy Supernet
- MaxShapley: Towards Incentive-compatible Generative Search with Fair Context Attribution
- Generalized Pearson correlation squares for capturing mixtures of bivariate linear dependences
- Benchmarking Scientific Understanding and Reasoning for Video Generation using VideoScience-Bench
- Process-Centric Analysis of Agentic Software Systems
- Nonequilibrium Thermodynamics in Measuring Carbon Footprints:\n Disentangling Structure and Artifact in Input-Output Accounting
- A Trainable Centrality Framework for Modern Data
- EfficientBERT: Progressively Searching Multilayer Perceptron via Warm-up Knowledge Distillation
- The life of central radio galaxies in clusters: AGN-ICM studies of eRASS1 clusters in the ASKAP fields
- Rethinking Test Time Scaling for Flow-Matching Generative Models
- MindEval: Benchmarking Language Models on Multi-turn Mental Health Support
- HardCoRe-NAS: Hard Constrained diffeRentiable Neural Architecture Search
- Differentially private testing for relevant dependencies in high dimensions
- Correlation-Aware Feature Attribution Based Explainable AI
- Pharos-ESG: A Framework for Multimodal Parsing, Contextual Narration, and Hierarchical Labeling of ESG Report
- CoSimGNN: Towards Large-scale Graph Similarity Computation
- Novel Tau-Informed Initialization for Maximum Likelihood Estimation of Copulas with Discrete Margins
- No-reference Image Denoising Quality Assessment
- Multivariate Rank-based Distribution-free Nonparametric Testing using Measure Transportation
- Symmetry-Aware Graph Metanetwork Autoencoders: Model Merging through Parameter Canonicalization
- Quantifying and Improving Adaptivity in Conformal Prediction through Input Transformations
- Model-oriented Graph Distances via Partially Ordered Sets
- Privacy-Preserving Explainable AIoT Application via SHAP Entropy Regularization
- An Extreme-Value Approach for Testing the Equality of Large U-Statistic Based Correlation Matrices
- Towards the cycle structures in complex network: A new perspective
- Knowledge-based anomaly detection for identifying network-induced shape artifacts
- SORTeD Rashomon Sets of Sparse Decision Trees: Anytime Enumeration
- Riemannian tangent space mapping and elastic net regularization for cost-effective EEG markers of brain atrophy in Alzheimer's disease
- How to select predictive models for decision-making or causal inference
- Maximizing spreading in complex networks with risk in node activation
- High Arctic vegetation lowers soil temperatures on sunny days despite low stature
- BARTScore: Evaluating Generated Text as Text Generation
- A new Gini correlation between quantitative and qualitative variables
- A Topic Modeling Approach to Ranking
- A mathematical perspective on hypothesis-driven model construction: A case study in pea
- Rising Heat, Rising Risks: Understanding the Nexus of Marine Heatwaves, Fishing Dependence, and Vulnerability to Coastal Communities
- Concordance and the Smallest Covering Set of Preference Orderings
- hyppo: A Multivariate Hypothesis Testing Python Package
- SimPol: Simulating polarisation in political belief networks in European countries
- On the quantification and efficient propagation of imprecise probabilities with copula dependence
- Robust Attribution Regularization
- MixPath: A Unified Approach for One-shot Neural Architecture Search
- A rate-dependent coreset selector for continual learning on time-varying data distributions
- Semantic Foggy Scene Understanding with Synthetic Data
- Modeling Uncertainty with Hedged Instance Embedding
- Comparison of Values of Pearson's and Spearman's Correlation Coefficients on the Same Sets of Data
- Deeper Insights into Weight Sharing in Neural Architecture Search
- How Well Can LLM Agents Simulate End-User Security and Privacy Attitudes and Behaviors?
- Enhancing Linguistic Competence of Language Models through Pre-training with Language Learning Tasks
- Testing Correlation in Graphs by Counting Bounded Degree Motifs
- Seeing Through the MiRAGE: Evaluating Multimodal Retrieval Augmented Generation
- Effect of micronutrients on fertility and aneuploidy rates in human conceptions: a systematic review and meta-analysis
- More Water, More of the Time: Spatial Changes in Flooding Over 83 Years in the Upper Mississippi River Floodplain and Relationships With Streamgage‐Derived Proxies
- MATCH: Task-Driven Code Evaluation through Contrastive Learning
- Implicit Modeling for Transferability Estimation of Vision Foundation Models
- AutoBench: Automating LLM Evaluation through Reciprocal Peer Assessment
- Human-Centred Evaluation of Text-to-Image Generation Models for Self-expression of Mental Distress: A Dataset Based on GPT-4o
- Chitchat with AI: Understand the supply chain carbon disclosure of companies worldwide through Large Language Model
- Efficient Algorithms for Computing Random Walk Centrality
- Nanoscale Mapping of Transition Metal Ordering in Individual LiNi0.5Mn1.5O4 Particles Using 4D-STEM ACOM Technique
- Human-Agent Collaborative Paper-to-Page Crafting
- SimBA: Simplifying Benchmark Analysis Using Performance Matrices Alone
- Assessing Monotone Dependence: Area Under the Curve Meets Rank Correlation
- Too Far From Relatives? Impact of the Genetic Distance on the Success of Exon Capture in Phylogenomics
- Exploring Structural Degradation in Dense Representations for Self-supervised Learning
- SafeSearch: Do Not Trade Safety for Utility in LLM Search Agents
- Readability Reconsidered: A Cross-Dataset Analysis of Reference-Free Metrics
- Capturing Context-Aware Route Choice Semantics for Trajectory Representation Learning
- Inferring network structure in non-normal and mixed discrete-continuous genomic data
- Signature in Code Backdoor Detection, how far are we?
- Understanding and Using the Relative Importance Measures Based on Orthogonalization and Reallocation
- Inter-rater Agreement on Sentence Formality
- Coding Inequity: Assessing GPT-4’s Potential for Perpetuating Racial and Gender Biases in Healthcare
- Quantifying consensus of rankings based on q-support patterns
- Identifying super-spreaders in information-epidemic coevolving dynamics on multiplex networks
- Efficient Bayesian Inference from Noisy Pairwise Comparisons
- MIRAGE: Runtime Scheduling for Multi-Vector Image Retrieval with Hierarchical Decomposition
- MaP: A Unified Framework for Reliable Evaluation of Pre-training Dynamics
- Consensus measure of rankings
- OPANAS: One-Shot Path Aggregation Network Architecture Search for Object Detection
- ISTA-NAS: Efficient and Consistent Neural Architecture Search by Sparse Coding
- SGAS: Sequential Greedy Architecture Search
- Climate change has increased global evaporative demand except in South Asia
- How Confident are Video Models? Empowering Video Models to Express their Uncertainty
- MLE-Smith: Scaling MLE Tasks with Automated Multi-Agent Pipeline
- Measures of Dependence based on Wasserstein distances
- Power-divergence copulas: A new class of Archimedean copulas, with an insurance application
- Signed network models for portfolio optimization
- Don't Pass@k: A Bayesian Framework for Large Language Model Evaluation
- From Absolute to Relative Code Comprehensibility Prediction
- Does higher interpretability imply better utility? A Pairwise Analysis on Sparse Autoencoders
- NSGANetV2: Evolutionary Multi-Objective Surrogate-Assisted Neural Architecture Search
- Deep Learning for Precipitation Nowcasting: A Benchmark and A New Model
- Transparent Reference-free Automated Evaluation of Open-Ended User Survey Responses
- QUARTZ : QA-based Unsupervised Abstractive Refinement for Task-oriented Dialogue Summarization
- Basic Cycle Ratio: Cost-Effective Ranking of Influential Spreaders from Local and Global Perspectives
- Human-MME: A Holistic Evaluation Benchmark for Human-Centric Multimodal Large Language Models
- The SAGA Survey. VI. The Size-Mass Relation for Low-Mass Galaxies Across Environments
- Partial analyses of humeral shape in sauropodomorph dinosaurs highlight a hidden modularity and the differential evolution of sauropod bauplan‐related traits
- LUQ: Layerwise Ultra-Low Bit Quantization for Multimodal Large Language Models
- Mapping Overlaps in Benchmarks through Perplexity in the Wild
- Algal enrichment process for newly formed sea ice in the Dalton polynya off East Antarctica during the late summer–early autumn
- Green Prompt Engineering: Investigating the Energy Impact of Prompt Design in Software Engineering
- Prompt-Aware Scheduling for Low-Latency LLM Serving
- Accelerate Creation of Product Claims Using Generative AI
- Inconsistency among evaluation metrics in link prediction
- Measuring Association on Topological Spaces Using Kernels and Geometric Graphs
- AutoSpec: An Agentic Framework for Automatically Drafting Patent Specification
- A New Correlation Coefficient for Aggregating Non-strict and Incomplete Rankings
- On the scale of heterogeneity in composite electrodes of batteries
- EVALUATE NODE IMPORTANCE BY DECOMPOSING NETWORK WITH A RECURSIVE PERCOLATION PROCESS
- Estimating Robot Strengths with Application to Selection of Alliance Members in FIRST Robotics Competitions
- Measuring Price Discrimination and Steering on E-commerce Web Sites
- Evaluation of pathological complete response as a surrogate endpoint for overall survival in resectable oesophageal cancer: integrated analysis of individual patient data from phase III trials
- Demographically-Inspired Query Variants Using an LLM
- Local Distance Constrained Bribery in Voting
- Spoiler Susceptibility in Party Elections
- A U-statistic Approach to Hypothesis Testing for Structure Discovery in Undirected Graphical Models
- Computational Socioeconomics
- Characterizing cycle structure in complex networks
- From Independence to Interaction: Speaker-Aware Simulation of Multi-Speaker Conversational Timing
- Analyzing and Mitigating Surface Bias in Code Evaluation Metrics
- Evaluating Compiler Optimization Impacts on zkVM Performance
- A Multi-To-One Interview Paradigm for Efficient MLLM Evaluation
- Prompt Stability in Code LLMs: Measuring Sensitivity across Emotion- and Personality-Driven Variations
- Magnetic Reconnection as a Potential Driver of X-ray Variability in Active Galactic Nuclei
- Learning Temporal Embeddings for Complex Video Analysis
- Sampling Permutations for Shapley Value Estimation
- Robust Functional Principal Component Analysis for Non-Gaussian Longitudinal Data
- Predicting Generalization in Deep Learning via Metric Learning -- PGDL Shared task
- Deep Active Learning by Leveraging Training Dynamics
- LLM-as-a-Judge: Rapid Evaluation of Legal Document Recommendation for Retrieval-Augmented Generation
- A key node identification method based on neighborhood-derived cluster method
- Permutation-Based Distances for Groups and Group-Valued Time Series
- Extreme Galaxy-scale Outflows Are Frequent among Luminous Early Quasars
- The Game is the Game: Dynamic network analysis and shifting roles in criminal networks
- Verbalized Algorithms: Classical Algorithms are All You Need (Mostly)
- The Lyα and Continuum Origins Survey. III. Investigating the link between galaxy morphology, merger properties and LyC escape
- Efficiently Ranking Software Variants with Minimal Benchmarks
- Measuring General Associations in Time Series: An Adaptation and Empirical Evaluation of the CODEC Coefficient in Determining Autoregressive Dynamics
- LogME: Practical Assessment of Pre-trained Models for Transfer Learning
- ProMQA-Assembly: Multimodal Procedural QA Dataset on Assembly
- Cryptocurrency market structure: connecting emotions and economics
- Degree-degree dependencies in directed networks with heavy-tailed degrees
- Empowering Large Language Model for Sequential Recommendation via Multimodal Embeddings and Semantic IDs
- Weighted H-index for identifying influential spreaders
- The Fools are Certain; the Wise are Doubtful: Exploring LLM Confidence in Code Completion
- CCE: Confidence-Consistency Evaluation for Time Series Anomaly Detection
- DeepResearch Arena: The First Exam of LLMs' Research Abilities via Seminar-Grounded Tasks
- Speaking at the Right Level: Literacy-Controlled Counterspeech Generation with RAG-RL
- Access Paths for Efficient Ordering with Large Language Models
- Human-AI Collaborative Bot Detection in MMORPGs
- SurGE: A Benchmark and Evaluation Framework for Scientific Survey Generation
- Streamlining the Development of Active Learning Methods in Real-World Object Detection
- Learning Probabilistic Ordinal Embeddings for Uncertainty-Aware Regression
- Assessing the Noise Robustness of Class Activation Maps: A Framework for Reliable Model Interpretability
- Discourse Structure in Machine Translation Evaluation
- Should one (be allowed to) replace the Cippolini's?
- Ranking and Tuning Pre-trained Models: A New Paradigm for Exploiting Model Hubs
- Match & Choose: Model Selection Framework for Fine-tuning Text-to-Image Diffusion Models
- On the Gaussian distribution of the Mann-Kendall tau in the case of autocorrelated data
- From Promise to Practical Reality: Transforming Diffusion MRI Analysis with Fast Deep Learning Enhancement
- Reducing the weight of low exam scores may raise average grades but does not appear to impact equity gaps
- Sensitivity Analysis to Unobserved Confounding with Copula-based Normalizing Flows
- A Weighted Generalization of the Graham-Diaconis Inequality for Ranked List Similarity
- Personalized Recommendations via Active Utility-based Pairwise Sampling
- Investigating the Reordering Capability in CTC-based Non-Autoregressive End-to-End Speech Translation
- Human-Alignment and Calibration of Inference-Time Uncertainty in Large Language Models
- Modulation of the association between blood glucose homeostasis and social hierarchy among co-housed mice by diet and amygdala activities
- Coverage correlation: detecting singular dependencies between random variables
- A Technique Based on Trade-off Maps to Visualise and Analyse Relationships Between Objectives in Optimisation Problems
- Agoran: An Agentic Open Marketplace for 6G RAN Automation
- High-dimensional Mixed Graphical Models
- Identify influential nodes in directed networks: A neighborhood entropy-based method
- User-Relatedness and Community Structure in Social Interaction Networks
- Trae Agent: An LLM-based Agent for Software Engineering with Test-time Scaling
- A Provable Smoothing Approach for High Dimensional Generalized Regression with Applications in Genomics
- Persona-Augmented Benchmarking: Evaluating LLMs Across Diverse Writing Styles
- Structural-Aware Key Node Identification in Hypergraphs via Representation Learning and Fine-Tuning
- Scaling in the global spreading patterns of pandemic Influenza A (H1N1) and the role of control: empirical statistics and modeling
- Continuous-Time User Modeling in the Presence of Badges: A Probabilistic Approach
- Commonsense Knowledge Mining from Term Definitions
- Some New Copula Based Distribution-free Tests of Independence among Several Random Variables
- Submodularity on Hypergraphs: From Sets to Sequences
- Accelerating Grasp Exploration by Leveraging Learned Priors
- Joint Source Selection and Data Extrapolation in Social Sensing for Disaster Response
- A metric for sets of trajectories that is practical and mathematically consistent
- Is (poly-) substance use associated with impaired inhibitory control? A mega-analysis controlling for confounders
- Role of calibration in uncertainty-based referral for deep learning
- Matchings under Preferences: Strength of Stability and Trade-offs
- Symbiotic Agents: A Novel Paradigm for Trustworthy AGI-driven Networks
- From Sorting Algorithms to Scalable Kernels: Bayesian Optimization in High-Dimensional Permutation Spaces
- mNARX+: A surrogate model for complex dynamical systems using manifold-NARX and automatic feature selection
- Fairness in Rankings and Recommendations: An Overview
- Learning user-specific latent influence and susceptibility from information cascades
- PONAS: Progressive One-shot Neural Architecture Search for Very Efficient Deployment
- Inferring the finest pattern of mutual independence from data
- A Theoretical Analysis of NDCG Type Ranking Measures
- On the Detection of Mutual Influences and Their Consideration in Reinforcement Learning Processes
- Data-Driven Differential Evolution in Tire Industry Extrusion: Leveraging Surrogate Models
- NavComposer: Composing Language Instructions for Navigation Trajectories through Action-Scene-Object Modularization
- Predicting the Reproducibility of Social and Behavioral Science Papers Using Supervised Learning Models
- Apps, Places and People: strategies, limitations and trade-offs in the physical and digital worlds
- Evaluating Generated Commit Messages with Large Language Models
- A predictive analytics framework for early detection of production halts and quality issues
- Multilayer Artificial Benchmark for Community Detection (mABCD)
- Large Language Models Have Intrinsic Meta-Cognition, but Need a Good Lens
- Asymmetry of cross correlations between intra-day and overnight volatilities
- Mapping Crisis-Driven Market Dynamics: A Transfer Entropy and Kramers-Moyal Approach to Financial Networks
- Does UMBRELA Work on Other LLMs?
- Extreme event propagation using counterfactual theory and vine copulas
- Composite Optimization with Indicator Functions: Stationary Duality and a Semismooth Newton Method
- Anthropomimetic Uncertainty: What Verbalized Uncertainty in Language Models is Missing
- Data Depth as a Risk
- A nonparametric distribution-free test of independence among continuous random vectors based on \texorpdfstringL1-norm
- Mallows Model with Learned Distance Metrics: Sampling and Maximum Likelihood Estimation
- Measuring Hypothesis Testing Errors in the Evaluation of Retrieval Systems
- Limited individual attention and online virality of low-quality information
- SCOOTER: A Human Evaluation Framework for Unrestricted Adversarial Examples
- Beyond Connectivity: Higher-Order Network Framework for Capturing Memory-Driven Mobility Dynamics
- Are Pretrained Multilingual Models Equally Fair Across Languages?
- Checklist Engineering Empowers Multilingual LLM Judges
- Metrics on Permutation Families Defined by a Restriction Graph
- Train-before-Test Harmonizes Language Model Rankings
- Incorporating Interventional Independence Improves Robustness against Interventional Distribution Shift
- Anti-Prompt: Image Protection against Text-Guided Image-to-Video Generation
- Robust reputation-based ranking on multipartite rating networks
- SAVVY: Student Attention Visualization for Video-based Learning Analysis
- Quantitative Storytelling in the Making of a Composite Indicator
- Learning to hash with semantic similarity metrics and empirical KL divergence
- VOCAL: Visual Odometry via ContrAstive Learning
- Comparing rankings by means of competitivity graphs: structural properties and computation
- On the power of Chatterjee rank correlation
- Helpful assistant or fruitful facilitator? Investigating how personas affect language model behavior
- Dynamic Contrastive Learning for Hierarchical Retrieval: A Case Study of Distance-Aware Cross-View Geo-Localization
- A case for data valuation transparency via DValCards
- How does Weight Correlation Affect the Generalisation Ability of Deep Neural Networks
- Video Perception Models for 3D Scene Synthesis
- Predicting missing links via correlation between nodes
- Deep Electromagnetic Structure Design Under Limited Evaluation Budgets
- Spotting Out-of-Character Behavior: Atomic-Level Evaluation of Persona Fidelity in Open-Ended Generation
- Re-Evaluating Code LLM Benchmarks Under Semantic Mutation
- Partial correlation analysis: Applications for financial markets
- Massive parallelization of projection-based depths
- Pattern-Based Graph Classification: Comparison of Quality Measures and Importance of Preprocessing
- Efficient space reduction techniques by optimized majority rules for the Kemeny aggregation problem and beyond
- NeurIPS 2025 E2LM Competition : Early Training Evaluation of Language Models
- References Matter: Investigating the Impact of Reference Set Variation on Summarization Evaluation
- Probabilistic patient risk profiling with pair-copula constructions
- SceneGram: Conceptualizing and Describing Tangrams in Scene Context
- When to Prune? A Policy towards Early Structural Pruning
- Fantastic Generalization Measures and Where to Find Them
- Towards Bridging Formal Methods and Human Interpretability
- Hide & Seek: Transformer Symmetries Obscure Sharpness & Riemannian Geometry Finds It
- Loss Functions for Predictor-based Neural Architecture Search
- Efficient Multistate Free-Energy Calculations with QM/MM Accuracy Using Replica-Exchange Enveloping Distribution Sampling
- Perfecting Depth: Uncertainty-Aware Enhancement of Metric Depth
- Efficient Computation of the Bergsma-Dassios Sign Covariance
- On the Non-degeneracy of Kendall's and Spearman's Correlation Coefficients
- Prioritized Architecture Sampling with Monto-Carlo Tree Search
- Recursive Neural Network Based Preordering for English-to-Japanese Machine Translation
- Literature Survey on Interplay of Topics, Information Diffusion and Connections on Social Networks
- Reuse, Temporal Dynamics, Interest Sharing, and Collaboration in Social Tagging Systems
- Marčenko-Pastur Law for Kendall's Tau
- Survey of Active Learning Hyperparameters: Insights from a Large-Scale Experimental Grid
- Analyzing a practitioner perspective on relevance of published empirical research in Requirements Engineering
- EssayBench: Evaluating Large Language Models in Multi-Genre Chinese Essay Writing
- Lessons Learned from the URGENT 2024 Speech Enhancement Challenge
- Maximum likelihood estimation for mechanistic network models
- Catching Stray Balls: Football, fandom, and the impact on digital discourse
- Judging the Judges: Evaluating the Performance of International Gymnastics Judges
- Lowest Degree Decomposition of Complex Networks
- E-Sports Talent Scouting Based on Multimodal Twitch Stream Data
- Tracking Large-Scale Video Remix in Real-World Events
- GreedyNASv2: Greedier Search with a Greedy Path Filter
- LegalEval-Q: A New Benchmark for The Quality Evaluation of LLM-Generated Legal Text
- Interpretable phenotyping of Heart Failure patients with Dutch discharge letters
- Learning Distributions over Permutations and Rankings with Factorized Representations
- EmotionRankCLAP: Bridging Natural Language Speaking Styles and Ordinal Speech Emotion via Rank-N-Contrast
- Threading the Needle: Reweaving Chain-of-Thought Reasoning to Explain Human Label Variation
- Intermittency in Interplanetary Coronal Mass Ejections Observed by Parker Solar Probe and Solar Orbiter
- Practical estimation of the optimal classification error with soft labels and calibration
- A Discussion on Solving Partial Differential Equations using Neural\n Networks
- Enhancing the Comprehensibility of Text Explanations via Unsupervised Concept Discovery
- Can LLMs Evaluate What They Cannot Annotate? Revisiting LLM Reliability in Hate Speech Detection
- Approximately Optimal Binning for the Piecewise Constant Approximation of the Normalized Unexplained Variance (nUV) Dissimilarity Measure
- LAMDA: A Longitudinal Android Malware Benchmark for Concept Drift Analysis
- Sampling from Conditional Distributions of Simplified Vines
- Locating influential nodes via dynamics-sensitive centrality
- An Experimental Exploration of Marsaglia's xorshift Generators, Scrambled
- Isotonic Bradley-Terry Model for Paired Comparison Data
- lmgame-Bench: How Good are LLMs at Playing Games?
- Angle-based Search Space Shrinking for Neural Architecture Search
- A Survey of Hierarchy Identification in Social Networks
- DSNAS: Direct Neural Architecture Search without Parameter Retraining
- ImaginE: An Imagination-Based Automatic Evaluation Metric for Natural Language Generation
- Binary stars take what they get: Evidence for Efficient Mass Transfer from Stripped Stars with Rapidly Rotating Companions
- Stable Geodesic Update on Hyperbolic Space and its Application to Poincare Embeddings
- A Copula Statistic for Measuring Nonlinear Multivariate Dependence
- GPU-accelerated Kendall distance computation for large or sparse data
- Pairwise Calibrated Rewards for Pluralistic Alignment
- Keep the Gradients Flowing: Using Gradient Flow to Study Sparse Network Optimization
- Unsupervised Inductive Graph-Level Representation Learning via Graph-Graph Proximity
- MRGRP: Empowering Courier Route Prediction in Food Delivery Service with Multi-Relational Graph
- Quantum hub and authority centrality measures for directed networks based on continuous-time quantum walks
- CARES: Comprehensive Evaluation of Safety and Adversarial Robustness in Medical LLMs
- MedGUIDE: Benchmarking Clinical Decision-Making in Large Language Models
- SEAL: Searching Expandable Architectures for Incremental Learning
- Cream of the Crop: Distilling Prioritized Paths For One-Shot Neural Architecture Search
- Phase transitions in quantum-circuit compilation
- Decreasing trends of particle number and black carbon mass concentrations at 16 observational sites in Germany from 2009 to 2018
- Adaptive Sequence Submodularity
- The Great Wave: Evidence of a large-scale vertical corrugation propagating outwards in the Galactic disc
- A test against trend in random sequences
- KOKKAI DOC: An LLM-driven framework for scaling parliamentary representatives
- Inequalities, Preferences and Rankings in US Sports Coach Hiring Networks
- Learning partially ranked data based on graph regularization
- Spectral Analysis of Fake News Propagation
- Probability of a Condorcet Winner for Large Electorates: An Analytic Combinatorics Approach
- Association and Independence Test for Random Objects
- Learning-based Efficient Graph Similarity Computation via Multi-Scale Convolutional Set Matching
- OODTE: A Differential Testing Engine for the ONNX Optimizer
- Exploring Communication Strategies for Collaborative LLM Agents in Mathematical Problem-Solving
- On the symmetry of evidential support
- Towards a practical measure of interference for reinforcement learning
- Assessing the quality of scientific conferences based on bibliographic citations
- Scalable and Interpretable Representation Alignment with Ordinal Similarity
- Comparison of manual and automatic daily sunshine duration measurements at German climate reference stations
- Competition structured a Late Cretaceous megaherbivorous dinosaur assemblage
- Lumos-Nexus: Efficient Frequency Bridging with Homogeneous Latent Space for Video Unified Models
- Leaderboard Incentives: Model Rankings under Strategic Post-Training
- How to Train Your Super-Net: An Analysis of Training Heuristics in Weight-Sharing NAS
- Tail behavior of dependent V-statistics and its applications
- Assessing potential effects of highway and urban runoff on receiving streams in total maximum daily load watersheds in Oregon using the stochastic empirical loading and dilution model
- A History of Distribution Sampling Prior to the Era of the Computer and its Relevance to Simulation
- There is Limited Correlation between Coverage and Robustness for Deep Neural Networks
- Can secure habits counter phishing? An exploration using a novel in-tray simulation
- Learning to Predict Combinatorial Structures
- Has the magnitude of floods across the USA changed with global CO2 levels?
- Investigating the characteristics of one-sided matching mechanisms under various preferences and risk attitudes
- MEMO: Memory-Augmented Model Context Optimization for Robust Multi-Turn Multi-Agent LLM Games
- Frustratingly Easy Transferability Estimation
- A Bayesian Model of node interaction in networks
- Robust rank correlation based screening
- Understanding Learning Dynamics for Neural Machine Translation
- Universal statistics of the knockout tournament
- Research and Teaching Efficiencies of Turkish Universities with Heterogeneity Considerations: Application of Multi-Activity DEA and DEA by Sequential Exclusion of Alternatives Methods
- The impact of class imbalance in logistic regression models for low-default portfolios in credit risk
- Could Large Language Models work as Post-hoc Explainability Tools in Credit Risk Models?
- Efficient Model Performance Estimation via Feature Histories
- Minimizing Time-to-Rank: A Learning and Recommendation Approach
- Paper2Rebuttal: A Multi-Agent Framework for Transparent Author Response Assistance
- Revealing dynamics of non-autonomous complex systems from data
- Ask Early, Ask Late, Ask Right: When Does Clarification Timing Matter for Long-Horizon Agents?
- LMME3DHF: Benchmarking and Evaluating Multimodal 3D Human Face Generation with LMMs
- Semantic Level of Detail for Knowledge Graphs: Discovering Abstraction Boundaries via Spectral Heat Diffusion
- Same Meaning, Different Scores: Lexical and Syntactic Sensitivity in LLM Evaluation
- Learning to Rank Critical Road Segments via Heterogeneous Graphs with OD Flow Integration
- Probabilistic Forecasting for Day-ahead Electricity Prices, Battery Trading Strategies and the Economic Evaluation of Predictive Accuracy
- Investigating Co-Constructive Behavior of Large Language Models in Explanation Dialogues
- Tidal amplification and salt intrusion in the Mekong Delta driven by anthropogenic sediment starvation
- Fluvial sediment supply to a mega-delta reduced by shifting tropical-cyclone activity
- ArrowFlow: Hierarchical Machine Learning in the Space of Permutations
- DOTS: Decoupling Operation and Topology in Differentiable Architecture Search
- Marginalized Frailty-Based Illness-Death Model: Application to the UK-Biobank Survival Data
- Inconsistencies in a schedule of paired comparisons
- Properties of sports ranking methods
- Ordinal Measures of Association
- A statistical analysis of the nulling pulsar population
- Searching for Sound-Meaning Collisions: Graph-Based Affordance Retrieval and Multi-Evaluator Ranking for Pun Translation at CLEF 2026 JOKER Task 2
- A Dual Evaluation for Music Transcription
- Aggregate Characterization of User Behavior in Twitter and Analysis of the Retweet Graph
- Density estimation with distribution element trees
- A New K-Shell Decomposition Method for Identifying Influential Spreaders of Epidemics on Community Networks
- The X-Ray-to-Optical Properties of Optically Selected Active Galaxies over Wide Luminosity and Redshift Ranges
- A Pathologist-Annotated Dataset for Validating Artificial Intelligence: A Project Description and Pilot Study
- TheHerschelExploitation of Local Galaxy Andromeda (HELGA)
- Identifying influential spreaders by weighted LeaderRank
- Accurate algorithms for identifying the median ranking when dealing with weak and partial rankings under the Kemeny axiomatic approach
- Automated Creativity Evaluation for Large Language Models: A Reference-Based Approach
- A high-resolution, dust-selected molecular cloud catalogue of M33, the Triangulum Galaxy
- Regression and Learning to Rank Aggregation for User Engagement Evaluation
- Quantum annealing versus classical machine learning applied to a simplified computational biology problem
- Dimensionality Reduction for Categorical Data
- Targeted Deep Learning: Framework, Methods, and Applications
- The X‐Ray Evolution of Early‐Type Galaxies in the Extended Chandra Deep Field–South
- A consistent test of independence based on a sign covariance related to Kendall’s tau
- What Determines the Local Metallicity of Galaxies: Global Stellar Mass, Local Stellar Mass Surface Density, or Star Formation Rate?
- Game of collusions
- DC3N observations towards high-mass star-forming regions
- DYNAMICAL INFERENCE FROM A KINEMATIC SNAPSHOT: THE FORCE LAW IN THE SOLAR SYSTEM
- Trends in Canadian Short‐Duration Extreme Rainfall: Including an Intensity–Duration–Frequency Perspective
- DVLTA-VQA: Decoupled Vision-Language Modeling with Text-Guided Adaptation for Blind Video Quality Assessment
- Non-Gaussian fluctuations for traces of squared sample correlation matrices in high dimensions
- Disparate Privacy Vulnerability: Targeted Attribute Inference Attacks and Defenses
- Standardization of Weighted Ranking Correlation Coefficients
- A Nature-Inspired Colony of Artificial Intelligence System with Fast, Detailed, and Organized Learner Agents for Enhancing Diversity and Quality
- Six Degrees of Epistasis: Statistical Network Models for GWAS. [europepmc]
- Group personality during collective decision-making: a multi-level approach. [europepmc]
- Histone modifications rather than the novel regional centromeres of Zymoseptoria tritici distinguish core and accessory chromosomes. [europepmc]
- Novel efficient genome-wide SNP panels for the conservation of the highly endangered Iberian lynx. [europepmc]
- Assessing mental stress from the photoplethysmogram: a numerical study. [europepmc]
- Breast MRI radiomics: comparison of computer- and human-extracted imaging phenotypes. [europepmc]
- Machine-learning-derived classifier predicts absence of persistent pain after breast cancer surgery with high accuracy. [europepmc]
- The Influence of Higher-Order Epistasis on Biological Fitness Landscape Topography. [europepmc]
- Neuronal Response Latencies Encode First Odor Identity Information across Subjects. [europepmc]
- Impact of Data Presentation on Physician Performance Utilizing Artificial Intelligence-Based Computer-Aided Diagnosis and Decision Support Systems. [europepmc]
- Why rankings of biomedical image analysis competitions should be interpreted with care. [europepmc]
- Identifying influential spreaders by gravity model. [europepmc]
- Large-Scale Benchmark of Exchange-Correlation Functionals for the Determination of Electronic Band Gaps of Solids. [europepmc]
- Measurement of Cyanobacterial Bloom Magnitude using Satellite Remote Sensing. [europepmc]
- Additive Dose Response Models: Defining Synergy. [europepmc]
- Impact of contouring variability on oncological PET radiomics features in the lung. [europepmc]
- Methods and open-source toolkit for analyzing and visualizing challenge results. [europepmc]
- Model-based assessment of replicability for genome-wide association meta-analysis. [europepmc]
- A topology-preserving dimensionality reduction method for single-cell RNA-seq data using graph autoencoder. [europepmc]
- Performance of early warning signals for disease re-emergence: A case study on COVID-19 data. [europepmc]
- PPI-Affinity: A Web Tool for the Prediction and Optimization of Protein-Peptide and Protein-Protein Binding Affinity. [europepmc]
- The Medical Segmentation Decathlon. [europepmc]
- Scoring Functions for Protein-Ligand Binding Affinity Prediction using Structure-Based Deep Learning: A Review. [europepmc]
- What Makes a Potent Nitrosamine? Statistical Validation of Expert-Derived Structure-Activity Relationships. [europepmc]
- Segmenting functional tissue units across human organs using community-driven development of generalizable machine learning algorithms. [europepmc]
Related