A Scale for the Measurement of the Psychological Magnitude Pitch
1937/01/01 by S. S. Stevens, J. Volkmann, Jens Volkmann +2 · 77 citations
Arts and Humanities · Computer Science · Mathematics · Neuroscience · Psychology · #Acoustics #Astrophysics #Audiology #Basilar membrane #Diverse Music Education Insights #Loudness #Magnitude (astronomy) #Mathematics #Music and Audio Processing #Musical #Neuroscience and Music Perception #Overtone #Physics #Pitch (Music) #Psychoacoustics #Psychology #Relative pitch #Scale (ratio) #Spectral line #Tone (literature)
paper · doi:10.1121/1.1915893
openalex publication_date 1937/01/01 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/31
Abstract
A subjective scale for the measurement of pitch was constructed from determinations of the half-value of pitches at various frequencies. This scale differs from both the musical scale and the frequency scale, neither of which is subjective. Five observers fractionated tones of 10 different frequencies at a loudness level of 60 db. From these fractionations a numerical scale was constructed which is proportional to the perceived magnitude of subjective pitch. In numbering the scale the 1000-cycle tone was assigned the pitch of 1000 subjective units (mels). The close agreement of the pitch scale with an integration of the differential thresholds (DL's) shows that, unlike the DL's for loudness, all DL's for pitch are of uniform subjective magnitude. The agreement further implies that pitch and differential sensitivity to pitch are both rectilinear functions of extent on the basilar membrane. The correspondence of the pitch scale and the experimentally determined location of the resonant areas of the basilar membrane suggests that, in cutting a pitch in half, the observer adjusts the tone until it stimulates a position half-way from the original locus to the apical end of the membrane. Measurement of the subjective size of musical intervals (such as octaves) in terms of the pitch scale shows that the intervals become larger as the frequency of the mid-point of the interval increases (except in the two highest audible octaves). This result confirms earlier judgments as to the relative size of octaves in different parts of the frequency range.
Cited by
- Human adults and human infants show a “perceptual magnet effect” for the prototypes of speech categories, monkeys do not
- StutterZero and StutterFormer: End-to-End Speech Conversion for Stuttering Transcription and Correction
- Beyond Lp clipping: Equalization-based Psychoacoustic Attacks against ASRs
- Dance Dance Convolution
- The Octave-Percept or Concept
- Mitigation of multi-path propagation artefacts in acoustic targets with adaptive cepstral filtering
- CLARGA: Multimodal Graph Representation Learning over Arbitrary Sets of Modalities
- Emotion Recognition from Speech
- Harmonic-Percussive Disentangled Neural Audio Codec for Bandwidth Extension
- HarmonicAttack: An Adaptive Cross-Domain Audio Watermark Removal
- Uncertainty Makes It Stable: Curiosity-Driven Quantized Mixture-of-Experts
- Towards End-to-End Prosody Transfer for Expressive Speech Synthesis with Tacotron
- Seeing Beyond Sound: Visualization and Abstraction in Audio Data Representation
- Distributed Learning of Deep Neural Networks using Independent Subnet Training
- BRIDGING THE GAP BETWEEN L2 SPEECH PERCEPTION RESEARCH AND PHONOLOGICAL THEORY
- Poolformer: Recurrent Networks with Pooling for Long-Sequence Modeling
- Kaleidoscope: An Efficient, Learnable Representation For All Structured Linear Maps
- Désentrelacement Fréquentiel Doux pour les Codecs Audio Neuronaux
- Soft Disentanglement in Frequency Bands for Neural Audio Codecs
- Video Object Segmentation-Aware Audio Generation
- Information processing pathway maps — A scalable framework for mapping cortical processing
- African elephants address one another with individually specific name-like calls
- WaveGuard: Understanding and Mitigating Audio Adversarial Examples
- Contrastive Learning of Musical Representations
- Characterizing Types of Convolution in Deep Convolutional Recurrent Neural Networks for Robust Speech Emotion Recognition
- Leveraging Multimodal Haptic Sensory Data for Robust Cutting
- Audio2Gestures: Generating Diverse Gestures from Speech Audio with Conditional Variational Autoencoders
- WaveLLDM: Design and Development of a Lightweight Latent Diffusion Model for Speech Enhancement and Restoration
- Putting a Face to the Voice: Fusing Audio and Visual Signals Across a Video to Determine Speakers
- SATEER: Subject-Aware Transformer for EEG-Based Emotion Recognition
- Perceiving Slope and Acceleration: Evidence for Variable Tempo Sampling in Pitch-Based Sonification of Functions
- nnAudio: An on-the-fly GPU Audio to Spectrogram Conversion Toolbox Using 1D Convolution Neural Networks
- MeanAudio: Fast and Faithful Text-to-Audio Generation with Mean Flows
- Road-pavement classification by artificial neural network model based on tire-pavement noise and road-surface image
- Audio Content Analysis
- Less Stress, More Privacy: Stress Detection on Anonymized Speech of Air Traffic Controllers
- Similarity in L2 Phonology: Evidence from L1 Spanish late-learners’ perception and lexical representation of English vowel contrasts
- DCTNet and PCANet for acoustic signal feature extraction
- Unsupervised Spoken Term Discovery on Untranscribed Speech
- Frequency-Weighted Training Losses for Phoneme-Level DNN-based Speech Enhancement
- AccEar: Accelerometer Acoustic Eavesdropping with Unconstrained Vocabulary
- From Spikes to Speech: NeuroVoc -- A Biologically Plausible Vocoder Framework for Auditory Perception and Cochlear Implant Simulation
- Real-time low-resource phoneme recognition on edge devices
- Mel-McNet: A Mel-Scale Framework for Online Multichannel Speech Enhancement
- 15,500 Seconds: Lean UAV Classification Using EfficientNet and Lightweight Fine-Tuning
- Face-to-Music Translation Using a Distance-Preserving Generative Adversarial Network with an Auxiliary Discriminator
- ESCAPE - Echo SCraper and ClAssifier of PErsons: A novel tool to facilitate using voice-controlled devices for research
- A scale of subjective brightness.
- On the psychophysical law.
- Optimisation of phonetic aware speech recognition through multi-objective evolutionary algorithms
- An Axiomatization of Utility Based on the Notion of Utility Differences
- Learning spectro-temporal representations of complex sounds with parameterized neural networks
- Vowel category boundaries enhance cortical and behavioral responses to speech feedback alterations. [europepmc]
- Individual identification via electrocardiogram analysis. [europepmc]
- Development of frequency discrimination at 250 Hz is similar for tone and /ba/ stimuli. [europepmc]
- Acoustic and linguistic factors affecting perceptual dissimilarity judgments of voices. [europepmc]
- Generating Natural, Intelligible Speech From Brain Activity in Motor, Premotor, and Inferior Frontal Cortices. [europepmc]
- Decoding speech from spike-based neural population recordings in secondary auditory cortex of non-human primates. [europepmc]
- Detecting Respiratory Pathologies Using Convolutional Neural Networks and Variational Autoencoders for Unbalancing Data. [europepmc]
- Emotional sounds of crowds: spectrogram-based analysis using deep learning. [europepmc]
- An Interoperable Architecture for the Internet of COVID-19 Things (IoCT) Using Open Geospatial Standards-Case Study: Workplace Reopening. [europepmc]
- Automatic Classification of Adventitious Respiratory Sounds: A (Un)Solved Problem? [europepmc]
- Machine and Deep Learning towards COVID-19 Diagnosis and Treatment: Survey, Challenges, and Future Directions. [europepmc]
- A New Acoustic-Based Pronunciation Distance Measure. [europepmc]
- Bioacoustic classification of avian calls from raw sound waveforms with an open-source deep learning architecture. [europepmc]
- Real-time synthesis of imagined speech processes from minimally invasive recordings of neural activity. [europepmc]
- Toward a Computational Neuroethology of Vocal Communication: From Bioacoustics to Neurophysiology, Emerging Tools and Future Directions. [europepmc]
- Under-resourced or overloaded? Rethinking working memory deficits in developmental language disorder. [europepmc]
- Dataset of Speech Production in intracranial.Electroencephalography. [europepmc]
- Movement variability can be modulated in speech production. [europepmc]
- Automatic Behavior Assessment from Uncontrolled Everyday Audio Recordings by Deep Learning. [europepmc]
- Development and Validation of a Single-Variable Comparison Stimulus for Matching Strained Voice Quality Using a Psychoacoustic Framework. [europepmc]
- Intelligent Fault Diagnosis of Industrial Bearings Using Transfer Learning and CNNs Pre-Trained for Audio Classification. [europepmc]
- Cross-modal correspondence enhances elevation localization in visual-to-auditory sensory substitution. [europepmc]
- Direct speech reconstruction from sensorimotor brain activity with optimized deep learning models. [europepmc]
- Online speech synthesis using a chronically implanted brain-computer interface in an individual with ALS. [europepmc]
- Consciousness Under the Spotlight: The Problem of Measuring Subjective Experience. [europepmc]
Related