Track GCI-anchored voice quality measures as a time series
trk_covarep_vq_gci.RdComputes five glottal voice quality parameters (NAQ, QOQ, H1H2, HRF, PSP) at
each glottal closure instant and interpolates them onto a regular 10 ms frame
grid. GCIs are detected automatically via SEDREAMS when not supplied. Use this
function for time-varying voice quality trajectories; for scalar summaries see
lst_covarep_vq().
Usage
trk_covarep_vq_gci(listOfFiles, gci_times = NULL, beginTime = 0, endTime = 0, toFile = FALSE, explicitExt = "vqg", outputDirectory = NULL, verbose = TRUE)Arguments
- listOfFiles
Character vector of audio file paths. Any format supported by av is accepted; non-native inputs are transcoded automatically.
- gci_times
Optional numeric vector (or list of vectors for multiple files) of GCI times in seconds. If
NULL(default), GCIs are computed internally via SEDREAMS.- toFile
Logical. If
TRUE, write SSFF output files and return the paths written invisibly. IfFALSE, return anAsspDataObj. DefaultFALSE.- explicitExt
Character. Output file extension. Default
"vqg".- beginTime
Start time for the extracted portion in seconds. Default: NULL (beginning of signal). Note: uses
beginTime/endTime(seconds) matching DSP function conventions, unlikeread_audio()which usesbegin/end.- endTime
The end time of the section of the sound files that should be analysed (in seconds). Use 0 for end of file.
- outputDirectory
The directory where the slice file should be stored. If not defiled (NULL), the sparse slice file will placed in the same folder as the media file.
- verbose
Logical. Show a progress bar (sequential path) or a progress-aware parallel apply (
pbapply/pbmcapply, if installed).
Value
If toFile = FALSE: an AsspDataObj with tracks:
naqFLOAT, Normalized Amplitude Quotient, 0–1, n_frames × 1. Low = creaky, high = breathy (typical range 0.5–0.9).
qoqFLOAT, Quasi-Open Quotient, 0–1, n_frames × 1. Fraction of pitch period in open phase (typical range 0.3–0.7).
h1h2FLOAT, H1–H2 spectral difference in dB, n_frames × 1. Positive = brighter/breathier, negative = darker/creakier.
hrfFLOAT, Harmonic Richness Factor, dimensionless, n_frames × 1. Higher values indicate more periodic phonation (typical range 0.5–0.95).
pspFLOAT, Parabolic Spectral Peak, dimensionless, n_frames × 1. Higher values indicate smoother spectral envelope (typical range 0.5–0.95).
Frame rate: 100 Hz (10 ms grid).
If toFile = TRUE: character vector of output file paths, returned invisibly.
Details
GCI detection uses SEDREAMS on the LPC residual. Glottal flow is obtained via IAIF. Parameter values are linearly interpolated to the 10 ms grid; frames with no voiced GCIs nearby are set to zero.
Examples
if (FALSE) { # \dontrun{
trk_covarep_vq_gci(
system.file("samples", "sustained", "a1.wav", package = "superassp"),
toFile = FALSE
)
} # }