Source author record

Jarmo Malinen

Jarmo Malinen appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

10works
10topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

10 published item(s)

preprint2020arXiv

Efficient solution of symmetric eigenvalue problems from families of coupled systems

Efficient solution of the lowest eigenmodes is studied for a family of related eigenvalue problems with common $2\times 2$ block structure. It is assumed that the upper diagonal block varies between different versions while the lower diagonal block and the range of the coupling blocks remains unchanged. Such block structure naturally arises when studying the effect of a subsystem to the eigenmodes of the full system. The proposed method is based on interpolation of the resolvent function after some of its singularities have been removed by a spectral projection. Singular value decomposition can be used to further reduce the dimension of the computational problem. Error analysis of the method indicates exponential convergence with respect to the number of interpolation points. Theoretical results are illustrated by two numerical examples related to finite element discretisation of the Laplace operator.

preprint2016arXiv

Post-processing speech recordings during MRI

We discuss post-processing of speech that has been recorded during Magnetic Resonance Imaging (MRI) of the vocal tract. Such speech recordings are contaminated by high levels of acoustic noise from the MRI scanner. Also, the frequency response of the sound signal path is not flat as a result of severe restrictions on recording instrumentation due to MRI technology. The post-processing algorithm for noise reduction is based on adaptive spectral filtering. The speech material consists of samples of prolonged vowel productions that are used for validation of the post-processing algorithm. The comparison data is recorded in anechoic chamber from the same test subject. Formant analysis is carried out for the post-processed speech and the comparison data. Artificially noise-contaminated vowel samples are used for validation experiments to determine performance of the algorithm where using true data would be difficult. The properties of recording instrumentation or the post-processing algorithm do not explain the consistent frequency dependent discrepancy between formant data from experiments during MRI and in anechoic chamber. It is shown that the discrepancy is statistically significant, in particular, where it is largest at 1 kHz and 2 kHz. The reflecting surfaces of the MRI head and neck coil are suspected to change the speech acoustics which results in "external formants" at these frequencies. However, the role of test subject adaptation to noise and constrained space acoustics during an MRI examination cannot be ruled out.

preprint2015arXiv

Modal locking between vocal fold and vocal tract oscillations: Experiments and statistical analysis

The human vocal folds are known to interact with the vocal tract acoustics during voiced speech production; namely a nonlinear source-filter coupling has been observed both by using models and in \emph{in vivo} phonation. These phenomena are approached from two directions in this article. We first present a computational dynamical model of the speech apparatus that contains an explicit filter-source feedback mechanism from the vocal tract acoustics back to the vocal folds oscillations. The model was used to simulate vocal pitch glideswhere the trajectory was forced to cross the lowest vocal tract resonance, i.e., the lowest formant $F_1$. Similar patterns produced by human participants were then studied. Both the simulations and the experimental results reveal an effect when the glides cross the first formant (as may happen in \textipa{[i]}). Conversely, this effect is not observed if there is no formant within the glide range (as is the case in \textipa{[\textscripta]}). The experiments show smaller effect compared to the simulations, pointing to an active compensation mechanism.

preprint2015arXiv

Spectral Study of the Vocal Tract in Vowel Synthesis: A Comparison between 1D and 3D Acoustic Analysis

A state-of-the-art 1D acoustic synthesizer has been previously developed, and coupled to speaker-specific biomechanical models of oropharynx in ArtiSynth. As expected, the formant frequencies of the synthesized vowel sounds were shown to be different from those of the recorded audio. Such discrepancy was hypothesized to be due to the simplified geometry of the vocal tract model as well as the one dimensional implementation of Navier-Stokes equations. In this paper, we calculate Helmholtz resonances of our vocal tract geometries using 3D finite element method (FEM), and compare them with the formant frequencies obtained from the 1D method and audio. We hope such comparison helps with clarifying the limitations of our current models and/or speech synthesizer.

preprint2014arXiv

A posteriori error estimates for Webster's equation in wave propagation

We consider a generalised Webster's equation for describing wave propagation in curved tubular structures such as variable diameter acoustic wave guides. Webster's equation in generalised form has been rigorously derived in a previous article starting from the wave equation, and it approximates cross-sectional averages of the propagating wave. Here, the approximation error is estimated by an a posteriori technique.

preprint2013arXiv

Acoustic wave guides as infinite-dimensional dynamical systems

We prove the unique solvability, passivity/conservativity and some regularity results of two mathematical models for acoustic wave propagation in curved, variable diameter tubular structures of finite length. The first of the models is the generalised Webster's model that includes dissipation and curvature of the 1D waveguide. The second model is the scattering passive, boundary controlled wave equation on 3D waveguides. The two models are treated in an unified fashion so that the results on the wave equation reduce to the corresponding results of approximating Webster's model at the limit of vanishing waveguide intersection.

preprint2013arXiv

Measurement of acoustic and anatomic changes in oral and maxillofacial surgery patients

We describe an arrangement for simultaneous recording of speech and geometry of vocal tract in patients undergoing surgery involving this area. Experimental design is considered from an articulatory phonetic point of view. The speech and noise signals are recorded with an acoustic-electrical arrangement. The vocal tract is simultaneously imaged with MRI. A MATLAB-based system controls the timing of speech recording and MR image acquisition. The speech signals are cleaned from acoustic MRI noise by a non-linear signal processing algorithm. Finally, a vowel data set from pilot experiments is compared with validation data from anechoic chamber as well as with Helmholtz resonances of the vocal tract volume.

preprint2012arXiv

How far are vowel formants from computed vocal tract resonances?

We compare numerically computed resonances of the human vocal tract with formants that have been extracted from speech during vowel pronunciation. The geometry of the vocal tract has been obtained by MRI from a male subject, and the corresponding speech has been recorded simultaneously. The resonances are computed by solving the Helmholtz partial differential equation with the Finite Element Method (FEM). Despite a rudimentary exterior space acoustics model, i.e., the Dirichlet boundary condition at the mouth opening, the computed resonance structure differs from the measured formant structure by $\approx$ 0.7 semitones for [i] and [u] having small mouth opening area, and by $\approx$ 3 semitones for vowels [a] and [ae] that have a larger mouth opening. The contribution of the possibly open velar port has not been taken into considaration at all which adds the discrepancy for [a] in the present data set. We conclude that by improving the exterior space model and properly treating the velar port opening, it is possible to computationally attain four lowest vowel formants with an error less than a semitone. The corresponding wave equation model on MRI-produced vocal tract geometries is expected to have a comparable accuracy.

preprint2012arXiv

Microspectral analysis of quasinilpotent operators

We develop a microspectral theory for quasinilpotent linear operators $Q$ (i.e., those with $σ(Q) = \{0}$) in a Banach space. When such $Q$ is not compact, normal, or nilpotent, the classical spectral theory gives little information, and a somewhat deeper structure can be recovered from microspectral sets in $\C$. Such sets describe, e.g., semigroup generation, resolvent properties, power boundedness as well as Tauberian properties associated to $zQ$ for $z \in \C$.