Source author record

Han Chen

Han Chen appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

7works
13topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

7 published item(s)

preprint2024arXiv

Evaluating and Personalizing User-Perceived Quality of Text-to-Speech Voices for Delivering Mindfulness Meditation with Different Physical Embodiments

Mindfulness-based therapies have been shown to be effective in improving mental health, and technology-based methods have the potential to expand the accessibility of these therapies. To enable real-time personalized content generation for mindfulness practice in these methods, high-quality computer-synthesized text-to-speech (TTS) voices are needed to provide verbal guidance and respond to user performance and preferences. However, the user-perceived quality of state-of-the-art TTS voices has not yet been evaluated for administering mindfulness meditation, which requires emotional expressiveness. In addition, work has not yet been done to study the effect of physical embodiment and personalization on the user-perceived quality of TTS voices for mindfulness. To that end, we designed a two-phase human subject study. In Phase 1, an online Mechanical Turk between-subject study (N=471) evaluated 3 (feminine, masculine, child-like) state-of-the-art TTS voices with 2 (feminine, masculine) human therapists' voices in 3 different physical embodiment settings (no agent, conversational agent, socially assistive robot) with remote participants. Building on findings from Phase 1, in Phase 2, an in-person within-subject study (N=94), we used a novel framework we developed for personalizing TTS voices based on user preferences, and evaluated user-perceived quality compared to best-rated non-personalized voices from Phase 1. We found that the best-rated human voice was perceived better than all TTS voices; the emotional expressiveness and naturalness of TTS voices were poorly rated, while users were satisfied with the clarity of TTS voices. Surprisingly, by allowing users to fine-tune TTS voice features, the user-personalized TTS voices could perform almost as well as human voices, suggesting user personalization could be a simple and very effective tool to improve user-perceived quality of TTS voice.

preprint2021arXiv

Few-shot Learning for CT Scan based COVID-19 Diagnosis

Coronavirus disease 2019 (COVID-19) is a Public Health Emergency of International Concern infecting more than 40 million people across 188 countries and territories. Chest computed tomography (CT) imaging technique benefits from its high diagnostic accuracy and robustness, it has become an indispensable way for COVID-19 mass testing. Recently, deep learning approaches have become an effective tool for automatic screening of medical images, and it is also being considered for COVID-19 diagnosis. However, the high infection risk involved with COVID-19 leads to relative sparseness of collected labeled data limiting the performance of such methodologies. Moreover, accurately labeling CT images require expertise of radiologists making the process expensive and time-consuming. In order to tackle the above issues, we propose a supervised domain adaption based COVID-19 CT diagnostic method which can perform effectively when only a small samples of labeled CT scans are available. To compensate for the sparseness of labeled data, the proposed method utilizes a large amount of synthetic COVID-19 CT images and adjusts the networks from the source domain (synthetic data) to the target domain (real data) with a cross-domain training mechanism. Experimental results show that the proposed method achieves state-of-the-art performance on few-shot COVID-19 CT imaging based diagnostic tasks.

preprint2016arXiv

Adjoint-based Gradient Estimation Using the Space-time Solutions of Unknown Conservation Law Simulations

Many control applications can be formulated as optimization constrained by conservation laws. Such optimization can be efficiently solved by gradient-based methods, where the gradient is obtained through the adjoint method. Traditionally, the adjoint method has not been able to be implemented in "gray-box" conservation law simulations. In gray-box simulations, the analytical and numerical form of the conservation law is unknown, but the space-time solution of relevant flow quantities is available. Without the adjoint gradient, optimization can be challenging for problems with many control variables. However, much information about the gray-box simulation is contained in its space-time solution, which motivates us to estimate the adjoint gradient by leveraging the space-time solution. This article considers a type of gray-box simulations where the flux function is partially unknown. A method is introduced to estimate the adjoint gradient at a cost independent of the number of control variables. The method firstly infers a conservation law, named the twin model, from the space-time solution, and then applies the adjoint method to the inferred twin model to estimate the gradient. The method is demonstrated to achieve good gradient estimation accuracies in several numerical examples. The main contributions of this paper are: a twin model method that enables the adjoint gradient computation for gray-box conservation law simulations; and an adaptive basis construction scheme that fully exploits the information of gray-box solutions.

preprint2016arXiv

Non-Convex Projected Gradient Descent for Generalized Low-Rank Tensor Regression

In this paper, we consider the problem of learning high-dimensional tensor regression problems with low-rank structure. One of the core challenges associated with learning high-dimensional models is computation since the underlying optimization problems are often non-convex. While convex relaxations could lead to polynomial-time algorithms they are often slow in practice. On the other hand, limited theoretical guarantees exist for non-convex methods. In this paper we provide a general framework that provides theoretical guarantees for learning high-dimensional tensor regression models under different low-rank structural assumptions using the projected gradient descent algorithm applied to a potentially non-convex constraint set $Θ$ in terms of its \emph{localized Gaussian width}. We juxtapose our theoretical results for non-convex projected gradient descent algorithms with previous results on regularized convex approaches. The two main differences between the convex and non-convex approach are: (i) from a computational perspective whether the non-convex projection operator is computable and whether the projection has desirable contraction properties and (ii) from a statistical upper bound perspective, the non-convex approach has a superior rate for a number of examples. We provide three concrete examples of low-dimensional structure which address these issues and explain the pros and cons for the non-convex and convex approaches. We supplement our theoretical results with simulations which show that, under several common settings of generalized low rank tensor regression, the projected gradient descent approach is superior both in terms of statistical error and run-time provided the step-sizes of the projected descent algorithm are suitably chosen.

preprint2014arXiv

The degenerative evolution from multicellularity to unicellularity during cancer

Theoretical reasoning suggests that human cancer may result from knocking down the genetic constraints evolved for maintenance of the metazoan multicellularity, which, however, requires a critical test. Using xenograft-based experimental evolution we characterized for the first time the full life history from initiation to metastasis of a tumor at the genomic and transcriptomic levels, and observed metastasis-driving positive selection for generally loss-of-function mutations on a set of multicellularity-related genes, which is further supported by large-scale exome data of clinical tumor samples. Subsequent expression analysis revealed mainly expression down-regulation of multicellularity-related genes, which form an evolving expression profile approaching that of embryonic stem cells, the cell type with the most characteristics of unicellular life. The theoretical conjecture predicts that genes born at the emergence of metazoan multicellularity tend to be cancer drivers, which we validated using a rigorous phylostratigraphy analysis on the birth rate of genes annotated by Cancer Gene Census. Also, the number of loss-of-function tumor suppressors often predominates over activated oncogenes in a typical tumor of human patients. These data collectively suggest that, different from typical organismal evolution in which gain of new genes is the mainstream, cancer represents a loss-of-function-driven degenerative evolution back to the unicellular ground state. This cancer evolution model may explain the enormous tumoral genetic heterogeneity in the clinic, underlie how distant-organ metastases originate in primary tumors despite distinct environmental requirements, and hold implications for designing effective cancer therapy.

preprint2013arXiv

Ground state of the double-well condensate for quantum metrology

We discuss theoretically the ground state of a Bose-Einstein condensate with attractive atom-atom interactions in a double-well trap as a starting point of Heisenberg-limited atom interferometry. The dimensionless parameter governing the quality of the ground state for this purpose is identified. The near-degeneracy between the ground state and the first excited state severely curtails the prospects of the thermally prepared ground state in quantum metrology.

preprint2012arXiv

Diffraction of transmission light through triangular apertures in array of retro-reflective micro-prisms

The array of micro-prisms was described by model of multi-period blazed gratings consisting of triangular apertures. The origins of hexagram-shaped diffraction patterns were interpreted based on multiple-beam interference and diffraction array theorem. The relation between zonal /line ghost fringes and imperfectly fabricated array structures was analyzed. Geometrical performances (e.g., the dihedral angle of micro-prism) were tested by measuring the features of diffraction patterns of samples from three retro-reflective sheeting manufacturers.