Catalog footprint

What is connected

33works
25topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

33 published item(s)

preprint2024arXiv

HOI4D: A 4D Egocentric Dataset for Category-Level Human-Object Interaction

We present HOI4D, a large-scale 4D egocentric dataset with rich annotations, to catalyze the research of category-level human-object interaction. HOI4D consists of 2.4M RGB-D egocentric video frames over 4000 sequences collected by 4 participants interacting with 800 different object instances from 16 categories over 610 different indoor rooms. Frame-wise annotations for panoptic segmentation, motion segmentation, 3D hand pose, category-level object pose and hand action have also been provided, together with reconstructed object meshes and scene point clouds. With HOI4D, we establish three benchmarking tasks to promote category-level HOI from 4D visual signals including semantic segmentation of 4D dynamic point cloud sequences, category-level object pose tracking, and egocentric action segmentation with diverse interaction targets. In-depth analysis shows HOI4D poses great challenges to existing methods and produces great research opportunities.

preprint2022arXiv

A stochastic analysis approach to lattice Yang--Mills at strong coupling

We develop a new stochastic analysis approach to the lattice Yang--Mills model at strong coupling in any dimension $d>1$, with t' Hooft scaling $βN$ for the inverse coupling strength. We study their Langevin dynamics, ergodicity, functional inequalities, large $N$ limits, and mass gap. Assuming $|β| < \frac{N-2}{32(d-1)N}$ for the structure group $SO(N)$, or $|β| < \frac{1}{16(d-1)}$ for $SU(N)$, we prove the following results. The invariant measure for the corresponding Langevin dynamic is unique on the entire lattice, and the dynamic is exponentially ergodic under a Wasserstein distance. The finite volume Yang--Mills measures converge to this unique invariant measure in the infinite volume limit, for which Log-Sobolev and Poincaré inequalities hold. These functional inequalities imply that the suitably rescaled Wilson loops for the infinite volume measure has factorized correlations and converges in probability to deterministic limits in the large $N$ limit, and correlations of a large class of observables decay exponentially, namely the infinite volume measure has a strictly positive mass gap. Our method improves earlier results or simplifies the proofs, and provides some new perspectives to the study of lattice Yang--Mills model.

preprint2022arXiv

A stochastic PDE approach to large N problems in quantum field theory: a survey

In this survey we review some recent rigorous results on large N problems in quantum field theory, stochastic quantization and singular stochastic PDEs, and their mean field limit problems. In particular we discuss the O(N) linear sigma model on two and three dimensional torus. The stochastic quantization procedure leads to a coupled system of N interacting $Φ^4$ equations. In d = 2, we show uniform in N bounds for the dynamics and convergence to a mean-field singular SPDE. For large enough mass or small enough coupling, the invariant measures (i.e. the O(N) linear sigma model) converge to the massive Gaussian free field, the unique invariant measure of the mean-field dynamics, in a Wasserstein distance. We also obtain tightness for certain O(N) invariant observables as random fields in suitable Besov spaces as $N\to \infty$, along with exact descriptions of the limiting correlations. In d = 3, the estimates become more involved since the equation is more singular. We discuss in this case how to prove convergence to the massive Gaussian free field. The proofs of these results build on the recent progress of singular SPDE theory and combine many new techniques such as uniform in N estimates and dynamical mean field theory. These are based on joint papers with Scott Smith, Rongchan Zhu and Xiangchan Zhu.

preprint2022arXiv

Analysis and Optimisation of Bellman Residual Errors with Neural Function Approximation

Recent development of Deep Reinforcement Learning (DRL) has demonstrated superior performance of neural networks in solving challenging problems with large or even continuous state spaces. One specific approach is to deploy neural networks to approximate value functions by minimising the Mean Squared Bellman Error (MSBE) function. Despite great successes of DRL, development of reliable and efficient numerical algorithms to minimise the MSBE is still of great scientific interest and practical demand. Such a challenge is partially due to the underlying optimisation problem being highly non-convex or using incomplete gradient information as done in Semi-Gradient algorithms. In this work, we analyse the MSBE from a smooth optimisation perspective and develop an efficient Approximate Newton's algorithm. First, we conduct a critical point analysis of the error function and provide technical insights on optimisation and design choices for neural networks. When the existence of global minima is assumed and the objective fulfils certain conditions, suboptimal local minima can be avoided when using over-parametrised neural networks. We construct a Gauss Newton Residual Gradient algorithm based on the analysis in two variations. The first variation applies to discrete state spaces and exact learning. We confirm theoretical properties of this algorithm such as being locally quadratically convergent to a global minimum numerically. The second employs sampling and can be used in the continuous setting. We demonstrate feasibility and generalisation capabilities of the proposed algorithm empirically using continuous control problems and provide a numerical verification of our critical point analysis. We outline the difficulties of combining Semi-Gradient approaches with Hessian information. To benefit from second-order information complete derivatives of the MSBE must be considered during training.

preprint2022arXiv

Large $N$ limit of the $O(N)$ linear sigma model in 3D

In this paper we study the large N limit of the $O(N)$-invariant linear sigma model, which is a vector-valued generalization of the $Φ^4$ quantum field theory, on the three dimensional torus. We study the problem via its stochastic quantization, which yields a coupled system of N interacting SPDEs. We prove tightness of the invariant measures in the large N limit. For large enough mass or small enough coupling constant, they converge to the (massive) Gaussian free field at a rate of order $1/\sqrt N$ with respect to the Wasserstein distance. We also obtain tightness results for certain $O(N)$ invariant observables. These generalize some of the results in \cite{SSZZ20} from two dimensions to three dimensions. The proof leverages the method recently developed by \cite{GH18} and combines many new techniques such as uniform in $N$ estimates on perturbative objects as well as the solutions.

preprint2022arXiv

Learning Category-Level Generalizable Object Manipulation Policy via Generative Adversarial Self-Imitation Learning from Demonstrations

Generalizable object manipulation skills are critical for intelligent and multi-functional robots to work in real-world complex scenes. Despite the recent progress in reinforcement learning, it is still very challenging to learn a generalizable manipulation policy that can handle a category of geometrically diverse articulated objects. In this work, we tackle this category-level object manipulation policy learning problem via imitation learning in a task-agnostic manner, where we assume no handcrafted dense rewards but only a terminal reward. Given this novel and challenging generalizable policy learning problem, we identify several key issues that can fail the previous imitation learning algorithms and hinder the generalization to unseen instances. We then propose several general but critical techniques, including generative adversarial self-imitation learning from demonstrations, progressive growing of discriminator, and instance-balancing for expert buffer, that accurately pinpoints and tackles these issues and can benefit category-level manipulation policy learning regardless of the tasks. Our experiments on ManiSkill benchmarks demonstrate a remarkable improvement on all tasks and our ablation studies further validate the contribution of each proposed technique.

preprint2022arXiv

Learning from Attacks: Attacking Variational Autoencoder for Improving Image Classification

Adversarial attacks are often considered as threats to the robustness of Deep Neural Networks (DNNs). Various defending techniques have been developed to mitigate the potential negative impact of adversarial attacks against task predictions. This work analyzes adversarial attacks from a different perspective. Namely, adversarial examples contain implicit information that is useful to the predictions i.e., image classification, and treat the adversarial attacks against DNNs for data self-expression as extracted abstract representations that are capable of facilitating specific learning tasks. We propose an algorithmic framework that leverages the advantages of the DNNs for data self-expression and task-specific predictions, to improve image classification. The framework jointly learns a DNN for attacking Variational Autoencoder (VAE) networks and a DNN for classification, coined as Attacking VAE for Improve Classification (AVIC). The experiment results show that AVIC can achieve higher accuracy on standard datasets compared to the training with clean examples and the traditional adversarial training.

preprint2021arXiv

Large $N$ Limit of the $O(N)$ Linear Sigma Model via Stochastic Quantization

This article studies large $N$ limits of a coupled system of $N$ interacting $Φ^4$ equations posed over $\mathbb{T}^{d}$ for $d=2$, known as the $O(N)$ linear sigma model. Uniform in $N$ bounds on the dynamics are established, allowing us to show convergence to a mean-field singular SPDE, also proved to be globally well-posed. Moreover, we show tightness of the invariant measures in the large $N$ limit. For large enough mass, they converge to the (massive) Gaussian free field, the unique invariant measure of the mean-field dynamics, at a rate of order $1/\sqrt{N}$ with respect to the Wasserstein distance. We also consider fluctuations and obtain tightness results for certain $O(N)$ invariant observables, along with an exact description of the limiting correlations.

preprint2021arXiv

Stochastic Ricci Flow on Compact Surfaces

In this paper we introduce the stochastic Ricci flow (SRF) in two spatial dimensions. The flow is symmetric with respect to a measure induced by Liouville Conformal Field Theory. Using the theory of Dirichlet forms, we construct a weak solution to the associated equation of the area measure on a flat torus, in the full "$L^1$ regime" $σ< σ_{L^1}=2\sqrtπ$ where $σ$ is the noise strength. We also describe the main necessary modifications needed for the SRF on general compact surfaces, and list some open questions.

preprint2020arXiv

3D Scene Geometry-Aware Constraint for Camera Localization with Deep Learning

Camera localization is a fundamental and key component of autonomous driving vehicles and mobile robots to localize themselves globally for further environment perception, path planning and motion control. Recently end-to-end approaches based on convolutional neural network have been much studied to achieve or even exceed 3D-geometry based traditional methods. In this work, we propose a compact network for absolute camera pose regression. Inspired from those traditional methods, a 3D scene geometry-aware constraint is also introduced by exploiting all available information including motion, depth and image contents. We add this constraint as a regularization term to our proposed network by defining a pixel-level photometric loss and an image-level structural similarity loss. To benchmark our method, different challenging scenes including indoor and outdoor environment are tested with our proposed approach and state-of-the-arts. And the experimental results demonstrate significant performance improvement of our method on both prediction accuracy and convergence efficiency.

preprint2020arXiv

CenterMask: single shot instance segmentation with point representation

In this paper, we propose a single-shot instance segmentation method, which is simple, fast and accurate. There are two main challenges for one-stage instance segmentation: object instances differentiation and pixel-wise feature alignment. Accordingly, we decompose the instance segmentation into two parallel subtasks: Local Shape prediction that separates instances even in overlapping conditions, and Global Saliency generation that segments the whole image in a pixel-to-pixel manner. The outputs of the two branches are assembled to form the final instance masks. To realize that, the local shape information is adopted from the representation of object center points. Totally trained from scratch and without any bells and whistles, the proposed CenterMask achieves 34.5 mask AP with a speed of 12.3 fps, using a single-model with single-scale training/testing on the challenging COCO dataset. The accuracy is higher than all other one-stage instance segmentation methods except the 5 times slower TensorMask, which shows the effectiveness of CenterMask. Besides, our method can be easily embedded to other one-stage object detectors such as FCOS and performs well, showing the generalization of CenterMask.

preprint2020arXiv

Dynamic Variational Autoencoders for Visual Process Modeling

This work studies the problem of modeling visual processes by leveraging deep generative architectures for learning linear, Gaussian representations from observed sequences. We propose a joint learning framework, combining a vector autoregressive model and Variational Autoencoders. This results in an architecture that allows Variational Autoencoders to simultaneously learn a non-linear observation as well as a linear state model from sequences of frames. We validate our approach on artificial sequences and dynamic textures.

preprint2016arXiv

$\ell_1$ Regularized Gradient Temporal-Difference Learning

In this paper, we study the Temporal Difference (TD) learning with linear value function approximation. It is well known that most TD learning algorithms are unstable with linear function approximation and off-policy learning. Recent development of Gradient TD (GTD) algorithms has addressed this problem successfully. However, the success of GTD algorithms requires a set of well chosen features, which are not always available. When the number of features is huge, the GTD algorithms might face the problem of overfitting and being computationally expensive. To cope with this difficulty, regularization techniques, in particular $\ell_1$ regularization, have attracted significant attentions in developing TD learning algorithms. The present work combines the GTD algorithms with $\ell_1$ regularization. We propose a family of $\ell_1$ regularized GTD algorithms, which employ the well known soft thresholding operator. We investigate convergence properties of the proposed algorithms, and depict their performance with several numerical experiments.

preprint2016arXiv

A central limit theorem for the KPZ equation

We consider the KPZ equation in one space dimension driven by a stationary centred space-time random field, which is sufficiently integrable and mixing, but not necessarily Gaussian. We show that, in the weakly asymmetric regime, the solution to this equation considered at a suitable large scale and in a suitable reference frame converges to the Hopf-Cole solution to the KPZ equation driven by space-time Gaussian white noise. While the limiting process depends only on the integrated variance of the driving field, the diverging constants appearing in the definition of the reference frame also depend on higher order moments.

preprint2016arXiv

Monitoring and Prediction in Smart Energy Systems via Multi-timescale Nexting

Reliable prediction of system status is a highly demanded functionality of smart energy systems, which can enable users or human operators to react quickly to potential future system changes. By adopting the multi-timescale nexting method, we develop an architecture of human-in-the-loop energy control system, which is capable of casting short-term predictive information about the specific smart energy system. The developed architecture does either require a system model nor additional acquisition of (sensor) data in the existing system configuration. Our first experiments demonstrate the performance of the proposed control architecture in an electrical heating system simulation. In the second experiment, we verify the effectiveness of our developed structure in simulating a heating system in a thermal model of a building, by employing natural EnergyPlus temperature data.

preprint2016arXiv

Optimal Water Heater Control in Smart Home Environments

In this work, we develop an optimal water heater control method for a smart home environment. It is important to notice that unlike battery storage systems, energy flow in a water heater control system is not reversible. In order to increase the eigen consumption of photovoltaic energy, i.e. direct consumption of generated energy in the house, we propose to employ a dynamic programming approach to optimize heating schedules using forecasted consumption and weather data. Simulation results demonstrate the capability of our proposed system in reducing the overall energy cost while maintaining residents' comfort.

preprint2016arXiv

Reinforcement Learning in Conflicting Environments for Autonomous Vehicles

In this work, we investigate the application of Reinforcement Learning to two well known decision dilemmas, namely Newcomb's Problem and Prisoner's Dilemma. These problems are exemplary for dilemmas that autonomous agents are faced with when interacting with humans. Furthermore, we argue that a Newcomb-like formulation is more adequate in the human-machine interaction case and demonstrate empirically that the unmodified Reinforcement Learning algorithms end up with the well known maximum expected utility solution.

preprint2015arXiv

Subsampled terahertz data reconstruction based on spatio-temporal dictionary learning

In this paper, the problem of terahertz pulsed imaging and reconstruction is addressed. It is assumed that an incomplete (subsampled) three dimensional THz data set has been acquired and the aim is to recover all missing samples. A sparsity-inducing approach is proposed for this purpose. First, a simple interpolation is applied to incomplete noisy data. Then, we propose a spatio-temporal dictionary learning method to obtain an appropriate sparse representation of data based on a joint sparse recovery algorithm. Then, using the sparse coefficients and the learned dictionary, the 3D data is effectively denoised by minimizing a simple cost function. We consider two types of terahertz data to evaluate the performance of the proposed approach; THz data acquired for a model sample with clear layered structures (e.g., a T-shape plastic sheet buried in a polythene pellet), and pharmaceutical tablet data (with low spatial resolution). The achieved signal-to-noise-ratio for reconstruction of T-shape data, from only 5% observation was 19 dB. Moreover, the accuracies of obtained thickness and depth measurements for pharmaceutical tablet data after reconstruction from 10% observation were 98.8%, and 99.9%, respectively. These results, along with chemical mapping analysis, presented at the end of this paper, confirm the accuracy of the proposed method.

preprint2015arXiv

Texture Retrieval via the Scattering Transform

This work studies the problem of content-based image retrieval, specifically, texture retrieval. It focuses on feature extraction and similarity measure for texture images. Our approach employs a recently developed method, the so-called Scattering transform, for the process of feature extraction in texture retrieval. It shares a distinctive property of providing a robust representation, which is stable with respect to spatial deformations. Recent work has demonstrated its capability for texture classification, and hence as a promising candidate for the problem of texture retrieval. Moreover, we adopt a common approach of measuring the similarity of textures by comparing the subband histograms of a filterbank transform. To this end we derive a similarity measure based on the popular Bhattacharyya Kernel. Despite the popularity of describing histograms using parametrized probability density functions, such as the Generalized Gaussian Distribution, it is unfortunately not applicable for describing most of the Scattering transform subbands, due to the complex modulus performed on each one of them. In this work, we propose to use the Weibull distribution to model the Scattering subbands of descendant layers. Our numerical experiments demonstrated the effectiveness of the proposed approach, in comparison with several state of the arts.

preprint2015arXiv

The strict-weak lattice polymer

We introduce the strict-weak polymer model, and show the KPZ universality of the free energy fluctuation of this model for a certain range of parameters. Our proof relies on the observation that the discrete time geometric q-TASEP model, studied earlier by A. Borodin and I. Corwin, scales to this polymer model in the limit q->1. This allows us to exploit the exact results for geometric q-TASEP to derive a Fredholm determinant formula for the strict-weak polymer, and in turn perform rigorous asymptotic analysis to show KPZ scaling and GUE Tracy-Widom limit for the free energy fluctuations. We also derive moments formulae for the polymer partition function directly by Bethe ansatz, and identify the limit of the free energy using a stationary version of the polymer model.

preprint2014arXiv

A renormalization group method by harmonic extensions and the classical dipole gas

In this paper we develop a new renormalization group method, which is based on conditional expectations and harmonic extensions, to study functional integrals related with small perturbations of Gaussian fields. In this new method one integrates Gaussian fields inside domains at all scales conditioning on the fields outside these domains, and by variation principle solves local elliptic problems. It does not rely on an a priori decomposition of the Gaussian covariance. We apply this method to the model of classical dipole gas on the lattice, and show that the scaling limit of the generating function with smooth test functions is the generating function of the renormalized Gaussian free field.

preprint2014arXiv

Parallel Interleaver Design for a High Throughput HSPA+/LTE Multi-Standard Turbo Decoder

To meet the evolving data rate requirements of emerging wireless communication technologies, many parallel architectures have been proposed to implement high throughput turbo decoders. However, concurrent memory reading/writing in parallel turbo decoding architectures leads to severe memory conflict problem, which has become a major bottleneck for high throughput turbo decoders. In this paper, we propose a flexible and efficient VLSI architecture to solve the memory conflict problem for highly parallel turbo decoders targeting multi-standard 3G/4G wireless communication systems. To demonstrate the effectiveness of the proposed parallel interleaver architecture, we implemented an HSPA+/LTE/LTE-Advanced multi-standard turbo decoder with a 45nm CMOS technology. The implemented turbo decoder consists of 16 Radix-4 MAP decoder cores, and the chip core area is 2.43 mm^2. When clocked at 600 MHz, this turbo decoder can achieve a maximum decoding throughput of 826 Mbps in the HSPA+ mode and 1.67 Gbps in the LTE/LTE-Advanced mode, exceeding the peak data rate requirements for both standards.

preprint2014arXiv

Rapidly reconfigurable radio-frequency arbitrary waveforms synthesized on a CMOS photonic chip

Photonic methods of radio-frequency waveform generation and processing provide performance and flexibility over electronic methods due to the ultrawide bandwidth offered by the optical carriers. However, they suffer from lack of integration and slow reconfiguration speed. Here we propose an architecture of integrated photonic RF waveform generation and processing, and implement it on a silicon chip fabricated in a semiconductor manufacturing foundry. Our device can generate programmable RF bursts or continuous waveforms with only the light source, electrical drives/controls and detectors being off chip. It turns on and off an individual pulse in the RF burst within 4 nanoseconds, achieving a reconfiguration speed three orders of magnitude faster than thermal tuning. The on-chip optical delay elements offers an integrated approach to accurately manipulate individual RF waveform features without constrains set by the speed and timing jitter of electronics, and should find broad applications ranging from high-speed wireless to defense electronics.

preprint2014arXiv

The dynamical sine-Gordon model

We introduce the dynamical sine-Gordon equation in two space dimensions with parameter $β$, which is the natural dynamic associated to the usual quantum sine-Gordon model. It is shown that when $β^2 \in (0,\frac{16π}{3})$ the Wick renormalised equation is well-posed. In the regime $β^2 \in (0,4π)$, the Da Prato-Debussche method applies, while for $β^2 \in [4π,\frac{16π}{3})$, the solution theory is provided via the theory of regularity structures (Hairer 2013). We also show that this model arises naturally from a class of $2+1$-dimensional equilibrium interface fluctuation models with periodic nonlinearities. The main mathematical difficulty arises in the construction of the model for the associated regularity structure where the role of the noise is played by a non-Gaussian random distribution similar to the complex multiplicative Gaussian chaos recently analysed by Lacoin, Rhodes and Vargas (2013).

preprint2013arXiv

An Adaptive Dictionary Learning Approach for Modeling Dynamical Textures

Video representation is an important and challenging task in the computer vision community. In this paper, we assume that image frames of a moving scene can be modeled as a Linear Dynamical System. We propose a sparse coding framework, named adaptive video dictionary learning (AVDL), to model a video adaptively. The developed framework is able to capture the dynamics of a moving scene by exploring both sparse properties and the temporal correlations of consecutive video frames. The proposed method is compared with state of the art video processing methods on several benchmark data sequences, which exhibit appearance changes and heavy occlusions.

preprint2013arXiv

Exact renormalization group analysis of turbulent transport by the shear flow

The exact renormalization group (RG) method initiated by Wilson and further developed by Polchinski is used to study the shear flow model proposed by Avellaneda and Majda as a simplified model for the diffusive transport of a passive scalar by a turbulent velocity field. It is shown that this exact RG method is capable of recovering all the scaling regimes as the spectral parameters of velocity statistics vary, found by Avellaneda and Majda in their rigorous study of this model. This gives further confidence that the RG method, if implemented in the right way instead of using drastic truncations as in the Yakhot-Orszag's approximate RG scheme, does give the correct prediction for the large scale behaviors of solutions of stochastic partial differential equations (PDE). We also derive the analog of the "large eddy simulation" models when a finite amount of small scales are eliminated from the problem.

preprint2012arXiv

Averaging Complex Subspaces via a Karcher Mean Approach

We propose a conjugate gradient type optimization technique for the computation of the Karcher mean on the set of complex linear subspaces of fixed dimension, modeled by the so-called Grassmannian. The identification of the Grassmannian with Hermitian projection matrices allows an accessible introduction of the geometric concepts required for an intrinsic conjugate gradient method. In particular, proper definitions of geodesics, parallel transport, and the Riemannian gradient of the Karcher mean function are presented. We provide an efficient step-size selection for the special case of one dimensional complex subspaces and illustrate how the method can be employed for blind identification via numerical experiments.

preprint2012arXiv

Influence of the substrate and precursor on the magnetic and magneto-transport properties in magnetite films

We have investigated the magnetic and transport properties of nanoscaled Fe3O4 films obtained from Chemical Vapor Deposition (CVD) technique using [FeIIFe2III(OBut)8] and [Fe2III(OBut)6] precursors. Samples were deposited on different substrates (i.e., MgO (001), MgAl2O4 (001) and Al2O3 (0001)) with thicknesses varying from 50 to 350 nm. Atomic Force Microscopy analysis indicated a granular nature of the samples, irrespective of the synthesis conditions (precursor and deposition temperature, Tpre) and substrate. Despite the similar morphology of the films, magnetic and transport properties were found to depend on the precursor used for deposition. Using [FeIIFe2III(OBut)8] as precursor resulted in lower resistivity, higher MS and a sharper magnetization decrease at the Verwey transition (TV). The temperature dependence of resistivity was found to depend on the precursor and Tpre. We found that the transport is dominated by the density of antiferromagnetic antiphase boundaries (AF-APB's) when [FeIIFe2III(OBut)8] precursor and Tpre = 363 K are used. On the other hand, grain boundary-scattering seems to be the main mechanism when [Fe2III(OBut)6] is used. The Magnetoresistance (MR(H)) displayed an approximate linear behavior in the high field regime (H > 796 kA/m), with a maximum value at room-temperature of \sim2-3% for H = 1592 kA/m, irrespective from the transport mechanism.

preprint2012arXiv

Uniqueness Analysis of Non-Unitary Matrix Joint Diagonalization

Matrix Joint Diagonalization (MJD) is a powerful approach for solving the Blind Source Separation (BSS) problem. It relies on the construction of matrices which are diagonalized by the unknown demixing matrix. Their joint diagonalizer serves as a correct estimate of this demixing matrix only if it is uniquely determined. Thus, a critical question is under what conditions a joint diagonalizer is unique. In the present work we fully answer this question about the identifiability of MJD based BSS approaches and provide a general result on uniqueness conditions of matrix joint diagonalization. It unifies all existing results which exploit the concepts of non-circularity, non-stationarity, non-whiteness, and non-Gaussianity. As a corollary, we propose a solution for complex BSS, which can be formulated in a closed form in terms of an eigenvalue and a singular value decomposition of two matrices.

preprint2011arXiv

Blind Source Separation with Compressively Sensed Linear Mixtures

This work studies the problem of simultaneously separating and reconstructing signals from compressively sensed linear mixtures. We assume that all source signals share a common sparse representation basis. The approach combines classical Compressive Sensing (CS) theory with a linear mixing model. It allows the mixtures to be sampled independently of each other. If samples are acquired in the time domain, this means that the sensors need not be synchronized. Since Blind Source Separation (BSS) from a linear mixture is only possible up to permutation and scaling, factoring out these ambiguities leads to a minimization problem on the so-called oblique manifold. We develop a geometric conjugate subgradient method that scales to large systems for solving the problem. Numerical results demonstrate the promising performance of the proposed algorithm compared to several state of the art methods.