Source author record

Igor Halperin

Igor Halperin appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

11works
13topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

11 published item(s)

preprint2022arXiv

Combining Reinforcement Learning and Inverse Reinforcement Learning for Asset Allocation Recommendations

We suggest a simple practical method to combine the human and artificial intelligence to both learn best investment practices of fund managers, and provide recommendations to improve them. Our approach is based on a combination of Inverse Reinforcement Learning (IRL) and RL. First, the IRL component learns the intent of fund managers as suggested by their trading history, and recovers their implied reward function. At the second step, this reward function is used by a direct RL algorithm to optimize asset allocation decisions. We show that our method is able to improve over the performance of individual fund managers.

preprint2022arXiv

Phases of MANES: Multi-Asset Non-Equilibrium Skew Model of a Strongly Non-Linear Market with Phase Transitions

This paper presents an analytically tractable and practically-oriented model of non-linear dynamics of a multi-asset market in the limit of a large number of assets. The asset price dynamics are driven by money flows into the market from external investors, and their price impact. This leads to a model of a market as an ensemble of interacting non-linear oscillators with the Langevin dynamics. In a homogeneous portfolio approximation, the mean field treatment of the resulting Langevin dynamics produces the McKean-Vlasov equation as a dynamic equation for market returns. Due to the strong non-linearity of the McKean-Vlasov equation, the resulting dynamics give rise to ergodicity breaking and first- or second-order phase transitions under variations of model parameters. Using a tractable potential of the Non-Equilibrium Skew (NES) model previously suggested by the author for a single-stock case, the new Multi-Asset NES (MANES) model enables an analytically tractable framework for a multi-asset market. The equilibrium expected market log-return is obtained as a self-consistent mean field of the McKean-Vlasov equation, and derived in closed form in terms of parameters that are inferred from market prices of S&P 500 index options. The model is able to accurately fit the market data for either a benign or distressed market environments, while using only a single volatility parameter.

preprint2020arXiv

G-Learner and GIRL: Goal Based Wealth Management with Reinforcement Learning

We present a reinforcement learning approach to goal based wealth management problems such as optimization of retirement plans or target dated funds. In such problems, an investor seeks to achieve a financial goal by making periodic investments in the portfolio while being employed, and periodically draws from the account when in retirement, in addition to the ability to re-balance the portfolio by selling and buying different assets (e.g. stocks). Instead of relying on a utility of consumption, we present G-Learner: a reinforcement learning algorithm that operates with explicitly defined one-step rewards, does not assume a data generation process, and is suitable for noisy data. Our approach is based on G-learning - a probabilistic extension of the Q-learning method of reinforcement learning. In this paper, we demonstrate how G-learning, when applied to a quadratic reward and Gaussian reference policy, gives an entropy-regulated Linear Quadratic Regulator (LQR). This critical insight provides a novel and computationally tractable tool for wealth management tasks which scales to high dimensional portfolios. In addition to the solution of the direct problem of G-learning, we also present a new algorithm, GIRL, that extends our goal-based G-learning approach to the setting of Inverse Reinforcement Learning (IRL) where rewards collected by the agent are not observed, and should instead be inferred. We demonstrate that GIRL can successfully learn the reward parameters of a G-Learner agent and thus imitate its behavior. Finally, we discuss potential applications of the G-Learner and GIRL algorithms for wealth management and robo-advising.

preprint2020arXiv

The Inverted Parabola World of Classical Quantitative Finance: Non-Equilibrium and Non-Perturbative Finance Perspective

Classical quantitative finance models such as the Geometric Brownian Motion or its later extensions such as local or stochastic volatility models do not make sense when seen from a physics-based perspective, as they are all equivalent to a negative mass oscillator with a noise. This paper presents an alternative formulation based on insights from physics.

preprint2013arXiv

USLV: Unspanned Stochastic Local Volatility Model

We propose a new framework for modeling stochastic local volatility, with potential applications to modeling derivatives on interest rates, commodities, credit, equity, FX etc., as well as hybrid derivatives. Our model extends the linearity-generating unspanned volatility term structure model by Carr et al. (2011) by adding a local volatility layer to it. We outline efficient numerical schemes for pricing derivatives in this framework for a particular four-factor specification (two "curve" factors plus two "volatility" factors). We show that the dynamics of such a system can be approximated by a Markov chain on a two-dimensional space (Z_t,Y_t), where coordinates Z_t and Y_t are given by direct (Kroneker) products of values of pairs of curve and volatility factors, respectively. The resulting Markov chain dynamics on such partly "folded" state space enables fast pricing by the standard backward induction. Using a nonparametric specification of the Markov chain generator, one can accurately match arbitrary sets of vanilla option quotes with different strikes and maturities. Furthermore, we consider an alternative formulation of the model in terms of an implied time change process. The latter is specified nonparametrically, again enabling accurate calibration to arbitrary sets of vanilla option quotes.

preprint2012arXiv

Pricing options on illiquid assets with liquid proxies using utility indifference and dynamic-static hedging

This work addresses the problem of optimal pricing and hedging of a European option on an illiquid asset Z using two proxies: a liquid asset S and a liquid European option on another liquid asset Y. We assume that the S-hedge is dynamic while the Y-hedge is static. Using the indifference pricing approach we derive a HJB equation for the value function, and solve it analytically (in quadratures) using an asymptotic expansion around the limit of the perfect correlation between assets Y and Z. While in this paper we apply our framework to an incomplete market version of the credit-equity Merton's model, the same approach can be used for other asset classes (equity, commodity, FX, etc.), e.g. for pricing and hedging options with illiquid strikes or illiquid exotic options.

preprint1998arXiv

Axion potential, topological defects and CP-odd bubbles in QCD

It follows on general grounds that the theta dependence in QCD is more complicated than suggested by the large N_c approach or instanton arguments. Generically, the vacuum energy E_{vac}(θ) is a multi-valued function of θadmitting the existence of metastable states. We discuss decays of such metastable vacua in the theory with and without the axion, and point out the potential relevance of this and related phenomena for constraining a dark matter axion. Based on the analysis of the axion potential, an idea for a new axion search experiment at RHIC is suggested. It is noted that the false vacuum decay proceeds with maximal violation of CP, even if θ= 0. We further speculate that the famous Sakharov criteria for baryogenesis could be satisfied at the QCD scale.

preprint1998arXiv

Can Theta/N Dependence for Gluodynamics be Compatible with 2 pi Periodicity in Theta ?

In a number of field theoretical models the vacuum angle θenters physics in the combination θ/N, where N stands generically for the number of colors or flavors, in an apparent contradiction with the expected 2 πperiodicity in θ. We argue that a resolution of this puzzle is related to the existence of a number of different θdependent sectors in a finite volume formulation, which can not be seen in the naive thermodynamic limit V -> \infty. It is shown that, when the limit V -> \infty is properly defined, physics is always 2 πperiodic in θfor any integer, and even rational, values of N, with vacuum doubling at certain values of θ. We demonstrate this phenomenon in both the multi-flavor Schwinger model with the bosonization technique, and four-dimensional gluodynamics with the effective Lagrangian method. The proposed mechanism works for an arbitrary gauge group.

preprint1997arXiv

B -> K eta' decay as unique probe of eta' meson

A theory of the B -> K η' decay is proposed. It is based on the Cabbibo favored b -> \bar{c} c s process followed by a direct materialization of the \bar{c} c pair into the η'. This mechanism works due to a non-valence Zweig rule violating c-quark component of the η', which is unique to its very special nature. This non-perturbative "intrinsic charm" content of the η' is evaluated using the Operator Product Expansion and QCD low energy theorems. Our results are consistent with an unexpectedly large Br(B -> K η') \simeq 7.8 \cdot 10^{-5} recently announced by CLEO.

preprint1995arXiv

Quantum KAM Technique and Yang-Mills Quantum Mechanics

We study a quantum analogue of the iterative perturbation theory by Kolmogorov used in the proof of the Kolmogorov-Arnold-Moser (KAM) theorem. The method is based on sequent canonical transformations with a "running" coupling constant $ \lm,\lm^{2},\lm^{4} $ etc. The proposed scheme, as its classical predecessor, is "superconvergent" in the sense that after the n-th step, a theory is solved to the accuracy of order $ \lm^{2^{n-1}} $. It is shown that the Kolmogorov technique corresponds to an infinite resummation of the usual perturbative series. The corresponding expansion is convergent for the quantum anharmonic oscillator due to the fact that it turns out to be identical to the Pade series. The method is easily generalizable to many-dimensional cases. The Kolmogorov technique is further applied to a non-perturbative treatment of Yang-Mills quantum mechanics. A controllable expansion for the wave function near the origin is constructed. For large fields, we build an asymptotic adiabatic expansion in inverse powers of the field. This asymptotic solution contains arbitrary constants which are not fixed by the boundary conditions at infinity. To find them, we approximately match the two expansions in an intermediate region. We also discuss some analogies between this problem and the method of QCD sum rules.