Source author record

Susama Agarwala

Susama Agarwala appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

12works
7topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

12 published item(s)

preprint2022arXiv

Eigenvalues of Autoencoders in Training and at Initialization

In this paper, we investigate the evolution of autoencoders near their initialization. In particular, we study the distribution of the eigenvalues of the Jacobian matrices of autoencoders early in the training process, training on the MNIST data set. We find that autoencoders that have not been trained have eigenvalue distributions that are qualitatively different from those which have been trained for a long time ($>$100 epochs). Additionally, we find that even at early epochs, these eigenvalue distributions rapidly become qualitatively similar to those of the fully trained autoencoders. We also compare the eigenvalues at initialization to pertinent theoretical work on the eigenvalues of random matrices and the products of such matrices.

preprint2022arXiv

Geometric instability of out of distribution data across autoencoder architecture

We study the map learned by a family of autoencoders trained on MNIST, and evaluated on ten different data sets created by the random selection of pixel values according to ten different distributions. Specifically, we study the eigenvalues of the Jacobians defined by the weight matrices of the autoencoder at each training and evaluation point. For high enough latent dimension, we find that each autoencoder reconstructs all the evaluation data sets as similar \emph{generalized characters}, but that this reconstructed \emph{generalized character} changes across autoencoder. Eigenvalue analysis shows that even when the reconstructed image appears to be an MNIST character for all out of distribution data sets, not all have latent representations that are close to the latent representation of MNIST characters. All told, the eigenvalue analysis demonstrated a great deal of geometric instability of the autoencoder both as a function on out of distribution inputs, and across architectures on the same set of inputs.

preprint2022arXiv

Structural Similarity for Improved Transfer in Reinforcement Learning

Transfer learning is an increasingly common approach for developing performant RL agents. However, it is not well understood how to define the relationship between the source and target tasks, and how this relationship contributes to successful transfer. We present an algorithm called Structural Similarity for Two MDPS, or SS2, that calculates a state similarity measure for states in two finite MDPs based on previously developed bisimulation metrics, and show that the measure satisfies properties of a distance metric. Then, through empirical results with GridWorld navigation tasks, we provide evidence that the distance measure can be used to improve transfer performance for Q-Learning agents over previous implementations.

preprint2021arXiv

Combinatorics of the geometry of Wilson loop diagrams I: equivalence classes via matroids and polytopes

Wilson loop diagrams are an important tool in studying scattering amplitudes of SYM $N=4$ theory and are known by previous work to be associated to positroids. We characterize the conditions under which two Wilson loop diagrams give the same positroid, prove that an important subclass of subdiagrams (exact subdiagrams) correspond to uniform matroids, and enumerate the number of different Wilson loop diagrams that correspond to each positroid cell. We also give a correspondence between those positroids which can arise from Wilson loop diagrams and directions in associahedra.

preprint2021arXiv

Combinatorics of the geometry of Wilson loop diagrams II: Grassmann necklaces, dimensions, and denominators

Wilson loop diagrams are an important tool in studying scattering amplitudes of SYM $N=4$ theory and are known by previous work to be associated to positroids. In this paper we study the structure of the associated positroids, as well as the structure of the denominator of the integrand defined by each diagram. We give an algorithm to derive the Grassmann necklace of the associated positroid directly from the Wilson loop diagram, and a recursive proof that the dimension of these cells is thrice the number of propagators in the diagram. We also show that the ideal generated by the denominator in the integrand is the radical of the ideal generated by the product of Grassmann necklace minors.

preprint2015arXiv

Generalizing the Connes Moscovici Hopf algebra to contain all rooted trees

This paper defines a generalization of the Connes-Moscovici Hopf algebra, $\mathcal{H}(1)$ that contains the entire Hopf algebra of rooted trees. A relationship between the former, a much studied object in non-commutative geometry, and the later, a much studied object in perturbative Quantum Field Theory, has been established by Connes and Kreimer. The results of this paper open the door to study the cohomology of the Hopf algebra of rooted trees.

preprint2015arXiv

The geometric $β$-function in curved space-time under operator regularization

In this paper, I compare the generators of the renormalization group flow, or the geometric $β$-functions for dimensional regularization and operator regularization. I then extend the analysis to show that the geometric $β$-function for a scalar field theory on a closed compact Riemannian manifold is defined on the entire manifold. I then extend the analysis to find the generator of the renormalization group flow for a conformal scalar-field theories on the same manifolds. The geometric $β$-function in this case is not defined.

preprint2015arXiv

Wilson Loop diagrams and Positroids

In this paper, we study a new application of the positive Grassmanian to Wilson loop diagrams (or MHV diagrams) for scattering amplitudes in N=4 Super Yang-Mill theory ($N=4$ SYM). There has been much interest in studying this theory via the positive Grassmanians using BCFW recursion. This is the first attempt to study MHV diagrams for planar Wilson loop calculations (or planar amplitudes) in terms of positive Grassmannians. We codify Wilson loop diagrams completely in terms of matroids. This allows us to apply the combinatorial tools in matroid theory used to identify positroids, (non-negative Grassmannians), to Wilson loop diagrams. In doing so, we find that certain non-planar Wilson loop diagrams define positive Grassmannians. While non-planar diagrams do not have physical meaning, this finding suggests that they may have value as an algebraic tool, and deserve further investigation.

preprint2012arXiv

Dynkin operators, renormalization and the geometric $β$ function

In this paper, I show a close connection between renormalization and a generalization of the Dynkin operator in terms of logarithmic derivations. The geometric $β$ function, which describes the dependence of a Quantum Field Theory on an energy scale defines is defined by a complete vector field on a Lie group $G$ defined by a QFT. It also defines a generalized Dynkin operator.

preprint2012arXiv

Geometrically relating momentum cut-off and dimensional regularization

The $β$ function for a scalar field theory describes the dependence of the coupling constant on the renormalization mass scale. This dependence is affected by the choice of regularization scheme. I explicitly relate the $β$-functions of momentum cut-off regularization and dimensional regularization on scalar field theories by a gauge transformation using the Hopf algebras of the Feynman diagrams of the theories.