Source author record

Nikhil Anand

Nikhil Anand appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

7works
11topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

7 published item(s)

preprint2026arXiv

ContextFocus: Activation Steering for Contextual Faithfulness in Large Language Models

Large Language Models (LLMs) encode vast amounts of parametric knowledge during pre-training. As world knowledge evolves, effective deployment increasingly depends on their ability to faithfully follow externally retrieved context. When such evidence conflicts with the model's internal knowledge, LLMs often default to memorized facts, producing unfaithful outputs. In this work, we introduce ContextFocus, a lightweight activation steering approach that improves context faithfulness in such knowledge-conflict settings while preserving fluency and efficiency. Unlike prior approaches, our solution requires no model finetuning and incurs minimal inference-time overhead, making it highly efficient. We evaluate ContextFocus on the ConFiQA benchmark, comparing it against strong baselines including ContextDPO, COIECD, and prompting-based methods. Furthermore, we show that our method is complementary to prompting strategies and remains effective on larger models. Extensive experiments show that ContextFocus significantly improves contextual-faithfulness. Our results highlight the effectiveness, robustness, and efficiency of ContextFocus in improving contextual-faithfulness of LLM outputs.

preprint2024arXiv

Dataset Difficulty and the Role of Inductive Bias

Motivated by the goals of dataset pruning and defect identification, a growing body of methods have been developed to score individual examples within a dataset. These methods, which we call "example difficulty scores", are typically used to rank or categorize examples, but the consistency of rankings between different training runs, scoring methods, and model architectures is generally unknown. To determine how example rankings vary due to these random and controlled effects, we systematically compare different formulations of scores over a range of runs and model architectures. We find that scores largely share the following traits: they are noisy over individual runs of a model, strongly correlated with a single notion of difficulty, and reveal examples that range from being highly sensitive to insensitive to the inductive biases of certain model architectures. Drawing from statistical genetics, we develop a simple method for fingerprinting model architectures using a few sensitive examples. These findings guide practitioners in maximizing the consistency of their scores (e.g. by choosing appropriate scoring methods, number of runs, and subsets of examples), and establishes comprehensive baselines for evaluating scores in the future.

preprint2021arXiv

Nonperturbative dynamics of (2+1)d $ϕ^4$-theory from Hamiltonian truncation

We use Lightcone Conformal Truncation (LCT) -- a version of Hamiltonian truncation -- to study the nonperturbative, real-time dynamics of $ϕ^4$-theory in 2+1 dimensions. This theory has UV divergences that need to be regulated. We review how, in a Hamiltonian framework with a total energy cutoff, renormalization is necessarily \emph{state-dependent}, and UV sensitivity cannot be canceled with standard local operator counterterms. To overcome this problem, we present a prescription for constructing the appropriate state-dependent counterterms for (2+1)d $ϕ^4$-theory in lightcone quantization. We then use LCT with this counterterm prescription to study $ϕ^4$-theory, focusing on the $\mathbb{Z}_2$ symmetry-preserving phase. Specifically, we compute the spectrum as a function of the coupling and demonstrate the closing of the mass gap at a (scheme-dependent) critical coupling. We also compute Lorentz-invariant two-point functions, both at generic strong coupling and near the critical point, where we demonstrate IR universality and the vanishing of the trace of the stress tensor.

preprint2020arXiv

Introduction to Lightcone Conformal Truncation: QFT Dynamics from CFT Data

We both review and augment the lightcone conformal truncation (LCT) method. LCT is a Hamiltonian truncation method for calculating dynamical quantities in QFT in infinite volume. This document is a self-contained, pedagogical introduction and "how-to" manual for LCT. We focus on 2D QFTs which have UV descriptions as free CFTs containing scalars, fermions, and gauge fields, providing a rich starting arena for LCT applications. Along our way, we develop several new techniques and innovations that greatly enhance the efficiency and applicability of LCT. These include the development of CFT radial quantization methods for computing Hamiltonian matrix elements and a new SUSY-inspired way of avoiding state-dependent counterterms and maintaining chiral symmetry. We walk readers through the construction of their own basic LCT code, sufficient for small truncation cutoffs. We also provide a more sophisticated and comprehensive set of Mathematica packages and demonstrations that can be used to study a variety of 2D models. We guide the reader through these packages with several examples and illustrate how to obtain QFT observables, such as spectral densities and the Zamolodchikov $C$-function. Specific models considered are finite $N_c$ QCD, scalar $ϕ^4$ theory, and Yukawa theory.

preprint2015arXiv

The Goldstone Equivalence Theorem and AdS/CFT

The Goldstone equivalence theorem allows one to relate scattering amplitudes of massive gauge fields to those of scalar fields in the limit of large scattering energies. We generalize this theorem under the framework of the AdS/CFT correspondence. First, we obtain an expression of the equivalence theorem in terms of correlation functions of creation and annihilation operators by using an AdS wave function approach to the AdS/CFT dictionary. It is shown that the divergence of the non-conserved conformal current dual to the bulk gauge field is approximately primary when computing correlators for theories in which the masses of all the exchanged particles are sufficiently large. The results are then generalized to higher spin fields. We then go on to generalize the theorem using conformal blocks in two and four-dimensional CFTs. We show that when the scaling dimensions of the exchanged operators are large compared to both their spins and the dimension of the current, the conformal blocks satisfy an equivalence theorem.

preprint2014arXiv

Model-independent Analyses of Dark-Matter Particle Interactions

A model-independent treatment of dark-matter particle elastic scattering has been developed, yielding the most general interaction for WIMP-nucleon low-energy scattering, and the resulting amplitude has been embedded in the nucleus, taking into account the selection rules imposed by parity and time-reversal. One finds that, in contrast to the usual spin-independent/spin-dependent (SI/SD) formulation, the resulting cross section contains six independent nuclear response functions, three of which are associated with possible velocity-dependent interactions. We find that current experiments are four orders of magnitude more sensitive to derivative couplings than is apparent in the standard SI/SD treatment, which necessarily associates such interactions with cross sections proportional to the square of the WIMP velocity relative to the nuclear center of mass.

preprint2013arXiv

Model-independent WIMP Scattering Responses and Event Rates: A Mathematica Package for Experimental Analysis

The community's reliance on simplified descriptions of WIMP-nucleus interactions reflects the absence of analysis tools that integrate general theories of dark matter with standard treatments of nuclear response functions. To bridge this gap, we have constructed a public-domain Mathematica package for WIMP analyses based on our effective theory formulation. Script inputs are 1) the coefficients of the effective theory, through which one can characterize the low-energy consequences of arbitrary ultraviolet theories of WIMP interactions; and 2) one-body density matrices for commonly used targets, the most compact description of the relevant nuclear physics. The generality of the effective theory expansion guarantees that the script will remain relevant as new ultraviolet theories are explored; the use of density matrices to factor the nuclear physics from the particle physics will allow nuclear structure theorists to update the script as new calculations become available, independent of specific particle-physics contexts. The Mathematica package outputs the resulting response functions (and associated form factors) and also the differential event rate, once a galactic WIMP velocity profile is specified, and thus in its present form provides a complete framework for experimental analysis. The Mathematica script requires no a priori knowledge of the details of the non-relativistic effective field theory or nuclear physics, though the core concepts are reviewed here and in arXiv:1203.3542.