Source author record

Tomasz Stebel

Tomasz Stebel appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

20works
8topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

20 published item(s)

preprint2026arXiv

Sampling two-dimensional spin systems with transformers

Autoregressive Neural Networks based on dense or convolutional layers have recently been shown to be a viable strategy for generating classical spin systems. Unlike these methods, sampling with transformers is commonly considered to be computationally inefficient. In this work, we propose a novel approach to transformer-based neural samplers in which we generate not a single spin per step but groups of spins. As an additional improvement, we construct a model of approximated probabilities, further improving the efficiency of the algorithm. Despite our approach being computationally heavier than dense networks or CNN-based approaches, we were able to sample larger systems of up to $180 \times 180$ spins in case of the Ising model. The Effective Sample Size of our sampler is $\sim 20$ times larger than that of the previous state-of-the-art neural sampler when trained for the $128 \times 128$ Ising model at critical temperature. Finally, we also test our algorithm on the 2D Edwards-Anderson model, where we train $64\times 64$ spin systems.

preprint2026arXiv

Variational Autoregressive Networks with probability priors

Monte Carlo methods are essential across diverse scientific fields, yet their efficiency is frequently hampered by critical slowing down-a sharp increase in autocorrelation times near phase transitions. Although deep learning approaches, such as neural-network-based samplers, have been proposed to alleviate this issue, they face another serious problem: the difficulty of training the models. This difficulty partially stems from the overly general nature of original machine-learning architectures, which often ignore underlying physical symmetries and force networks to relearn them from scratch. In this paper, we demonstrate that incorporating physical priors into the model significantly enhances performance. Building upon existing strategies that integrate spin-spin interactions, we propose a framework that utilizes a prior probability distribution as a starting point for training. Our results for the Ising model, as well as for the Edwards-Anderson spin glass model, suggest that moving away from `blank slate' models in favor of physics-informed priors reduces the training burden and facilitates the simulation of larger system sizes in discrete spin models.

preprint2023arXiv

Analysis of autocorrelation times in Neural Markov Chain Monte Carlo simulations

We provide a deepened study of autocorrelations in Neural Markov Chain Monte Carlo (NMCMC) simulations, a version of the traditional Metropolis algorithm which employs neural networks to provide independent proposals. We illustrate our ideas using the two-dimensional Ising model. We discuss several estimates of autocorrelation times in the context of NMCMC, some inspired by analytical results derived for the Metropolized Independent Sampler (MIS). We check their reliability by estimating them on a small system where analytical results can also be obtained. Based on the analytical results for MIS we propose a new loss function and study its impact on the autocorelation times. Although, this function's performance is a bit inferior to the traditional Kullback-Leibler divergence, it offers two training algorithms which in some situations may be beneficial. By studying a small, $4 \times 4$, system we gain access to the dynamics of the training process which we visualize using several observables. Furthermore, we quantitatively investigate the impact of imposing global discrete symmetries of the system in the neural network training process on the autocorrelation times. Eventually, we propose a scheme which incorporates partial heat-bath updates which considerably improves the quality of the training. The impact of the above enhancements is discussed for a $16 \times 16$ spin system. The summary of our findings may serve as a guidance to the implementation of Neural Markov Chain Monte Carlo simulations for more complicated models.

preprint2022arXiv

Gradient estimators for normalising flows

Recently a machine learning approach to Monte-Carlo simulations called Neural Markov Chain Monte-Carlo (NMCMC) is gaining traction. In its most popular form it uses neural networks to construct normalizing flows which are then trained to approximate the desired target distribution. In this contribution we present new gradient estimator for Stochastic Gradient Descent algorithm (and the corresponding \texttt{PyTorch} implementation) and show that it leads to better training results for $ϕ^4$ model. For this model our estimator achieves the same precision in approximately half of the time needed in standard approach and ultimately provides better estimates of the free energy. We attribute this effect to the lower variance of the new estimator. In contrary to the standard learning algorithm our approach does not require estimation of the action gradient with respect to the fields, thus has potential of further speeding up the training for models with more complicated actions.

preprint2021arXiv

Prompt photon production in proton collisions as a probe of parton scattering in high energy limit

We study the prompt photon hadroproduction at the LHC with the $k_T$-factorization approach and the $qg^* \to qγ$ and $g^*g^* \to q\bar qγ$ partonic channels, using three unintegrated gluon distributions which depend on gluon transverse momentum. They represent three different theoretical schemes which are usually considered in the $k_T$-factorization approach, known under the acronyms: KMR, CCFM and GBW gluon distributions. We find sensitivity of the calculated prompt photon transverse momentum distribution to the gluon transverse momentum distribution. The predictions obtained with the three approaches are compared to data, that allows to differentiate between them. We also discuss the significance of the two partonic channels, confronted with the expectations which are based on the applicability of the $k_T$-factorization scheme in the high energy approximation.

preprint2020arXiv

Associated top quark pair production with a heavy boson: differential cross sections at NLO+NNLL accuracy

We present theoretical predictions for selected differential cross sections for the process $pp \to t \bar{t} B$ at the LHC, where $B$ can be a Higgs ($H$), a $Z$ or a $W$ boson. The predictions are calculated in the direct QCD framework up to the next-to-next-leading logarithmic (NNLL) accuracy and matched to the complete NLO results including QCD and electroweak effects. Additionally, results for the total cross sections are provided. The calculations deliver a significant improvement of the theoretical predictions, especially for the $t \bar{t} H$ and the $t \bar{t} Z$ production. In these cases, predictions for both the total and differential cross sections are remarkably stable with respect to the central scale choice and carry a substantially reduced scale uncertainty in comparison with the complete NLO predictions.

preprint2020arXiv

Drell-Yan production with the CCFM-K evolution

We discuss the Drell-Yan dilepton production using the transverse momentum dependent parton distributions evolved with the Catani-Ciafaloni-Fiorani-Marchesini-Kwieciński (CCFM-K) equations in the single loop approximation. Such equations are obtained assuming angular ordering of emitted partons (coherence) for $x\sim 1$ and transverse momentum ordering for $x \ll 1$. This evolution scheme also contains the Collins-Soper-Sterman (CSS) soft gluon resummation. We make a comparison with a broad class of data on transverse momentum spectra of low mass Drell-Yan dileptons.

preprint2020arXiv

Sub-femtometer scale color charge correlations in the proton

Color charge correlations in the proton at moderately small $x\sim 0.1$ are extracted from its light-cone wave function. The charge fluctuations are far from Gaussian and they exhibit interesting dependence on impact parameter and on the relative transverse momentum (or distance) of the gluon probes. We provide initial conditions for small-$x$ Balitsky-Kovchegov evolution of the dipole scattering amplitude with impact parameter and $\hat r \cdot \hat b$ dependence, and with non-zero $C$-odd component due to three-gluon exchange. Lastly, we compute the (forward) Weizsaecker-Williams gluon distributions, including the distribution of linearly polarized gluons, up to fourth order in $A^+$. The correction due to the quartic correlator provides a transverse momentum scale, $q > 0.5$ GeV, for nearly maximal polarization.

preprint2020arXiv

Top Precision for Associated Top-Pair Production Processes at the LHC

The studies of the associated production processes of a top-quark pair with a colour-singlet boson, e.g. Higgs, W or Z, are among the highest priorities of the LHC programme. Correspondingly, improvements in precision of theoretical predictions for these processes are of central importance. In this talk, we review our latest results on resummation of soft gluon corrections. The resummation is carried out using the direct QCD Mellin space technique in three-particle invariant mass kinematics. We discuss the impact of the soft gluon corrections on predictions for total cross sections and differential distributions.

preprint2017arXiv

Twist decomposition of Drell-Yan structure functions: phenomenological implications

The forward Drell--Yan process in $pp$ scattering at the LHC at $\sqrt{S}=14$ TeV is considered. We analyze the Drell--Yan structure functions assuming the dominance of a Compton-like emission of a virtual photon from a fast quark scattering off the small $x$ gluons. The color dipole framework is applied to perform quantitatively the twist decomposition of all the Drell--Yan structure functions. Two models of the color dipole scattering are applied: the Golec-Biernat--Wüsthoff model and the dipole cross section obtained from the Balitsky--Fadin--Kuraev--Lipatov evolution equation. The two models have essentially different higher twist content and the gluon transverse momentum distribution and lead to different significant effects beyond the collinear leading twist description. It is found that the gluon transverse momentum effects are significant in the Drell--Yan structure functions for all Drell--Yan pair masses $M$, and the higher twist effects become important for $M \lesssim 10$ GeV. It is found that the structure function $W_{TT}$ related to the $A_2$ angular coefficient and the Lam--Tung observable $A_0 -A_2$ are particularly sensitive to the gluon $k_T$ effects and to the higher twist effects. A procedure is suggested how to disentangle the higher twist effects from the gluon transverse momentum effects.

preprint2016arXiv

Quantum Telescopes: feasibility and constrains

Quantum Telescope is a recent idea aimed at beating the diffraction limit of spaceborne telescopes and possibly also other distant target imaging systems. There is no agreement yet on the best setup of such devices, but some configurations have been already proposed. In this Letter we characterize the predicted performance of Quantum Telescopes and their possible limitations. Our extensive simulations confirm that the presented model of such instruments is feasible and the device can provide considerable gains in the angular resolution of imaging in the UV, optical and infrared bands. We argue that it is generally possible to construct and manufacture such instruments using the latest or soon to be available technology. We refer to the latest literature to discuss the feasibility of the proposed QT system design.

preprint2016arXiv

Soft gluon resummation at fixed invariant mass for associated $t\bar{t}H$ production at the LHC

In the following we present our results on resummation of invariant mass threshold corrections for the $2 \to 3$ type hadronic production processes in the Mellin moment space formalism. This method is applied to the associated Higgs boson production process $pp \to t \bar{t} H$ at the LHC. The results for the total cross section, the differential distribution with respect to the invariant mass and their uncertainties are presented.

preprint2016arXiv

Twist expansion of forward Drell-Yan process

We present a twist expansion of differential cross-sections of the forward Drell-Yan process at the high energies. The expansion of all invariant form-factors is performed assuming GBW saturation model and the saturation scale plays the role of the hadronic scale of OPE. Some expilicit predictions for LHC experiments are given. It is shown also how the Lam-Tung relation is broken at twist 4 what provides a sensitive probe for searching of higher twists.

preprint2015arXiv

Soft gluon resummation for associated $t \bar{t} H$ production at the LHC

We perform resummation of soft gluon corrections to the total cross section for the process $pp \to t\bar{t}H$. The resummation is carried out at next-to-leading-logarithmic (NLL) accuracy using the Mellin space technique, extending its application to the class of $2 \to 3$ processes. We present an analytical result for the soft anomalous dimension for a hadronic production of two coloured massive particles in association with a colour singlet. We discuss the impact of resummation on the numerical prediction for the associated Higgs boson production with top quarks at the LHC.

preprint2015arXiv

Twist expansion of Drell-Yan structure functions in color dipole approach

Forward Drell-Yan process at the LHC probes the proton structure at very small Bjorken-$x$ and moderate hard scales. In this kinematical domain higher twist effects may be significant and introduce sizeable corrections to the standard leading twist description. We study the forward Drell-Yan process beyond the leading twist approximation within the color dipole model framework that incorporates multiple scattering effects. We derive the Mellin representation of the forward Drell-Yan impact factors for fully differential cross-sections. These results combined with the color dipole cross-section of the saturation model are used to perform twist expansion of the Drell-Yan structure functions at arbitrary transverse momentum $q_T$ of the Drell-Yan pair and also of the structure functions integrated over $q_T$. We also investigate the Lam-Tung relation, find that it is broken at twist 4 and provide explicit estimates for the breaking term.

preprint2014arXiv

Twist expansion of differential cross-sections of forward Drell-Yan process

Forward Drell-Yan process at the LHC is a sensitive tool for investigating higher twist effects in QCD. The expansion of all Drell-Yan structure functions is performed assuming GBW saturation model and the saturation scale plays the role of the hadronic scale of OPE. We show that the Lam-Tung relation is broken at twist 4. The results open the way for a forthcoming analysis of multiple scattering and higher twist effects.

preprint2013arXiv

Quantitative analysis of Geometrical Scaling in Deep Inelastic Scattering

We analyze geometrical scaling in deep inelastic scattering using experimental data from HERA ep collider. Parallel analyses are performed for energy and Bjorken-x data binnings. In particular, value of parameter λwhich governs x-dependence of saturation scale is found for both binnings: λ_En = 0.352 \pm 0.008 and λ_Bj = 0.302 \pm 0.004. We use the following method: λ_En is found as a value for which ratios σ^{W1}_{γ*p}(τ)/σ^{W2}_{γ*p}(τ) are closest to 1 (σ^{Wi}_{γ*p} is a photon-proton cross section with definite energy W_i and τ= Q^2 x^λ is a scaling variable); λ_Bj is found analogously using ratios of cross sections with definite Bjorken-x variable. We show also that GS is present for x<0.2.

preprint2013arXiv

Quantitative Study of Different Forms of Geometrical Scaling in Deep Inelastic Scattering at HERA

We use recently proposed method of ratios to assess the quality of geometrical scaling in deep inelastic scattering for different forms of the saturation scale. We consider original form of geometrical scaling (motivated by the Balitski-Kovchegov (BK) equation with fixed coupling) studied in more detail in our previous paper, and four new hypotheses: phenomenologically motivated case with $Q^2$ dependent exponent $λ$ that governs small $x$ dependence of the saturation scale, two versions of scaling (running coupling 1 and 2) that follow from the BK equation with running coupling, and diffusive scaling suggested by the QCD evolution equation beyond mean field approximation. It turns out that more sophisticated scenarios: running coupling scaling and diffusive scaling are disfavored by the combined HERA data on $e^+p$ deep inelastic structure function $F_2$.

preprint2013arXiv

Quantitative Study of Geometrical Scaling in Charm Production at HERA

We apply the method of ratios to search for geometrical scaling in the charm production in deep inelastic scattering. To this end we use recent combined data from H1 and ZEUS experiments. Two forms of geometrical scaling are tested: originally proposed scaling which results from Golec-Biernat-Wusthoff model and scaling motivated by a dipole representation which takes into account charm mass. It turns out that in both cases some residual scaling is present and charm mass inclusion improves scaling quality.

preprint2013arXiv

Quantitative Study of Geometrical Scaling in Deep Inelastic Scattering at HERA

We propose a method to assess the quality of geometrical scaling in Deep Inelastic Scattering and apply it to the combined HERA data on $γ^{\ast}p$ cross-section. Using two different approaches based on Bjorken $x$ binning and binning in $γ^{\ast}p$ scattering energy $W$, we show that geometrical scaling in variable $τ\sim Q^{2} x^λ$ works well up to Bjorken $x$'s 0.1. The corresponding value of exponent $λ$ is 0.32 -- 0.34.