Source author record

Subhankar Ghosh

Subhankar Ghosh appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

9works
6topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

9 published item(s)

preprint2026arXiv

A Comprehensive Dataset for Human vs. AI Generated Image Detection

Multimodal generative AI systems like Stable Diffusion, DALL-E, and MidJourney have fundamentally changed how synthetic images are created. These tools drive innovation but also enable the spread of misleading content, false information, and manipulated media. As generated images become harder to distinguish from photographs, detecting them has become an urgent priority. To combat this challenge, We release MS COCOAI, a novel dataset for AI generated image detection consisting of 96000 real and synthetic datapoints, built using the MS COCO dataset. To generate synthetic images, we use five generators: Stable Diffusion 3, Stable Diffusion 2.1, SDXL, DALL-E 3, and MidJourney v6. Based on the dataset, we propose two tasks: (1) classifying images as real or generated, and (2) identifying which model produced a given synthetic image. The dataset is available at https://huggingface.co/datasets/Rajarshi-Roy-research/Defactify_Image_Dataset.

preprint2022arXiv

TIC: Text-Guided Image Colorization

Image colorization is a well-known problem in computer vision. However, due to the ill-posed nature of the task, image colorization is inherently challenging. Though several attempts have been made by researchers to make the colorization pipeline automatic, these processes often produce unrealistic results due to a lack of conditioning. In this work, we attempt to integrate textual descriptions as an auxiliary condition, along with the grayscale image that is to be colorized, to improve the fidelity of the colorization process. To the best of our knowledge, this is one of the first attempts to incorporate textual conditioning in the colorization pipeline. To do so, we have proposed a novel deep network that takes two inputs (the grayscale image and the respective encoded text description) and tries to predict the relevant color gamut. As the respective textual descriptions contain color information of the objects present in the scene, the text encoding helps to improve the overall quality of the predicted colors. We have evaluated our proposed model using different metrics and found that it outperforms the state-of-the-art colorization algorithms both qualitatively and quantitatively.

preprint2022arXiv

Towards a Tighter Bound on Possible-Rendezvous Areas: Preliminary Results

Given trajectories with gaps, we investigate methods to tighten spatial bounds on areas (e.g., nodes in a spatial network) where possible rendezvous activity could have occurred. The problem is important for reducing the onerous amount of manual effort to post-process possible rendezvous areas using satellite imagery and has many societal applications to improve public safety, security, and health. The problem of rendezvous detection is challenging due to the difficulty of interpreting missing data within a trajectory gap and the very high cost of detecting gaps in such a large volume of location data. Most recent literature presents formal models, namely space-time prism, to track an object's rendezvous patterns within trajectory gaps on a spatial network. However, the bounds derived from the space-time prism are rather loose, resulting in unnecessarily extensive post-processing manual effort. To address these limitations, we propose a Time Slicing-based Gap-Aware Rendezvous Detection (TGARD) algorithm to tighten the spatial bounds in spatial networks. We propose a Dual Convergence TGARD (DC-TGARD) algorithm to improve computational efficiency using a bi-directional pruning approach. Theoretical results show the proposed spatial bounds on the area of possible rendezvous are tighter than that from related work (space-time prism). Experimental results on synthetic and real-world spatial networks (e.g., road networks) show that the proposed DC-TGARD is more scalable than the TGARD algorithm.

preprint2014arXiv

Online Stroke and Akshara Recognition GUI in Assamese Language Using Hidden Markov Model

The work describes the development of Online Assamese Stroke & Akshara Recognizer based on a set of language rules. In handwriting literature strokes are composed of two coordinate trace in between pen down and pen up labels. The Assamese aksharas are combination of a number of strokes, the maximum number of strokes taken to make a combination being eight. Based on these combinations eight language rule models have been made which are used to test if a set of strokes form a valid akshara. A Hidden Markov Model is used to train 181 different stroke patterns which generates a model used during stroke level testing. Akshara level testing is performed by integrating a GUI (provided by CDAC-Pune) with the Binaries of HTK toolkit classifier, HMM train model and the language rules using a dynamic linked library (dll). We have got a stroke level performance of 94.14% and akshara level performance of 84.2%.

preprint2011arXiv

Concentration of measure for the number of isolated vertices in the Erdős-Rényi random graph by size bias couplings

A concentration of measure result is proved for the number of isolated vertices $Y$ in the Erdős-Rényi random graph model on $n$ edges with edge probability $p$. When $μ$ and $σ^2$ denote the mean and variance of $Y$ respectively, $P((Y-μ)/σ\ge t)$ admits a bound of the form $e^{-kt^2}$ for some constant positive $k$ under the assumption $p \in (0,1)$ and $np\rightarrow c \in (0,\infty)$ as $n \rightarrow \infty$. The left tail inequality $$ P(\frac{Y-μ}σ\le -t)&\le& \exp(-\frac{t^2σ^2}{4μ}) $$ holds for all $n \in {2,3,...},p \in (0,1)$ and $t \ge 0$. The results are shown by coupling $Y$ to a random variable $Y^s$ having the $Y$-size biased distribution, that is, the distribution characterized by $E[Yf(Y)]=μE[f(Y^s)] $ for all functions $f$ for which these expectations exist.

preprint2011arXiv

Concentration of measures via size biased couplings

Let $Y$ be a nonnegative random variable with mean $μ$ and finite positive variance $σ^2$, and let $Y^s$, defined on the same space as $Y$, have the $Y$ size biased distribution, that is, the distribution characterized by E[Yf(Y)]=μE f(Y^s) for all functions $f$ for which these expectations exist. Under a variety of conditions on the coupling of Y and $Y^s$, including combinations of boundedness and monotonicity, concentration of measure inequalities hold. Examples include the number of relatively ordered subsequences of a random permutation, sliding window statistics including the number of m-runs in a sequence of coin tosses, the number of local maximum of a random function on a lattice, the number of urns containing exactly one ball in an urn allocation model, the volume covered by the union of $n$ balls placed uniformly over a volume n subset of d dimensional Euclidean space, the number of bulbs switched on at the terminal time in the so called lightbulb process, and the infinitely divisible and compound Poisson distributions that satisfy a bounded moment generating function condition.

preprint2010arXiv

$L^p$ bounds for a central limit theorem with involutions

Let $E=((e_{ij}))_{n\times n}$ be a fixed array of real numbers such that $e_{ij}=e_{ji}, e_{ii}=0$ for $1\le i,j \le n$. Let the permutation group be denoted by $S_n$ and the collection of involutions with no fixed points by $Π_n$, that is, $Π_n=\{π\in S_n: π^2= id, π(i)\neq i\,\forall i\}$ with id denoting the identity permutation. For $π$ uniformly chosen from $Π_n$, let $Y_E=\sum_{i=1}^n e_{iπ(i)}$ and $W=(Y_E-μ_E)/σ_E$ where $μ_E=E(Y_E)$ and $σ_E^2= Var(Y_E)$. Denoting by $F_W$ and $Φ$ the distribution functions of $W$ and a $\mathcal{N}(0,1)$ variate respectively, we bound $||F_W-Φ||_p$ for $ 1\le p\le \infty$ using Stein's method and the zero bias transformation. Optimal Berry-Esseen or $L^\infty$ bounds for the classical problem where $π$ is chosen uniformly from $S_n$ were obtained by Bolthausen using Stein's method. Although in our case $π\in Π_n$ uniformly, the $L^p$ bounds we obtain are of similar form as Bolthausen's bound which holds for $p=\infty$. The difficulty in extending Bolthausen's method from $S_n$ to $Π_n$ arising due to the involution restriction is tackled by the use of zero bias transformations.

preprint2010arXiv

Multivariate concentration of measure type results using exchangeable pairs and size biasing

Let $(\mathbf{W,W'})$ be an exchangeable pair of vectors in $\mathbb{R}^k$. Suppose this pair satisfies \beas E(\mathbf{W}'|\mathbf{W})=(I_k-Λ)\mathbf{W}+\mathbf{R(W)}. \enas If $||\mathbf{W-W'}||_2\le K$ and $\mathbf{R(W)}=0$, then concentration of measure results of following form is proved for all $\mathbf{w}\succeq 0$ when the moment generating function of $\mathbf{W}$ is finite. \beas P(\mathbf{W}\succeq\mathbf{w}),P(\mathbf{W}\preceq -\mathbf{w})\le \exp(-\frac{||\mathbf{w}||_2^2}{2K^2ν_1}), \enas for an explicit constant $ν_1$, where $\succeq$ stands for coordinate wise $\ge$ ordering. This result is applied to examples like complete non degenerate U-statistics. Also, we deal with the example of doubly indexed permutation statistics where $\mathbf{R(W)}\neq 0$ and obtain similar concentration of measure inequalities. Practical examples from doubly indexed permutation statistics include Mann-Whitney-Wilcoxon statistic and random intersection of two graphs. Both these two examples are used in nonparametric statistical testing. We conclude the paper with a multivariate generalization of a recent concentration result due to Ghosh and Goldstein \cite{cnm} involving bounded size bias couplings.