Source author record

Guosheng Xu

Guosheng Xu appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Computation and Language Cryptography and Security Machine Learning physics.plasm-ph

Catalog footprint

What is connected

2works

4topics

4close collaborators

Actions

Connect this record

Open graph Browse works

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

preprint2024arXiv

Digger: Detecting Copyright Content Mis-usage in Large Language Model Training

Pre-training, which utilizes extensive and varied datasets, is a critical factor in the success of Large Language Models (LLMs) across numerous applications. However, the detailed makeup of these datasets is often not disclosed, leading to concerns about data security and potential misuse. This is particularly relevant when copyrighted material, still under legal protection, is used inappropriately, either intentionally or unintentionally, infringing on the rights of the authors. In this paper, we introduce a detailed framework designed to detect and assess the presence of content from potentially copyrighted books within the training datasets of LLMs. This framework also provides a confidence estimation for the likelihood of each content sample's inclusion. To validate our approach, we conduct a series of simulated experiments, the results of which affirm the framework's effectiveness in identifying and addressing instances of content misuse in LLM training processes. Furthermore, we investigate the presence of recognizable quotes from famous literary works within these datasets. The outcomes of our study have significant implications for ensuring the ethical use of copyrighted materials in the development of LLMs, highlighting the need for more transparent and responsible data management practices in this field.

preprint2019arXiv

Stabilizing effect of enhanced resistivity on peeling-ballooning instabilities on EAST

Previous stability analysis of NSTX equilibrium with lithium-conditioning demonstrates that the enhanced resistivity due to the increased effective charge number Zeff (i.e. increased impurity level) can provide a stabilizing effect on low-n edge localized modes (Banerjee et al 2017 Nucl. Fusion 24 054501). This paper extends the resistivity stabilizing effect to the intermediate-n peeling-ballooning (PB) instabilities with the linear stability analysis of EAST high-confinement mode equilibria in NIMROD two-fluid calculations. However, the resistivity stabilizing effect on PB instabilities in the EAST tokamak appears weaker than that found in NSTX. This work may give better insight into the physical mechanism behind the beneficial effects of impurity on the pedestal stability.