Source author record

Guimei Liu

Guimei Liu appears in the imported research catalog. Authorship, coauthor and topic links are available while profile ownership is still unclaimed.

ResearcherUnclaimed source record

Catalog footprint

What is connected

2works
2topics
4close collaborators

Actions

Connect this record

Log in to claim

Research graph

See the researcher in context

Open full explorer

Inspect adjacent papers, topics, institutions and collaborators without losing the researcher page.

Building this map preview

BZPEER is loading the nearby papers, people, topics and institutions for this page.

Published work

2 published item(s)

preprint2022arXiv

New Open Cluster candidates Found in Galactic Disk Using Gaia DR2/EDR3 Data

We report 541 new open cluster candidates in Gaia EDR3 through revisiting the cluster results from an earlier analysis of the Gaia DR2, which revealed nearly a thousand open cluster candidates in the solar neighborhood (mostly d < 3 kpc) resideing at Galactic latitudes |b| < 20 degrees. A subsequent comparison with lists of known clusters shows a large increases of the cluster samples within 2 kpc from the Sun. We assign membership probabilities to the stars through the open source pyUPMASK algorithm, and also estimate the physical parameters through isochrone fitting for each candidate. Most of the new candidates show small total proper motion dispersions and clear features in the color-magnitude diagrams. Besides, the metallicity gradient of the new candidates is consistent with those found in the literature. The cluster parameters and member stars are available at CDS via anonymous ftp to cdsarc.u-strasbg.fr(130.79.128.5) or via https://cdsarc.unistra.fr/viz-bin/cat/J/ApJS. The discovery of these new objects shows that the open cluster samples in Gaia data is still not complete, and more discoveries are expected in the future researches.

preprint2011arXiv

Controlling False Positives in Association Rule Mining

Association rule mining is an important problem in the data mining area. It enumerates and tests a large number of rules on a dataset and outputs rules that satisfy user-specified constraints. Due to the large number of rules being tested, rules that do not represent real systematic effect in the data can satisfy the given constraints purely by random chance. Hence association rule mining often suffers from a high risk of false positive errors. There is a lack of comprehensive study on controlling false positives in association rule mining. In this paper, we adopt three multiple testing correction approaches---the direct adjustment approach, the permutation-based approach and the holdout approach---to control false positives in association rule mining, and conduct extensive experiments to study their performance. Our results show that (1) Numerous spurious rules are generated if no correction is made. (2) The three approaches can control false positives effectively. Among the three approaches, the permutation-based approach has the highest power of detecting real association rules, but it is very computationally expensive. We employ several techniques to reduce its cost effectively.