Rule-based Cross-matching of Very Large Catalogs
Abstract
The NASA Extragalactic Database (NED) has deployed a new rule-based cross-matching algorithm called Match Expert (MatchEx), capable of cross-matching very large catalogs (VLCs) with >10 million objects. MatchEx goes beyond traditional position-based cross-matching algorithms by using other available data together with expert logic to determine which candidate match is the best. Furthermore, the local background density of sources is used to determine and minimize the false-positive match rate and to estimate match completeness. The logical outcome and statistical probability of each match decision is stored in the database and may be used to tune the algorithm and adjust match parameter thresholds. For our first production run, we cross-matched the GALEX All Sky Survey Catalog (GASC), containing nearly 40 million NUV-detected sources, against a directory of 180 million objects in NED. Candidate matches were identified for each GASC source within a 7''.5 radius. These candidates were filtered on position-based matching probability and on other criteria including object type and object name. We estimate a match completeness of 97.6% and a match accuracy of 99.75%. Over the next year, we will be cross-matching over 2 billion catalog sources to NED, including the Spitzer Source List, the 2MASS point-source catalog, AllWISE, and SDSS DR 10. We expect to add new capabilities to filter candidate matches based on photometry, redshifts, and refined object classifications. We will also extend MatchEx to handle more heterogenous datasets federated from smaller catalogs through NED's literature pipeline.
Additional Information
© 2015 Astronomical Society of the Pacific. The NASA/IPAC Extragalactic Database (NED) is operated by the Jet Propulsion Laboratory, California Institute of Technology, under contract with the National Aeronautics and Space Administration.Attached Files
Published - 495-0025.pdf
Submitted - 1503.01184v1.pdf
Files
Name | Size | Download all |
---|---|---|
md5:d1b33a0a4a4b22e0efe3f040a32dee78
|
971.5 kB | Preview Download |
md5:d86f595ce52be2839f4b8cc601335cd3
|
1.5 MB | Preview Download |
Additional details
- Alternative title
- Rule-based Cross-matching of Very Large Catalogs in NED
- Eprint ID
- 65643
- Resolver ID
- CaltechAUTHORS:20160324-081318970
- NASA/JPL/Caltech
- Created
-
2016-03-30Created from EPrint's datestamp field
- Updated
-
2023-06-02Created from EPrint's last_modified field
- Caltech groups
- Infrared Processing and Analysis Center (IPAC)
- Series Name
- Astronomical Society of the Pacific conference series
- Series Volume or Issue Number
- 495