Please use this identifier to cite or link to this item: http://dx.doi.org/10.25673/118442
Title: Duplicate detection of 2D-NMR Spectra
Author(s): Hinneburg, Alexander
Egert, Björn
Porzel, Andrea
Issue Date: 2007
Type: Article
Language: English
Abstract: 2D-Nuclear magnetic resonance (NMR) spectra are used in the (structural) analysis ofsmall molecules. In contrast to 1D-NMR spectra, 2D-NMR spectra correlate the chemicalshifts of1H and13C at the same time. A spectrum consists of several peaks in a two-dimensional space. The most important information of a peak is the location of its center,which captures the bonding relationships of hydrogen and carbon atoms. A spectrum con-tains much information about the chemical structure of a product, but in most cases thestructure cannot be read off in a simple and straightforward manner. Structure elucidationinvolves a considerable amount (manual) efforts.Using high-field NMR spectrometers, many 2D-NMR spectra can be recorded in shorttime. So the common situation is that a lab or company has a repository of 2D-NMRspectra, partially annotated with the structural information. For the remaining spectra thestructure in unknown. In case two research labs are collaborating, the repositories will bemerged and annotations shared.We reduce that problem to the task of finding duplicates in a given set of 2D-NMR spectra.Therefore, we propose a simple but robust definition of 2D-NMR duplicates, which allowsfor small measurement errors. We give a quadratic algorithm for the problem, which canbe implemented in SQL. Further, we analyze a more abstract class of heuristics, which arebased on selecting particular peaks. Such a heuristic works as a filter step on the pairs ofpossible duplicates and allows false positives. We compare all methods with respect totheir run time. Finally we discuss the effectiveness of the duplicate definition on real data.
URI: https://opendata.uni-halle.de//handle/1981185920/120401
http://dx.doi.org/10.25673/118442
Open Access: Open access publication
License: (CC BY-NC-ND 4.0) Creative Commons Attribution NonCommercial NoDerivatives 4.0(CC BY-NC-ND 4.0) Creative Commons Attribution NonCommercial NoDerivatives 4.0
Journal Title: Journal of integrative bioinformatics
Publisher: Walter de Gruyter GmbH
Publisher Place: Berlin
Volume: 4
Issue: 1
Original Publication: 10.1515/jib-2007-53
Appears in Collections:Open Access Publikationen der MLU

Files in This Item:
File Description SizeFormat 
10.1515_jib-2007-53.pdf718.56 kBAdobe PDFThumbnail
View/Open