Fast retrieval method of biomedical literature based on feature mining Online publication date: Tue, 17-Oct-2023
by Duo Long; Yunxin Long; Fahuan Xie; Ping Yu; Hui Yan
International Journal of Data Mining and Bioinformatics (IJDMB), Vol. 27, No. 4, 2023
Abstract: In order to solve the problems of large errors, low accuracy of feature mining and time-consuming traditional literature retrieval methods, this paper designs a fast retrieval method for biomedical literature based on feature mining. First, we simulate the document collection space, and collect documents according to the data centroid and probability density function. Secondly, the location of similar data is marked by mutual information method, and the hidden information of literature data is extracted after reducing the imbalance of dataset. Then, the Pearson correlation coefficient of the literature data is calculated and the key features of the literature are mined. Finally, we calculate the expected loss risk of literature data, design a fast retrieval algorithm for biomedical literature, and realise fast retrieval. The test results show that this method can reduce the retrieval error, improve the accuracy of document feature mining, and the retrieval time is shorter.
Existing subscribers:
Go to Inderscience Online Journals to access the Full Text of this article.
If you are not a subscriber and you just want to read the full contents of this article, buy online access here.Complimentary Subscribers, Editors or Members of the Editorial Board of the International Journal of Data Mining and Bioinformatics (IJDMB):
Login with your Inderscience username and password:
Want to subscribe?
A subscription gives you complete access to all articles in the current issue, as well as to all articles in the previous three years (where applicable). See our Orders page to subscribe.
If you still need assistance, please email subs@inderscience.com