Preprocessing in Biomedical Literature Mining Using Natural Language Processing

The number of biomedical literatures is growing rapidly, and biomedical literature mining is becoming essential. An approach for article processing in text preprocessing is proposed in order to improve the performance of biomedical literature mining. This approach combines the Web and corpus counts in order to eliminate the limitations of noise data of the Web. We experimentally showed that the performance of the combination models is the best comparing to the pure Web and corpus models. We achieve the best precision of 89.1% on all article forms and 88.7% article loss class.

[1]  William H. Majoros,et al.  Genomics and natural language processing , 2002, Nature Reviews Genetics.