Language Resources and Chemical Informatics

C.J. Rupp, Ann Copestake, Peter Corbett, Peter Murray-Rust, Advaith Siddharthan, Simone Teufel, Benjamin Waldron

Research output: Chapter in Book/Report/Conference proceedingPublished conference contribution

1 Citation (Scopus)


Chemistry research papers are a primary source of information about chemistry, as in any scientic eld. The presentation of the data is, predominantly, unstructured information, and so not immediately susceptible to processes developed within chemical informatics for carrying out chemistry research by information processing techniques. At one level, extracting the relevant information from research papers is a text mining task, requiring both extensive language resources and specialised knowledge of the subject domain. However,
the papers also encode information about the way the research is conducted and the structure of the eld itself. Applying language technology to research papers in chemistry can facilitate eScience on several different levels. The SciBorg project sets out to provide an extensive, analysed corpus of published chemistry research. This relies on the cooperation of several journal publishers to provide papers in an appropriate form. The work is carried out as a collaboration involving the Computer Laboratory, Chemistry Department and eScience Centre at Cambridge University, and is funded under the UK eScience programme.
Original languageEnglish
Title of host publicationProceedings of the 6th International Conference on Language Resources and Evaluation (LREC 2008)
Place of PublicationParis, France
Publication statusPublished - 2008
Event6th International Conference on Language Resources and Evaluation (LREC'2008) - Marrakesh, Morocco
Duration: 28 May 200830 May 2008


Conference6th International Conference on Language Resources and Evaluation (LREC'2008)


Dive into the research topics of 'Language Resources and Chemical Informatics'. Together they form a unique fingerprint.

Cite this