LitCovid-PD-CHEBI | | | 1.44 M | | admin | 2020-11-30 | Developing | |
PubMed-German-test | | A collection of PubMed abstracts which are written in German | 0 | | Jin-Dong Kim | 2020-12-13 | Developing | |
PubMed-French-test | | A collection of PubMed abstract written in French | 0 | | Jin-Dong Kim | 2020-12-13 | Developing | |
LitCovid-sample-Enju | | | 13.1 K | | Jin-Dong Kim | 2021-01-14 | Developing | |
LitCovid-sample-PD-MONDO | | | 1.21 K | | Jin-Dong Kim | 2021-01-14 | Developing | |
LitCovid-sample-PD-GlycoEpitope | | | 5 | | Jin-Dong Kim | 2021-01-14 | Developing | |
LitCovid-sample-PD-MAT | | | 251 | | Jin-Dong Kim | 2021-01-14 | Developing | |
LitCovid-docs | | A comprehensive literature resource on the subject of Covid-19 is collected by NCBI:
https://www.ncbi.nlm.nih.gov/research/coronavirus/
The LitCovid project@PubAnnotation is a collection of the titles and abstracts of the LitCovid dataset, for the people who want to perform text mining analysis. Please note that if you produce some annotation to the documents in this project, and contribute the annotation back to PubAnnotation, it will become publicly available together with contribution from other people.
If you want to contribute your annotation to PubAnnotation, please refer to the documentation page:
http://www.pubannotation.org/docs/submit-annotation/
The list of the PMID is sourced from here
The 6 entries of the following PMIDs could not be included because they were not available from PubMed:32161394,
32104909,
32090470,
32076224,
32161394
32188956,
32238946.
Below is a notice from the original LitCovid dataset:
PUBLIC DOMAIN NOTICE
National Center for Biotechnology Information
This software/database is a "United States Government Work" under the
terms of the United States Copyright Act. It was written as part of
the author's official duties as a United States Government employee and
thus cannot be copyrighted. This software/database is freely available
to the public for use. The National Library of Medicine and the U.S.
Government have not placed any restriction on its use or reproduction.
Although all reasonable efforts have been taken to ensure the accuracy
and reliability of the software and data, the NLM and the U.S.
Government do not and cannot warrant the performance or results that
may be obtained by using this software or data. The NLM and the U.S.
Government disclaim all warranties, express or implied, including
warranties of performance, merchantability or fitness for any particular
purpose.
Please cite the authors in any work or product based on this material :
Chen Q, Allot A, & Lu Z. (2020) Keep up with the latest coronavirus research, Nature 579:193
| 18 | | Jin-Dong Kim | 2021-01-16 | Developing | |
BLAH2015_Annotations_Adderall | | | 0 | nestoralvaro | nestoralvaro | 2015-03-15 | Testing | |
BioASQ-sample | | collection of PubMed articles which appear in the BioASQ sample data set. | 0 | BioASQ | Jin-Dong Kim | 2015-10-13 | Testing | |
SPECIES800_autotagged | | This project comprises the SPECIES800 corpus documents automatically annotated by the Jensenlab tagger.
Annotated entity types are:
Genes/proteins from the mentioned organisms (and any human ones)
PubChem Compound identifiers
NCBI Taxonomy entries
Gene Ontology cellular component terms
BRENDA Tissue Ontology terms
Disease Ontology terms
Environment Ontology terms
The SPECIES 800 (S800) comprises 800 PubMed abstracts. In its original form species mentions were manually identified and mapped to the corresponding NCBI Taxonomy identifiers.
Described in:
The SPECIES and ORGANISMS Resources for Fast and Accurate Identification of Taxonomic Names in Text.
Pafilis E, Frankild SP, Fanini L, Faulwetter S, Pavloudi C, et al. (2013). PLoS ONE, 2013, 8(6): e65390. doi:10.1371/journal.pone.0065390.
The manually annotated corpus is also available as a PubAnnotation project (see here).
| 0 | Evangelos Pafilis, Sampo Pyysalo, Lars Juhl Jensen | evangelos | 2015-11-20 | Testing | |
GlycoBiology-PACDB | | cGGDB-based annotation to GlycoBiology abstracts | 3.03 K | Toshihide Shikanai | shikanai | 2016-02-01 | Testing | |
GlycoBiology-GDGDB | | GDGDB-based annotation to GlycoBiology abstracts | 2.46 K | Toshihide Shikanai | shikanai | 2016-02-01 | Testing | |
GlycoBiology-cGGDB | | cGGDB-based annotation to GlycoBiology abstracts | 36 | Toshihide Shikanai | shikanai | 2016-02-01 | Testing | |
GlycoBiology-GO | | GO-based annotation to GlycoBiology abstracts | 0 | | Jin-Dong Kim | 2016-06-11 | Testing | |
EDAM-DFO | | annotation for EDAM terms for data, formats, and operations | 12.5 K | | Jin-Dong Kim | 2016-09-21 | Testing | |
EDAM-topics | | annotation for EDAM topics | 11.6 K | | Jin-Dong Kim | 2016-09-21 | Testing | |
Parkinson | | | 54 | | Jin-Dong Kim | 2017-03-11 | Testing | |
AlvisNLP-Async-Test | | Test for the asynchronous AlvisNLP/ML annotator family. | 0 | Robert Bossy | rbossy | 2017-04-07 | Testing | |
Test_economics | | test | 1 | YongHwanKim | kimyonghwan | 2017-07-13 | Testing | |