> top > users > Yue Wang
Yue Wang
User info
Collections
Name DescriptionUpdated at
1-3 / 3
AnEMthe largest manually annotated corpus on anatomical entities2019-04-03
DisGeNET5Associations obtained by text mining MEDLINE abstracts using the BeFree system2019-03-11
PIRProtein Information Resource (PIR)2019-03-12
Projects
NameTDescription# Ann.Updated atStatus
1-10 / 25 show all
0mytest1442023-11-29
bionlp-st-pc-2013-trainingThe training dataset from the pathway curation (PC) task in the BioNLP Shared Task 2013. The entity types defined in the PC task are simple chemical, gene or gene product, complex and cellular component.7.86 K2023-11-27Released
AIMedThe AIMed corpus is one of the most widely used corpora for protein-protein interaction extraction. The protein annotations are either parts of the protein interaction annotations, or are uninvolved in any protein interaction annotation. Publication: http://www.cs.utexas.edu/~ml/papers/bionlp-aimed-04.pdf4.04 K2023-11-27Testing
SCAI-TestA small corpus for the evaluation of dictionaries containing chemical entities. Publication: http://www.scai.fraunhofer.de/fileadmin/images/bio/data_mining/paper/kolarik2008.pdf Original source: https://www.scai.fraunhofer.de/en/business-research-areas/bioinformatics/downloads/corpora-for-chemical-entity-recognition.html1.21 K2023-11-28Released
bionlp-st-epi-2011-trainingThe training dataset from the Epigenetics and Post-translational Modifications (EPI) task in the BioNLP Shared Task 2011. The core entities of the task are genes and gene products (RNA and proteins), identified in the data simply as "Protein" annotations. 7.59 K2023-11-29Released
DisGeNET5_variant_diseaseThe file contains variant-disease associations obtained by text mining MEDLINE abstracts using the BeFree system, including the variant and disease off sets. 144 K2023-11-24Released
OryzaGPA dataset for Named Entity Recognition for rice gene29.1 K2023-11-24Uploading
PIR-corpus1The Protein Information Resource (PIR) is not biased towards any particular biomedical domain, and is expected to provide more diverse protein names in a given sample size. Annotation category: protein, compound-protein, acronym.4.44 K2023-11-27Released
funRiceGenes-exact8412023-11-28Developing
2_test145 M2023-11-24
Automatic annotators
Editors

none