species
OpenMed-NER-SpeciesDetect-ElectraMed-109MOpenMed-NER-SpeciesDetect-ModernMed-149MOpenMed-NER-SpeciesDetect-ModernClinical-395MOpenMed-NER-SpeciesDetect-PubMed-109MOpenMed-NER-SpeciesDetect-ElectraMed-560MOpenMed-NER-SpeciesDetect-BigMed-278MOpenMed-NER-SpeciesDetect-TinyMed-135MOpenMed-NER-SpeciesDetect-ModernMed-395M
Datasets
All datasets matching “species”vad-multi-species
Positive Transfer Of The Whisper Speech Transformer To Human And Animal Voice Activity Detection
We proposed WhisperSeg, utilizing the Whisper Transformer pre-trained for Automatic Speech Recognition (ASR) for both human and animal Voice Activity Detection (VAD). For more details, please refer to our paper
Positive Transfer of the Whisper Speech Transformer to Human and Animal Voice Activity Detection
Nianlong Gu, Kanghwi Lee, Maris Basha, Sumit Kumar Ram, Guanghao You, Richard H.… See the full description on the dataset page: https://huggingface.co/datasets/nccratliri/vad-multi-species.rare-species
Dataset Card for Rare Species Dataset
Dataset Description
Repository: Imageomics/bioclip
Paper: BioCLIP: A Vision Foundation Model for the Tree of Life (arXiv)
Dataset Summary
This dataset was generated alongside TreeOfLife-10M; data (images and text) were pulled from Encyclopedia of Life (EOL) to generate a dataset consisting of rare species for zero-shot-classification and more refined image classification tasks. Here, we use "rare species" to mean species… See the full description on the dataset page: https://huggingface.co/datasets/imageomics/rare-species.reefnet_species_images
ReefNet Species Images
Full-resolution coral reef images with species-level annotations from the
ReefNet dataset.
Dataset Description
This dataset contains 51,080 images from 54 CoralNet sources
with 365,030 annotation patches across 245 species.
Each annotation in metadata.parquet specifies a patch center (Row, Column)
within a full-resolution image. Crop a 224×224 patch centered at these coordinates
for classification tasks.
Splits
The curated… See the full description on the dataset page: https://huggingface.co/datasets/TsinghuaCorals/reefnet_species_images.species_800We have developed an efficient algorithm and implementation of a dictionary-based approach to named entity recognition,
which we here use to identifynames of species and other taxa in text. The tool, SPECIES, is more than an order of
magnitude faster and as accurate as existing tools. The precision and recall was assessed both on an existing gold-standard
corpus and on a new corpus of 800 abstracts, which were manually annotated after the development of the tool. The corpus
comprises abstracts from journals selected to represent many taxonomic groups, which gives insights into which types of
organism names are hard to detect and which are easy. Finally, we have tagged organism names in the entire Medline database
and developed a web resource, ORGANISMS, that makes the results accessible to the broad community of biologists.plant-multi-species-genomesDataset made of diverse genomes available on NCBI and coming from 48 different species.
Test and validation are made of 2 species each. The rest of the genomes are used for training.
Default configuration "6kbp" yields chunks of 6.2kbp (100bp overlap on each side). The chunks of DNA are cleaned and processed so that
they can only contain the letters A, T, C, G and N.cross-species-aging-data
Cross-Species Aging Data
This dataset repository stores the large matrices, metadata, and feature lists
required by the GitHub reproducibility package:
https://github.com/SimiaoZhao/cross-species-aging-resource
See data_manifest.tsv for the expected filenames and their analysis role.
During peer review, access may be restricted to reviewers and editors.
