datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Zebrafish_DNA_v0_tokenized_kmer6_stride1dice_color_sort_mergedThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"fps": 30,
"features": {
"action": {
"dtype": "float32",
"names": [
"shoulder_pan.pos",
"shoulder_lift.pos",
"elbow_flex.pos",
"wrist_flex.pos",
"wrist_roll.pos",
"gripper.pos"
],
"shape": [
6… See the full description on the dataset page: https://huggingface.co/datasets/k-chan-l/dice_color_sort_merged.Worm_DNA_v0_tokenized_kmer6_stride1Fruitfly_DNA_v0_tokenized_kmer6_stride1kmerArabidopsis_thaliana_DNA_v0_tokenized_kmer6_stride1Mouse_DNA_v0_tokenized_kmer6_stride1record-redlegoblock_20260715_133031This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"fps": 30,
"features": {
"action": {
"dtype": "float32",
"names": [
"shoulder_pan.pos",
"shoulder_lift.pos",
"elbow_flex.pos",
"wrist_flex.pos",
"wrist_roll.pos",
"gripper.pos"
],
"shape": [
6… See the full description on the dataset page: https://huggingface.co/datasets/kmerckae/record-redlegoblock_20260715_133031.record-test_20260711_134602This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"fps": 30,
"features": {
"action": {
"dtype": "float32",
"names": [
"shoulder_pan.pos",
"shoulder_lift.pos",
"elbow_flex.pos",
"wrist_flex.pos",
"wrist_roll.pos",
"gripper.pos"
],
"shape": [
6… See the full description on the dataset page: https://huggingface.co/datasets/kmerckae/record-test_20260711_134602.virus_dna_dedup_minihash_0.9_kmer_7
Dataset Card for "virus_dna_dedup_minihash_0.9_kmer_7"
More Information needed
Merge_datasetkmer_features_count_datasetpretty_name: "kmer_features_count_dataset"
tags:
flwr
genomics
A very simple dataset for testing and learning bioinformatics and federated learning .
Data is collected from katarinagresova/Genomic_Benchmarks_drosophila_enhancers_stark
then I apply very simple preprocessing like count kmer technuqes
dataset_info:
features:
- name: '0'
dtype: float64
- name: '1'
dtype: float64
- name: '2'
dtype: float64
- name: '3'
dtype: float64
- name: '4'
dtype:… See the full description on the dataset page: https://huggingface.co/datasets/EsaSinthia/kmer_features_count_dataset.kmersyeast_kmer_splits
