CoolFace
20 results

neuron

GT-Neuronext /human-motion-tracking-deeplabcutThis dataset is used to adapt DeepLabCut for Human motion tracking. Structure of the dataset videos contains 100+ videos of 4 candidates recorded during a game of darts. labeled-data contains labels on the corresponding frames of the videos. These labels are used to adapt DeepLabCut for human motion tracking. Under labeled-data there are 2 folders for every video. video_name has all the relevant frames extracted from the video, xy coordinates of the labels in the csv file and the… See the full description on the dataset page: https://huggingface.co/datasets/GT-Neuronext/human-motion-tracking-deeplabcut.image1K<n<10K0 likes1.6k downloads3y agoHugging Facemelof1001 /neuron-data1 likes1.1k downloads6mo agoHugging Faceneuronpedia-org /sae-evals0 likes628 downloads2y agoHugging Faceneurondb /postgresql-llm postgresql-llm A pure PostgreSQL dataset for training and evaluating LLMs on PostgreSQL SQL and PL/pgSQL. Every row is a (question, schema, SQL) triplet with rich metadata for filtering and analysis. Dataset Summary postgresql-llm is a pure PostgreSQL dataset: SQL and PL/pgSQL only, with metadata for difficulty, category, and source. Metric Value Total rows 211,539 PostgreSQL-specific rows 11,998 (5.7%) Schema fill rate 82.2% Explanation fill rate 17.8%… See the full description on the dataset page: https://huggingface.co/datasets/neurondb/postgresql-llm.tabulartext-generation100K<n<1M4 likes590 downloads7mo agoHugging FaceBrain2nd /NeuronSpark-V1 NeuronSpark-V1 Pretraining Dataset Bilingual (English + Chinese) pretraining corpus for NeuronSpark, a bio-inspired Spiking Neural Network language model. Dataset Summary Metric Value Total documents 17,174,734 Estimated tokens ~14.5B Languages English (55%), Chinese (42%), Bilingual Math (3%) Format Parquet (35 shards, ~39 GB) Columns text (string), source (string) Sources & Composition Source Documents Ratio Est. Tokens… See the full description on the dataset page: https://huggingface.co/datasets/Brain2nd/NeuronSpark-V1.texttext-generation10M<n<100M0 likes421 downloads6mo agoHugging FaceBrain2nd /NeuronSpark-Pretrain-v3 NeuronSpark-Pretrain-v3 Bilingual pretraining corpus for NeuronSpark v3, a bio-inspired Spiking Neural Network language model with selective PLIF neurons and dynamic per-token compute budget (PonderNet-v3). Composition Metric Value Total documents 18.2 M Estimated tokens ~20 B Format 37 Parquet shards (~1 GB each, zstd) Schema text: string, source: string Languages EN 55.6%, ZH 28.1%, code 16.3% Deduplication All source sampling is weighted so each… See the full description on the dataset page: https://huggingface.co/datasets/Brain2nd/NeuronSpark-Pretrain-v3.texttext-generation10M<n<100M0 likes407 downloads5mo agoHugging Face