datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Algorithm_and_Python_Source_CodeAlgorithm_and_Python_Source_Code
This dataset provides different algorithms and their corresponding source code in Python.
credits: Source codes given here are taken from "iamtarun/python_code_instructions_18k_alpaca" dataset in Hugging Face.
python-algorithm-sourcecode
Dataset Card for Dataset Name
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Dataset Details
Dataset Description
This dataset provides algorithms and corresponding Python source code which can be leveraged for any type of code conversion applications.
Curated by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]… See the full description on the dataset page: https://huggingface.co/datasets/annawleo/python-algorithm-sourcecode.Multilingual-Therapy-Dialogues
Dataset Summary
Multilingual Therapy Dialogues is a diverse and bilingual dataset consisting of paired dialogues between patients and therapists in both Persian and English.
Dataset Statistics
Number of samples: 7,179
English:
Average tokens per sentence: 101.30
Maximum tokens in a sentence: 939
Average characters per sentence: 567.85
Number of unique tokens: 32,968
Persian:
Average tokens per sentence: 100.06
Maximum tokens in a sentence: 1,413
Average… See the full description on the dataset page: https://huggingface.co/datasets/Algorithmic-Human-Development-Group/Multilingual-Therapy-Dialogues.africa-crypto-algorithm-inventory
Africa Cryptographic Algorithm Inventory | Africa (Electric Sheep Africa metadata inventory)
Size category: n<1K - Formats: csv - Sector: governance_security - Engineered by Electric Sheep Africa
TL;DR
This dataset is part of the Electric Sheep Africa catalog on Hugging Face. It is indexed for African data discovery with standardized metadata, loading guidance, provenance notes, and analyst-oriented context.
What This Dataset Covers
Public… See the full description on the dataset page: https://huggingface.co/datasets/electricsheepafrica/africa-crypto-algorithm-inventory.Sorting-Algorithms-Performance-Metrics
Sorting Algorithms Benchmark Dataset (Array Size: 1000)
A benchmark dataset comparing execution time, memory usage, and comparison counts of various sorting algorithms (Bubble Sort, Selection Sort, Insertion Sort, Merge Sort, Quick Sort, Heap Sort, Odd-Even Sort) on arrays of size 1000. Each algorithm was run 100 times with randomized inputs to ensure statistical significance.
Dataset Details
Columns
run: Trial number (1-100 per algorithm).
algorithm:… See the full description on the dataset page: https://huggingface.co/datasets/ismielabir/Sorting-Algorithms-Performance-Metrics.uk-algorithmic-transparency
UK Algorithmic Transparency Corpus
Every published record under the UK's Algorithmic Transparency Recording
Standard (ATRS), structured as an open corpus: which public bodies have
disclosed which algorithmic and AI tools, and when.
ATRS records
136
Distinct publishing bodies
73
Source
GOV.UK ATRS records
Licence
Open Government Licence v3.0
Why
The ATRS is the UK government's standard for publishing how the public sector
uses algorithmic… See the full description on the dataset page: https://huggingface.co/datasets/fabsssss/uk-algorithmic-transparency.Algorithm-Dataset
