datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
FINDER_API_KEY_AI_SEARCH_2023
FINDER_API_KEY_AI_SEARCH_2023
tags: data collection, machine learning, API performance
Note: This is an AI-generated dataset so its content may be inaccurate or false
Dataset Description:
The 'FINDER_API_KEY_AI_SEARCH_2023' dataset is designed to collect and analyze data from various AI search engines and their associated API performance metrics. The dataset focuses on the effectiveness of API key-based access in enhancing the search capabilities of AI systems and includes a… See the full description on the dataset page: https://huggingface.co/datasets/infinite-dataset-hub/FINDER_API_KEY_AI_SEARCH_2023.ai-api-pricing
AI API Pricing Dataset
This Hugging Face dataset is the machine-readable distribution of the public AI API pricing records published by AICostBudget. It is not a separately curated subset: train.csv, prices.csv, and prices.json are generated from the same Pricing V2 public projection used by the AICostBudget Dataset page and download APIs.
Prices change frequently. Verify production billing decisions against the provider pricing page, contract, billing dashboard, and invoice.… See the full description on the dataset page: https://huggingface.co/datasets/aicostbudget-ai/ai-api-pricing.BLaIR-Bench-APImozart-api-demo-pages
Dataset Card for Dataset Name
Dataset Summary
[More Information Needed]
Supported Tasks and Leaderboards
[More Information Needed]
Languages
[More Information Needed]
Dataset Structure
Data Instances
[More Information Needed]
Data Fields
[More Information Needed]
Data Splits
[More Information Needed]
Dataset Creation
Curation Rationale
[More Information Needed]
Source Data… See the full description on the dataset page: https://huggingface.co/datasets/DoctorSlimm/mozart-api-demo-pages.vulnerabilidades-ia-espanol
Vulnerabilidades CVE en sistemas de IA (espanol)
Corpus de advisories CVE/GHSA que afectan a paquetes y SDKs de IA (langchain, openai, anthropic, llamaindex, etc.) traducido al espanol por LaAutopsIA (ApisDom Intelligence Group). Fuente principal: GitHub Advisory Database (CC-BY-4.0).
Cifras del snapshot actual
14 vulnerabilidades publicadas en este snapshot.
Mes archivado: 2026-08.
Ultima edicion: 2026-09-01T21:35:37.468Z.
Frecuencia: sincronizacion mensual.… See the full description on the dataset page: https://huggingface.co/datasets/apisdom/vulnerabilidades-ia-espanol.indice-fallos-ia-espanol
Indice de Fallos IA en espanol
Snapshots mensuales del Indice de Fallos IA producido por el observatorio La AutopsIA (ApisDom Intelligence Group). Mide la fiabilidad de modelos LLM con benchmarks oficiales independientes, en formato citable y trazable.
Cifras del snapshot actual
765 mediciones en este snapshot.
Mes archivado: 2026-09.
Recomputado: 2026-09-01T21:34:26.058Z.
Frecuencia: sincronizacion mensual.
Para que sirve este dataset
Datos… See the full description on the dataset page: https://huggingface.co/datasets/apisdom/indice-fallos-ia-espanol.mozart-api
Dataset Card for Dataset Name
Dataset Summary
[More Information Needed]
Supported Tasks and Leaderboards
[More Information Needed]
Languages
[More Information Needed]
Dataset Structure
Data Instances
[More Information Needed]
Data Fields
[More Information Needed]
Data Splits
[More Information Needed]
Dataset Creation
Curation Rationale
[More Information Needed]
Source Data… See the full description on the dataset page: https://huggingface.co/datasets/DoctorSlimm/mozart-api.premier-league-2025-26-football-api-sample
Premier League 2025/26 Football API Sample
A compact, reproducible sample of Premier League fixtures from the opening weeks of the
2025/26 season. It includes results, half-time scores, corners, cards, opening and closing
1X2 prices, and opening/closing Asian handicap, goal and corner lines.
The sample is designed for tutorials, API evaluation and small data-analysis examples. It
is intentionally limited to 50 historical fixtures; it is not a substitute for a current
or complete… See the full description on the dataset page: https://huggingface.co/datasets/5dollarfootballapi/premier-league-2025-26-football-api-sample.covid_fact_checked_google_apiThis dataset was gathered from the Google Fact Checker API, using an automatic web scraper. 10,000 facts were pulled, but for the sake of simplicity, only ones were the ratings were singular words "false" or "true", were kept, which filtered it down to ~3000 fact checks, with about 90% of the facts being false.
annotations_creators:
expert-generated
language_creators:
crowdsourced
languages:
en-US
licenses:
unknown
multilinguality:
monolingual
pretty_name: polifact-covid-fact-checker… See the full description on the dataset page: https://huggingface.co/datasets/justinqbui/covid_fact_checked_google_api.ai-api-pricing-snapshot
AI API Model Pricing Snapshot — Qubax AI
Per-token public pricing for 211 ready models served by the Qubax AI API (OpenAI-compatible), exported from the public /v1/models endpoint.
Columns
Column
Description
model_id
API model identifier
model_name
Display name
owned_by
Publisher namespace
context_length
Max context window (tokens)
input_usd_per_1m_tokens
Input price, USD per 1M tokens
output_usd_per_1m_tokens
Output price, USD per 1M tokens… See the full description on the dataset page: https://huggingface.co/datasets/QubaxAI/ai-api-pricing-snapshot.api-oneshot-summaryToolusing-apiWolframAlpha_API_Questions
WolframAlpha_API_Questions
A curated collection of 276 mathematical, scientific and general‑knowledge questions that can be sent a queries to the Wolfram|Alpha API.The dataset is released under an MIT License and is hosted on Hugging Face for easy reuse in research, education and application development.
Repository URL: https://huggingface.co/datasets/abhibambhaniya/WolframAlpha_API_QuestionsLicense: MIT
Dataset Contents
id,category,sub_category,question,answer
1… See the full description on the dataset page: https://huggingface.co/datasets/abhibambhaniya/WolframAlpha_API_Questions.Nonsense-Internet-Niches
Dataset Card for Dataset Name
[this dataset is basically for internet-cultured LLMs or making a LLM more 21st century humane speech]
cool dataset that has SynthV, Vocaloids, Blender Niches, GTA, literature, Art, ill update it later to have more things soon!1!!
only have like 50s example because only 1 person operating this, the needed components are in file and versions, but you can add others and not mine if you want
In ListsForThings.txt, yes you can actually use it on serious… See the full description on the dataset page: https://huggingface.co/datasets/Apixhed/Nonsense-Internet-Niches.self-reflective-apis
Self-Reflective APIs Benchmark
Dataset accompanying the paper "Self-Reflective APIs: Enhancing AI Agent Efficiency Through Structured Semantic Feedback" (Canedo, Grama — Siemens DI SW).
Overview
This dataset contains the benchmark tasks, experiment results, and tidy analysis table used to produce every table and figure in the paper. It covers two experimental domains (recipe conversion and billing/refund policy) and three LLM conditions across adversarial… See the full description on the dataset page: https://huggingface.co/datasets/arquicanedo/self-reflective-apis.ysda-2026-seminar-yatasks-apiSO-Python_QA-API_Usage-tanh_score
Stack Overflow Python Q&A Dataset
Description
Filtered Python Q&A with API_Usage subcategory without:
Images
Links
Blocks of code
Scores in Q1-Q3 scaled with MaxAbsScaler. Tanh function applyed to joint Scores.
API_Documentation_dataset_alpaancosynthetic-api-traffic-anomalies
📊 Synthetic API Traffic Anomalies Dataset
A synthetic dataset containing normal API transaction records mixed with rate-limit bypasses and brute force attacks.
efeverde_5_cat_lemefeverde
AI_FINDER_API_Efficiency
AI_FINDER_API_Efficiency
tags: algorithm efficiency, machine learning, API optimization
Note: This is an AI-generated dataset so its content may be inaccurate or false
Dataset Description:
The 'AI_FINDER_API_Efficiency' dataset comprises a curated collection of articles, case studies, and technical papers that focus on the performance of AI-powered search and data retrieval systems. It evaluates the efficiency of various algorithms and machine learning models that are employed to… See the full description on the dataset page: https://huggingface.co/datasets/infinite-dataset-hub/AI_FINDER_API_Efficiency.KI_API_Classification_Dataset
KI_API_Classification_Dataset
tags: classification, machine learning, key API ki AI, security
Note: This is an AI-generated dataset so its content may be inaccurate or false
Dataset Description:
The 'KI_API_Classification_Dataset' contains texts that have been curated to evaluate the performance of a machine learning model in classifying documents based on the relevance to the keywords 'API KEY' and 'KI AI'. Each document has been manually labeled with one of two categories:… See the full description on the dataset page: https://huggingface.co/datasets/infinite-dataset-hub/KI_API_Classification_Dataset.genesys_api_http_alpacaCardiovascular_Diseaseapi-zeroshot-summaryexchange-api-latency
Crypto Exchange REST API Latency Benchmark
Open, reproducible latency benchmark for the public REST APIs of major crypto exchanges: Coinbase, Kraken, Gemini, Crypto.com, Bitfinex, and KuCoin.
Live site and full context: fillbench.com/exchange-api-latency
How it is measured
Each run holds one persistent keep-alive HTTPS connection per exchange, warms it up, then times back-to-back GET requests to a small public market-data endpoint (no API keys, no auth, no… See the full description on the dataset page: https://huggingface.co/datasets/Carlo-fillbench/exchange-api-latency.clean_newsClean News Spain
efeverdenoticias medioambiente
api-fifshot-summarynoticias
