datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
industrial-agent-benchmark
Industrial Agent Benchmark
Industrial Agent Benchmark (IAB) is an open benchmark for evaluating Industrial AI systems, Manufacturing AI assistants, and Industrial Agents.
This Dataset Card describes the Hugging Face Dataset release for Industrial Agent Benchmark v2.2.0 Japanese Canonical Normalization.
Repository:
https://github.com/masahirosakae/industrial-agent-benchmark
Hugging Face Dataset Repository:
https://huggingface.co/datasets/MSakae/industrial-agent-benchmark… See the full description on the dataset page: https://huggingface.co/datasets/MSakae/industrial-agent-benchmark.awesome-chatgpt-prompts
a.k.a. Awesome ChatGPT Prompts
This is a Dataset Repository mirror of prompts.chat — a social platform for AI prompts.
📢 Notice
This Hugging Face dataset is a mirror. For the latest prompts, features, and community contributions, please visit:
🌐 Website: prompts.chat
📦 GitHub: github.com/f/awesome-chatgpt-prompts
About
prompts.chat is an open-source platform where users can share, discover, and collect AI prompts from the community. The project can be… See the full description on the dataset page: https://huggingface.co/datasets/msantiiisocial/awesome-chatgpt-prompts.derja_to_msa_dataset
Dataset Card
MADAR (License): We construct 12,000 translation instructions using the Multi-Arabic Dialect Applications and Resources (MADAR) corpus is a collection of parallel sentences covering the dialects of 25 Arab cities. We select the dialect of Tunis city as Derja, along with MSA resulting into two translation directions.
ff-challenge-res
Tested Model Information
Model Name: SmolVLM-Base (2B Parameters)
Model Link: https://huggingface.co/HuggingFaceTB/SmolVLM-Base
Model Type: Multimodal Vision-Language Base Model
Loading Methodology & Python Code
I loaded the model using a Google Colab T4 GPU. To accommodate the 2B parameters within a 16GB VRAM limit,
the model was loaded in half precision like torch.float16 and mapped to the GPU using device_map="auto".
Inference was conducted using greedy decoding… See the full description on the dataset page: https://huggingface.co/datasets/msaleem-aisci/ff-challenge-res.splitter-dataset
Splitter Dataset
Bu dataset, matematiksel grafik sorularını adım adım çözüm yaklaşımıyla işlemek için tasarlanmıştır. Her görsel için, soruyu mantıksal alt sorulara (sub-questions) ayıran bir "Question Decomposer" sistemi geliştirilmiştir.
🎯 Amaç
Model, herhangi bir matematik grafiği sorusunu çözmeye çalışmadan önce, soruyu bir bütün olarak ele alıp mantıklı ve anlamlı adımlara ayırır. Bu yaklaşım, karmaşık soruların daha kolay anlaşılmasını ve çözülmesini sağlar.… See the full description on the dataset page: https://huggingface.co/datasets/Msalcann/splitter-dataset.
