granite
Datasets
All datasets matching “granite”ChartNet
ChartNet: A Million-Scale Multimodal Dataset for Chart Understanding
🌐 Homepage | 📖 arXiv
📝 Changelog
June 3, 2026 — Release of grounded_qa subset and completed reasoning subset (both subject to Notice Regarding Data Availability)
May 15, 2026 — Added link to 30K real-world charts and detailed captions dataset released by our collaborators Abaka AI/2077AI.
April 29, 2026 — Release of an additional 2.5 million row subset core_permissive (subject to… See the full description on the dataset page: https://huggingface.co/datasets/ibm-granite/ChartNet.GneissWeb
What is it?
Recipe for producing a state-of-the-art LLM pre-training dataset having 10+ Trillion tokens, derived from FineWeb V1.1.0
Evaluation results showing more than 2% avg improvement (with multiple random seeds) over FineWeb V1.1.0 tokens on common benchmarks for a 7B parameter ablation model
Data Prep Kit Notebook for reproducing the annotations and filters on top of FineWeb and Notebook for applying a bloom filter on FineWeb to quickly reproduce an approximate version of… See the full description on the dataset page: https://huggingface.co/datasets/ibm-granite/GneissWeb.granitewildchat-4.8m_1m_seed1_gemma_granite_metrics_extendedGranite-LLM-model-Fine-tuned-psychology-filosofi-and-romanceNeurvance Granite 30B – Psychology, Philosophy & Romance
Model Description
This model is a 30B-parameter Granite-based language model fine-tuned by Neurvance with a focus on:
Psychology
Philosophy
Romance and relationships
Human behavior
Emotional and reflective conversations
Deeper conversational reasoning
The goal of the model is to provide more nuanced, thoughtful and human-centered responses in conversations involving emotions, relationships, philosophical questions and psychological… See the full description on the dataset page: https://huggingface.co/datasets/WalkerDK/Granite-LLM-model-Fine-tuned-psychology-filosofi-and-romance.granite-decisions-synthetic
Granite Decisions synthetic datasets
Original, deterministic English fixtures for Adam Pippert's personal
Granite Decisions project.
The original default config has 162 examples: 54 train, 54 calibration, and 54 test.
These exercise the pipeline; they are not a representative quality benchmark.
Source and license
The source is the project's original template generator, published here as
make_smoke_data.py, from
release v0.1.0,
commit… See the full description on the dataset page: https://huggingface.co/datasets/adampippert/granite-decisions-synthetic.
