datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
shades_nationalityPossibly a placeholder dataset for the original here: https://huggingface.co/datasets/bigscience-catalogue-data/bias-shades
Data Statement for SHADES
How to use this document:
Fill in each section according to the instructions. Give as much detail as you can, but there's no need to extrapolate. The goal is to help people understand your data when they approach it. This could be someone looking at it in ten years, or it could be you yourself looking back at the data in two years.… See the full description on the dataset page: https://huggingface.co/datasets/bigscience-catalogue-data/shades_nationality.BigSolDBv2.1
BigSolDB v2.1
BigSolDB v2.1 contains 112,465 experimentally measured solubility values of 1,525 organic compounds in 218 solvents, reported in 1,687 peer-reviewed literature sources.
The dataset is designed for data-driven solubility modeling, benchmarking, and solvent selection tasks.
📊 Dataset Structure
The dataset consists of 12 columns, defined as follows:
SMILES_Solute — SMILES representation of the solute molecule
Temperature_K — temperature of the reported… See the full description on the dataset page: https://huggingface.co/datasets/levakrasnov/BigSolDBv2.1.AbstractReasoningbig_situations
