datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
blip-kd-results-mgd-fashion200kSlovAlpaca
SlovAlapca dataset
This dataset was created using machine translation (DeepL) of the original Alpaca dataset published here: https://github.com/tatsu-lab/stanford_alpaca
Here is an example of the first record...
[
{
"instruction": "Uveďte tri tipy, ako si udržať zdravie.",
"input": "",
"output": "1.Jedzte vyváženú stravu a dbajte na to, aby obsahovala dostatok ovocia a zeleniny. \n2. Pravidelne cvičte, aby ste udržali svoje telo aktívne a silné. \n3.… See the full description on the dataset page: https://huggingface.co/datasets/blip-solutions/SlovAlpaca.blip-kd-results-wsld-at-fashion200k-15kblip-kd-results-mocha-v3-progressive-fashion200k-15kamharic-blip-laionDataset used for pretraining clip alignment step of Amharic llava.
More details: https://arxiv.org/abs/2403.06354
blip-kd-resultsv2-mocha-v4-progressive-fashion200k-15kllava_pretrain_blip_laion_cc_sbu_558k_jablip-kd-results-fitnes-at-fashion200kBLIP3o-60kblip-kd-results-word-at-fashion200kblip-kd-student-attn-distill-fashion200k-15kBLIP-data-reproduceblip2-owlv2-dev-data
