AlhitawiMohammed22/words_hu_dict
This dataset was generated from a Hungarian dictionary, where 60345 sample given The command used to generate data : python3 run.py -i "dicts/hu.txt" -t 8 -f 64 -l hu -c 60345 -na 2 --output_dir "out/words/hu/" --font_dir fonts/hu/ -b 3 -al 0 TRDGHuMu is used for generating text: https://github.com/Mohammed20201991/TextRecognitionDataGeneratorHuMu23
015
upload splited data&renamed
refactor raw gen. data
Delete hu.txt
Delete train.parquet
Delete train.jsonl
Delete train.zip
Delete test.jsonl
Delete test.zip
Upload hu.txt
Upload train.parquet
Upload updated train jsonl file
Update README.md
Update README.md
Upload 2 files
Delete test.jsonl
Delete test.zip
Update test.zip
Update test.jsonl
Create test.zip
Create test.jsonl
Upload 2 files
initial commit
