CoolFace
21 results

vit

leduytho /vitra-ego4d-videovideo1K<n<10K4 likes7.4k downloads3mo agoHugging Facexincan /Llama-VITS_data Dataset Card for Llama-VITS_data The dataset repository contains data related with our work "Llama-VITS: Enhancing TTS Synthesis with Semantic Awareness", encapsulating: Filtered dataset EmoV_DB_bea_sem Filelists with semantic embeddings Model checkpoints Human evaluation templates Dataset Details Paper: Llama-VITS: Enhancing TTS Synthesis with Semantic Awareness Curated by: Xincan Feng, Akifumi Yoshimoto Funded by: CyberAgent Inc Repository:… See the full description on the dataset page: https://huggingface.co/datasets/xincan/Llama-VITS_data.text-to-speech2 likes4.7k downloads2y agoHugging FaceShiym /ViT-FineTuneimageimage-classification10K<n<100K0 likes4.2k downloads2y agoHugging Facetals /vitaminc Details Fact Verification dataset created for Get Your Vitamin C! Robust Fact Verification with Contrastive Evidence (Schuster et al., NAACL 21`) based on Wikipedia edits (revisions). For more details see: https://github.com/TalSchuster/VitaminC When using this dataset, please cite the paper: BibTeX entry and citation info @inproceedings{schuster-etal-2021-get, title = "Get Your Vitamin {C}! Robust Fact Verification with Contrastive Evidence", author =… See the full description on the dataset page: https://huggingface.co/datasets/tals/vitaminc.texttext-classification100K<n<1M11 likes3.3k downloads4y agoHugging FaceVITRA-VLA /VITRA-1M VITRA-1M: Human Hand V-L-A Dataset Dataset Summary VITRA-1M is a large-scale Human Hand Visual-Language-Action (V-L-A) dataset constructed as described in the paper Scalable Vision-Language-Action Model Pretraining for Robotic Manipulation with Real-Life Human Activity Videos. It contains 1.2 million short episodes with segmented language annotations, camera parameters (corrected intrinsics/extrinsics), and 3D hand reconstructions (left and right… See the full description on the dataset page: https://huggingface.co/datasets/VITRA-VLA/VITRA-1M.1M<n<10M29 likes2.9k downloads10mo agoHugging FaceMTDoven /ViTTiny1022 ViTTiny1022 The dataset for Scaling Up Parameter Generation: A Recurrent Diffusion Approach. Requirement Install torch and other dependencies conda install pytorch==2.3.1 torchvision==0.18.1 torchaudio==2.3.1 pytorch-cuda=12.1 -c pytorch -c nvidia pip install timm einops seaborn openpyxl Usage Test one checkpoint cd ViTTiny1022 python test.py ./chechpoint_test/0000_acc0.9613_class0314_condition_cifar10_vittiny.pth # python test.py… See the full description on the dataset page: https://huggingface.co/datasets/MTDoven/ViTTiny1022.1K<n<10K2 likes1.8k downloads2y agoHugging Face

Projects