CoolFace
13 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01jinaai /tweet-stock-synthetic-retrieval_beirThis is a copy of https://huggingface.co/datasets/jinaai/tweet-stock-synthetic-retrieval reformatted into the BEIR format. For any further information like license, please refer to the original dataset. Disclaimer This dataset may contain publicly available images or text data. All data is provided for research and educational purposes only. If you are the rights holder of any content and have concerns regarding intellectual property or copyright, please contact us at "support-data… See the full description on the dataset page: https://huggingface.co/datasets/jinaai/tweet-stock-synthetic-retrieval_beir.image1K<n<10K0 likes453 downloads1y agoHugging Face02StephanAkkerman /financial-tweets-crypto Financial Tweets - Cryptocurrency This dataset is part of the scraped financial tweets that I collected from a variety of financial influencers on Twitter, all the datasets can be found here: Crypto: https://huggingface.co/datasets/StephanAkkerman/financial-tweets-crypto Stocks (and forex): https://huggingface.co/datasets/StephanAkkerman/financial-tweets-stocks Other (Tweet without cash tags): https://huggingface.co/datasets/StephanAkkerman/financial-tweets-other Data… See the full description on the dataset page: https://huggingface.co/datasets/StephanAkkerman/financial-tweets-crypto.imagetext-classification10K<n<100K14 likes103 downloads2y agoHugging Face03jinaai /tweet-stock-synthetic-retrieval_deprecated Tweet Stock Document Retrieval This dataset is created from the original Kaggle Tweet Sentiment's Impact on Stock Returns dataset. The tables are rendered and queries created using templates. The text_description column contains OCR text extracted from the images using EasyOCR. This particular dataset is a subsample of at maximum 1000 random rows per language from the full dataset which can be found here. Disclaimer This dataset may contain publicly available images… See the full description on the dataset page: https://huggingface.co/datasets/jinaai/tweet-stock-synthetic-retrieval_deprecated.image10K<n<100K0 likes98 downloads1y agoHugging Face04jinaai /tweet-stock-synthetic-retrieval Tweet Stock Document Retrieval This dataset is created from the original Kaggle Tweet Sentiment's Impact on Stock Returns dataset. The tables are rendered and queries created using templates. The text_description column contains OCR text extracted from the images using EasyOCR. This particular dataset is a subsample of at maximum 1000 random rows per language from the full dataset which can be found here. Disclaimer This dataset may contain publicly available images… See the full description on the dataset page: https://huggingface.co/datasets/jinaai/tweet-stock-synthetic-retrieval.image10K<n<100K0 likes75 downloads1y agoHugging Face05StephanAkkerman /financial-tweets-stocksimage10K<n<100K6 likes60 downloads2y agoHugging Face06StephanAkkerman /financial-tweets Financial Tweets This dataset is a comprehensive collection of all the tweets from my Discord bot that keeps track of financial influencers on Twitter. The data includes a variety of information, such as the tweet and the price of the tickers in that tweet at the time of posting. This dataset can be used for a variety of tasks, such as sentiment analysis and masked language modelling (MLM). We used this dataset for training our FinTwitBERT model. Overview This… See the full description on the dataset page: https://huggingface.co/datasets/StephanAkkerman/financial-tweets.imagetext-classification100K<n<1M12 likes51 downloads2y agoHugging Face07Dhairya /trial-tweets Dataset Card for "trial-tweets" sample dataset of length 240000 image10K<n<100K2 likes49 downloads3y agoHugging Face08fdaudens /musk-tweetsimage1 likes41 downloads1y agoHugging Face09PersianML /persian-tweets-2024 Dataset Description This dataset contains high-engagement Persian language tweets collected from Twitter/X during 2024. The dataset includes comprehensive tweet metadata and user information, making it valuable for various NLP tasks, social media analysis, and Persian language processing research. Dataset Details Size: 900 tweets Language: Persian (Farsi) Time Period: 2024 Collection Criteria: Language: Persian Minimum Likes: 1,000+ Date Range: January 1, 2024… See the full description on the dataset page: https://huggingface.co/datasets/PersianML/persian-tweets-2024.imagetext-classificationn<1K0 likes26 downloads2mo agoHugging Face10mshojaei77 /persian-tweets-2024 Dataset Description This dataset contains high-engagement Persian language tweets collected from Twitter/X during 2024. The dataset includes comprehensive tweet metadata and user information, making it valuable for various NLP tasks, social media analysis, and Persian language processing research. Dataset Details Size: 900 tweets Language: Persian (Farsi) Time Period: 2024 Collection Criteria: Language: Persian Minimum Likes: 1,000+ Date Range: January 1, 2024 -… See the full description on the dataset page: https://huggingface.co/datasets/mshojaei77/persian-tweets-2024.imagetext-classificationn<1K1 likes21 downloads2y agoHugging Face11StephanAkkerman /financial-tweets-otherimage2 likes19 downloads2y agoHugging Face12dejp3 /Fitzwilliam-museum-tweetsimage10K<n<100K0 likes12 downloads11mo agoHugging Face13dejp3 /british-museum-pompeii-live-tweetsimage1K<n<10K0 likes3 downloads11mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.