datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ipfs_austria_laws_ir
Austria legislation IR (CID-keyed sparse GraphRAG)
Research retrieval release of endomorphosis/ipfs_austria_laws (revision 096bd65e89d3247debcfee4d51541b37d14b4def) packaged as
country-laws-ir-graphrag/v1 (layout family skillcenter-huggingface-release/v3 / publicus-ir).
Not legal advice. This is a research snapshot. The official gazette /
authentic source of Austria prevails over this corpus. Retrieved documents
and graph edges are retrieval evidence only. No legal text was… See the full description on the dataset page: https://huggingface.co/datasets/justicedao/ipfs_austria_laws_ir.ipfs_austria_laws
Austria RIS Bundesrecht (OGD)
Research snapshot of official national legislation from RIS Bundesrecht konsolidiert (data.bka.gv.at RIS OGD / ogd.ris.bka.gv.at / ris.bka.gv.at).
Not legal advice. The official gazette / authentic source prevails over this corpus.
Snapshot
Field
Value
Snapshot date
2026-09-24
Coverage
deepen complete (RIS OGD paragraph XML)
Source
RIS Bundesrecht konsolidiert (data.bka.gv.at RIS OGD / ogd.ris.bka.gv.at / ris.bka.gv.at)… See the full description on the dataset page: https://huggingface.co/datasets/endomorphosis/ipfs_austria_laws.NewsEye-Austrian-line
NewsEye Austrian - line level
Dataset Summary
The dataset comprises Austrian newspaper pages from 19th and early 20th century. The images were provided by the Austrian National Library.
Languages
The documents are in Austrian German with the Fraktur font.
Note that all images are resized to a fixed height of 128 pixels.
Dataset Structure
Data Instances
{
'image': <PIL.JpegImagePlugin.JpegImageFile image mode=RGB size=4300x128 at… See the full description on the dataset page: https://huggingface.co/datasets/Teklia/NewsEye-Austrian-line.BigEarthNet_S2_Austriaaustria-flat-prices-by-city-willhaben-2026
Austria Flat Prices by City — Willhaben.at, September 2026
Median asking prices for flats on sale in Austria, collected on 1 September 2026.
Two levels: 28 cities, and all of Vienna broken down by district.
What is inside
File
Rows
What it is
willhaben-at-market-summary-2026-09-01.csv
28
One row per city with at least 5 priced listings
willhaben-at-vienna-districts-2026-09-01.csv
21
Vienna by Bezirk — median price, price per m2, size, rooms… See the full description on the dataset page: https://huggingface.co/datasets/Anatolii2026/austria-flat-prices-by-city-willhaben-2026.austria-consumer-email-dataset
Austria Consumer Email Database — Market Intelligence Dataset
5,500,000 verified Austria contacts are available from LeadsBlue →. This open dataset provides the aggregate market intelligence behind that database — contact volume, benchmark open/reply rates, send timing, and compliance for the Austria segment.
At a glance: A research dataset describing the Austria Consumer Email Database market: verified-contact volume, industry distribution, outreach benchmarks, and… See the full description on the dataset page: https://huggingface.co/datasets/emailmarketingdataset/austria-consumer-email-dataset.liberty-archive-austrian-economics
The Liberty Archive: Austrian Economics Corpus
Austrian economics in bulk: 19,531,994 words across two splits.
Full text of the Creative Commons licensed books in the Liberty Archive, one row
per chapter. 2,582 chapters from 176 books, 7,132,091 words.
Machine transcripts of 2,701 recorded lectures, 12,399,903 words,
sit in a second split with a different and weaker rights basis than the books.
Read "Lecture rights: what is and is not established" before using them. Books and… See the full description on the dataset page: https://huggingface.co/datasets/davidveksler/liberty-archive-austrian-economics.austria_orthobutterflies-moths-austria
Butterflies & Moths Austria
Dataset Summary
This is a repackaged version of the Austria butterflies and moths dataset in PyTorch ImageFolder format.
Note: All credit goes to the original authors.
Compared to the original release, this upload:
Converts all images to WebP
Resizes each image so that the total number of pixels < 589,824 (=768×768), preserving aspect ratio
Pre-splits the dataset into train/validation/test using a 70:20:10 split
Packs the data in an… See the full description on the dataset page: https://huggingface.co/datasets/birder-project/butterflies-moths-austria.Prada.Product.prices.Austria
Prada web scraped data
About the website
The fashion industry, particularly the luxury fashion segment, exhibits a vast and dynamic scope in the Europe, Middle East, and Africa (EMEA) region, with Austria playing a crucial role in its positive trajectory. Prada, a prominent luxury fashion icon, continues to thrive in Austrias competitive market. The industry is significantly propelled by advancements in technology, leading to a surge in Ecommerce platforms. A recent… See the full description on the dataset page: https://huggingface.co/datasets/DBQ/Prada.Product.prices.Austria.Net.a.Porter.Product.prices.Austria
Net-a-Porter web scraped data
About the website
The EMEA industry, particularly in Austria, is a dynamic and competitive marketplace. It is characterized by a rapidly growing e-commerce sector, driven by technological advancements and changing consumer behavior. A significant player in this space is Net-a-Porter, a leading global online luxury fashion retailer. The observed dataset provides detailed Ecommerce product-list page (PLP) data on Net-a-Porter in Austria. The… See the full description on the dataset page: https://huggingface.co/datasets/DBQ/Net.a.Porter.Product.prices.Austria.algeria-vs-austria-2026-world-cup-100xAustria-Stock-Symbols-and-Metadata
Austria Stock Symbols & Company Metadata
This dataset contains stock symbols and basic company metadata for all listed companies in Austria.It is updated weekly if new changes are there.
📊 Dataset Contents
The dataset is provided as a CSV file with the following columns:
Column
Description
name
Full company name
ticker
Stock ticker symbol (e.g., AAPL, MSFT)
market
The exchange/market where the stock is listed
sector
The primary business sector of the… See the full description on the dataset page: https://huggingface.co/datasets/kjhq/Austria-Stock-Symbols-and-Metadata.austria-business-dataset
Austria Business Email List — Market Intelligence Dataset
300,000 verified Austria contacts are available from LeadsBlue →. This open dataset provides the aggregate market intelligence behind that database — contact volume, benchmark open/reply rates, send timing, and compliance for the Austria segment.
At a glance: This dataset provides aggregated market-intelligence on the Austria Business Email List segment — real business-population statistics, cold-email performance… See the full description on the dataset page: https://huggingface.co/datasets/emailmarketingdataset/austria-business-dataset.Austrian_Armed_Forces_Military_Dictionaries
[!NOTE]
Dataset origin: https://elrc-share.eu/repository/browse/austrian-armed-forces-military-dictionaries/de6c43db1e8611e6b68800155d020502bd63f01fb76f424697252f223684c93d/
Description
A collection of military dictionaries in the following language pairs:Austrian German-HungarianAustrian German-ItalianAustrian German-EnglishAustrian German-French
License
Used for resources that fall under the scope of PSI (Public Sector Information) regulations, and for which no… See the full description on the dataset page: https://huggingface.co/datasets/FrancophonIA/Austrian_Armed_Forces_Military_Dictionaries.austrian-passports
Disclaimer: All passport images and associated data in this dataset are synthetically generated and do not correspond to real individuals. Any names, numbers, or personal details are fictional and used solely for research and development purposes.
Introduction - Austria
The Synthetic Austria Passports Dataset compiles more than 1,000 AI-generated passport images created for training OCR and computer vision models on identity documents. Each record is fully synthetic, so the… See the full description on the dataset page: https://huggingface.co/datasets/ud-synthetic/austrian-passports.austria_data
Dataset Card for "austria_data"
More Information needed
Gucci.Product.prices.Austria
Gucci web scraped data
About the website
The luxury fashion industry within the EMEA region, particularly in Austria, has been greatly influenced by digitalization and the rise of e-commerce. The increased internet penetration and smartphone usage in this region have significantly boosted online luxury purchases. A notable player in this industry is the globally recognized Italian luxury brand, Gucci. Gucci operates within a very digital savvy and contemporary marketplace… See the full description on the dataset page: https://huggingface.co/datasets/DBQ/Gucci.Product.prices.Austria.austrian-german-instructions
AT-Instruct: Austrian German Instructions
500 instruction-response pairs written in Austrian German. Not translated from English or Bundesdeutsch — written from scratch with Austrian vocabulary, institutions, and perspective.
Why this exists
Every German instruction dataset I found was either translated from English (losing all cultural context) or written in Bundesdeutsch. If you fine-tune on those, your model will tell users to go to the "Bürgeramt" — which doesn't… See the full description on the dataset page: https://huggingface.co/datasets/Laborator/austrian-german-instructions.austrian-german-benchmark
AT-Bench: Austrian German Benchmark
300 multiple-choice questions testing whether an LLM actually understands Austrian German — not just German.
The problem
Every German benchmark treats German as one language. But ask GPT what "Obers" means and half the time it guesses wrong. Ask it about the Bezirksgericht and it describes the German court system. Austrian German is an official language variety spoken by 9 million people, and models consistently get it wrong.… See the full description on the dataset page: https://huggingface.co/datasets/Laborator/austrian-german-benchmark.austrian-german-dictionary
Austrian German — Standard German Dictionary
I built this because every time I tried to use a German NLP model on Austrian text, it choked on the basics. Erdapfel isn't Kartoffel, Obers isn't Sahne, and a Meldezettel sure as hell isn't a Meldebescheinigung. So I sat down and compiled 2000+ word pairs with examples in both varieties.
What's in it
2,042 entries across 7 categories. Each entry has the Austrian word, the Standard German equivalent, an English translation, and… See the full description on the dataset page: https://huggingface.co/datasets/Laborator/austrian-german-dictionary.
