THULab/niederschlag_station_4091_bavaria_germany
Niederschlag Station 4091, Bavaria, Germany (TsFile) This dataset is the Apache TsFile conversion of Kamilatr/niederschlag_station_4091_bavaria_germany. It contains the monthly precipitation-height observations published for station 4091. Modalities: Time-series. Overview Source dataset: Kamilatr/niederschlag_station_4091_bavaria_germany Source revision: 7445f38114e2cc07f63f4fb3d16fb53cd2fddf45 Source file: newdf.csv Observations: 60 monthly rows Converted file:… See the full description on the dataset page: https://huggingface.co/datasets/THULab/niederschlag_station_4091_bavaria_germany.
Niederschlag Station 4091, Bavaria, Germany (TsFile)
This dataset is the Apache TsFile conversion of `Kamilatr/niederschlag_station_4091_bavaria_germany`. It contains the monthly precipitation-height observations published for station 4091.
Modalities: Time-series.
Overview
- Source dataset: `Kamilatr/niederschlag_station_4091_bavaria_germany`
- Source revision:
7445f38114e2cc07f63f4fb3d16fb53cd2fddf45 - Source file:
newdf.csv - Observations: 60 monthly rows
- Converted file:
niederschlag_station_4091_bavaria_germany.tsfile(1,026 bytes) - Date range: 1986-01-01 through 1990-12-01
- Station: 4091 (Bavaria, Germany)
- Split:
train
The source card does not declare a license or a measurement unit; this card therefore does not add either. Values are copied exactly as published.
TsFile schema
The table is niederschlag_station_4091_bavaria_germany and has one device, the station TAG station_id = "4091".
Conversion notes
- The source
Date(YYYY-MM-DD) is parsed toTimeat UTC midnight. The redundant source string is not duplicated as a FIELD. - The dotted source measurement name is normalized to the TsFile-safe field
mo_rr_niederschlagshoehe; no value or unit conversion is performed. - All 60 observations are retained and sorted by
station_id, Time.
Read example
from tsfile import TsFileReader
path = "niederschlag_station_4091_bavaria_germany.tsfile"
with TsFileReader(path) as reader:
with reader.query_table(
"niederschlag_station_4091_bavaria_germany",
["mo_rr_niederschlagshoehe"],
batch_size=128,
) as result:
batch = result.read_arrow_batch()
if batch is not None:
print(batch.to_pandas())Source & license
- Original dataset: https://huggingface.co/datasets/Kamilatr/niederschlagstation4091bavariagermany
- Author / publisher: Kamilatr
- License: not declared by the original dataset; please defer to the original.
Usage
Install the Apache TsFile Python SDK (pip install tsfile) and read a converted file:
from pathlib import Path
from tsfile import TsFileReader
path = Path("niederschlag_station_4091_bavaria_germany.tsfile")
with TsFileReader(str(path)) as reader:
schemas = reader.get_all_table_schemas()
print("tables:", list(schemas))
table_name = next(iter(schemas))
table = schemas[table_name]
columns = [column.get_column_name() for column in table.get_columns()]
print("columns:", columns)
field_names = [
column.get_column_name()
for column in table.get_columns()
if column.get_column_name() not in {"Time", "time"}
]
if field_names:
with reader.query_table(table_name, field_names[:3], batch_size=1024) as result:
batch = result.read_arrow_batch()
if batch is not None:
print(batch.to_pandas().head())