dknoxon/nyc-tlc-yellow-2024-25
NYC TLC Yellow Taxi 2024 Monthly parquet files partitioned by year/month for easier notebook filtering. Layout year=2024/month=01/yellow_tripdata_2024-01.parquet year=2024/month=02/yellow_tripdata_2024-02.parquet ... year=2024/month=12/yellow_tripdata_2024-12.parquet Example usage from datasets import load_dataset ds = load_dataset( 'dknoxon/nyc-tlc-yellow-2024-25', data_files='year=2024/month=0[3-6]/*.parquet', split='train' ) Or… See the full description on the dataset page: https://huggingface.co/datasets/dknoxon/nyc-tlc-yellow-2024-25.
NYC TLC Yellow Taxi 2024
Monthly parquet files partitioned by year/month for easier notebook filtering.
Layout
year=2024/month=01/yellow_tripdata_2024-01.parquetyear=2024/month=02/yellow_tripdata_2024-02.parquet- ...
year=2024/month=12/yellow_tripdata_2024-12.parquet
Example usage
from datasets import load_dataset
ds = load_dataset(
'dknoxon/nyc-tlc-yellow-2024-25',
data_files='year=2024/month=0[3-6]/*.parquet',
split='train'
)Or with pandas/pyarrow directly using the raw file URLs from the Hub.
