CoolFace
Datasetpublic

shubhamg2208/lexicap

Lexicap contains the captions for every Lex Friedman Podcast episode. It it created by [Dr. Andrej Karpathy](https://twitter.com/karpathy). There are 430 caption files available. There are 2 types of files: - large - small Each file name follows the format `episode_{episode_number}_{file_type}.vtt`.

sourceHugging Faceupdated 4y agoView on Hugging Face
0likes47downloads
Dataset Card

Dataset Card for Lexicap

Table of Contents

Dataset Description

  • —

Dataset Structure

Data Instances

Train and test dataset. j

Data Fields

Dataset Creation

Curation Rationale

[More Information Needed]

Source Data

Initial Data Collection and Normalization

[More Information Needed]

Who are the source language producers?

[More Information Needed]

Annotations

Annotation process

[More Information Needed]

Who are the annotators?

[More Information Needed]

Personal and Sensitive Information

[More Information Needed]

Considerations for Using the Data

Social Impact of Dataset

[More Information Needed]

Discussion of Biases

[More Information Needed]

Other Known Limitations

[More Information Needed]

Additional Information

Dataset Curators

[More Information Needed]

Licensing Information

[More Information Needed]

Citation Information

Contributions