shaistaDev7/topic_modeling_on_UDC
06
tags:
- bertopic libraryname: bertopic pipelinetag: text-classification ---
urdutopicmodeling
This is a BERTopic model. BERTopic is a flexible and modular topic modeling framework that allows for the generation of easily interpretable topics from large datasets.
Usage
To use this model, please install BERTopic:
pip install -U bertopicYou can use the model as follows:
from bertopic import BERTopic
topic_model = BERTopic.load("shaistaDev7/urdu_topic_modeling")
topic_model.get_topic_info()Topic overview
- Number of topics: 5
- Number of training documents: 1008
<details> <summary>Click here for an overview of all topics.</summary>
</details>
Training hyperparameters
- calculate_probabilities: True
- language: urdu
- low_memory: True
- mintopicsize: 10
- ngramrange: (1, 1)
- nr_topics: None
- seedtopiclist: None
- topnwords: 10
- verbose: False
- zeroshotminsimilarity: 0.7
- zeroshottopiclist: None
Framework versions
- Numpy: 1.23.5
- HDBSCAN: 0.8.33
- UMAP: 0.5.5
- Pandas: 1.5.3
- Scikit-Learn: 1.2.2
- Sentence-transformers: 2.2.2
- Transformers: 4.35.2
- Numba: 0.58.1
- Plotly: 5.15.0
- Python: 3.10.12
