CoolFace
Modelpublic

asoria/transformers_issues_topics

sourceHugging Faceupdated 2y agoView on Hugging Face
0likes6downloads
Model Card

tags:

  • bertopic libraryname: bertopic pipelinetag: text-classification ---

transformersissuestopics

This is a BERTopic model. BERTopic is a flexible and modular topic modeling framework that allows for the generation of easily interpretable topics from large datasets.

Usage

To use this model, please install BERTopic:

pip install -U bertopic

You can use the model as follows:

python
from bertopic import BERTopic
topic_model = BERTopic.load("asoria/transformers_issues_topics")

topic_model.get_topic_info()

Topic overview

  • Number of topics: 30
  • Number of training documents: 9000

<details> <summary>Click here for an overview of all topics.</summary>

Topic IDTopic KeywordsTopic FrequencyLabel
-1pytorch - tensorflow - bert - tf - pretrained15-1pytorchtensorflowberttf
0bert - bertforsequenceclassification - berttokenizer - bart - batchencodeplus23210bertbertforsequenceclassificationberttokenizerbart
1cuda - memory - trainertrain - tensorflow - trainer15541cudamemorytrainertraintensorflow
2transformerscli - transformers - transformer - importerror - transformerxl8822transformersclitransformerstransformerimporterror
3modelcard - modelcards - card - model - models4903modelcardmodelcardscardmodel
4gpt2 - gpt2tokenizer - gpt2xl - gpt2tokenizerfast - gpt2model4624gpt2gpt2tokenizergpt2xlgpt2tokenizerfast
5attributeerror - typeerror - valueerror - runtimeerror - indexerror4375attributeerrortypeerrorvalueerrorruntimeerror
6typos - typo - doc - docstring - fix3366typostypodocdocstring
7t5 - t5model - t5base - tf - t5large2987t5t5modelt5basetf
8readmemd - readmetxt - readme - modelcard - file2708readmemdreadmetxtreadmemodelcard
9ci - testing - tests - test - speedup2549citestingteststest
10s2s - s2sdistill - s2t - s2strainer - exampless2s24510s2ss2sdistills2ts2strainer
11glue - gluepy - glueconvertexamplestofeatures - roberta - huggingfacetransformers21411gluegluepyglueconvertexamplestofeaturesroberta
12ner - pipeline - pipelines - nerpipeline - fillmaskpipeline15812nerpipelinepipelinesnerpipeline
13rag - ragtokenforgeneration - ragsequenceforgeneration - clean - tests15313ragragtokenforgenerationragsequenceforgenerationclean
14questionansweringpipeline - questionanswering - answering - tfalbertforquestionanswering - questionasnwering14314questionansweringpipelinequestionansweringansweringtfalbertforquestionanswering
15onnx - 04onnxexport - 04onnxexportipynb - aionnx - sphynx13115onnx04onnxexport04onnxexportipynbaionnx
16longformer - longformers - longform - longformerlayer - longformermodel10416longformerlongformerslongformlongformerlayer
17labelsmoothednllloss - label - labelsmoothingfactor - labels - labelsmoothing7617labelsmoothednlllosslabellabelsmoothingfactorlabels
18benchmark - benchmarking - benchmarks - accuracy - evaluation7318benchmarkbenchmarkingbenchmarksaccuracy
19wav2vec2 - wav2vec - wav2vec20 - wav2vec2forctc - wav2vec2xlrswav2vec26719wav2vec2wav2vecwav2vec20wav2vec2forctc
20flax - flaxelectraformaskedlm - flaxelectraforpretraining - flaxjax - flaxelectramodel5120flaxflaxelectraformaskedlmflaxelectraforpretrainingflaxjax
21configpath - configs - config - configuration - modelconfigs4921configpathconfigsconfigconfiguration
22logging - logs - log - logger - loghistory4022logginglogsloglogger
23cachedir - cache - cachedpath - caching - cached3823cachedircachecachedpathcaching
24wandbproject - wandb - sagemaker - sagemakertrainer - wandbcallback3624wandbprojectwandbsagemakersagemakertrainer
25notebook - notebooks - community - colab - t53325notebooknotebookscommunitycolab
26electra - electrapretrainedmodel - electraformaskedlm - electraformultiplechoice - electrafortokenclassification3026electraelectrapretrainedmodelelectraformaskedlmelectraformultiplechoice
27layoutlm - layout - layoutlmtokenizer - layoutlmbaseuncased - tf2527layoutlmlayoutlayoutlmtokenizerlayoutlmbaseuncased
28pplm - pr - deprecated - variable - ppl1528pplmprdeprecatedvariable

</details>

Training hyperparameters

  • calculate_probabilities: False
  • language: english
  • low_memory: False
  • mintopicsize: 10
  • ngramrange: (1, 1)
  • nr_topics: 30
  • seedtopiclist: None
  • topnwords: 10
  • verbose: True
  • zeroshotminsimilarity: 0.7
  • zeroshottopiclist: None

Framework versions

  • Numpy: 1.26.4
  • HDBSCAN: 0.8.38.post1
  • UMAP: 0.5.6
  • Pandas: 2.1.4
  • Scikit-Learn: 1.5.2
  • Sentence-transformers: 3.1.1
  • Transformers: 4.44.2
  • Numba: 0.60.0
  • Plotly: 5.24.1
  • Python: 3.10.12