Datasets › Experiment-data-for-UM-S-TM

Experiment-data-for-UM-S-TM

Introduced by Gangli Liu in Topic Model Supervised by Understanding Map12 Oct 2021 archive 2025-07-28

0.This is experiment data for the following article: @misc{liu2021topic, title={Topic Model Supervised by Understanding Map}, author={Gangli Liu}, year={2021}, eprint={2110.06043}, archivePrefix={arXiv}, primaryClass={cs.CL} }

  1. *.txt files are the data of Table 4 of the paper.

  2. The top lines of all the *.txt files are contents of the artificial documents. Column names are : "Topic", "Distance", "Topic-len", "alpha"/"Noise" , "doc concept-length", and "Votes counter".

3.Coding of file names of *.txt files see "Table 4: Discovered SCOM of six documents". "all_topic" means the candidate topic set is all the topics in a domain.

4.For the "300docs-mentioned-in-section3.2.xlsx" file, its name tells its contents.

Benchmarks archive 2025-07-28

No leaderboard in the archive resolves to this dataset.

Papers archive 2025-07-28

No paper in the archive has a leaderboard row on this dataset; the archive counts 1 paper for it but never published that list.

Dataset loaders archive 2025-07-28

No loader listed in the archive.

Tasks archive 2025-07-28

No task tagged in the archive.

License archive 2025-07-28

No licence recorded in the archive. Absence here is not a statement about the dataset's terms.

Modalities archive 2025-07-28

No modality tagged.

Languages archive 2025-07-28

No language tagged.

Variants archive 2025-07-28

  • Experiment-data-for-UM-S-TM

1 variant name, as the archive lists them.

Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections