arXiv AI By Raymond Li, Amirhossein Abaskohi, Chuyuan Li, Gabriel Murray, Giuseppe Carenini

Improving Topic Modeling by Distilling Soft Labels from Language Models

Read the original on arXiv AI →

arXiv:2602. 17907v3 Announce Type: replace-cross Abstract: Traditional neural topic models are typically optimized by reconstructing the document's Bag-of-Words (BoW) representations, overlooking contextual information and struggling with data sparsity.

Summary generated by The Flow from the publisher's feed. The full article lives at arXiv AI.