Topic extraction and categorization using a glove based technique for representing topics
preprint
OA: closed
CC-BY-4.0
Abstract
Topic extraction and categorization is an important task because by doing that it is easy to find out which are the topics most discussed by the users in their tweets or opinions and need to be analyzed. In this work, topics are extracted from positive and negative opinions and then categorized into different groups. For performing this, first a collection of opinions are divided into two sets- positive opinions and negative opinions by using a sentiment analyzer. Then a method is proposed to find out the most discussed topics in the set of positive opinions and negative opinions. For extracting the topics from a set of opinions the noun words are extracted from the set of the opinions. After extracting the topics, the similar topics have been combined by using synonymy relation. Then the frequent topic words are represented with the help of GloVe embedding technique. Finally, the topics are categorized by using a clustering algorithm by applying it on the frequent topic words. For the evaluation of the proposed method, tweets from a Twitter User dataset are used. The results obtained from the experiments by applying the proposed method on the dataset give promising result and provide interesting and meaningful clusters of topics. Moreover an analysis of the result obtained for both positive and negative opinions is also presented.
My notes (saved in your browser only)
Citation neighborhood (no data yet)
We don't have any in-corpus citations linked to this paper yet. The paper's references may be in our DB but unresolved to ``paper_id`` (resolution happens at ingest when the cited DOI matches a row we already have). Run the cross-source citation reconcile pass to retry.
Source provenance
- europepmc
- last seen: 2026-05-19T01:45:01.086888+00:00
- unpaywall
- last seen: 2026-05-26T02:00:01.498150+00:00
License: CC-BY-4.0