Beyond the Follow Button: Classifying User Intent via Issue Clusters and Influential Supporters

Follower Classification Based on User Behavior for Issue Clusters

2013-12-14
Kwang-Yong Jeong, Jae-Wook Seol, Kyung Soon Lee
Summary
Problem
Method
Results
Takeaways
Abstract

The paper introduces a novel follower classification framework that categorizes Twitter users into supporters, non-supporters, or neutrals relative to a target "authority" user. The core approach utilizes an "Issue Cluster" methodology combined with a modified HITS algorithm to identify influential supporters and track sentiment alignment over specific "trust periods."

TL;DR

On platforms like Twitter, a "follow" does not always equate to "support." This paper proposes a robust classification system that identifies whether followers are supporters, non-supporters, or neutral by analyzing Issue Clusters and Influential Supporters. By moving beyond simple text-based SVMs and looking at the network dynamics of retweets during specific "trust periods," the authors achieved a 20.3% increase in classification accuracy.

Background: The "Follower" Ambiguity

In the landscape of Social Network Services (SNS), authority users—such as politicians or celebrities—amass followers with vastly different agendas. Some follow to cheer, others to criticize, and many just to observe. Standard sentiment analysis often fails here because tweets are short and frequently lack explicit emotional keywords. The authors argue that to truly understand a follower, we must look at who they retweet and when they do it in relation to specific social issues.

Methodology: The Three Pillars of Classification

1. Identifying the Inner Circle (Influential Supporters)

The authors utilize a modified HITS (Hyperlink-Induced Topic Search) algorithm. They define "Influential Supporters" as users who consistently amplify the target user's message.

  • Logic: If User A retweets the target user frequently, and User B retweets User A, User A gains "Authority" and "Hub" scores.
  • Expansion: These influential supporters act as an "expanded concept" of the target user. If a random follower retweets an influential supporter, they are likely a supporter of the original target user.

2. The Issue Cluster & Trust Period

Public opinion is volatile. An issue relevant today might be forgotten in two weeks. The authors introduce the Trust Period:

  • Start: The first time a target user mentions a keyword.
  • End: The last time that keyword is retweeted by an influential supporter.

By clustering tweets within this window using a specialized tf-idf weight (multiplied by a "supporter frequency" score), the system identifies the "pulse" of an issue.

Issue Classification Logic

3. Resolving the "Co-Follower" Conflict

Many users follow multiple rival politicians. To solve this, the authors use a Bias Ratio: This mathematical approach quantifies relative loyalty, acknowledging that support is often comparative rather than absolute.

Experimental Results

The study analyzed 10,000 followers for five prominent Korean politicians (e.g., Moon J.I., Park G.H.).

MethodAccuracyPrecisionF1 Measure
SVM (Baseline)0.5650.6400.590
Partial (HITS only)0.6310.6810.648
Proposed (Issue Clusters)0.6800.7050.675

The results clearly indicate that incorporating Issue Clusters provides a significant boost over purely linguistic models (SVM). The network density (shown in Figure 3) illustrates how the Issue Cluster expansion captures a much wider array of classifiable interactions than traditional methods.

Network Expansion Comparison Fig 3. Visualization of the expanded user network using the proposed methodology.

Critical Insight & Conclusion

The brilliance of this work lies in its Inductive Bias: it assumes that user behavior (retweeting) within a specific temporal context (trust period) is a stronger signal of sentiment than the text itself.

Takeaway: For developers and researchers building social listening tools, the lesson is clear—don't just analyze what is said. Analyze who is amplifying whom during the peak of a specific controversy. While the 68% accuracy leaves room for improvement (potentially through modern LLMs), the framework of using influential cohorts to label data is a powerful strategy for handling sparse text data.

Limitations: The system relies heavily on retweet behavior. "Silent followers" (lurkers) remain difficult to classify. Future work would benefit from incorporating "like" patterns or dwell time if that data were available via API.

Find Similar Papers

Try Our Examples

  • Find recent papers that utilize modified HITS or PageRank algorithms for identifying influential users in modern decentralized social media platforms.
  • Which study first introduced the concept of "trust periods" or temporal decay in social media topic modeling, and how does this paper's implementation differ?
  • Examine how current Large Language Model (LLM) based sentiment analysis compares to the SVM and HITS-based methods for classifying political follower behavior.
Contents
Beyond the Follow Button: Classifying User Intent via Issue Clusters and Influential Supporters
1. TL;DR
2. Background: The "Follower" Ambiguity
3. Methodology: The Three Pillars of Classification
3.1. 1. Identifying the Inner Circle (Influential Supporters)
3.2. 2. The Issue Cluster & Trust Period
3.3. 3. Resolving the "Co-Follower" Conflict
4. Experimental Results
5. Critical Insight & Conclusion