Efficient Representation of Interaction Patterns with Hyperbolic Hierarchical Clustering for Classification of Users on Twitter
Abstract.
Social media platforms play an important role in democratic processes. During the 2019 General Elections of India, political parties and politicians widely used Twitter to share their ideals, advocate their agenda and gain popularity. Twitter served as a ground for journalists, politicians and voters to interact. The organic nature of these interactions can be upended by malicious accounts on Twitter, which end up being suspended or deleted from the platform. Such accounts aim to modify the reach of content by inorganically interacting with particular handles. These interactions are a threat to the integrity of the platform, as such activity has the potential to affect entire results of democratic processes. In this work, we design a feature extraction framework which compactly captures potentially insidious interaction patterns. Our proposed features are designed to bring out communities amongst the users that work to boost the content of particular accounts. We use Hyperbolic Hierarchical Clustering (HypHC) which represents the features in the hyperbolic manifold to further separate such communities. HypHC gives the added benefit of representing these features in a lower dimensional space – thus serving as a dimensionality reduction technique. We use these features to distinguish between different classes of users that emerged in the aftermath of the 2019 General Elections of India. Amongst the users active on Twitter during the elections, 2.8% of the users participating were suspended and 1% of the users were deleted from the platform. We demonstrate the effectiveness of our proposed features in differentiating between regular users (users who were neither suspended nor deleted), suspended users and deleted users. By leveraging HypHC in our pipeline, we obtain F1 scores of upto 93%.
1. Introduction
Discussions on Online Social Networks (OSNs) are a key player in huge democratic processes such as elections. Recently, the 2020 U.S. General Elections (Chowdhury et al. 2021) and the subsequent Capitol Riots (News 2021) have been heavily influenced by conversations on OSNs such as Twitter and Parler (Kumaraguru et al. 2021). Politicians and political handles are very active on OSNs, especially during times of democratic elections (Khan et al. 2020). The nature of their engagement on these platforms is often an important part of their campaign strategies (Katz et al. 2013). Particularly, OSNs have played an important part in the 2014 and 2019 General Elections in India, where studies show the success of the winning party was closely associated with their use of Twitter to engage with voters (Ahmed et al. 2016; Gupta et al. 2020b). Users of these OSNs can interact with the content shared by political parties. The more interaction such content gets, the higher the popularity and higher the chance that such content reaches more people (Boulianne and Larsson 2021). Higher engagement with posts can also lead to benefits such as a higher number of followers and more media coverage (Keller and von Königslöw 2018). This engagement could lead to a domino effect that can turn the tide of election results (Flores-Saviaga et al. 2018). Due to the effect of this engagement on the outcome of democratic processes, it is important for OSNs to be fair in the way content is boosted.
Twitter in particular aims to keep their algorithm fair by endorsing and boosting content that is of genuine interest to most users.11 1 https://help.twitter.com/en/rules-and-policies/platform-manipulation This is important to ensure its status as a reliable platform for all political parties as well as voters by preserving interests spanning across all users.22 2 https://help.twitter.com/en/rules-and-policies/election-integrity-policy However, a group of users can manipulate the system by closely working together to artificially engage with the posts of a particular account.33 3 https://www.cigionline.org/articles/how-bjp-used-technology-secure-modis-second-win/ During democratic processes like elections, collusive behaviour by users can be a serious threat to the integrity of a platform (Le et al. 2019). Identifying and moderating such malicious accounts thus becomes critical for these social media giants.
Accounts which are found to be participating in malicious engagement with posts with an intent to artificially boost their reach can be suspended from Twitter. The presence of such accounts could constitute large proportions of engagement with a politician’s account (Onuchowska et al. 2019). Along with suspended accounts, accounts that are ultimately deleted from the platform can also pose a threat. Media has reported the spread of misinformative and deceptive content by fraudulent accounts that end up being deleted from the platform (Volkova and Bell 2016). Accounts that end up being suspended or deleted can form significant portions of the aforementioned collusive groups (Onuchowska et al. 2019), and thus identifying such accounts becomes very important, especially in the context of democratic processes like elections.
During the Indian General Elections of 2019, we observed three major classes of users. 1) Regular users, 2) Suspended users and 3) Deleted users. The class suspended consists of those users who were identified to be violating the Twitter policy and were suspended by the Twitter platform. The class deleted consists of users who were once active on Twitter but whose accounts were later deleted. Regular user accounts are Twitter users who were neither suspended nor deleted from the platform.
In the real world, any given OSN itself does not have information on the motivation behind a user’s behaviour. The only information such platforms have regarding the users is the engagement of the users with the platform. To distinguish between the different classes of users, we make use of this same information available publicly. Given that Twitter is only able to observe user activity but not the ultimate outcome (here, suspension or deletion of the account), we design a set of features which are class independent. We then use these features to train models to distinguish between different classes of users.
User interactions in OSNs often emulate trees with users forming communities (Ganea et al. 2018b; Ganea et al. 2018a; Chami et al. 2019). The communities of users which participate in artificial engagement can have deep hierarchies.3 Capturing such extreme hierarchies in the traditional Euclidean space is hard and inefficient (Ganea et al. 2018b). So, for our work we use hyperbolic manifolds (Ratcliffe et al. 1994) which are geometric structures with a negative curvature. This representation space can also be perceived as a continuous version of a tree with area increasing as we move away from the origin (point of observation) of the plane. This nature of geometry is useful to capture the tree-like user communities (Chami et al. 2020).
We use Hyperbolic Hierarchical Clustering - HypHC as the hyperbolic representation technique to capture communities and reduce dimensionality (Chami et al. 2020). HypHC provides the inherent advantage of community detection with information encoding. The reduced dimensions in the hyperbolic manifold ensures a computationally efficient downstream task. Additionally, representations obtained using HypHC can be treated just like Euclidean embeddings. We reduce the dimensionality of our features by a factor of ten without compromising on performance by leveraging hyperbolic manifolds via HypHC, all the while keeping every other component in our classification pipeline unchanged from its original Euclidean implementation.
The contributions of our work are as follows:
- (1)
A feature engineering framework that can help OSNs engineer efficient representations which capture the interaction patterns with top handles.
- (2)
Demonstrating the application of hyperbolic manifolds on real world social media data to reduce the computational and space complexity while effectively separating the regular, suspended and deleted accounts.
2. Related Work
Our work covers two major domains: the study around classification of accounts on Twitter; and hyperbolic manifolds and their applications.
2.1. Classification of Users in OSNs
Earlier works that studied spam on Twitter (Mccord and Chuah 2011) leveraged user characteristics such as number of followers, and tweet content to generate features. These features were used to train a random forest classifier to detect spamming accounts. Follow-up works (Lee and Kim 2014) aimed to identify malicious accounts created in a short period of time by using account names. They compare algorithm-made names with man-made names by clustering accounts sharing similar name-based features. The work by Wei et al. (Wei et al. 2016) uses temporal sentiment analysis to differentiate suspended users from non-suspended users with the help of statistical techniques like Naive Bayes classifier and SVM.
With claims that Twitter had been influencing voter sentiment during the U.S. elections, there have been focused attempts at characterizing Twitter’s part in democratic processes in many nations (Strandberg 2013; Knight 2012; Dzisah 2018). Notable studies on characterizing users based on Twitter’s moderation decisions (Le et al. 2019) show that the malicious communities suspended by Twitter exhibit a considerable difference from regular accounts. This class of works combines Twitter and elections to understand different user groups on the platform, spread of sentiment and news on the network and echo chambers and their effects, to minimize the effect Twitter has on democratic processes. In other works (Chowdhury et al. 2021), authors group Twitter users into communities based on their retweet and mention networks and analyze different characteristics such as popular tweeters, domains, and hashtags. They found that malicious and regular accounts participate in communities which exhibit significant differences in terms of popular account and hashtag usage.
Work has also been done on identifying and characterizing deleted users on social media platforms. Authors in (Bastos 2021) found that a significant proportion of accounts involved in political discourse on Twitter revolving around the Brexit referendum campaign were deleted from the platform. Volkova et al. used profile, network and behavior clues, sentiment and emotion features, text embeddings and topics to detect accounts deleted from Twitter which were active in the context of the Russian-Ukranian crisis (Volkova and Bell 2016). Twitter itself removed over 2 million accounts from the platform which were suspected to be fake. These accounts were allegedly giving misleading follower counts for some accounts (Dave 2018).
2.2. Hyperbolic manifolds and applications
Hyperbolic Hierarchical Clustering (HypHC) (Chami et al. 2020) is a key component of our work. We use this technique to reduce the dimensionality of our user representations for efficient computation. The following sub-section introduces previous research work on hyperbolic manifolds and their applications.
In the work (Nickel and Kiela 2017) authors claim that the ability of embeddings to model complex patterns is bounded by the dimensionality of the embedding space, because of which it is impossible to extract embeddings of large graph-structured data without any loss of information. So, to increase the representation capacity of embedding methods, they used hyperbolic spaces. Representations in hyperbolic spaces are capable of effectively capturing the underlying hierarchical structures in data at lower dimensions (Feng et al. 2020).
Feng et al. (Feng et al. 2020) leveraged this property of hyperbolic spaces to build a hyperbolic metric emebedding model, which projects location-based check-in data of users from social media into the hyperbolic space to predict the next POI (point of interest) for a user. Wang et al. (Wang et al. 2020b) proposed a hyperbolic geometry representation learning model to link user identities across different social network platforms.
Hyperbolic manifolds have also been used to embed knowledge graphs (Wang et al. 2020a), images (Khrulkov et al. 2020) and words (Tifrea et al. 2018) in far lower dimensions. Question-answering (Tay et al. 2018) and clustering (Monath et al. 2019) models have also leveraged hyperbolic manifolds to improve their performance.
3. Data description
In this section we introduce and describe the dataset used and the definitions of the different classes of users in our dataset, i.e. regular, suspended and deleted users.
For our work we use the ‘Analysis of General Elections 2019 in India’ (AGE2019) dataset (Gupta et al. 2020a). This dataset consists of tweets that span from February 5th 2019 to June 25th 2019. This covers a time period starting from two months prior to the first polling in the elections upto one month after the results of the elections were declared. The tweets are collected by querying the Twitter data collection APIs to retrieve tweets having hashtags related to the 2019 Indian General Elections. A total of 45.6 million tweets made by 2.2 million unique users are collected in the dataset.
The AGE2019 dataset also consists of two lists of user ids - deleted users and suspended users. These are users who were found to be either deleted or suspended from the platform as of June 29th, 2019. A total of 56,927 of these users were identified to be suspended as they returned error code 63 on querying the Twitter API. These users form our suspended class. Additionally, 21,083 users returned error code 50, signifying that these accounts had been deleted from the platform. These users form our deleted class. The AGE2019 dataset also shared a list of 100,000 users randomly sampled from the remaining users (neither suspended nor deleted). This set of users forms the regular class of users in our study. To generate features to separate the deleted, suspended and regular classes of users, we use the election-related tweets by each user, that are provided as a part of the AGE2019 dataset.
4. Feature Engineering
In this section, we describe our feature engineering process. Past research works have used features derived from the content of the tweets, extensive graphs of interactions between the users and more to extract features like sentiment, emotion, lexical features etc. to distinguish between classes on Twitter (Volkova and Bell 2017; Volkova and Bell 2016). However, these can become complicated to extract given the large amounts of data involved and complicated multi-lingual nature of the tweets. We use features that can easily be extracted based on information captured from the user’s profile and tweet activity, and do not delve into the actual content of the tweets. Section 6.1.1 demonstrates the importance of our features by comparing the same with some user-level features that past works have used to model the account characteristics. We use this section to explain the motivation and design of our proposed interaction features to capture the nature of interactions of each user with top profiles that emerged during the General Elections.
4.1. Interaction features
During the 2019 General Elections in India, contesting parties have been known to maintain an IT cell (Chattopadhyay 2020; Campbell-Smith and Bradshaw 2019). These IT cells, along with a dedicated team of supporters work to push the agenda of their parties as much as possible. As mentioned in (Campbell-Smith and Bradshaw 2019), these IT cells are known to propagate particular agendas on OSNs by engaging with posts by a certain account. So, artificial engagement in the context of the General Elections revolves around boosting the popularity and pushing the ideologies of a particular leader or party.
To identify the top leaders, we curate a list of users whose content is widely shared across the platform. We call them Influencers in our work because of the effect these users have on the content shared on the platorm. We particularly look at the engagement on the platform as a result of their tweets. Influencers need not be political leaders, but in the context of the 2019 General Elections, we observed that most Influencers are political leaders or political party handles. We use these Influencers to generate our interaction features. Table 1 summarizes the notations used henceforth.
| Notation | Meaning |
|---|---|
| Set of all Influencers | |
| Number of Influencers in | |
| Number of users retweeting a particular | |
| Influencer | |
| Set of all Tweets by Influencer | |
| Users engaging with the Influencers in | |
| All retweets by user of Influencer | |
| th Tweet by Influencer | |
| User ’s retweet of tweet | |
| Delay in retweeting by user | |
| Set of all delays | |
| Median of all delays in | |
| Number of times retweeted | |
| Interaction features for |
.
Retweets are a great way to quickly engage and amplify the reach of a tweet. Those retweets without any text of their own are just endorsements (Designer 2021; Cork and Eddy 2017; Kim and Yoo 2012). We take all retweets from the data and curate a list of user profiles whose tweets have been retweeted. For each such user in , we keep track of how many users in the dataset retweeted their tweets (i.e. ). We then consider the top user profiles having the highest number of retweets as Influencers to form the set .
The set of users interacting with Influencers in is given by where each user in has engaged with at least one Influencer in by retweeting at least one of their tweets. For each Influencer, we define to be the set of all tweets by Influencer . For each pair where user and Influencer , we look at those tweets by user which are retweets of any tweet by Influencer , i.e. retweet of any tweet in . We thus have a set of retweets set of all where is a retweet of Influencer ’s tweet by user . If a user has never retweeted a tweet by Influencer , then . In order to quantify a particular user’s () interaction with any Influencer (), we use two values:
- (1)
The first value we use is the delay, or time lag in retweeting. For each retweet in , we define the corresponding delay to be the absolute value of difference in seconds between time of the original tweet () and the time of retweet of that tweet (), i.e. (time of retweet - time of original tweet ). Figure 1 shows an illustrative demonstration of how each such is calculated. Thus, for each pair we have a set which is the set of all . forms the set of delays for each retweet by user of Influencer . We take the median of the delays in to get the final delay .
- (2)
The second value that we use is the number of times the user has retweeted that Influencer’s tweets. We define to be the number of elements in , i.e. the number of times user retweeted a tweet by Influencer .
Thus for each pair , we have a two-element vector which has two values: delay and the number of retweets . We form the interaction feature vector for each user by concatenating all such vectors obtained for each corresponding Influencer where . In case a user has never interacted with an Influencer (i.e. ), we set to a large negative value, and to 0. We choose a large negative in this case because such a value would never appear if was not empty, thus achieving a good separation in the feature space. Algorithm 1 describes the above explained feature engineering process in a pseudo-code format.
potential Influencers ;
foreach user in Users do
foreach Influencer in Influencers do
let be the set of tweets by Influencer ;
+ end foreach
With the interaction feature vector , we quantify each user’s interaction with the top handles. If there is a group of users colluding to interact with an Influencer account, their interaction feature vectors would look similar. Thus our features will help to capture such groups of collusive accounts.
In order to generate interaction features we have to choose the number of Influencers to include in the set . We observed the trends of number of unique users added to the set for each new Influencer added to the set . We wish to choose high enough that a large number of users are incorporated. A larger set of users means different subgroups of these users interact with the tweets of different Influencers, thus capturing a larger number of collusive groups. However, at the same time we do not wish to have a large because we only want the top tweeting Influencers, i.e. the most popular ones, who have a relatively large number of users retweeting their tweets. In order to achieve this balance, we study the relationship between the number of Influencers and number of users as depicted in Figure 2. In Figure 2(a), each point on the x-axis represents a particular Influencer, and the y-axis depicts the number of unique users retweeting that particular Influencer (i.e. ). This indicates how many unique users are interacting with each Influencer. Higher the value, more popular the Influencer.
To get a sense of the diversity of users incorporated by choosing a particular , we plot Figure 2(b). In Figure 2(b), the x-axis represents the number of Influencers , and the y-axis represents the size of (or ) for a particular value of . It is important to note that Figure 2(b) is not merely a cumulative plot of Figure 2(a) because there will be users who retweet the tweets of multiple Influencers. This graph is increasing, because as we add more Influencers to the set , we incorporate more users in the set . The amount of increase in the y-value of the graph as the x-value changes from to is the number of users added in when we increase the number of Influencers by one. To better visualise the effect of increase in on we plot Figure 2(c).
Each value in Figure 2.c indicates the additional number of users that are included for each new Influencer added to I. We reiterate that there will be users who retweet the tweets of multiple Influencers, and thus Figure 2(c) is different from Figure 2(a). By increasing the value of our final chosen , we include more spikes from Figure 2(c), which means we incorporate a more diverse set of users in . In Figure 2(c), we observe that after around 300 Influencers, the additional number of users that are included for each new Influencer added reduces. We hypothesize this to be the optimal value of . To verify the validity of this hypothesis, we chose values of at intervals of 50 between 100 and 600 and found best results on our classifiers (the same classifiers described in Section 6.1) with =300. This confirmed our hypothesis. By choosing (as depicted with a vertical line in Figure 2), we end up with containing 59,260 users. The class distribution of our final is depicted in Table 2. Note that we refer to our proposed user interaction features as henceforth.
| Class | Number of Users |
|---|---|
| Deleted | 8,078 |
| Regular | 32,386 |
| Suspended | 18,796 |
| Total | 59,260 |
5. Dimensionality reduction using Hyperbolic Hierarchical Clustering
Hyperbolic Hierarchical Clustering (HypHC) is a similarity based clustering method (Chami et al. 2020). First, a binary tree with leaves is constructed. Each leaf node denotes a Twitter user that needs encoding to a lower dimension. From these leaf nodes, intermediate nodes that connect closer nodes are formed. If two leaf nodes are found to potentially belong to the same cluster, they have a least common ancestor (LCA). Each sub-tree denotes a potential cluster. The goal of HypHC is to cluster nodes in such a way that the pairwise similarity of the data is captured and used to form clusters. The binary tree is built such that the pair-wise similarity between each pair of nodes is preserved. Once the binary tree is created, the Dasgupta cost is calculated on the tree. A good tree with distinct clusters is characterised by a low cost . Minimizing the Dasgupta cost merges similar nodes in the hierarchy, resulting in a tree with nodes clustered into appropriate communities. If is a binary tree, and is a pair of nodes with similarity between the two nodes, HypHC minimizes the Dasgupta cost amongst all possible binary trees as shown in the equation:
With the objective function in place, HypHC minimizes the cost via a continuous constrained optimization problem. A continuous tree representation is built with the help of leaf nodes which are initialized with random embeddings. The leaf nodes should ultimately contain enough information to recover the full tree. All the nodes are pushed towards the boundary of a Poincaré disk. A Poincaré is the hyperbolic geometric model HypHC uses. It has a negative curvature of -1. The curvature dictates how the geometry differs from a Euclidean plane. Negative curvature makes hyperbolic manifolds behave like continuous trees. In a hyperbolic manifold, under the Poincaré model, the distance between two points is defined by a geodesic (Chami et al. 2020).
The shortest path between any two nodes must pass through their least common ancestor (LCA), which in turn aids in constructing the whole binary tree from just the boundary nodes on the Poincaré disk. The HypHC algorithm gives us the binary tree with minimum Dasgupta cost. From this binary tree we extract embeddings of the leaf nodes, which are our final user embeddings. Figure 3 shows the application of HypHC in our work. We feed the interaction features to HypHC. After reducing the dimensionality of the 600 dimensional interaction features with HypHC, we get a 60 dimensional vector for each user.
| Deleted vs (Suspended + Regular) | Suspended vs (Regular + Deleted) | Regular vs (Deleted + Suspended) | |||||||||||||
| Model | U | U+F | HypHC | SE | FA | U | U+F | HypHC | SE | FA | U | U+F | HypHC | SE | FA |
| RFC | 90.03 | 92.25 | 93.84 | 90.29 | 87.40 | 85.82 | 89.01 | 87.50 | 84.23 | 84.16 | 77.31 | 79.69 | 79.03 | 76.34 | 74.53 |
| LGBM | 90.97 | 91.14 | 92.73 | 90.23 | 87.77 | 86.68 | 88.49 | 87.91 | 85.77 | 83.53 | 78.18 | 79.43 | 79.31 | 76.88 | 75.62 |
| XGB | 87.17 | 88.49 | 89.16 | 86.52 | 84.40 | 83.86 | 82.91 | 84.37 | 82.34 | 83.40 | 76.61 | 76.87 | 78.50 | 75.23 | 73.68 |
| GBC | 87.16 | 88.46 | 89.04 | 86.12 | 84.99 | 83.57 | 82.90 | 83.88 | 83.18 | 83.81 | 76.52 | 76.88 | 78.58 | 74.98 | 74.02 |
| DNN | 72.43 | 85.46 | 88.42 | 87.56 | 77.72 | 81.20 | 88.03 | 84.17 | 82.43 | 81.24 | 75.63 | 81.57 | 78.72 | 74.07 | 73.81 |
| LSTM | 80.33 | 89.40 | 92.27 | 87.13 | 76.39 | 67.69 | 81.98 | 77.81 | 76.94 | 74.97 | 61.44 | 68.21 | 71.38 | 76.38 | 74.85 |
| Deleted vs Suspended | Suspended vs Regular | Regular vs Deleted | |||||||||||||
| Model | U | U+F | HypHC | SE | FA | U | U+F | HypHC | SE | FA | U | U+F | HypHC | SE | FA |
| RFC | 83.34 | 86.42 | 85.34 | 85.14 | 81.98 | 81.51 | 87.29 | 86.07 | 83.33 | 82.55 | 85.50 | 88.52 | 88.01 | 81.93 | 83.61 |
| LGBM | 83.64 | 85.80 | 84.89 | 84.20 | 81.87 | 85.78 | 87.27 | 87.25 | 82.34 | 83.33 | 87.20 | 88.01 | 88.91 | 83.25 | 84.45 |
| XGB | 81.50 | 82.35 | 82.27 | 80.03 | 79.11 | 82.84 | 82.37 | 84.06 | 81.39 | 79.45 | 84.77 | 85.71 | 86.02 | 80.89 | 81.93 |
| GBC | 81.22 | 82.33 | 82.76 | 81.85 | 80.87 | 83.21 | 82.80 | 84.31 | 81.43 | 79.85 | 84.56 | 84.70 | 84.37 | 82.15 | 81.64 |
| DNN | 78.15 | 84.46 | 83.81 | 81.39 | 80.23 | 81.70 | 85.13 | 83.70 | 82.42 | 81.05 | 73.36 | 81.67 | 86.59 | 72.58 | 74.29 |
| LSTM | 76.99 | 82.73 | 81.17 | 78.92 | 78.52 | 81.20 | 85.87 | 83.28 | 80.23 | 81.48 | 69.76 | 80.72 | 86.38 | 74.29 | 71.48 |
6. Results
In this section we present our experiments to evaluate the effectiveness of our features in segregating the three classes, and analysis of the same. We also evaluate the effectiveness of HypHC as a dimensionality reduction technique.
6.1. Classifier Results
While comparing the different classes of users, we trained each model to classify the users in a one vs two fashion by training for suspended vs (regular + deleted), regular vs (deleted + suspended) and deleted vs (regular + suspended) separations. We also trained each model to classify the users in a one vs one (i.e. binary) fashion by training for suspended vs deleted, suspended vs regular and deleted vs regular separations.
We use two types of classifier models: deep learning based and tree based. For the deep learning based classifiers, we use a deep neural network and an LSTM. The results are presented after appropriate hyperparameter tuning in the loss function, activation function, the optimizer used, number of layers and the number of epochs. For the tree based classifiers, we use lightGBM (LGBM), XGBoost (XGB), Gradient Boosting Classifier (GBC) and the Random Forest Classifier (RFC). We use Grid Search to find the best set of hyperparameters for the tree based models. To account for class imbalance, we balance our training dataset using SMOTE (Chawla et al. 2002).
6.1.1. Comparison with standard feature engineering processes:
To evaluate the performance of our proposed features, we calculate 13 additional features for each user which are total number of tweets, number of tweets that are retweets, number of friends, number of followers, total likes, friends to follower ratio, time since account creation, lengths of screen name and bio (in characters and words) and average length of the tweet (in characters and words). We refer to these as user-level features. These features have been used in previous works to distinguish between deleted, suspended and regular users on Twitter (Volkova and Bell 2017; Volkova and Bell 2016). We do not use psycholinguistic features for comparison for reasons discussed in Section 4.
To establish the benefit of our proposed features, we use various feature sets, two of which are:
- (1)
U: 13 dimensional user-level features as described above.
- (2)
U+F: 613 dimensional features formed by appending the user-level features () to the 600 dimensional interaction features ().
Table 3 clearly shows that there is an increase in F1 scores in all the cases when the user-level features are appended with the interaction features (compare columns and ), showing that our features help the various models to achieve better separation.
6.1.2. Comparison with other dimensionality reduction techniques:
To establish our choice of HypHC as a dimensionality reduction technique, we compared its performance with popular unsupervised dimensionality reduction methods like Principal Component Analysis (PCA), t-distributed Stochastic Neighbor Embedding (t-SNE), Spectral Embedding (SE) and Feature Agglomeration (FA). Spectral Embedding and t-SNE are dimensionality reduction techniques based on manifold learning. Feature Agglomeration applies hierarchical clustering.
We reduced the 600 dimensional features () to 30, 60, 80 and 100 dimensions using HypHC, PCA, t-SNE, SE and FA, and appended the 13 dimensional user-level features (). The results observed after reducing to the dimensions mentioned above followed the same pattern – which was that HypHC outperformed all the reduction techniques. For the sake of brevity, we only present and discuss the results obtained after reducing the dimensions to 60 (which gave the highest F1 scores overall) and with the Spectral Embedding and Feature Agglomeration techniques, which performed the best out of our comparison techniques.
We thus obtain our next set of features:
- (1)
HypHC features: 73 dimensional features formed by reducing the 600 dimensional features () to 60 dimensions using HypHC and appending the 13 dimensional user-level features ().
- (2)
SE features: 73 dimensional features formed by reducing the 600 dimensional features () to 60 dimensions using SE and appending the 13 dimensional user-level features ().
- (3)
FA features: 73 dimensional features formed by reducing the 600 dimensional features () to 60 dimensions using FA and appending the 13 dimensional user-level features ().
HypHC, as mentioned in Section 5, takes advantage of hierarchical community detection to perform unsupervised dimensionality reduction to better separate classes. We evaluate the effectiveness of HypHC using two methods. First, we compare the performance of features with the original features. We find that in most cases, features perform at par with the features. Moreover, there are cases where features perform better than the features (compare columns and ). This shows that HypHC is an effective dimensionality reduction technique that is able to preserve the separation between classes even at a much lower dimension.
Second, we compare the performance of HypHC with established unsupervised dimensionality reduction techniques like SE and FA. Barring a few cases in rows 3 and 6 in Table 3(a), we find that the features outperform both features and features. This shows that HypHC is the superior dimensionality reduction technique.
6.2. Interclass Distances
To further evaluate the performance of HypHC, we compare the distances between the centers of the three classes after performing dimensionality reduction of the interaction features using HypHC, SE and FA. We take each of these three representations, and standardize each feature by removing the mean and scaling to unit variance. Table 4 shows the cosine distances between the centers of each of the three classes in the 60 dimensional space that was generated by each of the dimensionality reduction methods. It is immediately obvious that HypHC is able to achieve the highest and most uniform separation between the classes. This ratifies our claim that HypHC is able to obtain the best separation between the classes.
| HypHC | SE | FA | |
|---|---|---|---|
| Deleted vs Suspended | 10.9165 | 3.1636 | 2.9524 |
| Suspended vs Regular | 10.9197 | 3.4279 | 2.1796 |
| Regular vs Deleted | 10.9175 | 3.9857 | 3.2614 |
Our observations from Sections 6.1 and 6.2 show the effectiveness of our interaction features, and the valuable advantage of using HypHC in our pipeline. On comparing the results of all the classifiers for the dimensionality reduced () and high dimensional () features, we can see that the HypHC features often outperform the original features despite being at a much lower dimension. The added bonus of using lower dimensional data is decreased storage space and lower computation cost and time.
7. Conclusion
Ensuring that interactions between politicians and voters remain organic is critical to the fair functioning of any OSN, especially during democratic processes like elections. To do so, we capture these interaction patterns through our designed features. These interaction features are able to distinguish between the three classes effectively. To ensure that the model can run efficiently and take up as little space as possible, it is important to reduce the dimensionality of the features. To this end, we leverage HypHC, a novel unsupervised dimensionality reduction technique. We show that HypHC performs better than other established dimensionality reduction techniques at separating the classes. Since our interaction features are OSN-agnostic, we plan to carry out these same analyses on other platforms.
References
- (1)
- Ahmed et al. (2016) Saifuddin Ahmed, Kokil Jaidka, and Jaeho Cho. 2016. The 2014 Indian elections on Twitter: A comparison of campaign strategies of political parties. Telematics and Informatics 33, 4 (Nov. 2016), 1071–1087. https://doi.org/10.1016/j.tele.2016.03.002
- Bastos (2021) Marco Bastos. 2021. This Account Doesn’t Exist: Tweet Decay and the Politics of Deletion in the Brexit Debate. American Behavioral Scientist 65, 5 (Jan. 2021), 757–773. https://doi.org/10.1177/0002764221989772
- Boulianne and Larsson (2021) Shelley Boulianne and Anders Olof Larsson. 2021. Engagement with candidate posts on Twitter, Instagram, and Facebook during the 2019 election. New Media & Society (April 2021), 146144482110095. https://doi.org/10.1177/14614448211009504
- Campbell-Smith and Bradshaw (2019) Ualan Campbell-Smith and Samantha Bradshaw. 2019. Global cyber troops country profile: India.
- Chami et al. (2020) Ines Chami, Albert Gu, Vaggos Chatziafratis, and Christopher Ré. 2020. From Trees to Continuous Embeddings and Back: Hyperbolic Hierarchical Clustering. In Advances in Neural Information Processing Systems, H. Larochelle, M. Ranzato, R. Hadsell, M. F. Balcan, and H. Lin (Eds.), Vol. 33. Curran Associates, Inc., 15065–15076. https://proceedings.neurips.cc/paper/2020/file/ac10ec1ace51b2d973cd87973a98d3ab-Paper.pdf
- Chami et al. (2019) Ines Chami, Rex Ying, Christopher Ré, and Jure Leskovec. 2019. Hyperbolic graph convolutional neural networks. Advances in neural information processing systems 32 (2019), 4869.
- Chattopadhyay (2020) Aditi Chattopadhyay. 2020. Killing Democracy Tweet By Tweet: How IT Cells Of Political Parties Wage Propaganda War. https://thelogicalindian.com/exclusive/bjp-it-cell-amit-malviya-congress-it-cell-seed-accounts-social-media-twitter-19712 Accessed: 21-03-2021.
- Chawla et al. (2002) N. V. Chawla, K. W. Bowyer, L. O. Hall, and W. P. Kegelmeyer. 2002. SMOTE: Synthetic Minority Over-sampling Technique. Journal of Artificial Intelligence Research 16 (Jun 2002), 321–357. https://doi.org/10.1613/jair.953
- Chowdhury et al. (2021) Farhan Asif Chowdhury et al. 2021. Examining Factors Associated with Twitter Account Suspension Following the 2020 U.S. Presidential Election. CoRR abs/2101.09575 (2021). arXiv:2101.09575 https://arxiv.org/abs/2101.09575
- Cork and Eddy (2017) B Colin Cork and Terry Eddy. 2017. The retweet as a function of electronic word-of-mouth marketing: A study of athlete endorsement activity on Twitter. International Journal of Sport Communication 10, 1 (2017), 1–16.
- Dave (2018) Paresh Dave. 2018. Twitter cuts suspect users from follower counts again, blames bug. https://theacademicdesigner.com/2020/likes-and-retweets-are-endorsements/ Accessed: 01-11-2021.
- Designer (2021) The Academic Designer. 2021. Likes and Retweets Are Endorsements on Social Media. https://theacademicdesigner.com/2020/likes-and-retweets-are-endorsements/ Accessed: 20-02-2021.
- Dzisah (2018) Wilberforce S Dzisah. 2018. Social media and elections in Ghana: Enhancing democratic participation. African Journalism Studies 39, 1 (2018), 27–47.
- Feng et al. (2020) Shanshan Feng, Lucas Vinh Tran, Gao Cong, Lisi Chen, Jing Li, and Fan Li. 2020. Hme: A hyperbolic metric embedding approach for next-poi recommendation. In Proceedings of the 43rd International ACM SIGIR Conference on Research and Development in Information Retrieval. 1429–1438.
- Flores-Saviaga et al. (2018) Claudia Flores-Saviaga, Brian Keegan, and Saiph Savage. 2018. Mobilizing the Trump Train: Understanding Collective Action in a Political Trolling Community. In ICWSM.
- Ganea et al. (2018a) Octavian Ganea, Gary Bécigneul, and Thomas Hofmann. 2018a. Hyperbolic entailment cones for learning hierarchical embeddings. In International Conference on Machine Learning. PMLR, 1646–1655.
- Ganea et al. (2018b) Octavian-Eugen Ganea, Gary Bécigneul, and Thomas Hofmann. 2018b. Hyperbolic neural networks. (2018).
- Gupta et al. (2020a) Saurabh Gupta, Agarwal Anant, Suryatej Reddy Vyalla, Arun Balaji Buduru, and Ponnurangam Kumaraguru. 2020a. #IVoted to #IGotPwned: Studying Voter Privacy Leaks in Indian Lok Sabha Elections on Twitter. (2020).
- Gupta et al. (2020b) Saurabh Gupta, Asmit Kumar Singh, Arun Balaji Buduru, and Ponnurangam Kumaraguru. 2020b. Hashtags Are (Not) Judgemental: The Untold Story of Lok Sabha Elections 2019. 216–220. https://doi.org/10.1109/BigMM50055.2020.00038
- Katz et al. (2013) James Katz, Anshul Jain, and Michael Barris. 2013. The Social Media President: Barack Obama and the Politics of Digital Engagement.
- Keller and von Königslöw (2018) Tobias R. Keller and Katharina Kleinen von Königslöw. 2018. Followers, Spread the Message! Predicting the Success of Swiss Politicians on Facebook and Twitter. Social Media + Society 4, 1 (Jan. 2018), 205630511876573. https://doi.org/10.1177/2056305118765733
- Khan et al. (2020) Asif Khan et al. 2020. Predicting Politician’s Supporters’ Network on Twitter Using Social Network Analysis and Semantic Analysis. Scientific Programming 2020 (09 2020), 1–17. https://doi.org/10.1155/2020/9353120
- Khrulkov et al. (2020) Valentin Khrulkov, Leyla Mirvakhabova, Evgeniya Ustinova, Ivan Oseledets, and Victor Lempitsky. 2020. Hyperbolic image embeddings. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition. 6418–6428.
- Kim and Yoo (2012) Jihie Kim and Jaebong Yoo. 2012. Role of sentiment in message propagation: Reply vs. retweet behavior in political communication. In 2012 International Conference on Social Informatics. IEEE, 131–136.
- Knight (2012) Megan Knight. 2012. Journalism as usual: The use of social media as a newsgathering tool in the coverage of the Iranian elections in 2009. Journal of Media Practice 13, 1 (2012), 61–74.
- Kumaraguru et al. (2021) Ponnurangam Kumaraguru et al. 2021. Capitol (Pat) riots: A comparative study of Twitter and Parler. arXiv preprint arXiv:2101.06914 (2021).
- Le et al. (2019) Huyen Le, GR Boynton, Zubair Shafiq, and Padmini Srinivasan. 2019. A postmortem of suspended Twitter accounts in the 2016 US presidential election. In 2019 IEEE/ACM International Conference on Advances in Social Networks Analysis and Mining (ASONAM). IEEE, 258–265.
- Lee and Kim (2014) Sangho Lee and Jong Kim. 2014. Early filtering of ephemeral malicious accounts on Twitter. Computer Communications 54 (2014), 48–57.
- Mccord and Chuah (2011) Michael Mccord and M Chuah. 2011. Spam detection on twitter using traditional classifiers. In international conference on Autonomic and trusted computing. Springer, 175–186.
- Monath et al. (2019) Nicholas Monath, Manzil Zaheer, Daniel Silva, Andrew McCallum, and Amr Ahmed. 2019. Gradient-based hierarchical clustering using continuous representations of trees in hyperbolic space. In Proceedings of the 25th ACM SIGKDD International Conference on Knowledge Discovery & Data Mining. 714–722.
- News (2021) BBC News. 2021. Capitol riots timeline: The evidence presented against Trump. https://www.bbc.com/news/world-us-canada-56004916 Accessed: 20-02-2021.
- Nickel and Kiela (2017) Maximilian Nickel and Douwe Kiela. 2017. Poincaré Embeddings for Learning Hierarchical Representations. CoRR abs/1705.08039 (2017). arXiv:1705.08039 http://arxiv.org/abs/1705.08039
- Onuchowska et al. (2019) A. Onuchowska, D. Berndt, and Sagar Samtani. 2019. Rocket ship or Blimp? - Implications of Malicious Accounts removal on Twitter. In ECIS.
- Ratcliffe et al. (1994) John G Ratcliffe, S Axler, and KA Ribet. 1994. Foundations of hyperbolic manifolds. Vol. 149. Springer.
- Strandberg (2013) Kim Strandberg. 2013. A social media revolution or just a case of history repeating itself? The use of social media in the 2011 Finnish parliamentary elections. New Media & Society 15, 8 (2013), 1329–1347. https://doi.org/10.1177/1461444812470612 arXiv:https://doi.org/10.1177/1461444812470612
- Tay et al. (2018) Yi Tay, Luu Anh Tuan, and Siu Cheung Hui. 2018. Hyperbolic representation learning for fast and efficient neural question answering. In Proceedings of the Eleventh ACM International Conference on Web Search and Data Mining. 583–591.
- Tifrea et al. (2018) Alexandru Tifrea, Gary Bécigneul, and Octavian-Eugen Ganea. 2018. Poincaré GloVe: Hyperbolic Word Embeddings. CoRR abs/1810.06546 (2018). arXiv:1810.06546 http://arxiv.org/abs/1810.06546
- Volkova and Bell (2016) Svitlana Volkova and Eric Bell. 2016. Account Deletion Prediction on RuNet: A Case Study of Suspicious Twitter Accounts Active During the Russian-Ukrainian Crisis. 1–6. https://doi.org/10.18653/v1/W16-0801
- Volkova and Bell (2017) Svitlana Volkova and Eric Bell. 2017. Identifying effective signals to predict deleted and suspended accounts on twitter across languages. In Proceedings of the International AAAI Conference on Web and Social Media, Vol. 11.
- Wang et al. (2020b) Feiyang Wang, Li Sun, and Zhongbao Zhang. 2020b. Hyperbolic User Identity Linkage across Social Networks. In GLOBECOM 2020 - 2020 IEEE Global Communications Conference. 1–6. https://doi.org/10.1109/GLOBECOM42002.2020.9322242
- Wang et al. (2020a) Shen Wang et al. 2020a. H2KGAT: Hierarchical Hyperbolic Knowledge Graph Attention Network. In Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing (EMNLP). 4952–4962.
- Wei et al. (2016) Wei Wei, Kenneth Joseph, Huan Liu, and Kathleen M Carley. 2016. Exploring characteristics of suspended users and network stability on Twitter. Social network analysis and mining 6, 1 (2016), 1–18.