ValueGraph: Value-Signal Guided Graph Pre-training for Contextualized
User Representation
Abstract
Value signals are aggregated user-level moral representations that capture users’ inferred value-related tendencies from their online discourse. User behavior on social media is shaped not only by what users say or whom they interact with, but also by the value signal through which they express attitudes. Existing user representation methods largely miss this value-relevant dimension. We propose ValueGraph, a graph pre-training framework that uses automatically inferred moral-value signals as noisy auxiliary signals for contextualized user representation. From post-reply graphs, ValueGraph learns semantic and structural representations and further aligns users through relative value similarity with contrastive and clustering objectives. Rather than treating inferred values as gold psychological labels, ValueGraph uses them as soft constraints for representation learning. Experiments on stance detection and twitter bot detection show consistent gains over strong text-based, graph-based, and text-only LLM baselines, highlighting value-signal guidance as a useful inductive bias for socially informed user modeling.11 1 Code is released at https://github.com/HanYiton/ValueGraph
1School of Computing and Information Systems, Singapore Management University, Singapore
2Institute of Advanced Intelligence and Computing, A*STAR, Singapore
{ythan,weigao,yizhao,fzzeng,mohammada}@smu.edu.sg, prasanta_bhattacharya@a-star.edu.sg
Introduction
Understanding user behavior is central to social media analytics, recommendation systems, and personalized AI. User embedding models map user content and interactions to dense representations for downstream tasks such as profiling, recommendation, stance detection, and social influence analysis, and must generalize to unseen users and evolving social contexts (Pan and Ding 2019).
Most existing methods learn user representations from text (Benton et al. 2016; Ding et al. 2017), visual content (Do et al. 2018), or network structure (Rahimi et al. 2015; Pick et al. 2022), and optimize them with self-supervised or graph-based objectives (Perozzi et al. 2014; Sun et al. 2020b; Grover and Leskovec 2016; Donnat et al. 2018). These objectives often assume that users with similar content exposure or graph neighborhoods should have similar embeddings. However, similarity in interaction does not necessarily indicate similarity in the underlying factors that drive user behaviors. Users may engage with the same event because of shared exposure, controversy, or platform dynamics while expressing substantially different motivations and attitudes. When social contexts shift, such surface-level similarity can become brittle (Zhao et al. 2021; Hafidi et al. 2020; Hassani and Khasahmadi 2020; Sun et al. 2020a). For example, news recommendation systems that rely on interaction-based similarity may mistake shared exposure or controversy-driven interactions for shared preferences or values. Users who engage with the same event may hold opposing attitudes, but such spurious behavioral associations can cause the system to repeatedly recommend similar content to these users, reinforcing homogeneous information exposure and potentially amplifying polarization or extremist dynamics (Ribeiro et al. 2020). These limitations suggest that effective user representations should capture not only observable behaviors, such as what users say or where they interact, but also latent factors underlying these behaviors.
To address these challenges, we adopt Moral Foundations Theory (MFT) (Haidt et al. 2007) as our theoretical foundation. MFT provides a framework for understanding moral judgments reflected in social discourse and defines interpretable moral dimensions that have been widely studied in computational social science (Graham et al. 2013). Building upon recent advances in text-based moral value prediction, we derive 10-dimensional user-level moral feature vectors by aggregating post-level moral scores output from MoralBERT (Preniqi et al. 2024), and refer to these vectors as value signals. Without explicitly modeling such value signals, conventional embedding models may struggle to distinguish users who exhibit similar interaction patterns but hold different stances and behavioral tendencies. This motivates us to leverage value signals as an auxiliary supervisory cue to learn more robust user representations.
We therefore propose ValueGraph, a contextualized user representation framework that incorporates inferred value signals to regularize graph pre-training. In this work, contextualized refers to representations learned by jointly modeling users’ textual content and their interaction context in conversation graphs. Concretely, ValueGraph operates on a post-reply graph and learns user representations in two stages. First, it uses masked graph autoencoding to learn semantic and structural post representations from conversation graphs. Second, it aggregates post embeddings into user embeddings and constructs positive and negative user pairs according to inferred value similarity, which is derived from psychology and computational ethics literature (Preniqi et al. 2024; Nguyen et al. 2024; Guo et al. 2023). A user-level contrastive loss aligns users with similar inferred value profiles, while a clustering loss encourages coherent and separated regions in the embedding space. These objectives guide the encoder to represent what users engage with alongside the value signals associated with their behavior.
We evaluate ValueGraph on stance detection and Twitter bot detection. Results show consistent gains over graph pre-training methods, Pre-trained Language Model (PLM) encoders, inductive GNNs, and text-only LLM baselines, while ablations confirm that the improvements come from jointly modeling semantic, structural, and value-signal information. Our main contributions are summarized as follows:
- •
We propose ValueGraph, a value-signal guided graph pre-training framework that leverages inferred value signals as noisy auxiliary supervision for contextualized user representation learning.
- •
We design a hierarchical objective that combines masked graph autoencoding with contrastive learning and clustering to jointly capture semantic, structural, and value information.
- •
We provide a mechanism-level analysis of the proposed objectives, characterizing value-signal similarity preservation and cluster compactness/separation in the learned embedding space.
- •
Experiments show that ValueGraph achieves the best performance among compared user representation methods on both stance detection and bot detection tasks, validating the effectiveness of incorporating value signals into graph-based user representation learning.
Related Work
User Representation.
Social media user representations map high-dimensional user features into dense embeddings (Pan and Ding 2019). Prior work learns such representations from text (Benton et al. 2016; Ding et al. 2017), visual content (Do et al. 2018), network structure (Rahimi et al. 2015), or multi-view fusion (Zhang et al. 2017; Zhang et al. 2018; Ribeiro et al. 2018). These methods are effective, but rarely use value-relevant moral signals as explicit inductive bias.
Graph Pre-training.
GNN pre-training improves representation generalization by exploiting graph structure and neighborhood similarity (Perozzi et al. 2014; Sun et al. 2020b; Grover and Leskovec 2016; Donnat et al. 2018; Zhang et al. 2019; Tang et al. 2015; Zhao et al. 2021). Later work explores similar-domain and cross-domain pre-training to improve transferability (Hafidi et al. 2020; Hassani and Khasahmadi 2020; Sun et al. 2020a; Zhu et al. 2020; Hu et al. 2020a; You et al. 2020; Hu et al. 2020b; Lu et al. 2021; Qiu et al. 2020). However, existing graph pre-training methods rarely model value signal or motivational dimensions in user behavior.
AI and Human Values.
Recent work detects moral-values in text using transformer-based models (Preniqi et al. 2024; Nguyen et al. 2024; Guo et al. 2023; Zangari et al. 2025), while value-alignment research focuses on instruction following and preference-based reward modeling (Ouyang et al. 2022; Christiano et al. 2017). Instead of focusing on text-level moral-value detection or language-model alignment, we use automatically inferred value signals as auxiliary signals for graph-based user representation learning.
Problem Formulation
We are given a dataset of social conversation graphs. Each graph represents a conversation thread, where is the set of posts and is the set of directed reply edges between posts22 2 We focus on post-reply relations, since follow/friendship relations are often platform-dependent and unavailable across datasets. This design enables a more general evaluation of value-signal guided graph pre-training.. Let denote all users who authored posts in . Since each post is authored by exactly one user, we define an authorship mapping , where indicates that post is authored by user . Let be the set of all posts. We define as a post aggregation function, where is the set of posts authored by .
The goal of user representation learning is to learn a parameterized encoder that maps each user to a compact embedding . The embedding is generated by aggregating semantic information from the user’s posts and structural context from post-post reply relations. The resulting representations are expected to capture semantic, relational, and value-relevant behavioral cues that support downstream behavior and content understanding (Pan and Ding 2019), such as stance detection and Twitter bot detection (Section Experiments and Results).
Methodology
Our ValueGraph consists of two stages: Foundation GNN Pre-training and Value-Guided Pre-training. In the first stage, we pre-train a foundation GNN on the post-reply graph with a masked autoencoding paradigm to obtain noise-resilient representations. In the second stage, we refine these representations using value signals derived from MoralBERT. Figure 1 shows an overview of our ValueGraph framework.
Stage 1: Foundation GNN Pre-training.
We first perform unsupervised pre-training of a foundation GNN on the constructed graph dataset. Each post is encoded using ModernBERT (Warner et al. 2025), chosen for robustness to noisy social media text. Reply interactions define graph edges, enabling multi-hop message passing to capture local and global conversational patterns. Pre-training follows the masked autoencoding paradigm of GraphMAE2 (Hou et al. 2023). Node features are masked before encoding, and the model reconstructs the original features. This objective encourages noise-resilient representations that serves as a strong initialization for the subsequent value-guided pre-training.
Stage 2: Value-Guided Pre-training.
In this stage, we refine the pre-trained representations with a hierarchical contrastive learning framework. Specifically, we infer a value profile for each user, form contrastive user pairs between users who have similar and dissimilar values, and combine a user-level contrastive objective with a clustering objective , which fully exploit the inter-user relations encoded in value similarity and remain robust to the noise in the signal.
Human-Value Grounding for Training Data.
Human values play a central role in shaping social interactions, moral reasoning, and ideological expression. Several frameworks model human values, including Hofstede’s Cultural Dimensions Theory (Hofstede 2001), MFT (Haidt et al. 2007), and Schwartz’s Theory of Basic Human Values (Schwartz et al. 2012). These frameworks have been widely adopted to analyze value-driven language and behavior. We adopt MFT as our primary value framework for three reasons. First, it captures moral judgments expressed in everyday language, making it well suited for social media analysis (Johnson and Goldwasser 2018; Mooijman et al. 2018). Second, its effectiveness in extracting moral perspectives from text has been empirically validated (Lin et al. 2018; Mokhberian et al. 2020). Third, it provides structured and interpretable moral dimensions aligned with our objective of modeling value-related signals (Araque et al. 2020; Zhang et al. 2025). More details about MFT are provided in Appendix: Moral Foundations Background.
To obtain value signals for pre-training, we employ MoralBERT (Trager et al. 2022), a fine-tuned language model for capturing moral-values in social discussions33 3 https://github.com/vjosapreniqi/MoralBERT. Specifically, we use ten independently fine-tuned classifiers, each corresponding to one of ten moral foundations defined by MFT, i.e., care, harm, fairness, cheating, loyalty, betrayal, authority, subversion, purity, degradation. Given a post , each classifier produces a probability score for each foundation . To obtain user-level signal, we aggregate scores across posts authored by a user:
| (1) |
The resulting 10-D vector is not treated as gold user values; it serves as a noisy but more robust moral-framing signal, whose usefulness comes from relative differences across users rather than exact post-level predictions.
Contrastive User Pairs Construction.
Given inferred value vectors derived from MFT, we sample user pairs according to a similarity function . We adopt a symmetric percentile-based strategy, where the -th percentile (e.g., ) defines positive pairs and the -th percentile defines negative pairs:
| (2) | ||||
For each user , we define positive and negative sets as and . Assuming an approximately symmetric similarity distribution (see Appendix: Distribution of Sampled User Similarities), this strategy yields balanced positive and negative sets. If either set is underpopulated for a given user, we iteratively relax the threshold to the median similarity until a minimum set size is reached.
Similarity Computation.
User similarity is computed from the Euclidean distance between inferred value vectors:
| (3) |
which is converted into a similarity score through a Gaussian radial basis function kernel (Li et al. 2021):
| (4) |
where is set to the median of the sampled distances.
Loss Functions.
We treat the value vectors from text as a noisy auxiliary signal that defines relative constraints among users. This places a twofold requirement on our training objective: it should both fully exploit the inter-user relations encoded in value similarity and remain robust to the noise in the signal. Based on this, we design two complementary losses. The full training is illustrated in Appendix: Algorithm.
(1) User-level contrastive loss : For each seed user , we encourage its embedding to be close to value-similar and distant from value-dissimilar users . For each positive pair , we apply the InfoNCE loss:
| (5) |
where is a temperature and denotes cosine similarity. The overall contrastive loss is:
| (6) |
where
| (7) |
This loss constructs positive and negative pairs according to the inferred value similarity and aligns user representations at the semantic level, pulling users with similar inferred value-related textual patterns together and pushing users with dissimilar values apart.
(2) Clustering loss : To regularize the global structure of the user embedding space, we periodically perform -means clustering on the user embeddings every epochs. Such periodic cluster assignment follows the alternating optimization paradigm commonly adopted in deep clustering methods, where cluster assignments are periodically updated while the encoder progressively refines representations (Caron et al. 2018). In our framework, clustering serves as a global embedding-space regularizer rather than the primary learning objective. Let denote the centroid of cluster and be the cluster assignment of user . Following maximum-margin clustering, we first define a compactness term that encourages each embedding to stay close to its assigned centroid:
| (8) |
We further define a margin term that enforces a minimum distance between distinct cluster centroids:
| (9) |
Combining the two, the clustering loss is defined as:
| (10) |
where weights the margin term. For epochs without clustering, is set to zero. This loss imposes a global geometric constraint on the user embeddings, with the compactness term encouraging intra-cluster cohesion and the margin term encouraging inter-cluster separation.
(3) The overall loss : Each loss on its own covers only half of the training objective. is a purely pairwise constraint that only specifies which users should be close to which, without constraining the global geometry of the embedding space. , in contrast, is a self-reinforcing objective that clusters the current embeddings and pulls them toward centroids: it amplifies whatever structure already exists in the representations, without distinguishing whether that structure reflects genuine user differences. Consequently, using either loss in isolation could be unstable. We therefore design the final objective as a combination of the two:
| (11) |
where controls the contribution of the clustering term. The two losses act as mutual regularizers, in which provides with a meaningful clustering axis aligned along value-relevant dimensions, while provides with global stabilization and denoising by aggregating large numbers of users into groups. Representations that simultaneously satisfy local (pairwise) value consistency and global (group) structural consistency are preserved, while noisy structure is filtered out. Appendix: Training Loss and Hyperparameter Tuning gives training loss curves and hyperparameter tuning details.
Theoretical Analysis
We provide a mechanism-level characterization of how the training objectives shape the embedding space, which formalizes how value-signal contrastive learning pulls users with similar inferred profiles closer, while clustering promotes compact and separated user groups. Full Proofs are provided in Appendix: Full Proof of Theorem 1 and 2.
Theorem 1 (Value-Signal Similarity Preservation).
Let denote the inferred value vector of user , and let be the user embedding learned by ValueGraph. Suppose converges under fixed positive and negative sets constructed from . For sampled user pairs, the learned embedding similarity is encouraged to preserve the ordering induced by value-signal similarity:
Theorem 2 (Cluster Compactness and Separation).
Let user embeddings be optimized with , and let denote the resulting cluster centroids. The compactness term minimizes for users assigned to cluster , while the margin term penalizes centroid pairs with distance below and therefore encourages inter-cluster separation.
Experiments and Results
Pre-training Corpus
We construct a large-scale pre-training corpus from social media conversations collected from Reddit and Twitter. Reddit dataset44 4 https://convokit.cornell.edu/documentation/subreddit.html provides diverse community discussions across subreddits, where we follow Trager et al. (2022) to select communities reflecting diverse moral concerns. Twitter datasets include widely used rumor detection benchmarks, including PHEME (Zubiaga et al. 2016), Twitter16 (Ma et al. 2016), BEARD (Zeng and Gao 2022), and Twitter-Covid (Lin et al. 2022), which contain conversations involving rumors, controversies, and value-related debates. These datasets provide rich textual and reply-based interaction signals for learning contextualized user representations. After preprocessing, the combined corpus consists of 461,198 conversation graphs with approximately 13.6 million nodes and 39.9 million edges.
| LLM | Graph Pre-training | PLM | GNN | ValueGraph | ||||||||
| Model | Dataset | Metric | GPT-5.4 | GraphCL | GraphMAE2 | SimCSE | ModernBERT | Bertweet | GCN | GAT | GTN | |
| GLAN | MT_CSD | Acc. | 0.58 | 0.41 | 0.60 | 0.60 | 0.56 | 0.58 | 0.40 | 0.42 | 0.57 | 0.63∗ |
| MacF1 | 0.56 | 0.38 | 0.55 | 0.56 | 0.53 | 0.51 | 0.35 | 0.33 | 0.56 | 0.58 | ||
| BrLSTM | RumourEval19 | Acc. | 0.73 | 0.59 | 0.67 | 0.74 | 0.74 | 0.65 | 0.61 | 0.31 | 0.71 | 0.77∗ |
| MacF1 | 0.47 | 0.38 | 0.47 | 0.53 | 0.59 | 0.56 | 0.25 | 0.18 | 0.45 | 0.71∗ | ||
Stance Detection
Stance detection classifies users’ attitudes toward a target as support, oppose, or neutral. Because socio-political stances are often associated with value signal and ethical orientation (Al-Khatib et al. 2020; Durmus and Cardie 2018), ValueGraph provides value signal that serve as a useful inductive bias for stance inference. We evaluate ValueGraph on two benchmarks, MT_CSD (Niu et al. 2024) and RumourEval19 (Gorrell et al. 2019). We integrate our post-level embeddings into two established stance models. Details are shown in Appendix: Stance Detection.
(1) On MT_CSD, we adopt the graph-prompt framework of Zhao et al. (2024), which performs event-aware prompt tuning over conversation graphs. We replace its GNN node embeddings with our post embeddings and use GLAN (Yuan et al. 2019) for prediction.
(2) On RumourEval19, we use BrLSTM (Yue et al. 2022), which models tree-structured reply dependencies, and replace its initial token-level inputs with precomputed post embeddings. This plug-in strategy provides contextual and relational representations without modifying downstream architectures or adding supervision.
Baselines.
We compare against four categories of strong baseline encoders.
(1) Graph-based pre-training: GraphCL (Hafidi et al. 2020), which applies contrastive learning over augmented graph views, and GraphMAE2 (Hou et al. 2023), a masked graph autoencoder. Both are pre-trained on the same datasets using an identical graph encoder and subsequently fine-tuned for stance classification.
(2) Pre-trained language models: BERTweet (Nguyen et al. 2020), SimCSE (Gao et al. 2021), and ModernBERT (Warner et al. 2025). All these pre-trained language models are fine-tuned end-to-end on the stance detection task.
(3) Inductive GNNs: GCN (Kipf and Welling 2016), GTN (Yun et al. 2019), and GAT (Velickovic et al. 2017), trained from scratch on the downstream task to test whether graph structure alone, without graph pre-training or value-signal guidance, is sufficient.
(4) Text-only LLM: GPT-5.4, prompted with the target, available user posts, and thread text with no graph serialization.55 5 GPT-5.4 does not receive explicit graph; structure-aware graph serialization for LLMs is nontrivial and left to future work.
For fair comparison, trainable models use identical data splits, optimizer configurations, and early-stopping criteria based on validation loss. GPT-5.4 is evaluated on the same test splits with fixed prompts and deterministic decoding. Fixed hyperparameters and prompt templates are provided in Appendix: Hyperparameter Settings and GPT-5.4 Prompt Templates.
Evaluation Metrics.
We report accuracy (acc) and macro F1 (macF1). While accuracy provides an overall measure, macF1 is more informative under class imbalance, which is prevalent in stance detection datasets.
Results.
As shown in Table 1, ValueGraph consistently outperforms the non-LLM and LLM baselines on both benchmarks; GPT-5.4 performs comparably with the non-LLM baselines. On MT_CSD, ValueGraph attains 63% acc and 58% macF1, yielding relative gains of 5% in acc and 5.5% in macF1 over the strongest pre-trained baseline (GraphMAE2, 60% acc and 55% macF1). On RumourEval19, ValueGraph achieves 77% acc and 71% macF1, corresponding to a relative macF1 improvement of 20.3% over the strongest PLM baseline (ModernBERT, 59%) and 57.8% over the best graph-based baseline (GTN, 45%).
Twitter Bot Detection
| Method | Setting | Type | Accuracy | MacF1 | Precision | Recall | MCC |
| GPT-5.4 | Text-only | T | 56.9 | 52.1 | 57.1 | 54.9 | 11.8 |
| RoBERTa | Baseline | T | 50.7 | 54.8 | 48.8 | 62.5 | – |
| ValueGraph | T | 56.9∗∗ | 64.3∗∗ | 53.3∗∗ | 81.0∗ | – | |
| Baseline | UT | 63.1 | 66.2 | 58.9 | 75.6 | – | |
| ValueGraph | UT | 65.9∗∗ | 68.9∗ | 61.2 | 78.7 ∗ | – | |
| BotGCN (T5) | Baseline | TG | 56.8 | 64.4 | 53.7 | 82.1 | 17.9 |
| ValueGraph | TG | 70.2∗∗ | 67.9∗∗ | 69.8∗ | 66.4∗∗ | 40.3∗∗ | |
| Baseline | FTUG | 70.5 | 70.4 | 67.1 | 74.2 | 41.3 | |
| ValueGraph | FTUG | 71.1∗∗ | 71.1 | 68.9∗∗ | 72.2∗∗ | 42.5 | |
| BotGAT (T5) | Baseline | TG | 47.4 | 64.3 | 47.5 | 99.8 | –0.6 |
| ValueGraph | TG | 73.1∗∗ | 66.0∗∗ | 82.5∗∗ | 55.1∗∗ | 47.8∗∗ | |
| Baseline | FTUG | 73.2 | 70.9 | 74.6 | 69.3 | 47.3 | |
| ValueGraph | FTUG | 74.7∗∗ | 73.0∗ | 74.2 | 72.1 | 49.3∗∗ | |
| BotRGCN (T5) | Baseline | TG | 69.2 | 69.1 | 66.2 | 72.6 | 38.9 |
| ValueGraph | TG | 70.3 | 69.1∗∗ | 68.5∗ | 70.0∗∗ | 40.7 | |
| Baseline | FTUG | 71.0 | 73.5 | 65.0 | 84.8 | 44.7 | |
| ValueGraph | FTUG | 73.0∗ | 74.2 | 68.0∗ | 82.0∗ | 47.0 |
Detecting twitter bots is critical for maintaining the integrity of online discourse. However, most existing detectors rely primarily on surface-level textual features or network structure. We evaluate whether replacing or augmenting such cues with ValueGraph user embeddings guided by inferred value signals can improve bot detection performance across strong baseline models. We use TwiBot-22 (Feng et al. 2022) dataset for evaluation (Dataset details are in Appendix: Twitter Bot Detection).
Baselines.
We select high-performing models from the benchmark study by Feng et al. (2022), covering different data modalities. We include RoBERTa (Liu et al. 2019) as a text-based baseline, together with BotRGCN (Feng et al. 2021), BotGAT (Lei et al. 2022), and BotGCN (Feng et al. 2021), which use T5 encoder (Raffel et al. 2020) to encode text. We also add a GPT-5.4 baseline using sampled user posts and profile text with no graph serialization. Each non-LLM model is evaluated under configurations using F (user metadata or engineered features), T (posts), U (profile description), and G (network structure). We examine whether substituting T/G-based representations with ValueGraph embeddings improves performance.
Setup and Metrics.
All trainable encoders retain their original architectures and default settings. We perform grid search over hyperparameters (see Appendix: Hyperparameter Settings), run each experiment five times with different random seeds, and report Accuracy, MacF1, Precision, Recall, and Matthews Correlation Coefficient (MCC).
Results.
Table 2 shows that ValueGraph-based embeddings consistently outperform non-LLM baseline representations across models. For instance, RoBERTa in the text-only configuration improves macF1 from 54.8% to 64.3%, with recall increasing from 62.5% to 81.0%. While GPT-5.4 achieves acc generally on par with ValueGraph in the text-only setting, its macF1 is substantially lower. These gains indicate that ValueGraph captures complementary semantic and structural cues for bot-human discrimination. While the improvements are statistically significant in text+network settings, significance tends to decrease when full multimodal inputs are available, particularly for macF1 in FTUG configurations, where strong profile and graph features already dominate the decision boundary. Nevertheless, because the ValueGraph encoder is not trained with bot labels or downstream fine-tuning, these gains indicate that value-signal guidance adds useful information even in feature-rich settings.
Ablation Study
| Task | Model/data | Configuration | Accuracy | MacF1 |
| Stance | GLAN (MT_CSD) | Only Stage 1 | 41.2 | 40.7 |
| Stage 1 + | 50.4 | 34.3 | ||
| Stage 1 + | 45.3 | 32.6 | ||
| Full | 63.0 | 58.0 | ||
| BrLSTM (RumourEval19) | Only Stage 1 | 53.7 | 55.8 | |
| Stage 1 + | 62.7 | 52.8 | ||
| Stage 1 + | 53.9 | 45.7 | ||
| Full | 77.1 | 71.2 | ||
| Bot | BotGAT (TwiBot-22) | Only Stage 1 | 73.1 | 70.3 |
| Stage 1 + | 72.9 | 72.6 | ||
| Stage 1 + | 72.6 | 72.0 | ||
| Full | 74.7 | 73.0 |
To assess the contribution of value-signal information, we conduct controlled ablation experiments. Since Stage 1 pre-training is essential for capturing reply structure and producing stable representations, it is retained in all settings. We evaluate four variants: (1) Stage 1 pre-training only; (2) Stage 1 ; (3) Stage 1 ; and (4) the full model.
Table 3 shows that neither nor alone yields consistent improvements, which supports our conjecture in the Loss Function section. Notably, a single loss can raise accuracy while lowering macF1. For example, on GLAN (MT_CSD), adding improves accuracy from to but drops macF1 from to . In contrast, the full ValueGraph model, which jointly integrates structural pre-training, clustering supervision, and value-signal guidance, achieves consistent improvements across all tasks. These results validate the complete ValueGraph design (see Appendix: Supplementary Results of the Ablation Experiments for full results).
User Clustering Analysis
We use t-SNE (see Appendix: t-SNE Setting for parameter setup) to visualize how different embedding models organize users for Twitter bot detection. As shown in Figure 2, ValueGraph yields clearer bot-human separation than the compared text and graph encoders. This visualization provides dataset-specific qualitative evidence of improved representation separability.
A content-level comparison helps explain this separation. In the evaluated dataset, some bot clusters contain repetitive and polarized value signal. For example, the bot account (author_id anonymized) posted an accusatory narrative: “Paul Vickers died suddenly 3 months after I exposed his #Establishment lies … contributed to his death … #FactCheck …”, relying on blame attribution, repeated hashtags, and a one-sided moral stance. In contrast, a human account (author_id anonymized) wrote: “Is that what the point is? Being angry with the one we love is not the same as unloving them, is it?”, showing more dialogic and context-sensitive expression. These examples suggest that ValueGraph captures value signals useful for bot detection.
Conclusion
We presented ValueGraph, a value-signal guided graph pre-training framework for contextualized user representation. Grounded in MFT, ValueGraph combines semantic and structural graph pre-training, contrastive learning over inferred value similarity, and clustering to capture textual, relational, and value-relevant behavioral cues. Experiments on stance detection and Twitter bot detection show consistent gains over strong baselines. These results suggest that noisy but theory-informed value signals provide useful auxiliary guidance for socially informed user modeling without predicting users’ true values.
Ethics Statement
This work uses publicly available social-media datasets under their original access conditions and licenses. We do not infer users’ true psychological values; MoralBERT-derived vectors are used only as noisy aggregate signals for representation learning. Results should be interpreted as dataset-specific and should not be used for individual-level profiling or decision-making without appropriate safeguards. We will release code and processing scripts but will not redistribute raw social-media content or personally identifiable information. The released artifacts are intended for research purposes only.
Acknowledgment
This research is supported by the SMU-A*STAR Joint Lab in Social and Human-Centered Computing (SMU grant no.: SAJL-2022-CSS02, SAJL-2022-CSS003). This research is supported by A*STAR (C232918004, C232918005).
References
- Exploiting personal characteristics of debaters for predicting persuasiveness. In ACL, Cited by: Stance Detection.
- MoralStrength: exploiting a moral lexicon and embedding similarity for moral foundations prediction. Knowledge-Based Systems 191, pp. 105184. Cited by: Human-Value Grounding for Training Data..
- Learning multiview embeddings of Twitter users. In ACL, Cited by: Introduction, User Representation..
- Deep clustering for unsupervised learning of visual features. In Proceedings of the European Conference on Computer Vision (ECCV), Cited by: Loss Functions..
- Deep reinforcement learning from human preferences. In NIPS, Cited by: AI and Human Values..
- Multi-view unsupervised user feature embedding for social media-based substance use prediction. In EMNLP, Cited by: Introduction, User Representation..
- Twitter user geolocation using deep multiview learning. In ICASSP, Cited by: Introduction, User Representation..
- Learning structural node embeddings via diffusion wavelets. In SIGKDD, Cited by: Introduction, Graph Pre-training..
- Exploring the role of prior beliefs for argument persuasion. In NAACL, Cited by: Stance Detection.
- TwiBot-22: towards graph-based twitter bot detection. In NeurIPS, Cited by: Appendix F, Baselines., Twitter Bot Detection.
- BotRGCN: twitter bot detection with relational graph convolutional networks. In ASONAM, Cited by: Appendix F, Baselines..
- SimCSE: simple contrastive learning of sentence embeddings. In EMNLP, Cited by: Baselines..
- SemEval-2019 task 7: RumourEval, determining rumour veracity and support for rumours. In SemEval, Cited by: Appendix F, Stance Detection.
- Chapter two - moral foundations theory: the pragmatic validity of moral pluralism. P. Devine and A. Plant (Eds.), Advances in Experimental Social Psychology, Vol. 47, pp. 55–130. External Links: ISSN 0065-2601, Document, Link Cited by: Introduction.
- Liberals and conservatives rely on different sets of moral foundations. Journal of personality and social psychology 96, pp. 1029–46. Cited by: Appendix C.
- Node2vec: scalable feature learning for networks. In SIGKDD, Cited by: Introduction, Graph Pre-training..
- A data fusion framework for multi-domain morality learning. In ICWSM, Cited by: Introduction, AI and Human Values..
- GraphCL: contrastive self-supervised learning of graph representations. ArXiv abs/2007.08025. Cited by: Introduction, Graph Pre-training., Baselines..
- The moral mind: how five sets of innate intuitions guide the development of many culture-specific virtues, and perhaps even modules. The innate mind 3, pp. 367–391. Cited by: Introduction, Human-Value Grounding for Training Data..
- Contrastive multi-view representation learning on graphs. In ICML, Cited by: Introduction, Graph Pre-training..
- Culture’s consequences: comparing values, behaviors, institutions and organizations across nations. Sage Publications. Cited by: Human-Value Grounding for Training Data..
- GraphMAE2: a decoding-enhanced masked self-supervised graph learner. In WWW, Cited by: Stage 1: Foundation GNN Pre-training., Baselines..
- Strategies for pre-training graph neural networks. In ICLR, Cited by: Graph Pre-training..
- GPT-gnn: generative pre-training of graph neural networks. In SIGKDD, Cited by: Graph Pre-training..
- Classification of moral foundations in microblog political discourse. In ACL, Cited by: Human-Value Grounding for Training Data..
- Semi-supervised classification with graph convolutional networks. arXiv:1609.02907. Cited by: Baselines..
- BIC: twitter bot detection with text-graph interaction and semantic consistency. arXiv:2208.08320. External Links: Link Cited by: Appendix F, Baselines..
- Self-supervised learning with kernel dependence maximization. In NeurIPS, Cited by: Similarity Computation..
- Detect rumors in microblog posts for low-resource domains via adversarial contrastive learning. In Findings of NAACL, Cited by: Appendix F, Table 5, Pre-training Corpus.
- Acquiring background knowledge to improve moral value prediction. In ASONAM, Cited by: Human-Value Grounding for Training Data..
- Roberta: a robustly optimized bert pretraining approach. arXiv:1907.11692. Cited by: Baselines..
- Learning to pre-train graph neural networks. In AAAI, Cited by: Graph Pre-training..
- Detecting rumors from microblogs with recurrent neural networks. In IJCAI, Cited by: Appendix F, Table 5, Pre-training Corpus.
- Moral framing and ideological bias of news. In SocInfo, Cited by: Human-Value Grounding for Training Data..
- Moralization in social networks and the emergence of violence during protests. Nature human behaviour 2 (6), pp. 389–396. Cited by: Human-Value Grounding for Training Data..
- BERTweet: a pre-trained language model for English tweets. In EMNLP, Cited by: Baselines..
- Measuring moral dimensions in social media with mformer. In ICWSM, Cited by: Introduction, AI and Human Values..
- A challenge dataset and effective models for conversational stance detection. In LREC-COLING, Cited by: Appendix F, Stance Detection.
- Training language models to follow instructions with human feedback. In NeurIPS, Cited by: AI and Human Values..
- Social media-based user embedding: a literature review. In IJCAI, Cited by: Introduction, User Representation., Problem Formulation.
- DeepWalk: online learning of social representations. In SIGKDD, Cited by: Introduction, Graph Pre-training..
- Stem: unsupervised structural embedding for stance detection. In AAAI, Cited by: Introduction.
- MoralBERT: a fine-tuned language model for capturing moral values in social discussions. In GoodIT, Cited by: Introduction, Introduction, AI and Human Values..
- GCC: graph contrastive coding for graph neural network pre-training. In SIGKDD, Cited by: Graph Pre-training..
- Exploring the limits of transfer learning with a unified text-to-text transformer. J. Mach. Learn. Res. 21 (1). Cited by: Baselines..
- Twitter user geolocation using a unified text and network prediction model. In ACL-IJCNLP, Cited by: Introduction, User Representation..
- Characterizing and detecting hateful users on twitter. In ICWSM, Cited by: User Representation..
- Auditing radicalization pathways on youtube. In FAT, Cited by: Introduction.
- Refining the theory of basic individual values. Journal of Personality and Social Psychology 103 (4), pp. 663. Cited by: Human-Value Grounding for Training Data..
- InfoGraph: unsupervised and semi-supervised graph-level representation learning via mutual information maximization. In ICLR, Cited by: Introduction, Graph Pre-training..
- Multi-stage self-supervised learning for graph convolutional networks. In AAAI, Cited by: Introduction, Graph Pre-training..
- LINE: large-scale information network embedding. In WWW, Cited by: Graph Pre-training..
- The moral foundations reddit corpus. arXiv:2208.05545. Cited by: Appendix F, Human-Value Grounding for Training Data., Pre-training Corpus.
- Graph attention networks. stat 1050 (20), pp. 10–48550. Cited by: Baselines..
- Smarter, better, faster, longer: a modern bidirectional encoder for fast, memory efficient, and long context finetuning and inference. In ACL, Cited by: Stage 1: Foundation GNN Pre-training., Baselines..
- Graph contrastive learning with augmentations. In NeurIPS, Cited by: Graph Pre-training..
- Jointly embedding the local and global relations of heterogeneous graph for rumor detection. External Links: 1909.04465, Link Cited by: Stance Detection.
- Multi-target stance detection with bi-directional recursive encoding and classification. In IJCAI, Cited by: Stance Detection.
- Graph transformer networks. In NeurIPS, Cited by: Baselines..
- ME2-bert: are events and emotions what you need for moral foundation prediction?. In COLING, Cited by: AI and Human Values..
- Early rumor detection using neural Hawkes process with a new benchmark dataset. In NAACL, Cited by: Appendix F, Table 5, Pre-training Corpus.
- User profile preserving social network embedding. In IJCAI, Cited by: User Representation..
- Enhancing stance classification on social media using quantified moral foundations. In ASONAM, Cited by: Human-Value Grounding for Training Data..
- ProNE: fast and scalable network representation learning. In IJCAI, Cited by: Graph Pre-training..
- User-guided hierarchical attention network for multi-modal social image popularity prediction. In WWW, Cited by: User Representation..
- Data augmentation for graph neural networks. In AAAI, Cited by: Introduction, Graph Pre-training..
- Graph prompt learning for stance detection. arXiv:2403.11145. Cited by: Stance Detection.
- Deep graph contrastive representation learning. ArXiv abs/2006.04131. Cited by: Graph Pre-training..
- PHEME: a dataset for fine-grained stance detection. arXiv:1610.07363. Cited by: Appendix F, Table 5, Pre-training Corpus.
Appendix
Appendix A Full Proof of Theorem 1
Proof.
The gradient of the InfoNCE loss with respect to the embedding takes the form
| (12) |
where, for simplicity, all embeddings are assumed to be normalized so that cosine similarity can be treated as a dot product. Define
| (13) | ||||
Then the loss can be written as
| (14) | ||||
Differentiating the first term gives
| (15) |
Differentiating the second term gives
| (16) |
Define the softmax probabilities as
| (17) | ||||
The gradient of the second term is therefore
| (18) |
and the full gradient becomes
| (19) |
By construction, positive pairs have high inferred value-signal similarity, while negative pairs have low inferred value-signal similarity. The InfoNCE objective therefore pulls toward embeddings of value-similar users and pushes it away from value-dissimilar users. Under gradient descent, the update for is
| (20) |
When and are close, the positive pair dominates the softmax, so the update primarily pulls toward . This yields
| (21) |
Thus, the learned embedding similarity is encouraged to preserve the ordering induced by inferred value-signal similarity. ∎
Appendix B Full Proof of Theorem 2
Proof.
The compactness loss
| (22) |
minimizes the squared Euclidean distance between each user embedding and its assigned centroid . If deviates from , the loss increases; optimization therefore pulls embeddings toward their assigned centroids and encourages intra-cluster compactness.
The margin loss is
| (23) |
For any pair of distinct centroids and , if , the corresponding loss term is zero and no further separation force is applied. If , the pair contributes
| (24) |
Let . The gradient with respect to is
| (25) |
with a symmetric expression for . Thus, when , the gradient pushes the centroids apart until their distance reaches the margin.
The total loss is
| (26) |
where includes both compactness and margin terms. Therefore, minimizing jointly encourages embeddings within the same cluster to remain close while pushing different cluster centroids apart. This establishes the intended compactness and separation properties. ∎
Appendix C Moral Foundations Background
Moral Foundations Theory (MFT): To assess individuals’ moral foundations, Graham et al. (2009) developed survey-based questions using factor analysis. We summarize the core moral foundations below:
- •
Care/Harm: This foundation reflects the tendency to form emotional bonds and experience distress at others’ suffering. It emphasizes kindness, compassion, and nurturance, while discouraging harm.
- •
Fairness/Cheating: Rooted in reciprocal altruism, this foundation promotes justice, equity, proportionality, and autonomy.
- •
Loyalty/Betrayal: This foundation reflects loyalty to one’s in-group or community. It supports patriotism and self-sacrifice, but in extreme cases can lead to nepotism or favoritism.
- •
Authority/Subversion: This foundation reflects respect for leadership, hierarchy, tradition, and established social structures.
- •
Sanctity/Degradation: This foundation reflects the inclination to uphold purity and avoid contamination. It is often tied to religious and cultural ideals of moral elevation.
Appendix D Distribution of Sampled User Similarities
As shown in Figure 3, similarity scores are concentrated at lower values, with a long tail toward higher similarities. This indicates that most user pairs in the inferred value-profile space have weak affinity, while a small subset shows strong alignment.
Based on this distribution, we use percentile-based sampling, selecting positive pairs from the 99.5-th percentile and negative pairs from the 0.5-th percentile. This strategy yields discriminative contrastive examples, reduces ambiguity from moderately similar pairs, and preserves enough samples for stable training.
Appendix E Training Loss and Hyperparameter Tuning
To optimize ValueGraph, we conduct a grid search over key hyperparameters in the loss formulation. The objective combines the user-level contrastive loss and the clustering loss . Final hyperparameters are selected based on convergence behavior and downstream validation performance. The tuned hyperparameters are:
- •
Temperature : Controls the sharpness of similarity distributions in the InfoNCE loss. Smaller emphasizes hard negatives, while larger produces smoother gradients and more stable optimization.
- •
Margin coefficient : Scales the margin term in the clustering loss. Larger enforces stronger inter-cluster separation, while smaller allows more flexible cluster boundaries.
- •
Number of augmented users : The number of users sampled per seed user during contrastive training. Larger increases training diversity but also computational cost.
- •
Number of clusters : Controls the granularity of user grouping in the embedding space.
- •
Margin : Specifies the minimum distance encouraged between cluster centroids.
We report convergence plots for representative configurations in Figure 4. The final configuration is: temperature , margin coefficient , augmented users , clusters , clustering weight , and margin .
Appendix F Evaluation Tasks and Datasets
Details of Pre-training Datasets
We provide additional statistics of the pre-training datasets, including source domains, topic distributions, and graph statistics.
Reddit Corpus.
Our Reddit corpus is constructed from the Reddit Corpus provided by ConvoKit66 6 https://convokit.cornell.edu/documentation/subreddit.html, following the dataset construction procedure of Trager et al. (2022). It contains conversations collected from multiple subreddits representing diverse social communities and discussion topics. The selected subreddits cover political discussions, interpersonal relationships, social issues, and general-interest communities. We retain conversation threads with at least 50 comments to ensure sufficient interaction structures.
After preprocessing, including removing inaccessible comments, filtering isolated nodes, and preserving reply-based interactions, the Reddit corpus contains approximately 9.6 million comments. The detailed distribution across subreddits is shown in Table 4.
| Subreddit | Comments |
| Politics | 5,626,191 |
| Neoliberal | 2,315,326 |
| Relationship Advice | 861,704 |
| Confession | 477,353 |
| Conservative | 165,182 |
| Worldnews | 100,194 |
| AmItheAsshole | 83,624 |
| Geopolitics | 41,700 |
| Antiwork | 452 |
| Nostalgia | 378 |
| Total | 9,672,104 |
Twitter Corpus.
Our Twitter corpus combines several publicly available social media conversation datasets, including PHEME (Zubiaga et al. 2016), Twitter16 (Ma et al. 2016), BEARD (Zeng and Gao 2022), and Twitter-Covid (Lin et al. 2022). These datasets cover diverse scenarios such as rumor propagation, public health discussions, and event-driven social interactions.
During preprocessing, we remove isolated nodes without parent or child interactions and preserve only users participating in conversational structures. The final Twitter corpus contains over 5 million posts. The detailed statistics of each source dataset are summarized in Table 5.
| Dataset | Posts |
| BEARD (Zeng and Gao 2022) | 3,393,578 |
| Twitter-Covid (Lin et al. 2022) | 474,710 |
| PHEME (Zubiaga et al. 2016) | 59,387 |
| Twitter16 (Ma et al. 2016) | 1,101,985 |
| Total | 5,029,660 |
Stance Detection
MT_CSD (Niu et al. 2024) is a benchmark for Multi-Target Conversational Stance Detection, constructed from authentic Reddit interactions. It contains 15,876 human-annotated instances over extended conversation threads, with approximately 76% involving more than three reply turns. The dataset includes Reddit posts, popularity metrics, and discussions centered on targets such as Tesla, SpaceX, Donald Trump, Joe Biden, and Bitcoin.
RumourEval19 (Gorrell et al. 2019) Task A is a Twitter benchmark for rumor stance classification in conversational threads. It contains 15,874 tweets and provides tweet IDs and reply-to IDs, enabling reconstruction of tree- or graph-structured conversations.
| Split | Label | Count |
| Train | Against | 3,349 |
| Favor | 2,383 | |
| None | 5,441 | |
| Valid | Against | 710 |
| Favor | 481 | |
| None | 1,115 | |
| Test | Against | 728 |
| Favor | 470 | |
| None | 1,197 |
| Split | Label | Count |
| Train | Comment | 2,734 |
| Deny | 333 | |
| Query | 330 | |
| Support | 841 | |
| Dev | Comment | 173 |
| Deny | 11 | |
| Query | 28 | |
| Support | 69 |
Twitter Bot Detection
TwiBot-22 (Feng et al. 2022) is a graph-based bot detection benchmark and one of the largest annotated Twitter account datasets. It models a heterogeneous graph with user, tweet, and hashtag nodes connected by multiple edge types. We select representative baselines from TwiBot-22 and categorize them by modality: F (user metadata), T (tweet and description content), and G (Twitter network structure).
For BotRGCN (Feng et al. 2021), BotGAT (Lei et al. 2022), and BotGCN (Feng et al. 2021), we concatenate our 512-dimensional GNN-generated user embeddings with the original user features, including numerical properties (num_prop5), categorical properties (cat_prop3), tweets (tweet), and user descriptions (des).
Appendix G Additional Experimental Details
Algorithm
Algorithm 1 summarizes the training procedure of ValueGraph. At each epoch, we first sample seed users and construct an augmented user batch by retrieving similar and dissimilar users according to their inferred value profiles. The corresponding posts are encoded by the graph encoder, and user embeddings are obtained by aggregating post representations. The value-guided contrastive loss is then computed to align users with similar value profiles while separating dissimilar ones. To further regularize the global structure of the embedding space, we periodically perform -means clustering every epochs and compute the clustering loss . The final objective combines both losses to update the encoder parameters.
Hyperparameter Settings
We report the hyperparameter configurations used in our experiments. Tables 8 and 9 and summarize the settings for stance detection and Twitter bot detection, respectively. Hyperparameters are tuned for each model and dataset. Additional search may further improve performance for specific embedding configurations.
| Model | Hidden Dim | Learning Rate | Weight Decay | Dropout | Batch Size | Max Epoch |
| GLAN (MT_CSD) | 512 | 1e-5 | 1e-4 | 0.2 | 32 | 100 |
| BrLSTM (RumourEval19) | 300 | 1e-4 | 1e-4 | 0.5 | 16 | 100 |
| Model | Hidden Dim | Learning Rate | Weight Decay | Dropout | Batch Size | Max Epoch |
| RoBERTa (T) | 128 | 1e-4 | 1e-5 | 0.5 | 96 | 60 |
| RoBERTa (T,U) | 128 | 1e-4 | 1e-5 | 0.5 | 96 | 60 |
| BotRGCN (T,G) | 256 | 5e-4 | 5e-5 | 0.5 | 128 | 300 |
| BotRGCN (F,T,U,G) | 128 | 5e-4 | 1e-3 | 0.3 | 128 | 300 |
| GAT (T,G) | 96 | 2e-4 | 5e-5 | 0.5 | 128 | 300 |
| GAT (F,T,U,G) | 96 | 2e-4 | 5e-5 | 0.5 | 256 | 300 |
| GCN (T,G) | 256 | 2e-4 | 1e-4 | 0.3 | 128 | 300 |
| GCN (F,T,U,G) | 256 | 5e-4 | 1e-3 | 0.3 | 128 | 300 |
Supplementary Results of the Ablation Experiments
We report supplementary ablation results in Table 10, which are omitted from the main text due to space limitations.
Table 10 evaluates ValueGraph under different training configurations. Using only Stage 1 yields limited performance across tasks. Adding the clustering loss generally improves task-aware discrimination, while incorporating the user-level consistency loss further helps capture user interaction structure and value-aware relational information. Overall, the full model achieves the strongest and most stable performance across models and tasks, supporting the effectiveness and generalizability of ValueGraph.
| Task | Model/dataset | Configuration | Accuracy | F1-score |
| Stance Detection | GLAN (MT_CSD) | Only Stage 1 | 41.2 | 40.7 |
| Stage 1 + | 50.4 | 34.3 | ||
| Stage 1 + | 45.3 | 32.6 | ||
| Full | 63.0 | 58.0 | ||
| BrLSTM (RumourEval19) | Only Stage 1 | 53.7 | 55.8 | |
| Stage 1 + | 62.7 | 52.8 | ||
| Stage 1 + | 53.9 | 45.7 | ||
| Full | 77.1 | 71.2 | ||
| Twitter Bot Detection | RoBERTa (Twibot-22) | Only Stage 1 | 63.6 | 68.9 |
| Stage 1 + | 64.9 | 67.8 | ||
| Stage 1 + | 64.7 | 67.5 | ||
| Full | 65.9 | 68.9 | ||
| BotGCN (Twibot-22) | Only Stage 1 | 70.3 | 71.0 | |
| Stage 1 + | 71.2 | 71.8 | ||
| Stage 1 + | 70.9 | 71.7 | ||
| Full | 71.1 | 71.1 | ||
| BotRGCN (Twibot-22) | Only Stage 1 | 71.6 | 73.3 | |
| Stage 1 + | 72.8 | 73.4 | ||
| Stage 1 + | 72.5 | 73.1 | ||
| Full | 73.0 | 74.2 | ||
| BotGAT (Twibot-22) | Only Stage 1 | 73.1 | 70.3 | |
| Stage 1 + | 72.9 | 72.6 | ||
| Stage 1 + | 72.6 | 72.0 | ||
| Full | 74.7 | 73.0 |
GPU Hours
All models are run with fixed random seeds for reproducibility. Training and evaluation are conducted on a compute cluster with two NVIDIA H100 GPUs. The full suite of experiments consumes approximately 1,000 GPU hours, including two-stage ValueGraph pre-training and downstream evaluations under both baseline and ValueGraph configurations.
ValueGraph training consists of Stage 1, which uses textual and network signals to learn initial representations, and Stage 2, which incorporates inferred value signals to produce final user embeddings. These embeddings replace the corresponding conventional features in downstream evaluations.
t-SNE Setting
For consistency across models, we apply t-SNE to project high-dimensional embeddings into two dimensions, using n_components=2, perplexity=30, n_iter=1000, and random_state=42. This configuration preserves local neighborhood structure while ensuring reproducibility.
GPT-5.4 Prompt Templates
We include GPT-5.4 as a text-only LLM baseline for stance detection and Twitter bot detection. GPT-5.4 receives only textual inputs, including the target, user posts, thread text, or profile text when available; it does not receive explicit reply-graph adjacency or message-passing structure. We use deterministic decoding with temperature set to 0. The prompts used for the two tasks are shown below.
Stance Detection Prompt.
For stance detection, the model is given the target, the source post or thread context, and the user’s available posts. It is asked to predict one of the task labels. In our default setting (text_mode=leaf), MT-CSD uses the discussion topic together with the leaf comment, and RumourEval19 uses the reply text only.
Task: Determine the user’s stance toward the given target.
Target: {target}
Conversation context: {thread_text}
User posts: {user_posts}
Choose exactly one label from: Favor, Against, None.
Return only the label.
In practice, for MT-CSD we instantiate {target} with the discussion topic and use the leaf comment as the textual input; for RumourEval19, we use the reply text as the input and omit separate thread or user-history fields in the default configuration. The implemented MT-CSD prompt is:
You are annotating stance in a social-media discussion about the topic: {topic}.
Classify the STANCE of the following comment toward the discussion target. Use exactly one label: - favor: supports/agrees with the target stance - against: opposes/disagrees with the target stance - none: neutral, unrelated, or unclear stance
Comment: {text}
Reply with exactly one word: favor, against, or none.
For RumourEval19, we replace the label set with the dataset-specific labels:
Choose exactly one label from: Support, Deny, Query, Comment.
Return only the label.
The implemented RumourEval19 prompt is:
You are annotating stance toward a rumor in social media.
Classify the STANCE of the following reply (SDQC scheme). Use exactly one label: - support: supports/agrees the rumor is true - deny: refutes or disagrees with the rumor - query: asks for evidence or clarification - comment: neutral or unrelated to rumor veracity
Reply text: {text}
Reply with exactly one word: support, deny, query, or comment.
Twitter Bot Detection Prompt.
For bot detection, the model is given the user’s profile text and sampled posts. It is asked to classify the account as human or bot. In our TwiBot experiments, profile text is not available; we therefore provide only the user’s test-split posts, aggregated at the account level.
Task: Determine whether the following Twitter account is operated by a human or a bot.
User profile: {profile_text}
User posts: {sampled_posts}
Choose exactly one label from: Human, Bot.
Return only the label.
In practice, we set {profile_text} to empty and construct {sampled_posts} by concatenating all available tweets from the same user in the test split, separated by \n---\n. The implemented prompt is:
You are detecting whether a Twitter/X account is operated by a human or a bot.
Read the following tweets posted by ONE account (may be truncated). Classify the account type. Use exactly one label: - human: likely a real person - bot: likely automated, spam, or bot-like
Tweets from this account: {text}
Reply with exactly one word: human or bot.
For all prompts, long inputs are truncated to fit the model context window while preserving the target, profile text, source post, and the most recent or most relevant user posts. In our implementation, the truncation limit is 12,000 characters. No examples from the test set are included in the prompt.