<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Conversations on Twitter: Structure, Pace, Balance</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Danica Vukadinovi´c Greetham</string-name>
          <email>d.v.greetham@reading.ac.uk</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Jonathan A. Ward</string-name>
          <email>j.a.ward@leeds.ac.uk</email>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Centre for the Mathematics of Human Behaviour Department of Mathematics and Statistics University of Reading</institution>
          ,
          <country country="UK">UK</country>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>Department of Applied Mathematics University of Leeds</institution>
          ,
          <country country="UK">UK</country>
        </aff>
      </contrib-group>
      <fpage>99</fpage>
      <lpage>110</lpage>
      <abstract>
        <p>Twitter is both a micro-blogging service and a platform for public conversation. Direct conversation is facilitated in Twitter through the use of @'s (mentions) and replies. While the conversational element of Twitter is of particular interest to the marketing sector, relatively few data-mining studies have focused on this area. We analyse conversations associated with reciprocated mentions that take place in a data-set consisting of approximately 4 million tweets collected over a period of 28 days that contain at least one mention. We ignore tweet content and instead use the mention network structure and its dynamical properties to identify and characterise Twitter conversations between pairs of users and within larger groups. We consider conversational balance, meaning the fraction of content contributed by each party. The goal of this work is to draw out some of the mechanisms driving conversation in Twitter, with the potential aim of developing conversational models.</p>
      </abstract>
      <kwd-group>
        <kwd>Twitter mentions networks</kwd>
        <kwd>conversations models</kwd>
        <kwd>maximal cliques</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>-</title>
      <p>
        The rapid uptake of online social media, combined with consumer behavioural
changes around television and news broadcasting, has instigated a sea change
in attitudes within the advertising and marketing sectors. A frequently
encountered adage is that “everything is about conversation and not about
broadcasting” [
        <xref ref-type="bibr" rid="ref10 ref6">10,6</xref>
        ]. By facilitating public addressability through the @ sign (so called
‘mentions’) and enabling private messages, Twitter has confirmed their
intention to function as a communication channel as well as a broadcasting tool.
Access to large quantities of data produced by Twitter users has resulted in a
surge of interest from the academic community [
        <xref ref-type="bibr" rid="ref20">20</xref>
        ], who have largely focused
on Twitter’s information flow and retweet behaviour, and hence implicitly the
underlying network of ‘followers’ (e.g. [
        <xref ref-type="bibr" rid="ref21 ref22">22,21</xref>
        ]). While broadcasting short
messages, or micro-blogging, remains an important component of Twitter use, to
our knowledge comparatively little work has addressed the mining of (public)
conversations on a large scale [
        <xref ref-type="bibr" rid="ref14 ref19 ref3">3,19,14</xref>
        ]. Consequently, we focus in this paper
on analysing the network of communication patterns resulting from mentions in
Twitter.
      </p>
      <p>
        Although it may not always be clear, even from message content, what
intention a user had in mind when posting—information seeking or information
sharing, broadcasting or conversation—we have tried to specifically extract
conversations by focusing our data-analysis on reciprocated tweets. Moreover, we
have completely ignored the content of conversations and concentrated on
structural and dynamic properties of the underlying mentions network. Our main
objective was to mine actionable insights that could inform our knowledge of
conversational mechanisms and the frequency/timings of tweets. Our hope is
that empirical observations and quantifiable insights from this analysis could
inform a simple, data driven model of the timing and structure of Twitter
conversations. One possible application would be for automated recommendations
of conversation trends, as discussed in [
        <xref ref-type="bibr" rid="ref1 ref3">3,1</xref>
        ].
      </p>
      <p>
        A large number of registered Twitter accounts are operated by automated
software scripts, known as bots [
        <xref ref-type="bibr" rid="ref18">18</xref>
        ]. While such accounts are encouraged for
the purpose of developing applications and services, bots whose functions
violate Twitter policy (e.g. spammers) are common. The analysis of conversational
patterns and the development of associated models have potential application
for those trying to develop algorithms that can identify nuisance bots.
Furthermore, the identification of groups of Twitter users who, through conversational
behaviour, are particularly influential on a specific topic would be particularly
attractive in the marketing sector. Thus, understanding conversational
structure could impact the design and implementation of social media campaigns
and potentially provide a quantitative comparison between Twitter discourse
and other channels of communication, such as face-to-face, telephone, SMS,
forums or email. In addition, curating and recommending conversational trends,
for both Twitter and more generally in online social media, is crucial for social
networking sites as it is one of the main characteristics of user experience. We
believe that a better understanding of the structure, dynamics and balance of
multi-user conversation is key to improving such automated curation systems.
Ultimately, we hope that studying Twitter conversation can ultimately improve
user experience.
      </p>
      <p>In Section 2, we give an account of previous work in this space. Our results
of pairwise and multiple conversations and the Twitter dataset we used are
presented in Section 3. Finally, in Section 4 we summarise and describe possible
directions of future work.
2</p>
    </sec>
    <sec id="sec-2">
      <title>Previous work</title>
      <p>
        The phenomenal uptake of Twitter over the last few years has resulted in a
rapidly growing interest in mining Twitter data and particularly sentiment
analysis of tweets. A recent study analyzing a large amount of Twitter and
Facebook data [
        <xref ref-type="bibr" rid="ref12">12</xref>
        ] found correlations between friendship/follower relations and
positive/negative moods of Twitter users. Diurnal and seasonal mood rhythms that
are common across di↵erent cultures have also been identified in cross-cultural
Twitter data [
        <xref ref-type="bibr" rid="ref5">5</xref>
        ], shedding light on the dynamics of positive and negative a↵ect.
      </p>
      <p>
        A study of conversations within a sample of 8.5k tweets collected over an hour
long period [
        <xref ref-type="bibr" rid="ref9">9</xref>
        ] found that the @ sign appeared in about 30% of the collected
sample, its function was mostly for addressing (as intended) and it was relatively well
reciprocated—around 30% of messages containing an @ were reciprocated within
an hour. The majority of these conversations were short, coherent exchanges
between two people, but longer exchanges did occur, sometimes consisting of up
to 10 people. They found that
      </p>
      <p>
        “...Tweets with @ signs are more focused on an addressee, more likely
to provide information for others, and more likely to exhort others to do
something—in short, their content is more interactive. ”
Twitter conversations also contain both momentarily salient or ‘peaky’ topics,
signified by increased word-use frequency of specific terms, as well as more
‘persistent conversations’, in which less salient terms recur over longer periods [
        <xref ref-type="bibr" rid="ref14">14</xref>
        ].
In addition, words that relate to negative emotions are less persistent [
        <xref ref-type="bibr" rid="ref22">22</xref>
        ].
      </p>
      <p>
        In [
        <xref ref-type="bibr" rid="ref3">3</xref>
        ], several algorithms for recommending conversations based on the lengths,
topic and ‘tie-strength’3 of conversations were compared. Their results showed
that the di↵erent uses of Twitter (social vs. informational) had a big influence
on the algorithm’s performance — recommendations based on tie strength were
preferred by social users, whilst those based on topic were preferred by
informational users. Related work considered automated curation of online conversations
to present discussion threads of interest to users in e.g. Facebook and Google+.
[
        <xref ref-type="bibr" rid="ref1">1</xref>
        ]. Key to this was the prediction of conversation length around a topic and
re-entry of interlocutors. In another work concerning Twitter conversation [
        <xref ref-type="bibr" rid="ref13">13</xref>
        ],
a relatively large corpus and content (topic) analysis of 1.3 million tweets was
used to develop an unsupervised model of dialogue from open-topic data.
      </p>
      <p>In our work we completely ignore content, instead focusing on timing,
structure and balance of conversation between pairs of individuals as well as multi-user
conversations. Our contribution is an attempt to map the structure of Twitter
exchanges over a relatively large dataset, while o↵ering some new methods to
mine conversation data and improve statistical models of dialogue.
3
3.1</p>
    </sec>
    <sec id="sec-3">
      <title>Analysis</title>
      <p>Data
The Twitter data-set investigated in this paper was collected on our behalf
by Datasift, a certified Twitter partner, allowing us to access the full Twitter
3 Tie-strength is an increasing function of the number of exchanged messages between
two people and the number of messages exchanged between them and their mutual
friends.
firehose rather than being rate-limited by the API. The data-set consists of all
UK based4 Twitter users that sent tweets with at least one mention between
8 Dec 2011 and 4 Jan 2012 (28 days in total). In the remainder of the paper,
use of the word ‘tweet’ will specifically mean tweets containing at least one
mention. Mentions are messages that include an @ followed by a username.
Thus if person a puts “@b”, it designates that a is addressing the tweet to b
specifically. Mentions are not private messages and can be read by anyone who
searches for them. A tweet can be addressed to several users simultaneously using
@ repetitively. Any Twitter user can mention any other Twitter user, they don’t
have to be related in any way. Since conversational characteristics are influenced
by many factors, including language, culture, community membership etc., one
has to keep in mind the natural limitations of the results of our analysis.</p>
      <p>We preprocessed the data, removing empty mentions and self-addressing5
and created a directed multigraph, or mentions network, containing 3, 614, 705
timestamped arcs (individual mentions) from a total of 819, 081 distinct
usernames, or nodes. Of these distinct usernames, 732, 043 were “receivers”, i.e. to
whom a message was addressed, and 137, 184 were “tweeters”, i.e. people who
tweeted a message with a mention. There were approximately 50k nodes that
appeared both as tweeters and receivers. Note that our graph is a multigraph,
meaning that multiple arcs are allowed between pairs of nodes, each having a
direction and timestamp.
3.2</p>
      <p>
        Conversations
An important feature of both face-to-face conversation [
        <xref ref-type="bibr" rid="ref15 ref16">16,15</xref>
        ] and
computermediated communication [
        <xref ref-type="bibr" rid="ref8">8</xref>
        ], is the process of turn-taking. Thus in sequences of
mentions between pairs of users, say a and b, we might expect that sequences
like ABABAB would be more common than say AAABBB, where we use A
to denote that party a mentions party b and likewise B to denote that party b
mentions party a.
      </p>
      <p>To establish if this is the case, we assume the null hypotheses that
contributions are independent events with probability PA that party a contributes to
a conversation and thus probability PB = 1 PA that party b contributes. For
a given interaction sequence of length N between parties a and b, we are
interested in the number of occurrences of B following A and vice-versa. We call these
transitions, thus the sequence ABAABBA of length N = 7, has 4 transitions.
Note that we focus on reciprocated interactions, meaning that each party makes
at least one contribution and consequently that there is by default at least one
transition in all interactions that we consider. We call the remaining transitions
the excess transitions. For any sequence of length N , the maximum possible
number of excess transitions is clearly N 2. Under the null hypotheses, excess
4 All Twitter users appearing in our data-set had selected the UK as their location.
5 Self-mentioning was surprisingly common in the data-set: 12,680 die↵rent users
created a total of 44,319 self-mentions, with the maximum being 5,586 from an
automated service that advertises itself at the end of each tweet.
transitions occur with probability PT = 2PA(1 PA). Since we assume that
transitions are independent, the probability distribution of a given number of
excess transitions is binomial, and thus the expected number is ET = (N 2)PT
with variance VT = (N 2)PT (1 PT ).</p>
      <p>To test the null hypothesis, we consider all reciprocated pairwise interaction
sequences in our Twitter data-set. For each sequence having nX contributions
from party X 2 { A, B}, we assume that the probability of party a contributing
is simply nA/(nA + nB). This does not yield any problematic probabilities (i.e.
0 or 1) since both parties always make at least one contribution.</p>
      <p>Each sequence may have a di↵erent number of interactions and a di↵erent
transition probability, but assuming that the pairwise interactions are
independent, the expectation and variance of the ensemble is simply equal to the sum of
the interaction expectations and variances respectively. Doing this, we find that
the expected number of transitions is 85,390 with a standard deviation of 226.3,
but we observe 88,758 transitions in practice, more than 15 standard deviations
above the expected value. We take this as strong evidence that we can reject
the null hypothesis and thus infer that the data contains a significant level of
turn-taking and hence conversation.</p>
      <p>Each sequence of pairwise interactions may constitute a number of di↵erent
conversations, but ascertaining when one conversation ends and another begins
may be an extremely dicult task, especially when the goal is to apply an
automated processes to a large data-set. Instead of using a time-intensive lexical
analysis, we investigate whether we can detect conversations by applying a simple
threshold rule to the time gap between responses, where we assume that a time
gap that is larger than the threshold indicates the start of a new conversation.</p>
      <p>This method requires that we can identify a suitable threshold. To achieve
this, we divide each sequence of pairwise interactions up according to a given
threshold, then define distinct conversations to be reciprocated sub-sequences,
i.e. sequences containing a contribution from both parties. Thus the number of
sub-sequences nI is always larger than the number of distinct conversations nC.
In Fig. 1(a) and (b) we plot the mean number of sub-sequences and the mean
number of distinct conversations respectively over a range of threshold values.
The number of distinct conversations nC has a peak value at approximately
9hrs. This peak is expected, since we only count reciprocated interactions as
distinct conversations. Thus small threshold values, which split an interaction
sequence up into a large number of short sub-sequences (see Fig. 1(a)), result in
relatively few distinct conversations because many of the sub-sequences feature
contributions from only one party. High threshold values also result in a small
number of conversations, but this is simply because they do not split the sequence
up into many sub-sequences. Thus the maximum at 9hrs is a natural choice
of threshold and corresponds to one’s intuition that conversations may reflect
diurnal patterns.</p>
      <p>The mean and median number of tweets during conversations were 13.09 and
4 respectively, but the distribution was heavy tailed (see Fig. 2).
5.5
5
4
3.5
30
1.35
1.3</p>
      <p>We now consider whether the number of contributions from each party are
similar, or ‘balanced’ within pairwise interactions and conversations. For a given
sequence of tweets, there are two ways to compute balance, we can either
consider the ratio of means b = hmax(nA, nB)i/hmin(nA, nB)i or the mean of
ratios = hmax(nA, nB)/ min(nA, nB)i. We will use the subscripts ‘I’ and ‘C’
to denote whether these have been calculated for interactions or conversations
respectively. Since we only consider reciprocated interactions, both quantities
are well-defined and we would generally expect b &lt; . For the total number of
interactions between pairs, we find that bI = 2.424 and I = 3.457. Thus on
average, one party contributes around 3 times as much as the other. For the
sub-set of conversations, we find that bC = 1.148 and C = 1.425. These are
much closer to 1, and hence more what we would expect from typical, balanced
conversations. The distribution of conversation contribution ratios is plotted in
Fig. 3(a), which illustrates that conversations are most likely to be balanced,
but some extremely unbalanced conversations do occur. In Fig. 3(b), for each
100
x
a
m
n
50
00
101
Balance
102
50 nmin 100
150
minimum conversation contribution nmin = 1, 2, 3, . . . , we compute the mean of
the maximum contribution nmax. There is a roughly linear trend (the grey line
is nmax = 1.148nmin + 1), which further illustrates conversational balance.
By allowing multiple @ signs in one message, a Twitter user could send a tweet
to several recipients simultaneously, facilitating multi-user conversations or
multicasting. Note that because of the 140 character limit there is a physical limit
on how many users each message can be multicast to.</p>
      <p>In this part of analysis, our aim is to
– Identify multi-users exchanges;
– Determine how many users typically engage in them;
– Identify their time-frame, pace and how balanced they are.</p>
      <p>In addition, are all users equally involved, or do some dominate the discussion?
Are the same people at the heart of di↵erent multi-user conversations? What
are the enablers and inhibitors of conversation flowing in the sense of pauses
between consecutive contributions?
3.4</p>
      <p>
        Identification of multi-users conversations
The reciprocated mentions data represents a directed multi-graph G (where
an edge from A to B implies at least one edge from B to A), thus
multiuser exchanges correspond to strongly-connected6 subgraphs of G with k &gt; 2
participants. We ran a non-recursive version of Tarjan’s algorithm [
        <xref ref-type="bibr" rid="ref11 ref17">17,11</xref>
        ], as
6 A directed graph is called strongly-connected if there is a path from each vertex in
the graph to every other vertex. This means that for two vertices a and b there is
a path in both directions, i.e. from a to b and also from b to a. Strongly-connected
components of a graph are maximal subgraphs that are strongly-connected.
implemented in NetworkX [
        <xref ref-type="bibr" rid="ref7">7</xref>
        ], to get a list of the strongly-connected
components of G. Pairwise conversations were discussed in Section 3.2, so we excluded
all strongly-connected components of size 2 from the present analysis. Each
strongly-connected component of at least three vertices was then transformed
into an undirected multi-graph and we ran the NetworkX implementation of the
modified Bron’s algorithm [
        <xref ref-type="bibr" rid="ref2">2</xref>
        ] to find all maximal cliques7. We then disregarded
all cliques of size two. We found in total 2190 cliques of size 3, 4, 5 and 6. The
total number of users in these cliques was 3275 which is around 20% of users
who reciprocated mentions.
      </p>
      <p>In order to take the time elapsed between consecutive messages into account,
we use the same threshold method explained in subsection 3.2, this time
demanding for an exchange to be a“conversation” that there is a contribution from all
parties and got relatively similar results (see Fig 4). The number of exchanges
which had contribution from all parties was at peak around 9 and 11 hours. We
took a threshold of 9 hours which gave us 334 multiuser conversations of sizes
3, 4, and 5 (see Fig 5a).</p>
      <p>Most users (out of 646) in our dataset were involved in just one multi-user
conversation, but a small number were involved in multiple conversations. The
users’ involvement in multi-user conversation is illustrated in Fig 5b.</p>
      <p>When examining the time-frame of multi-user exchanges, we found that the
correlation coecient between the total number of exchanges between clique
members and the average di↵erence between consecutive exchanges was 0.244
(see Fig 6a). This was not surprising, since we would expect lively conversations
(with lots of exchanged messages) to have a relatively fast pace, in contrast to
a casual exchange of messages with longer di↵erences inside our chosen 9 hour
time-window. The same picture is obtained from looking at the median time
di↵erences between consecutive messages across di↵erent clique sizes (see Fig
7 Maximal cliques are the largest complete subgraphs containing a given node.
6b). We also investigated how balanced multi-user exchanges were, although
this situation is more complicated than in the pairwise case.</p>
      <p>Firstly, we looked at the di↵erence between the number of tweets received and
sent by individual clique members. For each node, we computed the di↵erence
of their in-degree and out-degree. We summed up the positive values8 and to
normalise, we divided by the total number of exchanged messages. In this way,
we obtained a percentage of ‘unreciprocated’ messages, where reciprocity is not
8 Clearly the number of sent and received messages within a group are equal, thus
summing the di↵erences between in- and out-degree over individual members in the
group is by definition equal to zero.
toward a sender but toward a whole group. We show the histograms for the
di↵erent sizes of cliques in Fig. 7a. Across all clique sizes and in most of the
multi-user conversations around 30% messages were unreciprocated. In a small
number of conversations of 3 or 4 users a larger percentage were unreciprocated,
i.e. they were dominated by certain members, but also a large number of cliques
were very balanced (with unreciprocated messages at 0 10%), meaning every
individual received and sent a similar number of tweets.</p>
      <p>
        Finally, we looked at so-called ‘floor-gaining’ [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ], i.e. how much input each
user had over the course of a group exchange9. We compared the out-degree of
each user within a clique, (remember that each clique is a directed multigraph)
with the mean number of edges r = |nE|/|nV|, where nE is the total number of
edges within the clique and nV is the total number of vertices within the clique. In
a ‘round robin’ group conversation, with balanced turn taking, each user would
send out r messages, i.e. be responsible for an equal percentage p = 100r/e of
the total number nE of exchanged messages. For each clique size, we looked at
how many users’ representations were greater than or equal to p, i.e. those users
who ‘dominate’ the conversation. On Fig 7b below, we present the histogram for
a number of dominant users in the cliques of size 3, 4 and 5. This shows that in
most of the cliques of size 3 and 4, one user was responsible for the majority of
communication, whilst in cliques of size five, 2 users were dominant. However in
about 13% of all cliques of size 3 no users dominated, confirming that Twitter
is used for multi-user conversations and not just pairwise conversations.
9 We argue that the action of tweeting in multiuser exchanges can be regarded as
floorgaining, since tweets with mentions can in principal be read by a wider audience than
the group conversing.
      </p>
    </sec>
    <sec id="sec-4">
      <title>Conclusions</title>
      <p>We looked at conversations in Twitter, based on the underlying structure and
timings in approximately 4 million UK tweets with mentions over a period of
28 days. We structured the data as a multigraph to make use of graph
algorithms. We proposed a simple method of identifying conversations between pairs
of users, based on a time-threshold on the time-to-next tweet, and found
evidence that a threshold of 9hrs gives a good indication of distinct conversations.
We observed that the conversations detected using this method appeared to be
balanced, meaning that each party involved contributed approximately equally
to the conversation. This was not the case within more general interactions, in
which one agent typically contributed around three times as much as the other.</p>
      <p>Although finding cliques in graphs is computationally demanding, because
of the sparsity of interactions patterns within the data-set, extracting
multiuser exchanges was feasible and relatively fast. We were able to find all cliques
within the graph and, using the threshold method, identify conversations for
up to a maximum of 5 users. Most of those exchanges were fast-paced. We also
found that the number of messages in multi-user exchanges was reciprocal to the
average time di↵erence between them. When looking at the balance of multi-user
conversations, we found that most exchanges are dominated by just one or two
users, with some evidence of well-balanced group exchanges in between 3 users.
Regarding the number of received and sent messages by each individual in a
group, we found that some were dominated by one or two users, but also some
were well balanced.</p>
      <p>Further work needs to be done using content information to explore how
topics flow through multi-user exchange and if there is any relationship between
time-di↵erences between messages and topic. We hope that the insights gained
from our analysis could help to develop an understanding of the mechanisms and
dynamics of Twitter conversations, with potential scope for generating models
of micro-blogging behaviour.</p>
    </sec>
    <sec id="sec-5">
      <title>Acknowledgment</title>
      <p>This work is partially funded by the RCUK Digital Economy programme via
EPSRC grant EP/G065802/1 ‘The Horizon Hub’ and EPSRC MOLTEN EP/I016031/1.
We would like to thank Datasift for the provision of the data analysed, and to
Colin Singleton and Bruno Gon¸calves for very useful feedback and comments.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          1.
          <string-name>
            <given-names>L.</given-names>
            <surname>Backstrom</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Kleinberg</surname>
          </string-name>
          ,
          <string-name>
            <given-names>L.</given-names>
            <surname>Lee</surname>
          </string-name>
          , and
          <string-name>
            <given-names>C.</given-names>
            <surname>Danescu-Niculescu-Mizil</surname>
          </string-name>
          .
          <article-title>Characterizing and curating conversation threads: expansion, focus</article-title>
          , volume, re-entry.
          <source>In Proceedings of the sixth ACM international conference on Web search and data mining</source>
          ,
          <source>WSDM '13</source>
          , pages
          <fpage>13</fpage>
          -
          <lpage>22</lpage>
          , New York, NY, USA,
          <year>2013</year>
          . ACM.
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          2.
          <string-name>
            <given-names>C.</given-names>
            <surname>Bron</surname>
          </string-name>
          and
          <string-name>
            <given-names>J.</given-names>
            <surname>Kerbosch</surname>
          </string-name>
          . Algorithm 457:
          <article-title>finding all cliques of an undirected graph</article-title>
          .
          <source>Commun. ACM</source>
          ,
          <volume>16</volume>
          (
          <issue>9</issue>
          ):
          <fpage>575</fpage>
          -
          <lpage>577</lpage>
          , Sept.
          <year>1973</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          3.
          <string-name>
            <given-names>J.</given-names>
            <surname>Chen</surname>
          </string-name>
          ,
          <string-name>
            <given-names>R.</given-names>
            <surname>Nairn</surname>
          </string-name>
          , and
          <string-name>
            <given-names>E.</given-names>
            <surname>Chi</surname>
          </string-name>
          .
          <article-title>Speak little and well: recommending conversations in online social streams</article-title>
          .
          <source>In Proceedings of the SIGCHI Conference on Human Factors in Computing Systems, CHI '11</source>
          , pages
          <fpage>217</fpage>
          -
          <lpage>226</lpage>
          , New York, NY, USA,
          <year>2011</year>
          . ACM.
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          4.
          <string-name>
            <given-names>C.</given-names>
            <surname>Edelsky</surname>
          </string-name>
          .
          <article-title>Who's got the floor?</article-title>
          <source>Language in Society</source>
          ,
          <volume>10</volume>
          :
          <fpage>383</fpage>
          -
          <lpage>421</lpage>
          ,
          <year>1981</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          5.
          <string-name>
            <given-names>S.</given-names>
            <surname>Golder</surname>
          </string-name>
          and
          <string-name>
            <given-names>M. W.</given-names>
            <surname>Macy</surname>
          </string-name>
          .
          <article-title>Diurnal and seasonal mood tracks work sleep and daylength across diverse cultures</article-title>
          .
          <source>Science</source>
          ,
          <volume>333</volume>
          :
          <fpage>1878</fpage>
          -
          <lpage>1881</lpage>
          ,
          <year>2011</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          6.
          <string-name>
            <given-names>D.</given-names>
            <surname>Graham</surname>
          </string-name>
          .
          <article-title>Twitter: value-added conversation is more important than broadcasts, rts</article-title>
          . memeburn,
          <year>August 2012</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          7.
          <string-name>
            <given-names>A.</given-names>
            <surname>Hagberg</surname>
          </string-name>
          ,
          <string-name>
            <given-names>D.</given-names>
            <surname>Schult</surname>
          </string-name>
          , and
          <string-name>
            <given-names>P.</given-names>
            <surname>Swart</surname>
          </string-name>
          .
          <article-title>Exploring network structure, dynamics, and function using networkx</article-title>
          .
          <source>In Proceedings of the 7th Python in Science Conference (SciPy2008)</source>
          , Passadena, CA USA, pages
          <fpage>11</fpage>
          -
          <lpage>15</lpage>
          ,
          <year>August 2008</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          8.
          <string-name>
            <given-names>S.</given-names>
            <surname>Herring</surname>
          </string-name>
          .
          <article-title>Computer-mediated discourse</article-title>
          . In D. Tannen, D. Schir↵in, and H. Hamilton, editors,
          <source>Handbook of Discourse Analysis. Oxford: Blackwell</source>
          ,
          <year>2001</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          9.
          <string-name>
            <given-names>C.</given-names>
            <surname>Honeycutt</surname>
          </string-name>
          and
          <string-name>
            <given-names>S. C.</given-names>
            <surname>Herring</surname>
          </string-name>
          .
          <article-title>Beyond microblogging: Conversation and collaboration via twitter</article-title>
          .
          <source>In Proceedings of the Forty-Second Hawai'i International Conference on System Sciences, Los Alamitos</source>
          ,
          <year>2009</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          10. G. Lebo↵.
          <article-title>Marketing is a conversation? oh yeah</article-title>
          ...
          <source>about what? The Marketing Donut</source>
          ,
          <year>October 2011</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          11.
          <string-name>
            <given-names>E.</given-names>
            <surname>Nuutila</surname>
          </string-name>
          and
          <string-name>
            <given-names>E.</given-names>
            <surname>Soisalon-Soinen</surname>
          </string-name>
          .
          <article-title>On finding the strongly connected components in a directed graph</article-title>
          .
          <source>Information Processing Letters</source>
          ,
          <volume>49</volume>
          (
          <issue>1</issue>
          ):
          <fpage>9</fpage>
          -
          <lpage>14</lpage>
          ,
          <year>1994</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          12.
          <string-name>
            <given-names>D.</given-names>
            <surname>Quercia</surname>
          </string-name>
          ,
          <string-name>
            <given-names>L.</given-names>
            <surname>Capra</surname>
          </string-name>
          , and
          <string-name>
            <surname>J. Crowcroft.</surname>
          </string-name>
          <article-title>The social world of twitter: Topics, geography, and emotions</article-title>
          .
          <source>In The 6th international AAAI Conference on weblogs and social media</source>
          , Dublin,
          <year>2012</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          13.
          <string-name>
            <given-names>A.</given-names>
            <surname>Ritter</surname>
          </string-name>
          ,
          <string-name>
            <given-names>C.</given-names>
            <surname>Cherry</surname>
          </string-name>
          , and
          <string-name>
            <given-names>B.</given-names>
            <surname>Dolan</surname>
          </string-name>
          .
          <article-title>Unsupervised modeling of twitter conversations</article-title>
          .
          <source>In HLT-NAACL</source>
          <year>2010</year>
          ,
          <year>2010</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          14.
          <string-name>
            <given-names>D. A.</given-names>
            <surname>Shamma</surname>
          </string-name>
          ,
          <string-name>
            <given-names>L.</given-names>
            <surname>Kennedy</surname>
          </string-name>
          , and
          <string-name>
            <given-names>E. F.</given-names>
            <surname>Churchill</surname>
          </string-name>
          .
          <article-title>Peaks and persistence: modeling the shape of microblog conversations</article-title>
          .
          <source>In Proceedings of the ACM 2011 conference on Computer supported cooperative work</source>
          ,
          <source>CSCW '11</source>
          , pages
          <fpage>355</fpage>
          -
          <lpage>358</lpage>
          , New York, NY, USA,
          <year>2011</year>
          . ACM.
        </mixed-citation>
      </ref>
      <ref id="ref15">
        <mixed-citation>
          15.
          <string-name>
            <given-names>J.</given-names>
            <surname>Sidnell</surname>
          </string-name>
          .
          <source>Conversation Analysis: An Introduction. Blackwell</source>
          ,
          <year>2010</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref16">
        <mixed-citation>
          16.
          <string-name>
            <given-names>D.</given-names>
            <surname>Starkey</surname>
          </string-name>
          .
          <article-title>Some signals and rules for taking speaking turns in conversations</article-title>
          .
          <source>Journal of Personality and Social Psychology</source>
          ,
          <volume>23</volume>
          (
          <issue>2</issue>
          ):
          <fpage>283</fpage>
          -
          <lpage>292</lpage>
          ,
          <year>1972</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref17">
        <mixed-citation>
          17.
          <string-name>
            <given-names>R. E.</given-names>
            <surname>Tarjan</surname>
          </string-name>
          .
          <article-title>Depth-first search and linear graph algorithms</article-title>
          .
          <source>SIAM J. Comput.</source>
          ,
          <volume>1</volume>
          (
          <issue>2</issue>
          ):
          <fpage>146</fpage>
          -
          <lpage>160</lpage>
          ,
          <year>1972</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref18">
        <mixed-citation>
          18. K. Thomas,
          <string-name>
            <given-names>C.</given-names>
            <surname>Grier</surname>
          </string-name>
          ,
          <string-name>
            <given-names>V.</given-names>
            <surname>Paxson</surname>
          </string-name>
          , and
          <string-name>
            <given-names>D.</given-names>
            <surname>Song</surname>
          </string-name>
          .
          <article-title>Suspended accounts in retrospect: an analysis of twitter spam</article-title>
          .
          <source>In Proceedings of the 2011 ACM SIGCOMM conference on Internet measurement conference</source>
          , pages
          <fpage>243</fpage>
          -
          <lpage>258</lpage>
          . ACM,
          <year>2011</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref19">
        <mixed-citation>
          19.
          <string-name>
            <surname>C. Wang</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          <string-name>
            <surname>Ye</surname>
            , and
            <given-names>B. A.</given-names>
          </string-name>
          <string-name>
            <surname>Huberman</surname>
          </string-name>
          .
          <article-title>From user comments to on-line conversations</article-title>
          .
          <source>In Proceedings of the 18th ACM SIGKDD international conference on Knowledge discovery and data mining</source>
          ,
          <source>KDD '12</source>
          , pages
          <fpage>244</fpage>
          -
          <lpage>252</lpage>
          , New York, NY, USA,
          <year>2012</year>
          . ACM.
        </mixed-citation>
      </ref>
      <ref id="ref20">
        <mixed-citation>
          20.
          <string-name>
            <given-names>S.</given-names>
            <surname>Williams</surname>
          </string-name>
          ,
          <string-name>
            <given-names>M.</given-names>
            <surname>Terras</surname>
          </string-name>
          , and
          <string-name>
            <surname>Warwick</surname>
          </string-name>
          .
          <article-title>What people study when they study twitter: Classifying twitter related academic papers</article-title>
          .
          <source>Journal of Documentation</source>
          ,
          <volume>69</volume>
          ,
          <year>2013</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref21">
        <mixed-citation>
          21. S. Wu,
          <string-name>
            <given-names>J.</given-names>
            <surname>Hofman</surname>
          </string-name>
          ,
          <string-name>
            <given-names>W.</given-names>
            <surname>Mason</surname>
          </string-name>
          , and
          <string-name>
            <given-names>D.</given-names>
            <surname>Watts</surname>
          </string-name>
          .
          <article-title>Who says what to whom on twitter</article-title>
          .
          <source>In WWW</source>
          ,
          <year>2011</year>
          ,
          <year>2011</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref22">
        <mixed-citation>
          22. S. Wu,
          <string-name>
            <given-names>C.</given-names>
            <surname>Tan</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Kleinberg</surname>
          </string-name>
          , and
          <string-name>
            <given-names>M.</given-names>
            <surname>Macy</surname>
          </string-name>
          .
          <article-title>Does bad news go away faster?</article-title>
          <source>In 5th International AAAI Conference on Weblogs and Social Media</source>
          ,
          <year>2011</year>
          ,
          <year>2011</year>
          .
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>