<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Cooperative Authorship Social Network</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Giseli Rabello Lopes</string-name>
          <email>grlopes@inf.ufrgs.br</email>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Mirella M. Moro</string-name>
          <email>mirella@dcc.ufmg.br</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Leandro Krug Wives</string-name>
          <email>wives@inf.ufrgs.br</email>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Jose Palazzo Moreira de Oliveira</string-name>
          <email>palazzo@inf.ufrgs.br</email>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Universidade Federal de Minas Gerais - UFMG Belo Horizonte</institution>
          ,
          <country country="BR">Brazil</country>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>Universidade Federal do Rio Grande do Sul - UFRGS Porto Alegre</institution>
          ,
          <country country="BR">Brazil</country>
        </aff>
      </contrib-group>
      <abstract>
        <p>This paper introduces a set of challenges for developing a dissemination service over a Web collaborative network. We de ne speci c metrics for working on a co-authorship research social network. As a case study, we build such a network using those metrics and compare it to a manually built one. Speci cally, once we build a collaborative network and verify its quality, the overall e ectiveness of the dissemination services will also be improved.</p>
      </abstract>
      <kwd-group>
        <kwd>Social Networks</kwd>
        <kwd>Dissemination Systems</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>Introduction</title>
      <p>Web 2.0 is the second generation of communities and services characterized by
providing techniques for the personal publication, sharing, collaboration, and
organization of information on the World Wide Web. In this perspective, not only
the technological and content aspects but also the social interactions and its
relational aspects must be considered. In this context, the web-based communities,
hosted services, and web applications emerged, including Social Networks.</p>
      <p>The Social Network Analysis (SNA) is based on the assumption that the
relationship's importance between interaction units is a central point to the
evaluation and analysis of social interaction. Some fundamental concepts used
on SNA include actors and relational ties [1]. Actors are social entities that have
social linkages modeled by the Social Network (SN). Actors are linked to other
actors by relational ties.</p>
      <p>The increasing interest in researching in SNA was encouraged by the
popularization of online social networks, which are very interesting Web applications.
Another example of such concepts application is a co-authorship social network
representing a scienti c collaboration network. In this network, actors represent
authors and relational ties represent the relationships between pairs of authors.
The presence of at least one co-authored paper between two authors determines a
? This research is partially supported by CNPq (Brazil), and is part of the InWeb
research project.
relational tie between them. Some examples of data sources for the construction
of this kind of networks are DBLP, Google Scholar, CiteSeer, among others.</p>
      <p>The relational tie between authors may help to identify long term
collaborations, common research interests, preferred conferences, research groups under
formation, among others. Furthermore, as the social ties evolve, new research
interests and new collaborations will be identi ed. Any person who wants to
keep updated about such an evolution can be noti ed of such novel aspects by
adding a dissemination service to the social network.</p>
      <p>A dissemination service is formed by data producers and consumers.
Specifically, consumers subscribe to the service by de ning a pro le, which is usually
composed of di erent queries. As the producers inject the system with new
data, usually through messages, the dissemination service evaluates each
message against the pro les. Once there is a match between a pro le and a message,
the service sends that message to the pro le's consumer [2].</p>
      <p>The contributions of this paper are twofold. First, we introduce a set of
challenges for developing a dissemination service over a Web collaborative network.
Then, we tackle the challenges from the SN perspective. Speci cally, we present
an architecture for such a dissemination service over a collaborative network.
The architecture is formed by di erent layers, from the Web to digital libraries,
social network, and the dissemination service. Based on the architecture, we were
able to identify research challenges that are innovative to the SN area. We de ne
speci c metrics for working on a co-authorship research SN. Then, we build a
network using those metrics and compare it to a manually built one. Speci cally,
once we build a collaborative network and verify its quality, the e ectiveness of
the dissemination services will also be improved. Therefore, based on such an
evaluation, the dissemination service can identify (and recommend) the more
pertinent publications as well as identify possible hidden collaboration nets.</p>
      <p>The paper is organized as follows. Section 2 describes the general context
of dissemination services and de nes the base architecture. Section 3 introduces
the metrics to determine the weights of relational ties of a co-authorship Social
Network. Section 4 presents a case study that shows the construction of
collaboration Social Network. It also evaluates the metrics employed to analyze the
SN. Section 5 presents some related work. Section 6 concludes this paper.
2</p>
    </sec>
    <sec id="sec-2">
      <title>Dissemination Service in Social Network Context</title>
      <p>Content-based dissemination is a form of data delivery that di ers from
traditional communications since the messages are delivered according to their
content rather than the IP address of their destination. There is a
continuous stream of messages from data producers to consumers, without any of the
human parties having knowledge of the other [2, 3]. This form of communication
is widely employed by dissemination services, which may be employed within
publish/subscribe systems (pub/sub for short).</p>
      <p>
        In order to clarify how a dissemination service can work on a Web
collaborative network, we present a case study based on the academic eld. It exempli es
a service that disseminates new publications and research connections. Speci
cally, individuals (or organizations) can subscribe to research topics or researcher
names, for example. Once a new publication or a new collaboration is detected,
this information is disseminated to those subscribers whose keywords match such
new data. It is important to notice that not only publications are recommended
but also (and more important) new possible cooperations among researchers
are identi ed and suggested. The whole process is composed by six phases, as
illustrated in Figure 1. Each step of this process works as follows.
(
        <xref ref-type="bibr" rid="ref1">1</xref>
        ) The information about researchers is
mined from the Web or provided by
individuals or organizations. Their actual
publications or their curricula vitae are
organized in semi-structured data. (
        <xref ref-type="bibr" rid="ref2">2</xref>
        ) A
Digital Library (DL) stores and allows to
manage such data. (
        <xref ref-type="bibr" rid="ref3">3</xref>
        ) A DL interactive
process feeds relevant information to build a
social-research network. (
        <xref ref-type="bibr" rid="ref4">4</xref>
        ) The
dissemination service evaluates this huge volume
of connected data and identi es the
resulting, ltered, quali ed data. (
        <xref ref-type="bibr" rid="ref5">5</xref>
        ) This
resulting information is delivered to the
individuals (researchers, students,
professionals) and organizations (educational,
governmental, and industrial), and (
        <xref ref-type="bibr" rid="ref6">6</xref>
        ) published
Fig. 1. Dissemination service over Web col-back to the Web, providing universal access
laborative network and visibility to the research network data.
      </p>
      <p>
        The dissemination service from Figure 1 illustrates tasks with challenges to
di erent Computer Science areas. Speci cally, Information Retrieval techniques
may be employed along with Data Mining algorithms in order to recover the
researchers' data from the Web (
        <xref ref-type="bibr" rid="ref1">1</xref>
        ). Moreover, Web Management issues become
critical when considering that the data will be extracted from the Web (for
example privacy, security, provenance, and credibility). The Digital Library
maintenance presents new challenges due to the interactive nature of the framework
(
        <xref ref-type="bibr" rid="ref2">2</xref>
        ), where individuals and organization will access the data through the
dissemination service, and not through the Digital Library interface as usual. Social
Network's mechanisms are necessary for de ning the collaborative network (
        <xref ref-type="bibr" rid="ref3">3</xref>
        ).
Then, the challenges appear on the Dissemination service level, which also
include Network Management (
        <xref ref-type="bibr" rid="ref4">4</xref>
        ). Finally, the actual dissemination and evaluation
of data involve Document Management, Distributed Systems, Parallel
Computing, Security and Networks as well (
        <xref ref-type="bibr" rid="ref5 ref6">5, 6</xref>
        ).
      </p>
      <p>It is important to notice that each of those disciplines is complex by
nature. Instead of discussing each of such areas, the focus of this paper is on the
social networks challenges. Speci cally, with the increasing interest in Social
Networks, the interaction of the parties (data producers and consumers) within
the dissemination service will soon conquer the spotlight. In social networks, it
is important to qualify and quantify how individuals (people and organizations)
are connected, how tightly (or loosely) they interact, and what their common
interests are. Due to the large volume of data involved and the high complexity
of those connections, the development of an automatic mechanism capable of
e ciently identifying and analyzing such interactions is imperative.
3</p>
    </sec>
    <sec id="sec-3">
      <title>Social Scienti c Networks Analysis</title>
      <p>Social Networks are based on the assumption of the relationship's importance
between interaction units. The weights of the relational ties in a social network
aim to measure the importance of the ties between actors. It is necessary to
establish approaches to automatically determine these weights based on
information available about the actor's relationships.</p>
      <p>In this paper, we employ a scienti c collaboration network as base example.
We present approaches to determine two types of associations namely
Collaboration in Co-authorship and Collaboration in Research Areas. These associations
were chosen because they cover certain facets of the relational ties of the
collaboration network. According to Newman [4], that studied scienti c collaboration
networks in which two scientists are considered connected if they have
coauthored a paper, this seems a reasonable de nition of scienti c acquaintance.
3.1</p>
      <p>Collaboration-based association - Co-authorship (Ca)
Formally, a Social Network SN of a co-author relationship a is a pair: SNa =
(N; E) where N and E are the set of N odes and Edges. Each edge e 2 E is a
tuple of the form hai; t; w; aj i, where the edge is directed from ai to aj , t denotes
the type of association between ai and aj , and w denotes the weight a ected
to the association. This weight is a numerical value between 0 and 1. In our
approach, the equation 1 determines the Collaboration in Co-authorship weight.
wtCa(ai!aj) = jaj co authorshipj
jaiauthorj
{ wtCa(ai!aj) corresponds to the weight of the recommendation based on the
co-author relationship. The weight is di erent according to the relation
direction (the weight in the direction ai ! aj is di erent than in aj ! ai);
{ jaj co authorshipj corresponds to the number of times that the author aj
was a co-author of a paper with author ai;
{ jaiauthorj corresponds to the total number of papers of the author ai.</p>
      <p>
        In other words, the higher this weight is, the more relevant is the relationship
with author aj to the author ai. The use of Ca metric implies that there is a
graph with 0 or 2 links between two authors. The weights represent the degree of
collaboration in co-authorship between the authors. This metric is an asymmetric
variant of the Jaccard Coe cient and it was applied in the context of Social
Networks by other works as [5, 6].
(
        <xref ref-type="bibr" rid="ref1">1</xref>
        )
      </p>
      <p>Collaboration-based association - Research Areas (Ra)
In this case, we consider the same de nition of Social Network SN of co-author
relationship (as de ned in the previous section). However, each edge e 2 E is
a tuple of the form hai; t; r; w; aj i, where the edge is directed from ai to aj , t
denotes the type of association between ai and aj , r denotes the research area
associated to the relationship represented, and w denotes the weight a ected to
the association. This weight is a numerical value between 0 and 1. The equation
2 provides the Collaboration in Research Areas weight.</p>
      <p>wtRa(ai!aj) =</p>
      <p>
        Crresearch areas(ai;aj)
jresearch areasai j
co authorshipresearch area rx(ai;aj)
co authorshipresearch areas(ai;aj)
(
        <xref ref-type="bibr" rid="ref2">2</xref>
        )
{ wtRa(ai!aj) corresponds to the weight of the recommendation based on the
coauthor relationship according to research areas. Again, the weight is di erent
according to the relation direction;
{ Crresearch areas(ai;aj) corresponds to the number of research areas in which
the authors ai and aj published co-authored papers;
{ jresearch areasai j corresponds to the total number of research areas in
which author ai published;
{ co authorshipresearch area rx(ai;aj) corresponds to the number of co-author
relationship between authors ai and aj in the x area;
{ co authorshipresearch areas(ai;aj) is the total number of co-author
relationship between authors ai and aj in every research areas in which they
published together.
      </p>
      <p>The use of Ra metric implies that there are 2n links between two authors,
being that n indicates the number of research areas in which the authors published
together. Each link has a direction, a research area and a weight associated. The
higher this weight is, the more relevant is the relationship with author aj to
the author ai in the research area x. In such an approach, we have the idea of
collaboration in research areas.
4</p>
    </sec>
    <sec id="sec-4">
      <title>Case Study</title>
      <p>This paper proposes an approach to construct a social network for collaborative
research. The complete work is under development as research project of the
InWeb (MCT/CNPq Grant Number 573871/2008-6), the Brazilian National
Institute of Science and Technology for the Web. In fact, we have built a collaborative
social network based on the publications of the researchers associated to INWeb.
The Institute is formed by 27 researchers and their students. All researchers are
professors in a major education institution (namely UFMG, UFRGS, UFAM,
and CEFET-MG) with graduate program in Computer Science.</p>
      <sec id="sec-4-1">
        <title>UFMG</title>
        <p>Berthier A.
Ribeiro-Neto
Mirella
M. Moro
Clodoveu
A. Davis</p>
        <p>Alberto
Laender</p>
        <p>Marcos
A. Gonçalves
Raquel
O. Prates</p>
        <p>Jussara
M. Almeida
Arnaldo
A. Araújo
Gisele
L. Pappa
Renato
Ferreira</p>
      </sec>
      <sec id="sec-4-2">
        <title>UFAM</title>
        <p>Altigran S. da Silva
Edleno S. de Moura
João M. B.</p>
        <p>Cavalcanti</p>
      </sec>
      <sec id="sec-4-3">
        <title>UFRGS</title>
        <p>Viviane M. Orengo</p>
        <p>Carlos A. Heuser</p>
        <p>Renata M. Galante
José Palazzo M. de Oliveira
Nivio</p>
        <p>Ziviani
Virgílio A.</p>
        <p>F. Almeida
Leandro K. Wives</p>
      </sec>
      <sec id="sec-4-4">
        <title>CEFET/MG</title>
        <p>Evandrino G. Barros</p>
        <p>Fabiano Botelho
Cristina Murta</p>
        <p>Building the Social Network: Manually and Automatically
Initially, this group of researchers was manually analyzed by a specialist. The
resulting network can be visualized in Figure 2. This network is used as baseline.</p>
        <p>Rede Co-Autoria: UFMG + UFAM, UFRGS, CEFET
Genaína Nunes Rodrigues</p>
        <p>DGo.rNgeivtaol MWeaigraneJrr.</p>
        <p>Adriano</p>
        <p>C. M. Pereira</p>
        <p>For validating our metrics, we have implemented a tool to automatically
generate a Social Network. This SN was build using information about authors
provided by the DBLP digital library. It is important to notice that this library
is exported as an XML document. Instead of using the whole dataset, we
extracted from the library just the papers written by the considered researchers and
published in conferences proceedings and in journals (as elements inproceedings
and article). Such a subset was chosen because this information is signi cantly
important for representing the co-author relationship between authors and,
consequently, to determine the research collaborations among them.</p>
        <p>The actors of the SN can be chosen and they are a subset of authors with
scienti c papers indexed by the DBLP. The relational ties between actors are the
relationships between pairs of authors. These social ties represent the co-author
relationships. The weights of the linkages are determined by equation 1. In that
equation, jaiauthorj corresponds to the total number of papers of the author ai,
and it considers all papers to this author ai indexed at DBLP, including papers
that are not co-authored by authors in the SN who will be graphically presented.</p>
        <p>The resultant INWeb Social Network constructed automatically is presented
in Figure 3. The data used in this case was collected from the DBLP repository
on January 21, 2009. This data gathering process summed up 677,345 authors;
692,431 conference proceedings papers and 432,663 journal articles.</p>
        <p>After building them, we compared the two Social Networks: the manually
constructed SN (called Manual INWeb SN) and the automatically generated one
Legend:
1-Adriano M. Pereira
2-Alberto H. F. Laender
3-Altigran Soares da Silva
4-Arnaldo de Albuquerque Araujo
5-Berthier A. Ribeiro-Neto
6-Carlos A. Heuser
7-Clodoveu A. Davis
8-Cristina D. Murta
9-Dorgival Olavo Guedes Neto
10-Edleno Silva de Moura
11-Evandrino G. Barros
12-Fabiano C. Botelho
13-Gena na Nunes Rodrigues
14-Gisele L. Pappa
15-Joa~o M. B. Cavalcanti
1O6l-ivJeoisrea Palazzo Moreira de
17-Jussara M. Almeida
18-Leandro Krug Wives
19-Marcos Andre Gonalves
20-Mirella Moura Moro
21-Nivio Ziviani
22-Raquel Oliveira Prates
23-Renata de Matos Galante
24-Renato Ferreira
25-Virg lio A. F. Almeida
26-Viviane Moreira Orengo
27-Wagner Meira Jr.
(called Automatic INWeb SN). Comparing them against each other, we observed
that the Manual INWeb SN covers 93.44% of the Automatic INWeb SN. The
Automatic INWeb SN covers 83.82% of the Manual INWeb SN. Furthermore, if
we consider that the ideal network (144 edges) is the union between the edges of
the Manual INWeb SN (136 edges, considering that each linkage was reciprocal)
and the edges of the Automatic INWeb SN (122 edges), we have the following
results. The Manual INWeb SN recall is 94.44% and the Automatic INWeb SN
recall is 84.72%. The ideal network was considered the union because the Manual
INWeb SN was carefully developed by a specialist and the Automatic INWeb
SN was based on an occurrence of a co-authorship between two authors for the
establishment of the relational ties.</p>
        <p>The main goal of this comparative analysis between the two networks was to
validate the Social Network constructed automatically by our system using the
DBLP dataset. The results obtained demonstrate that the DBLP digital library
is a good data source that considerably covers the co-authorship relations in
Computer Science, more speci cally in Information Systems research area.
4.2</p>
        <p>Analysis of the Automatic Co-authorship Network
In this section, we further analyze the Automatic INWeb Social Network. The
goal is to use other metrics to understand the properties of the Social Network on
this case study. In the next subsections, we present the metrics considered and
discuss the results obtained (observation: the results of the metrics were plotted
in decreasing order of the values obtained in all graphics and the authors were
represented by numbers in the range of 1 to 27 into accordance to the ascending
order of the full names (see Legend of Figure 3)).</p>
        <p>Clustering Metrics. Clustering is a process that aims to identify subsets or
clusters of \similar" elements (or data items). The goal of clustering algorithms
is to create groups that are coherent internally, but clearly di erent from each
other. Thus, elements within a cluster should be as similar as possible; and
elements in one cluster should be as dissimilar as possible from elements in other
clusters [7]. In order to evaluate the clusters generated by those algorithms,
we can employ internal quality measures that require no human intervention,
such as cohesion and coupling [8]. Cohesion is the average pairwise similarity
of elements within the cluster. Coupling is the average pairwise similarity of
elements in which one element belongs to cluster C and the other does not.</p>
        <p>The clustering metrics were adapted for evaluating our case study. We
considered each group constituted by an author and all his co-authors as a cluster.
For each cluster (each author), we calculated the respective cluster metrics. The
similarity values for the metrics calculation are the weights of the relational ties
between authors. In our case, the best results will be that whose cohesion and
coupling measure high values. Such result is important because each cluster is a
subnet of the social network being analyzed.</p>
        <p>The cohesion metric was adapted to consider two similarity values between
each pair of authors. This was necessary because our SN is represented by a
directional graph. The new equation is de ned as follows (Equation 3).
m 1 m 1</p>
        <p>
          X X
cohesion(C) = i=1 j=i+1
wt(ai!aj) + wt(aj!ai)
m(m
1)
(
          <xref ref-type="bibr" rid="ref3">3</xref>
          )
where, m corresponds to the total number of authors in the group considered
(m=1(author)+n(total number of his/her INWeb co-authors)).
        </p>
        <p>In this case, the similarity values used (wt) in the calculation were the weights
wtCa . Figure 4 presents the cohesion results obtained to each cluster formed by
one author and all his INWeb co-authors. The results obtained show the average
of importance between all pairs of authors in each cluster considered. The more
cohesive groups are those formed by authors with high number of collaborations
whose weights indicate a high importance in these co-authorships.</p>
        <p>As Figure 4 illustrates, some clusters formed by few authors have the best
results. This probably happened because these clusters are formed by young
authors whose importance weights in relation to their co-authors are high. Some
senior authors formed clusters with low cohesion values. This probably happened
because those worked with many co-authors over time and/or have a much larger
collaboration (cooperation) network that the one formed by INWeb authors.</p>
        <p>Figure 5 presents the results for coupling metric. This graphic plots the
authors in x axis and the coupling values obtained for each cluster (formed by
the author and his co-authors) in y axis. Equation 4 was used. This metric was
evaluated by using the output weights to the author ai whose cluster C is being
analyzed as similarity value. Indeed, C is the cluster formed by an author and his
co-authors; m is the number of elements in the cluster C; and n is the number of
elements outside the cluster C belonging to a cluster Q formed by the co-authors
of ai and all co-authors of these co-authors of ai (including ai). In this case, the
similarity values used in the calculation are the weights wtCa(ai!aj) where ai was
Authors
Authors
the author been analyzed and aj varies among each author of the cluster Q.
coupling(C) = i;j</p>
        <p>X sim(ci; qj)
m
n
Note that the nonzero similarity values are between ai and his co-authors, and
between ai and ai himself. On the equation, the weight between the author and
himself was considered 1. This shows the coupling among the group of researchers
formed by each author and his co-authors. The results show that some young
researchers that have \good" publications present high coupling. This probably
occurred because such researchers work in more \condensed" groups while the
others have a larger network and/or work in several groups.</p>
        <p>Complementary Analysis. This subsection presents other analysis
performed on the Automatic INWeb Social Network.</p>
        <p>First, Figure 6 presents the percentage of INWeb Co-authors in relation of
the total Co-authors indexed by DBLP, for each author. This metric prioritizes
authors that have high number of his total co-authors within the INWeb Social
Network. The results show higher values to the authors that have his co-author
relationships represented more signi cantly by the INWeb partnerships.
0,600
0,500
n0,400
iso0,300
e
ho0,200
C0,100
0,000
35,00%
rs30,00%
o
th25,00%
u
-a20,00%
o
fC15,00%
eo10,00%
g
ta 5,00%
n
rce 0,00%
e
P</p>
        <p>Authors</p>
        <p>Authors
n</p>
        <p>
          X wtaj!ai
In Avg Imp(ai) = j=1
(
          <xref ref-type="bibr" rid="ref5">5</xref>
          )
        </p>
        <p>Out Avg Imp(ai) = j=1
n n
where ai corresponds to the author being analysed, aj varies among the
coauthors of ai, and n corresponds to the total number of co-authors of ai in the
Social Network being considered.</p>
        <p>
          The graph in Figure 8 plots the authors in x axis and the input average
importance values obtained for each author in y axis. For calculating the
importance (wtCa(ai!aj) ), it considered the DBLP Social Network (i.e., all publications
indexed by DBLP were considered, whether they are co-authored by an INWeb
author or not). However, the co-authors considered were only those belonging
to the INWeb Network. Figure 8 also illustrates the relative importance of each
author to the others. The result shows that the equation prioritizes authors
who have a high average importance value to his co-authors. Some authors that
have few co-authors but have a meaningful importance value to his co-authors
overcame other authors that have a high number of co-authors.
n
X wtai!aj
(
          <xref ref-type="bibr" rid="ref6">6</xref>
          )
0,350
Itrcgaeepnom I-trcTChoooN )trsho0000,,,,122350500000
rae tu au0,100
ItvnpuA (fraom 00,,000500
1,000
caen tso 0,800
tr r
o o
m au )r 0,600
p th
Iree -TCo auho00,,240000
ga c t
tvA IN
tpu rom 0,000
u (f
        </p>
        <p>O
Authors</p>
        <p>Authors
This section overviews some work related to recommender systems (a type of
dissemination system) and social networks.</p>
        <p>Weng and Chang [9] propose a recommender method that employs ontologies
and the spreading activation model The ontologies are employed for de ning
user pro les, being the basis to reason about the users' interests. The spreading
activation model is used to search for other in uential users in a Social Network</p>
        <p>Golbeck et al. [10] present a website that integrates Social Networks on the
Semantic Web context and the trust concept for the generation of movies'
recommendations. The Social Networks then indicate the trust ratings between users
by considering the path length between them.</p>
        <p>Aleman-Meza et al. [5] de ne a solution for the Con ict of Interest (COI)
problem using Social Networks. The goal is to detect COI relationships among
authors of scienti c papers and potential reviewers of these papers. Moreover,
rules are established to determine a possible degree of COI among the authors
based on the Social Networks built and the relationship's weights between them.</p>
        <p>Jeh et al. [11] propose a measure of structural-context similarity, called
SimRank. The recommender systems were used as motivation. The base idea of the
model is that two objects are similar if they are related to similar objects.</p>
        <p>Zaiane et al. [12] explore a Social Network coded within the DBLP database.
It considers a new random walk approach to reveal interesting knowledge about
the research community and even to recommend collaborations.</p>
        <p>Menezes et al. [13] developed a geographical analysis of knowledge
production in Computer Science. They analyzed co-authorship Social Networks of the
Computer Science area.</p>
        <p>Ganev et al. [14] developed a set of tools for building, exploring and querying
academic Social Networks. They proposed a measure reputation called visibility
as an adjusted PageRank applied on the Social Network context.</p>
        <p>Our paper is related to all those since it focuses on solutions for Social
Networks. However, we presented a case study to clarify how a dissemination service
can work on top of a Web collaborative network. We presented an approach to
construct a Social Network for collaborative research that considers new
metrics. Our paper also adapts evaluation metrics to analyze the quality of the social
network obtained using the proposed approach.
6</p>
      </sec>
    </sec>
    <sec id="sec-5">
      <title>Concluding Remarks</title>
      <p>The section 4 analyzed the Automatic INWeb Social Network. In the future, we
plan to analyze the evolution of these results. We will also be able to compare
them against new analysis from other Social Networks. Regarding the
dissemination service, these results will also be useful. Speci cally, once we build a
collaborative network and verify its quality (using the aforementioned metrics),
the quality of the dissemination services will also be improved. In other words,
the evaluation of the relational ties among the researchers (authors) ensures
better quality to the dissemination service. Therefore, based on such an
evaluation, the dissemination service can identify (and recommend) the more pertinent
publications as well as identify possible hidden collaboration nets.</p>
      <p>As dissemination systems have recently grown from topic-based systems to
XML-enabled systems, we believe that the next step is for them to follow the data
technology and support any type of data uniformly (e.g. relational and XML).
Moreover, considering all the aspects involved from the other research areas,
we believe that the database technology must evolve to consider uniformly and
seamlessly any type of data there exist with extensible and Web-scalable features.
This complex scenario brings new and exciting issues to be handled by many
di erent Computer Science communities. Our nal goal is to have a working
system that integrates our research groups. The results will be evaluated, at the
end of a four year period, by the access patterns and users evaluation of the
quality of the disseminated papers and, more important, by the increase in the
cooperation pattern among inter-institutional researchers. From the social point
of view, those features are the fundamental element to the integration to the
access of the content available at INWeb.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          1.
          <string-name>
            <surname>Wasserman</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Faust</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          :
          <article-title>Social Network Analysis: methods and applications</article-title>
          . Cambridge University Press (
          <year>1994</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          2.
          <string-name>
            <surname>Diao</surname>
            ,
            <given-names>Y.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Rizvi</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Franklin</surname>
            ,
            <given-names>M.J.</given-names>
          </string-name>
          :
          <article-title>Towards an internet-scale xml dissemination service</article-title>
          .
          <source>In: VLDB</source>
          . (
          <year>2004</year>
          )
          <volume>612</volume>
          {
          <fpage>623</fpage>
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          3.
          <string-name>
            <surname>Moro</surname>
            ,
            <given-names>M.M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Vagena</surname>
            ,
            <given-names>Z.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Tsotras</surname>
          </string-name>
          , V.J.:
          <article-title>Recent Advances and Challenges in XML Document Routing</article-title>
          . In: Open and
          <article-title>Novel Issues in XML Database Applications: Future Directions and Advanced Technologies</article-title>
          .
          <source>IGI Global</source>
          (
          <year>2009</year>
          )
          <volume>136</volume>
          {
          <fpage>150</fpage>
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          4.
          <string-name>
            <surname>Newman</surname>
            ,
            <given-names>M.E.J.:</given-names>
          </string-name>
          <article-title>The structure and function of complex networks</article-title>
          .
          <source>SIAM Review</source>
          <volume>45</volume>
          (
          <year>2003</year>
          )
          <volume>167</volume>
          {
          <fpage>256</fpage>
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          5.
          <string-name>
            <surname>Aleman-Meza</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Nagarajan</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ding</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Sheth</surname>
            ,
            <given-names>A.P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Arpinar</surname>
            ,
            <given-names>I.B.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Joshi</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Finin</surname>
            ,
            <given-names>T.W.</given-names>
          </string-name>
          :
          <article-title>Scalable semantic analytics on social networks for addressing the problem of con ict of interest detection</article-title>
          .
          <source>TWEB</source>
          <volume>2</volume>
          (
          <issue>1</issue>
          ) (
          <year>2008</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          6.
          <string-name>
            <surname>Mika</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          :
          <article-title>Social networks and the semantic web</article-title>
          .
          <source>In: WI '04</source>
          , Washington, DC, USA, IEEE Computer Society (
          <year>2004</year>
          )
          <volume>285</volume>
          {
          <fpage>291</fpage>
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          7.
          <string-name>
            <surname>Manning</surname>
            ,
            <given-names>C.D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Raghavan</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          , Schutze, H.: Introduction to Information Retrieval. Cambridge University Press (
          <year>July 2008</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          8.
          <string-name>
            <surname>Kunz</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Black</surname>
            ,
            <given-names>J.P.</given-names>
          </string-name>
          :
          <article-title>Using automatic process clustering for design recovery and distributed debugging</article-title>
          .
          <source>IEEE Trans. Softw. Eng</source>
          .
          <volume>21</volume>
          (
          <issue>6</issue>
          ) (
          <year>1995</year>
          )
          <volume>515</volume>
          {
          <fpage>527</fpage>
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          9.
          <string-name>
            <surname>Weng</surname>
            ,
            <given-names>S.S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Chang</surname>
            ,
            <given-names>H.L.</given-names>
          </string-name>
          :
          <article-title>Using ontology network analysis for research document recommendation</article-title>
          .
          <source>Expert Syst. Appl</source>
          .
          <volume>34</volume>
          (
          <issue>3</issue>
          ) (
          <year>2008</year>
          )
          <year>1857</year>
          {
          <fpage>1869</fpage>
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          10.
          <string-name>
            <surname>Golbeck</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Hendler</surname>
          </string-name>
          , J.:
          <article-title>Filmtrust: movie recommendations using trust in webbased social networks</article-title>
          .
          <source>In: IEEE CCNC - Consumer Communications and Networking Conference</source>
          . Volume
          <volume>1</volume>
          . (
          <year>2006</year>
          )
          <volume>282</volume>
          {
          <fpage>286</fpage>
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          11.
          <string-name>
            <surname>Jeh</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Widom</surname>
          </string-name>
          , J.:
          <article-title>Simrank: a measure of structural-context similarity</article-title>
          .
          <source>In: ACM SIGKDD</source>
          . (
          <year>2002</year>
          )
          <volume>538</volume>
          {
          <fpage>543</fpage>
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          12.
          <string-name>
            <surname>Zaiane</surname>
            ,
            <given-names>O.R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Chen</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Goebel</surname>
          </string-name>
          , R.:
          <article-title>Dbconnect: mining research community on dblp data</article-title>
          . In: WebKDD/SNA - Workshop on Web Mining and
          <string-name>
            <surname>Social Network Analysis.</surname>
          </string-name>
          (
          <year>2007</year>
          )
          <volume>74</volume>
          {
          <fpage>81</fpage>
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          13.
          <string-name>
            <surname>Menezes</surname>
            ,
            <given-names>G.V.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ziviani</surname>
            ,
            <given-names>N.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Laender</surname>
            ,
            <given-names>A.H.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Almeida</surname>
          </string-name>
          , V.:
          <article-title>A geographical analysis of knowledge production in computer science</article-title>
          . In: WWW. (
          <year>2009</year>
          )
          <volume>1041</volume>
          {
          <fpage>1050</fpage>
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          14.
          <string-name>
            <surname>Ganev</surname>
            ,
            <given-names>V.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Guo</surname>
            ,
            <given-names>Z.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Serrano</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Tansey</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Barbosa</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Stroulia</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          :
          <article-title>An environment for building, exploring and querying academic social networks</article-title>
          .
          <source>In: MEDES '09</source>
          , New York, NY, USA, ACM (
          <year>2009</year>
          )
          <volume>282</volume>
          {
          <fpage>289</fpage>
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>