<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Dealing with Goal Models Complexity using Topological Metrics and Algorithms</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Lucía Méndez Tapia</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Lidia López</string-name>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Claudia P. Ayala</string-name>
          <email>cayala@essi.upc.edu</email>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Universidad del Azuay (UDA) Av. 24 de Mayo 7-77 y Francisco Moscoso</institution>
          ,
          <addr-line>Cuenca</addr-line>
          ,
          <country country="EC">Ecuador</country>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>Universitat Politècnica de Catalunya (UPC) c/Jordi Girona</institution>
          ,
          <addr-line>1-3, E-08034 Barcelona</addr-line>
          ,
          <country country="ES">Spain</country>
        </aff>
      </contrib-group>
      <abstract>
        <p>The inherent complexity of business goal-models is a challenge for organizations that has to analyze and maintaining them. Several approaches are developed to reduce the complexity into manageable limits, either by providing support to the modularization or designing metrics to monitor the complexity levels. These approaches are designed to identify an unusual complexity comparing it among models. In the present work, we expose two approaches based on structural characteristics of goal-model, which do not require these comparisons. The first one ranks the importance of goals to identify a manageable set of them that can be considered as a priority; the second one modularizes the model to reduce the effort to understand, analyze and maintain the model.</p>
      </abstract>
      <kwd-group>
        <kwd>i* Framework</kwd>
        <kwd>iStar</kwd>
        <kwd>Complexity</kwd>
        <kwd>Metrics</kwd>
        <kwd>PageRank</kwd>
        <kwd>Clustering</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>-</title>
      <p>
        The envisioned state that all organizations desire to achieve, is represented by a set of
strategic goals, which in turn are related to each other through semantic links that
denote the participation that a specific goal as a support of others. The particular goal
arrangement and goal relationships of an organization, constitute its business goal
model. It is well-known the extensiveness and complexity inherent to business
goalmodels [
        <xref ref-type="bibr" rid="ref1">1</xref>
        ]. And according to [
        <xref ref-type="bibr" rid="ref2">2</xref>
        ] its complexity can be seen from a general point of
view as “the difficultly of handling a system, as it is hard to estimate the outcome of an
action”, that involves specific properties [
        <xref ref-type="bibr" rid="ref3">3</xref>
        ] like understandability (it is difficult to
understand and verify) and high interaction among its components. Hence, it is crucial
to managing the complexity in an effective way [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ].
      </p>
      <p>
        Several approaches has been developed to address the complexity problem, among
them we refer to [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ], where the authors defines the set of metrics to evaluate the
accidental complexity (originated by the modeling way) of KAOS goal models while
building those models; the StarGro approach [
        <xref ref-type="bibr" rid="ref5">5</xref>
        ] that contains three requirements
management metrics which also be applied to goal-model complexity. In [
        <xref ref-type="bibr" rid="ref6">6</xref>
        ], the
authors propose a metrics suite to take advantage of the modularity given by the actor's
boundaries in i* models. The metrics of all of these approaches generate a set of values
that must be compared with datasets of other models, in order to identify if they are an
‘unusual behaviors’ or if they are ‘normal’. On the other hand, the work presented in
[
        <xref ref-type="bibr" rid="ref7">7</xref>
        ] proposes different types of modules associated with a specific semantic (Data
Warehouse domain), and the work of [
        <xref ref-type="bibr" rid="ref8">8</xref>
        ] shows 3 types of Strategic Rationale modules
(task-decomposition, means-end, and contribution) as a composition of elements.
      </p>
      <p>To reduce the complexity, we propose two approaches from the Graph Theory
perspective, which are based on topological characteristics of the model, and unlike the
aforementioned, they do not require to be compared with any dataset and are not
associated with a specific semantic or based on goals relationship. We apply these
approaches to an organization’s goal model created to support the analysis of OSS
adoption implications. The complexity hinders the analysis and management of the
model. Our first proposed approach is the Ranking, which identifies a manageable set
of goals that are relevant for a specific analysis; this approach allows us to focus the
effort on goals that can be considered as high priority. Our second proposed approach
is the Clustering, which seeks to decrease the complexity creating modules of goals
(clusters) that can integrate a hierarchy with different levels of abstraction; this
hierarchy facilitates the analysis and maintenance tasks because the effort is centered
in one subset of goals at a time.</p>
      <p>The rest of the paper is structured as follows: Section 2 introduces the characteristics
of our goal model; Section 3 presents the ranking approach; Section 4 presents the
clustering approach; finally Section 5 shows the conclusions.
2</p>
    </sec>
    <sec id="sec-2">
      <title>The goal model</title>
      <p>
        With the business goals catalogs presented in our previous work [
        <xref ref-type="bibr" rid="ref9">9</xref>
        ], we built a Strategic
Rationale diagram that represents the goal model of a software-intensive organization
(who develops software and/or offers services related to software), that incorporates
Open Source Software (OSS) as part of its customer offer. These business goals have
been extended including the strategic goals related to the OSS Integration adoption
strategy defined by [10], characterized by the active participation of the organization
in an OSS community in order to share and co-create OSS. The complexity of our
diagram can be appreciated in Fig. 1, it is hard to visualize, manage and maintain a
model with 80 goals and more than 120 links.
      </p>
      <p>In the context of our research, we need to analyze the importance of the goals from
the organization’s point of view. The resulting model contains a unique root element
representing the organization’s vision (1BG01 Vision, the main business goal to reach)
located at the upper level; from this root are disaggregated all other goals. Our example
only includes those business goals that are involved in OSS adoption.
As aforementioned, the large number of goals and its relationships increases the
complexity of the model and, therefore, the effort and resources required for its
analysis. For this reason, a selective analysis is more efficiently that an exhaustive one,
because the first one allows focusing on a manageable set of highly impacted goals.</p>
      <p>With this perspective, the first of our approaches proposes to identify this
manageable set through a goal importance ranking. This ranking allows us to know the
business goals that receive more cumulative impact from its offspring (all its sub-goals
down to OSS adoption strategy goals). This ranking also considers the total size of the
goal model, because, for example, a goal does not have the same importance if it
belongs to a model of 200 goals or if it belongs to a model of 20,000 goals, even if its
offspring is the same. It is important to emphasize that our analysis is topologic, not
semantic, and therefore do not consider the type of link.</p>
      <p>In graph theory, the centrality concept manages the importance of a node in the
network. From several centrality metrics, we decided to apply PageRank [11] because
it calculates the importance value for each node based on topological characteristics of
the model (number of goals and links among them) and works with a unique ‘root’
node. An excerpt of PageRank (PR) values for the goal model of our example is
presented in Table 1. They are obtained using Gephi tool (https://gephi.org/) with a
damping factor set in 1 (a value less than one and greater than or equal to zero is
assigned to damping factor when this algorithm is applied to web navigation graphs).
As we appreciate, the most impacted goals are in the first places of the ranking, that is,
goals which achievement depends on the achievement of a major number of sub-goals.
This is the case, for instance, of To ensure that income (revenue streams from the s/p/f)
are obtained as planned, that is the 3rd goal in the ranking with an importance value of
0.0586 and depends on 44 sub-goals, against the goal To offer the p/s/f required, that is
the 25th goal in the ranking with an importance value of 0.0084 and depends on 15
subgoals.</p>
      <p>This ranking may be used to know the most impacted node among nodes that have
the same detail level. For example, at the 4th level of detail, the importance value of To
establish a patent scheme (0.0020), is less than To incorporate external innovation
inputs into the business offering (0.0204): the difference is caused by the number of
sub-goals each has.
4</p>
    </sec>
    <sec id="sec-3">
      <title>Discovering Goal Clusters</title>
      <p>As we mentioned in the Introduction, an appropriate management of the goal model’s
complexity is a critical success factor to improve the analysis and understanding of goal
model. One way to deal with this issue is to modularize in order to divide an extensive
model into small, more manageable modules that can be analyzed and maintained as a
unit. In this sense, our Clustering approach groups the goals applying a clustering
algorithm to find, if possible, two or more community structures that could constitute
modules. A community structure is a set of nodes that has more connections between
its members than to the remainder of the network [12].</p>
      <p>We apply three clustering algorithms: Clauset-Newman-Moore (CNM) [13],
Wakita-Tsurumi (WT) [14], and Girvan-Newman (GN) [15]. In Table 2 we present the
synthesis of results. The CNM algorithm found 6 clusters: Offer &amp; Innovation, Strategy
&amp; Law compliance, Incomings, Oss Community, Human Talent, and Quality. In this
last one, the membership of 4 of its goals it is not quite clear; these goals are: To manage
customer relationships (establish, maintain and expand them), To ensure the output
logistic (customer delivery), To choose a compatible license, and ACQ-Leg (To acquire
legal skills). Over the others clusters, there are not doubts about its members. In the
Fig. 2 we show the clusters identified by the CNM algorithm. For the processes of
clustering and visualization, we use NodeXL Excel Template
(http://www.smrfoundation.org/).</p>
      <p>The Wakita-Tsurumi algorithm found 10 clusters, 2 of which are the same as those
found by the CNM algorithm (Human Talent and Community); 3 of them are very
similar (Offer &amp; Innovation, Strategy &amp; Law compliance, and Incomings; they have 3,
4 and 2 goals less than the correspondent CNM groups, respectively); 3 of them are
about Quality (component integration, component selection, and customer issues,
which in total have 4 goals less than CNM Quality group); 1 of them is new: Offer
Delivery; the last group comprises 6 goals (about market, offering, use of OSS
component, working practices) without a clear relationship.</p>
      <p>The Girvan-Newman algorithm found 7 clusters where the most relevant issues with
regard to CNM classification are: Quality is divided into 2 clusters (component
integration, and component selection); the Human Talent and OSS Community goals
are grouped into a single cluster; and, Legal goals are placed in a cluster with the goal
about to the shareholder value model sustainability.
5</p>
    </sec>
    <sec id="sec-4">
      <title>Conclusions</title>
      <p>In the present work, we have proposed two approaches to managing the complexity of
goal-oriented models, based on its topological characteristics. The first approach
generates a ranking of the importance that each goal has like part of an entire model
(without considering a goal in isolation); the highest values in the ranking correspond
to the goals with major relative importance, which can be selected to perform a deeper
analysis. The second approach seeks to identify groups of goals that can become
modules; thus, based on the application results of Clauset-Newman-Moore,
WakitaTsurumi, and Girvan-Newman algorithms, we found that the adequate goals grouping
is performed by the first of them; this algorithm generates modules which goals have
more affinity.</p>
      <p>Acknowledgments. This work is a result of the Q-Rapids project, which has
received funding from the European Union’s Horizon 2020 research and innovation
program under grant agreement N° 732253. Lucía Méndez’s work is supported by a
SENESCYT (Secretaría de Educación Superior, Ciencia, Tecnología e Innovación)
grant from the Ecuatorian Government.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          1.
          <string-name>
            <surname>Langermeier</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Saad</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Bauer</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          :
          <article-title>Adaptive Approach for Impact Analysis in Enterprise Architectures</article-title>
          . B.
          <string-name>
            <surname>Shishkov</surname>
          </string-name>
          (Ed.):
          <source>Business Modeling and Software Design 4th International Symposium BMSD</source>
          <year>2014</year>
          , pp.
          <fpage>22</fpage>
          --
          <lpage>42</lpage>
          (
          <year>2014</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          2.
          <string-name>
            <surname>Lee</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          :
          <article-title>Complexity Theory in Axiomatic Design</article-title>
          .
          <source>Ph.D. Thesis</source>
          , Massachusetts Institute of Technology,
          <year>2003</year>
          . Cambridge, MA: Massachusetts Institute of Technology (
          <year>2003</year>
          ).
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          3.
          <string-name>
            <surname>Kreimeyer</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Lindemann</surname>
            ,
            <given-names>U.</given-names>
          </string-name>
          : Complexity Metrics in Engineering Design. Springer (
          <year>2011</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          4.
          <string-name>
            <surname>Espada</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Goulão</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Araújo</surname>
            ,
            <given-names>J.:</given-names>
          </string-name>
          <article-title>A Framework to Evaluate Complexity and Completeness of KAOS Goal Models</article-title>
          .
          <source>In: 25th International Conference on Advanced Information Systems Engineering</source>
          , CAiSE'13, Springer-Verlag, Valencia, Spain, pp.
          <fpage>562</fpage>
          -
          <lpage>577</lpage>
          . (
          <year>2013</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          5.
          <string-name>
            <surname>Colomer</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Franch</surname>
            ,
            <given-names>X.</given-names>
          </string-name>
          :
          <article-title>Stargro: building i* metrics for agile methodologies</article-title>
          , in: Dalpiaz,
          <string-name>
            <given-names>F.</given-names>
            ,
            <surname>Horkoff</surname>
          </string-name>
          ,
          <string-name>
            <surname>J.</surname>
          </string-name>
          , (Eds.),
          <source>7th International i* Workshop (iStar2014)</source>
          ,
          <source>CEUR Workshop Proceedings</source>
          , vol
          <volume>1157</volume>
          (
          <year>2014</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          6.
          <string-name>
            <surname>Gralha</surname>
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Araújo</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Goulão</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          :
          <article-title>Metrics for measuring complexity and completeness for social goal models</article-title>
          .
          <source>Information Systems</source>
          , vol
          <volume>53</volume>
          , pp.
          <fpage>346</fpage>
          --
          <lpage>362</lpage>
          (
          <year>2015</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          7.
          <string-name>
            <surname>Maté</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Trujillo</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Franch</surname>
            ,
            <given-names>X.</given-names>
          </string-name>
          ;
          <article-title>Adding semantic modules to improve goal-oriented analysis of datawarehouses using I-star</article-title>
          .
          <source>The Journal of Systems and Software</source>
          ,
          <volume>88</volume>
          , pp.
          <fpage>102</fpage>
          --
          <lpage>111</lpage>
          (
          <year>2014</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          8.
          <string-name>
            <surname>Franch</surname>
            ,
            <given-names>X.</given-names>
          </string-name>
          :
          <article-title>Incorporating modules into the i* framework</article-title>
          .
          <source>In: CAiSE</source>
          , Vol.
          <source>6051of LNCS</source>
          . Springer, Berlin, pp.
          <fpage>439</fpage>
          -
          <lpage>454</lpage>
          (
          <year>2010</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          9.
          <string-name>
            <surname>Méndez</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>López</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ayala</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Annosi</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          :
          <source>Towards an OSS Adoption Business Impact Assessment, Lecture Notes in Business Information Processing</source>
          , vol
          <volume>235</volume>
          , pp
          <fpage>289</fpage>
          --
          <lpage>305</lpage>
          (
          <year>2015</year>
          )
          <fpage>10</fpage>
          .
          <string-name>
            <surname>López</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Costal</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ayala</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Franch</surname>
            ,
            <given-names>X.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Annosi</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Glott</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Haaland</surname>
          </string-name>
          , R.:
          <article-title>Adoption of OSS components: a goal-oriented approach</article-title>
          . In Data &amp; Knowledge
          <string-name>
            <surname>Engineering</surname>
          </string-name>
          (
          <year>2015</year>
          )
          <fpage>11</fpage>
          .
          <string-name>
            <surname>Newman</surname>
            ,
            <given-names>M. E. J.</given-names>
          </string-name>
          :
          <article-title>Networks - An introduction</article-title>
          . University of Michigan and Santa Fe Institute. Oxford University Press (
          <year>2010</year>
          )
          <fpage>12</fpage>
          .
          <string-name>
            <surname>Leskovec</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Lang</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Dasgupta</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Mahoney</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          :
          <article-title>Statistical Properties of Community Structure in Large Social and Information Networks</article-title>
          .
          <source>In: Proceedings of the 17th international conference on World Wide Web</source>
          , pp
          <fpage>695</fpage>
          -
          <lpage>704</lpage>
          (
          <year>2008</year>
          )
          <fpage>13</fpage>
          .
          <string-name>
            <surname>Clauset</surname>
            .,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Newman</surname>
            ,
            <given-names>M. E. J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Moore</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          :
          <article-title>Finding community structure in very large networks</article-title>
          .
          <source>Physical Review E</source>
          ,
          <volume>70</volume>
          :
          <fpage>066111</fpage>
          , (
          <year>2004</year>
          )
          <fpage>14</fpage>
          .
          <string-name>
            <surname>Wakita</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Tsurumi</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          :
          <article-title>Finding community structure in mega-scale social networks</article-title>
          .
          <source>Poster Paper on Proceedings of the 16th international conference on World Wide Web</source>
          , pp.
          <fpage>1275</fpage>
          -
          <lpage>1276</lpage>
          (
          <year>2007</year>
          )
          <fpage>15</fpage>
          .
          <string-name>
            <surname>Girvan</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Newman</surname>
            ,
            <given-names>M. E. J.:</given-names>
          </string-name>
          <article-title>Community structure in social and biological networks</article-title>
          .
          <source>Proceedings of the National Academy of Sciences of the United States of America PNAS</source>
          . vol
          <volume>99</volume>
          no.
          <issue>12</issue>
          , pp
          <fpage>7821</fpage>
          -
          <lpage>7826</lpage>
          (
          <year>2002</year>
          )
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>