<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Using ontologies to build a database to obtain strategic information in decision making</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>E´rica F. Souza</string-name>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Leandro E. Oliveira</string-name>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Ricardo A. Falbo</string-name>
          <email>falbo@inf.ufes.br</email>
        </contrib>
        <contrib contrib-type="author">
          <string-name>N. L. Vijaykumar</string-name>
        </contrib>
        <contrib contrib-type="author">
          <string-name>erica.souza</string-name>
        </contrib>
        <contrib contrib-type="author">
          <string-name>vijay}@lac.inpe.br</string-name>
        </contrib>
        <contrib contrib-type="author">
          <string-name>leandro.oliveira</string-name>
        </contrib>
        <contrib contrib-type="author">
          <string-name>@fatec.sp.gov.br</string-name>
        </contrib>
      </contrib-group>
      <fpage>200</fpage>
      <lpage>205</lpage>
      <abstract>
        <p>The manipulation of ontologies in databases can represent gains in the recovery of strategic information in decision-making process within the software development organizations. The software testing processes are strategic elements to develop projects and to the quality of the final product. Thus, this study investigates strategies to promote data handling of testing processes that are generated from a testing ontology. For this, a knowledge database is structured in a dimensional model for Data Warehouse to support storage and processing of data to obtain strategic information that can facilitate decision making. Resumo. A manipulac¸a˜o de ontologias em bancos de dados pode representar ganhos na recuperac¸a˜o de informac¸o˜es estra´tegicas no processo de tomada de decisa˜o dentro das organizac¸o˜es de desenvolvimento de software. Os processos de teste de software sa˜o elementos estrate´gicos para a conduc¸a˜o de projetos de desenvolvimento e qualidade do produto final. Diante disso, este trabalho tem como objetivo investigar estrate´gias que possam promover a manipulac¸a˜o de dados de processos de teste que sa˜o gerados a partir de uma ontologia de teste de software. Para isso, estrutura-se uma base de conhecimento em um modelo dimensional de Data Warehouse que apoie o armazenamento e o processamento dos dados para obtenc¸a˜o de informac¸o˜es estrate´gicas que podem facilitar a tomada de decisa˜o.</p>
      </abstract>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>1. Introduction</title>
      <p>With the exponential growth of data from several different sources of knowledge within an
organization, it becomes necessary to provide automatized support for tasks of acquiring,
processing, analyzing and disseminating knowledge. Organizations need to effectively
manage the information generated in its production environment to promote the
improvement of the processes used to generate knowledge and also support future decisions. Such
data can provide important information for decision making, involving the identification
and implementation of corrective actions.</p>
      <p>
        One of the characteristics of software engineering projects is to deal with a great
deal of information that are generated and manipulated. People involved in the project
face problems, such as: organize in a systematic way the information generated through
the software process; reuse the knowledge generated from one project to another; loss
of intellectual capital of the organization due to better opportunities; and no knowledge
representation [
        <xref ref-type="bibr" rid="ref1">Andrade et al. 2010</xref>
        ].
      </p>
      <p>
        In the area of software development, testing is a critical factor in product quality,
and thus there is a greater concern with related research. Studies indicate that the quality
of the software product is strongly dependent on the quality of the processes that are
part of the project, especially the software testing process. However, finding relevant
information (knowledge) in these processes can be a difficult and complex task, and it is
related mainly to the lack of semantics associated with the large volume of information.
There is a need to represent knowledge, to make it affordable and manageable. In this
context, ontologies have been pointed out as an important way for representing knowledge
[
        <xref ref-type="bibr" rid="ref13">Rios 2005</xref>
        ].
      </p>
      <p>
        The manipulation of big ontologies with a high number of instances in the form
of text files has a number of disadvantages, such as processing and query
optimization [
        <xref ref-type="bibr" rid="ref8">Filho et al. 2010</xref>
        ]. Because of this, ontologies can be incorporated into a
knowledge base to facilitate its management and access. Depending on the structure the
database is created, the analysis of large data volumes and mining strategic
information can facilitate decision making. Related work is found in [
        <xref ref-type="bibr" rid="ref2">Astrova et al. 2007</xref>
        ],
[
        <xref ref-type="bibr" rid="ref15">Vysniauskas and Nemuraite 2006</xref>
        ], but they propose approaches that transform ontology
representation into a relational database, but not in a dimensional model as this enables
dealing with large volumes of information.
      </p>
      <p>Given the above context, this paper aims to investigate strategies that can promote
the manipulation of data generated from large ontologies. For this, a knowledge base in a
dimensional model for Data Warehouse is structured to support the storage and processing
of data to obtain strategic information that can facilitate decision making. A software
testing ontology is being developed to support this work and instances of this ontology
is used as a data source for the structure created. Section 2 briefly discusses ontologies
and storage structures. Section 3 presents the proposed structure to store ontology data.
Section 4 presents conclusions and future directions to follow.</p>
    </sec>
    <sec id="sec-2">
      <title>2. Ontologies and structures for storage</title>
      <p>
        During the last decades, ontologies have been shown useful in the field of Computer
Science [
        <xref ref-type="bibr" rid="ref10">Guizzardi 2005</xref>
        ]. In a nutshell, an ontology is a formal specification of a shared
conceptualization [
        <xref ref-type="bibr" rid="ref9">Gruber 1993</xref>
        ], i.e., a description of concepts and relationships that may
exist for an agent or an agents community. Representation of a shared conceptualization
requires a representation language. There are many representation languages. Some are
defined based on the syntax of the eXtensible Markup Language (XML), like Resource
Description Framework (RDF), Ontology Interchange Language (OIL) and Web
Ontology Language (OWL). There is also graphical language for ontologies; an example is the
Graphical Language for Expression Ontologies (LINGO) [
        <xref ref-type="bibr" rid="ref7">Falbo et al. 1998</xref>
        ].
      </p>
      <p>
        Literature provides several tools to store and manipulate content from an
ontology in a database and several strategies can be found [
        <xref ref-type="bibr" rid="ref8">Filho et al. 2010</xref>
        ]. Most tools use
relational databases. Depending on the number of instances of an ontology, it becomes
necessary to create structures that support high volumes of data and allow employing
techniques to find relationships among these data. An alternative is the use of Data
Warehouse (DW) as the storage structure. It is a large repository of integrated data obtained
from several sources for the specific purpose of data analysis [
        <xref ref-type="bibr" rid="ref4">Christian et al. 2010</xref>
        ]. A
DW may take on different models: Cube and Star. The difference is in the Database
Management System (DBMS). When the dimensional model is implemented in a relational
database it is implemented as a Star architecture and when implemented in a
multidimensional database it is known as Cube. Considering that DW stores a large volume of data, it
optimizes and reduces the complexity of consultations, thus decreasing the response time
and gaining in performance.
      </p>
    </sec>
    <sec id="sec-3">
      <title>3. Proposed structure for storage of ontology content</title>
      <p>
        The ontology used in this paper is in the context of the Ph.D. thesis of the first author
[
        <xref ref-type="bibr" rid="ref14">Souza 2011</xref>
        ]. The ontology aims at software testing for Knowledge Management (KM).
Test activity is incorporated into a process as it consists of several steps to add quality to
the final products [
        <xref ref-type="bibr" rid="ref3">Bastos et al. 2007</xref>
        ]. Since the test activity is a process, the
improvement can be incorporated with the use of KM and ontologies are identified today as being
crucial in the KM in processes improvement.
      </p>
      <p>
        Given the complexity of the software testing domain, an ontology and its
subontologies were created and used. Currently the main ontology created has: Steps,
Techniques, Types, Artifacts and Environment. SABiO (Systematic Approach for Building
Ontologies) was adopted to develop the software testing ontology [
        <xref ref-type="bibr" rid="ref6">Falbo 2004</xref>
        ].
      </p>
      <p>As the ontology is still under development it may suffer some changes. However,
though in the initial phase, it is already capable of describing execution phase of the tests,
and from this premise that DW will be created to store ontology data. This phase contains
data related to dates, for example, date of test case execution, date of defect submission
and date of defect correction.</p>
      <p>
        Instances of the software test ontology were extracted from an actual project
developed at the (Technological Institute of Aeronautics - ITA) - (Project of Amazon
Integration and Cooperation for Modernization of Hydrological Monitoring - ICA-MMH B)
[
        <xref ref-type="bibr" rid="ref5">Cunha 2010</xref>
        ]. We used test data generated from Organizational Testing Management
Maturity Model (OTM3) testing process [
        <xref ref-type="bibr" rid="ref12">Lamas et al. 2010</xref>
        ].
      </p>
      <p>For converting the content of an ontology written in RDF into a DW architecture,
we follow the process shown in Figure 1.</p>
      <p>
        To enable reading the ontology in the Pentaho Data Integration tool it was
necessary to transform the RDF/OWL in a simple XML and export the data of the instances for
the DW model created. For this, we used the
        <xref ref-type="bibr" rid="ref11">Jena framework [Jena 2012</xref>
        ] that provides an
Application Programming Interface (API) in Java allowing writing, reading and
extracting the description of fields, classes and instances from a file in the RDF/OWL in XML
format. With the support of Jena framework, a Java class that converts the data into a pure
XML was developed, that is, using only the native tags of XML without the RDF/OWL
specific tags. We call thissimple XML, as shown in step three in Figure 1. The XML with
simple pattern is to be read by Pentaho.
      </p>
      <p>From the simple XML Pentaho is used to extract, transform and load the data
for DW using ETL (Extract Transform Load). The data is loaded in DW table of facts.
Star model was chosen to create the DW. The model consists of table ft test execution
which is the table of facts, and five dimensions, namely: (i) three dimensions of time:
lk execution time contains the date of the test execution, lk submission time refers to the
date of a defect submission, and lk last change time refers to the date of the last change
made in the request to repair a bug; the lk team dimension contains information about
the tester, such as the name of the tester and his or her level within the team; and ii) the
lk project dimension contains information with respect to the project for which the test
was performed. Figure 2 shows the Star model created with its respective tables.</p>
      <p>In order to see the contribution of the methodology, some questions (that might
be important for decision makers) were posed. The idea is whether the table of facts
can, in fact, answer such questions in a satisfied manner and useful to decision makers or
managers. Therefore, just as an example, the following questions were defined to exercise
the table of facts and the different dimensions of the DW:
1. What is the average time between the report of a defect and the test run to check if
the bug was repaired? To this question the result obtained was an average of 5.85
days. This information can be useful for the manager or responsible for project to
analyze whether the time between a request and the execution of the tests are on
schedule.
2. What is the percentage of tests executed per profile of the team? Figure 3(a) shows
the results obtained for this question. Most tests were executed by team members
with the position of tester, which is logical since the rest of the tests were executed
by analysts and the main function of the analyst is to develop test plans and not
to execute them. The result is consistent with the reality of a team of software
testing.
3. How severe are the defects reported? Most defects reported have minimal severity
and only 3% of the defects are of large severity. This suggests that the process of
system development is a reasonable quality control (Figure 3(b)).
4. What is the relation between testing and defect solving? Most of the reported
defects have been solved so far. Only 2% of the requests were reopened indicating
a recurrence of an already reported defect, and only 1% of duplicate requests that
are still under correction (Figure 3(c)).
(a) Number of Executions by
profile.</p>
      <p>(b) Number of tests by defect
severity.</p>
      <p>(c) Number of tests by resolution of
the defect.</p>
    </sec>
    <sec id="sec-4">
      <title>4. Conclusions</title>
      <p>This paper presented structuring and storing of ontology content in a DW model. We used
a preliminary ontology of software testing process that is under development. The data
analyzed refers to the test execution phase. The DW model created is expandable and
may include other phases within the process of software testing.</p>
      <p>The model facilitates queries and can also provide strategic information that can
be used to improve the development process as well as for the process of software testing.
The model can be enriched and provide more information that can be used in decision
making. Some difficulties were encountered as lack of support for the Jena framework
and real case studies with test procedures clearly defined.</p>
      <p>Future directions include the automation of data entry in the ontology from other
sources, model expansion to store data from other phases of the software testing process
and integration with the software development process.</p>
    </sec>
    <sec id="sec-5">
      <title>Acknowledgements</title>
      <p>FAPESP and CNPq (PIBIC) for the financial support. ITA, Brazilian Water
Agency (ANA), Brazilian Agency of Research and Projects Financing (FINEP) and the
Casimiro Montenegro Filho Foundation (FCMF) for providing the data of Project FINEP
5206/06 for this work.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          <string-name>
            <surname>Andrade</surname>
          </string-name>
          , M. T. T.,
          <string-name>
            <surname>Ferreira</surname>
            ,
            <given-names>C. V.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Pereira</surname>
            ,
            <given-names>H. B. B.</given-names>
          </string-name>
          (
          <year>2010</year>
          ).
          <article-title>Uma ontologia para a gesta˜o do conhecimento no processo de desenvolvimento de produto</article-title>
          .
          <source>Gesta˜o e Produc¸a˜o</source>
          ,
          <volume>17</volume>
          :
          <fpage>537</fpage>
          -
          <lpage>551</lpage>
          . http://dx.doi.org/10.1590/
          <fpage>S0104</fpage>
          -530X2010000300008.
          <source>Access in: Ago</source>
          <year>2012</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          <string-name>
            <surname>Astrova</surname>
            ,
            <given-names>I.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Korda</surname>
            ,
            <given-names>N.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Kalja</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          (
          <year>2007</year>
          ).
          <article-title>Rule-based transformation of sql relational databases to owl ontologies</article-title>
          .
          <source>In: Proceedings of the 2nd International Conference on Metadata &amp; Semantics Research.</source>
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          <string-name>
            <surname>Bastos</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Rios</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Cristalli</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Moreira</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          (
          <year>2007</year>
          ). Base de conhecimento em testes de software.
          <source>Martins Editora Livraria, Sa˜o Paulo</source>
          ,
          <volume>2</volume>
          <fpage>edition</fpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          <string-name>
            <surname>Christian</surname>
            ,
            <given-names>S. J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Pedersen</surname>
          </string-name>
          , T. B., and
          <string-name>
            <surname>Thomsen</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          (
          <year>2010</year>
          ).
          <source>Multidimensional Databases and Data Warehousing</source>
          , volume
          <volume>2</volume>
          . Morgan and Claypool,
          <volume>1</volume>
          <fpage>edition</fpage>
          .
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          <string-name>
            <surname>Cunha</surname>
            ,
            <given-names>A. M.</given-names>
          </string-name>
          (
          <year>2010</year>
          ).
          <source>Relato´rio Te´cnico do 5 ◦ Semestre do Projeto FINEP 5206/06. Technical report, Sa˜o Jose´ dos Campos.</source>
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          <string-name>
            <surname>Falbo</surname>
            ,
            <given-names>R. A.</given-names>
          </string-name>
          (
          <year>2004</year>
          ).
          <article-title>Experiences in using a method for building domain ontologies</article-title>
          . In: International Workshop on Ontology in Action, pages
          <fpage>474</fpage>
          -
          <lpage>477</lpage>
          . Banff, Canada.
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          <string-name>
            <surname>Falbo</surname>
            ,
            <given-names>R. A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Menezes</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Rocha</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          (
          <year>1998</year>
          ).
          <article-title>A systematic approach for building ontologies</article-title>
          .
          <source>In: In Proceedings of the 6th Ibero-American Conference on AI, IBERAMIA98</source>
          . Lisbon, Portugal.
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          <string-name>
            <surname>Filho</surname>
            ,
            <given-names>S. N. V.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Moura</surname>
            ,
            <given-names>A. M. C.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Cavalcanti</surname>
            ,
            <given-names>M. C. R.</given-names>
          </string-name>
          (
          <year>2010</year>
          ).
          <article-title>Armazenamento e manipulac¸a˜o de ontologias utilizando sistemas gerenciadores de banco de dados</article-title>
          .
          <source>Technical report</source>
          , Instituto Militar de Engenharia (IME), Rio de Janeiro.
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          <string-name>
            <surname>Gruber</surname>
            ,
            <given-names>T. R.</given-names>
          </string-name>
          (
          <year>1993</year>
          ).
          <article-title>Toward principles for the design of ontologies used for knowledge sharing</article-title>
          .
          <source>In: Formal Ontology in Conceptual Analysis and Knowledge Representation. Padova</source>
          , Italy.
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          <string-name>
            <surname>Guizzardi</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          (
          <year>2005</year>
          ).
          <article-title>Ontological foundations for structural conceptual models</article-title>
          .
          <source>Telematica Institute Fundamental Research Series, The Netherlands. ISBN 90-75176-81-3.</source>
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          <string-name>
            <surname>Jena</surname>
          </string-name>
          (
          <year>2012</year>
          ).
          <source>Apache Software Foundation - Jena</source>
          . http://jena.apache.org/documentation. Access in: Aug.
          <year>2012</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          <string-name>
            <surname>Lamas</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Souza</surname>
            ,
            <given-names>E. F.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Nascimento</surname>
            ,
            <given-names>M. R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Dias</surname>
            ,
            <given-names>L. A. V.</given-names>
          </string-name>
          , and
          <string-name>
            <surname>Silveira</surname>
            ,
            <given-names>F. F.</given-names>
          </string-name>
          (
          <year>2010</year>
          ).
          <article-title>Organizational testing management maturity model for a software product line</article-title>
          .
          <source>In: Seventh International Conference on Information Technology, ITNG'2010</source>
          , pages
          <fpage>1026</fpage>
          -
          <lpage>1031</lpage>
          . Las Vegas, Nevada, USA.
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          <string-name>
            <surname>Rios</surname>
            ,
            <given-names>J. A.</given-names>
          </string-name>
          (
          <year>2005</year>
          ).
          <article-title>Ontologias: alternativa para a representac¸a˜o do conhecimento expl´ıcito organizacional</article-title>
          . In:
          <string-name>
            <surname>Proceedings CINFORM - Encontro Nacional de Cieˆncia da Informac</surname>
          </string-name>
          <article-title>¸a˜o VI. Salvador, Bahia</article-title>
          .
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          <string-name>
            <surname>Souza</surname>
            ,
            <given-names>E. F.</given-names>
          </string-name>
          (
          <year>2011</year>
          ). Estrate´gias de reuso para melhoria de processo de teste de
          <article-title>software baseado em ontologias</article-title>
          .
          <source>Technical report</source>
          , INPE, Sa˜o Jose´ dos Campos/SP. http://urlib.net/8JMKD3MGP7W/3BFFA9H. Access in:
          <year>June 2012</year>
          .
        </mixed-citation>
      </ref>
      <ref id="ref15">
        <mixed-citation>
          <string-name>
            <surname>Vysniauskas</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          and
          <string-name>
            <surname>Nemuraite</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          (
          <year>2006</year>
          ).
          <article-title>Transforming ontology representation from owl to relational database</article-title>
          .
          <source>In: Information Technology and Control.</source>
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>