<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Ontologies, ICTs and Law The International Ontojuris Project</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Bibiana Luz Clara</string-name>
          <email>bluzclara@ufasta.edu.ar</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Ana Haydée Di Iorio</string-name>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Roberto Giordano Lerena</string-name>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>: Abogada. Profesora Investigadora de la Facultad de Ingeniería de la Universidad FASTA. Presidente del Instituto de Derecho Informático del Colegio de Abogados de Mar del Plata.</institution>
          <country country="AR">Argentina</country>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>: Ingeniero en Sistemas. Profesor Investigador y Decano de la Facultad de Ingeniería de la Universidad FASTA.</institution>
          <country country="AR">Argentina</country>
        </aff>
      </contrib-group>
      <pub-date>
        <year>2010</year>
      </pub-date>
      <fpage>95</fpage>
      <lpage>102</lpage>
      <abstract>
        <p>This article presents the experience of the International Ontojuris Project, modeled and developed to search and retrieve multilingual legal information based on ontologies and on the Universal Networking Language (UNL). It also presents the issue of multilingual information management, the importance of data processing from the semantic point of view and the possibility of semantic interoperability between systems, basically on Web search engines. In the beginning, the development of legal informatics captured the attention of law practitioners to improve their working practices and increase accessibility to information and documents [1]. Due to the large amount of legal information in existence, it was necessary to find a support to facilitate access to this information, both to legal practitioners and citizens. The Documentary Legal Informatics would thus develop aiming at the automatic processing of legal information sources: legislation, jurisprudence and doctrine [2]. In Argentina, the best example of Documentary Legal Informatics is the Sistema Argentino de Informática Jurídica, SAIJ (Argentine System of Legal Informatics) created in 1979. The SAIJ is a government agency supervised by the Dirección Técnica de Formación e Información JurídicoLegal (Technical Office for Legal Information), under the Subsecretaría de Justicia (Justice Subsecretariat) in the Ministerio de Justicia, Seguridad y Derechos Humanos (Ministry of Justice, Security and Human Rights). It provides normative, jurisprudential and doctrinaire information, whether national or provincial, taken from official sources.</p>
      </abstract>
      <kwd-group>
        <kwd>Ontology</kwd>
        <kwd>Law</kwd>
        <kwd>Artificial Intelligence</kwd>
        <kwd>UNL</kwd>
        <kwd>Ontojuris</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>-</title>
      <p>The SAIJ1 also coordinates the National Network of Legal Informatics,
established in 1995. This net is constituted by all provinces that have
signed agreements with the entity. Each province has a cooperation center
in charge of providing and updating the provincial legal information2.
In the Argentine legal system, jurisprudence is a formal source of law. For
this reason, when a law practitioner carries out a jurisprudential search, he
is seeking to reinforce the interpretation of standards or a personal point of
view. In short, he attempts to present, by reference to verdicts, persuasive
arguments to influence the judge’s reasoning towards his side.
In addition to being limited by the syntactic search, most of these legal
information search systems require the thorough knowledge of the verdict
which the operator is trying to find: the year, actors involved, the court, and
subject matter. At the moment of search, both in government initiatives and
in private ones related to legal publishers, the legal practitioner frequently
retrieves a large amount of irrelevant data that should be refined repeatedly
until obtaining the desired result.</p>
      <p>
        Many Artificial Intelligence (AI) techniques related to the representation of
knowledge have tried to solve this problem. Among them, the
representation by “ontologies” is noted; it refers to the formulation of a
conceptual scheme within a given domain to allow the search of knowledge
through meaning. This “ontological” representation is the basis for a real
“Semantic Web” [
        <xref ref-type="bibr" rid="ref3">3</xref>
        ][
        <xref ref-type="bibr" rid="ref4">4</xref>
        ] by which the legal practitioner will be able to
retrieve information from concepts, semantically, or by obtaining the exact
data related to the search, all these independently from the possibility that
in the referred text the specific term could be used at the moment of query
[
        <xref ref-type="bibr" rid="ref5">5</xref>
        ][
        <xref ref-type="bibr" rid="ref6">6</xref>
        ][
        <xref ref-type="bibr" rid="ref7">7</xref>
        ].
      </p>
      <p>In addition to the usual problems of legal search, the globality of law
appears together with the complexity of varied conceptualization related
not only to language but also to the particular cultures to which the concept
refers.</p>
    </sec>
    <sec id="sec-2">
      <title>Integration in a multilingual world</title>
      <p>In a globalized, multicultural and multilingual world, access to resources is
limited by multiple barriers, among them language and those originated in
the interpretation of the real world to be conceptualized.
1 Source www.saij.jus.gov.ar – Institutional
2 There are also numerous initiatives of legislative, executive and judicial
entities linked to the publication of legal information. Namely, the JUBA system of
the Supreme Court of Justice of the Province of Buenos Aires, now accessible via
the Web. It includes summaries and complete veredicts in the Province of Buenos
Aires; the FANA system, also accessible on the Web including nationwide
summaries and veredicts.
The use of ontologies, and conceptualization in itself, is complex, no matter
which culture or language it is dealt with (or conceptualized). A higher
level of complexity occurs when expanding the coverage of the proposed
solution and the necessary conceptualization, to a multicultural and
multilingual world.</p>
      <p>Culture, terms, concepts, relationships between terms and term-concept
relationships differ from one place to another, from one language to
another. Whatever the domain, a given conceptualization that is valid in a
certain language may not be acceptable in a different one. Thus, it is not
possible to automate the process to a strict automatic translation, and, above
all, it is impossible to homologate terms and concepts in different
languages. There are words which cannot be translated in any language,
simply because its strict meaning and use in its place of origin (in that
specific culture) does not have a strict equivalent in another culture. Each
term is the representation of a concept in a given language, and it may turn
out that this concept will not have its equivalent in another language; hence,
a possible word to represent that concept in that second language will not
exist. Even in the same language (Spanish, for example), the same word
may be used to represent different concepts in different countries or
regions, and the same concept may be represented by different words in
different regions using the same language.</p>
      <p>The cultural complexity of languages and multilingualism are then
transformed in a barrier to communication. In an interconnected and
globalized world, it is urgent and necessary to undertake these issues from a
technological point of view in order to facilitate intercultural
communication.</p>
      <p>Information on the web is growing daily and there is an urgent need to find
“intelligent” searchers capable to work with semantics in the language and
place where the search is performed. Users need to search by concept, not
by term. Users think according to concepts, but must search the web for
terms. The search is usually syntactic, not semantic. Browsers retrieve and
return web pages containing search terms as they are spelt, textually, and
not semantically. Some of them even propose pages in different languages
where the specific term appears (having a different meaning in that
language), without the ability to discriminate or prioritize on behalf of the
concept.</p>
      <p>There is also specific terminology in each domain or specialty which makes
certain terms (words or set of words) have different meanings in a
language, being the same country and culture. There are also concepts built
on the basis of words that separately, have a certain meaning, but with a
composition that does not mean the composition of those meanings.
Traditional translators do not recognize this kind of compositions, namely,
the representation of new concepts.
Users around the world need to see web pages from other countries, but on
the basis of a specific concept, not terms representing it in their languages.
This requires the development of search engines capable to understand and
process the concept associated to the indicated term; with that concept (or
meaning), engines should find pages having any of those terms or
expressions representing it in their respective languages. These are known
as intelligent search engines. They index and retrieve on-line meanings or
concepts instead of words or terms. They include a conceptual
infrastructure and ontological relationships that allow such management of
search.</p>
      <p>Given the importance of the term (actually, the concept) of the query, its
vital correct interpretation, and since the above mentioned must be limited
to the scope of law, the problem is increased. Misconception of a legal
document is a very high risk that law practitioners cannot take;
consequently, they need the support of technological tools to collaborate
with their work and guarantee the correct interpretation of data, terms,
information and documents involved in their decisions.</p>
    </sec>
    <sec id="sec-3">
      <title>The UNL Program</title>
      <p>In 1996, the United Nations General Assembly crated the Universal
Networking Language Program (UNLP) as a project of the United Nations
University (UNU).</p>
      <p>The aim and activities of the UNLP are to develop and promote platforms
and communication and information tools that will provide every nation the
same opportunities to access, share and exchange scientific, cultural, social,
and economic resources available in the global village. The UNLP has a
flexible and dynamic net of persons and institutions devoted to developing,
expanding, improving and multiplying the UNL System www.fi.unl.upm.es
as a means of overcoming linguistic barriers; it is also a platform to collect
and multiply human knowledge among people speaking different
languages3.</p>
      <p>Ultimately, the project aims at allowing any person to share and retrieve
information in their own language, no matter the language originating it.
The project counted with an initial participation of 15 languages: German,
Arabic, Chinese, Spanish, French, Hindi, Indonesian, English, Italian,
Japanese, Latvian, Mongolian, Portuguese, Russian and Thai. The UNL
System basically consists of UNL servers, UNL editors, and UNL viewers.
The UNL language consists in UNL relationships and their attributes,
universal terms, and a knowledge database4.
The Centro de Lengua Española (Spanish Language Center) and their
group are working under the assistance of the Universidad Politécnica de
Madrid (UPM) and they represent the Spanish language, not only in Spain,
but also in every country sharing this language. Consequently, with this
pretended universal program, the UNL language is adopted as the platform
for the development of a multilingual legal server based on ontologies to
pursue the International Ontojuris Project.</p>
    </sec>
    <sec id="sec-4">
      <title>The Ontojuris Project</title>
      <p>The International Ontojuris Project aims at facilitating a multilingual access
to information about legal documents in the areas of Intellectual Property
Law, Consumer Rights and Informatics Law. A consortium was then
formed by researchers from Argentina, Brazil and Spain. Argentina was
represented by Universidad FASTA; Brazil, by Instituo I3G; and Spain, by
Univesidad Politécninca de Madrid. Experts from Universidad de Chile are
also collaborating with the project.</p>
      <p>
        The overall objective of the program consists in the research and
development of an intelligent multilingual system based on ontologies, for
the retrieval of legal information, limited in a first stage to the domains of
Intellectual Property Law, Consumer Rights and Informatics Law [
        <xref ref-type="bibr" rid="ref8">8</xref>
        ].
Broadly, the stages in this project are as follows:
a) Selection of texts related to legislation, jurisprudence and doctrine
of the domains involved in order to generate ontologies associated
to each domain.
b) Identification and definition of terms and patterns of relationship.
      </p>
      <p>Construction of specific ontologies5.
c) English determination of Headword6 for each term.
d) Conversion of each term to the UNL and construction of domain
terms inexistent in dictionary7.
5 The Project is developed under its own ontology editor, provided by I3G
which anticipates the definition of some relationships.</p>
      <p>The definition of ontologies was based on the identification of terms and
by linking them after the following relationships: Synonymy, Type of (category or
class), and Part (fraction or component).
6 The UW (Universal Words) constitute the vocabulary of the UNL. They
are concept labels, syntactic and semantic units that combine to form the UNL
expression. Each UW represents a concept. A UW is formed by a Headword and a
list of restrictions. The Headword may be a word, a compound word or a phrase in
English. The list of restrictions is associated to the Headword to disambiguate and
add specifications.
7 The retrieval of universal words to fulfil the Ontology was achieved by
referring to the UNL dictionary available at the Centro de Lengua Española. Those
not included in the data base were created.
e) Definition of measures for indexing ontologies8.
f) Definition of parameters allowing flow of ontologies.
g) Development of procedures for the integration of the ontology
editor with applications and tools for the web search.
h) Modelling of the tool interphase.
i) Integration of the UNL to the ontology editor.
j) Expansion of search through the Universal Word (UW)9.
k) Specification of the result presentation.</p>
      <p>In developing the tool, the methodology of Knowledge Engineering, based
on the semi-formal ontology description, was used to support the process of
ontology engineering. In this methodology, instances of representation do
not include the description of objects, only their relationships within a
given domain.</p>
      <p>The editor was designed to support the task and experience of the
Knowledge Engineers when constructing the multilingual ontologies. It is a
complex structure which connects terms taking into account concepts in the
knowledge of their specific application. This allows the editor to determine
the context of the documents in the query: they are contextualized.
The basic components of the ontology editor are: classes (taxonomically
organized) and relationships (representing the type of interaction between
concepts in a given domain). The ontology representation dos not use
axioms or instances.</p>
    </sec>
    <sec id="sec-5">
      <title>Discussion. Aiming at the future.</title>
      <p>Having concluded the Ontojuris prototype, it is now time to verify its
utility “in the field”, with law practitioners from different countries
validating the system.
8 The Ontojuris system uses the Ontology created by the editor to index
and retrieve information in the specified legal documents (laws, decrees, doctrine).
The terms established by the method of creation of Ontology are used in the
indexing process. The terms considered as relevant in a given document are added
to the list of terms. On the other hand a list of words is generated from a dictionary
of the natural language of the document. Hence, each document is labeled wiith the
indices of terms and the words it contains.
9 This phase, not yet completed, suggests a new expansion method based
on domain ontologies and UW in order to retrieve multilingual information. This is
an ongoing study based on the possibility of relying on a domain dictionary of
UW for each natural language of the original documents. Each document to be
searched is converted into a term vector and a word vector; besides it is mapped
with a UW vector by which it is also indexed. That is to say, during indexing, the
system converts each term into its corresponding UW, and it converts the original
term into each different language associated to that UW.
Future steps must also be discussed. On the one hand, the expansion of the
project to other disciplines given that this type of browser may be adjusted
to a domain in any field provided that experts accomplish an accurate
selection of ontologies.</p>
      <p>On the other hand, members of other languages will be invited to the
consortium to expand multilingual competences to other languages.
In the field of system integration, and given the level of knowledge
embodied in the ontology of Law, it is necessary to work for an on-line
integration with Law systems in Latin America certifying the semantic
inter-operability among the systems.</p>
      <p>Finally, it must be noted that the participation of the consortium
Universities in the future course UNESCO “TECLIN” – Linguistic
Technologies for Children Education in Aboriginal Communities will allow
methodologies and the developed technology to extrapolate to other fields
and contribute to the fulfilment of the United Nations goals for the
millennium.</p>
      <p>There is already a pilot project in Argentina for the recovery of endangered
languages (particularly, the Quechua) which enables a “dialogue” between
modern languages, such as Spanish, and aboriginal languages (declared
World Heritage Site) favoring conservation. The project is developed on
the same methodological and technological basis carried out for interaction
between different languages in the field of Law. With very encouraging
preliminary results, the Centro de Investigación CIPCO (CIPCO Research
Center), La Buhardilla Foundation, in Tucuman, Argentina, is working in
this direction with the support of the Ontojuris Universities.</p>
    </sec>
    <sec id="sec-6">
      <title>Conclusion</title>
      <p>After overcoming issues like the availability of digital information,
connectivity and technical interoperability, cultural diversity and linguistics
appear to be the real problems to reach a global knowledge society. It is at
this point where technology of information has a fundamental role and an
enthralling challenge.</p>
      <p>The Ontojuris project reveals the potential of technological tools available
to add up to the socialization of knowledge in the great “Global Village”.</p>
    </sec>
    <sec id="sec-7">
      <title>Acknowledgements</title>
      <p>The authors are grateful to the members of the Ontojuris project of
Universidad FASTA and to the members of the Ontojuris Project in Spain
and Brazil. Also, to Tania Bueno, Sonali Bedin, Hugo Hoeschl, Cesar
Stradiotto and Jesús Cardeñosa.
We would like to thank Ana Inés Cosulich for her assessment in the
English version of this paper.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          [1]
          <string-name>
            <given-names>Peñaranda</given-names>
            <surname>Quintero</surname>
          </string-name>
          (
          <year>2001</year>
          ),
          <article-title>Iuscibernética, interrelación entre el derecho y la informática</article-title>
          , Ed. Miguel García e hijo, Caracas, Venezuela.
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          [2]
          <string-name>
            <given-names>Luz</given-names>
            <surname>Clara</surname>
          </string-name>
          ,
          <string-name>
            <surname>Bibiana</surname>
          </string-name>
          (
          <year>2001</year>
          ), Manual de Derecho Informático, Ed. Nova Tesis, Rosario, Argentina.
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          [3]
          <string-name>
            <surname>Castells</surname>
          </string-name>
          , Pablo, La web semántica, disponible en http://arantxa.ii.uam.es/~castells/publications/castells-uclm03.
          <fpage>pdf</fpage>
          . (accedida 2 de Mayo de
          <year>2009</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          [4]
          <string-name>
            <surname>Castells</surname>
          </string-name>
          , Pablo, Búsqueda semántica basada en conocimiento, disponible en http://nets.ii.uam.es/publications/castells-fds08.pdf (accedida 15 de Abril de
          <year>2009</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          [5]
          <string-name>
            <given-names>Pompeu</given-names>
            <surname>Casanovas</surname>
          </string-name>
          (
          <year>2005</year>
          ),
          <article-title>Ontologías jurídicas profesionales</article-title>
          .
          <source>Sobre conocer y representar el Derecho</source>
          , disponible en http://www.leibnizsociedad.org/secciones/mater/pon/textos/ontolo gias_pompeu.pdf. (accedida el 17 de Junio de
          <year>2009</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          [6]
          <string-name>
            <surname>Abian</surname>
            ,
            <given-names>Miguel</given-names>
          </string-name>
          <string-name>
            <surname>Ángel</surname>
          </string-name>
          (
          <year>2005</year>
          ),
          <article-title>Ontologías, que son y para que sirven</article-title>
          , disponible en http://www.wshoy.sidar.org/index.php?
          <year>2005</year>
          /12/09/30- ontologias
          <article-title>-que-son-y-para-que-sirven. (accedida el 4</article-title>
          de Octubre de
          <year>2009</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          [7]
          <string-name>
            <given-names>Burners</given-names>
            <surname>Lee</surname>
          </string-name>
          ,
          <string-name>
            <surname>Tim</surname>
          </string-name>
          (
          <year>2000</year>
          ),
          <article-title>Conference on the Semantic Web disponible en http://www</article-title>
          .w3.org/2000/Talks/1206-xml2k-tbl/slide10-
          <fpage>0</fpage>
          .html. (accedida el 20 de Octubre de
          <year>2009</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          [8]
          <string-name>
            <given-names>Proyecto</given-names>
            <surname>Ontojuris</surname>
          </string-name>
          :
          <article-title>Disponible en www</article-title>
          .i3g.org.br/ontojuris/sistema.html
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>