<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta>
      <journal-title-group>
        <journal-title>October</journal-title>
      </journal-title-group>
    </journal-meta>
    <article-meta>
      <title-group>
        <article-title>MULTIAGENT INFORMATION TECHNOLOGIES IN SYSTEM ANALYSIS</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>V.A. Inkina</string-name>
          <email>vainkina@mephi.ru</email>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>E.V. Antonov</string-name>
          <email>eantonov@kaf65.ru</email>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>A.A. Artamonov</string-name>
          <email>aaartamonov@mephi.ru</email>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>K.V. Ionkina</string-name>
          <email>ionkinakristina@gmail.com</email>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>E.S. Tretyakov</string-name>
          <email>estretyakov@mephi.ru</email>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>A.I. Cherkasskiy</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>National Research Nuclear University MEPhI (Moscow Engineering Physics Institute)</institution>
          ,
          <addr-line>Moscow</addr-line>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>Vera Inkina</institution>
          ,
          <addr-line>Evgenii Antonov, Alexey Artamonov, Kristina Ionkina, Evgeny Tretyakov, Andrey Cherkasskiy</addr-line>
        </aff>
      </contrib-group>
      <pub-date>
        <year>2019</year>
      </pub-date>
      <volume>4</volume>
      <issue>2019</issue>
      <fpage>195</fpage>
      <lpage>199</lpage>
      <abstract>
        <p>Agent technologies currently play an increasingly important role in the information technology industry given its ability to learn and evolve, to solve information management problems, to employ data visualization and many other benefits. As a computer program, an agent deals with a challenge Internet users face every single day: to obtain reliable and effective data in the specific thematic field. Multiagent system consists of two or more autonomous agents and aimed at solving complex problems, such as Big Data, Data mining, primary structured and unstructured information processing (including text, numbers and multimedia types of data).</p>
      </abstract>
      <kwd-group>
        <kwd>agent</kwd>
        <kwd>multiagent technology</kwd>
        <kwd>agent system</kwd>
        <kwd>Big Data</kwd>
        <kwd>intellectual agent</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>1. Introduction</title>
      <p>Science and technology is one of the most significant development directions within the scope of
which technologies are created, solutions are implemented that effectively respond to the grand challenges.
Problem solving in this area requires specific complex knowledge on a great number of research areas in
conditions of incomplete information, limited resources and lack of time.</p>
      <p>Big Data and Data mining class of problems describe knowledge acquisition problem from a
numerous multidirectional information resources.</p>
      <p>
        The emergence of information retrieval systems (IRS) is a natural reaction of mankind to solve the
Big Data problem [
        <xref ref-type="bibr" rid="ref1">1</xref>
        ]. But IRS are not responsible for the data reliability and relevance. Its development led
to the inability of a user to master the found amount of data.
      </p>
      <p>IRS filling is conducted by agent technologies - technologies using special autonomous programs for
solving various problems.</p>
    </sec>
    <sec id="sec-2">
      <title>2. Urgency</title>
      <p>Nowadays in the development of intelligent information agent technologies, multi -agent
technologies have been established to solve the problems of regular automatic collection and processing
of structured and unstructured information in a certain thematic direction from various network sources.</p>
      <p>There are three defining properties of Big Data – volume, variety and velocity. Referring to the
amount of data, types of data and to the speed of data processing respectively. According to this 3Vs model,
the challenges of big data management result from the expansion of all three properties, rather than just the
volume alone - the sheer amount of data to be managed.</p>
      <p>
        An agent system is a system created by an intelligent agent and is one of the Big Data proble m
solving [
        <xref ref-type="bibr" rid="ref2 ref3">2, 3</xref>
        ]. An agent itself in information technology or a software agent is a computer program that
performs a task entrusted by a user and acts on behalf of that user.
      </p>
    </sec>
    <sec id="sec-3">
      <title>3. Modern agent technologies</title>
      <p>There are invested additional properties in the modern definition of an intellectual agent:
autonomy, sociality, reactivity and proactivity.</p>
      <p>Autonomy refers to the ability of an agent to act purposefully to achieve a result without external
control by systems. He has control over his actions and the state of internal variables.</p>
      <p>The sociability of an agent lies in the interaction with other agents to accomplish a common task, and
for consistency in resolving conflicts to reach a consensus.</p>
      <p>Reactivity is a reaction to receiving an event from the external environment and adjusting its
behavior according to the knowledge gained.</p>
      <p>Proactivity means a desire to achieve the goal, improving the characteristics of the internal state
nonstop.</p>
      <p>
        Multiagent system (MAS) is the system of autonomous intelligent agents united in a common
network and their activity is also aimed at solving the tasks set by the system [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ]. The use of MAS is one of
the key approaches to the tasks of reducing the time for processing information and reducing the cost of
cloud services, which allows autonomously, make decisions and allocate decentralized resources.
      </p>
      <p>Due to its multitasking and structure, multi-agent technologies are used in many different areas:
management of distributed and network enterprises, the educational process in systems of distance learning,
complex and multifunctional logistics, etc.</p>
      <p>
        A demonstrative example is the social network LinkedIn, which uses intelligent agent
technologies in recruiting to search for people based on their career information and research
activities.[
        <xref ref-type="bibr" rid="ref5">5</xref>
        ].
      </p>
      <p>Projects in the field of mass online education, like Coursera, use agent technologies to inform
users about events within a certain course.</p>
    </sec>
    <sec id="sec-4">
      <title>4. Agent acquisition of information specifics</title>
      <p>
        When working in a global network, agents have to interact with diverse types of infor mation
sources [
        <xref ref-type="bibr" rid="ref6">6</xref>
        ]. The formation of knowledge in the global network is disordered; therefore, the type of data
contained can classify information sources.
      </p>
      <p>Structured data: Data containing a defined data type, format, and structure (that is, transaction
data, online analytical processing data cubes, traditional RDBMS, and even simple spreadsheets).</p>
      <p>Semi-structured data: Textual data files with a discernible pattern that enabl es parsing (such as
XML, data files that are self-describing and defined by an XML schema).</p>
      <p>Quasi-structured data: Textual data with erratic data formats that can be formatted with effort,
tools, and time (for instance, web clickstream data that may contai n inconsistencies in data values and
formats).</p>
      <p>Unstructured data has no inherent structure, may include text documents, PDFs, images, and
video.</p>
      <p>The degree of structuredness of the collected data is directly determined by the variety and
intellectual abilities of the agents. When collecting a news article from an information portal, an agent
can collect the whole source code along with ad units and other irrelevant information, and can
highlight the components of a publication and collect data in its pure form, e.g.</p>
      <p>The development of an agent for each information resource has some difficulties because of the
lack of strictly regulated standards for constructing information resources and because of high
flexibility of presenting information using html code. Therefore, different algorithms and standards are
created for the placement of information for its correct identification: HTML 5.2, Selectors Level 3,
ITS Version 2.0, etc.</p>
      <p>Time passed by, and a problem arose in the use of agent technologies because the collection of
data may occur against the desire of the internet resources owners, thereby the agent, without realizing
it, could infringe copyrights.</p>
      <p>In U.S. court practice; there have been cases where information portal owners have brought
cases against collecting agents. For example, in case №17-cv-03301-EMC of United States District
Court Northern District of California hiQ Labs inc. indicted LinkedIn for the fact that the social
network has built protection against the collection of publicly available p rofile information in reference
from United States Antitrust law. This case was initiated in response to the hiQ company's early
accusation of illegitimate automated collection of LinkedIn profile information in reference from
Computer Fraud and Abuse Act. Subsequently, a case judge Edward M. Chen obliged LinkedIn to
remove the previously established protection on the collection of publicly available information by
autonomous intelligent agents.</p>
      <p>Afterwards the use of agents to collect information led to the practice of implementing the
Robots Exclusion Standard (robots.txt, 1994). This is a file with a list of access restrictions for robots
to content of an http server. The file should have a path relative to the site name /robots.txt. It reflects
the availability of information to be collected by an agent.</p>
    </sec>
    <sec id="sec-5">
      <title>5. IT environment research</title>
      <p>
        To build multi-agent systems, it is necessary, among other things, to conduct a study of the
information environment[
        <xref ref-type="bibr" rid="ref7">7</xref>
        ]. It allows user to determine the structure of the informat ion source, the
object of interest and the list of necessary software. Accordingly, three stages can be distinguished.
      </p>
      <sec id="sec-5-1">
        <title>5.1 Reconnaissance</title>
        <p>At this stage, it is necessary to determine the available information about the object of interest
and the legitimacy of its collection. It is also important to identify the available attributes of the object
of interest in order to build the corresponding data model and relationships with other objects of the
information source.</p>
        <p>For example, in social networks where the person is the object of interest, communities, audio
and video recordings can play the role of objects associated with it. In such cases, it would be
necessary not only to collect information about objects of interest, but also to take into account the
relationship with other objects, since in the future they can act as characteristics during intellectual
processing.</p>
      </sec>
      <sec id="sec-5-2">
        <title>5.2 Crawling</title>
      </sec>
      <sec id="sec-5-3">
        <title>5.3 Configuration</title>
        <p>The stage includes the study of mechanisms for placing information about objects of interest into
an information source and the definition of an agent behavior algorithm for collecting information with
maximum completeness. Any site is a set of interconnected web pages and its structure can be
represented in the form of a hierarchical model tree, where the top of the tree is the ini tial page of the
site and pages with the necessary information is the basis of the tree.</p>
        <p>The purpose of the configuration stage is to list third-party software solutions to expand the
functions of the agent. In particular, the Competitive Systems Analysis Department at MEPhI uses
Python 3.7 as the programming language: on which the agent is implemented.</p>
        <p>The essence of this method of collecting information is to obtain hypertext documents in html
format using the http / https protocol and further identify the necessary data from the received
documents. To develop such an agent modules are required, that allow you to access the information
resource using the http / https protocol and a module for recognizing the contents of an html document.
At this stage, the source of information considered fully explored.</p>
      </sec>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          [1]
          <string-name>
            <surname>Garcia</surname>
            ,
            <given-names>A.J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Toril</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Oliver</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Luna-Ramirez</surname>
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Garcia</surname>
            <given-names>R</given-names>
          </string-name>
          .
          <article-title>Big Data Analytics for Automated QoE</article-title>
          Management in Mobile Networks // IEEE Communications Magazine Volume
          <volume>57</volume>
          ,
          <string-name>
            <surname>Issue</surname>
            <given-names>8</given-names>
          </string-name>
          ,
          <string-name>
            <surname>August</surname>
            <given-names>2019</given-names>
          </string-name>
          , Pages
          <fpage>91</fpage>
          -
          <lpage>97</lpage>
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          [2]
          <string-name>
            <surname>Xiao</surname>
            <given-names>Y.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Liu</surname>
            <given-names>Q</given-names>
          </string-name>
          .
          <source>Application of Big Data Processing Method in intelligent Manufacturing // Proceedings of 2019 IEEE International Conference on Mechatronics and Automation</source>
          ,
          <string-name>
            <surname>ICMA</surname>
          </string-name>
          <year>2019</year>
          ,
          <year>Pages 1895</year>
          -1900
          <source>16th IEEE International Conference on Mechatronics and Automation</source>
          ,
          <string-name>
            <surname>ICMA</surname>
          </string-name>
          <year>2019</year>
          ;
          <article-title>Tianjin; China; 4 August 2019 - 7 August 2019</article-title>
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          [3]
          <string-name>
            <surname>Kulik</surname>
            <given-names>S.D.</given-names>
          </string-name>
          <article-title>Model for evaluating the effectiveness of search operations //</article-title>
          <source>Journal of ICT Research and Applications</source>
          Volume
          <volume>9</volume>
          ,
          <string-name>
            <surname>Issue</surname>
            <given-names>2</given-names>
          </string-name>
          ,
          <year>2015</year>
          , Pages
          <fpage>177</fpage>
          -
          <lpage>196</lpage>
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          [4]
          <string-name>
            <surname>Ananieva</surname>
            ,
            <given-names>A.G.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Artamonov</surname>
            ,
            <given-names>A.A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Galin</surname>
            ,
            <given-names>I.U.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Tretyakov</surname>
            ,
            <given-names>E.S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kshnyakov</surname>
            ,
            <given-names>D.O.</given-names>
          </string-name>
          <article-title>Algorithmization of search operations in multiagent information -</article-title>
          analytical
          <source>systems // Journal of Theoretical and Applied Information Technology Volume 81, Issue</source>
          <volume>1</volume>
          ,
          <issue>10</issue>
          <year>November 2015</year>
          , Pages
          <fpage>11</fpage>
          -
          <lpage>17</lpage>
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          [5]
          <string-name>
            <surname>Artamonov</surname>
            <given-names>A.A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ionkina</surname>
            <given-names>K.V.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kirichenko</surname>
            <given-names>A.V.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Lopatina</surname>
            <given-names>E.O.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Tretyakov</surname>
            <given-names>E.S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Cherkasskiy</surname>
            <given-names>A.I.</given-names>
          </string-name>
          <article-title>Agent-based search in social networks //</article-title>
          <source>International Journal of Civil Engineering and Technology</source>
          Volume
          <volume>9</volume>
          ,
          <string-name>
            <surname>Issue</surname>
            <given-names>13</given-names>
          </string-name>
          ,
          <string-name>
            <surname>December</surname>
            <given-names>2018</given-names>
          </string-name>
          , Pages
          <fpage>28</fpage>
          -
          <lpage>35</lpage>
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          [6]
          <string-name>
            <surname>Onykiy</surname>
            <given-names>B.N.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Artamonov</surname>
            <given-names>A.A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Tretyakov</surname>
            <given-names>E.S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ionkina</surname>
            <given-names>K.V.</given-names>
          </string-name>
          <article-title>Visualization of large samples of unstructured information on the basis of specialized thesauruses</article-title>
          // Scientific Visualization Volume
          <volume>9</volume>
          ,
          <string-name>
            <surname>Issue</surname>
            <given-names>5</given-names>
          </string-name>
          ,
          <year>2017</year>
          , Pages
          <fpage>54</fpage>
          -
          <lpage>58</lpage>
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          [7]
          <string-name>
            <surname>Kulik</surname>
            <given-names>S.D.</given-names>
          </string-name>
          <article-title>Factographic information retrieval for semiconductor physics</article-title>
          , micro - And nanosystems // IOP Conference Series: Materials Science and Engineering Volume
          <volume>498</volume>
          ,
          <string-name>
            <surname>Issue</surname>
            <given-names>1</given-names>
          </string-name>
          , 16 April 2019
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>