<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta>
      <journal-title-group>
        <journal-title>December</journal-title>
      </journal-title-group>
      <issn pub-type="ppub">1613-0073</issn>
    </journal-meta>
    <article-meta>
      <title-group>
        <article-title>Understanding Tables in Financial Documents</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Yasutomo Kimura</string-name>
          <email>kimura@res.otaru-uc.ac.jp</email>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Eisaku Sato</string-name>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Kazuma Kadowaki</string-name>
          <email>kadowaki.kazuma@jri.co.jp</email>
          <xref ref-type="aff" rid="aff2">2</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Hokuto Ototake</string-name>
          <email>ototake@fukuoka-u.ac.jp</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Workshop</string-name>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Fukuoka University</institution>
          ,
          <addr-line>Fukuoka</addr-line>
          ,
          <country country="JP">Japan</country>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>Otaru University of Commerce</institution>
          ,
          <addr-line>Hokkaido</addr-line>
          ,
          <country country="JP">Japan</country>
        </aff>
        <aff id="aff2">
          <label>2</label>
          <institution>The Japan Research Institute</institution>
          ,
          <addr-line>Limited, Tokyo</addr-line>
          ,
          <country country="JP">Japan</country>
        </aff>
      </contrib-group>
      <pub-date>
        <year>2024</year>
      </pub-date>
      <volume>12</volume>
      <issue>2024</issue>
      <fpage>0000</fpage>
      <lpage>0003</lpage>
      <abstract>
        <p>This paper presents a framework for the “NTCIR-18 U4” and “SIG-FIN UFO-2024” shared tasks, which focus on tables within annual securities reports. Annual securities reports are critical documents that provide insights into a company's financial status and business performance. However, challenges remain in accurately and eficiently analyzing the data they contain. To address these issues, we propose two sub-tasks for the above shared tasks: Table Retrieval and Table QA tasks, which utilize datasets from TOPIX100 and TOPIX500 annual securities reports. Participants are tasked with developing systems (programs) that automatically process data for the two tasks and compete for top performance on a leaderboard. Accuracy scores and rankings are determined by submitting the task's output, in JSON format, to the leaderboard. Through these shared tasks, we aim to enhance the utility of annual securities reports and advance natural language processing technologies for financial data analysis.</p>
      </abstract>
      <kwd-group>
        <kwd>annual securities report</kwd>
        <kwd>shared task</kwd>
        <kwd>table retrieval</kwd>
        <kwd>table question-answering</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>1. Introduction</title>
      <p>In recent years, financial disclosures have become essential for investors seeking to make informed
decisions based on reliable corporate data. In Japan, listed companies are required to submit an annual
securities report, a statutory disclosure document that provides comprehensive information on business
operations, financial data, risk factors, corporate governance, and shareholder information. These
reports, accessible via the Electronic Disclosure for Investors’ NETwork (EDINET)1, serve as a critical
information source for investors aiming to compare companies efectively.</p>
      <p>These securities reports are structured in XBRL (eXtensible Business Reporting Language), an
XMLbased format designed to standardize and facilitate the production, distribution, and reuse of financial
information. By incorporating “taxonomies” that define the structure and meaning of data, XBRL
enables automated processing, potentially streamlining financial analysis.</p>
      <p>However, practical challenges arise due to the presence of untagged data and the existence of
unique taxonomies created by diferent report submitters, complicating the identification of comparable
elements across reports.</p>
      <p>To this end, we propose two tasks that aim to facilitate cross-company comparisons by focusing on
the tables and text within annual securities reports. The first task is the NTCIR-18 U4 task, adopted
by Japan’s National Institute of Informatics (NII) as part of NTCIR-182. The second is the SIG-FIN
UFO-2024 task, organized by the Financial Informatics Study Group (SIG-FIN) under the Japanese
Society for Artificial Intelligence (JSAI). The former focuses on TOPIX100 annual securities reports
submitted between April 1, 2020, and March 31, 2021, while the latter focuses on TOPIX500 annual
securities reports submitted between July 1, 2023, and June 30, 2024.
∗Corresponding author.</p>
      <p>CEUR</p>
      <p>ceur-ws.org</p>
      <p>
        We organized these shared tasks in collaboration with the NII Testbeds and Community for
Information Access Research (NTCIR) [
        <xref ref-type="bibr" rid="ref1">1</xref>
        ], which specializes in information retrieval, and the SIG-FIN group of
the JSAI, which focuses on financial technology. Through these initiatives, we aim to attract researchers
and practitioners interested in these fields and contribute to further advancing technologies at the
intersection of finance and information retrieval.
      </p>
      <p>In each shared task, we conducted two sub-tasks: Table Retrieval, which involves searching for tables,
and Table Question Answering (Table QA), which involves answering questions by identifying the
target cells within the tables, as illustrated in Figure 1. We designed each sub-task and constructed
datasets for each task.</p>
      <p>The contributions of this study are as follows:
• Design of two tasks, Table Retrieval and Table QA, targeting securities reports.
• Construction of datasets for Table Retrieval and Table QA, and their release on GitHub3.
• Organization of the NTCIR-18 U4 task and the SIG-FIN UFO-2024 task.</p>
    </sec>
    <sec id="sec-2">
      <title>2. Related Work</title>
      <sec id="sec-2-1">
        <title>2.1. Research on Tables</title>
        <p>
          A table is a data format with a two-dimensional structure used to organize and manage knowledge
or information, and it is widely utilized in various contexts. However, not all tables have a highly
structured database format; they are often represented as semi-structured data. Furthermore, the data
contained in table cells is not limited to numerical values; it frequently includes strings and other
non-numerical data. Numerous methods have been proposed to accommodate such diverse table data.
Zhang and Balog [
          <xref ref-type="bibr" rid="ref2">2</xref>
          ] surveyed on tables on the web and classified approaches to accessing table data
into six main categories.
        </p>
        <p>1. Table extraction
2. Table interpretation
3. Table search
4. Table question answering
5. Knowledge base augmentation
6. Table augmentation
1https://disclosure2.edinet-fsa.go.jp/
2https://research.nii.ac.jp/ntcir/ntcir-18/index-en.html
3https://github.com/nlp-for-japanese-securities-reports/ntcir18-u4,
https://github.com/nlp-for-japanese-securities-reports/ufo-2024</p>
        <p>
          In addition, table-related tasks include table fact verification [
          <xref ref-type="bibr" rid="ref3 ref4">3, 4</xref>
          ], table detection (searching for tables
within documents) [
          <xref ref-type="bibr" rid="ref5 ref6">5, 6</xref>
          ], spreadsheet manipulation [
          <xref ref-type="bibr" rid="ref7">7</xref>
          ], column type annotation [
          <xref ref-type="bibr" rid="ref8">8</xref>
          ], and entity linking
(linking to knowledge bases) [
          <xref ref-type="bibr" rid="ref10 ref9">9, 10</xref>
          ]. These tasks are critical in information retrieval and data analysis
based on table data, and they are particularly anticipated in fields where handling large-scale data
and automation are required. Recently, approaches utilizing large language models (LLMs) and visual
language models (VLMs) have been increasing, and research on learning methods, prompt engineering,
and agents is also gaining attention [
          <xref ref-type="bibr" rid="ref11">11</xref>
          ]. Our shared tasks (NTCIR-18 U4 and the SIG-FIN UFO-2024)
are related to table search, table detection, and table question answering (Table QA).
        </p>
      </sec>
      <sec id="sec-2-2">
        <title>2.2. Table Retrieval and Table QA</title>
      </sec>
      <sec id="sec-2-3">
        <title>2.3. Tables in the Financial Domain</title>
        <p>Hybrid data, which includes both tables and text, such as in financial reports, is quite prevalent in the
real world [16]. Zhu et al. constructed a question-answering benchmark dataset focused on the hybrid
content of tabular and textual data in the financial domain [ 15].</p>
        <p>Pan et al. proposed CLTR, an architecture for end-to-end table retrieval at the cell level [17]. While
CLTR can be applied to open-domain datasets, including finance and healthcare ones, its performance
specifically within the financial domain has not been clarified, nor does it target the Japanese language.</p>
        <p>One of the tasks focused on Japanese financial table structure analysis is the UFO (Understanding
of non-Financial Objects in Financial Reports) task [18]. The UFO task aims to extract structured
information from tables and text found in annual securities reports and consists of two sub-tasks: the
Table Data Extraction (TDE) task and the Text-to-Table Relationship Extraction (TTRE) task. The
TDE task classifies cells in tables into four categories with the goal of identifying the type of each
cell: metadata, header, attribute, and data [19]. The main focus of TDE was on cell classification, and
additional processing to enable inter-company comparisons remained unexplored.</p>
      </sec>
    </sec>
    <sec id="sec-3">
      <title>3. NTCIR-18 U4 and SIG-FIN UFO-2024 Tasks</title>
      <p>Both the NTCIR-18 U4 and SIG-FIN UFO-2024 tasks consist of two sub-tasks: the Table Retrieval
task, which involves searching for tables, and the Table Question Answering (Table QA) task, which
involves answering questions by identifying the target cells within the tables [20]. Figure 1 illustrates
the concept of these two sub-tasks.</p>
      <sec id="sec-3-1">
        <title>3.1. Table Retrieval: Table Search Task</title>
        <p>Table Retrieval is a task that involves searching for a “table” containing the values that answer a given
question from the tables included in a company’s annual securities report. On average, a company’s
annual securities report contains 221.9 tables [21], and it is necessary to identify the specific table that
contains the answer to the question needs to be identified. The input, output, and evaluation criteria
for this task are as follows.</p>
        <p>For the input, HTML files downloaded from EDINET are used. Each table element ( &lt;table&gt;) in the
HTML files is assigned a unique Table ID, and when outputting the table that answers the question,
this Table ID is used as the output. In the output example above, the Table ID “S100ISF1-0101010-tab2”
refers to the second table in the “S100ISF1-0101010.html” file, with “-tab{ table number }” appended to
the file name.</p>
        <p>The metric used for evaluation is accuracy, which is calculated by dividing the number of correct
outputs by the total number of inputs in the test dataset.</p>
        <p>We evaluated a few baseline methods using our validation datasets, which contain 3,131 and 1,533
questions for NTCIR-18 U4 and SIG-FIN UFO-2024, respectively. The results are shown in Table 1. For
the NTCIR-18 U4 task, the highest accuracy of 0.2111 was achieved by using the text-embedding-3-small
model to create embeddings based on Cell Text. Similarly, for the SIG-FIN UFO-2024 task, a top accuracy
of 0.1937 was obtained using the text-embedding-3-large model for Cell Text embeddings.</p>
      </sec>
      <sec id="sec-3-2">
        <title>3.2. Table QA: Table Question Answering Task</title>
        <p>The Table QA task, given a target table, identifies the “value” that answers the question. To accurately
determine the answer, complex tables included in annual securities reports need to be handled [22].
The input, output, and evaluation criteria for this task are as follows:</p>
        <sec id="sec-3-2-1">
          <title>Output</title>
          <p>Evaluation</p>
        </sec>
        <sec id="sec-3-2-2">
          <title>1. Question</title>
          <p>2. Target table (Table ID)
3. HTML file of the annual securities report
Value, Cell ID</p>
          <p>Accuracy (value), Accuracy (cell ID)</p>
          <p>For input data, HTML files and Table IDs are provided, allowing the system to extract the range
enclosed by the &lt;table&gt; tag from the HTML file. Additionally, if necessary, the system can utilize the
surrounding context of the table.</p>
          <p>Similar to Table IDs, each cell (&lt;th&gt; and &lt;td&gt; tags) within the table in the HTML file is assigned a
unique Cell ID. When outputting the value corresponding to the answer to the given question, this Cell
ID is used. In the output example above, the Cell ID is the Table ID of the table containing the cell, with
“-r{row number }c{column number }” appended, so the Cell ID “S100ISF1-0101010-tab2-r8c1” refers to the
cell in the 8th row and 1st column of the table “100ISF1-0101010-tab2”.</p>
          <p>For evaluation, similar to the Table Retrieval task, accuracy is calculated by dividing the number of
correct outputs by the total number of inputs in the test dataset. However, discrepancies between the
value contained in the HTML cell and the expected answer are frequently observed. For example, if the
expected answer is “4448000000”, the corresponding cell in the HTML might contain the string “4,448”,
while another cell, such as in the top right or column name, might indicate “(in millions of yen)”. In this
case, the system answering the task must reference both cells to generate the answer “4,448 million
yen”. While this is equivalent to the correct answer, to compare it accurately, the system must replace
the string “million yen” with “000000” and remove the comma.</p>
          <p>Due to this, in this task, both the response and the correct answer are normalized before
calculating accuracy. The normalization specification was continually revised during the “dry run” period,
considering feedback from participants.</p>
          <p>We evaluated a few baseline methods using our validation datasets, which contain 3,132 and 1,534
questions for NTCIR-18 U4 and SIG-FIN UFO-2024, respectively. These baseline methods involved
converting the target table into text format and inputting it, along with the question, into an LLM to
generate the desired values. The results are shown in Table 2. For the NTCIR-18 U4 task, the highest
accuracy of 0.7471 was achieved by using the Claude 3 Opus model. Similarly, for the SIG-FIN UFO-2024
task, a top accuracy of 0.5750 was obtained using the GPT-4o model.</p>
        </sec>
      </sec>
    </sec>
    <sec id="sec-4">
      <title>4. Dataset</title>
      <sec id="sec-4-1">
        <title>4.1. Securities Reports Used in Our Dataset</title>
        <p>The NTCIR-18 U4 task focuses on analyzing securities reports from companies in the TOPIX100 index.
The dataset consists of securities reports from companies that are part of the TOPIX100, submitted
between April 1, 2020, and March 31, 2021.</p>
        <p>The SIG-FIN UFO-2024 task, on the other hand, focuses on analyzing securities reports from companies
in the TOPIX500 index. The annual securities reports used in this task are drawn from the TOPIX500,
which represent publicly listed companies with high market capitalization and liquidity. For this task,
we target the annual securities reports of 497 companies4 constituting the TOPIX 500 as of April 30,
4Despite the name TOPIX500, as of 30 April 2024, it only includes 497 companies. To be more precise, the TOPIX500 is a</p>
        <p>Baseline methods</p>
        <p>Accuracy (value)
2024. The dataset includes 494 financial statements submitted to EDINET between July 1, 2023, and
June 30, 2024.</p>
        <p>To account for diferences in the structure of annual securities reports across industries, we ensure
that the dataset is balanced by industry. The annual securities reports are distributed across the training,
validation, and test sets with minimal industry bias. Specifically, we use the ten major categories
from the Tokyo Stock Exchange’s 33 industry classifications (service industry, transportation and
communications, finance and insurance, construction, mining, commerce, fisheries, agriculture and
forestry, manufacturing, electricity and gas, and real estate). The data is divided such that the ratio of
train:validation:test is approximately 6:1:3 within each industry category. This results in 289 companies’
reports being used for training, 52 for validation, and 153 for testing.</p>
        <p>We retrieve the financial data using the EDINET API v2, utilizing the XBRL, HTML, and CSV files
available through the API. The XBRL files contain tabular data, such as taxonomies and instances,
referred to as “XBRL information” below, which is also embedded in the corresponding HTML files.
The CSV files, referred to below as “annual securities report CSVs,” provide a more accessible format for
the XBRL data for easier handling in the study.</p>
      </sec>
      <sec id="sec-4-2">
        <title>4.2. Question Creation</title>
        <p>Questions are created using annual securities report CSVs and question templates. In the annual
securities report CSV, each row represents data, and each column shows the corresponding XBRL
information (element ID, item name, context ID, relevant year, consolidated or individual, period or
point in time, unit ID, unit, value). Among this XBRL information, the element ID and context ID are
crucial for data extraction. The element ID indicates what the data represents, but it is not unique
within a single annual securities report. Therefore, combining the element ID with the context ID,
which represents the period and dimension, enables data within a report to be uniquely identified and
the desired information to be extracted. Thus, the question must include both element ID and context
ID.</p>
        <p>On the basis of this, the initial version of the question is created as follows:
Question (Initial Version)</p>
        <sec id="sec-4-2-1">
          <title>What is the value of “{Element ID}” for {Company Name} in {Context ID}?</title>
          <p>However, if the element ID and context ID are used as they appear in the annual securities report
CSV, the question will not be meaningful in Japanese as they are simply IDs. Therefore, the context ID
is represented using the relative year, consolidated or individual, and period or point in time, while the
element ID is expressed as the item name. The final version of the question is defined as follows:
stock price index composed of the TOPIX Core30, TOPIX Large70 and TOPIX Mid400, but of these, only 397 companies are
included in the TOPIX Mid400.</p>
          <p>Question (Detailed Version)
What is the value of “{Item Name}” in the {Year } {Period or Point in Time} {Consolidated or Individual (optional)}
annual securities report of {Company Name} for {Member Element (optional)}?
The explanations for each part are as follows:
• Year: Calculated based on the basis of the relevant year, and the string is included in the question.
• Period or Point in Time: If it is a point in time, the word “point” is added right after the year.
• Consolidated or Individual: If it is consolidated or individual, the corresponding string is included;
otherwise, “annual securities report of” is omitted.
• Member Element: If the context ID contains a member element, the string is included. This
element is not translated into Japanese to ensure uniqueness and is used as-is from the annual
securities report CSV (ensuring uniqueness is a future challenge).</p>
          <p>• Item Name: This is essentially the Japanese translation of the element ID, so the string is included.</p>
        </sec>
        <sec id="sec-4-2-2">
          <title>An example of a question created using the template is as follows: Example Created with Question Template</title>
          <p>What is the value of “Building (net amount)” in the 2020 individual annual securities report of Daiwa House
Industry Co., Ltd. for NonConsolidatedMember?</p>
          <p>When creating questions for the SIG-FIN UFO-2024 dataset, we also performed data sampling to
avoid generating too many similar questions. For data sampling, a unique list of item names is created
for each company, and random sampling is performed so that 1/10th of the entire dataset is selected.
Additionally, the number of samples per item name is adjusted on the basis of the number of data entries
for each item name5.</p>
          <p>As a result of these procedures, we constructed the NTCIR-18 U4 dataset consisting of 32,587 entries
for the Table Retrieval task and 32,589 entries for the Table QA task, and the SIG-FIN UFO-2024 dataset
consisting of 14,410 entries for the Table Retrieval task and 14,412 entries for the Table QA task, as
shown in Table 36.</p>
        </sec>
      </sec>
    </sec>
    <sec id="sec-5">
      <title>5. Schedule</title>
      <p>To encourage broad participation from those interested in finance, the organizers have introduced two
complementary tasks: the NTCIR-18 U4 and the SIG-FIN UFO-2024. The SIG-FIN community includes
researchers and practitioners actively engaged in finance, while NTCIR attracts participants interested
in shared tasks, especially those with a focus on information retrieval and natural language processing.
5If there is only one data entry for an item name, one sample is taken; if there are two to five, two samples are taken; if there
are six or more, three samples are randomly selected. Data with submitter-specific taxonomies that do not include an item
name in the annual securities report CSV, or data not from tables (i.e., data not in HTML’s &lt;td&gt; tags) are excluded.
6These are the breakdowns of the initial datasets used for the Dry Run. The datasets used in the Formal Run phase have been
modified to fix a couple of issues, and as a result, they contain a slightly diferent number of entries.</p>
      <p>By running similar tasks across these two distinct communities, we aim to foster interaction among
participants with diverse expertise and perspectives on finance, creating an opportunity for knowledge
exchange and collaboration. We look forward to welcoming a diverse group of participants to build a
comprehensive and impactful competition.</p>
      <p>The schedule for each task is outlined in Table 4, showing the parallel timelines and key phases for
both NTCIR-18 U4 and SIG-FIN UFO-2024. As illustrated in the table, both tasks share similar phases
such as a dry run, formal run, and evaluation period, which will allow participants to apply and refine
their approaches across both tasks seamlessly. This alignment ensures that participants will have the
opportunity to benefit from complementary insights across the two tasks and fosters collaborative
learning within the community.
5.1. NTCIR-18 U4 Task
The schedule for the NTCIR-18 U4 task is as follows: The dataset for the NTCIR-18 U4 shared task
was released in July 2024, followed by an online briefing session on July 20, 2024, where participants
received essential information about the task. The dry run phase ran from July 2024 to October 31, 2024,
during which participants worked on the dataset and refined their methods. Any issues identified in
the dataset during this phase were addressed and resolved to ensure a smooth formal run. The formal
run phase is scheduled from November 1, 2024, to December 28, 2024. Throughout the NTCIR-18 U4
task, a leaderboard will be used to provide participants with real-time feedback on their performance.
Similar to the SIG-FIN UFO-2024 task, the NTCIR-18 U4 leaderboard will display a Public score on the
basis of a subset of the test data during the task period, allowing participants to gauge their progress.
Evaluation results and final rankings are scheduled to be returned to participants on February 1, 2025,
along with a partial publication of the task overview paper summarizing key outcomes.
5.2. SIG-FIN UFO-2024 Task
The schedule for the SIG-FIN UFO-2024 task is as follows: The dataset for this shared task was released
on August 15, 2024, with the dry run phase extending until October 31, 2024. During this phase,
participants worked on developing their methods using the dataset, and any data issues identified
during this period were addressed and corrected. The formal run phase is scheduled from November 1,
2024, to December 28, 2024.</p>
      <p>The shared task ranking will be determined on the basis of the evaluation method used in Kaggle7,
incorporating both Public and Private scores. Throughout the shared task, the Public score (calculated
from a subset of the test data) will be displayed on the leaderboard. After the shared task concludes, the
Private score (evaluated on the remaining portion of the test data) will be calculated. The final results,
based on the Private score, are scheduled to be announced at the 34th SIG-FIN in March 2025.</p>
    </sec>
    <sec id="sec-6">
      <title>6. Conclusion</title>
      <p>This paper proposed a framework for two shared tasks, NTCIR-18 U4 and SIG-FIN UFO-2024, which
focus on tables within annual securities reports. In these shared tasks, two sub-tasks are conducted:
Table Retrieval and Table Question Answering (Table QA), which target the annual securities reports of
companies belonging to the TOPIX 100 or TOPIX 500 indexes.</p>
    </sec>
    <sec id="sec-7">
      <title>Acknowledgments</title>
      <p>This research was supported by JSPS KAKENHI Grant Number 21H03769. We would also like to express
our gratitude to everyone at the National Institute of Informatics, Japan, the NTCIR Co-chairs, the
members of the SIG-FIN Research Group, and our corporate sponsor, Preferred Networks, Inc., for their
valuable cooperation in planning these shared tasks.
Proceedings of the 5th Workshop on NLP for Conversational AI (NLP4ConvAI 2023), 2023, pp.
59–70. doi:10.18653/v1/2023.nlp4convai-1.6.
[13] N. Jin, J. Siebert, D. Li, Q. Chen, A survey on table question answering: Recent advances, in:
Knowledge Graph and Semantic Computing: Knowledge Graph Empowers the Digital Economy,
Springer Nature Singapore, 2022, pp. 174–186. doi:10.1007/978-981-19-7596-7_14.
[14] Z. Chen, W. Chen, C. Smiley, S. Shah, I. Borova, D. Langdon, R. Moussa, M. Beane, T.-H. Huang,
B. Routledge, W. Y. Wang, FinQA: A dataset of numerical reasoning over financial data, in:
Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing, 2021,
pp. 3697–3711. doi:10.18653/v1/2021.emnlp-main.300.
[15] F. Zhu, W. Lei, Y. Huang, C. Wang, S. Zhang, J. Lv, F. Feng, T.-S. Chua, TAT-QA: A question
answering benchmark on a hybrid of tabular and textual content in finance, in: Proceedings of the
59th Annual Meeting of the Association for Computational Linguistics and the 11th International
Joint Conference on Natural Language Processing (Volume 1: Long Papers), 2021, pp. 3277–3287.
doi:10.18653/v1/2021.acl-long.254.
[16] N. Romanus Myrberg, S. Danielsson, Question-Answering in the Financial Domain, Master’s thesis,
Department of Computer Science, Lund University, 2023. URL: http://lup.lub.lu.se/student-papers/
record/9126226.
[17] F. Pan, M. Canim, M. Glass, A. Gliozzo, P. Fox, CLTR: An end-to-end, transformer-based system
for cell-level table retrieval and table question answering, in: Proceedings of the 59th Annual
Meeting of the Association for Computational Linguistics, 2021, pp. 202–209. doi:10.18653/v1/
2021.acl-demo.24.
[18] Y. Kimura, T. Kondo, K. Kadowaki, M. Kato, UFO: Proposal for an information extraction task for
tables in annual securities reports (in Japanese), in: JSAI Technical Report, Type 2 SIG, volume
FIN029, The Japanese Society for Artificial Intelligence, 2022, pp. 32–38. doi: 10.11517/jsaisigtwo.
2022.FIN-029_32.
[19] K. Kadowaki, Y. Kimura, M. Kato, T. Kondo, H. Ototake, Toward the construction of a dataset
for table structure analysis for annual securities reports (in Japanese), in: JSAI Technical Report,
Type 2 SIG, volume FIN-030, The Japanese Society for Artificial Intelligence, 2023, pp. 100–105.
doi:10.11517/jsaisigtwo.2023.FIN-030_100.
[20] E. Sato, Y. Kimura, Creating a question-answering dataset for securities reports and evaluation
of the method using LLM (in Japanese), in: IEICE Technical Report, volume 124, no. 173, The
Institute of Electronics, Information and Communication Engineers, 2024, pp. 93–98. URL: https:
//www.ieice.org/publications/search/summary.php?id=132450&amp;tbl=ken&amp;lang=jp.
[21] E. Sato, Y. Kaji, Y. Kimura, Analysis of tabular data contained in the TOPIX100 annual
securities report (in Japanese), The 21st Forum on Information Technology (FIT2022) E-021 (2022).
URL: https://www.ieice.org/publications/conferences/summary.php?id=FIT0000015362&amp;ConfCd=
F&amp;conf_type=F&amp;year=2022.
[22] K. Okuyama, Y. Kimura, Analysis of machine-unreadable table structures in securities reports (in
Japanese), The 30th Annual Meeting of the Association for Natural Language Processing (NLP2024)
P3-20 (2024). URL: https://www.anlp.jp/proceedings/annual_meeting/2024/pdf_dir/P3-20.pdf.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          [1]
          <string-name>
            <given-names>T.</given-names>
            <surname>Sakai</surname>
          </string-name>
          ,
          <string-name>
            <given-names>D. W.</given-names>
            <surname>Oard</surname>
          </string-name>
          ,
          <string-name>
            <given-names>N.</given-names>
            <surname>Kando</surname>
          </string-name>
          ,
          <source>Evaluating Information Retrieval and Access Tasks: NTCIR's Legacy of Research Impact, The Information Retrieval Series</source>
          , Springer Nature,
          <year>2021</year>
          . doi:
          <volume>10</volume>
          .1007/
          <fpage>978</fpage>
          -981-15-5554-1.
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          [2]
          <string-name>
            <given-names>S.</given-names>
            <surname>Zhang</surname>
          </string-name>
          , K. Balog,
          <article-title>Web table extraction, retrieval, and augmentation: A survey</article-title>
          ,
          <source>in: ACM Transactions on Intelligent Systems and Technology (TIST)</source>
          , volume
          <volume>11</volume>
          , issue 2,
          <year>2020</year>
          , pp.
          <fpage>1</fpage>
          -
          <lpage>35</lpage>
          . doi:
          <volume>10</volume>
          .1145/3372117.
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          [3]
          <string-name>
            <given-names>W.</given-names>
            <surname>Chen</surname>
          </string-name>
          ,
          <string-name>
            <given-names>H.</given-names>
            <surname>Wang</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Chen</surname>
          </string-name>
          ,
          <string-name>
            <given-names>Y.</given-names>
            <surname>Zhang</surname>
          </string-name>
          ,
          <string-name>
            <given-names>H.</given-names>
            <surname>Wang</surname>
          </string-name>
          ,
          <string-name>
            <given-names>S.</given-names>
            <surname>Li</surname>
          </string-name>
          ,
          <string-name>
            <given-names>X.</given-names>
            <surname>Zhou</surname>
          </string-name>
          , W. Y. Wang,
          <article-title>TabFact: A largescale dataset for table-based fact verification</article-title>
          ,
          <source>in: 8th International Conference on Learning Representations, ICLR</source>
          <year>2020</year>
          ,
          <year>2020</year>
          . URL: https://openreview.net/forum?id=rkeJRhNYDH.
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          [4]
          <string-name>
            <given-names>F.</given-names>
            <surname>Wang</surname>
          </string-name>
          ,
          <string-name>
            <given-names>K.</given-names>
            <surname>Sun</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Pujara</surname>
          </string-name>
          ,
          <string-name>
            <given-names>P.</given-names>
            <surname>Szekely</surname>
          </string-name>
          ,
          <string-name>
            <given-names>M.</given-names>
            <surname>Chen</surname>
          </string-name>
          ,
          <article-title>Table-based fact verification with salience-aware learning</article-title>
          ,
          <source>in: Findings of the Association for Computational Linguistics: EMNLP</source>
          <year>2021</year>
          ,
          <year>2021</year>
          , pp.
          <fpage>4025</fpage>
          -
          <lpage>4036</lpage>
          . doi:
          <volume>10</volume>
          .18653/v1/
          <year>2021</year>
          .findings-emnlp.
          <volume>338</volume>
          .
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          [5]
          <string-name>
            <given-names>L.</given-names>
            <surname>Chen</surname>
          </string-name>
          ,
          <string-name>
            <given-names>C.</given-names>
            <surname>Huang</surname>
          </string-name>
          ,
          <string-name>
            <given-names>X.</given-names>
            <surname>Zheng</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Lin</surname>
          </string-name>
          ,
          <string-name>
            <surname>X. Huang,</surname>
          </string-name>
          <article-title>TableVLM: Multi-modal pre-training for table structure recognition, in: Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics</article-title>
          (Volume
          <volume>1</volume>
          :
          <string-name>
            <surname>Long</surname>
            <given-names>Papers)</given-names>
          </string-name>
          ,
          <year>2023</year>
          , pp.
          <fpage>2437</fpage>
          -
          <lpage>2449</lpage>
          . doi:
          <volume>10</volume>
          .18653/v1/
          <year>2023</year>
          .
          <article-title>acl-long</article-title>
          .
          <volume>137</volume>
          .
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          [6]
          <string-name>
            <given-names>D.</given-names>
            <surname>Prasad</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Gadpal</surname>
          </string-name>
          ,
          <string-name>
            <given-names>K.</given-names>
            <surname>Kapadni</surname>
          </string-name>
          ,
          <string-name>
            <given-names>M.</given-names>
            <surname>Visave</surname>
          </string-name>
          , K. Sultanpure,
          <article-title>CascadeTabNet: An approach for end to end table detection and structure recognition from image-based documents</article-title>
          ,
          <source>in: 2020 IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW)</source>
          ,
          <year>2020</year>
          , pp.
          <fpage>2439</fpage>
          -
          <lpage>2447</lpage>
          . doi:
          <volume>10</volume>
          .1109/CVPRW50498.
          <year>2020</year>
          .
          <volume>00294</volume>
          .
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          [7]
          <string-name>
            <given-names>Z.</given-names>
            <surname>Ma</surname>
          </string-name>
          ,
          <string-name>
            <given-names>B.</given-names>
            <surname>Zhang</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Zhang</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Yu</surname>
          </string-name>
          ,
          <string-name>
            <given-names>X.</given-names>
            <surname>Zhang</surname>
          </string-name>
          ,
          <string-name>
            <given-names>X.</given-names>
            <surname>Zhang</surname>
          </string-name>
          ,
          <string-name>
            <given-names>S.</given-names>
            <surname>Luo</surname>
          </string-name>
          ,
          <string-name>
            <given-names>X.</given-names>
            <surname>Wang</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Tang</surname>
          </string-name>
          , SpreadsheetBench: Towards challenging real world spreadsheet manipulation,
          <year>2024</year>
          . doi:
          <volume>10</volume>
          .48550/arXiv.2406.14991.
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          [8]
          <string-name>
            <given-names>P.</given-names>
            <surname>Li</surname>
          </string-name>
          ,
          <string-name>
            <given-names>Y.</given-names>
            <surname>He</surname>
          </string-name>
          ,
          <string-name>
            <given-names>D.</given-names>
            <surname>Yashar</surname>
          </string-name>
          ,
          <string-name>
            <given-names>W.</given-names>
            <surname>Cui</surname>
          </string-name>
          ,
          <string-name>
            <given-names>S.</given-names>
            <surname>Ge</surname>
          </string-name>
          ,
          <string-name>
            <given-names>H.</given-names>
            <surname>Zhang</surname>
          </string-name>
          ,
          <string-name>
            <given-names>D. Rifinski</given-names>
            <surname>Fainman</surname>
          </string-name>
          ,
          <string-name>
            <given-names>D.</given-names>
            <surname>Zhang</surname>
          </string-name>
          , S. Chaudhuri,
          <article-title>TableGPT: Table fine-tuned GPT for diverse table tasks</article-title>
          ,
          <source>in: Proceedings of the ACM on Management of Data</source>
          , volume
          <volume>2</volume>
          , issue 3,
          <year>2024</year>
          , pp.
          <fpage>1</fpage>
          -
          <lpage>28</lpage>
          . doi:
          <volume>10</volume>
          .1145/3654979.
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          [9]
          <string-name>
            <given-names>X.</given-names>
            <surname>Deng</surname>
          </string-name>
          ,
          <string-name>
            <given-names>H.</given-names>
            <surname>Sun</surname>
          </string-name>
          ,
          <string-name>
            <given-names>A.</given-names>
            <surname>Lees</surname>
          </string-name>
          ,
          <string-name>
            <given-names>Y.</given-names>
            <surname>Wu</surname>
          </string-name>
          ,
          <string-name>
            <surname>C.</surname>
          </string-name>
          <article-title>Yu, TURL: table understanding through representation learning</article-title>
          ,
          <source>in: Proceedings of the VLDB Endowment</source>
          , volume
          <volume>14</volume>
          , issue 3,
          <year>2020</year>
          , p.
          <fpage>307</fpage>
          -
          <lpage>319</lpage>
          . doi:
          <volume>10</volume>
          .14778/ 3430915.3430921.
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          [10]
          <string-name>
            <given-names>T.</given-names>
            <surname>Zhang</surname>
          </string-name>
          ,
          <string-name>
            <given-names>X.</given-names>
            <surname>Yue</surname>
          </string-name>
          ,
          <string-name>
            <given-names>Y.</given-names>
            <surname>Li</surname>
          </string-name>
          ,
          <string-name>
            <given-names>H.</given-names>
            <surname>Sun</surname>
          </string-name>
          ,
          <article-title>TableLlama: Towards open large generalist models for tables, in: Proceedings of the 2024 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 1: Long Papers</article-title>
          ),
          <year>2024</year>
          , pp.
          <fpage>6024</fpage>
          -
          <lpage>6044</lpage>
          . doi:
          <volume>10</volume>
          .18653/v1/
          <year>2024</year>
          .
          <article-title>naacl-long</article-title>
          .
          <volume>335</volume>
          .
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          [11]
          <string-name>
            <given-names>W.</given-names>
            <surname>Lu</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Zhang</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            <surname>Fan</surname>
          </string-name>
          ,
          <string-name>
            <given-names>Z.</given-names>
            <surname>Fu</surname>
          </string-name>
          ,
          <string-name>
            <given-names>Y.</given-names>
            <surname>Chen</surname>
          </string-name>
          ,
          <string-name>
            <given-names>X.</given-names>
            <surname>Du</surname>
          </string-name>
          ,
          <article-title>Large language model for table processing: A survey</article-title>
          ,
          <year>2024</year>
          . doi:
          <volume>10</volume>
          .48550/arXiv.2402.05121.
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          [12]
          <string-name>
            <given-names>A. S.</given-names>
            <surname>Sundar</surname>
          </string-name>
          , L. Heck, cTBLS:
          <article-title>Augmenting large language models with conversational tables</article-title>
          , in:
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>