<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Typographic Sets: Labeled Set Elements with Font Attributes</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Richard Brath</string-name>
          <email>brathr@lsbu.ac.uk</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Ebad Banissi</string-name>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>London South Bank University</institution>
          ,
          <addr-line>London</addr-line>
          ,
          <country country="UK">U.K</country>
        </aff>
      </contrib-group>
      <abstract>
        <p>We show that many different set visualization techniques can be extended with the addition of labeled elements using font attributes. Elements labeled with font attributes can: uniquely identify elements; encode membership in ten sets; use size to indicate proportions among set relations; can scale to thousands on clearly labeled elements; and use intuitive mappings to facilitate decoding. The approach can be applied to many different set visualization layouts, including Venn and Euler diagrams, graphs, mosaic plots and cartograms.</p>
      </abstract>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>Introduction</title>
      <p>Alsallakh’s STAR report provides a broad collection of set visualizations, tasks and
applications. It identifies 26 set analysis tasks, of which 7 are element related (e.g.
A1: find elements that belong to a specific set) and 5 more are related to attributes on
elements (e.g. C1: find the attribute values of a certain element). The other 14 tasks are
about set and set relations, some of which are also about elements, such as summaries of
elements (e.g. B10: comparison of set intersection cardinalities) or element exclusivity
(e.g. B12: does one set contain more exclusive elements than another set).</p>
      <sec id="sec-1-1">
        <title>Review of Element Representations</title>
        <p>Given the importance of elements to set analysis, it is useful to consider how elements
are represented. This includes whether they uniquely identify the element and how they
convey multiple data attributes. In the STAR report there are 28 examples with elements.
Those elements are represented as:</p>
        <p>
          Dots. Simple dots are often used to represent one or two data attributes, such as
color or brightness to identify set membership. For example, TwitterVenn [
          <xref ref-type="bibr" rid="ref10">10</xref>
          ] (fig.
1 far left) uses dots to indicate search results in a Venn diagram.
        </p>
        <p>
          Labels. Labels can uniquely identify items. In all cases additional attributes were
not encoded in labels, but rather other visual attributes such as the background
color, lines connecting labels, etc. e.g. ComED [
          <xref ref-type="bibr" rid="ref22">22</xref>
          ] fig. 1 (2nd from left).
Glyphs. In 5 cases, glyphs (e.g. pies, bars, and icons) are used to encode attributes
such as set membership (e.g. fig . 1 (3rd from left) [
          <xref ref-type="bibr" rid="ref4">4</xref>
          ]). Sometimes glyphs are used
together with plain labels.
        </p>
        <p>
          Images. In 3 cases, images are used. In two of these the images supplement
otherwise undifferentiated labels (BubbleSets [
          <xref ref-type="bibr" rid="ref12">12</xref>
          ] and Vizster [
          <xref ref-type="bibr" rid="ref15">15</xref>
          ]).
        </p>
        <p>Shape. In one case (fig. 1 right) unique shapes of countries identify the individual
elements, assuming that the viewer has a reasonable geographic literacy.</p>
        <p>In no case do the element representations illustrated in the STAR report use more
than two visual attributes. For example EulerGlyphs uses hue and outline to encode
memberships. Although there is some use of labels and/or a visual attribute or two to
indicate elements in sets, the existing use suggest that much more could be done:
Identifiable Elements. Uniquely identifiable elements adds information and
context regarding the members of the sets which may be relevant to the task. Viewers
can use their existing knowledge regarding the specific elements to augment their
understanding of the sets.</p>
        <p>Multi-attribute Elements. Adding more data through different visual attributes can
help identify elements, membership in sets or other data.
2.2</p>
      </sec>
      <sec id="sec-1-2">
        <title>Glyphs in visualization</title>
        <p>While identifying individual elements with multiple data attributes per element may not
occur frequently in set visualization, other examples can be found more broadly in data
visualization.</p>
        <p>
          Borgo et al.’s state-of-the-art report on glyphs [
          <xref ref-type="bibr" rid="ref3">3</xref>
          ] characterizes the design space of
glyphs including common visualization attributes (e.g. [
          <xref ref-type="bibr" rid="ref18 ref2">2,18</xref>
          ]) such as size, color,
intensity, opacity, and shape; as well as semantic attributes such as text, symbols, icons,
pictograms [
          <xref ref-type="bibr" rid="ref9">9</xref>
          ]. However, the guidelines surveyed focus on traditional visual attributes
(color, shape, size, orientation, texture and opacity) although there is some discussion
regarding metaphoric pictograms. There is no broader discussion for specifically
identifying a large number of unique items.
        </p>
        <p>
          Brath [
          <xref ref-type="bibr" rid="ref5">5</xref>
          ] itemizes approaches for uniquely encoding a high number of categories in
glyphs as 1) icons; 2) geometric shapes, e.g. circles, stars; 3) textures, i.e. imagery
combining shape and color, e.g. logos and flags; and 4) text labels. Icons, shapes and textures
can be difficult to create for a large number of categories, without pre-existing libraries
of glyphs (e.g. logos, symbols, images), or automated techniques e.g.[
          <xref ref-type="bibr" rid="ref24">24</xref>
          ]. However,
abstract concepts (e.g. GDP, CPI) may be difficult to encode; some glyphs may be
ambiguous (e.g. Clarus the dog-cow); while other glyphs are difficult to add attributes (e.g.
a French flag rotated 90 degrees is confused with Netherlands’ flag).
3
        </p>
      </sec>
    </sec>
    <sec id="sec-2">
      <title>Text and font attributes</title>
      <p>Given the challenges of glyphs, text is an alternative that can be readily used to uniquely
identify hundreds to thousands of elements. In cartography, fonts have been used for
centuries to uniquely identify geographic features. Text offers some interesting
additional capabilities beyond glyphs which we will discuss in the following subsections.
3.1</p>
      <sec id="sec-2-1">
        <title>Small Type Sizes</title>
        <p>
          In paper-based cartography, minimum font sizes are defined as 3 or 4 point [
          <xref ref-type="bibr" rid="ref16 ref23">16,23</xref>
          ],
with guidelines recommending 5 or 6 point as a minimums (one point = 1/72 inch).
Similarly, charts in print have small minimum point sizes, e.g. 4 point [
          <xref ref-type="bibr" rid="ref8">8</xref>
          ].
        </p>
        <p>Historically, text in visualizations was limited due to low resolution displays (72-96
pixels per inch). More recent devices, including phones, tablets and PCs, have
significantly increased pixel densities. Current Apple guidelines recommend a minimum size
of 13 “CSS points” which renders physically at 3.1 points on an iPhone 6plus.</p>
        <p>Text in print environments can go even smaller: microtext is extremely small sized
text and originated as an anti-counterfeiting device for banknotes and official
government documents as well as archival texts. Microtext can be clear at extremely small
sizes, although the ability to decipher the text depends on the eyesight of the viewer.
3.2</p>
      </sec>
      <sec id="sec-2-2">
        <title>Even Type Density</title>
        <p>
          Text can be perceived as an even textural tone due to a unique feature of typography
referred to as color by typographers. A well-designed font has an even distribution of
text ink across a sequence of letters regardless if the letters are sparse (e.g. i,v) or dense
(e.g. m, e). In figure 2, the apparent shading behind the large word CANADA in the
center of the banknote can be perceived as text on close inspection. This even density
of type is used in Typographic Maps to indicate lines and regions with blocks of type.
Another unique feature of typography is the use of font-specific attributes, such as bold
and italic, to indicate additional data in text labels such as set membership. In 1920’s
Ordance Survey maps [
          <xref ref-type="bibr" rid="ref17">17</xref>
          ] (fig. 3) cities are indicated as follows:
        </p>
        <sec id="sec-2-2-1">
          <title>Text indicates the literal name of the city.</title>
          <p>Case differentiates between town (uppercase) vs. village (lowercase).</p>
          <p>Italics are used to indicate an administrative centre, i.e. a county town.
Font size is used to indicate population category.</p>
          <p>Font family indicates country: serif for U.K., slab-serif or serif variant for Scotland.</p>
          <p>In text visualization, additional visual attributes are used on text in less than half
of 249 peer-reviewed text visualizations at Text Visualization Browser (textvis.lnu.
se) as of 01/22/2016. 40 have no text, 103 have plain text and 106 use text with visual
attributes to encode additional data. Of the 106 using visual attributes, most use either
color or hue: typographic specific attributes such as bold, italic and case occur in only
15 examples, and typically only a single additional attribute (e.g. highlight with bold).</p>
          <p>
            Brath and Banissi (e.g. [
            <xref ref-type="bibr" rid="ref6">6</xref>
            ]) itemize font attributes based on cross-disciplinary
review, and identify the following font-specific attributes for encoding data:
          </p>
        </sec>
        <sec id="sec-2-2-2">
          <title>Alphanumeric glyphs (A,B,C,1,2,3) literally encode data.</title>
          <p>Symbols (e.g. ¥; 8; [; ~; \). Unlike alphanumerics, symbols are not orderable.
Weight may include variants such as black, bold, book, light or extra light.
Italic and Oblique are both sloped fonts but italics have different letterforms. Italics
may occur at different slope angles, including reverse (sometimes used on maps).
CASE includes UPPER, lower, Mixed and SMALL CAPS.</p>
          <p>Typeface indicates font family, e.g.: sans, blackletter, script, source, etc.
Underline also has many variants, e.g. w::a:v:e, .d.o.t., dash, double.</p>
          <p>Width may be adjusted with font variants (e.g. condensed, expanded), by scaling
type (not recommended by typographers), or adjusting inter-character spacing.
Baseline shift. Can used together with size change creates subscript and superscript.
“Paired delimiters” evoke enclosure by pairing shapes around text, e.g. ({},“ ”,* *)
Furthermore, these font attributes are distinct and can be used together in any
combination, for example: italic + CASE + UNDERLINE + BOLD + WIDE. However,
the authors do not provide any examples with more than 2-4 font attributes in prior work
nor application to sets.
4</p>
        </sec>
      </sec>
    </sec>
    <sec id="sec-3">
      <title>Using Text and Font-Attributes to Represent Elements in Sets</title>
      <p>There are many potential benefits to using text and font attributes in set visualizations:
1. Labels can uniquely depict elements. Furthermore, supplemental graphics, e.g.</p>
      <p>dots, image, etc. are not needed (e.g. ComEd in fig. 1).
2. Font Attributes can be used to encode set membership, thereby providing
clarification to some types of set visualizations where element membership may be
otherwise ambiguous (e.g. graphs, or maps as in fig. 3). Given the large number of
font attributes (10) plus traditional visual attributes (e.g. hue, intensity, orientation),
a large number of sets can potentially be encoded. This is a unique contribution.
3. Area Proportions can be represented with text-based set elements. At a
macrolevel, a block of text can be seen as a texture covering an area. At a micro-level the
individual text elements can be read (as in fig. 2). This is a unique contribution.
4. Layout Agnostic. Labeled, font-attribute elements can be used across a wide
variety of set visualization approaches, including Venn and Euler diagrams, mosaic
plots, graphs, maps and so on.</p>
      <p>The following examples show different set visualizations techniques extended with
labeled font-attribute elements.
4.1</p>
      <sec id="sec-3-1">
        <title>Typographic Venn Diagram of the U.S. Senate</title>
        <sec id="sec-3-1-1">
          <title>Text shows the name of each senator.</title>
          <p>Slope indicates the political party membership. Right-leaning text indicates
Republicans, while left-leaning text indicates Democrats. Independents are represented
with no leaning at all.</p>
          <p>Bold indicates senators who have served more than one term.</p>
          <p>Underline indicates senators who have a graduate or professional degree.</p>
          <p>Hue indicates gender: blue for male, magenta for female.</p>
          <p>Beyond the four sets depicted by the Venn diagram, additional data is encoded:</p>
        </sec>
        <sec id="sec-3-1-2">
          <title>Case indicates age. Those over 65 are indicated in uppercase.</title>
          <p>Font family indicates ethnicity. Most senators are plain Caucasians (in a sans serif
font), with a couple Latinos (in a curvy font), an Asian-American (in a serif) and a
couple African-Americans (in a rectangular font).</p>
          <p>At the level of individual elements, the names of individual senators are readable.
Set memberships can be seen, either by assessing the containment of elements relative
to the set outlines, or by the font attributes. For example, MAZIE HIRONO is a female
(purple), Democrat (left-leaning italic), over age 65 (all caps), first term senator (not
bold), with an advanced degree (underline), and is an Asian-American (serif). BERNIE
SANDERS is male (blue), independent (no italics), over age 65 (all caps), multi-term
senator (bold), with no advanced degree (no underline), and is Caucasian (plain sans
serif font).</p>
          <p>At a macro-level, the use of stacked text elements allows stacks to be visually
compared similar to bars in a bar chart. The viewer can attend to the stacks without regard
to the individual names. Many visual comparisons of quantities can be done at the level
of set relations. e.g.</p>
        </sec>
        <sec id="sec-3-1-3">
          <title>There are far more men than women senators. There are more Democratic women senators than Republican women senators. There are more Democratic women senators with advanced degrees than corresponding Republican women.</title>
          <p>There are no first-term Democratic women without an advanced degree.</p>
          <p>
            Instead of using an area-proportional Venn diagram, the use of stacked text elements
allows for the separation of the depiction of logical relations (i.e. the curved lines and
fills depicting each set) from the quantities of elements (i.e. stacked text). Issues with
attempting to algorithmically size areas of Venn outlines so that areas represent
quantities are easily side-stepped, e.g. [
            <xref ref-type="bibr" rid="ref26">26</xref>
            ]. Each stack can be ordered too: in this example
alphabetic order facilitates visual search within subsets. A similar example of the U.S.
House of Representatives is available online and also as an interactive demo.
4.2
          </p>
        </sec>
      </sec>
      <sec id="sec-3-2">
        <title>Typographic Mosaic Plot of Titanic Survivorship</title>
        <p>A larger scale example is the data repository encyclopedia-titanica.org, which
provides detailed biographies of 1308 passengers of the Titanic. However, one must
either search or browse through lists of names: a macro view of the passengers is not
available on this site.</p>
        <p>Figure 5 is a typographic mosaic plot of Titanic passengers indicating class
(horizontal bands 1,2,3), survivors (vertically: green survive, red death), and differentiates
between men (cranberry/forest green) vs. women and children (orange/chartreuse) in
horizontal slices. In a mosaic plot, areas represent quantities. In this plot the higher
death rate of third class passengers is immediately obvious.</p>
        <p>Within each box of the mosaic plot are the passenger names. At a micro-level the
names of individual passengers are visible. A small but readable six point font links to
biographies. Font attributes redundantly encode data: italics for women and children;
plain roman for men. Sans serifs for survivors and serifs for the dead. Figure 6 shows a
closeup near the center of the plot.</p>
        <p>One problem using text labels to create quantitative areas is the potential bias in
sizes that can occur if there is a concentration of particularly long labels in an area. For
example, surviving first class women and children names are on average 36.7 characters
long while deceased third class men average only 22.0 characters - the names former
segment are 67% longer than the latter. In this example, names have been shortened
to a familiar name and a surname, so that all passengers are reduced to similar visual
length. For example the passenger Cardeza, Mrs. James Warburton Martinez (Charlotte
Wardle Drake) is recorded with eight words in the passenger list, and is reduced in the
visualization to Charlotte Cardeza; while Bird, Miss Ellen is visualized as Ellen Bird.
When the passenger name length is reduced to two words, the above two segments are
13.9 and 13.6 characters respectively - a 2% difference.</p>
        <p>Another consideration is box size. The algorithm used here nudges rectangle sizes
larger to fit text. As a result, there is a margin of error between the areas if represented
accurately compared to areas adjusted to fit text. This is most acute at the smallest
sizes (as shown in the fourth column of table 1). For example, the thin orange box in
fig. 5 (representing deceased first class women and children) has only 7 members and
the height of the box is 8 points tall instead of 6.5 points tall - making this rectangle
approximately 20% larger in area than it should be.</p>
        <p>There are other considerations as well for using text in the areas. The minimum
height for a box is related to the height of a single line of text; whereas the minimum
width for a box is related to the width of a word - which is a much larger size than
height. The first version of the Titanic passenger mosaic plot (fig. 7 right) created splits
in the opposite orientation resulting in many tall narrow boxes. These tall narrow boxes
when adjusted for minimum widths, resulted in higher error rates, as shown in the sixth
column of table 1.</p>
      </sec>
      <sec id="sec-3-3">
        <title>Typographic Graph of Word-Emotion Association Lexicon</title>
        <p>
          Node-link diagrams (i.e. graphs) are sometimes used to represent sets. In anchored
maps [
          <xref ref-type="bibr" rid="ref20">20</xref>
          ] a circular layout is used to depict sets as nodes around a circle, and elements
as free-floating nodes connected to respective sets. Using a physics-based graph layout
model, element nodes are pulled based on set relations: elements belonging to only a
single set are pushed outside the circle close to their set, other elements pulled to a
position somewhere between their sets.
        </p>
        <p>Node-link diagrams have scalability issues. With few elements, links can clearly
show which set an element is connected to. However, when the number of elements is
high, the many overlapping links makes it difficult distinguish membership.</p>
        <p>Instead, element membership in sets can be encoded in font attributes. Links can
be de-emphasized to reduce clutter leaving visible clusters of labels. Clusters can be
visually inspected. If none of the elements stands-out from other elements in the
cluster, all the font attributes are the same across elements and the cluster is homogenous.
Furthermore, font attributes can be used to decode memberships.</p>
        <p>
          The typographic graph in figure 8 depicts 4463 words associated with 8 emotions
(based on [
          <xref ref-type="bibr" rid="ref21">21</xref>
          ]). Words can be associated with more than one emotion. The eight large
clusters around the perimeter are the words belonging to a single emotion. Each word
is encoded with font attributes to indicate its set membership. Starting at 10am and
proceeding clockwise: blackletter for anger, underline for fear, added exclamation for
surprise!, spacing for t r u s t, baseline shift for joy, small caps for ANTICIPATION,
lightweight for sadness, and italic for disgust. Word hue indicates sentiment
membership. Positive sentiment is green, negative sentiment is red, neither is blue, both is
amber. Overall, membership in 10 sets is encoded. The largest cluster (at the perimeter at
2am) is exclusively in the set trust.
        </p>
        <p>There are 140 unique set relations out of 256 possible. A portion of the graph
interior is shown in figure 9. Near the bottom left is a small cluster of words such as
SIMMER, SUSPICIOUS and SHARPEN - this entire cluster is homogenous
with blackletter caps indicating angry anticipation. To the right is a cluster
immediately visible as heterogeneous; with words such as WORRY, WILDERNESS and PLEA
in underline lightweight caps indicating fear, sadness and anticipation. However, two
other words positionally close to these words have different very memberships as
indicated by their attributes: p i o u s, in spaced italics indicates both disgust and trust,
while liquor in a blackletter with a shifting baseline indicates both anger and joy - the
latter two words both being singlular elements in these particular set intersections.</p>
        <p>Unlike the Titanic example, variance in word length was not addressed. The average
word length is 7.63 characters. For the 24 largest clusters (each more than 50 words),
the standard deviation is 0.36 characters. However as clusters become smaller, there can
be wider variation: the 10 words in angry surprise average only 6 characters.</p>
        <p>Instead of focusing on area accuracy, the design attempts intuitive encodings. Anger
words are in a blackletter font - sometimes associated with angry heavy metal bands.
Surprise adds an exclamation mark - literally a mark indicating astonishment. Fear uses
an underline - as the word line is associated with emotion fear in the lexicon. Joy uses
a baseline shift - making the word appear bouncy. In addition, interactive techniques,
such as mouseover links or tooltips, can be used to clearly indicate memberships.
4.4</p>
      </sec>
      <sec id="sec-3-4">
        <title>Typographic Cartogram of Country Risks</title>
        <p>Set elements are sometimes depicted on a map, such as geographic regions (e.g.
countries, states, counties) or points (e.g. gas stations, hospitals, etc.) Set membership might
be indicated by textures, icons, overlaid blobs, etc. Figure 10 is a closeup view of a
map indicating many different types of risk associated with each country via icons.
Long lines of icons require leader lines in congested areas. Countries are coloured by
summaries, but tiny countries are teeny dots difficult to see when zoomed out.</p>
        <p>Encoding a high number of set memberships can also be done with positional
encoding. If labels are constrained to a fixed length (e.g. country ISO codes), then font
attributes can be applied to each character separately. For example, USA, CAN, DEU,
etc have bold applied to the first, second or third character independently to indicate set
membership in three different sets.</p>
        <p>Extending this across a variety of font and other visual attributes suggests a possible
20 or more unique indications of set membership into a single label. Figure 11 shows a
typographic cartogram, where each country label indicates set membership by the
formatting of each character. In this example, nine set attributes are conveyed in the labels,
plus the background colour and the literal label. Countries with no risks are plain (e.g.</p>
        <p>USA), countries with all risks are completely bold, italic and underline (e.g. Ethiopia
ETH) and those in-between have only some combination of attributes (e.g.Brazil BRA).
5</p>
      </sec>
    </sec>
    <sec id="sec-4">
      <title>Additional Considerations</title>
      <p>The use of text elements raises additional issues not relevant to simple glyphs.
5.1</p>
      <sec id="sec-4-1">
        <title>Representing Quantities</title>
        <p>Because labels can be different lengths, there can be challenges attempting to use labels
to indicate counts. Various strategies can be used:</p>
        <p>
          Height. Typographic Venn (fig. 4) uses stacked labels making perception of
quantities a visual comparison of heights. Height comparison outperforms area estimation
tasks [
          <xref ref-type="bibr" rid="ref11 ref14">11,14</xref>
          ]. The approach sidesteps the issue of variable string lengths.
Fixed Length Codes. Typographic Cartogram (fig. 11) reduces string lengths to
consistent 3 letter codes. A fixed width font ensures consistent physical length.
Processed Labels. Typographic Mosaic (fig. 5) uses area to indicate quantities,
which requires consideration of string length. A simple algorithm reduces strings
to similar lengths thereby making the resulting areas directly comparable.
Accept Some Error. Emotion Words (fig. 8) accepts error and even uses
encodings such as spacing (making strings wider) and baseline shifts (making strings
taller) thereby increasing the error in the areas. This may be somewhat offset by
differences in density of characters.
        </p>
        <p>Other possible techniques to make string lengths more uniform could include
adjusting spacing between characters; using font variants with different width; or
padding with additional characters.
5.2</p>
      </sec>
      <sec id="sec-4-2">
        <title>Intuitive Mappings</title>
        <p>
          Many authors recognize the importance of establishing a metaphoric association
between a visual attribute and the concept encoded (e.g.[
          <xref ref-type="bibr" rid="ref19 ref7">7,19</xref>
          ]). This makes it easier for
users to infer meaning from an encoding with less effort required to learn and remember
them. In Typographic Venn (fig. 4), gender uses familiar color encodings, party
affiliations are indicated by text literally leaning left and right, added lines indicated added
degrees. In Typographic Mosaic (fig. 5), female is represented by italic, considered by
some to be more feminine than roman type [
          <xref ref-type="bibr" rid="ref13">13</xref>
          ]. As discussed earlier, Typographic
Graph (fig. 8) also uses various metaphoric cues.
5.3
        </p>
      </sec>
      <sec id="sec-4-3">
        <title>Differentiation vs. Decoding</title>
        <p>Gestalt principles indicate that visually similar items are perceived as belonging
together. The viewer immediately understands whether the two items are similar or
different if their font attributes are similar or different. Showing all attributes
simultaneously can be useful: for example, in fig. 9 the use of font attributes readily distinguishes
between adjacent elements belonging to the same set relations (homogenous attributes)
or different relations (heterogenous attributes).</p>
        <p>
          Decoding font attributes is a second step. Short-term working memory only holds
a small amount of information (3-7 items, e.g. [
          <xref ref-type="bibr" rid="ref25">25</xref>
          ]). This implies cognitive challenges
when encoding 10 memberships such as Typographic Graph (fig. 8). Furthermore, some
typographic attributes rely on the same low level visual channel (e.g. case and font
family both use shape) overloading the visual channel.
5.4
        </p>
      </sec>
      <sec id="sec-4-4">
        <title>Layout Considerations</title>
        <p>Unlike dots or icons, labels have an aspect ratio that is wider than tall. This impacts
layout algorithms. Typographic Mosaic generated more accurate areas when favouring
wider boxes over taller boxes. The initial Typographic Graph used a collision detection
algorithm which assumed square proportions of elements and tended to push text apart
vertically rather than horizontally, necessitating adjustments to the layout algorithm.</p>
      </sec>
      <sec id="sec-4-5">
        <title>5.5 Interaction</title>
        <p>Encoding many different font attributes simultaneously onto a label may make it
difficult for the viewer to decode all the memberships. Tooltips can be used to immediately
show memberships. If the task is focused on a subset of set memberships, font attributes
can be toggled on/off. Search can be used if the task requires locating a particular
element. Using SVG and javascript, toggling attributes can be done in a single line of code
and search can simply use the browser’s find feature (ctrl+f).
All examples shown are English language. Some languages do not have the same font
attributes (e.g. case). Also, languages that represent words with single glyphs will have
different considerations: inter character spacing is not available, encoding separate
characters within a word is not feasible, and so on.</p>
        <p>Evaluation is non-trivial and there are many confounding factors. Evaluation studies
should be done to assess ability to notice differences, ability to decode; ability to recall
a font mapping; change in perception of areas when font attributes are manipulated;
how readability is impacted by layout; how readability is impacted by the application
of multiple simultaneous font attributes; and limitations across different languages.
6</p>
      </sec>
    </sec>
    <sec id="sec-5">
      <title>Conclusions</title>
      <p>We have shown that font attributes can be used to extend set visualization techniques,
to enhance labeled elements with additional data such as set membership. These
attributes can aid detecting differences between elements, encode membership or other
data attributes, scale to at least 10 sets and thousands of items. Future work can include
extensions to other set layouts, evaluation studies, higher scalability (more elements,
more sets) and better techniques for integrating word semantics into encoding schemes.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          1.
          <string-name>
            <surname>Alsallakh</surname>
            ,
            <given-names>B.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Micallef</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Aigner</surname>
            ,
            <given-names>W.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Hauser</surname>
            ,
            <given-names>H.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Miksch</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Rodgers</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          :
          <article-title>Visualizing sets and set-typed data: State-of-the-art and future challenges</article-title>
          .
          <source>In: Eurographics conference on Visualization (EuroVis</source>
          )
          <article-title>-State of The Art Reports</article-title>
          . pp.
          <fpage>1</fpage>
          -
          <lpage>21</lpage>
          (
          <year>2014</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          2.
          <string-name>
            <surname>Bertin</surname>
            ,
            <given-names>J.: Sémiologie</given-names>
          </string-name>
          <string-name>
            <surname>Graphique.</surname>
          </string-name>
          Gauthier-Villars, Paris (
          <year>1967</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          3.
          <string-name>
            <surname>Borgo</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kehrer</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Chung</surname>
            ,
            <given-names>D.H.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Maguire</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Laramee</surname>
            ,
            <given-names>R.S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Hauser</surname>
            ,
            <given-names>H.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Ward</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Chen</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          :
          <article-title>Glyph-based visualization: Foundations, design guidelines, techniques and applications</article-title>
          .
          <source>In: Eurographics State of the Art Reports</source>
          . pp.
          <fpage>39</fpage>
          -
          <lpage>63</lpage>
          . EG STARs, Eurographics Association (May
          <year>2013</year>
          ), http://diglib.eg.org/EG/DL/conf/EG2013/stars/039-
          <fpage>063</fpage>
          .pdf
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          4.
          <string-name>
            <surname>Brase</surname>
            ,
            <given-names>G.L.</given-names>
          </string-name>
          :
          <article-title>Pictorial representations in statistical reasoning</article-title>
          .
          <source>Applied Cognitive Psychology</source>
          <volume>23</volume>
          (
          <issue>3</issue>
          ),
          <fpage>369</fpage>
          -
          <lpage>381</lpage>
          (
          <year>2009</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          5.
          <string-name>
            <surname>Brath</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          :
          <article-title>High category glyphs in industry</article-title>
          .
          <source>In: Visualization in Practice at 2015 IEEE Symposium on Information Visualization (VisWeek</source>
          <year>2015</year>
          ). IEEE (
          <year>2015</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          6.
          <string-name>
            <surname>Brath</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Banissi</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          :
          <article-title>The design space of typeface</article-title>
          .
          <source>Proceedings of the 2014 IEEE Symposium on Information Visualization (VisWeek</source>
          <year>2014</year>
          ) (
          <year>2014</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          7.
          <string-name>
            <surname>Brath</surname>
            ,
            <given-names>R.K.</given-names>
          </string-name>
          :
          <article-title>Effective Information Visualizaiton: Guidelines and Metrics for 3D Interactive Representations of Business Data</article-title>
          .
          <source>Ph.D. thesis</source>
          , University of Toronto (
          <year>1999</year>
          ), http://www. collectionscanada.gc.ca/obj/s4/f2/dsk1/tape7/PQDD_0006/MQ45944.pdf
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          8.
          <string-name>
            <surname>Brinton</surname>
          </string-name>
          , W.C.: Graphic Presentation. Brinton Associates (
          <year>1939</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          9.
          <string-name>
            <surname>Chen</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Floridi</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          :
          <article-title>An analysis of information in visualization</article-title>
          .
          <source>Synthese</source>
          (
          <year>2013</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          10.
          <string-name>
            <surname>Clark</surname>
          </string-name>
          , J.:
          <article-title>Twitter venn (</article-title>
          <year>2008</year>
          ), http://www.neoformix.com/2008/TwitterVenn.html, accessed
          <volume>04</volume>
          /02/2016
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          11.
          <string-name>
            <surname>Cleveland</surname>
            ,
            <given-names>W.S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>McGill</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          :
          <article-title>Graphical perception: Theory, experimentation, and application to the development of graphical methods</article-title>
          .
          <source>Journal of the American statistical association</source>
          <volume>79</volume>
          (
          <issue>387</issue>
          ),
          <fpage>531</fpage>
          -
          <lpage>554</lpage>
          (
          <year>1984</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          12. Collins,
          <string-name>
            <given-names>C.</given-names>
            ,
            <surname>Penn</surname>
          </string-name>
          ,
          <string-name>
            <given-names>G.</given-names>
            ,
            <surname>Carpendale</surname>
          </string-name>
          ,
          <string-name>
            <surname>S.</surname>
          </string-name>
          :
          <article-title>Bubble sets: Revealing set relations with isocontours over existing visualizations</article-title>
          .
          <source>IEEE Transactions on Visualization and Computer Graphics</source>
          <volume>15</volume>
          (
          <issue>6</issue>
          ),
          <fpage>1009</fpage>
          -
          <lpage>1016</lpage>
          (
          <year>2009</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          13.
          <string-name>
            <surname>Cousins</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          :
          <article-title>Design stereotypes: Masculine and feminine design techniques (</article-title>
          <year>2012</year>
          ), http: //designmodo.com/masculine-feminine-designs, accessed
          <volume>03</volume>
          /30/2016
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          14.
          <string-name>
            <surname>Heer</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Bostock</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          :
          <article-title>Crowdsourcing graphical perception: Using mechanical turk to assess visualization design</article-title>
          .
          <source>In: ACM Human Factors in Computing Systems (CHI)</source>
          . pp.
          <fpage>203</fpage>
          -
          <lpage>212</lpage>
          (
          <year>2010</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref15">
        <mixed-citation>
          15.
          <string-name>
            <surname>Heer</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Boyd</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          : Vizster:
          <article-title>Visualizing online social networks</article-title>
          .
          <source>In: IEEE Symposium on Information Visualization</source>
          ,
          <year>2005</year>
          .
          <source>INFOVIS 2005</source>
          . pp.
          <fpage>32</fpage>
          -
          <lpage>39</lpage>
          . IEEE (
          <year>2005</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref16">
        <mixed-citation>
          16.
          <string-name>
            <surname>Hodges</surname>
            ,
            <given-names>E.R.S.:</given-names>
          </string-name>
          <article-title>The Guild Handbook of Scientific Illustration</article-title>
          . John Wiley (
          <year>2003</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref17">
        <mixed-citation>
          17.
          <string-name>
            <surname>Hodson</surname>
            ,
            <given-names>Y.</given-names>
          </string-name>
          :
          <article-title>Popular maps: The ordnance survey popular edition one-inch map of england and wales,</article-title>
          <year>1919</year>
          -
          <fpage>1926</fpage>
          . (
          <year>1999</year>
          ), http://www.davidrumsey.com/luna/servlet/s/ s15w94, accessed:
          <volume>04</volume>
          /02/2016
        </mixed-citation>
      </ref>
      <ref id="ref18">
        <mixed-citation>
          18.
          <string-name>
            <surname>MacKinlay</surname>
          </string-name>
          , J.:
          <article-title>Automating the design of graphical representations</article-title>
          .
          <source>ACM Transactions on Graphics 5</source>
          ,
          <fpage>27</fpage>
          -
          <lpage>34</lpage>
          (
          <year>2006</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref19">
        <mixed-citation>
          19.
          <string-name>
            <surname>Maguire</surname>
          </string-name>
          , E.:
          <article-title>Systematising glyph design for visualization</article-title>
          .
          <source>Ph.D. thesis</source>
          , Oxford University (
          <year>2015</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref20">
        <mixed-citation>
          20.
          <string-name>
            <surname>Misue</surname>
            ,
            <given-names>K.</given-names>
          </string-name>
          :
          <article-title>Drawing bipartite graphs as anchored maps</article-title>
          .
          <source>In: Proceedings of the 2006 AsiaPacific Symposium on Information Visualisation-</source>
          Volume
          <volume>60</volume>
          . pp.
          <fpage>169</fpage>
          -
          <lpage>177</lpage>
          . Australian Computer Society, Inc. (
          <year>2006</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref21">
        <mixed-citation>
          21.
          <string-name>
            <surname>Mohammad</surname>
            ,
            <given-names>S.M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Turney</surname>
          </string-name>
          , P.D.:
          <article-title>Emotions evoked by common words and phrases: Using mechanical turk to create an emotion lexicon</article-title>
          .
          <source>In: Proceedings of the NAACL HLT</source>
          <year>2010</year>
          <article-title>workshop on computational approaches to analysis and generation of emotion in text</article-title>
          . pp.
          <fpage>26</fpage>
          -
          <lpage>34</lpage>
          . Association for Computational Linguistics (
          <year>2010</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref22">
        <mixed-citation>
          22.
          <string-name>
            <surname>Riche</surname>
            ,
            <given-names>N.H.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Dwyer</surname>
            ,
            <given-names>T.</given-names>
          </string-name>
          :
          <article-title>Untangling Euler diagrams</article-title>
          .
          <source>Visualization and Computer Graphics, IEEE Transactions on 16(6)</source>
          ,
          <fpage>1090</fpage>
          -
          <lpage>1099</lpage>
          (
          <year>2010</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref23">
        <mixed-citation>
          23.
          <string-name>
            <surname>Robinson</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Morrison</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Muehrcke</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kimerling</surname>
            ,
            <given-names>A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Guptill</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          : Elements of Cartography. John Wiley &amp; Sons, New York, NY (
          <year>1995</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref24">
        <mixed-citation>
          24.
          <string-name>
            <surname>Setlur</surname>
            ,
            <given-names>V.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Mackinlay</surname>
            ,
            <given-names>J.D.</given-names>
          </string-name>
          :
          <article-title>Automatic generation of semantic icon encodings for visualizations</article-title>
          .
          <source>In: Proceedings of the SIGCHI Conference on Human Factors in Computing Systems</source>
          . pp.
          <fpage>541</fpage>
          -
          <lpage>550</lpage>
          . ACM (
          <year>2014</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref25">
        <mixed-citation>
          25.
          <string-name>
            <surname>Ware</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          : Information Visualization: Perception for Design. Springer-Verlag (
          <year>2000</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref26">
        <mixed-citation>
          26.
          <string-name>
            <surname>Wilkinson</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          :
          <article-title>Exact and approximate area-proportional circular Venn and Euler diagrams</article-title>
          .
          <source>Visualization and Computer Graphics, IEEE Transactions on 18(2)</source>
          ,
          <fpage>321</fpage>
          -
          <lpage>331</lpage>
          (
          <year>2012</year>
          )
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>