<!DOCTYPE article PUBLIC "-//NLM//DTD JATS (Z39.96) Journal Archiving and Interchange DTD v1.0 20120330//EN" "JATS-archivearticle1.dtd">
<article xmlns:xlink="http://www.w3.org/1999/xlink">
  <front>
    <journal-meta />
    <article-meta>
      <title-group>
        <article-title>Predicting the In uence of Urban Vacant Lots on Neighborhood Property Values</article-title>
      </title-group>
      <contrib-group>
        <contrib contrib-type="author">
          <string-name>Muhammad Fazalul Rahman</string-name>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Pradeep Murukannaiah</string-name>
          <email>p.k.murukannaiah@tudelft.nl</email>
          <xref ref-type="aff" rid="aff0">0</xref>
        </contrib>
        <contrib contrib-type="author">
          <string-name>Naveen Sharma</string-name>
          <xref ref-type="aff" rid="aff1">1</xref>
        </contrib>
        <aff id="aff0">
          <label>0</label>
          <institution>Delft University of Technology</institution>
          ,
          <addr-line>Delft</addr-line>
          ,
          <country country="NL">Netherlands</country>
        </aff>
        <aff id="aff1">
          <label>1</label>
          <institution>Rochester Institute of Technology</institution>
          ,
          <addr-line>Rochester NY 14623</addr-line>
          ,
          <country country="US">USA</country>
        </aff>
      </contrib-group>
      <abstract>
        <p>Vacant lots are municipally-owned land parcels which were acquired post-abandonment or due to tax foreclosures. With time, failure to sell or nd alternate uses for vacant lots results in them causing adverse e ects on the health and safety of residents, and cost the city both directly and indirectly. Although existing research has tried to dene these impacts, cities need quanti able evidence from within the city to make planning decisions based on these studies. Moreover, trying to understand the impact of vacant lots in an uncontrolled setting makes it di cult to perform A key problem with existing methodologies is that they tend to look at the city as a whole, while ignoring the diverse socioeconomic factors at play. Altogether, city planners are left with little or no actionable information to prioritize conversion of vacant lots. In contrast, for our research we try to model the city as blocks, census tracts and neighborhoods while using relevant features to capture key demographic, economic and geographic characteristics. In addition, we build a deep learning model to quantify the impact of vacant lots on changing property values so as to recommend conversions that yields the maximum bene t through property value tax increase. Our results indicate that our model is able to capture the relationship between vacant lots and property values better than conventionally used algorithms and data models. Further, our model speci cally caters to small and mid size cities, which are often neglected in the mainstream urban computing research.</p>
      </abstract>
      <kwd-group>
        <kwd>Urban computing</kwd>
        <kwd>deep learning</kwd>
        <kwd>Gaussian processes</kwd>
        <kwd>spatiotemporal data</kwd>
        <kwd>computational social science</kwd>
        <kwd>vacant lots</kwd>
      </kwd-group>
    </article-meta>
  </front>
  <body>
    <sec id="sec-1">
      <title>INTRODUCTION</title>
      <p>
        In the past century, cities in the United States have undergone signi cant changes.
While some cities improved, with job opportunities that came with the
establishment of new and relocated industries and increased immigration, others su ered
from depopulation and job losses [
        <xref ref-type="bibr" rid="ref16">16</xref>
        ]. This led to properties in the latter cities
Copyright c 2020 for this paper by its authors. Use permitted under Creative
Commons License Attribution 4.0 International (CC BY 4.0).
getting abandoned or foreclosed due to tax delinquencies [
        <xref ref-type="bibr" rid="ref1 ref20 ref7">20, 7, 1</xref>
        ]. The structures
in these properties have to be demolished if found to be in hazardous conditions.
Such parcels of city-owned real estate without any known uses are known as
vacant lots, which comprise about an average of 15% land area across seventy
US cities according to a 2000 Brookings Institutions study [
        <xref ref-type="bibr" rid="ref17">17</xref>
        ].
      </p>
      <p>
        While vacant lots seem harmless to cities and residents, they are found to
have major impacts on the stability of neighborhoods and lives of neighborhood
residents. First, they become dumping grounds for litter and other solid wastes,
and eventually, health hazards if left unchecked. Since inspection for city-owned
properties are done seasonally, it becomes the responsibility of neighboring
residents to call the respective city o cials and report such conditions, which leads
to code inspections and corrective actions. Further, studies have shown that
the presence of abandoned properties and vacant lots can increase crime in the
neighboring vicinity [
        <xref ref-type="bibr" rid="ref15 ref6">6, 15</xref>
        ]. Moreover, vacant parcels can be perceived as a sign
of neglect and distress, which can drive down the values of neighboring
properties [
        <xref ref-type="bibr" rid="ref14">14</xref>
        ]. Property value depreciation further erodes the tax base of the already
declining budgets of cities, and incur major losses to these cities over time.
      </p>
      <p>To understand the situation within city o ces, we conducted meetings with
the city planners, property assessment o cers and other city o cials, which led
do the following observations.
1. Currently, the monetary impact of vacant lot is xed at USD 6 per lot, which
is the amount paid to contractors to perform the seasonal cleaning.
Indirect costs like cost of increased police surveillance that result from increased
crime, cost of unscheduled cleanups and property tax depreciation are often
overlooked.
2. City planners and assessors were unaware of the literature outlining the impact
of vacant lots on di erent neighborhood conditions including health, property
values, crime, etc.
3. Although di erent studies show varying levels of impact of vacant lots on
property values, they would not be able to consider these when making policy
changes since none of the studies were conducted in Rochester.
4. City policy allowed the leasing of vacant land by city residents for community
gardening. The length of the leases are set for 10 months, after which residents
can request for renewal.
5. City assessors use real-estate sales values for reassessing the property values
every four to ve years. Neighborhood conditions are not accounted for
directly in the assessments, but real-estate values can be heavily a ected by
these neighborhood factors.
6. While certain departments employ data scientists, their numbers are few and
they are primarily occupied with other responsibilities and would not be able
to focus their time and attention for machine learning applications using
urban data.</p>
      <p>City residents who are themselves interested to lease and convert these lots
are hindered by the policies in place since they need to demonstrate that the
bene ts of conversion outweigh the cost incurred on the city by the lots.</p>
      <p>In order to make informed decisions about vacant lot conversion, urban
planners and city administrators need data-driven models. However, as we described
earlier, there is a dearth of such models for small and medium-sized cities.
Models trained for bigger cities would not necessarily work well for smaller cities. In
an e ort to ll this gap, we develop a data-driven model of vacant lots and their
impact on neighborhoods for Rochester NY, a medium-sized, rust-belt city.</p>
      <p>In order to measure the impact, we consider the in uence of vacant lots on
neighborhood property values. Our discussions with city o cials suggest that
the property value tax depreciation is a primary source of revenue loss for a city.
Accordingly, demonstrating any relationship between vacant lots and property
values makes a strong case for policy changes toward vacant lot conversions.</p>
      <p>Our approach, rst, de nes a data model that takes into consideration
multiple hierarchies within a city|blocks, census tracts, and neighborhoods. We then
extract features relevant to our analysis from each layer in the city hierarchy
along with the characteristics of individual property parcels. Finally, we provide
the data model thus generated as input to a deep learning framework. Our
analysis shows that our model gives much better precision compared to conventional
methodologies used for predicting the impact of vacant lots on property values.</p>
      <p>The remainder of this paper is organized as follows. In Section 2, we discuss
related works. We formally de ne the problem and describe our data framework
in Section 3. The approaches we used are described in Section 4 followed by the
results obtained in Section 5. We conclude this paper in Section 7.
2</p>
    </sec>
    <sec id="sec-2">
      <title>Related Work</title>
      <p>
        The growing number of vacant lots in cities and their e ects on neighborhoods
have been well explored in the social science literature, with studies dating back
to mid-1950s [
        <xref ref-type="bibr" rid="ref4">4</xref>
        ]. These earlier works dealt with the cost of public works that
arise due to large areas being left vacant, which in turn increased in cost for
installation and maintenance for electric poles, cables, water mains, and so on.
However, it was in the late 1900s that population shift started being a more
signi cant issue for smaller cities [
        <xref ref-type="bibr" rid="ref16">16</xref>
        ]. This was the time when development in
metropolitan areas accelerated much faster than smaller cities, leading to housing
abandonment. Burchell and Listokin [
        <xref ref-type="bibr" rid="ref5">5</xref>
        ] describe abandonment as both a
symptom and disease; one that not only indicated urban decline, but also provided
the feedback mechanism to accelerate and perpetuate it. With the increase in
housing supplies due to abandonment, rental properties become unable to cover
taxes and related costs from the income they produce. The high supply, and
consequently declining demand, also make it nearly impossible for landlords to
sell them, increasing the pool of abandoned properties even further [
        <xref ref-type="bibr" rid="ref15 ref21 ref9">9, 15, 21</xref>
        ].
      </p>
      <p>
        Once the problem was well-de ned, further studies tried to quantify the exact
impact abandoned and vacant properties have on neighborhood dynamics. [
        <xref ref-type="bibr" rid="ref15">15</xref>
        ]
model the e ects of vacant lots on crime. Although the results showed
appreciably positive correlation, they were not statistically signi cant. In a later study
by [
        <xref ref-type="bibr" rid="ref8">8</xref>
        ], rather than just using data, the authors conducted a randomized control
trial by greening a vacant lot cluster to understand changes in crime and safety
relating to conversion. Another cluster was used as control group. The results
showed an insigni cant decrease in violent crime. However, a follow-up study
by [
        <xref ref-type="bibr" rid="ref3">3</xref>
        ] on 541 randomly sampled vacant lots showed signi cant improvement in
actual and perceived safety in the neighborhoods.
      </p>
      <p>
        Multiple studies have tried to model the relationship between vacant lots on
neighborhood property values. Immergluck and Smith [
        <xref ref-type="bibr" rid="ref14">14</xref>
        ] studied the impact
of housing foreclosures on property values in Chicago using a hedonic regression
model, and found statistically signi cant relationship between the two. An
overall estimated loss of $598 billion in property values was valuated using average
property value in the city. However, this model is not adequate enough to make
conversion decisions as it has high error rates. A study by [
        <xref ref-type="bibr" rid="ref10">10</xref>
        ], in addition to
modeling the impact of abandoned properties as a function of distance, used
weighted repeated sales model to estimate the impact that duration of
abandonment has on property values. The results indicate that both distance from
abandoned properties and duration of abandonment have signi cant impact on
property values. However, the use of repeated sales model requires sale prices of
the same property during di erent time periods. Such data would be sparse in
historic property value records, and the number of examples available would be
too low.
      </p>
      <p>
        Other studies have focused on the positive impacts of greening vacant land [
        <xref ref-type="bibr" rid="ref2">2</xref>
        ].
[
        <xref ref-type="bibr" rid="ref11 ref12">12, 11</xref>
        ] performed di erence-in-di erence analysis to understand the impact of
converting vacant land into green spaces. Her results suggest that while property
values tend to increase all over the city, the properties surrounding converted
vacant lots enjoy a greater increase in value. However, the results also indicated
that the impact was more pronounced in some parts of the city than others.
This shows the need for treating a city not as a whole but as a collection of
sub-populations.
      </p>
      <p>
        To the best of our knowledge, there is no existing work on modeling vacant
lots in the computing literature. Although there are computational models of
vacant lots in the social science literature, they tend to use simplistic regression
models. In line with recent advances in urban computing [
        <xref ref-type="bibr" rid="ref25">25</xref>
        ], we seek to use
state-of-the-art machine learning techniques for modeling the vacant lot problem.
      </p>
      <p>
        We address a novel problem. However, our solution borrow ideas from
several recent works on urban computing. For example, [
        <xref ref-type="bibr" rid="ref22">22</xref>
        ] de ne a model to
understand urban migrant mobility within the new city long with housing price
information, geographic information, call behavior and social connections. They
then used these features to model the problem of understand migrant churn as
a classi cation task. Similarly, we include multiple features to model vacant lots
and develop a baseline classi er.
      </p>
      <p>
        Huang et.al.[
        <xref ref-type="bibr" rid="ref13">13</xref>
        ] use a deep neural network architecture to build a crime
prediction framework which can capture the dynamic crime patterns and their
inter-dependencies. Their framework models the multidimensional interactions
between crime categories and regions over di erent time periods. [
        <xref ref-type="bibr" rid="ref24">24</xref>
        ] also built a
deep-learning based model to predict crowd ow within cities. We use a similar
deep learning approach to model the characteristics that conventional regression
models have not been able to successfully predict. We hypothesize that deep
neural networks are more e cient in capturing higher dimensional inter-dependent
features than linear regression models.
3
      </p>
    </sec>
    <sec id="sec-3">
      <title>The Vacant Lot Problem</title>
      <p>In this section, we outline our research questions and describe the dataset we
build to answer those questions.</p>
      <p>The vacant lot problem is a well-de ned and well explored problem in the
realm of urban sciences. However based on existing results, it is not possible
to drive data-driven decision making in cities. We therefore need to build a
city-speci c model that can capture the direct impact of vacant lots on nearby
properties, rather than use the same approximate impact percentage for every
property. Such a model that can explain e ects in the micro-level can help
urban planners make informed decisions, and can help improve the conditions in
neighborhoods. This can also help prioritize neighborhoods within the city that
require immediate attention so as to allocate limited budgets more e ciently.
Our objective through this research is to answer two key research questions:
RQ1 What impacts do vacant lots in a neighborhood have on the property
values of that neighborhood?
RQ2 How can we choose vacant lots to convert so as to maximize the bene ts
of conversion (minimize the negative impacts on property values if the lots
are not converted)?</p>
      <p>By answering RQ1, we hope to provide a monetary assessment that can help
city o cials in understanding the \silent impact" vacant lots have on the city,
and in particular, on individual properties within the city. Once we have a model
trained using data from a city, we can identify vacant lots within the city that
will have the most impact on property values if converted. For our model, this
would be based on the depreciation in property value the vacant lots cause.
Then, we can use this information to answer RQ2, i.e., to prioritize vacant lots
to convert based on budget constraints.
3.1</p>
      <sec id="sec-3-1">
        <title>Data Collection</title>
        <p>Data from larger cities are being made publicly available more often now, while
smaller cities neither have the resources nor the manpower to do the same.
When it comes to vacant lots, this has led to the lack of research about how
to manage vacant land within tight budgets. The common solution provided for
the problems caused by vacant lot is to convert them to green spaces, which is
not always feasible for the rust-belt cities. For our research, we therefore chose
to use one such city for analysis. Since the city of Rochester ts the pro le of
a declining city, and because their data was available for analysis, we chose to
center our research around Rochester, NY. Table 1 shows the number of vacant
lots in the city.</p>
        <p>The data about property parcels, 311 calls, crime and demographics for
Rochester are publicly available. While some of the data was in GIS format,
others varied depending on the software used in the respective o ces from which the
data has to be collected. All the data was made GIS-compatible by geo-coding
addresses and using unique IDs assigned by the city.</p>
        <p>Although we focus our research on Rochester, NY, our models and
methodologies are generic. We conjecture that our methods work for other similar sized
cities, too, with appropriate tuning.
In order to approach the vacant lot problem as a data problem, it is necessary to
capture the complex relationships that occur in an urban settings that can
contribute towards the di erent outcomes that can be observed in cities. Therefore,
http://www.cityofrochester.gov/innovation/
approaching a complex setting naively might lead to loss of critical information
about the underlying hierarchical structures and the inter- and intra-hierarchical
relationships between them. To this end, we model a city as three layers as shown
in Figure 1 to capture the essential characteristics of each level.</p>
        <p>Level 1: City Block. The lowest level in the vacant lot model hierarchy,
the city block is the smallest area that is surrounded by streets. Each city block
contains one or more property parcels, and is used to obtain neighborhood
characteristics which are mostly distance-sensitive.</p>
        <p>Level 2: Neighborhood. Neighborhoods form larger geographic boundaries
within cities, and are sometimes given o cial or semi-o cial statuses through
resident associations or watch groups. While neighborhood data fails to capture
distance-sensitive features, trends and policies that are usually similar for each
neighborhood can be acquired in this level.</p>
        <p>Level 3: City. At the city level most diverse characteristics of residents are
lost; however, aggregated city data can help di erentiate between cities and can
help adapt models based on city-speci c characteristics. City level aggregated
data also helps understand where neighborhood and city blocks stand in terms
of features like property values and crime.
3.3</p>
      </sec>
      <sec id="sec-3-2">
        <title>Feature Modeling</title>
        <p>Based on the hierarchies de ned above, we collected features for each of the three
levels of the hierarchy. The features can be subdivided into three categories:</p>
        <p>Spatio-Temporal Featuresinclude crime incidents, 311 calls, code
violations and property values. The examples in these datasets, with the exclusion of
property values, occur at a location mostly only once; therefore, they are
aggregated for each year for each level in the hierarchy. Property values (per square
feet) data is available for multiple years, although the ranges of available data
vary depending on the city. For example, for the city of Philadelphia, property
value data is available for every year from 2012 to 2017, whereas for Rochester,
data is available quadrennially from 1990 to 2017.</p>
        <p>Spatial Features include property parcel information and locations of parks,
schools, libraries and city facilities. Distances and densities of these features can
help model the block or neighborhood characteristics. Distance to the nearest
vacant lot and density of vacant lots in the city block are also used to incorporate
any impacts they might have on the models.</p>
        <p>
          Hedonic Features are the broken down constituent parts of a component
like real estate or consumer electronics that can be used to predict a dependent
variable [
          <xref ref-type="bibr" rid="ref19">19</xref>
          ]. For property parcels, the hedonic features include the area of the
lot and any units in the lot, number of rooms, stories, bathrooms, etc. This can
particularly help when we try to t regression models so that the cost identi ed
with these hedonic features can be accounted for, while being able to identify
the costs associated with vacant lot related features.
        </p>
        <p>Demographic Features are collected from various census datasets and
incorporated to model the social dynamics. This includes population, average
income, education attainment and market demand data. In addition, a diversity
index is included which shows the probability that two people chosen at random
belong to the same ethnic/racial group. The complete list of features extracted
from the hierarchical model has been listed in Table 2
Once the required data has been identi ed and collected as mentioned in the
previous sections, they have to be modeled to preserve neighborhood and block
characteristics, while avoiding the need to have multiple models within the city.
We therefore collect data about individual properties (parcel area, distance
to facilities, property prices, etc.) and join them with block level data (crime
count, demographics, property count, etc.) before scaling them among
neighborhoods. To this end, we consider a set of neighborhoods within the City of
Rochester (i.e., N = (N1; :::; NI )). Each neighborhood contains multiple census
tracts (C1; :::; Cj ) 2 Ni and street blocks (B1; :::; Bk) 2 Cj . Within these blocks,
there are multiple non-vacant property parcels (Px) and vacant property parcels
(Vy). We seek to model the e ects of vacant lots (Vy 2 Bk) on the properties
within the same block (Px 2 Bk). We rst construct a neighborhood matrix
(NS F ), where each sample S~ contains F features (as mentioned in the previous
section) for every property in the neighborhood collected from its corresponding
layers Ni; Cj and Bj as well as the individual parcel's characteristics (P~i) and
vacant lot characteristics (V~ ). That is,</p>
        <p>S~i = B~ k a C~j a P~x a V~y</p>
        <p>Once we have the neighborhood matrix, we standardize the features among
each neighborhood as opposed to standardizing the data for the entire city. This
can be represented as follows:
where F is the original feature vector, F is the mean of the feature vector and
is its standard deviation within the neighborhood.
4</p>
      </sec>
    </sec>
    <sec id="sec-4">
      <title>Approaches</title>
      <p>In this section we outline the approaches we explore using the data models
described in the previous section.
4.1</p>
      <sec id="sec-4-1">
        <title>Gaussian Process Regression</title>
        <p>
          Gaussian Processes are supervised non-parametric learning approaches in which
we consider the predictions to be probabilistic [
          <xref ref-type="bibr" rid="ref18">18</xref>
          ]. This is with the underlying
assumption that the probability distribution of a set of arbitrary points from
the dataset is jointly Gaussian with some mean and covariance. Just like in any
supervised learning models where we assume that the target variables are similar
for similar predictor variables, Gaussian Processes also follow the same
assumption and use a covariance matrix (x) to de ne the similarities. A characteristic
length scale ( l) is used to de ne the maximum distance between input values,
beyond which the target values become uncorrelated.
        </p>
        <p>For our analysis, we provided the set of features S~i as independent
variables and the change in property value value P as the dependent variable. In
addition, linear regression is also used as a baseline to compare with more
complicated analyses, since ordinary least square and linear regression is generally
used in social science literature to identify correlations.
4.2</p>
      </sec>
      <sec id="sec-4-2">
        <title>Arti cial Neural Network</title>
        <p>While conventional or Gaussian process regression works for most cases, they are
not always able to capture key relationships within the data. Especially when
it comes to a hierarchical framework like our model where there are multiple
parameters a ecting each other, regression might not be able to provide us with
the best possible t. To overcome this issue, we experiment with the application
of neural networks. We build a neural network that optimizes on di erent
hyperparameters like the optimizer, activation functions, batch sizes and epochs. Since
there is no one size ts all, we believe that this is essential to nd the best t,
given the unknown underlying relationships within and across the di erent layers
in our city hierarchy.</p>
        <p>The input layer to the neural network architecture consists of the features
from di erent hierarchies that we had discussed earlier (S~i). We de ne the
expected output to be the change in property values over the years ( P ). We
de ne the rst hidden layer to contain the same number of neurons as the input
layer, while the number of neurons in the layers after the rst is chosen
dynamically to optimize the loss function. We use mean squared error (MSE) as the
loss function for our experiment. We also perform parameter tuning using these
di erent activation functions, optimizers, and data models and try to nd the
right combination that optimizes our result.
4.3</p>
      </sec>
      <sec id="sec-4-3">
        <title>Conversion Prioritization</title>
        <p>One we have a model that captures the cost associated with the vacant features
for each property, we can modify the data in a way which would re ect the
changes in the neighborhood if all the vacant lots are converted. This gives
us the pre- and post-conversion property values, helping us determine which
vacant lots are causing the most impact in the neighborhood. As mentioned
earlier, conversion of all the lots in the city is infeasible; therefore, choosing lots
to convert is based on budget constraints.</p>
        <p>
          This then becomes similar to a bin packing problem [
          <xref ref-type="bibr" rid="ref23">23</xref>
          ], where the number
of bins would be the number of vacant lots that can be converted based on
the city's constraints and selecting the vacant lots from the entire pool so as to
optimize pro t is the goal. Having to check every possible combination of vacant
lots to convert is an NP-hard problem. However, it is possible to sort the vacant
lots in the decreasing order of property value impact and then the rst t for
the budget can be found.
5
        </p>
      </sec>
    </sec>
    <sec id="sec-5">
      <title>Results and Discussion</title>
      <p>In this section we perform experiments to evaluate how well our data and
machine learning models perform on real-world datasets collected from the City of
Rochester. Our aim is to identify how well our models perform compared to
currently used techniques used for the purpose of property value impact prediction.</p>
      <p>We begin by visualizing mean changes in property values with respect to the
number of vacant lots in each block. That is, we group the properties based on
the number of vacant lots that are present in its block and then average the
property values within the groups. As it can be seen from Figure 3, the average
property values tend to be showing intuitive results, with properties without
any vacant lots nearby having the highest property values, and as the number
of vacant lots increase, the average property values seem to be lower. However,
the variance was very high for these averages, and therefore it is not possible to
directly understand the changes in property values using averages.</p>
      <p>Next, we plot similar graphs with the number of 311 calls (non-informational)
made, number of code violations, and vacant lots, and found similar results, but
they too su ered the same problem with high variance. However, when we tried
plotting the relationship between the average number of crime incidents in the
block and number of vacant lots, the graphs did not show any correlation.
5.1</p>
      <sec id="sec-5-1">
        <title>Gaussian Process Regression</title>
        <p>This section discusses the results from the experiments described in Section
4.2. We used the hierarchical data model to run linear and Gaussian Process
Regression (GPR) algorithms to understand whether the independent variables
are able to predict the change in property prices. We tried using di erent kernels
for GPR to nd the best t for our data. We used data from the city of Rochester,
with the target variable being the ratio of property values in the year 2018 to
property values in 2014. These two years were chosen in particular as the housing
market prices, after having su ered a sudden drop during the nancial crisis, has
shown signs of improvement in recent years. As it can be observed in Figure 3,
the mean property prices have been improving signi cantly between 2013 and
2014, and then becomes steadier after that.</p>
        <p>The results obtained, as shown in Table 3, show that Gaussian process
regression models are able to provide slightly better results compared to basic linear
regression. However, this improvement is not su cient to correctly predict
property value changes as on an average, the Matern 5/2 kernel would give an error
of approximately 11%. This shows that even with a non-parametric probabilistic
model, the relationships in a social setting is di cult to establish, showing the
need for more complex multi-layer models.
Similar to GPR, we used the same hierarchical model as input to our neural
networks. We gave this as input to a network with a single hidden layer having 34
neurons (equal to the number of features). More hidden layers were incrementally
added and number of neurons reduced to optimize the errors. Upon tuning these
two parameters, the optimal results were obtained with 4 hidden layers- three
with 30 neurons each and the last layer with a single neuron.</p>
        <p>We then trained the model by using di erent combinations of the optimizers
and activation functions to nd the model with the least mean squared error
(MSE). As shown in Figure 4, the results for sigmoid and relu activation functions
were close to each other, while the use of tanh as activation function gives much
higher error comparatively. The best results were obtained using sigmoid as the
activation function and Adagrad as the optimizer. The error seems to atten out
as the number of epochs approach 1000, with the MSE for this combination being
0.00493, which translates to an average error of 6-7%. This model outperforms
other regression-based models commonly used in social science literature, and
provides better estimates about the e ects of vacant lots on nearby property
values. This con rms our intuition that the use of deep learning for predicting
property value prices in the neighborhood shows much more promising results
than linear and Gaussian process regressions.</p>
        <p>These observations lead us to answer our research questions.</p>
        <p>RQ1: What impacts do vacant lots in a neighborhood have on the property
values of that neighborhood? The process of modeling social relationships is
complex, and it is likely that even with all the available data, key neighborhood
characteristics are lost in the modeling process. However, by exploiting
existing literature, we were able to include some of the most relevant feature. We
used neighborhood aggregation, which was not used in prior literature, to
compare properties within a neighborhood rather than comparing with an entire
city. While di erent studies have shown varying results, the results we obtained
using conventional regression models were inadequate to make property value
predictions based on vacant lot features.</p>
        <p>With the use of neural networks, the performance seems to improve
signi cantly, allowing for better predictions. Based on experimental results, we
conclude that a hierarchical data model combined with a deep neural network
architecture can be used to capture neighborhood characteristics and perform
property value predictions with low error margins. We used the model thus
trained and changed the data to re ect the conditions that would occur if all the
vacant lots in the city are converted. That is, almost all of the features related
to vacant lots would become zero. We use this as the input to our model for
prediction. Based on the results obtained, by converting every vacant lot in the
city, the total property value increase in Rochester is approximately 1.54 million.</p>
        <p>RQ2: How can we choose vacant lots to convert so as to maximize the bene ts
of conversion (minimize the negative impacts on property values if the lots are
not converted)? While it is not feasible for most cities to re-purpose every vacant
lot, based on the data obtained from the model, it is possible to sort out the
vacant lots that have the highest impact on nearby property values. However,
it is not necessary to convert every single vacant lot near a property to observe
improvement in property values. That is, the e ects of vacant lots are observed
when there is a cluster of such lots near a property. Converting even a couple of
these lots can bring about changes to the property values.</p>
        <p>To optimally chose vacant lots to convert, it is necessary to iteratively change
vacant lot density data for each property to re ect conversion and test the change
in property value with those parameters. If the budget allotted by the city for
vacant lot conversion is x, it is necessary to try every possible combination of vacant
lots that can be converted, and the corresponding total change it would bring
to property values. This can then help order the vacant lots in the descending
order of impact and the top x vacant lots can be selected for intervention.
5.3</p>
      </sec>
      <sec id="sec-5-2">
        <title>Social Implications</title>
        <p>As it was demonstrated in this study, we were able to use data that was publicly
available to build models that can predict the impact of vacant lots on
neighborhood property values. However, the key motivation for understanding this
impact was to show that it is possible to gather evidence from within the city
to drive changes in city policy. While this study was restricted to one use case
and one city, similar implementations can help both city o cials and residents
derive evidence for other urban problems as well.</p>
        <p>With the prioritization of vacant lots based on impact, it becomes possible
for city o cials to nd locations where interventions can bring about the most
impact. These interventions can be in the form of incentives to residents for
fostering conversions of these lots, or through investments or subsidization by
the city that might make these lots more desirable for purchase. While di erent
studies have shown the impact of conversions of these lots, the data about these
impacts is di cult to acquire. With such data, it would also become possible to
recommend actions that yield the best outcome.</p>
        <p>However, unlike conventional machine learning applications, the use of data
for urban planning decision making can have implications on the lives of cities'
residents. As demonstrated in this paper, it is possible to apply optimization
techniques on social problems and minimize for errors, but without thoroughly
understanding the reasoning behind why a model has given a particular result
or recommendation, it would be risky to deploy it for decision making.</p>
        <p>With the vacant lot impact assessment tool, the same problem arises. While
the model was able to provide better performance compared to Gaussian
process or linear regression, it isn't apparent what led the model to made these
conclusions. Since the key set of features included demographics, it is possible
that the model might have learned with inherent biases. Another problem that
might arise could be gentri cation. In an ideal scenario where all vacant lots get
converted or sold, it is likely that the real estate demand would go up in an area.
This can further lead to increase in property value assessments and subsequently,
higher property value taxes, leading to gentri cation. Although this is
speculative, since it a ects the lives of citizens, it is always better to err on the side of
caution. We therefore believe that it is necessary to improve the model to (1)
provide better explainability before deployment, and (2) conduct a longitudinal
impact study about the impact of conversions on di erent neighborhood factors.
6</p>
      </sec>
    </sec>
    <sec id="sec-6">
      <title>Limitations</title>
      <p>Firstly, with the large number of unknown variables that might occur in a
social setting, it becomes almost impossible to do a intervention-control study to
understand the causal relationships behind the impact of vacant land on
neighborhood property values. Such causal studies have been performed in social
sciences in various cities, and therefore we make the assumption that the same
causal relationships exist in Rochester as well.</p>
      <p>Secondly, while we have done our best to ensure that the features collected
and used are as accurate as possible, there still exists the possibility that the
changes in property values were also tied to some unknown variable. Based on the
discussion with the city assessor, it was evident that vacant lots do not directly
impact property assessments, but rather have a more indirect impact through
real estate values. In fact, it was mentioned during the meeting that no
neighborhood factors (crime, blight, demographics, etc.) are taken into consideration
when assessing properties.</p>
    </sec>
    <sec id="sec-7">
      <title>Conclusion</title>
      <p>Urban data is often under-utilized yet highly valuable in making informed policy
and urban planning decisions. One domain where the exploitation of such data
can bring about positive change is the vacant lot problem. While the e ects
of property abandonment and subsequent generation of vacant lots have been
extensively discussed, the models built using these methodologies are inaccurate
to be used to understand the e ects vacant lots have on a smaller scale. Moreover,
no tools and models for using this knowledge to make informed decisions exist.
With our research, we propose a novel way of representing data so as to capture
key neighborhood characteristics, and also understand the impact of vacant lots
on neighborhood property values. We also propose a deep learning framework
that can predict changes in property values with respect to a set of
vacantlot-related features. We then show experimental evidence that our model shows
better results compared to baseline methods. Unlike other models, our model
caters to small and mid-sized cities,making it easier to make informed policy
decisions while taking budget constraints into consideration.</p>
      <p>Notwithstanding the improvement and accuracy obtained, some directions
exist for future work. Firstly, this framework could further be improved by using
recurrent neural networks and using time series crime and 311 call data. Secondly,
only limited data was made available to us about property values. With data
spanning longer periods of time along with observable conversion of vacant lots
for residential or public use, it would be possible to generate better estimates
about the impact of conversion depending on what the lot is being re-purposed
into. Lastly, with data from multiple cities, a model learned from one city can be
transferred to another without the need for re-training using transfer learning.</p>
    </sec>
  </body>
  <back>
    <ref-list>
      <ref id="ref1">
        <mixed-citation>
          1.
          <string-name>
            <surname>Accordino</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          , Johnson, G.T.:
          <article-title>Addressing the vacant and abandoned property problem</article-title>
          .
          <source>Journal of Urban A airs 22(3)</source>
          ,
          <volume>301</volume>
          {
          <fpage>315</fpage>
          (
          <year>2000</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref2">
        <mixed-citation>
          2.
          <string-name>
            <surname>Branas</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Cheney</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>M MacDonald</surname>
          </string-name>
          ,
          <string-name>
            <given-names>J.</given-names>
            ,
            <surname>Tam</surname>
          </string-name>
          ,
          <string-name>
            <given-names>V.</given-names>
            ,
            <surname>Jackson</surname>
          </string-name>
          ,
          <string-name>
            <surname>T.</surname>
          </string-name>
          ,
          <string-name>
            <given-names>R Ten</given-names>
            <surname>Have</surname>
          </string-name>
          ,
          <string-name>
            <surname>T.</surname>
          </string-name>
          :
          <article-title>A di erence-in-di erences analysis of health, safety, and greening vacant urban space</article-title>
          .
          <source>American journal of epidemiology 174</source>
          ,
          <volume>1296</volume>
          {
          <volume>306</volume>
          (11
          <year>2011</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref3">
        <mixed-citation>
          3.
          <string-name>
            <surname>Branas</surname>
            ,
            <given-names>C.C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>South</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Kondo</surname>
            ,
            <given-names>M.C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Hohl</surname>
            ,
            <given-names>B.C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Bourgois</surname>
            ,
            <given-names>P.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Wiebe</surname>
            ,
            <given-names>D.J.</given-names>
          </string-name>
          , MacDonald,
          <string-name>
            <surname>J.M.:</surname>
          </string-name>
          <article-title>Citywide cluster randomized trial to restore blighted vacant land and its e ects on violence, crime, and fear</article-title>
          .
          <source>Proceedings of the National Academy of Sciences</source>
          (
          <year>2018</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref4">
        <mixed-citation>
          4.
          <string-name>
            <surname>Brown</surname>
            ,
            <given-names>E.R.:</given-names>
          </string-name>
          <article-title>The vacant lot problem in american cities</article-title>
          .
          <source>American Journal of Economics and Sociology</source>
          <volume>17</volume>
          (
          <issue>1</issue>
          ),
          <volume>41</volume>
          {
          <fpage>42</fpage>
          (
          <year>1957</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref5">
        <mixed-citation>
          5.
          <string-name>
            <surname>Burchell</surname>
            ,
            <given-names>R.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Listokin</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          :
          <article-title>Property abandonment in the united states</article-title>
          . In: The Adaptive Reuse Handbook: Procedures to Inventory, Control, Manage, and
          <article-title>Reemploy Surplus Municipal Properties</article-title>
          . Rutgers University, Center for Urban Policy Research (
          <year>1981</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref6">
        <mixed-citation>
          6.
          <string-name>
            <surname>Cui</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Walsh</surname>
          </string-name>
          , R.:
          <article-title>Foreclosure, vacancy and crime</article-title>
          .
          <source>Journal of Urban Economics</source>
          <volume>87</volume>
          ,
          <issue>72</issue>
          {
          <fpage>84</fpage>
          (
          <year>2015</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref7">
        <mixed-citation>
          7.
          <string-name>
            <surname>Garvin</surname>
            ,
            <given-names>E.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Branas</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Keddem</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Sellman</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Cannuscio</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          :
          <article-title>More than just an eyesore: local insights and solutions on vacant land and urban health</article-title>
          .
          <source>Journal of Urban Health</source>
          <volume>90</volume>
          (
          <issue>3</issue>
          ),
          <volume>412</volume>
          {
          <fpage>426</fpage>
          (
          <year>2013</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref8">
        <mixed-citation>
          8.
          <string-name>
            <surname>Garvin</surname>
            ,
            <given-names>E.C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Cannuscio</surname>
            ,
            <given-names>C.C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Branas</surname>
            ,
            <given-names>C.C.</given-names>
          </string-name>
          :
          <article-title>Greening vacant lots to reduce violent crime: a randomised controlled trial</article-title>
          .
          <source>Injury Prevention</source>
          <volume>19</volume>
          (
          <issue>3</issue>
          ),
          <volume>198</volume>
          {
          <fpage>203</fpage>
          (
          <year>2013</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref9">
        <mixed-citation>
          9.
          <string-name>
            <surname>Goldstein</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Jensen</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Reiskin</surname>
          </string-name>
          , E.:
          <article-title>Urban vacant land redevelopment: Challenges and progress (</article-title>
          <year>2001</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref10">
        <mixed-citation>
          10. Han,
          <string-name>
            <surname>H.S.:</surname>
          </string-name>
          <article-title>The impact of abandoned properties on nearby property values</article-title>
          .
          <source>Housing Policy Debate</source>
          <volume>24</volume>
          (
          <issue>2</issue>
          ),
          <volume>311</volume>
          {
          <fpage>334</fpage>
          (
          <year>2014</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref11">
        <mixed-citation>
          11.
          <string-name>
            <surname>Heckert</surname>
            ,
            <given-names>M.</given-names>
          </string-name>
          :
          <article-title>Access and equity in greenspace provision: A comparison of methods to assess the impacts of greening vacant land</article-title>
          .
          <source>Transactions in GIS 17(6)</source>
          ,
          <volume>808</volume>
          {
          <fpage>827</fpage>
          (
          <year>2012</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref12">
        <mixed-citation>
          12.
          <string-name>
            <surname>Heckert</surname>
            ,
            <given-names>M.:</given-names>
          </string-name>
          <article-title>A spatial di erence-in-di erences approach to studying the e ect of greening vacant land on property values</article-title>
          .
          <source>Cityscape</source>
          <volume>17</volume>
          (
          <issue>1</issue>
          ),
          <volume>51</volume>
          (
          <year>2015</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref13">
        <mixed-citation>
          13.
          <string-name>
            <surname>Huang</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Zhang</surname>
          </string-name>
          , J.,
          <string-name>
            <surname>Zheng</surname>
            ,
            <given-names>Y.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Chawla</surname>
            ,
            <given-names>N.V.</given-names>
          </string-name>
          :
          <article-title>Deepcrime: Attentive hierarchical recurrent networks for crime prediction</article-title>
          .
          <source>In: Proceedings of the 27th ACM International Conference on Information and Knowledge Management</source>
          . pp.
          <volume>1423</volume>
          {
          <fpage>1432</fpage>
          . CIKM '18,
          <string-name>
            <surname>ACM</surname>
          </string-name>
          , New York, NY, USA (
          <year>2018</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref14">
        <mixed-citation>
          14.
          <string-name>
            <surname>Immergluck</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Smith</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          :
          <article-title>The external costs of foreclosure: The impact of singlefamily mortgage foreclosures on property values</article-title>
          .
          <source>Housing Policy Debate</source>
          <volume>17</volume>
          (
          <issue>1</issue>
          ),
          <volume>57</volume>
          {
          <fpage>79</fpage>
          (
          <year>2006</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref15">
        <mixed-citation>
          15.
          <string-name>
            <surname>Immergluck</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Smith</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          :
          <article-title>The impact of single-family mortgage foreclosures on neighborhood crime</article-title>
          .
          <source>Housing Studies</source>
          <volume>21</volume>
          (
          <issue>6</issue>
          ),
          <volume>851</volume>
          {
          <fpage>866</fpage>
          (
          <year>2006</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref16">
        <mixed-citation>
          16.
          <string-name>
            <surname>Inman</surname>
            ,
            <given-names>R.P.</given-names>
          </string-name>
          :
          <article-title>Making cities work: Prospects and policies for urban America</article-title>
          . Princeton University Press (
          <year>2009</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref17">
        <mixed-citation>
          17.
          <string-name>
            <surname>Pagano</surname>
            ,
            <given-names>M.A.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Bowman</surname>
            ,
            <given-names>A.O.</given-names>
          </string-name>
          :
          <article-title>Vacant land in cities: An urban resource</article-title>
          .
          <source>Brookings Institution</source>
          , Center on Urban and Metropolitan Policy Washington, DC (
          <year>2000</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref18">
        <mixed-citation>
          18.
          <string-name>
            <surname>Rasmussen</surname>
            ,
            <given-names>C.E.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Nickisch</surname>
          </string-name>
          , H.:
          <article-title>Gaussian processes for machine learning (gpml) toolbox</article-title>
          .
          <source>Journal of machine learning research 11(Nov)</source>
          ,
          <volume>3011</volume>
          {
          <fpage>3015</fpage>
          (
          <year>2010</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref19">
        <mixed-citation>
          19.
          <string-name>
            <surname>Rosen</surname>
            ,
            <given-names>S.</given-names>
          </string-name>
          :
          <article-title>Hedonic prices and implicit markets: Product di erentiation in pure competition</article-title>
          .
          <source>Journal of Political Economy</source>
          <volume>82</volume>
          (
          <issue>1</issue>
          ),
          <volume>34</volume>
          {
          <fpage>55</fpage>
          (
          <year>1974</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref20">
        <mixed-citation>
          20.
          <string-name>
            <surname>Schilling</surname>
            ,
            <given-names>J.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Logan</surname>
          </string-name>
          , J.:
          <article-title>Greening the rust belt: A green infrastructure model for right sizing america's shrinking cities</article-title>
          .
          <source>Journal of the American Planning Association</source>
          <volume>74</volume>
          (
          <issue>4</issue>
          ),
          <volume>451</volume>
          {
          <fpage>466</fpage>
          (
          <year>2008</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref21">
        <mixed-citation>
          21.
          <string-name>
            <surname>Sternlieb</surname>
            ,
            <given-names>G.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Burchell</surname>
            ,
            <given-names>R.W.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Hughes</surname>
            ,
            <given-names>J.W.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>James</surname>
            ,
            <given-names>F.J.:</given-names>
          </string-name>
          <article-title>Housing abandonment in the urban core</article-title>
          .
          <source>Journal of the American Institute of Planners</source>
          <volume>40</volume>
          (
          <issue>5</issue>
          ),
          <volume>321</volume>
          {
          <fpage>332</fpage>
          (
          <year>1974</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref22">
        <mixed-citation>
          22.
          <string-name>
            <surname>Yang</surname>
            ,
            <given-names>Y.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Liu</surname>
            ,
            <given-names>Z.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Tan</surname>
            ,
            <given-names>C.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Wu</surname>
            ,
            <given-names>F.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Zhuang</surname>
            ,
            <given-names>Y.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Li</surname>
            ,
            <given-names>Y.</given-names>
          </string-name>
          :
          <article-title>To stay or to leave: Churn prediction for urban migrants in the initial period</article-title>
          . CoRR abs/
          <year>1802</year>
          .09734 (
          <year>2018</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref23">
        <mixed-citation>
          23.
          <string-name>
            <surname>Yao</surname>
            ,
            <given-names>A.C.C.</given-names>
          </string-name>
          :
          <article-title>New algorithms for bin packing</article-title>
          .
          <source>J. ACM</source>
          <volume>27</volume>
          (
          <issue>2</issue>
          ),
          <volume>207</volume>
          {227 (Apr
          <year>1980</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref24">
        <mixed-citation>
          24.
          <string-name>
            <surname>Zhang</surname>
          </string-name>
          , J.,
          <string-name>
            <surname>Zheng</surname>
            ,
            <given-names>Y.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Qi</surname>
            ,
            <given-names>D.</given-names>
          </string-name>
          :
          <article-title>Deep spatio-temporal residual networks for citywide crowd ows prediction</article-title>
          .
          <source>CoRR abs/1610</source>
          .00081 (
          <year>2016</year>
          )
        </mixed-citation>
      </ref>
      <ref id="ref25">
        <mixed-citation>
          25.
          <string-name>
            <surname>Zheng</surname>
            ,
            <given-names>Y.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Capra</surname>
            ,
            <given-names>L.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Wolfson</surname>
            ,
            <given-names>O.</given-names>
          </string-name>
          ,
          <string-name>
            <surname>Yang</surname>
          </string-name>
          , H.:
          <article-title>Urban computing: Concepts, methodologies, and applications</article-title>
          .
          <source>ACM Transaction on Intelligent Systems and Technology (October</source>
          <year>2014</year>
          )
        </mixed-citation>
      </ref>
    </ref-list>
  </back>
</article>