How to read this note
The note works through one investment decision with one main method and a set of checks. It starts with the investment committee's argument for holding a Long Island City parcel, shows how to build a fair stand-in for what home values would have done without Amazon's announcement, and applies it to the two ZIP codes where Amazon announced sites. The checks test the result a second way, at metropolitan level, with rents, with neighboring ZIPs, under alternative setups and against event dates. A section on statistical power explains why the New York test is weak, and the last sections take the committee's five claims apart and say what the committee could and could not conclude.
Terms are explained where they first appear and again in the glossary. The text reports gaps in percent. Figures and tables report them in log points (100 times the natural log of a ratio), which are close to percent for gaps this small.
What the analysis found
After Amazon's announcement, home values in Crystal City moved above a matched comparison and home values in Long Island City did not. From November 2018 to February 2020 the typical home value in Crystal City (ZIP 22202) averaged 3.7 percent above its synthetic control, a weighted blend of other Washington-area ZIPs built to track 22202 closely before the announcement. Among 22202 and 304 placebo ZIPs (ZIPs that received no announcement), 22202 has the largest gap relative to its pre-announcement fit. It ranks 1 of 305, with a p-value of 0.003.
Two cautions apply. The Crystal City gap opens slowly. It averaged 0.5 percent in the first three months and reached 6.9 percent by February 2020, after several later milestones. And the three neighboring Arlington ZIPs show an average gap of 2.5 percent, so the Crystal City estimate bundles the announcement with wider-area movement and later events.
In Long Island City (ZIP 11101) the early gap is +0.01 percent, so values did not jump when the news came. The gap turns negative in the value dated 2019-01, a month before the cancellation of 2019-02-14, and averages -3.7 percent from November 2018 to February 2020. Placebo ZIPs in New York produce a gap that large about 15 percent of the time, so the gap is inside the range of chance. The smoothed condo series shows an early gap of +1.9 percent that fades to +0.3 percent by February 2020, also inside that range.
The metropolitan-level comparison finds nothing for either site. Against the 12 other finalist metropolitan areas, the Washington metro's typical home value moved -0.2 percent on average and the New York metro's -1.1 percent, and both are inside the range of chance. Washington's null sits next to a ZIP-level result that ranks first of 305, because a change in one ZIP is diluted in a metropolitan average.
The rent (ZORI) fits are poor, with a typical monthly miss before the announcement of 1.42, 1.69 and 2.65 log points for the three rent ZIPs. The rent results are reported and not relied on.
| Estimate | Early gap (percent) | Average gap (percent) | p-value |
|---|---|---|---|
| Crystal City (22202) against its synthetic control | +0.53 | +3.71 | 0.003 ratio; 0.033 size |
| Long Island City (11101), raw series | +0.01 | -3.68 | 0.362 ratio; 0.152 size |
| Long Island City (11101), smoothed condo series | +1.89 | +1.22 | 0.854 ratio; 0.702 size |
| Washington metro against the other finalist metros | -0.01 | -0.21 | 0.846 |
| New York metro against the other finalist metros | -0.07 | -1.14 | 0.462 |
Table 1. The main estimates, in percent. The early gap averages November 2018 to January 2019 and the average gap November 2018 to February 2020. The ratio p-value ranks the post-announcement miss relative to the pre-announcement miss among placebo units; the size p-value ranks the average gap itself, ignoring sign. The metro rows are event-study coefficients with one p-value, ranked by size. Exact values in log points are in Table 5 and Table 8.
The committee's argument for holding the parcel rests on five claims, and each fails for a different reason: a before-and-after change with no comparison, a premium assumed for a ZIP whose raw series shows none, a loss dated from a smoothed series, a unit of analysis too large for the shock, and a co-movement the data do not show. The section on the five claims gives the number that exposes each. All results are reported as computed, including nulls and poor fits.
What you should be able to do after this note
- Explain why a before-and-after change in one ZIP says little, and build a comparison group that shows what the ZIP would have done without the news.
- Read a synthetic control, a placebo test and a p-value well enough to explain each to a colleague.
- Match the unit of analysis to the size of the shock, and say why a metropolitan average cannot see a change confined to one neighborhood.
- Decide whether a test could have seen an effect of the plausible size before treating a null result as evidence of no effect.
- Say what monthly, smoothed and restated home-value indexes can and cannot date.
- Turn an imprecise estimate into a sell or hold recommendation, fix the date, comparison, outcome and test before looking, and name the evidence that would change the recommendation.
The decision and the argument for holding
An investment committee holds a Long Island City development parcel, bought in December 2018 weeks after Amazon named the neighborhood for half of its second headquarters. In early March 2020 it must decide whether to sell. A memo argues for holding in five steps (Table 2). Each step rests on a number that is correct as arithmetic. The rest of the note asks whether the number supports the step.
| Step | The memo's claim | The number the memo cites |
|---|---|---|
| 1 | Amazon's announcement was worth about 10 percent to Crystal City. | The typical home value in ZIP 22202 rose from $589,285 in October 2018 to $648,353 in October 2019, a gain of 10.0 percent. |
| 2 | Long Island City was offered identical terms (25,000 jobs and $2.5 billion of investment), would have earned the same premium, and lost it when Amazon cancelled on February 14, 2019. | ZIP 11101 fell from $1,076,866 in October 2018 to $1,005,333 in February 2020, 6.6 percent below its October 2018 level. |
| 3 | The loss followed the cancellation. | The condo index for 11101, which Zillow smooths and adjusts for season, peaked in December 2018 and was down only 0.2 percent in January 2019. |
| 4 | The cancellation did not hurt New York. | The New York metropolitan area differed from the average of the other 12 finalist metropolitan areas by -1.1 percent over the 16 months from November 2018 to February 2020. |
| 5 | ZIP 11101 follows its metropolitan area, so it will recover with it and a sale now locks in a loss. | From January 2014 to October 2018 the ZIP rose 45.3 percent and the metropolitan area 25.3 percent; the metropolitan area is up 2.4 percent since October 2018. |
Table 2. The five steps of the memo's argument for holding, with the number each one cites. Percent changes are in typical home value (ZHVI), raw series.
Three ways to build the comparison
The memo compares each ZIP with itself a year earlier. A fair comparison asks what this ZIP's home values would have done without the announcement. That path is the counterfactual. It cannot be observed, so it is replaced by a stand-in built from a comparison group, ZIPs that the announcement did not reach. The stand-in is trustworthy only if it moved like the treated ZIP before the announcement. This section builds three stand-ins for each of the two treated ZIPs, Crystal City (22202, in Arlington, Virginia) and Long Island City (11101, in Queens), and grades each on that test.
All three use the ZIP's typical home value (Zillow's ZHVI, raw series) turned into an index equal to 100 times the natural log of the value, set to zero at its average over the twelve months from November 2017 to October 2018. The gap in a month is the ZIP's index minus the stand-in's index. Because 100 times a log difference is close to a percent difference, a gap of +3.7 means the ZIP sits about 3.7 percent above its stand-in. Table 3 defines the summaries used from here on.
| Name | Definition | The question it answers |
|---|---|---|
| Gap | The ZIP's index minus its stand-in's index in a given month | How far above or below its stand-in is the ZIP this month? |
| Early gap | Average gap over the three values dated 2018-11 to 2019-01, all before the cancellation | Did values jump when the news came? |
| Average gap | Average gap over the 16 values dated 2018-11 to 2020-02 | How far from its stand-in did the ZIP sit, on average, after the announcement? |
| Give-back | Average gap over 2019-03 to 2020-02 minus the early gap | Did a gap, once open, shrink after the cancellation? A negative value means it shrank or reversed. |
| Fit error | The typical monthly miss before the announcement: the root mean squared gap over the 58 months 2014-01 to 2018-10, in log points (close to percent) | How closely did the stand-in follow the ZIP before the news? |
Table 3. Quantities reported for each treated ZIP. The tables label them "early gap", "average gap", "give-back" and "fit error"; the design files and output tables call them E1, E2, E3 and pre-period RMSPE.
The first stand-in, the pool average, is the unweighted mean of every ZIP in the same metropolitan area with a complete series (304 ZIPs in Washington, 815 in New York). ZIPs the announcement could have reached are left out. These are the ring and co-treatment ZIPs described later and one low-confidence ZIP. The pool average is easy to explain and fits poorly. Before the announcement its typical monthly miss is 3.85 log points for 22202 and 6.54 for 11101.
The second stand-in, the nearest ten, averages only the ten pool ZIPs whose paths from 2014-01 to 2018-10 stayed closest to the treated ZIP's, as measured by the root mean squared difference between the two paths. The typical monthly miss before the announcement falls to 0.26 and 2.13.
The third stand-in, the synthetic control, blends 40 pool ZIPs (the donors, screened as the nearest to the treated ZIP) with weights that are never negative and add to one. The weights are chosen so that before the announcement the blend moves as closely as possible like the treated ZIP. Its typical monthly miss is 0.15 and 0.84.
For Crystal City the average gap is positive under all three stand-ins, between +3.7 and +4.2 percent. For Long Island City it is negative under all three but varies more, from -6.4 to -3.5 percent. A stand-in that did not track the ZIP before the announcement cannot be trusted after it, and the spread follows the fit error. Ranking each average gap against placebo ZIPs (last column of Table 4; the ranking is explained in the next section) puts Crystal City 7th or 8th of 305 under every stand-in and Long Island City between 135th and 170th of 816. The Washington result survives the choice of comparison, and the New York result is a null under all three.
Figure 1. Crystal City's average gap is +3.6 to +4.1 log points whichever comparison is used; Long Island City's runs from -6.6 to -3.6 and widens as the comparison fits worse
Show the data behind this chart
| Month | 22202: pool average | 22202: nearest ten | 22202: synthetic control | 11101: pool average | 11101: nearest ten | 11101: synthetic control |
|---|---|---|---|---|---|---|
| 2015-01 | +5.28 | +0.16 | +0.02 | -5.00 | +2.28 | +1.20 |
| 2016-01 | +3.04 | +0.48 | +0.03 | -0.28 | +1.77 | +0.07 |
| 2017-01 | +1.39 | +0.10 | +0.02 | +2.58 | +1.92 | -0.11 |
| 2018-01 | -0.04 | -0.14 | -0.01 | +0.70 | +0.09 | -0.01 |
| 2018-10 | -0.49 | +0.12 | -0.14 | -1.26 | -0.36 | -0.46 |
| 2018-11 | -0.25 | +0.34 | +0.04 | -0.40 | +0.77 | +0.49 |
| 2018-12 | +0.16 | +0.60 | +0.38 | +0.20 | +1.27 | +0.81 |
| 2019-01 | +1.35 | +1.41 | +1.15 | -2.14 | -0.94 | -0.92 |
| 2019-02 | +1.98 | +1.78 | +1.57 | -4.16 | -2.77 | -2.96 |
| 2019-03 | +2.72 | +2.18 | +2.06 | -4.47 | -2.36 | -2.78 |
| 2019-06 | +4.20 | +3.40 | +3.38 | -8.04 | -5.33 | -6.75 |
| 2019-12 | +6.56 | +5.97 | +6.03 | -9.03 | -6.65 | -3.71 |
| 2020-02 | +7.51 | +6.39 | +6.67 | -10.52 | -6.97 | -3.21 |
| ZIP | Comparison | Comparison ZIPs | Fit error before the announcement | Early gap | Average gap | Gap in 2020-02 | Rank of the average gap by size (p) |
|---|---|---|---|---|---|---|---|
| 22202 | Pool average | all 304 comparison ZIPs | 3.85 | +0.42 | +4.13 | +7.51 | 8 of 305 (0.026) |
| 22202 | Nearest ten ZIPs | 20191, 20833, 20895, 22044, 22181, 22205, 22207, 22209, 22304, 22308 | 0.26 | +0.78 | +3.68 | +6.39 | 7 of 305 (0.023) |
| 22202 | Synthetic control | 40 screened donors, weighted | 0.15 | +0.52 | +3.64 | +6.67 | 8 of 305 (0.026) |
| 11101 | Pool average | all 815 comparison ZIPs | 6.54 | -0.78 | -6.59 | -10.52 | 135 of 816 (0.165) |
| 11101 | Nearest ten ZIPs | 07030, 08701, 10009, 10026, 10035, 10044, 11228, 11366, 11411, 11423 | 2.13 | +0.37 | -4.18 | -6.97 | 170 of 816 (0.208) |
| 11101 | Synthetic control | 40 screened donors, weighted | 0.84 | +0.13 | -3.59 | -3.21 | 135 of 816 (0.165) |
Table 4. Three comparisons for each treated ZIP, raw ZHVI, index of 100 x log value with the twelve months 2017-11 to 2018-10 as base (log points). The early gap is the mean gap over 2018-11 to 2019-01 and the average gap the mean over 2018-11 to 2020-02. The rank is that of the treated ZIP's average gap, ignoring sign, among the treated ZIP and every comparison ZIP treated in turn as if it were the treated ZIP (p = rank / (N + 1)).
What the ZIP-level comparison shows
Synthetic controls for Crystal City and Long Island City
The main analysis gives each treated ZIP its own synthetic control. The blend draws on the 40 pool ZIPs nearest to the treated ZIP on its 2014-01 to 2018-10 path (58 months), taken from its metropolitan area's pool of ZIPs with complete data (304 in Washington, 815 in New York), and its weights are chosen to make the squared monthly gap over those months as small as possible. The fit error is 0.15 log points for 22202 and 0.84 for 11101. Neither is above the 90th percentile of the fit errors of its placebo ZIPs (it sits at percentile 16 for 22202 and 82 for 11101), so neither fit is flagged as poor.
Figure 2 shows each ZIP with its synthetic control. Crystal City and its stand-in coincide until 2018 and separate afterwards; Long Island City and its stand-in stay close until early 2019. Figure 3 plots the difference between the two lines and compares it with the gaps that placebo ZIPs produce.
Figure 2. Crystal City pulled away from its synthetic control during 2019; Long Island City did not show an announcement bump
Show the data behind this chart
| Month | 22202 | 22202: synthetic | 11101 | 11101: synthetic |
|---|---|---|---|---|
| 2015-01 | -8.7 | -8.9 | -22.3 | -24.0 |
| 2016-01 | -6.7 | -6.9 | -11.8 | -12.5 |
| 2017-01 | -5.5 | -5.7 | -5.5 | -5.9 |
| 2018-01 | -3.3 | -3.4 | -2.8 | -3.3 |
| 2018-10 | +0.0 | +0.0 | +0.0 | +0.0 |
| 2018-11 | -0.1 | -0.3 | +0.5 | -0.4 |
| 2018-12 | -0.0 | -0.5 | +0.8 | -0.4 |
| 2019-01 | +0.9 | -0.4 | -1.7 | -1.2 |
| 2019-02 | +1.8 | +0.1 | -3.7 | -1.2 |
| 2019-03 | +3.4 | +1.2 | -3.8 | -1.5 |
| 2019-06 | +7.6 | +4.0 | -5.1 | +1.3 |
| 2019-12 | +10.1 | +3.9 | -5.3 | -2.0 |
| 2020-02 | +11.6 | +4.8 | -6.9 | -4.0 |
Figure 3. Crystal City's gap rises above the range of placebo gaps in April 2019 and stays there; Long Island City's gap falls below a band 2.4 times as wide for 4 months in 2019 and ends inside it
Show the data behind this chart
| Month | 22202 gap | 22202 placebo 5th to 95th | 11101 gap | 11101 placebo 5th to 95th |
|---|---|---|---|---|
| 2015-01 | +0.02 | -0.7 to +0.7 | +1.20 | -1.1 to +0.9 |
| 2016-01 | +0.03 | -0.6 to +0.8 | +0.06 | -0.9 to +1.0 |
| 2017-01 | +0.02 | -0.7 to +0.6 | -0.12 | -1.1 to +0.9 |
| 2018-01 | -0.01 | -0.5 to +0.6 | -0.08 | -1.0 to +0.9 |
| 2018-10 | -0.14 | -0.9 to +0.8 | -0.57 | -1.3 to +1.6 |
| 2018-11 | +0.05 | -0.9 to +1.1 | +0.37 | -1.6 to +1.8 |
| 2018-12 | +0.39 | -1.2 to +1.1 | +0.68 | -1.9 to +2.2 |
| 2019-01 | +1.16 | -1.4 to +1.4 | -1.03 | -2.7 to +2.8 |
| 2019-02 | +1.58 | -1.8 to +1.8 | -3.08 | -3.2 to +3.6 |
| 2019-03 | +2.06 | -1.7 to +2.1 | -2.90 | -4.0 to +4.2 |
| 2019-06 | +3.39 | -1.8 to +2.5 | -6.89 | -4.9 to +5.5 |
| 2019-12 | +6.05 | -2.5 to +3.4 | -3.90 | -6.7 to +7.2 |
| 2020-02 | +6.69 | -2.7 to +3.4 | -3.49 | -7.1 to +7.4 |
Placebo tests: how large a gap chance produces
A gap of a few percent could still be chance, because home values wander and no blend tracks a ZIP perfectly. A placebo test measures how large a gap chance produces. Pretend that each of the other ZIPs in the metropolitan area had received the announcement, give each its own synthetic control, and compute the same gap. The 304 Washington placebo gaps (and 815 for New York) show what the method produces when nothing special happened. The treated ZIP is then ranked among them. A rank of 1 of 305 means the biggest of 305, and the p-value is the rank divided by the number of ZIPs, 0.003 in that case.
Two statistics can be ranked. The ratio test divides a ZIP's typical monthly miss after the announcement by its miss before; it is the main test in the design file, and the size test is an additional ranking. A ZIP that was tracked tightly before the announcement gets a large ratio from a modest gap. The size test ranks the average gap itself, ignoring sign, and is the stricter reading for a ZIP with a tight fit. Table 5 gives both.
Crystal City ranks 1 of 305 on the ratio test (p = 0.003; 0.004 after dropping placebo ZIPs whose fit error before the announcement is more than 5 times Crystal City's own). On the size test it ranks 10 of 305 (p = 0.033). Its early gap is not unusual on the size test, which fits a gap that builds slowly. Dropping any one donor with weight of at least 0.01 moves the average gap only between +3.6 and +4.1 percent.
Long Island City ranks 295 of 816 on the ratio test (p = 0.362), and the size-test p-values for its average gap and give-back are 0.152 and 0.092. Its give-back is -4.6 percent. After the cancellation the gap moved down, and the fall is inside the range of chance. None of the Long Island City p-values is below the usual 0.05 threshold.
| Statistic | 22202 | 11101 |
|---|---|---|
| Comparison pool (ZIPs); donor ZIPs in the blend | 304; 40 | 815; 40 |
| Fit error before the announcement, log points (percentile among placebo ZIPs) | 0.15 (16) | 0.84 (82) |
| Early gap, 2018-11 to 2019-01: log points (percent) | +0.53 (+0.53%) | +0.01 (+0.01%) |
| Average gap, 2018-11 to 2020-02: log points (percent) | +3.65 (+3.71%) | -3.75 (-3.68%) |
| Give-back (average gap 2019-03 to 2020-02 minus early gap): log points (percent) | +4.07 (+4.15%) | -4.76 (-4.65%) |
| Gap in 2019-02, 2019-05, 2019-11 and 2020-02, log points | +1.58, +3.12, +5.89, +6.69 | -3.08, -4.88, -4.43, -3.49 |
| Ratio of the post-announcement miss to the pre-announcement miss | 27.62 | 5.15 |
| Rank of that ratio among the treated ZIP and its placebo ZIPs; p | 1 of 305; 0.003 | 295 of 816; 0.362 |
| p when ranking by size: early gap, average gap, give-back | 0.377, 0.033, 0.013 | 0.994, 0.152, 0.092 |
| Placebo ZIPs, average gap: 5th to 95th percentile | -1.72 to +2.23 | -4.66 to +4.89 |
| Placebo ZIPs, early gap: 5th to 95th percentile | -1.08 to +1.11 | -1.88 to +2.14 |
| Range of the average gap when any one donor is dropped | +3.57 to +4.02 | -4.49 to -3.45 |
| Largest weight; donors with weight of at least 0.01; effective number of donors | 0.26; 10; 6.2 | 0.52; 7; 3.2 |
| Optimizer converged (share of placebo fits that converged) | yes (0.984) | yes (0.946) |
Table 5. ZIP-level comparison (synthetic control), raw all-homes ZHVI, base setup. Gaps are the treated ZIP minus its synthetic control in log points x100, with the percent equivalent in brackets. The early gap is the mean gap over 2018-11 to 2019-01 and the average gap the mean over 2018-11 to 2020-02. A p-value is the share of ZIPs (the treated ZIP plus its placebo ZIPs) whose statistic is at least as large as the treated ZIP's, that is rank divided by the number of ZIPs. The fit error is the typical monthly miss before the announcement (the root mean squared gap), and its percentile compares it with the placebo ZIPs' fit errors.
Which ZIPs make up each blend
Although 40 donors are available, the weights concentrate on a few. The largest single weight is 0.26 for 22202 and 0.52 for 11101. Counting each ZIP by the square of its weight, the blend is equivalent to 6.2 equally weighted ZIPs for 22202 and 3.2 for 11101 (the effective number of donors). Table 6 lists the largest weights.
| Rank | Donor ZIP for 22202 | Weight | Donor ZIP for 11101 | Weight |
|---|---|---|---|---|
| 1 | 22207 (Arlington) | 0.258 | 10044 (New York) | 0.518 |
| 2 | 20191 (Fairfax) | 0.222 | 11416 (Queens) | 0.136 |
| 3 | 20855 (Montgomery) | 0.136 | 10027 (New York) | 0.121 |
| 4 | 20015 (District of Columbia) | 0.107 | 11220 (Kings) | 0.095 |
| 5 | 22304 (Alexandria City) | 0.076 | 11211 (Kings) | 0.074 |
| 6 | 20815 (Montgomery) | 0.067 | 11205 (Kings) | 0.036 |
| 7 | 20833 (Montgomery) | 0.060 | 11379 (Queens) | 0.020 |
| 8 | 20861 (Montgomery) | 0.032 | 11207 (Kings) | 0.000 |
Table 6. The eight largest donor weights for each treated ZIP (ZIP-level comparison, raw series, base setup). The county is in brackets. A weight is the share of the synthetic control taken from that ZIP.
Checks on the ZIP-level result
Event study: a second method gives the same answer
The first check changes the method. An event study follows the treated ZIP month by month, from 2016-11 to 2020-02, and estimates how far it moved relative to the simple average of the comparison ZIPs in its metropolitan area, with October 2018 as zero. "Stacked" means each treated ZIP gets its own block of data (the ZIP plus the comparison ZIPs of its metropolitan area), with all dates counted from the announcement. Both treated ZIPs share one announcement date, so no ZIP that has already been treated serves as a comparison.
For Crystal City the average coefficient after the announcement is +4.7 percent (p = 0.003), the same sign and a similar size to the synthetic-control average gap. For Long Island City it is -5.2 percent (p = 0.148), with a give-back of -7.2 percent (p = 0.069). The coefficients for the months before the announcement have root mean square 1.32 (22202) and 2.80 (11101) log points, with placebo p-values of 0.452 and 0.496, so neither pre-announcement trend stands out.
The conventional standard-error intervals in the replication output are narrower than the placebo bands. With one treated ZIP per block they are not reliable, so the placebo bands are the measure of uncertainty here.
Figure 4. The event study repeats the pattern: the average coefficient after the announcement is +4.6 log points for Crystal City and -5.3 for Long Island City
Show the data behind this chart
| Month | DC coefficient | DC placebo 2.5th to 97.5th | NY coefficient | NY placebo 2.5th to 97.5th |
|---|---|---|---|---|
| 2016-11 | +1.89 | -6.8 to +5.3 | +2.26 | -15.2 to +10.5 |
| 2017-06 | +2.00 | -5.7 to +4.3 | +3.49 | -11.2 to +7.8 |
| 2018-01 | +0.45 | -2.9 to +2.9 | +1.96 | -6.6 to +4.8 |
| 2018-09 | +0.33 | -0.6 to +0.5 | +0.01 | -1.3 to +1.1 |
| 2018-10 | +0.00 | +0.0 to +0.0 | +0.00 | +0.0 to +0.0 |
| 2018-11 | +0.24 | -0.6 to +0.8 | +0.86 | -2.2 to +1.4 |
| 2018-12 | +0.65 | -0.9 to +1.0 | +1.46 | -2.6 to +2.3 |
| 2019-01 | +1.83 | -1.4 to +1.4 | -0.88 | -4.9 to +3.7 |
| 2019-02 | +2.47 | -2.0 to +1.7 | -2.91 | -7.1 to +5.6 |
| 2019-06 | +4.69 | -2.7 to +2.9 | -6.78 | -11.0 to +8.8 |
| 2019-12 | +7.05 | -4.0 to +4.1 | -7.77 | -12.1 to +12.2 |
| 2020-02 | +7.99 | -4.3 to +4.2 | -9.26 | -15.7 to +16.0 |
| Statistic | 22202 | 11101 |
|---|---|---|
| Early coefficient, average of 2018-11 to 2019-01 | +0.91 | +0.48 |
| Average coefficient, 2018-11 to 2020-02 | +4.61 | -5.33 |
| Give-back (average of 2019-03 to 2020-02 minus early coefficient) | +4.81 | -7.46 |
| p when ranking by size: early coefficient, average coefficient, give-back | 0.052, 0.003, 0.003 | 0.614, 0.148, 0.069 |
| Months before the announcement, 2016-11 to 2018-09: root mean square of the coefficients (largest absolute value) | 1.32 (2.10) | 2.80 (4.87) |
| Placebo p for that root mean square | 0.452 | 0.496 |
| Comparison ZIPs in the block | 304 | 815 |
Table 7. Event study, raw series. Coefficients are in log points x100 and measure the treated ZIP's movement relative to the average of its comparison ZIPs. A p-value ranks the coefficient by size among the placebo ZIPs plus the treated ZIP.
Metropolitan level: a null that cannot see a one-ZIP change
The unit of analysis is the thing being compared. In the ZIP-level comparison it is a ZIP. In the metropolitan-level comparison it is a whole metropolitan area. The typical home value of the Washington (or New York) metro is compared with the average of the other finalist metros, again from October 2018, using the same event-study and synthetic-control methods and the same placebo logic with metros in place of ZIPs. The raw comparison has 12 comparison metros because Los Angeles is missing a month; the smoothed comparison has 13.
Neither metro moves. The average event-study coefficients are -0.2 percent for Washington (p = 0.846) and -1.1 percent for New York (p = 0.462), and the synthetic-control gaps are -2.5 and -1.0 percent. Against 80 large non-finalist metros the New York coefficient is -1.3 percent (p = 0.346), the same picture. With only 12 comparison metros, the smallest p-value any placebo ranking can reach is 0.077, so this check cannot produce strong evidence in either direction.
The Washington result shows why the unit matters. The same metropolitan area contains a ZIP whose gap ranks 1 of 305, yet its metropolitan coefficient is indistinguishable from zero. The arithmetic is simple. If Crystal City sits 3.7 percent above where it would otherwise be and no other ZIP moves, the unweighted average of the 320 ZIPs in the Washington metropolitan area moves by 0.012 percent. That is the by-count dilution; the Zillow files give no housing units, so the share of the housing stock is not computed. A change in one neighborhood is invisible in a metropolitan average, and a metropolitan null is therefore weak evidence of no neighborhood effect.
Figure 5. At the metropolitan scale the average coefficients are -0.2 (Washington) and -1.1 (New York) log points and stay inside the placebo band
Show the data behind this chart
| Month | DC coefficient | DC placebo 2.5th to 97.5th | NY coefficient | NY placebo 2.5th to 97.5th |
|---|---|---|---|---|
| 2016-11 | +5.07 | -4.5 to +4.6 | +1.76 | -4.5 to +4.6 |
| 2017-06 | +3.61 | -3.9 to +3.7 | +0.67 | -3.9 to +3.7 |
| 2018-01 | +2.36 | -2.5 to +1.9 | +1.16 | -2.5 to +1.9 |
| 2018-09 | +0.07 | -0.6 to +0.6 | +0.02 | -0.6 to +0.6 |
| 2018-10 | +0.00 | +0.0 to +0.0 | +0.00 | +0.0 to +0.0 |
| 2018-11 | -0.05 | -0.6 to +0.5 | -0.07 | -0.6 to +0.5 |
| 2018-12 | -0.06 | -1.1 to +0.9 | -0.15 | -1.1 to +0.9 |
| 2019-01 | +0.08 | -1.5 to +1.1 | +0.01 | -1.5 to +1.1 |
| 2019-02 | +0.27 | -1.7 to +1.2 | -0.12 | -1.7 to +1.2 |
| 2019-06 | -0.12 | -1.7 to +2.0 | -1.53 | -1.7 to +2.0 |
| 2019-12 | -0.60 | -3.3 to +2.9 | -1.75 | -3.3 to +2.9 |
| 2020-02 | -0.32 | -4.0 to +2.6 | -2.00 | -4.0 to +2.6 |
| Treated metro | Comparison metros | Number | Event study: early coefficient | Event study: average coefficient | Event study: p (by size) | Event study: give-back | Event study: root mean square before announcement | Synthetic control: early gap | Synthetic control: average gap | Synthetic control: p (ratio) | Synthetic control: fit error |
|---|---|---|---|---|---|---|---|---|---|---|---|
| Washington | other finalist metros, raw | 12 | -0.01 | -0.21 | 0.846 | -0.29 | 3.27 | -2.33 | -2.49 | 0.615 | 1.30 |
| New York | other finalist metros, raw | 12 | -0.07 | -1.15 | 0.462 | -1.43 | 1.20 | +0.10 | -1.02 | 0.538 | 0.53 |
| Washington | other finalist metros, smoothed | 13 | -0.24 | -0.30 | 0.786 | -0.07 | 3.31 | -1.83 | -2.27 | 0.571 | 1.25 |
| New York | other finalist metros, smoothed | 13 | -0.35 | -0.71 | 0.500 | -0.47 | 1.58 | +0.16 | -0.33 | 0.571 | 0.32 |
| Washington | large non-finalist metros, raw | 80 | -0.05 | -0.37 | 0.827 | -0.44 | 3.42 | -0.58 | -0.79 | 0.259 | 0.20 |
| New York | large non-finalist metros, raw | 80 | -0.11 | -1.31 | 0.346 | -1.58 | 1.30 | +0.36 | -0.36 | 0.716 | 0.35 |
Table 8. Metropolitan-level comparison. For each treated metro: the event-study coefficients (early: 2018-11 to 2019-01; average: 2018-11 to 2020-02; give-back: the average of 2019-03 to 2020-02 minus the early coefficient), the placebo p-value of the average coefficient, and the synthetic-control gaps, each against the other finalist metros and against large non-finalist metros (S15). Gaps are in log points x100.
Rents: the fits are too poor to use
Rent (Zillow's observed rent index, ZORI) is a second outcome. All three rent fits are poor. The typical monthly miss before the announcement is 1.42, 1.69 and 2.65 log points for 22202, 11101 and 11109, each above the 90th percentile of the placebo ZIPs' misses. For comparison, the home-value fits miss by 0.15 and 0.84. For 22202 the average rent gap is -0.8 percent with a ratio p of 1.000.
The two Long Island City ZIPs show positive rent gaps (average gap +3.7 percent for 11101 and +3.3 for 11109) and positive give-backs (+2.3 and +4.0), so the rent gap does not reverse after the cancellation. The two p-values for 11101 disagree (0.386 on the ratio test against 0.018 by size) because a poor fit before the announcement inflates the denominator of the ratio. Anchoring on October 2018 alone instead of the twelve-month average moves the 11101 rent gap from +3.7 to +1.0 percent. ZIP 11109 has no home-value series, so rents are the only outcome available for the waterfront part of the site. ZORI is smoothed and not seasonally adjusted. The rent results carry little weight and should not be used to time anything.
Figure 6. Rent synthetic controls fit poorly: the typical monthly miss before the announcement is 1.42, 1.69 and 2.65 log points, each above the 90th percentile of placebo ZIPs
Show the data behind this chart
| Month | 22202 | 22202: synthetic | 11101 | 11101: synthetic | 11109 | 11109: synthetic |
|---|---|---|---|---|---|---|
| 2016-06 | -7.9 | -7.5 | -4.7 | -2.0 | -0.8 | -0.3 |
| 2017-01 | -9.2 | -7.6 | -7.5 | -3.7 | -5.2 | -3.0 |
| 2018-01 | -9.3 | -4.8 | -10.7 | -4.9 | -11.9 | -5.8 |
| 2018-10 | +0.0 | +0.0 | +0.0 | +0.0 | +0.0 | +0.0 |
| 2018-11 | -2.3 | +0.1 | -1.8 | -1.1 | -1.7 | -1.0 |
| 2019-01 | -3.0 | +0.5 | -3.5 | -1.6 | -3.1 | -2.1 |
| 2019-02 | -1.9 | +0.3 | -2.9 | -1.2 | -3.0 | -1.9 |
| 2019-06 | +0.2 | +3.1 | +3.0 | +1.5 | +6.6 | +1.7 |
| 2019-12 | +1.2 | +2.8 | +3.4 | +2.2 | +4.3 | +1.7 |
| 2020-02 | +2.0 | +3.1 | +3.3 | +2.4 | +4.3 | +1.4 |
| Statistic | 22202 | 11101 | 11109 |
|---|---|---|---|
| Comparison pool (ZIPs); donor ZIPs in the blend | 22; 22 | 56; 40 | 76; 28 |
| Months before and after the announcement in the data | 43, 16 | 46, 16 | 30, 16 |
| Fit error before the announcement, log points (percentile among placebo ZIPs) | 1.42 (95) | 1.69 (95) | 2.65 (100) |
| Poor fit (error above the placebo 90th percentile) | yes | yes | yes |
| Early gap, 2018-11 to 2019-01, log points | -1.80 | +1.93 | +0.29 |
| Average gap, 2018-11 to 2020-02, log points | -0.80 | +3.64 | +3.22 |
| Give-back, log points | +1.27 | +2.32 | +3.93 |
| Average gap with the base month set to 2018-10 only (S05), log points | -2.08 | +0.97 | +1.98 |
| Rank of the post-to-pre miss ratio among the treated ZIP and its placebo ZIPs; p | 23 of 23; 1.000 | 22 of 57; 0.386 | 66 of 77; 0.857 |
| p when ranking by size: early gap, average gap, give-back | 0.130, 0.478, 0.261 | 0.088, 0.018, 0.070 | 0.818, 0.039, 0.013 |
Table 9. Rent check, ZORI, base setup. All three fits are poor. Gaps are in log points x100.
Nearby ZIPs: Arlington's neighbors move with Crystal City
The announcement could have reached ZIPs next to the sites, so those ZIPs are kept out of every comparison pool and are estimated as treated ZIPs of their own, against the same placebo set as the treated ZIP of their metropolitan area. Ring ZIPs are the neighbors the announcement could plausibly have reached: 22201, 22204 and 22206 around Crystal City and 7 ZIPs in Queens and Brooklyn around Long Island City. Co-treatment ZIPs are two Alexandria ZIPs tied to Virginia's package through the Virginia Tech Innovation Campus: 22301, and 22305, which was named as the campus site only from 2019-06-10.
The Arlington ring ZIPs open gaps like Crystal City's. Their unweighted average gap is +2.5 percent, which is 68 percent of Crystal City's average gap, and ZIP 22301 shows +2.2 percent. A gap of this sign and size next to 22202 means that the 22202 result does not isolate the HQ2 footprint. It is consistent with a spillover from the announcement, a county-wide Arlington trend, or a comparison pool that does not match northern Virginia. The three Arlington ring ZIPs rank high on the ratio test (p = 0.003, 0.007 and 0.007) but not on the size test, where all three p-values are above 0.05, so the two rankings disagree for them.
The 7 Long Island City ring ZIPs do not move together. Their average gap is -1.0 percent, individual values run from negative to positive, and none has a size-test p-value below 0.10. ZIP 11104 has an early gap of +2.8 percent but an average gap of +1.1, so its early gap is not sustained. ZIP 22314 is a low-confidence ZIP that is kept out of the comparison pool and not estimated; returning it to the pool leaves the 22202 average gap unchanged within optimizer tolerance (see the sensitivity table).
Figure 7. The Arlington ring ZIPs rise alongside Crystal City: their average gap is +2.5 log points against +3.6 for 22202
Show the data behind this chart
| Month | 22202 (core) | 22201 | 22204 | 22206 | 22301 | 22305 |
|---|---|---|---|---|---|---|
| 2018-10 | -0.14 | +0.10 | -0.25 | -0.46 | -0.32 | +0.12 |
| 2018-11 | +0.05 | +0.14 | +0.17 | -0.27 | -0.06 | +0.61 |
| 2019-01 | +1.16 | +1.17 | +0.30 | +0.15 | +0.29 | +0.79 |
| 2019-02 | +1.58 | +1.32 | +0.03 | +0.21 | +0.64 | +0.42 |
| 2019-03 | +2.06 | +1.68 | +0.29 | +0.50 | +1.12 | +0.06 |
| 2019-06 | +3.39 | +2.61 | +1.31 | +2.25 | +1.67 | +0.26 |
| 2019-09 | +5.05 | +4.24 | +2.37 | +3.79 | +2.70 | +0.53 |
| 2019-12 | +6.05 | +4.74 | +4.07 | +5.38 | +4.08 | +1.36 |
| 2020-02 | +6.69 | +5.09 | +5.07 | +5.89 | +5.78 | +3.90 |
| Metro | ZIP | Role | Fit error | Early gap | Average gap | Give-back | p (ratio) | p (size of average gap) |
|---|---|---|---|---|---|---|---|---|
| Washington | 22201 | ring | 0.13 | +0.57 | +2.88 | +3.02 | 0.003 | 0.056 |
| Washington | 22204 | ring | 0.23 | +0.23 | +1.95 | +2.31 | 0.007 | 0.105 |
| Washington | 22206 | ring | 0.29 | -0.01 | +2.62 | +3.49 | 0.007 | 0.072 |
| Washington | 22301 | cotreat | 0.32 | +0.17 | +2.20 | +2.67 | 0.026 | 0.085 |
| Washington | 22305 | cotreat | 0.25 | +0.80 | +0.87 | +0.13 | 0.167 | 0.380 |
| Washington | Ring average | 3 ring ZIPs | +0.26 | +2.48 | +2.94 | |||
| New York | 11102 | ring | 0.58 | -0.08 | -4.51 | -5.73 | 0.100 | 0.112 |
| New York | 11103 | ring | 0.47 | -1.01 | -3.27 | -2.90 | 0.172 | 0.192 |
| New York | 11104 | ring | 0.93 | +2.76 | +1.06 | -2.27 | 0.689 | 0.608 |
| New York | 11105 | ring | 0.44 | -0.82 | +0.43 | +1.58 | 0.812 | 0.819 |
| New York | 11106 | ring | 1.46 (poor) | +0.59 | -1.32 | -2.47 | 0.938 | 0.536 |
| New York | 11222 | ring | 0.69 | +1.29 | +4.30 | +3.94 | 0.192 | 0.118 |
| New York | 11377 | ring | 0.85 | +0.48 | -3.95 | -5.55 | 0.336 | 0.138 |
| New York | Ring average | 7 ring ZIPs | +0.46 | -1.04 | -1.92 |
Table 10. Ring and co-treatment ZIPs, raw series, base setup, gaps in log points x100. Ring averages are unweighted means of the ZIP estimates and carry no pooled p-value. A fit error marked poor is above the placebo 90th percentile.
Sensitivity: the sign survives nearly every alternative setup
The design file lists alternative setups for the ZIP-level comparison: different donor sets (20, 80 or all pool ZIPs, or urban-core donors only), a different base month, a fit window that starts in 2016, shorter post-announcement windows, ring ZIPs allowed as donors, and different versions of the home-value series. Across the 11 setups that keep the raw series and the full post-announcement window, Crystal City's average gap runs from +3.1 percent (full pool) to +4.4 percent (pre-period from 2016-01). Long Island City's runs from -5.6 percent (shortlist month) to -3.0 percent (full pool).
Counting the changed series (condo homes, smoothed all-homes) and the shortened windows but not the placebo in time, the average gap keeps the sign of the base estimate in 15 of 15 setups for 22202 and 12 of 13 for 11101. The exception for 11101 is the condo series, whose average gap is +1.2 percent. The placebo in time uses a fake announcement date of 2016-11, when nothing happened, and gives fake average gaps of -0.4 percent for 22202 and -1.7 percent for 11101. Of the 141 estimation fits, 4 did not converge; they carry asterisks in the table and are listed in Table 16.
Figure 8. The average gap stays between +3.1 and +4.3 log points for Crystal City and between -5.8 and -3.1 for Long Island City across 11 alternative setups that keep the raw series and the full period
Show the data behind this chart
| Specification | 22202 | 11101 |
|---|---|---|
| Base specification | +3.65 | -3.75 |
| Treatment month moved to the shortlist, 2018-01 (S01) | +4.16 | -5.81 |
| Condo homes instead of all homes (S02) | +2.00 | +1.22 |
| Smoothed all-homes series (S03) | +2.77 | -3.50 |
| Donors from the urban core only (S04) | +3.73 | -3.75 |
| Base month is 2018-10 only (S05) | +3.75 | -3.50 |
| Base period is the last twelve months (S05b) | +3.64 | -3.59 |
| Fit window starts 2016-01 (S06) | +4.30 | -4.00 |
| Post-announcement window ends 2019-12 (S07a) | +3.23 | -3.76 |
| Post-announcement window ends 2019-02 (S07b) | +0.79 | |
| Blend drawn from 20 screened ZIPs (S08) | +3.99 | -5.01 |
| Blend drawn from 80 screened ZIPs, capped (S08) | +3.70 | -4.16 |
| Blend drawn from the full pool (S08) | +3.07 | -3.09 * |
| Ring and co-treatment ZIPs allowed as donors (S11) | +3.29 | -3.75 |
| ZIP 22314 allowed as a donor (S12) | +3.65 | |
| Placebo in time: fake treatment month 2016-11 (S10) | -0.42 | -1.75 |
| Specification | 22202: early gap / average gap (p, ratio) | 11101: early gap / average gap (p, ratio) |
|---|---|---|
| Base specification | +0.53 / +3.65 (p 0.003) | +0.01 / -3.75 (p 0.362) |
| Treatment month moved to the shortlist, 2018-01 (S01) | +0.79 / +4.16 (p 0.003) | -1.72 / -5.81 (p 0.504) |
| Condo homes instead of all homes (S02) | -0.62 / +2.00 (p 0.008) | +1.88 / +1.22 (p 0.854) |
| Smoothed all-homes series (S03) | +0.18 / +2.77 (p 0.003) | +0.41 / -3.50 (p 0.371) |
| Donors from the urban core only (S04) | +0.58 / +3.73 (p 0.007) | +0.01 / -3.75 (p 0.478) |
| Base month is 2018-10 only (S05) | +0.71 / +3.75 (p 0.003) | +0.55 / -3.50 (p 0.311) |
| Base period is the last twelve months (S05b) | +0.52 / +3.64 (p 0.003) | +0.13 / -3.59 (p 0.380) |
| Fit window starts 2016-01 (S06) | +0.64 / +4.30 (p 0.003) | +0.11 / -4.00 (p 0.377) |
| Post-announcement window ends 2019-12 (S07a) | +0.53 / +3.23 (p 0.003) | +0.01 / -3.76 (p 0.298) |
| Post-announcement window ends 2019-02 (S07b) | +0.53 / +0.79 (p 0.013) | n/a |
| Blend drawn from 20 screened ZIPs (S08) | +0.69 / +3.99 (p 0.003) | +0.00 / -5.01 (p 0.186) |
| Blend drawn from 80 screened ZIPs, capped (S08) | +0.55 / +3.70 (p 0.003) | -0.13 / -4.16 (p 0.314) |
| Blend drawn from the full pool (S08) | +0.46 / +3.07 (no placebos run) | +0.94 / -3.09 (no placebos run) * |
| Ring and co-treatment ZIPs allowed as donors (S11) | +0.47 / +3.29 (p 0.003) | +0.01 / -3.75 (p 0.361) |
| ZIP 22314 allowed as a donor (S12) | +0.53 / +3.65 (p 0.003) | n/a |
| Placebo in time: fake treatment month 2016-11 (S10) | -0.36 / -0.42 (p 0.639) | -0.92 / -1.75 (p 0.918) |
Table 11. ZIP-level comparison under alternative setups (raw series except where the label says otherwise). Cells show the early gap / average gap in log points with the ratio p-value in brackets; an asterisk marks a fit whose optimizer did not converge. For the placebo in time the cells are fake-window values. The code in each label is the setup's name in the design file.
Timing: Crystal City's gap builds after the announcement window
The Crystal City gap is not an announcement-window jump. The early gap is +0.5 percent, 15 percent of the average gap in log points, and the gap in 2020-02 is 4.2 times its value in 2019-02. Ending the post-announcement window at 2019-02 leaves 22 percent of the average gap (+0.8 percent; ratio p 0.013, size p 0.252). Most of the gap accumulates in a period that holds the County Board vote (2019-03-16), the Virginia Tech site move (2019-06-10), the Phase 1 site plan (2019-12-14) and the start of demolition (2020-01-22). The data do not say which of these, if any, moved values.
A second timing check moves the treatment month back to the January 2018 shortlist and measures the gap over 2018-01 to 2018-10, before the November 2018 reports. It is +0.0 percent for 22202 and -1.0 percent for 11101, so neither ZIP shows a positive run-up of meaningful size.
For Long Island City the raw gap is +0.4 percent in 2018-11 and +0.7 in 2018-12, then turns negative in the January 2019 value (-1.0) and reaches -3.0 in 2019-02. The 2019-02 value straddles the cancellation, which falls on day 14 of a 28-day month. Monthly data cannot separate a news channel (values fell before the cancellation) from an index-timing channel (Zillow's January estimate reflects later information).
Zillow's smoothed series is slower. Across 1,123 ZIPs, smoothed monthly changes line up best with raw changes 1 month later. For 11101 the smoothed all-homes gap is +0.1 percent in 2019-01 and -1.4 in 2019-02. The smoothed condo gap rises to +2.6 percent in 2019-01 and is +0.3 by 2020-02, which looks like a premium given back. It ranks 269 of 816 (ratio p 0.854), and its average gap ranges from -1.7 to +1.2 percent (-1.74 to +1.15 log points) when any one donor is dropped.
Figure 9. In Long Island City the raw gap turns negative in the month dated 2019-01, before the cancellation; the smoothed condo gap peaks at +2.6 log points that month and fades
Show the data behind this chart
| Month | Raw | Smoothed, all homes | Smoothed, condo |
|---|---|---|---|
| 2018-01 | -0.08 | -0.39 | -0.21 |
| 2018-06 | +0.08 | +0.00 | -0.33 |
| 2018-10 | -0.57 | +0.33 | +0.71 |
| 2018-11 | +0.37 | +0.48 | +1.16 |
| 2018-12 | +0.68 | +0.63 | +1.90 |
| 2019-01 | -1.03 | +0.12 | +2.57 |
| 2019-02 | -3.08 | -1.43 | +2.31 |
| 2019-03 | -2.90 | -2.95 | +2.26 |
| 2019-06 | -6.89 | -5.41 | +1.31 |
| 2019-12 | -3.90 | -3.95 | +0.51 |
| 2020-02 | -3.49 | -4.02 | +0.30 |
| ZIP | Early gap (2018-11 to 2019-01) | Gap 2019-02 | Gap 2019-05 | Gap 2019-11 | Gap 2020-02 | Average gap (to 2020-02) | Average gap if the window ends 2019-02 (S07b) | Average gap if the window ends 2019-12 (S07a) | Treatment month moved to 2018-01: mean gap 2018-01 to 2018-10 (S01) | Treatment month moved to 2018-01: average gap (S01) |
|---|---|---|---|---|---|---|---|---|---|---|
| Crystal City (22202) | +0.53 | +1.58 | +3.12 | +5.89 | +6.69 | +3.65 | +0.79 | +3.23 | +0.01 | +4.16 |
| Long Island City (11101) | +0.01 | -3.08 | -4.88 | -4.43 | -3.49 | -3.75 | n/a | -3.76 | -1.05 | -5.81 |
Table 12. Timing of the gaps, raw series, log points x100. S07b ends the post-announcement window at 2019-02, before the County Board vote. S01 moves the treatment month to 2018-01; its run-up window is 2018-01 to 2018-10.
A study of prices near the sites finds earlier and larger premia
Chen, Wilkoff and Yoshida [1] study the same episode with housing prices near the sites. Their abstract reports a premium of 4.9 percent near the Virginia headquarters months before the decision, and a premium for New York that reaches 17.5 percent before the decision and disappears on cancellation, with no significant effects in the other finalist cities. A premium here means higher prices than comparable homes elsewhere. These two figures are literature values taken from the abstract, not results of this analysis.
The ZIP-level numbers here differ. The mean gap over 2018-01 to 2018-10, before the November 2018 reports, is +0.0 percent for 22202 and -1.0 percent for 11101. The raw early gap for 11101 is +0.01 percent and the condo early gap is +1.9 percent, the largest positive announcement-window figure for 11101 in this analysis. The give-back for 11101 is -4.6 percent in the raw series and -0.9 in the condo series, with size-test p-values of 0.092 and 0.752. For 22202 the first three months average +0.5 percent. The ZIP-level data show no anticipation premium in either ZIP, and no raw announcement-window premium in Long Island City.
The two can differ for reasons that this analysis cannot sort out, and the following are possibilities, not findings. Prices near a site can respond while a ZIP-wide, model-based index barely moves. Of the 109 sampled points of the Long Island City lots, 69 fall in 11101 and 40 in 11109, which has no ZHVI, and ZHVI describes homes in the 35th to 65th percentile band of a whole ZIP.
Zillow's smoothed series spreads and lags a change by about 1 month, which can blur a short window. And the comparison groups differ, because here the comparison is a weighted set of ZIPs matched on the pre-announcement path, which can differ from the comparison in a study of prices near the sites. A question for discussion is which design could have detected a premium confined to nearby blocks.
A Crystal City-sized gain would not have stood out in the New York test
A test is informative only if it could have found what it is looking for. Statisticians call this power. A New York p-value of 0.362 would be evidence of no effect only if a real gain of Crystal City's size would have produced a small p-value. It would not have, because the placebo gaps that set the benchmark are much more spread out in New York than in Washington.
Before 2018-10 the width of the middle half of New York placebo gaps (the interquartile range) averages 0.59 log points; after it, 3.08, a factor of 5.2. Washington's goes from 0.38 to 1.34. In 2020-02 the middle 90 percent of New York placebo gaps spans -6.9 to +7.7 percent, 2.4 times the width of Washington's span. The monthly 11101 gap is below the placebo 5th percentile in 4 months (2019-05 to 2019-08) and is back inside the 90 percent band by 2020-02.
The size test makes the same point. The 95th percentile of the placebo average gap, ignoring sign, is 6.28 log points in New York and 2.94 in Washington. A gap the size of Crystal City's average gap (3.7 percent) would be matched or exceeded in size by 130 of 815 New York placebo ZIPs (16 percent) and by 9 of 304 Washington placebo ZIPs (3 percent). A gain of Crystal City's size would not have produced a significant result in Long Island City, so the p-value says little about whether one occurred.
Two further features weaken the New York test. The 11101 fit is looser than the Crystal City fit (fit error 0.84 against 0.15) and rests on 3.2 effective donors. And three-month changes in 11101 are noisy. Relative to the middle New York ZIP, their standard deviation before the announcement is 2.41 log points, at percentile 93 of the pool. The metropolitan-level check is weaker still, because a change confined to one ZIP is diluted in the average and 12 comparison metros cannot produce a p-value below 0.077.
Figure 10. The middle half of New York placebo gaps spans 3.1 log points after 2018-10, 2.3 times the width in Washington
Show the data behind this chart
| Month | New York | Washington |
|---|---|---|
| 2014-06 | 0.60 | 0.35 |
| 2016-01 | 0.61 | 0.38 |
| 2018-01 | 0.58 | 0.36 |
| 2018-10 | 0.76 | 0.44 |
| 2018-11 | 1.06 | 0.65 |
| 2019-01 | 1.58 | 0.92 |
| 2019-06 | 3.17 | 1.31 |
| 2019-12 | 4.20 | 1.91 |
| 2020-02 | 4.60 | 1.95 |
| Metro | Placebo ZIPs | Spread (interquartile range), average over months to 2018-10 | Spread, average over 2018-11 to 2020-02 | Spread in 2020-02 | Ratio, after to before | Placebo gap in 2020-02, 5th to 95th percentile | 95th percentile of the placebo average gap, ignoring sign | Placebo ZIPs with an average gap at least as large as 22202's, ignoring sign |
|---|---|---|---|---|---|---|---|---|
| New York | 815 | 0.59 | 3.08 | 4.60 | 5.2 | -7.1 to +7.4 | 6.28 | 130 (16.0%) |
| Washington | 304 | 0.38 | 1.34 | 1.95 | 3.6 | -2.7 to +3.4 | 2.94 | 9 (3.0%) |
Table 13. Spread of placebo gaps by metro, raw series, log points x100. The spread is the interquartile range (the width of the middle half) of the placebo gaps in a month, averaged over the months shown.
The five claims and the number that exposes each
Each of the five claims in the memo fails on a different point. Table 14 lists them; the subsections give the arithmetic.
| Claim | What the data show | Number |
|---|---|---|
| 1. The announcement was worth about 10 percent to Crystal City | The synthetic control and the median comparison ZIP rose over the same twelve months | Synthetic control +4.1 percent, median comparison ZIP +3.2 percent |
| 2. Long Island City would have earned the same premium | No gap in the announcement window in the raw series; the change from 2018-10 to 2019-10 is negative | Early gap +0.01 percent; ZIP change -4.7 percent |
| 3. The decline began after the cancellation | The raw series fell in the value dated before the cancellation; the smoothed condo series barely moved | Raw -2.5 percent, condo -0.2 percent, December 2018 to January 2019 |
| 4. The cancellation did not hurt New York | A metropolitan average cannot detect a change in one ZIP; Washington shows the same null beside a ZIP-level rank of 1 of 305 | New York -1.1 percent (p 0.462); Washington -0.2 percent (p 0.846) |
| 5. Long Island City follows the metropolitan area | 11101 grew about 1.8 times as fast as the metropolitan area and its changes move against it; since 2018-10 the two diverge | Correlation of 12-month changes -0.56; ZIP -6.6 percent, metro +2.4 percent |
Table 14. The memo's five claims and the number that exposes each. Percent changes are in typical home value from 2018-10 unless stated.
Claim 1: a before-and-after change with no comparison
The 10.0 percent compares Crystal City with itself a year earlier, from 2018-10 to 2019-10. Over the same twelve months its synthetic control rose 4.1 percent and the median of the 304 other Washington-area comparison ZIPs rose 3.2 percent, so comparable ZIPs were rising too. In log terms the synthetic control accounts for 42 percent of the ZIP's rise. What is left after the comparison is 5.7 percent against the synthetic control and 6.6 percent against the median ZIP.
That remainder is not the announcement alone. The ring ZIPs average +2.5 percent, and the County Board vote of 2019-03-16 and the Virginia Tech site move of 2019-06-10 fall inside the 2018-10 to 2019-10 window, which the claim that nothing else changed ignores. The base month is not the problem. The mean gap over 2018-01 to 2018-10 is +0.0 percent, so 22202 shows no run-up before the announcement.
Figure 11. The 10.0 percent rise in Crystal City from October 2018 to October 2019 is 4.1 percent for its synthetic control and 3.2 percent for the median Washington-metro ZIP
Show the data behind this chart
| Series | Change 2018-10 to 2019-10 (%) | Same change in log points | Middle half of donor ZIPs (%) |
|---|---|---|---|
| 22202 (Crystal City) | +10.0 | +9.55 | |
| Synthetic control for 22202 | +4.1 | +4.01 | |
| Median Washington-metro ZIP | +3.2 | +3.19 | +2.2 to +4.2 |
| 11101 (Long Island City) | -4.7 | -4.84 | |
| Synthetic control for 11101 | -0.3 | -0.31 | |
| Median New York-metro ZIP | +2.7 | +2.65 | -0.3 to +5.5 |
Claim 2: a premium assumed for Long Island City
The raw series for 11101 shows no announcement-window gain. The early gap is +0.01 percent against a placebo range of -1.9 to +2.2 percent (5th to 95th percentile), with a size-test p-value of 0.994. The ZIP stood +0.5 and +0.8 percent from 2018-10 in 2018-11 and 2018-12, against -0.3 and -0.6 percent for the median New York ZIP.
From 2018-10 to 2019-10 the ZIP changed by -4.7 percent against the +10.0 percent that the claimed premium implies, and its synthetic control changed by -0.3 percent. The claim is also inconsistent with itself. A premium of about 10 percent removed by the cancellation would return the ZIP to its October 2018 level, and the ZIP is 6.6 percent below it. The same terms of 25,000 jobs and $2.5 billion do not guarantee the same response. Crystal City's gap reached +1.6 percent three months in and +6.9 percent after fifteen, and the milestones behind that growth did not exist for Long Island City.
Claim 3: timing read from a smoothed series
The memo dates the loss to the cancellation using the smoothed condo index. The raw series for 11101 falls 2.5 percent from December 2018 to January 2019, in a value dated before the cancellation. The smoothed condo index falls 0.2 percent over the same month and the smoothed all-homes index 0.4 percent, because across 1,123 ZIPs the smoothed series follows raw changes 1 month late.
The raw fall of 2.50 log points is large against the one-month standard deviation of 0.64 log points for New York ZIPs relative to their metropolitan middle, so it is not noise alone. Monthly data cannot say whether it reflects news before 2019-02-14 or the way Zillow's estimates are timed. They also cannot separate the press reports of 2018-11-02 to 2018-11-05 from the official announcement of 2018-11-13, which all fall between the 2018-10-31 and 2018-11-30 observations.
Claim 4: the wrong unit of analysis
The memo's comparison is the New York metropolitan coefficient, -1.1 percent (p = 0.462). The same check gives -0.2 percent for Washington (p = 0.846), where the ZIP-level result for 22202 ranks 1 of 305. A metropolitan null cannot separate no change from a change in a small part of the metropolitan area. ZIP 11101 is one of 838 New York-metro ZIPs, 0.12 percent by count. With 12 comparison metros the smallest p-value the placebo ranking can give is 0.077.
The claim also answers a different question from the investor's. The parcel's neighbors are the ZIP and its 7 ring ZIPs, not the metropolitan area.
Claim 5: Long Island City does not follow the metropolitan area
Both rose from 2014-01 to 2018-10, ZIP 11101 by 45.3 percent and the metropolitan area by 25.3 percent, so 11101 grew 1.8 times as fast and ranks at percentile 91 of New York-metro ZIPs for growth over 2014-01 to 2017-12. Their changes do not move together. Over the months to 2018-10 the correlation of 12-month log changes between 11101 and the metropolitan area is -0.56 (46 overlapping windows, so far fewer independent observations) and the correlation of monthly changes is -0.20. Since 2018-10 the ZIP is -6.6 percent and the metropolitan area +2.4 percent.
The synthetic control matches 11101's own path, not the metropolitan average, and puts the average gap at -3.7 percent. The data do not show the ZIP following the metropolitan area, so they give no support to the claim that it will recover with it.
What an investor could and could not conclude
An investor could conclude that the typical home value in 11101 shows no announcement-window gain over its synthetic control (early gap +0.01 percent), and that the gap after January 2019 is negative (-3.7 percent on average) but no larger than gaps that placebo ZIPs show in 15 percent of cases (size p = 0.152; ratio p = 0.362). The smoothed condo series shows an early gain of +1.9 percent that fades, which fits a premium given back, and it also lies inside the placebo range (ratio p = 0.854).
An investor could not conclude that Long Island City gave back an HQ2 premium, and could not conclude that it did not. The test lacks the power to see a gain of Crystal City's size in a single New York ZIP. The placebo range for the early gap runs from -1.9 to +2.2 percent. The metropolitan check shows nothing at metropolitan scale, which is the expected result for a shock this size. Of the two treated ZIPs only Arlington's beats its placebos, and it cannot be separated from a wider Arlington trend or from later milestones.
A sell or hold decision on the parcel should not rest on these estimates alone. The home-value index describes existing homes, while the parcel's value depends on what can be built on it and at what cost; the estimates bear on the first, not the second. Evidence that would change the recommendation includes the parcel's own appraisals and comparable land trades, a transaction-level analysis of the ZIP and its ring, and rents or leasing activity for the site's neighbors.
What these data cannot show
ZHVI in small ZIPs is mostly model output. Zillow's size rank is 4,267 for 22202 and 1,464 for 11101, so these series reflect Zillow's model estimates for homes that did not sell, not observed sales. History is restated every month, so the values here are the 2026-10-01 estimate of 2014 to 2020, not what an investor saw at the time.
ZIP boundaries do not follow the sites. Of 109 sampled points in the announced Long Island City lots, 69 fall in 11101 and 40 in 11109, which has no ZHVI. The condo share of each ZIP is unknown. The all-homes value sits at position 0.55 between the condo and single-family values in 22202 and 0.07 in 11101, but that is a position between two values, not a share. The window ends with the U.S. national emergency of 2020-03-13 [2], so no estimate here says what the cancellation did to values after the pandemic began, or what Arlington's later construction did.
Evidence on individual homes near the sites is outside these data; the section on prices near the sites sets one study of such prices [1] next to the ZIP-level estimates. Greenstone, Hornbeck and Moretti compare counties that won a large plant with the runners-up for the same plant [3]; the analogue here would be the sites Amazon considered last, which the Zillow files do not identify. Abadie, Diamond and Hainmueller [4] and Abadie [5] set out the synthetic control method and its inference by placebo.
How the fits were audited
Each synthetic control is found by an optimizer (sequential least squares programming) that searches for the best weights, which must be non-negative and add to one. An optimizer can stop short of the best answer, so every fit records whether it converged. Across the Washington and New York fits, 501 of the 13,406 placebo fits (3.7%) did not converge. A fit that did not converge is reported and flagged, not replaced. The share is higher for condo series, which have fewer donors and thinner pools.
The two main fits, for 22202 and 11101, both converged. A second, independent solver (non-negative least squares) reproduces them, with a largest difference in the average gap of -2.3e-06 log points for 22202 and -1.4e-04 for 11101. The design (dates, comparison pools, outcome series, tests and the list of alternative setups) was written to a design file and fingerprinted before any gap was computed; the fingerprint and the other audit details are in the methods note in the appendix.
| Outcome series | Metro | Placebo fits | Share not converged |
|---|---|---|---|
| raw all-homes | New York | 815 | 5.4% |
| condo smoothed | New York | 314 | 16.9% |
| smoothed all-homes | New York | 805 | 3.4% |
| raw all-homes | Washington | 304 | 1.6% |
| condo smoothed | Washington | 129 | 10.9% |
| smoothed all-homes | Washington | 303 | 1.0% |
| ZORI | 6 placebo sets | 262 | 2.7% |
Table 15. Placebo fits whose optimizer did not converge, by series and metro (base setup). Fits that did not converge are kept and flagged, not replaced.
| Fit | Specification | Series | ZIP |
|---|---|---|---|
| ring ZIP | S04 | raw all-homes | 11377 |
| ZIP-level comparison | S08_full | raw all-homes | 11101 |
| rent check | S04 | ZORI | 11101 |
| rent check | S05b | ZORI | 11101 |
Table 16. Estimation fits, not placebo fits, whose optimizer did not converge. They carry asterisks in the sensitivity table.
References
- Chen, Y., Wilkoff, S. and Yoshida, J. (2024). Amazon is coming to town: sequential information revelation in the housing market. Real Estate Economics 52(2), 277-323. https://doi.org/10.1111/1540-6229.12457
- White House (2020-03-13). Proclamation on declaring a national emergency concerning the novel coronavirus disease (COVID-19) outbreak. https://trumpwhitehouse.archives.gov/presidential-actions/proclamation-declaring-national-emergency-concerning-novel-coronavirus-disease-covid-19-outbreak/
- Greenstone, M., Hornbeck, R. and Moretti, E. (2010). Identifying agglomeration spillovers: evidence from winners and losers of large plant openings. Journal of Political Economy 118(3), 536-598. https://doi.org/10.1086/653714
- Abadie, A., Diamond, A. and Hainmueller, J. (2010). Synthetic control methods for comparative case studies: estimating the effect of California's tobacco control program. Journal of the American Statistical Association 105(490), 493-505. https://www.jstor.org/stable/29747059
- Abadie, A. (2021). Using synthetic controls: feasibility, data requirements, and methodological aspects. Journal of Economic Literature 59(2), 391-425. https://www.aeaweb.org/articles?id=10.1257/jel.20191450
- Zillow Research. Housing data (ZHVI and ZORI downloads). https://www.zillow.com/research/data/
- Zillow Research. Zillow Home Value Index methodology, 2019 revision: getting under the hood. https://www.zillow.com/research/zhvi-methodology-2019-deep-26226/
- Zillow Research. Methodology: Zillow Observed Rent Index (ZORI). https://www.zillow.com/research/methodology-zori-repeat-rent-27092/
- Amazon (2017-09-07). Amazon opens search for Amazon HQ2, a second headquarters city in North America. Press release. https://press.aboutamazon.com/2017/9/amazon-opens-search-for-amazon-hq2-a-second-headquarters-city-in-north-america
- Amazon (2018-01-18). Amazon announces candidates for HQ2. Press release. https://press.aboutamazon.com/news-releases/news-release-details/amazon-announces-candidates-hq2/
- Amazon (2018-11-13). Amazon selects New York City and Northern Virginia for new headquarters. Company news. https://www.aboutamazon.com/news/company-news/amazon-selects-new-york-city-and-northern-virginia-for-new-headquarters
- Amazon (2019-02-14). Update on plans for New York City headquarters. Company news. https://www.aboutamazon.com/news/company-news/update-on-plans-for-new-york-city-headquarters