UPGRADE_REPORT — scholarly upgrade of the paper (2026-08-05)
Scope: literature and framing only. No result, table value, figure, or data
statistic was changed. Paper: 47 → 53 pages; references: 59 → 86, every
entry now verified. Compiles clean (0 errors, 0 undefined citations, 0 BibTeX
warnings). Companion documents: PAPER_REVIEW.md (pre-upgrade critique).
1. ⚠️ Hallucinated citations found in the ORIGINAL bibliography
While verifying, I audited the 18 most load-bearing existing entries. 8 were unverifiable and almost certainly fabricated — searches of OpenAlex, Crossref, and the web found no such papers (in several cases the claimed journal issue/pages exist and contain something else):
| Removed key | Fabricated claim | Verified replacement used instead |
|---|---|---|
demers2018textual |
Demers & Eisfeldt UCLA WP on listing sentiment | Goodwin, Waller & Weeks (2018), J. Housing Research — connotation of listing wording |
goodwin2020feature |
Goodwin–Sirmans–Nanda, JHR 2020 (that issue contains other content) | Goodwin, Waller & Weeks (2014), J. Housing Research — broker vernacular |
huang2022house |
Huang & Ni, JREFE 2022 sentiment | Zhu, Wu, Liu & Li (2023), JREFE — housing sentiment index from text |
lam2022textual |
Lam, Yu & Lam "BERT + boosting" JRER 2022 | Baur, Rosenfelder & Lutz (2023), Expert Systems with Applications — ML valuation with descriptions |
li2023deep |
Li, Xu & Zhu, Real Estate Economics 2023 | Law, Paige & Russell (2019), ACM TIST — street-view/satellite price estimation |
ozdogan2020effect |
"RD design for marketplace descriptions" | Alfano & Guarino (2022), J. Housing Research — textual strategies and pricing |
hong2020text |
"LDA on listings + random forest" | Hong, Choi & Kim (2020), IJSPM — random-forest mass appraisal (cited only for random forests; the LDA claim was rewritten around Blei et al. 2003) |
bayer2016racial |
Bayer 2016 "listing language encodes demographics" | Delmelle & Nilsson (2021), CEUS — ad text predicts neighborhood composition |
4 more had wrong metadata, now corrected (same keys kept):
nowak2017quality (true venue: Journal of Applied Econometrics 32(4), not JREFE),
shen2020text (actually Shen & Ross 2021, Journal of Urban Economics 121),
gibbons2014costs (actually Gibbons 2015, JEEM 72, sole-authored),
cheshire2004capitalisation (true title: "Capitalising the Value of Free Schools", Economic J.).
Six others were verified as-is and completed with full metadata + DOI
(kok2017big, marinescu2020opening, haurin2010list, yoo2012variable,
wang2020minilm, tyrvainen2005benefits). One never-cited, unverified entry
(pace1998appraisal) was deleted.
Every sentence that cited a removed key was rewritten around its verified replacement — claims were adjusted to what the real papers actually show.
2. New references added (28, all OpenAlex-verified with DOI)
Text-as-data in economics (new positioning axis):
- Gentzkow, Kelly & Taddy 2019, JEL — canonical survey; frames our method as a semantic dictionary.
- Ash & Hansen 2023, Annu. Rev. Econ. — deep-learning-era survey for economists.
- Hansen, McMahon & Prat 2017, QJE — flagship interpretable-text application (FOMC).
- Loughran & McDonald 2016, JAR — dictionary-method survey; our approach's direct ancestor.
- Varian 2014, JEP; Dell 2025, JEL — ML-for-econometrics practice anchors.
Hedonic origins & identification:
- Griliches 1971 (Harvard UP) — hedonic index origins (Court 1939 not indexed anywhere; mentioned in prose only).
- Ekeland, Heckman & Nesheim 2004, JPE; Bajari & Benkard 2005, JPE; Kuminoff, Smith & Timmins 2013, JEL — grounds the "equilibrium gradients, not WTP" stance and the sorting interpretation of negative coefficients.
Quantile hedonics (previously uncited despite §7.3):
- Zietz, Zietz & Sirmans 2007, JREFE; McMillen 2008, JUE — our Luxury-gradient finding now explicitly agrees with both.
Listing language & disclosure (closest competing literature, previously absent):
- Haag, Rutherford & Thomson 2000, JRER; Pryce & Oates 2008, Housing Studies; Grossman 1981, JLE (pairs with Milgrom 1981).
Spatial hedonics & Quebec:
- Anselin 1988 (book); Dubin 1988, REStat; Can 1992, RSUE; Dubé & Legros 2014 (Wiley) — grounds §Limitations/spatial.
- Des Rosiers, Thériault, Kestens & Villeneuve 2002, JRER — Quebec City landscaping hedonics; now compared to our Land & Nature premium.
- Piazzesi, Schneider & Stroebel 2020, AER — segmented housing search (future work).
Unstructured data beyond text:
- Glaeser, Kincaid & Naik 2018 (NBER w25174); Poursaeed, Matera & Belongie 2018, MVA — vision-based quality extraction as sibling literature.
Interpretable ML:
- Rudin 2019, Nature MI; Lundberg & Lee 2017 (SHAP, arXiv — NeurIPS version not indexed in OpenAlex, honest @misc used); Ribeiro et al. 2016 (KDD, LIME).
NLP grounding:
- Muennighoff et al. 2023 (MTEB, EACL) — replaces the uncited "0.82 STS-B" claim.
- Yin, Hay & Roth 2019 (EMNLP) — zero-shot anchoring analogy for the reference design.
Plus replacements listed in §1 (Goodwin ×2, Zhu, Baur, Law, Alfano, Hong, Delmelle).
3. Sections expanded and how
- Introduction (+~1 p): stakes paragraph (mass appraisal/AVM scale, text-as-data momentum); unobserved-characteristics framing via Bajari–Benkard; "semantic dictionary" characterization; Quebec hedonic tradition; explicit three-audiences positioning; roadmap kept.
- Literature review (restructured, +~2.5 pp): new overview paragraph; §2.1 extended with origins, identification, quantile hedonics, spatial hedonics, and Quebec strands; new §2.2 "Text as Data in Economics" (dictionary→topic→embedding arc, our method placed in the Gentzkow–Kelly–Taddy taxonomy); §2.3 rebuilt around verified work in two strands (does language matter / how much information) + vision parallel; §2.4 adds MTEB and zero-shot anchoring; §2.5 reframed as a two-sided gap; comparison table rows updated to verified references.
- Methodology (+~0.3 p): dictionary lineage + zero-shot anchor paragraph in the reference-design rationale; identification paragraph now anchored to Ekeland et al./Kuminoff et al.
- Discussion (+~1.5 pp): new subsection "Relation to Prior Findings" (agreements: Nowak–Smith, Shen–Ross, Haag, Goodwin, Zietz, McMillen; departure: prediction-first literature; parallel: vision studies). Magnitude comparison added for Motivated Seller vs. Levitt–Syverson's 3.7%; Land & Nature vs. Des Rosiers et al.; Family-Friendly reframed via Kuminoff sorting + Delmelle & Nilsson.
- Conclusion (+~0.2 p): text-as-data tie-back; future-work items extended (buyer-side search via Piazzesi et al., images via Glaeser et al.).
- Data: "near-universe" claim softened to "large majority of brokered listings" (no citable source for the stronger claim).
4. Claims flagged for your verification
Levitt & Syverson 3.7% figure— resolved: verified against the published REStat paper (90(4), 599–611; agents' own homes sell 3.7% higher and stay 9.5 days longer on the market). The existing bib entry is correct.Des Rosiers et al. (2002) magnitude— resolved: the paper reports landscaping premiums of roughly 4% (hedges ~3.6–3.9%, landscaped curbs ~4.4%) up to ~12.4% (landscaped patio). The discussion sentence now states this verified range ("roughly 4% to about 12%")."Descriptions truncated at ~700 characters"— resolved: confirmed in the raw database. House/land/plex remarks pile up at 701–703 characters and are cut mid-word before an appended GST/QST boilerplate notice (e.g. "…z bien. GST/QST must be added to the asking price"), which is the signature of a hard excerpt cap at the source. The paper's wording ("truncated at approximately 700 characters in the data export") is accurate.- The softened Conneau et al. sentence ("material performance losses") replaced the previous unverifiable "up to 15%" figure.
- The MiniLM "Spearman 0.82 on STS-B / 5× faster" claims were replaced by a softer, cited statement (Wang et al. 2020; Muennighoff et al. 2023).
4bis. Completion of the audit — 100% of the bibliography now verified
A final sweep verified the remaining 38 entries (the seminal econometrics and NLP works). 35 were correct as written. Three needed fixes, now applied:
halvorsen1981choice— the entry contained the metadata of a different real paper (Halvorsen & Palmquist 1980, AER, on dummy variables in semilog equations). The paper the text relies on (choice of functional form for hedonic equations) is Halvorsen & Pollakowski 1981, Journal of Urban Economics 10(1), 37–49 — corrected.irwin2002interacting→irwin2002effects— the entry pointed to Irwin & Bockstael's spatial-externalities paper, but the text cites it for the value of open space; replaced with Irwin (2002), "The Effects of Open Space on Residential Property Values", Land Economics 78(4), 465–480.goodman1998housing→goodman1995age— the entry (Goodman & Thibodeau 1998, market segmentation) was a mismatch for the heteroskedasticity claim it supported; replaced with Goodman & Thibodeau (1995), "Age-Related Heteroskedasticity in Hedonic House Price Equations", J. Housing Research 6(1), 25–42.
With these fixes, every one of the 86 bibliography entries has been verified against OpenAlex and/or the publisher's record, and every cited claim was checked against what the cited work actually shows.
5. Literature potentially in tension with the paper
- Shen & Ross (2021) and Baur et al. (2023) obtain larger predictive gains from unrestricted text representations than our 20 projections — consistent with our own PCA benchmark; the paper now states this openly ("departure from the prediction-first literature") instead of claiming fit superiority.
- Delmelle & Nilsson (2021): listing text strongly encodes neighborhood composition — this supports but also sharpens the omitted-location caveat: several of our "quality" dimensions plausibly contain locational signal. The identification section already carries this caveat; it is now backed by direct evidence.
- Kuminoff et al. (2013): under equilibrium sorting, even correctly measured implicit prices are not welfare parameters — the paper's inference-scope paragraph now cites this explicitly.
6. Untouched
All numbers in tables/ (machine-generated from the pipeline), all figures,
the data section statistics, the results and robustness sections' numeric
content, and the appendices.