{"status":"ok","message-type":"work","message-version":"1.0.0","message":{"indexed":{"date-parts":[[2026,6,30]],"date-time":"2026-06-30T06:32:18Z","timestamp":1782801138818,"version":"3.54.5"},"reference-count":93,"publisher":"Association for Computing Machinery (ACM)","issue":"2","license":[{"start":{"date-parts":[[2026,4,30]],"date-time":"2026-04-30T00:00:00Z","timestamp":1777507200000},"content-version":"vor","delay-in-days":0,"URL":"https:\/\/creativecommons.org\/licenses\/by\/4.0\/legalcode"}],"content-domain":{"domain":["dl.acm.org"],"crossmark-restriction":true},"short-container-title":["Proc. ACM Hum.-Comput. Interact."],"published-print":{"date-parts":[[2026,4,30]]},"abstract":"<jats:p>People are increasingly using Large Language Models (LLMs) for a sense of \u201clocalness,\u201d yet their ability to accurately and equitably represent local knowledge remains unexamined. To investigate this, we conducted a large-scale evaluation using a benchmark of over 12,000 question-answer pairs spanning structured census data, local news, and social media. Our results show that performance is strongly shaped by data modality: structured tasks expose deep limitations in numerical reasoning and calibration, while open-ended prompts reveal a clear performance hierarchy favoring informal user-generated content over professionally edited prose. Our primary finding is the existence of deep, context-dependent disparities that affect communities differently. We uncover a dual geographic bias: in formal news contexts, models exhibit a strong \u201curban advantage,\u201d leaving rural areas systematically underrepresented with lower semantic depth. Conversely, in social media data, models suffer an \u201curban penalty,\u201d struggling to navigate the conversational complexity and slang of high-density areas. This indicates that while rural locales face a \u201cpoverty of data,\u201d highly documented urban centers face a \u201cpoverty of precision.\u201d We also identify a domain bias: models are more adept at handling concrete, physical questions but consistently struggle to capture the nuanced relational and cognitive dimensions of a community. This work provides the first systematic audit of localness disparities in LLMs, revealing how they reflect and risk amplifying real-world inequities. Achieving equitable local representation requires moving beyond passive evaluation to active intervention. We call for a concerted effort from the CSCW community to build richer and more ethical datasets, design interfaces that prioritize user verification over blind trust, and architect AI systems for deeper and more just engagement with place.<\/jats:p>","DOI":"10.1145\/3788058","type":"journal-article","created":{"date-parts":[[2026,5,20]],"date-time":"2026-05-20T15:29:14Z","timestamp":1779290954000},"page":"1-45","update-policy":"https:\/\/doi.org\/10.1145\/crossmark-policy","source":"Crossref","is-referenced-by-count":2,"title":["Is Your Chatbot a Tourist or a Townie? Quantifying Geographic and Localness Disparities in LLM Representations of Place CSCW022"],"prefix":"10.1145","volume":"10","author":[{"ORCID":"https:\/\/orcid.org\/0009-0004-5207-5511","authenticated-orcid":false,"given":"Zihan","family":"Gao","sequence":"first","affiliation":[{"name":"Information School","place":["Madison, USA"]},{"name":"University of Wisconsin-Madison","place":["Madison, USA"]}],"role":[{"vocabulary":"crossref","role":"author"}]},{"ORCID":"https:\/\/orcid.org\/0000-0003-1569-4466","authenticated-orcid":false,"given":"Jacob","family":"Thebault-Spieker","sequence":"additional","affiliation":[{"name":"Information School","place":["Madison, USA"]},{"name":"University of Wisconsin - Madison","place":["Madison, USA"]}],"role":[{"vocabulary":"crossref","role":"author"}]}],"member":"320","published-online":{"date-parts":[[2026,5,20]]},"reference":[{"key":"e_1_3_3_2_2","unstructured":"Penelope\u00a0Muse Abernathy. 2018. The expanding news desert. https:\/\/www.usnewsdeserts.com\/reports\/expanding-news-desert\/. Accessed: 2025-11-24."},{"key":"e_1_3_3_3_2","unstructured":"Thales\u00a0Sales Almeida Giovana\u00a0Kerche Bon\u00e1s Jo\u00e3o Guilherme\u00a0Alves Santos Hugo Abonizio and Rodrigo Nogueira. 2025. TiEBe: Tracking Language Model Recall of Notable Worldwide Events Through Time. arxiv:https:\/\/arXiv.org\/abs\/2501.07482\u00a0[cs.CL] https:\/\/arxiv.org\/abs\/2501.07482"},{"key":"e_1_3_3_4_2","doi-asserted-by":"publisher","unstructured":"Marianne Aubin Le\u00a0Qu\u00e9r\u00e9 Mor Naaman and Jenna Fields. 2024. Not Quite Filling the Void: Comparing the Perceptions of Local Online Groups and Local Media Pages on Facebook. Proc. ACM Hum.-Comput. Interact. 8 CSCW1 Article 100 (April 2024) 22\u00a0pages. 10.1145\/3637377","DOI":"10.1145\/3637377"},{"key":"e_1_3_3_5_2","volume-title":"Reddit News Users More Likely to Be Male, Young and Digital in Their News Preferences","author":"Barthel Michael","year":"2016","unstructured":"Michael Barthel, Galen Stocking, Jesse Holcomb, and Amy Mitchell. 2016. Reddit News Users More Likely to Be Male, Young and Digital in Their News Preferences. Technical Report. Pew Research Center, Journalism & Media. https:\/\/www.pewresearch.org\/journalism\/2016\/02\/25\/reddit-news-users-more-likely-to-be-male-young-and-digital-in-their-news-preferences\/ Accessed: 2025-11-22."},{"key":"e_1_3_3_6_2","doi-asserted-by":"publisher","unstructured":"Cillian Berragan Alex Singleton Alessia Calafiore and Jeremy Morley. 2024. Mapping Great Britain\u2019s semantic footprints through a large language model analysis of Reddit comments. Computers Environment and Urban Systems 110 (2024) 102121. 10.1016\/j.compenvurbsys.2024.102121","DOI":"10.1016\/j.compenvurbsys.2024.102121"},{"key":"e_1_3_3_7_2","volume-title":"First Conference on Language Modeling","author":"Blakeney Cody","year":"2024","unstructured":"Cody Blakeney, Mansheej Paul, Brett\u00a0W. Larsen, Sean Owen, and Jonathan Frankle. 2024. Does your data spark joy? Performance gains from domain upsampling at the end of training. In First Conference on Language Modeling. https:\/\/openreview.net\/forum?id=vwIIAot0ff"},{"key":"e_1_3_3_8_2","first-page":"1877","volume-title":"Advances in Neural Information Processing Systems","author":"Brown Tom","year":"2020","unstructured":"Tom Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared\u00a0D Kaplan, Prafulla Dhariwal, Arvind Neelakantan, Pranav Shyam, Girish Sastry, Amanda Askell, Sandhini Agarwal, Ariel Herbert-Voss, Gretchen Krueger, Tom Henighan, Rewon Child, Aditya Ramesh, Daniel Ziegler, Jeffrey Wu, Clemens Winter, Chris Hesse, Mark Chen, Eric Sigler, Mateusz Litwin, Scott Gray, Benjamin Chess, Jack Clark, Christopher Berner, Sam McCandlish, Alec Radford, Ilya Sutskever, and Dario Amodei. 2020. Language Models are Few-Shot Learners. In Advances in Neural Information Processing Systems, H.\u00a0Larochelle, M.\u00a0Ranzato, R.\u00a0Hadsell, M.F. Balcan, and H.\u00a0Lin (Eds.), Vol.\u00a033. Curran Associates, Inc., 1877\u20131901. https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2020\/file\/1457c0d6bfcb4967418bfb8ac142f64a-Paper.pdf"},{"key":"e_1_3_3_9_2","doi-asserted-by":"publisher","DOI":"10.1145\/3025453.3025627"},{"key":"e_1_3_3_10_2","unstructured":"Aili Chen Xuyang Ge Ziquan Fu Yanghua Xiao and Jiangjie Chen. 2024. TravelAgent: An AI Assistant for Personalized Travel Planning. arxiv:https:\/\/arXiv.org\/abs\/2409.08069\u00a0[cs.AI] https:\/\/arxiv.org\/abs\/2409.08069"},{"key":"e_1_3_3_11_2","doi-asserted-by":"publisher","DOI":"10.1145\/3025453.3025495"},{"key":"e_1_3_3_12_2","doi-asserted-by":"publisher","DOI":"10.1145\/2858036.2858573"},{"key":"e_1_3_3_13_2","doi-asserted-by":"publisher","DOI":"10.1609\/icwsm.v6i1.14278"},{"key":"e_1_3_3_14_2","volume-title":"The Atlas of AI: Power, Politics, and the Planetary Costs of Artificial Intelligence","author":"Crawford Kate","year":"2021","unstructured":"Kate Crawford. 2021. The Atlas of AI: Power, Politics, and the Planetary Costs of Artificial Intelligence. Yale University Press."},{"key":"e_1_3_3_15_2","doi-asserted-by":"publisher","DOI":"10.7551\/mitpress\/9780262015554.001.0001"},{"key":"e_1_3_3_16_2","doi-asserted-by":"publisher","DOI":"10.1145\/3708359.3712111"},{"key":"e_1_3_3_17_2","doi-asserted-by":"publisher","unstructured":"Sebastian Farquhar Jannik Kossen Lorenz Kuhn and Yarin Gal. 2024. Detecting hallucinations in large language models using semantic entropy. Nature 630 8017 (2024) 625\u2013630. 10.1038\/s41586-024-07421-0","DOI":"10.1038\/s41586-024-07421-0"},{"key":"e_1_3_3_18_2","doi-asserted-by":"publisher","unstructured":"Casey Fiesler Michael Zimmer Nicholas Proferes Sarah Gilbert and Naiyan Jones. 2024. Remember the Human: A Systematic Review of Ethical Considerations in Reddit Research. Proc. ACM Hum.-Comput. Interact. 8 GROUP Article 5 (Feb. 2024) 33\u00a0pages. 10.1145\/3633070","DOI":"10.1145\/3633070"},{"key":"e_1_3_3_19_2","doi-asserted-by":"publisher","unstructured":"A\u00a0Stewart Fotheringham and David\u00a0WS Wong. 1991. The Modifiable Areal Unit Problem in Multivariate Statistical Analysis. Environment and Planning A: Economy and Space 23 7 (1991) 1025\u20131044. arXiv:10.1068\/a23102510.1068\/a231025","DOI":"10.1068\/a231025"},{"key":"e_1_3_3_20_2","doi-asserted-by":"publisher","unstructured":"Devin Gaffney and J.\u00a0Nathan Matias. 2018. Caveat Emptor Computational Social Science: Large-scale Missing Data in a Widely-published Reddit Corpus. PLOS ONE 13 7 (07 2018) 1\u201313. 10.1371\/journal.pone.0200162","DOI":"10.1371\/journal.pone.0200162"},{"key":"e_1_3_3_21_2","doi-asserted-by":"publisher","unstructured":"Zihan Gao Justin Cranshaw and Jacob Thebault-Spieker. 2024. Journeying Through Sense of Place with Mental Maps: Characterizing Changing Spatial Understanding and Sense of Place During Migration for Work. Proc. ACM Hum.-Comput. Interact. 8 CSCW2 Article 503 (Nov. 2024) 31\u00a0pages. 10.1145\/3687042","DOI":"10.1145\/3687042"},{"key":"e_1_3_3_22_2","unstructured":"Zihan Gao Justin Cranshaw and Jacob Thebault-Spieker. 2025. A Turing Test for \u201dLocalness\u201d: Conceptualizing Defining and Recognizing Localness in People and Machines. arxiv:https:\/\/arXiv.org\/abs\/2505.07282\u00a0[cs.HC] https:\/\/arxiv.org\/abs\/2505.07282"},{"key":"e_1_3_3_23_2","doi-asserted-by":"publisher","DOI":"10.1145\/3715070.3749277"},{"key":"e_1_3_3_24_2","doi-asserted-by":"crossref","unstructured":"Zihan Gao Yifei Xu and Jacob Thebault-Spieker. 2026. LocalBench: Benchmarking LLMs on County-Level Local Knowledge and Reasoning. Proceedings of the AAAI Conference on Artificial Intelligence (2026).","DOI":"10.1609\/aaai.v40i45.41190"},{"key":"e_1_3_3_25_2","volume-title":"NeurIPS 2025 Workshop on Algorithmic Collective Action","author":"Gao Zihan","year":"2025","unstructured":"Zihan Gao, Mohsin Y.\u00a0K. Yousufi, and Jacob Thebault-Spieker. 2025. Collective Narrative Grounding: Community-Coordinated Data Contributions to Improve Local AI Systems. In NeurIPS 2025 Workshop on Algorithmic Collective Action. https:\/\/openreview.net\/forum?id=2ZRwKlGSDa"},{"key":"e_1_3_3_26_2","doi-asserted-by":"publisher","DOI":"10.1145\/2858036.2858094"},{"key":"e_1_3_3_27_2","doi-asserted-by":"publisher","unstructured":"Timnit Gebru Jamie Morgenstern Briana Vecchione Jennifer\u00a0Wortman Vaughan Hanna Wallach Hal\u00a0Daum\u00e9 III and Kate Crawford. 2021. Datasheets for datasets. Commun. ACM 64 12 (Nov. 2021) 86\u201392. 10.1145\/3458723","DOI":"10.1145\/3458723"},{"key":"e_1_3_3_28_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2024.naacl-long.366"},{"key":"e_1_3_3_29_2","doi-asserted-by":"publisher","unstructured":"Thomas\u00a0F. Gieryn. 2000. A Space for Place in Sociology. Annual Review of Sociology 26 Volume 26 2000 (2000) 463\u2013496. 10.1146\/annurev.soc.26.1.463","DOI":"10.1146\/annurev.soc.26.1.463"},{"key":"e_1_3_3_30_2","doi-asserted-by":"publisher","unstructured":"Martyna Gliniecka. 2023. The Ethics of Publicly Available Data Research: A Situated Ethics Framework for Reddit. Social Media + Society 9 3 (2023) 20563051231192021. arXiv:10.1177\/2056305123119202110.1177\/20563051231192021","DOI":"10.1177\/20563051231192021"},{"key":"e_1_3_3_31_2","doi-asserted-by":"publisher","unstructured":"Sarah\u00a0E. Gollust Laura\u00a0M. Baum Jeff Niederdeppe Colleen\u00a0L. Barry and Erika\u00a0Franklin Fowler. 2017. Local Television News Coverage of the Affordable Care Act: Emphasizing Politics Over Consumer Information. American Journal of Public Health 107 5 (2017) 687\u2013693. arXiv:10.2105\/AJPH.2017.30365910.2105\/AJPH.2017.303659 PMID: 28207336.","DOI":"10.2105\/AJPH.2017.303659"},{"key":"e_1_3_3_32_2","doi-asserted-by":"publisher","DOI":"10.2307\/j.ctv272452n"},{"key":"e_1_3_3_33_2","unstructured":"Atharva Gundawar Mudit Verma Lin Guan Karthik Valmeekam Siddhant Bhambri and Subbarao Kambhampati. 2024. Robust Planning with LLM-Modulo Framework: Case Study in Travel Planning. arxiv:https:\/\/arXiv.org\/abs\/2405.20625\u00a0[cs.AI] https:\/\/arxiv.org\/abs\/2405.20625"},{"key":"e_1_3_3_34_2","doi-asserted-by":"publisher","unstructured":"Jose\u00a0A. Guridi Cristobal Cheyre and Qian Yang. 2025. Thoughtful Adoption of NLP for Civic Participation: Understanding Differences Among Policymakers. Proc. ACM Hum.-Comput. Interact. 9 2 Article CSCW193 (May 2025) 27\u00a0pages. 10.1145\/3711091","DOI":"10.1145\/3711091"},{"key":"e_1_3_3_35_2","volume-title":"The Twelfth International Conference on Learning Representations","author":"Gurnee Wes","year":"2024","unstructured":"Wes Gurnee and Max Tegmark. 2024. Language Models Represent Space and Time. In The Twelfth International Conference on Learning Representations. https:\/\/openreview.net\/forum?id=jE8xbmvFin"},{"key":"e_1_3_3_36_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2020.acl-main.740"},{"key":"e_1_3_3_37_2","doi-asserted-by":"publisher","unstructured":"Keith Hampton and Barry Wellman. 2003. Neighboring in Netville: How the Internet Supports Community and Social Capital in a Wired Suburb. City & Community 2 4 (2003) 277\u2013311. arXiv:https:\/\/onlinelibrary.wiley.com\/doi\/pdf\/10.1046\/j.1535-6841.2003.00057.x10.1046\/j.1535-6841.2003.00057.x","DOI":"10.1046\/j.1535-6841.2003.00057.x"},{"key":"e_1_3_3_38_2","doi-asserted-by":"publisher","DOI":"10.1145\/3301019.3323906"},{"key":"e_1_3_3_39_2","doi-asserted-by":"publisher","DOI":"10.1145\/240080.240193"},{"key":"e_1_3_3_40_2","doi-asserted-by":"publisher","unstructured":"Brent Hecht and Monica Stephens. 2014. A Tale of Cities: Urban Biases in Volunteered Geographic Information. Proceedings of the International AAAI Conference on Web and Social Media 8 1 (May 2014) 197\u2013205. 10.1609\/icwsm.v8i1.14554","DOI":"10.1609\/icwsm.v8i1.14554"},{"key":"e_1_3_3_41_2","doi-asserted-by":"publisher","DOI":"10.1145\/1718918.1718962"},{"key":"e_1_3_3_42_2","first-page":"29217","volume-title":"Advances in Neural Information Processing Systems","author":"Henderson Peter","year":"2022","unstructured":"Peter Henderson, Mark Krass, Lucia Zheng, Neel Guha, Christopher\u00a0D Manning, Dan Jurafsky, and Daniel Ho. 2022. Pile of Law: Learning Responsible Data Filtering from the Law and a 256GB Open-Source Legal Dataset. In Advances in Neural Information Processing Systems, S.\u00a0Koyejo, S.\u00a0Mohamed, A.\u00a0Agarwal, D.\u00a0Belgrave, K.\u00a0Cho, and A.\u00a0Oh (Eds.), Vol.\u00a035. Curran Associates, Inc., 29217\u201329234. https:\/\/proceedings.neurips.cc\/paper_files\/paper\/2022\/file\/bc218a0c656e49d4b086975a9c785f47-Paper-Datasets_and_Benchmarks.pdf"},{"key":"e_1_3_3_43_2","doi-asserted-by":"publisher","unstructured":"Benjamin\u00a0D. Horne Maur\u00edcio Gruppi Kenneth Joseph Jon Green John\u00a0P. Wihbey and Sibel Adal\u0131. 2022. NELA-Local: A Dataset of U.S. Local News Articles for the Study of County-Level News Ecosystems. Proceedings of the International AAAI Conference on Web and Social Media 16 1 (May 2022) 1275\u20131284. 10.1609\/icwsm.v16i1.19379","DOI":"10.1609\/icwsm.v16i1.19379"},{"key":"e_1_3_3_44_2","doi-asserted-by":"publisher","unstructured":"Lei Huang Weijiang Yu Weitao Ma Weihong Zhong Zhangyin Feng Haotian Wang Qianglong Chen Weihua Peng Xiaocheng Feng Bing Qin and Ting Liu. 2025. A Survey on Hallucination in Large Language Models: Principles Taxonomy Challenges and Open Questions. ACM Trans. Inf. Syst. 43 2 Article 42 (Jan. 2025) 55\u00a0pages. 10.1145\/3703155","DOI":"10.1145\/3703155"},{"key":"e_1_3_3_45_2","doi-asserted-by":"publisher","unstructured":"Kee\u00a0Moon Jang Junda Chen Yuhao Kang Junghwan Kim Jinhyung Lee Fabio Duarte and Carlo Ratti. 2024. Place identity: a generative AI\u2019s perspective. Humanities and Social Sciences Communications 11 1 (2024) 1156. 10.1057\/s41599-024-03645-7","DOI":"10.1057\/s41599-024-03645-7"},{"key":"e_1_3_3_46_2","doi-asserted-by":"publisher","unstructured":"Kee\u00a0Moon Jang and Youngchul Kim. 2019. Crowd-sourced cognitive mapping: A new way of displaying people\u2019s cognitive perception of urban space. PLOS ONE 14 6 (06 2019) 1\u201318. 10.1371\/journal.pone.0218590","DOI":"10.1371\/journal.pone.0218590"},{"key":"e_1_3_3_47_2","doi-asserted-by":"publisher","unstructured":"Andrew Jenkins Arie Croitoru Andrew\u00a0T. Crooks and Anthony Stefanidis. 2016. Crowdsourcing a Collective Sense of Place. PLOS ONE 11 4 (04 2016) 1\u201320. 10.1371\/journal.pone.0152932","DOI":"10.1371\/journal.pone.0152932"},{"key":"e_1_3_3_48_2","doi-asserted-by":"publisher","DOI":"10.1145\/2858036.2858123"},{"key":"e_1_3_3_49_2","doi-asserted-by":"publisher","DOI":"10.1145\/2858036.2858122"},{"key":"e_1_3_3_50_2","doi-asserted-by":"publisher","unstructured":"Bradley\u00a0S. Jorgensen and Richard\u00a0C. Stedman. 2001. Sense of Place as An Attitude: Lakehouse Owners Attitudes Toward Their Properties. Journal of Environmental Psychology 21 3 (2001) 233\u2013248. 10.1006\/jevp.2001.0226","DOI":"10.1006\/jevp.2001.0226"},{"key":"e_1_3_3_51_2","doi-asserted-by":"publisher","DOI":"10.1145\/3173574.3173839"},{"key":"e_1_3_3_52_2","unstructured":"Gali Katz Hai Sitton Guy Gonen and Yohay Kaplan. 2025. Beyond the Surface: Uncovering Implicit Locations with LLMs for Personalized Local News. arxiv:https:\/\/arXiv.org\/abs\/2502.14660\u00a0[cs.LG] https:\/\/arxiv.org\/abs\/2502.14660"},{"key":"e_1_3_3_53_2","doi-asserted-by":"publisher","DOI":"10.1145\/3630106.3658941"},{"key":"e_1_3_3_54_2","doi-asserted-by":"publisher","DOI":"10.1145\/3706598.3714020"},{"key":"e_1_3_3_55_2","volume-title":"The Thirteenth International Conference on Learning Representations","author":"Leng Jixuan","year":"2025","unstructured":"Jixuan Leng, Chengsong Huang, Banghua Zhu, and Jiaxin Huang. 2025. Taming Overconfidence in LLMs: Reward Calibration in RLHF. In The Thirteenth International Conference on Learning Representations. https:\/\/openreview.net\/forum?id=l0tg0jzsdL"},{"key":"e_1_3_3_56_2","doi-asserted-by":"publisher","unstructured":"Charis Lengen and Thomas Kistemann. 2012. Sense of place and place identity: Review of neuroscientific evidence. Health & Place 18 5 (2012) 1162\u20131171. 10.1016\/j.healthplace.2012.01.012","DOI":"10.1016\/j.healthplace.2012.01.012"},{"key":"e_1_3_3_57_2","doi-asserted-by":"publisher","unstructured":"Laura Lentini and Fran\u00e7oise Decortis. 2010. Space and places: when interacting with and in physical space becomes a meaningful experience. Personal and Ubiquitous Computing 14 5 (2010) 407\u2013415. 10.1007\/s00779-009-0267-y","DOI":"10.1007\/s00779-009-0267-y"},{"key":"e_1_3_3_58_2","unstructured":"Q.\u00a0Vera Liao and Jennifer\u00a0Wortman Vaughan. 2023. AI Transparency in the Age of LLMs: A Human-Centered Research Roadmap. arxiv:https:\/\/arXiv.org\/abs\/2306.01941\u00a0[cs.HC] https:\/\/arxiv.org\/abs\/2306.01941"},{"key":"e_1_3_3_59_2","volume-title":"The Thirteenth International Conference on Learning Representations","author":"Liu Hongfu","year":"2025","unstructured":"Hongfu Liu, Hengguan Huang, Xiangming Gu, Hao Wang, and Ye Wang. 2025. On Calibration of LLM-based Guard Models for Reliable Content Moderation. In The Thirteenth International Conference on Learning Representations. https:\/\/openreview.net\/forum?id=wUbum0nd9N"},{"key":"e_1_3_3_60_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2024.naacl-long.26"},{"key":"e_1_3_3_61_2","unstructured":"Shayne Longpre Robert Mahari Anthony Chen Naana Obeng-Marnu Damien Sileo William Brannon Niklas Muennighoff Nathan Khazam Jad Kabbara Kartik Perisetla Xinyi Wu Enrico Shippole Kurt Bollacker Tongshuang Wu Luis Villa Sandy Pentland and Sara Hooker. 2023. The Data Provenance Initiative: A Large Scale Audit of Dataset Licensing & Attribution in AI. http:\/\/arxiv.org\/abs\/2310.16787 arXiv:https:\/\/arXiv.org\/abs\/2310.16787 [cs]."},{"key":"e_1_3_3_62_2","unstructured":"Matt MacVey. 2022. AI & Local News newsletter issue 12. NYU Tandon School of Engineering NYC Media Lab. https:\/\/engineering.nyu.edu\/news\/ai-local-news-newsletter-issue-12 Accessed: 2025-07-16."},{"key":"e_1_3_3_63_2","doi-asserted-by":"publisher","unstructured":"Momin Malik Hemank Lamba Constantine Nakos and J\u00fcrgen Pfeffer. 2021. Population Bias in Geotagged Tweets. Proceedings of the International AAAI Conference on Web and Social Media 9 4 (Aug. 2021) 18\u201327. 10.1609\/icwsm.v9i4.14688","DOI":"10.1609\/icwsm.v9i4.14688"},{"key":"e_1_3_3_64_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2023.acl-long.546"},{"key":"e_1_3_3_65_2","series-title":"(ICML\u201924)","volume-title":"Proceedings of the 41st International Conference on Machine Learning","author":"Manvi Rohin","year":"2024","unstructured":"Rohin Manvi, Samar Khanna, Marshall Burke, David Lobell, and Stefano Ermon. 2024. Large Language Models are Geographically Biased. In Proceedings of the 41st International Conference on Machine Learning (Vienna, Austria) (ICML\u201924). JMLR.org, Article 1409, 16\u00a0pages."},{"key":"e_1_3_3_66_2","volume-title":"ICLR","author":"Manvi Rohin","year":"2024","unstructured":"Rohin Manvi, Samar Khanna, Gengchen Mai, Marshall Burke, David Lobell, and Stefano Ermon. 2024. GeoLLM: Extracting Geospatial Knowledge from Large Language Models. In ICLR. https:\/\/openreview.net\/forum?id=TqL2xBwXP3"},{"key":"e_1_3_3_67_2","doi-asserted-by":"publisher","unstructured":"Ninareh Mehrabi Fred Morstatter Nripsuta Saxena Kristina Lerman and Aram Galstyan. 2021. A Survey on Bias and Fairness in Machine Learning. ACM Comput. Surv. 54 6 Article 115 (July 2021) 35\u00a0pages. 10.1145\/3457607","DOI":"10.1145\/3457607"},{"key":"e_1_3_3_68_2","doi-asserted-by":"publisher","DOI":"10.1145\/3630106.3658967"},{"key":"e_1_3_3_69_2","doi-asserted-by":"publisher","unstructured":"Stein Monteiro. 2024. Searching for Settlement Information on Reddit. International Migration 62 3 (2024) 100\u2013119. arXiv:https:\/\/onlinelibrary.wiley.com\/doi\/pdf\/10.1111\/imig.1326110.1111\/imig.13261","DOI":"10.1111\/imig.13261"},{"key":"e_1_3_3_70_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2025.emnlp-main.626"},{"key":"e_1_3_3_71_2","doi-asserted-by":"publisher","unstructured":"Helen Nissenbaum. 2011. A Contextual Approach to Privacy Online. Daedalus 140 4 (2011) 32\u201348. 10.1162\/DAED_a_00113","DOI":"10.1162\/DAED_a_00113"},{"key":"e_1_3_3_72_2","doi-asserted-by":"crossref","DOI":"10.18574\/nyu\/9781479833641.001.0001","volume-title":"Algorithms of oppression","author":"Noble Safiya\u00a0Umoja","year":"2018","unstructured":"Safiya\u00a0Umoja Noble. 2018. Algorithms of Oppression: How Search Engines Reinforce Racism. In Algorithms of oppression. New York university press."},{"key":"e_1_3_3_73_2","unstructured":"Will Orr and Kate Crawford. 2024. Building Better Datasets: Seven Recommendations for Responsible Design from Dataset Creators. arxiv:https:\/\/arXiv.org\/abs\/2409.00252\u00a0[cs.LG] https:\/\/arxiv.org\/abs\/2409.00252"},{"key":"e_1_3_3_74_2","doi-asserted-by":"publisher","DOI":"10.1007\/978-3-031-88708-6_1"},{"key":"e_1_3_3_75_2","doi-asserted-by":"publisher","DOI":"10.1145\/2628363.2628407"},{"key":"e_1_3_3_76_2","doi-asserted-by":"publisher","unstructured":"Nicholas Proferes Naiyan Jones Sarah Gilbert Casey Fiesler and Michael Zimmer. 2021. Studying Reddit: A Systematic Overview of Disciplines Approaches Methods and Ethics. Social Media + Society 7 2 (2021) 20563051211019004. arXiv:10.1177\/2056305121101900410.1177\/20563051211019004","DOI":"10.1177\/20563051211019004"},{"key":"e_1_3_3_77_2","doi-asserted-by":"publisher","unstructured":"Yao Qu and Jue Wang. 2024. Performance and Biases of Large Language Models in Public Opinion Simulation. Humanities and Social Sciences Communications 11 1 (2024) 1095. 10.1057\/s41599-024-03609-x","DOI":"10.1057\/s41599-024-03609-x"},{"key":"e_1_3_3_78_2","volume-title":"Place and placelessness","author":"Relph Edward","year":"1976","unstructured":"Edward Relph. 1976. Place and placelessness. Vol.\u00a067. Pion London."},{"key":"e_1_3_3_79_2","unstructured":"J Riley and H Cowart. 2021. The Reddit Oasis: Analyzing the potential role of location-based subreddits in the alleviation of news deserts. Community Journalism 9 1 (2021)."},{"key":"e_1_3_3_80_2","doi-asserted-by":"publisher","unstructured":"Emily Sun. 2021. The Importance of Play in Digital Placemaking. Proceedings of the International AAAI Conference on Web and Social Media 9 2 (Aug. 2021) 23\u201325. 10.1609\/icwsm.v9i2.14680","DOI":"10.1609\/icwsm.v9i2.14680"},{"key":"e_1_3_3_81_2","doi-asserted-by":"publisher","DOI":"10.1145\/3173574.3174155"},{"key":"e_1_3_3_82_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2024.naacl-long.18"},{"key":"e_1_3_3_83_2","doi-asserted-by":"publisher","DOI":"10.1145\/3630106.3658992"},{"key":"e_1_3_3_84_2","doi-asserted-by":"publisher","DOI":"10.18653\/v1\/2024.emnlp-industry.104"},{"key":"e_1_3_3_85_2","doi-asserted-by":"publisher","DOI":"10.1145\/2702123.2702558"},{"key":"e_1_3_3_86_2","doi-asserted-by":"publisher","DOI":"10.1145\/3173574.3173722"},{"key":"e_1_3_3_87_2","doi-asserted-by":"publisher","DOI":"10.1145\/3148330.3148350"},{"key":"e_1_3_3_88_2","doi-asserted-by":"publisher","unstructured":"Emily Tseng Rosanna Bellini Yeuk-Yu Lee Alana Ramjit Thomas Ristenpart and Nicola Dell. 2024. Data Stewardship in Clinical Computer Security: Balancing Benefit and Burden in Participatory Systems. Proc. ACM Hum.-Comput. Interact. 8 CSCW1 Article 39 (April 2024) 29\u00a0pages. 10.1145\/3637316","DOI":"10.1145\/3637316"},{"key":"e_1_3_3_89_2","volume-title":"Space and place: The perspective of experience","author":"Tuan Yi-Fu","year":"1977","unstructured":"Yi-Fu Tuan. 1977. Space and place: The perspective of experience. U of Minnesota Press."},{"key":"e_1_3_3_90_2","doi-asserted-by":"publisher","DOI":"10.1145\/2598510.2598523"},{"key":"e_1_3_3_91_2","volume-title":"Forty-second International Conference on Machine Learning","author":"Xu Yifei","year":"2025","unstructured":"Yifei Xu, Tusher Chakraborty, Emre Kiciman, Bibek Aryal, Srinagesh Sharma, Songwu Lu, and Ranveer Chandra. 2025. RLTHF: Targeted Human Feedback for LLM Alignment. In Forty-second International Conference on Machine Learning. https:\/\/openreview.net\/forum?id=ATUfSZayVo"},{"key":"e_1_3_3_92_2","unstructured":"Susan Zhang Stephen Roller Naman Goyal Mikel Artetxe Moya Chen Shuohui Chen Christopher Dewan Mona Diab Xian Li Xi\u00a0Victoria Lin Todor Mihaylov Myle Ott Sam Shleifer Kurt Shuster Daniel Simig Punit\u00a0Singh Koura Anjali Sridhar Tianlu Wang and Luke Zettlemoyer. 2022. OPT: Open Pre-trained Transformer Language Models. arxiv:https:\/\/arXiv.org\/abs\/2205.01068\u00a0[cs.CL] https:\/\/arxiv.org\/abs\/2205.01068"},{"key":"e_1_3_3_93_2","series-title":"Proceedings of Machine Learning Research","first-page":"12697","volume-title":"Proceedings of the 38th International Conference on Machine Learning","author":"Zhao Zihao","year":"2021","unstructured":"Zihao Zhao, Eric Wallace, Shi Feng, Dan Klein, and Sameer Singh. 2021. Calibrate Before Use: Improving Few-shot Performance of Language Models. In Proceedings of the 38th International Conference on Machine Learning(Proceedings of Machine Learning Research, Vol.\u00a0139), Marina Meila and Tong Zhang (Eds.). PMLR, 12697\u201312706. https:\/\/proceedings.mlr.press\/v139\/zhao21c.html"},{"key":"e_1_3_3_94_2","doi-asserted-by":"publisher","DOI":"10.63317\/2ptonpvxx2s2"}],"container-title":["Proceedings of the ACM on Human-Computer Interaction"],"original-title":[],"language":"en","link":[{"URL":"https:\/\/dl.acm.org\/doi\/pdf\/10.1145\/3788058","content-type":"unspecified","content-version":"vor","intended-application":"similarity-checking"}],"deposited":{"date-parts":[[2026,5,20]],"date-time":"2026-05-20T16:53:50Z","timestamp":1779296030000},"score":1,"resource":{"primary":{"URL":"https:\/\/dl.acm.org\/doi\/10.1145\/3788058"}},"subtitle":[],"short-title":[],"issued":{"date-parts":[[2026,4,30]]},"references-count":93,"journal-issue":{"issue":"2","published-print":{"date-parts":[[2026,4,30]]}},"alternative-id":["10.1145\/3788058"],"URL":"https:\/\/doi.org\/10.1145\/3788058","relation":{},"ISSN":["2573-0142"],"issn-type":[{"value":"2573-0142","type":"electronic"}],"subject":[],"published":{"date-parts":[[2026,4,30]]},"assertion":[{"value":"2026-05-20","order":3,"name":"published","label":"Published","group":{"name":"publication_history","label":"Publication History"}}]}}