Friday, September 11, 2026
Science
No Result
View All Result
  • Login
  • HOME
  • SCIENCE NEWS
  • CONTACT US
  • HOME
  • SCIENCE NEWS
  • CONTACT US
No Result
View All Result
Scienmag
No Result
View All Result
Home Science News Social Science

Multimodal LLMs show geographic and perceptual bias across 200 cities

September 11, 2026
in Social Science
Courtney Benton
By Courtney Benton Scienmag Editorial Profile - Science and Technology Policy
Reading Time: 5 mins read
0
Multimodal LLMs show geographic and perceptual bias across 200 cities

Multimodal LLMs show geographic and perceptual bias across 200 cities

65
SHARES
587
VIEWS
Share on FacebookShare on Twitter
ADVERTISEMENT

The rapid global spread of multimodal large language models has raised a deceptively simple question: when these systems look at a photograph of a city street, do they see it the same way regardless of where in the world that street happens to be? A new study published in npj Urban Sustainability suggests that the answer is a resounding no. Researchers led by H. Kim, X. Li and M. Quintana have assembled one of the most geographically comprehensive test beds to date for probing the visual and perceptual capacities of multimodal artificial intelligence, drawing on imagery from more than 200 cities across every inhabited continent. Their findings reveal systematic geographic and perceptual biases that mirror, and in some cases amplify, longstanding inequities in the digital data on which these models are trained.

The research team’s central concern was not merely whether multimodal large language models, or MLLMs, can identify a landmark or read a street sign. Rather, they set out to determine whether the perceptual judgments these systems make about urban environments — assessments of safety, walkability, beauty, wealth and order — are applied consistently across cities in different countries, income brackets and cultural contexts. Such perceptual scores increasingly matter in the real world. Urban planners, technology companies and civic institutions have begun experimenting with AI-assisted tools that evaluate neighborhoods, rank streetscapes and inform investment decisions. If the underlying models carry hidden geographic biases, those biases could quietly shape policy and perception at scale.

To investigate, the authors curated a global dataset spanning more than 200 cities, deliberately reaching well beyond the heavily documented urban cores of North America and Western Europe. The dataset incorporates street-level and urban imagery paired with perception labels, allowing the researchers to compare how state-of-the-art multimodal models rate visual attributes of city environments against established human benchmarks derived from perception studies. The design enabled a kind of controlled stress test: hold the perceptual question constant, vary the geography, and observe where model performance holds and where it fractures.

The results were striking. The team documented consistent disparities in how accurately the models perceived urban attributes depending on where a city is located. In many regions of the world — particularly in parts of Africa, Latin America and South and Southeast Asia — the models’ perceptual judgments diverged more sharply from human ground truth than they did for cities in North America, Europe and East Asia’s wealthiest urban centers. The pattern constitutes what the researchers describe as a geographic bias: a systematic degradation of perceptual reliability that is correlated with location rather than with any inherent property of the places themselves.

Beneath that headline pattern, the study unpacked a more granular form of distortion the authors characterize as perceptual bias. Even when a model could roughly identify what it was looking at, its judgments along specific perceptual dimensions were skewed. Some attributes were systematically overestimated in certain regions and underestimated in others. A street in one city might be rated as cleaner, safer or more affluent than an objectively comparable street elsewhere, purely because of the visual and cultural priors embedded in the model’s training data. These are not random errors; they are directional errors, and directional errors at scale become a form of algorithmic stereotyping.

The technical roots of the problem trace back to how multimodal models are built. Contemporary MLLMs combine a vision encoder, typically a transformer-based architecture trained on enormous collections of web-scraped images, with a language model that interprets and verbalizes the visual representations. Both components inherit the statistical fingerprints of their training corpora. Web imagery is not uniformly distributed across the globe: images from wealthier, English-speaking and heavily digitized countries are dramatically overrepresented, while many cities in the Global South appear only sparsely. The result is a vision system that has effectively “seen” far more of some parts of the world than others, and a language layer that has absorbed uneven cultural narratives about what different places look like and mean.

The new study’s contribution is to quantify these effects with global scope and methodological rigor. By correlating model error with indicators such as geographic region, economic development and data availability, the authors were able to dissociate several candidate explanations. They found that the biases are not simply a function of image quality or resolution, nor can they be dismissed as artifacts of labeling noise. Instead, the evidence points to deep-seated representational gaps: for underrepresented cities, the models appear to substitute learned priors — generalized impressions of what a “developing” or “non-Western” urban environment supposedly looks like — for accurate perception of the actual scene.

The implications extend well beyond the laboratory. Multimodal AI is increasingly positioned as an engine for automated assessment of the built world, from satellite- and street-level analytics platforms to smart-city dashboards that promise real-time monitoring of infrastructure and quality of life. If a model systematically underestimates the safety or aesthetic quality of streets in Lagos, Jakarta or Lima relative to comparable streets in Amsterdam or Toronto, any downstream application built on those scores inherits the distortion. Real-estate analytics, insurance pricing, tourism ranking and urban planning support tools could all, without anyone intending it, channel attention and resources toward places the model already favors, reinforcing existing cycles of advantage and neglect.

The study also speaks to a broader tension in contemporary AI governance. Efforts to audit large models for bias have concentrated heavily on demographic attributes such as race and gender, often measured within Western contexts. Geographic bias is more diffuse and harder to operationalize, because the relevant variable — where a place is — entangles everything from lighting conditions and vegetation to signage, architecture and cultural conventions of visual order. The global dataset assembled by Kim, Li, Quintana and colleagues offers a template for making this variable tractable, and demonstrates that perceptual fairness is a measurable property, not merely an intuition.

What can be done? The authors point toward several complementary remedies. Training and fine-tuning data can be rebalanced to include far denser coverage of underrepresented regions, ideally collected with local participation and local contextual grounding. Evaluation practices should become geographically stratified, with performance reported not only as a global average but as a breakdown across regions and city types, so that a model cannot pass an audit on the strength of high scores in familiar territory alone. And downstream users should treat perceptual outputs for unfamiliar geographies with explicit caution, since the models’ confidence in these regions is not matched by their accuracy.

There is also a subtler lesson embedded in the findings. Human perception of urban environments is itself culturally situated — people from different backgrounds rate the same street differently — and any perception benchmark encodes some cultural standpoint. The challenge for AI developers is not to eliminate standpoint altogether, an impossible goal, but to make the standpoints plural and the uncertainty visible. A model that says “this street feels unsafe, but I am far less certain for cities in this region” is a fundamentally more trustworthy instrument than one that projects confident, homogeneous judgments across an unevenly understood planet.

As multimodal AI systems continue their rapid diffusion into everyday tools and professional workflows, the study stands as a timely warning wrapped in a rigorous empirical package. More than 200 cities’ worth of evidence shows that the artificial eyes now increasingly used to interpret the urban world do not see all of it equally. Making them see more fairly is not a niche technical concern; it is a prerequisite for any serious claim that AI can help build smarter, more equitable cities.

Subject of Research: Geographic and perceptual bias in multimodal large language models (MLLMs) when evaluating urban imagery from more than 200 cities worldwide, revealing systematic regional disparities in perceptual judgments.

Subject of Research: Social Science

Article Title: Geographic and perceptual bias in multimodal LLMs: evidence from a global dataset spanning more than 200 cities

Article References: Kim, H., Li, X., Quintana, M., Biljecki, F., & Lee, S. (2026). Geographic and perceptual bias in multimodal LLMs: evidence from a global dataset spanning more than 200 cities. npj Urban Sustainability. https://doi.org/10.1038/s42949-026-00466-2

Image Credits: AI Generated

DOI: 10.1038/s42949-026-00466-2

Keywords: multimodal large language models, geographic bias, perceptual bias, urban sustainability, global cities, street-level imagery, AI fairness, algorithmic stereotyping, urban analytics, training data bias, smart cities, AI auditing

Cite Scienmag News

Courtney Benton. (September 11, 2026). Multimodal LLMs show geographic and perceptual bias across 200 cities. Scienmag. https://scienmag.com/multimodal-llms-show-geographic-and-perceptual-bias-across-200-cities/

Courtney Benton. "Multimodal LLMs show geographic and perceptual bias across 200 cities." Scienmag, 11 September 2026, https://scienmag.com/multimodal-llms-show-geographic-and-perceptual-bias-across-200-cities/. Accessed 11 September 2026.

Courtney Benton. "Multimodal LLMs show geographic and perceptual bias across 200 cities." Scienmag. September 11, 2026. https://scienmag.com/multimodal-llms-show-geographic-and-perceptual-bias-across-200-cities/

Tags: AI perception of beauty and order in citiesAI-based urban beauty and wealth perceptionbiases in multimodal AI understanding of urban environmentscity safety and walkability assessment by AIcity street image recognition biasescross-cultural AI bias in urban imagerycross-cultural perception in artificial intelligencedigital data inequities in AI trainingdigital data inequities in multimodal modelsgeographic disparities in AI urban analysisgeographic disparities in AI urban imageryglobal city image dataset for AI trainingglobal city image recognition biasesimpact of training data on urban AI perceptionimplications of AI perceptual biases for urban sustainabilityinfluence of digital data on AI urban judgmentsinfluence of socioeconomic factors on AI urban perceptionMultimodal large language models geographic biasperceptual bias in AI urban analysissystematic bias in multimodal AI modelsurban perception bias in AIurban safety and walkability assessment AIvisual perception in AI across cities
Share26Tweet16
Previous Post

Separating cumulative and differential land subsidence with machine learning

Next Post

How oblique bedding gives rockslides lateral resistance: Shanyang case study

Related Posts

Separating cumulative and differential land subsidence with machine learning
Social Science

Separating cumulative and differential land subsidence with machine learning

September 11, 2026
Family Involvement Shapes New Standards for Early Childhood Care Quality
Social Science

Family Involvement Shapes New Standards for Early Childhood Care Quality

September 11, 2026
Green Spaces in South Korea Buffer Negative Emotions, Study Finds
Social Science

Green Spaces in South Korea Buffer Negative Emotions, Study Finds

September 11, 2026
Why Teens Freeze or Fight Back: Inside the Mind of the Bullying Bystander
Social Science

Why Teens Freeze or Fight Back: Inside the Mind of the Bullying Bystander

September 11, 2026
What Doctors Wear Shapes Patient Trust and Infection Fears in Sri Lanka
Social Science

What Doctors Wear Shapes Patient Trust and Infection Fears in Sri Lanka

September 11, 2026
AI Panel Helps Build a Readiness Test for Adaptive Moodle Courses
Social Science

AI Panel Helps Build a Readiness Test for Adaptive Moodle Courses

September 11, 2026
Next Post
How oblique bedding gives rockslides lateral resistance: Shanyang case study

How oblique bedding gives rockslides lateral resistance: Shanyang case study

  • Mothers who receive childcare support from maternal grandparents show more optimized

    Mothers who receive childcare support from maternal grandparents show more parental warmth, finds NTU Singapore study

    27656 shares
    Share 11059 Tweet 6912
  • University of Seville Breaks 120-Year-Old Mystery, Revises a Key Einstein Concept

    1061 shares
    Share 424 Tweet 265
  • Bee body mass, pathogens and local climate influence heat tolerance

    682 shares
    Share 273 Tweet 171
  • Researchers record first-ever images and data of a shark experiencing a boat strike

    546 shares
    Share 218 Tweet 137
  • Groundbreaking Clinical Trial Reveals Lubiprostone Enhances Kidney Function

    531 shares
    Share 212 Tweet 133
Science

Embark on a thrilling journey of discovery with Scienmag.com—your ultimate source for cutting-edge breakthroughs. Immerse yourself in a world where curiosity knows no limits and tomorrow’s possibilities become today’s reality!

RECENT NEWS

  • How oblique bedding gives rockslides lateral resistance: Shanyang case study
  • Multimodal LLMs show geographic and perceptual bias across 200 cities
  • Separating cumulative and differential land subsidence with machine learning
  • Density separation recovers microplastics from soil despite aging effects

Categories

  • Agriculture
  • Anthropology
  • Archaeology
  • Athmospheric
  • Biology
  • Biotechnology
  • Blog
  • Bussines
  • Cancer
  • Chemistry
  • Climate
  • Earth Science
  • Editorial Policy
  • Marine
  • Mathematics
  • Medicine
  • Pediatry
  • Policy
  • Psychology & Psychiatry
  • Science Education
  • Social Science
  • Space
  • Technology and Engineering

Subscribe to Blog via Email

Enter your email address to subscribe to this blog and receive notifications of new posts by email.

Join 5,151 other subscribers

© 2025 Scienmag - Science Magazine

Welcome Back!

Login to your account below

Forgotten Password?

Retrieve your password

Please enter your username or email address to reset your password.

Log In
No Result
View All Result
  • HOME
  • SCIENCE NEWS
  • CONTACT US

© 2025 Scienmag - Science Magazine

Discover more from Science

Subscribe now to keep reading and get access to the full archive.

Continue reading