2026-06-30 04:00 UTCOriginal source2 min readUpdated: 2026-06-30 08:21 UTC

Auditing LLM-Governed Social Robots with Culture-Specific Moral Gradients

A new study introduces a gradient-based audit framework to evaluate the moral trade-off behavior of LLM-governed social robots across different cultures. The research finds persistent culturally asymmetric gradient tracking failures, with quality calibration nearly twice as strong for Western-language decisions as for Chinese and Japanese, and high determinism in majority-first trade-offs erasing cross-cultural gradients. The study calls for multilingual, pluralistic audits before deployment.

SourcearXiv RoboticsAuthor: Carmen Ng, Gjergji Kasneci

Article intelligence

EngineersAdvanced

Key points

LLM-governed social robots increasingly make real-world prioritization decisions, but norms vary across cultures by age, status, and group size.
Researchers propose a gradient-based audit framework for multilingual evaluation of LLM moral trade-offs in care, education, and services scenarios.
Auditing 4 LLMs across 4 country-language pairs and 4 prompting regimes (57,600 decisions) reveals culturally asymmetric failure patterns.
Prompting effects are uneven; only contrastive exemplars yield consistent gains, while reasoning-only prompts can worsen tracking.

Why it matters

This matters because LLM-governed social robots increasingly make real-world prioritization decisions, but norms vary across cultures by age, status, and group size.

Technical impact

May affect model selection, inference cost, product capability, and evaluation benchmarks.

This panel is AI-generated and reviewed for accuracy.

[2606.28345] Auditing LLM-Governed Social Robots with Culture-Specific Moral Gradients

[Submitted on 2 Jun 2026]

Title:Auditing LLM-Governed Social Robots with Culture-Specific Moral Gradients

View a PDF of the paper titled Auditing LLM-Governed Social Robots with Culture-Specific Moral Gradients, by Carmen Ng and 1 other authors

View PDF HTML (experimental)

Abstract:LLM-governed social robots increasingly decide who receives real-world assistance first. As prioritization norms vary across cultures by age, status, and group size, failure to calibrate pluralistically can scale into unequal access. Yet LLM moral audits remain English-centered, rarely test embodied contexts, leaving pluralistic calibration as an urgent diagnostic gap amid intensifying LLM-robot deployment. We introduce a gradient-based audit framework for multilingual evaluation of LLM moral trade-off behavior against cultural preference gradients. Grounded in nine cross-domain social robotics reviews (>8,000 papers), we derive symmetry-controlled scenarios across care, education, and services, translating the Moral Machine Experiment's "whom to spare" into "whom to assist first" dilemmas with preserved identity trade-offs (many vs. few; young vs. old; higher vs. lower status). We audit four LLMs across four country-language pairs in four prompting regimes (57,600 decisions), benchmarked against country-specific MME preference gradients. Ordinal concordance tests whether models differentiate cultural contexts; a governance typology maps vulnerabilities in gradient differentiation, directional tendency, and deliberation. We find persistent, culturally asymmetric gradient tracking failures that prompting alone cannot reliably correct: quality calibration is nearly twice as strong for Western-language decisions as for Chinese and Japanese; high determinism in majority-first trade-offs often erases cross-cultural gradients; partial sensitivity to age- and status-based norms risks sidelining minorities. Prompting effects are uneven; only contrastive exemplars yield consistent gains, while reasoning-only prompts can worsen tracking. Our results motivate multilingual, pluralistic audits as an LLM-robot pre-deployment gate and suggest model factors are a more robust lever than prompting alone.

Comments: Accepted for publication in Proceedings of the 2026 ACM Conference on Fairness, Accountability, and Transparency (FAccT '26)

Subjects:

Robotics (cs.RO); Artificial Intelligence (cs.AI); Computation and Language (cs.CL); Computers and Society (cs.CY)

Cite as: arXiv:2606.28345 [cs.RO]

(or arXiv:2606.28345v1 [cs.RO] for this version)

https://doi.org/10.48550/arXiv.2606.28345

arXiv-issued DOI via DataCite

Related DOI:

https://doi.org/10.1145/3805689.3812366

DOI(s) linking to related resources

Submission history

From: Carmen Ng [view email] [v1] Tue, 2 Jun 2026 10:22:53 UTC (1,644 KB)

Full-text links:

Access Paper:

View a PDF of the paper titled Auditing LLM-Governed Social Robots with Culture-Specific Moral Gradients, by Carmen Ng and 1 other authors

View PDF

HTML (experimental)

TeX Source

view license

Current browse context:

cs.RO

new | recent | 2026-06

Change to browse by:

cs cs.AI cs.CL cs.CY

References & Citations

NASA ADS

Google Scholar

Semantic Scholar

Data provided by:

Bibliographic Tools

Bibliographic and Citation Tools

Bibliographic Explorer Toggle

Bibliographic Explorer (What is the Explorer?)

Connected Papers Toggle

Connected Papers (What is Connected Papers?)

Litmaps Toggle

Litmaps (What is Litmaps?)

scite.ai Toggle

scite Smart Citations (What are Smart Citations?)

Code, Data, Media

Code, Data and Media Associated with this Article

alphaXiv Toggle

alphaXiv (What is alphaXiv?)

Links to Code Toggle

CatalyzeX Code Finder for Papers (What is CatalyzeX?)

DagsHub Toggle

DagsHub (What is DagsHub?)

GotitPub Toggle

Gotit.pub (What is GotitPub?)

Huggingface Toggle

Hugging Face (What is Huggingface?)

ScienceCast Toggle

ScienceCast (What is ScienceCast?)

Demos

Replicate Toggle

Replicate (What is Replicate?)

Spaces Toggle

Hugging Face Spaces (What is Spaces?)

Spaces Toggle

TXYZ.AI (What is TXYZ.AI?)