Back to Research papers
Research paper index

Benchmarking Identity-Sensitive LLM Outputs for Surveillance and Security Robots

Nneka Hyman, Jasmine Khan, Raj Korpan

arXiv:2608.16030Published August 17, 20260 citations
  • cs.RO
  • cs.CY
  • action
  • robot

Abstract

Large language models (LLMs) are increasingly used to generate textual robot design specifications, interaction policies, and risk assessments during early-stage robot development. Such outputs may influence how surveillance and security robots are conceptualized, documented, and ultimately implemented. This paper evaluates whether identity-conditioned prompts produce systematic differences in LLM-generated surveillance and security robot design descriptions. Using 236 demographic identity labels across single-label and model-augmented prompt conditions, we analyze readability as an initial benchmark for evaluating accessibility and identity-conditioned variation in generated robot design descriptions. The results show significant differences in readability across prompt conditions, design dimensions, and demographic identities. Although readability cannot determine whether an output is fair or socially appropriate, it provides an interpretable baseline within a broader benchmarking framework that also includes lexical, semantic, sentiment, syntactic, and fairness-focused analyses.

Read the original paper

This page indexes public paper metadata. The manuscript remains with its original publisher and authors.