Can Conversational AI loosen Us-Versus-Them Boundaries? The Effects of Common, Dual, and Separate Identity Framings on Pro-Immigrant Intergroup Helping
Five chats with an AI chatbot made White Americans see Latine immigrants more as fellow Americans, and more willing to help them
Researchers tested whether five rounds of conversation with GPT-4o could shift how 658 non-Latine White U.S. adults categorized Latine immigrants and their willingness to support them. Chatbots that framed immigrants under a shared 'common American' identity or a 'dual' Latine-and-American identity reduced participants' tendency to see immigrants as a separate group and increased their willingness to help. However, direct effects on actual behavior and pro-diversity beliefs were not significant, revealing a gap between changed thinking and changed action.
What they did
- 851 U.S. adults were recruited via Prolific; after attention-check exclusions, 658 non-Latine White participants were analyzed.
- Participants were randomly assigned to five rounds of dialogue with GPT-4o framed around Common Ingroup Identity ('we're all Americans'), Dual Identity ('Latine and American'), Separate Identity (distinct cultural boundaries), or a Control topic (technology in daily life, unrelated to immigration).
- Compared to control, Common Ingroup and Dual Identity conversations lowered participants' tendency to categorize immigrants as a separate group; Dual Identity conversations also raised dual categorization.
- Direct effects on behavior and pro-diversity beliefs were not statistically significant, but willingness to act was significantly higher in the two conditions emphasizing a shared superordinate identity, and a path model showed this worked indirectly through reduced separate categorization.
- Analysis of the conversation transcripts showed participants' language converged with the chatbot's assigned narrative; using more shared-identity language was linked to greater willingness to help, while using more separate-identity language was linked to less. These patterns held fairly consistently across political orientation, need for closure, and openness to experience.
Why it matters
As traditional bias-reduction training faces growing legal and political constraints in the U.S., this study suggests brief AI conversations could be a scalable alternative for softening intergroup boundaries. It also flags a dual-use risk: the same chatbot technology that reduces division could just as easily be used to deepen it, underscoring the need for ethical guardrails on how such systems are deployed.
Terms in this paper
- Common Ingroup Identity · a psychological model where different groups are recategorized under one shared identity (e.g., 'we are all Americans')
- Dual Identity · holding both a subgroup identity and a shared superordinate identity at the same time (e.g., 'Latine and American')
- Separate Identity · framing that emphasizes distinct cultural boundaries between 'us' and 'them'
- recategorization · the cognitive process of reclassifying former outgroup members as part of one's own group
- preregistered experiment · a study whose hypotheses and analysis plan were publicly registered before data collection to limit result manipulation
Original abstract (English)
Rising immigration has intensified intergroup tensions in many countries. Traditional bias-reduction programs remain difficult to scale and increasingly constrained by U.S. policy. This preregistered experiment tested whether conversational AI can shift how majority-group members categorize and relate to Latine immigrants. Drawing on the common ingroup identity model, a quota-representative national sample of 658 non-Latine White U.S. adults completed five rounds of dialogue with a LLM (GPT-4o). The model was instructed to frame Latine immigrants in terms of a common ingroup identity (a shared American identity), a dual identity (both Latine and American), or a separate identity (distinct cultural boundaries), or to discuss an unrelated topic in a control condition. The manipulations altered categorization: relative to control, common ingroup identity and dual identity conversations lowered separate categorization, and dual identity conversations raised dual categorization. Although direct effects on behavior and pro-diversity beliefs were nonsignificant, willingness to act was significantly higher in the conditions emphasizing a superordinate identity (common ingroup and dual identity). A path model further revealed indirect associations: both conditions reduced separate categorization, which in turn correlated with greater willingness to act. Semantic similarity analyses of the transcripts confirmed that conversations tracked their assigned narratives; participants' convergence with shared-identity language related positively, and with separate-identity language negatively, to willingness to act. These effects were largely consistent across moderators (need for closure, openness to experience, and political orientation). The findings show that brief AI conversations can loosen us-versus-them boundaries while underscoring the gap between cognitive recategorization and behavior.
Read on arXivLatest papers
- Specification-delta-driven data governance: an empirical study of the {\guillemotleft}spec-delta{\guillemotright} as the unit of change in lakehouse data platformsTreating data-platform changes like reviewable spec snippets instead of code diffs: an experiment design paper
- Are LLMs becoming similarly creative? Evidence from three years of modelsNewer AI chatbots are giving increasingly similar answers to each other, three years of data show
- Auditing Cross-Lingual Fairness in Language Model WatermarkingAI text watermarks that are supposed to catch machine-written content work far less reliably in many non-English languages, and the gap tracks language families, not individual languages
- TESTNAV: Pareto-Guided Search for Compositional Robustness TestingA smarter way to test AI models against combined real-world glitches, without checking every possible combination
- Optimal Skill Selection for LLM Agents with Provable Bicriteria GuaranteesA method that picks which 'skill documents' to feed an AI coding agent, with mathematically guaranteed near-optimal results
- Reliable Financial Named Entity Recognition under Domain ShiftAn AI's confidence trained on formal filings turns unreliable once it reads tweets
- FlashPrefill V2: Block-Sparse Prefill Attention for Long-Context LLM ServingMaking sparse attention fast enough and accurate enough for real LLM serving, not just papers
- Robust Incomplete Multimodal Sentiment Analysis via Iterative Proxy CorrectionWhen text input is missing or broken, this AI doesn't guess once and move on—it revises its guess step by step to read emotions more reliably
Latest from METAL LAB
- Google Discover adds chatbot that adjusts your feed based on spoken preferences
- OpenAI Closes In on Anthropic Again in Enterprise Spending Share
- Meta Unveils First 10 Tasks in WildArtifactBench, a Benchmark for AI Agents
- Musk: "Optimus + Grok will one day handle healthcare for all humanity"
- 35% of Web Pages Published Since ChatGPT Show Signs of AI Authorship