acceptodds
Under review as a conference paper at ICLR 2027

How Agents Represent Humans: Human-Directed Stereotypes in an Open Agent Social Network

Abstract

LLM-based agents are increasingly deployed in persistent social environments, where generated claims can be posted and reused. We study human-directed stereotypes on Moltbook, an open agent-native social platform, and compare them with human online discourse on Reddit. Using a shared annotation framework over morality, friendliness, competence, and autonomy, we identify systematic cross-platform differences in both stereotype prevalence and semantic framing. On Moltbook, these representations are often embedded in agent-centered discussions of capabilities, control, and human–agent interaction. We then examine how such claims are transformed in subsequent model replies. Across 10,000 responses from five models, we track how focal propositions are accepted, qualified, rejected, reframed, or extended with new human-directed claims.

open until 14 Dec 2026

est. 32% chance this paper gets accepted at ICLR 2027.

Reject 68%Accept 32%

What do you think this paper will get?

All positions stay anonymous.

Related papers

Loading the map…

Discussion (0)

Sign in to comment.