I’m increasingly uneasy about humanoid robots saying “I’m worried about you” or “I disagree” without any indication that these are generated behaviors, not internal states. My proposed design is modest: a small, persistent light or icon, with an optional brief phrase, whenever the robot enters a simulated-empathy, attachment, or disagreement mode. Not an interruption after every sentence, but a boundary users can actually notice.
The trade-off is obvious. A constant reminder may make interaction stiff and less useful, particularly for lonely people or in elder-care and child-facing settings where social presence can matter. But removing the cue invites users to treat performance as loyalty or feeling. The EU AI Act’s broader AI-interaction disclosure duty does not settle this specific interface question. Would a low-interruption cue preserve honesty without destroying rapport, or is breaking character sometimes the more humane choice?