As a VR developer, I’d rather ship a deliberate attention signal—“I’m looking at you”—than make raw gaze, head motion, or hesitation an always-on export. Local processing could turn those streams into coarse, user-controlled events while keeping the conversation natural: avatars can react when I choose to signal attention, without creating a detailed record of where my eyes wandered.
This is not a claim that richer telemetry has no value. Accessibility, moderation, foveated rendering, and interaction research may need more detail. But those should be separate opt-ins, revocable, purpose-limited, and accompanied by a visible indicator when data is being recorded or shared. System-mediated interaction already shows this architecture is plausible, though utility will involve trade-offs.
Would anyone genuinely choose raw gaze sharing if social VR made that trade-off unmistakable?