In remote multilingual meetings, I’d rather hear a slightly delayed translation than a perfectly fluent sentence that quietly changed the meaning. A delay tells me the system is still working; polished audio invites me to treat ambiguity, omitted words, or a technical term as settled.
That matters for requirements, incident response, and especially consent. One wrong phrase can redirect an implementation or make agreement sound clearer than it was. Google Meet already treats a few seconds of delay as a completeness trade-off and lets people review earlier translated captions, but I’d like the audio to expose uncertainty too: optional captions, phrase-level markers, an omission warning, or a “verify source” control that replays the relevant original segment. Not one reassuring score for the whole sentence.
The usability cost is obvious: verification prompts and pauses interrupt turn-taking. Would you prefer visible uncertainty, a source-replay control, or a fast mode that hides those signals? Developers and multilingual users: what real example would change your mind?