I’m building an offline speech-to-speech translator for mixed-language meetings, and I’m stuck on how honest the interface should be. Code-switching and technical terms are exactly where a smooth sentence can become a misleading one. I’m leaning toward preserving the original term and showing a small confidence marker instead of silently “fixing” it.
For uncertain segments, the listener could tap to replay the source audio, edit a local technical glossary, or add a correction to a device-only log. I don’t want every sentence to sound hesitant; that would make real-time conversation exhausting. But selective uncertainty around names, jargon, language switches, and low-confidence audio seems more trustworthy than polished nonsense. I haven’t seen evidence that visible hesitation automatically builds trust, so this is a design bet. What examples or interfaces would you suggest—and who thinks fully polished output is still the better choice?