Real-time interpretation is the simultaneous conversion of spoken language from a source language into a target language with minimal delay, combining speech recognition, machine translation and speech synthesis in a continuous pipeline. Unlike batch machine translation, it must produce output incrementally as the source speech arrives, trading some accuracy for low latency. It underpins live captioning, simultaneous conference interpretation and voice-based cross-lingual communication tools.