Represent and generate speech with acoustic features, neural codecs, text-to-speech models, and speech-to-speech systems. Assess intelligibility, speaker similarity, latency, consent, and voice-cloning risks.